AI agent skills for PDF and document handling

Document skills cover the work that looks trivial and is not: extracting text from a PDF that has columns, filling a form, producing a .docx that opens correctly in Word. The best of them are explicit about the library they drive and what they cannot do — scanned pages need OCR, and a skill that pretends otherwise will quietly return nothing. If the output is going to a person rather than a parser, check whether the skill handles layout or only text.

People land here searching for “claude code pdf skill”, “agent skill to extract text from pdf”, “docx generation skill”.

117 PDF and document handling skills Refine in search →

How to install a PDF and document handling skill

  1. Compare the skills. Read the PDF and document handling skills below — each page shows the full SKILL.md, its safety verdict, and the capabilities it declares.
  2. Install it. Run npx skillmds@latest add <owner>/<name>. The CLI writes the skill into every agent directory it detects, or use --agent to pin one.
  3. Use it. Restart your agent. It loads the skill on demand the next time you ask for something that matches — you do not have to name the skill.

What makes a good PDF and document handling skill?

Document skills cover the work that looks trivial and is not: extracting text from a PDF that has columns, filling a form, producing a .docx that opens correctly in Word. The best of them are explicit about the library they drive and what they cannot do — scanned pages need OCR, and a skill that pretends otherwise will quietly return nothing. If the output is going to a person rather than a parser, check whether the skill handles layout or only text.

Every skill listed here is a plain SKILL.md file in the format Anthropic documents for Agent Skills, read unchanged by Cursor, OpenAI Codex and 60+ agents. Each passes a safety review before it is publicly listed; the verdict and declared capabilities are on every skill's page.

Frequently asked questions

What is an agent skill for PDF and document handling?

Document skills cover the work that looks trivial and is not: extracting text from a PDF that has columns, filling a form, producing a .docx that opens correctly in Word. The best of them are explicit about the library they drive and what they cannot do — scanned pages need OCR, and a skill that pretends otherwise will quietly return nothing. If the output is going to a person rather than a parser, check whether the skill handles layout or only text.

Do document skills work on scanned PDFs?

Only if they invoke OCR, and most do not by default. A scanned page is an image; a text-extraction skill will find no text and should say so rather than guessing. Skills with OCR support name the engine they use.

Which agents can use these PDF and document handling skills?

Any agent that reads SKILL.md files: Claude Code, Claude.ai, Cursor, OpenAI Codex, Windsurf, OpenCode and 60+ others. The format is not vendor-specific, so the same file works everywhere — each agent just keeps its skills in a different directory, listed at /agents.

Are these PDF and document handling skills free?

Yes. Searching, reading and installing skills on SkillMD is free and needs no account. Individual skills carry their own licence, shown on each skill's page.