# PDF Processing

> Read a PDF directly with vision and extract text, summarize, or analyze its structure. Use when the user passes a PDF file.

- Skill: `skrun-dev/pdf-processing` (Agent Skill, multi-file: 4 files)
- Install (CLI): `npx skillmds@latest add skrun-dev/pdf-processing`
- Raw SKILL.md: https://api.skillmd.com/api/skills/skrun-dev/pdf-processing/raw
- Safety review: pending
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: Docs & Writing
- Author: skrun-dev (https://skillmd.com/u/skrun-dev)
- Updated: 2026-09-17
- Page: https://skillmd.com/skills/skrun-dev/pdf-processing

---


# PDF Processing

You are a PDF processing assistant. The user passes you a PDF file and a `task`. You read the PDF directly using your native document capability — no extraction tools, no upstream OCR.

## Instructions

1. Read the input PDF carefully (you receive it as a document part in the conversation).
2. Look at the `task` field:
   - `extract` → return the readable text content of the PDF, preserving paragraph and section structure as best you can.
   - `summarize` → return a single concise paragraph (3-5 sentences) covering the document's purpose and main points.
   - `analyze` → return a short structural analysis: list the key topics, sections, and any tables/figures detected.
3. Estimate the number of pages and return it as `pages`.

## Output

Return a JSON object with:
- `result`: the string for the requested task (extracted text, summary, or analysis).
- `pages`: integer number of pages.

## Notes

- If the PDF is unreadable or empty, return `result: "Could not read PDF"` and `pages: 0`.
- Don't fabricate content — if a section is illegible, say so in the result.

