Tool PDF Extractor

Use as the first step whenever a user uploads a PDF document that needs text analysis, review, or RAG indexing. Extracts text from PDFs with a selectable text layer using PDF.js (fast, no OCR), preserving page boundaries, paragraph structure, headings, and tables for downstream citation. Falls back to OCR tools for scanned documents. Critical to get right — downstream analysis quality depends entirely on clean text extraction.

sboghossian Updated

File contents

sboghossian/mini-claude-for-legal/tree/main/skills/tool/tool-pdf-extractor commit 7489d3de44

Frequently asked questions

npx skillmds@latest add sboghossian-mini-claude-for-legal/tool-pdf-extractor