PDF Text Extractor

Extract text from PDF files with intelligent chunking and metadata preservation. For batch extraction (1K+ PDFs), use pdftotext (poppler) via subprocess — see pdf skill Tool Selection table. For single-doc quality, use Codex or PyMuPDF. Supports technical documents, standards libraries, research papers, or any PDF collection.

vamseeachanta 9bd6b97 3.4 KB Updated

File contents

vamseeachanta/workspace-hub/tree/main/.agents/skills/data/documents/pdf/text-extractor commit 9bd6b97a68

Frequently asked questions

npx skillmds@latest add vamseeachanta/pdf-text-extractor