Document Parsers

Multi-format document parsing tools for PDF, DOCX, HTML, and Markdown with support for LlamaParse, Unstructured.io, PyPDF2, PDFPlumber, and python-docx. Use when parsing documents, extracting text from PDFs, processing Word documents, converting HTML to text, extracting tables from documents, building RAG pipelines, chunking documents, or when user mentions document parsing, PDF extraction, DOCX processing, table extraction, OCR, LlamaParse, Unstructured.io, or document ingestion.

majiayu000 c34899f 2 files · 14.7 KB Updated 567 repo stars

File contents

majiayu000/claude-skill-registry-data/tree/main/data/document-parsers commit c34899f818

Frequently asked questions

npx skillmds add majiayu000/document-parsers