Ocr And Documents

Extract text from PDFs and scanned documents. Use web_extract for remote URLs, pymupdf for local text-based PDFs, marker-pdf for OCR/scanned docs. For DOCX use python-docx, for PPTX see the powerpoint skill.

graniet 9e2bf6e 4 files · 11.8 KB Updated

File contents

graniet/kheish/tree/main/skills/productivity/ocr-and-documents commit 9e2bf6ead4

Frequently asked questions

npx skillmds@latest add graniet/ocr-and-documents