Ocr And Documents

Extract text from PDFs and scanned documents. Use web_extract for remote URLs, pymupdf for local text-based PDFs, marker-pdf for OCR/scanned docs. For DOCX use python-docx, for PPTX see the powerpoint skill.

math-inc cf2e656 4 files · 10.4 KB Updated

File contents

math-inc/opengauss/tree/main/skills/productivity/ocr-and-documents commit cf2e656bde

Frequently asked questions

npx skillmds@latest add math-inc/ocr-and-documents