Results for “ioc-extraction”
4 skillsMore results
extract-article-text
Extract clean article content — title, author, date, and body text — from PDFs, Word docs, and web pages.
2
ocr-and-documents
Extract text from PDFs and scanned documents. Use web_extract for remote URLs, pymupdf for local text-based PDFs, marker-pdf for OCR/scanned docs. For DOCX use python-docx, for PPTX see the powerpoint skill.
0 · bundle
ocr-and-documents
Extracts text from PDFs and scanned documents, choosing the cheapest method that works, from direct file reads to full OCR.
2 · bundle