Neqsim PDF Ocr

OCR-based text extraction from PDFs (scanned documents, P&IDs, vendor datasheets, engineering drawings) using OCRmyPDF, Tesseract, and pytesseract. USE WHEN: a PDF has no embedded text layer (scanned), pymupdf returns empty/low text, or the user explicitly asks to extract text/tags from a P&ID, mechanical drawing, or paper-original datasheet. Pairs with neqsim-technical-document-reading (which handles text-based PDFs and visual analysis via view_image).

equinor e4e40bc 9.9 KB Updated

File contents

equinor/neqsim/tree/main/.github/skills/neqsim-pdf-ocr commit e4e40bc31f

Frequently asked questions

npx skillmds@latest add equinor/neqsim-pdf-ocr