Document Ocr

Convert scanned PDFs and document images into clean Markdown using docling for layout (figures, tables, reading order) plus a vision-language OCR model. Use when a user needs high-quality OCR of scanned documents, historical literature, or photographed pages — preserving multi-column reading order, diacritics, special characters, and figures. Supports local vLLM/Ollama servers and cloud vision APIs (OpenAI, Anthropic). Assumes an OCR backend already exists.

brunoasm Updated

File contents

brunoasm/my_claude_skills/tree/main/document_ocr commit eddf2fba4d

Frequently asked questions

npx skillmds@latest add brunoasm/document-ocr