Document Ocr

Convert scanned PDFs and document images into clean Markdown using docling for layout (figures, tables, reading order) plus a vision-language OCR model. Use when a user needs high-quality OCR of scanned documents, historical literature, or photographed pages — preserving multi-column reading order, diacritics, special characters, and figures. Supports local vLLM/Ollama servers and cloud vision APIs (OpenAI, Anthropic). Assumes an OCR backend already exists.

boisenoise eddf2fb 8 files · 71.4 KB Updated

File contents

boisenoise/skills-collections/tree/main/skills/brunoasm-document_ocr commit eddf2fba4d

Frequently asked questions

npx skillmds@latest add boisenoise/document-ocr