Ocr

Optical Character Recognition for scanned documents and images. Covers Tesseract, EasyOCR, PaddleOCR, AWS Textract, Azure Document Intelligence, Google Document AI, and Anthropic Claude vision. Preprocessing (deskew, denoise, binarize). USE WHEN: user mentions "OCR", "scanned PDF", "Tesseract", "Textract", "EasyOCR", "PaddleOCR", "Document AI", "Claude vision OCR", "image to text" DO NOT USE FOR: born-digital PDFs with selectable text - use `pdf-extraction`; table-only extraction - use `table-extraction`

claude-dev-suite Updated 28 repo stars

File contents

claude-dev-suite/claude-dev-suite/tree/main/skills/document-processing/ocr commit 5b077ea361

Frequently asked questions

npx skillmds@latest add claude-dev-suite/ocr