Ocr

OCR skill for extracting text from images and PDFs. Use when you need to read text from screenshots, photos, scanned documents, or any image file. Supports Chinese, English, and 100+ languages.

mr-shaper 10544d0 7 files · 24.1 KB Updated

File contents

OCR Skill

Usage

To extract text from an image or PDF, run:

python3 "/Users/mrshaper/Library/Application Support/com.differentai.openwork/workspaces/starter/.opencode/skills/paddle-ocr/scripts/ocr.py" "/path/to/image.png"

Options

Option Description
--prompt "text" Custom prompt (e.g., "Extract table as markdown")
--fast Use faster PaddleOCR instead of DeepSeek-OCR
--json Output as JSON format

Examples

# Basic OCR
python3 scripts/ocr.py image.png

# Extract table as markdown
python3 scripts/ocr.py table.png --prompt "Extract this table as markdown"

# Fast mode
python3 scripts/ocr.py image.png --fast

# PDF OCR
python3 scripts/ocr.py document.pdf

Supported Formats

Images: PNG, JPG, JPEG, BMP, GIF, WEBP, TIFF Documents: PDF

mr-shaper/opencode-skill-hybrid-ocr/tree/main/ commit 10544d0f87

Frequently asked questions

npx skillmds add mr-shaper/ocr