# Ocr

> OCR skill for extracting text from images and PDFs. Use when you need to read text from screenshots, photos, scanned documents, or any image file. Supports Chinese, English, and 100+ languages.

- Skill: `mr-shaper/ocr` (Agent Skill, multi-file: 7 files)
- Install (CLI): `npx skillmds add mr-shaper/ocr`
- Raw SKILL.md: https://api.skillmd.com/api/skills/mr-shaper/ocr/raw
- Safety review: pending
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: Coding & Dev Tools
- Author: mr-shaper (https://skillmd.com/u/mr-shaper)
- Updated: 2026-09-09
- Page: https://skillmd.com/skills/mr-shaper/ocr

---


# OCR Skill

## Usage

To extract text from an image or PDF, run:

```bash
python3 "/Users/mrshaper/Library/Application Support/com.differentai.openwork/workspaces/starter/.opencode/skills/paddle-ocr/scripts/ocr.py" "/path/to/image.png"
```

## Options

| Option | Description |
|--------|-------------|
| `--prompt "text"` | Custom prompt (e.g., "Extract table as markdown") |
| `--fast` | Use faster PaddleOCR instead of DeepSeek-OCR |
| `--json` | Output as JSON format |

## Examples

```bash
# Basic OCR
python3 scripts/ocr.py image.png

# Extract table as markdown
python3 scripts/ocr.py table.png --prompt "Extract this table as markdown"

# Fast mode
python3 scripts/ocr.py image.png --fast

# PDF OCR
python3 scripts/ocr.py document.pdf
```

## Supported Formats

Images: PNG, JPG, JPEG, BMP, GIF, WEBP, TIFF
Documents: PDF

