Ocr

OCR image files using the Qwen VL model via the bundled scripts/ocr.py (lives in this skill's directory, NOT your cwd). Use when the user asks to extract text from an image, perform OCR on a photo or screenshot, or recognize characters in an image file (.jpg, .jpeg, .png, .gif, .webp). Requires QWEN_API_KEY and QWEN_BASE_URL (env vars, or a .env in the skill directory).

jin-bo Updated

File contents

jin-bo/agentao/tree/main/examples/skills/ocr commit 9b8ae027d5

Frequently asked questions

npx skillmds@latest add jin-bo/ocr