Youcansee

Read, OCR, describe, or answer questions about local images when the current model cannot see images natively or native vision fails. Use the bundled scripts/see_image.py helper, which supports an OpenAI-compatible vision API, four quality tiers (max/high/medium/low), per-tier provider failover, local VLMs, and Windows OCR. Triggers include 读图、看图、OCR、截图内容、图片描述、图片里是什么; English triggers include read image、look at image、OCR、image text、screenshot content、 describe image、what is in this image、image question、visual question answering; tier selectors include 最强神眼/全力/完全体/最高档/max、最强、天花板, 高档/high/stronger/clearer、 中档/medium/normal/quick/balanced、 低档/low/local/cheap/save money/offline/fast。 Do not use this skill for image generation or editing.

Dithob 331e0bf 8 files · 17.1 KB Updated

File contents

Dithob/Youcansee/tree/main/ commit 331e0bf7f3

Frequently asked questions

npx skillmds@latest add dithob/youcansee