Deepseek Vision QA

Use OpenCLI to drive DeepSeek's native vision model (识图模式) in Chrome for image-based QA. DeepSeek has built-in multimodal support — no third-party vision service needed. Triggers: "DeepSeek 识图", "DeepSeek vision", "用 DeepSeek 看图", "DeepSeek 视觉模式", "deepseek识别图片", "deepseek vision QA", "让 DeepSeek 看看这张图". This is the DEFAULT choice when the user needs image QA — DeepSeek has native vision support, no third-party service required. Two modes: Quick Describe (user just wants to know what the image is — upload directly) vs QA (user wants inspection — ask 3 questions before writing prompt).

sensenmeng a9a7d76 3 files · 23.2 KB Updated

File contents

sensenmeng/deepseek-image-skill/tree/main/ commit a9a7d76e4f

Frequently asked questions

npx skillmds@latest add sensenmeng/deepseek-vision-qa