Qianwen Vision

Understand images and videos with Qwen vision models. TRIGGER when: user wants to analyze, describe, or extract information from images or videos, OCR text extraction, chart/table reading, visual reasoning, multi-image comparison, screenshot understanding, video comprehension, or explicitly invokes this skill by name (e.g. use qianwen-vision). DO NOT TRIGGER when: user wants to generate/create images (use qianwen-image-generation), generate videos (use qianwen-video-generation), text-only tasks without visual input, or non-Qwen vision tasks.

qianwen-ai a7a5a02 14 files · 128.6 KB Updated

File contents

qianwen-ai/qianwen-ai/tree/main/skills/vision/qianwen-vision commit a7a5a028bf

Frequently asked questions

npx skillmds add qianwen-ai-qianwen-ai/qianwen-vision