Ovis U1 Unified Multimodal

A 3B unified model combining image understanding, text-to-image generation, and image editing end-to-end rather than as separate frozen components. Use when you need a single efficient model for multiple vision-language tasks without the overhead of separate specialized systems.

adu2021 Updated

File contents

adu2021/skillxiv/tree/main/skills/skillxiv-v0.0.2-claude-opus-4.6/ovis-u1-unified-multimodal commit 83031f0cd0

Frequently asked questions

npx skillmds@latest add adu2021/ovis-u1-unified-multimodal