Unicom Compressed Multimodal Representations

Compress visual embeddings into compact latent space for unified image understanding and generation. Combines attention-based compression with diffusion decoding to bridge comprehension and generation through a shared semantic bottleneck.

adu2021 e4de81a 11.4 KB Updated

File contents

adu2021/skillxiv/tree/main/skills/skillxiv-v0.0.2-claude-opus-4.6/unicom-compressed-multimodal-representations commit e4de81ae31

Frequently asked questions

npx skillmds@latest add adu2021/unicom-compressed-multimodal-representations