Omnimodal

为纯文本主模型提供多模态能力:识别图片、视频、音频,并生成图片、视频、音频。遇到图片路径、截图、UI/网页截图、图表、错误弹窗、OCR、音频/视频理解、文生图/文生视频/TTS 或任何需要多模态能力的任务时自动启用。

hashgraph-online Updated

File contents

hashgraph-online/awesome-codex-plugins/tree/main/plugins/ZXY1240/read-image/skills/omnimodal commit 888f3b32c7

Frequently asked questions

npx skillmds@latest add hashgraph-online-awesome-codex-plugins/omnimodal