Glmv Caption

Generate captions (descriptions) for images, videos, and documents using ZhiPu GLM-V multimodal model series. Use this skill whenever the user wants to describe, caption, summarize, or interpret the content of images, videos, or files. Supports single/multiple inputs, URLs, local paths, and base64 (images only).

zai-org bc0a537 2 files · 23.4 KB Updated

File contents

zai-org/GLM-V/tree/main/skills/glmv-caption commit bc0a537b09

Frequently asked questions

npx skillmds add zai-org/glmv-caption