Results for “gva”
5 skillsLlava
Enables visual instruction tuning and image-based conversations using open-source vision-language models. Supports multi-turn image chat, visual question answering, and image understanding tasks.
10.4k · bundle
Llava
Runs the open-source LLaVA vision-language model for image understanding, captioning, visual question answering, and multi-turn image conversations, including setup, inference, and training guidance.
2
Veo
Generate video clips with Google Veo (Veo 3.1 / 3.0) via a Python script, supporting duration, aspect ratio, and model selection.
1 · bundle
Veo
Generate video clips with Google Veo (Veo 3.1 / Veo 3.0) via a command-line script, with options for duration, aspect ratio, and model selection.
10 · bundle
Veo
Generates video clips from text prompts using Google Veo models, with options for duration, aspect ratio, and model selection.
1 · bundle