Image & Video Generation
189 skillsSynapse Image Describe
Provides detailed, structured image descriptions covering objects, people, colors, text, and scene context, with an overview and interpretation.
14
AI Media Generation
Generates optimized prompts for AI video and image generation across multiple platforms, using cinematic styles and behavioral frameworks for commercial outcomes.
2 · bundle
Imagen
Generates images from text prompts using Google Gemini's image generation model, saving them as PNG files for use in UI, documentation, and design assets.
2
Videodb
Ingests video and audio from files, URLs, RTSP feeds, or desktop capture; indexes and searches moments with timestamps; transcodes, edits timelines, generates media assets, and creates real-time alerts for live streams.
0 · bundle
Fal AI Media
Generates images, videos, and audio using fal.ai models via MCP, covering text-to-image, text/image-to-video, text-to-speech, and video-to-audio.
0
Comfyui
Generate images, video, and audio with ComfyUI — install, launch, manage nodes/models, run workflows with parameter injection. Uses the official comfy-cli for lifecycle and direct REST/WebSocket API for execution.
2
Clip
Enables zero-shot image classification, image-text matching, and cross-modal retrieval using OpenAI's CLIP model, with code for semantic search, content moderation, and vector database integration.
2
Llava
Runs the open-source LLaVA vision-language model for image understanding, captioning, visual question answering, and multi-turn image conversations, including setup, inference, and training guidance.
2
Fal AI
Generates and edits images and videos via fal.ai's queue-based API, supporting models like Flux, Gemini image, and Kling video-to-video, with automatic polling.
1 · bundle
Fal
Search, explore, and run fal.ai generative AI models for image, video, audio, and 3D generation, including queue management and file uploads.
1 · bundle
Video Gen
Produz vídeo curto com roteiro, storyboard, prompts por cena e geração via Runway quando há chave, ou especificação manual para Higgsfield.
2
Fal AI Media
Generates images, videos, and audio using fal.ai models via MCP, covering text-to-image, text/image-to-video, text-to-speech, and video-to-audio.
1
Imagen
Generates images from text prompts using Google Gemini's image generation model, saving them as PNG files for use in UI, documentation, and design assets.
0 · bundle
Vox Explainer
Produces a complete narrated, subtitled, scored explainer video from a single topic prompt using a six-stage pipeline with script, voiceover, keyframes, animation, music, and local assembly.
17 · bundle
AI Media Generator
Generates high-quality prompts for AI image, video, and music generation platforms, with optional browser automation to submit them to target sites.
17 · bundle
Vgl
Define every visual attribute as structured VGL JSON for deterministic, reproducible image generation with Bria FIBO models, covering objects, lighting, camera settings, composition, and style.
17 · bundle
Video Claw
Generates complete AI videos through a 6-stage pipeline (script, character/scene design, storyboard, reference images, video generation, post-production) or one-shot pipelines for short videos, action transfer, and digital human dubbing, all running on local servers.
17 · bundle
Bria AI
Generates, edits, and transforms images via the Bria API, including text-to-image, background removal, product photography, and batch processing for e-commerce catalogs.
17 · bundle
Drug Photo
Identifies a medication from a photo and generates a genotype-informed dosage card using CPIC guidelines and real 23andMe data.
17 · bundle
Comfyui Skill Openclaw
Runs ComfyUI image-generation workflows from an AI agent via a CLI, handling imports, dependencies, execution, and history.
17 · bundle
Videodb
Ingest, index, search, edit, and generate video and audio assets from files, URLs, RTSP feeds, or desktop capture, with real-time alerts and stream links.
1 · bundle
Sora
Generates, remixes, and manages short video clips via OpenAI's Sora API for cinematic shots, b-roll, and rapid concept iteration.
1
Imagen
Generates images with Google Gemini's image generation API for UI mockups, icons, illustrations, and visual assets.
1
Fal AI Media
Generates images, videos, and audio using fal.ai models via MCP, covering text-to-image, text/image-to-video, text-to-speech, and video-to-audio.
1
Video Gen
Generates short AI videos from text or images across Runway, Kling, Sora, Pika, Seedance, Grok Imagine, and Veo, with provider-specific APIs and prompt guidance.
10
Geminigen AI
Unified multimedia generation API for images, videos, and text-to-speech, replacing separate providers for a single workflow.
10
Xiaohongshu
Generates Xiaohongshu (Little Red Book) posts — titles, body copy, and cover images — and manages publishing, searching, and interacting with the platform through a local MCP service.
1 · bundle
Veo
Generates video clips from text prompts using Google Veo models, with options for duration, aspect ratio, and model selection.
1 · bundle
Vgl
Generates structured VGL JSON for Bria FIBO models, giving deterministic control over objects, lighting, camera, composition, and style instead of natural language prompts.
1 · bundle
Imagen
Generates images from text prompts using Google Gemini's image generation model, saving them as PNG files for use in UI, documentation, and design projects.
3
Videodb
Ingests video and audio from files, URLs, and live streams, builds searchable visual and spoken indexes, edits timelines with subtitles and overlays, and generates real-time alerts.
3 · bundle
Videodb
Ingest, index, search, and edit video and audio content with timestamps, subtitles, overlays, and live-stream alerts.
5 · bundle
Atxp
Access ATXP's paid API tools for web search, AI image generation, music creation, video generation, and X/Twitter search via CLI or programmatic client.
2 · bundle
Sora
Generates, remixes, polls, lists, downloads, and deletes Sora videos via OpenAI's video API using a bundled CLI.
0 · bundle
Mmx CLI
Generate text, images, video, speech, and music via the MiniMax AI platform using the mmx CLI.
42.4k
Blockrun
Pays for external capabilities like image generation, real-time X/Twitter data, and alternative LLMs via micropayments without requiring API keys.
42.4k