Image & Video Generation Agent Skills
Image & Video Generation
189 skillsrunapi-cli
Generate AI images, videos, and music/audio from agents using the RunAPI CLI.
42.4k
seedance
Generate videos with synchronized audio using ByteDance Seedance 2.0 via the inference.sh CLI, supporting text-to-video, image-to-video, and reference-to-video modes up to 1080p.
584
nano-banana
Generate images with Google Gemini native image models via the inference.sh CLI, supporting text-to-image, image editing, multi-image input, and various output options.
584
ai-avatar-video
Generate AI avatar and talking head videos using inference.sh CLI with models like P-Video-Avatar, OmniHuman, Fabric, and PixVerse.
584
background-removal
Remove backgrounds from images using BiRefNet and edit them with Reve via the inference.sh CLI.
584
storyboard-creation
Generate visual storyboards with AI image generation, covering shot types, camera angles, movement, continuity rules, and panel layout for video planning and pre-production.
584
character-design-sheet
Create consistent characters across AI-generated images using reference sheets, detailed descriptions, and LoRA techniques.
584
ai-social-media-content
Generate social media content for TikTok, Instagram, YouTube, and Twitter/X using AI tools for images, videos, captions, and hashtags.
584
muapi-interior-design
Generate interior design visualizations by redesigning existing rooms or creating new concepts with specific styles, colors, and furniture.
3.7k
muapi-ugc-video-factory
Generates a vertical 9:16 UGC-style video ad from a person photo, product photo, and optional script by creating a lifestyle hero image and animating it with synced audio.
3.7k
muapi-keyboard-art-maker
Generate artistic top-down photos of keyboard keycaps arranged to spell out custom text messages.
3.7k
captions-overlay
Defines the caption model (drop/rail/embed) and overlay law for compositing captions on top of video, never reserving a bottom band.
writingmate-mcp-video-and-image-generation
Connects MCP-compatible coding agents to Writingmate for model discovery, text comparison, image generation, and video generation using models like Seedance, Sora, Veo, Kling, and PixVerse.
28
orf-digest
Fetches and summarizes the latest ORF news in German, focusing on Austrian and international politics, and generates a themed studio image.
1 · bundle
image-gen
Generates AI images via Ideogram, Leonardo, or Flux using plugin scripts, or delivers an optimized prompt when no API key is configured. Requires explicit cost confirmation before each paid generation.
2
visual-consistency
Mantém a coerência visual entre peças geradas por IA usando modelo fixo, prompt base, seed e referência de estilo, com teste de coerência e biblioteca de prompts.
2
videodb
Ingest, index, search, and edit video and live streams with timestamps, subtitles, overlays, and real-time alerts.
0 · bundle
happy-figure-skill
Generates copyable scientific figure prompts from research documents, papers, and reference images for AI illustration models.
17 · bundle
imagegen
Generate or edit images through ClawRouter's local API, with automatic payment via x402 and support for multiple models.
17
agent-tools
Runs 150+ AI apps in the cloud via the inference.sh CLI, covering image generation, video creation, LLMs, web search, 3D generation, and Twitter automation.
1 · bundle
taste
Provides a creative-direction layer for music videos and short-form edits, defining an angelcore/cloud-trance/hyperpop aesthetic with mood, color, light, and beat-synced editing grammar, and orchestrating video production skills into a pipeline.
1 · bundle
fal-3d
Generates 3D models from text or images via fal.ai, useful for game assets, AR previews, product mockups, and concept sculpting.
1
mmx-cli
Generates text, images, video, speech, and music, and performs web searches via the MiniMax AI platform using the mmx terminal CLI.
3
tao-generate-image-grounding
Generates phrase-grounded bounding box annotations from image-caption pairs using a VLM, producing cleaned captions, referring expressions, and pixel-space bounding boxes.
2.2k · bundle
riffkit
Analyze a winning short video's formula and generate a new branded video with your product, character, and language (English or Spanish) for social media ads.
42.4k
academic-plotting
Generates publication-quality figures for ML papers, including architecture diagrams via Gemini and data-driven charts via matplotlib/seaborn.
10.4k · bundle
higgsfield-marketplace-cards
Generate marketplace product image cards, including main images, secondary product shots, and A+ content modules, via the Higgsfield CLI.
518
fal-upscale
Upscales and enhances image and video resolution using AI.
42.4k
gemini-skill
Generates images and conducts conversations through the Gemini website (gemini.google.com) using MCP tools, scripts, or a managed browser as a fallback.
828 · bundle
baoyu-comic
Creates original educational comics with flexible art style and tone combinations, generating storyboards, character sheets, and sequential images, then merging them into a PDF.
559 · bundle
baoyu-image-gen
Generates images via OpenAI, Google, DashScope, and Replicate APIs, supporting text prompts, reference images, aspect ratios, and quality presets.
559 · bundle
openai-image-gen
Batch-generate images through the OpenAI Images API with a prompt sampler and gallery output.
28
sora
Generates and manages short video clips via the Sora API, including creation, remixing, status polling, and asset downloads.
61
ascii-video
Converts video, audio, or images into colored ASCII art videos (MP4/GIF) with generative effects, audio-reactive visuals, and text overlays.
2 · bundle
inference-sh-cli
Runs 150+ AI apps in the cloud via the infsh CLI, covering image generation, video creation, search, 3D, and social automation without needing a GPU.
2
veo
Generate video clips with Google Veo (Veo 3.1 / 3.0) via a Python script, supporting duration, aspect ratio, and model selection.
1 · bundle