curated_sonic_talking_avatar_landscape_1024x576px
Curated workflow skill generated from Sonic Talking Avatar Landscape 1024x576px.json.
Capability Family
video_t2v_i2v_avatar
Inputs
- Optional runtime overrides supported by
run(...):promptnegative_promptwidth,heightseed,steps,cfgsampler_name,scheduler,denoiseserver,headers,api_prefix
Outputs
- Returns JSON with:
statusprompt_idoutput_images(includes image/video entries reported by Comfy history)
Model Requirements
- None detected from loader nodes.
Custom Node Requirements
comfyui-easy-usecomfyui-videohelpersuitepr-was-node-suite-comfyui-47064894
Links Extracted From Workflow Notes
- https://discord.com/invite/gggpkVgBf3
- https://drive.google.com/drive/folders/1QIIDvCDU-rp1ZB8qDA6NQqVn8F9WYMhE
- https://drive.google.com/drive/folders/1jI32B-2JX17seSGG0-MnZgUhCMHCEZlx
- https://drive.google.com/drive/folders/1oe8VTPUy0-MHHW2a_NJ1F8xL-0VN5G7W
- https://github.com/smthemex/ComfyUI_Sonic
- https://huggingface.co/openai/whisper-tiny/tree/main
- https://huggingface.co/stabilityai/stable-video-diffusion-img2vid-xt-1-1/tree/main
- https://huggingface.co/stabilityai/stable-video-diffusion-img2vid-xt/tree/main
- https://www.youtube.com/@pixaroma
Source
- Original:
comfy-data/workflows/Sonic Talking Avatar Landscape 1024x576px.json
Routing Metadata
- Family:
video_t2v_i2v_avatar - Input modalities:
audio, image - Output modalities:
image/png, video/mp4 - Model families:
other - Node count:
13 - Complexity score:
7 - Resource profile:
high - Estimated runtime:
slow (often 2-6 min depending on model/server load) - Max latent resolution hint:
NonexNone - Max sampler steps hint:
None
Detected Models
- None detected.
Detected Custom Nodes
comfyui-easy-usecomfyui-videohelpersuitepr-was-node-suite-comfyui-47064894
Runtime Warnings
- Audio generation may take longer on CPU-only or low-VRAM servers.
- Uses custom nodes; missing nodes can cause validation/runtime failures.
- Video workflow: usually slower and VRAM-intensive than still-image workflows.