Local Lipsync Pipeline

Chains local-GPU lip-sync for stylized or non-photoreal faces (anime, cartoons, CG avatars, puppets, VTubers): Wav2Lip, MuseTalk, LivePortrait, SadTalker, GFPGAN/CodeFormer paste-back, detector fallbacks, feathered mouth composite, CFR conform, and FFmpeg remux. Use when a real-face detector rejects stylized input, jaw clips, mouth seams, audio drift, or VRAM OOMs while chaining models. Not for photoreal talking-head SaaS or generic ffmpeg without sync; never load Wav2Lip, MuseTalk, and LivePortrait in one process on under 12 GB VRAM.

Kayforkind 23dd9c7 35.5 KB Updated

File contents

Kayforkind/skill-slice commit 23dd9c7c5a

Frequently asked questions

npx skillmds@latest add kayforkind/local-lipsync-pipeline