Lip To Speech Eval

Evaluates a model's ability to synthesize high-fidelity, intelligible speech directly from visual lip movements. It probes perceptual audio quality, content accuracy, and speaker identity preservation in a cross-dataset generalization setting. Use when the user wants to benchmark on LRS3-TED, LRS2-BBC, or asks about evaluating this task. Reports WER.

qhjqhj00 2553dc2 4.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/lip-to-speech-eval commit 2553dc2a31

Frequently asked questions

npx skillmds add qhjqhj00/lip-to-speech-eval