Zero Shot Voice Synthesis Eval

Evaluates zero-shot speech synthesis and novel voice generation by measuring how well models can produce intelligible, natural, and speaker-similar audio for unseen speakers using only conditioning embeddings. Use when the user wants to benchmark on English Multi-Accent Dataset (VCTK + Internal), or asks about evaluating this task. Reports WER.

qhjqhj00 003b884 3.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/zero-shot-voice-synthesis-eval commit 003b884484

Frequently asked questions

npx skillmds add qhjqhj00/zero-shot-voice-synthesis-eval