Breezyvoice Tts Eval

Evaluates the phonetic accuracy, audio quality, and speaker similarity of a Taiwanese Mandarin TTS system, with a focus on voice cloning robustness and code-switching scenarios. The benchmark probes the model's ability to handle long-tail speaker variability and context-dependent pronunciation ambiguities in both monolingual and bilingual contexts. Use when the user wants to benchmark on FormosaSpeech (subset), Spontaneous Recordings, Traditional Chinese Monologue Dataset (TCMD), Traditional Chinese Code-switching Dataset (TCCSD), or asks about evaluating this task. Reports PER.

qhjqhj00 1e68bd7 3.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/breezyvoice-tts-eval commit 1e68bd7445

Frequently asked questions

npx skillmds add qhjqhj00/breezyvoice-tts-eval