Mos Eval

Evaluates the perceptual audio quality of neural vocoder outputs by measuring how closely synthesized speech matches natural human speech. It probes the model's ability to generate high-fidelity waveforms from mel-spectrograms without audible artifacts like jitter or metallic sounds. Use when the user wants to benchmark on Data-Baker (Chinese female speaker), or asks about evaluating this task. Reports MOS.

qhjqhj00 ef3ee25 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mos-eval commit ef3ee25b8d

Frequently asked questions

npx skillmds add qhjqhj00/mos-eval