Mos Bench Eval

This benchmark evaluates the out-of-domain generalization and robustness of subjective speech quality assessment (SSQA) models. It probes whether models trained on single or multiple datasets can accurately predict human-perceived quality scores across diverse conditions, including different languages, speech types (TTS, voice conversion, enhancement, noisy), and sampling frequencies. Use when the user wants to benchmark on MOS-Bench, or asks about evaluating this task. Reports Best score difference.

qhjqhj00 e55535b 4.2 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mos-bench-eval commit e55535b3bd

Frequently asked questions

npx skillmds add qhjqhj00/mos-bench-eval