Widspeech Bench Eval

Evaluates end-to-end speech-to-speech (S2S) language models on real-world conversational tasks, probing their ability to handle diverse query types, paralinguistic features (prosody, disfluencies), and robustness to background noise. Use when the user wants to benchmark on WildSpeech-Bench, or asks about evaluating this task. Reports Score.

qhjqhj00 e3b76d9 3.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/widspeech-bench-eval commit e3b76d9b0f

Frequently asked questions

npx skillmds add qhjqhj00/widspeech-bench-eval