Voiceagentbench Eval

Evaluates speech language models and ASR-LLM pipelines on agentic speech tasks. It probes single/multi-tool orchestration, multi-turn dialogue, and safety refusal capabilities across multiple languages, including English, Hindi, and five Indic languages. Use when the user wants to benchmark on VoiceAgentBench, or asks about evaluating this task. Reports PF (Parameter Filling).

qhjqhj00 b4b826c 3.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/voiceagentbench-eval commit b4b826c3b1

Frequently asked questions

npx skillmds add qhjqhj00/voiceagentbench-eval