Full Duplex Bench Eval

Evaluates real-time interactive behaviors in full-duplex spoken dialogue models. It specifically probes turn-taking, pause handling, backchanneling, and interruption management capabilities without relying on human studies. Use when the user wants to benchmark on Full-Duplex-Bench, or asks about evaluating this task. Reports descriptive metrics.

qhjqhj00 124d986 2.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/full-duplex-bench-eval commit 124d986f0f

Frequently asked questions

npx skillmds add qhjqhj00/full-duplex-bench-eval