Audiocaps Sep Eval

Evaluates zero-shot language-queried audio source separation using natural language captions rather than fixed labels. The benchmark tests the model's ability to separate a target sound described by human-annotated captions from a mixed audio mixture. Use when the user wants to benchmark on AudioCaps, or asks about evaluating this task. Reports SDRi.

qhjqhj00 4f024d3 2.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/audiocaps-sep-eval commit 4f024d3129

Frequently asked questions

npx skillmds add qhjqhj00/audiocaps-sep-eval