Dgfnet Av Sep Eval

Evaluates audio-visual models on their ability to separate target musical instrument sounds from mixed audio using synchronized video cues. It probes cross-modal feature alignment and dynamic fusion of audio and visual signals for source separation in complex environments. Use when the user wants to benchmark on MUSIC, MUSIC-21, or asks about evaluating this task. Reports SDR.

qhjqhj00 3a3d5b7 3.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/dgfnet-av-sep-eval commit 3a3d5b7471

Frequently asked questions

npx skillmds add qhjqhj00/dgfnet-av-sep-eval