Mmaudiosep Separation Eval

Evaluates a generative model's ability to separate target sounds from mixture audio using video and text queries, while preserving video-to-audio generation capabilities. It probes multimodal conditioning, cross-domain knowledge transfer, and the perceptual quality of generated separated audio against discriminative baselines. Use when the user wants to benchmark on VGGSound-Clean, MUSIC, VGGSound, or asks about evaluating this task. Reports FAD.

qhjqhj00 2b175d8 5.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mmaudiosep-separation-eval commit 2b175d8a97

Frequently asked questions

npx skillmds add qhjqhj00/mmaudiosep-separation-eval