Asd And Se Eval

Evaluates a unified audio-visual model's ability to detect which speaker is actively speaking in multi-person video scenes and to enhance speech signals by removing background noise and interference. Use when the user wants to benchmark on AVA-ActiveSpeaker, LRS2, TalkSet, Columbia, MUSAN, or asks about evaluating this task. Reports mAP.

qhjqhj00 b66fb21 3.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/asd-and-se-eval commit b66fb2161e

Frequently asked questions

npx skillmds add qhjqhj00/asd-and-se-eval