Audiomarathon Eval

Evaluates long-context audio understanding and inference efficiency across speech, sound, and music domains. It probes temporal dependency modeling, multi-hop reasoning, and memory/token-pruning scalability in Large Audio Language Models. Use when the user wants to benchmark on AudioMarathon, or asks about evaluating this task. Reports F1-score.

qhjqhj00 6204f5d 3.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/audiomarathon-eval commit 6204f5d94e

Frequently asked questions

npx skillmds add qhjqhj00/audiomarathon-eval