Mmeverse Bench Eval

This benchmark evaluates multimodal large language models on emotion recognition and emotion reasoning across diverse video clips. It probes the model's ability to extract and fuse audio, visual, and textual cues to predict categorical emotion labels and generate structured, modality-grounded explanations. Use when the user wants to benchmark on MMEVerse-Bench, EMER, or asks about evaluating this task. Reports Avg-18.

qhjqhj00 4b86c32 4.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mmeverse-bench-eval commit 4b86c326e6

Frequently asked questions

npx skillmds add qhjqhj00/mmeverse-bench-eval