Mvbench Eval

Evaluates multi-modal large language models' ability to understand video content, with a strong focus on temporal perception and static-to-dynamic task transformation across 20 diverse categories ranging from basic perception to complex reasoning. Use when the user wants to benchmark on MVBench, or asks about evaluating this task. Reports accuracy.

qhjqhj00 bc08ed5 2.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mvbench-eval commit bc08ed5be4

Frequently asked questions

npx skillmds add qhjqhj00/mvbench-eval