Mmsi Bench Eval

This benchmark evaluates multimodal large language models' ability to perform multi-image spatial reasoning. It probes capabilities such as tracking object and camera motion, reconstructing scenes from multiple views, and inferring spatial logic across image sequences. Use when the user wants to benchmark on MMSI-Bench, or asks about evaluating this task. Reports accuracy.

qhjqhj00 65f7f10 2.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mmsi-bench-eval commit 65f7f10e6b

Frequently asked questions

npx skillmds add qhjqhj00/mmsi-bench-eval