Adaptmmbench Eval

Evaluates Vision-Language Models' ability to dynamically select between text-only and tool-augmented reasoning modes, and assesses the quality, efficiency, and final accuracy of their reasoning processes across multimodal domains. Use when the user wants to benchmark on AdaptMMBench, or asks about evaluating this task. Reports Accuracy.

qhjqhj00 84807a4 3.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/adaptmmbench-eval commit 84807a4115

Frequently asked questions

npx skillmds add qhjqhj00/adaptmmbench-eval