Vlm Interaction Reasoning Eval

Evaluates vision-language models on general visual understanding, spatial/relational reasoning, and specifically interactional reasoning in dynamic scenes using a suite of standard VQA and scene understanding benchmarks. Use when the user wants to benchmark on VQAv2, VizWiz, TextVQA, GQA, VSR, RealWorldQA, MMT-Bench, SEEDBench, A-Bench, or asks about evaluating this task. Reports accuracy.

qhjqhj00 7475cf1 2.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/vlm-interaction-reasoning-eval commit 7475cf1d0b

Frequently asked questions

npx skillmds add qhjqhj00/vlm-interaction-reasoning-eval