Scientific Figure Mcqa Eval

Evaluates a model's ability to perform high-level visual reasoning and domain-specific knowledge grounding on scientific figures within a multiple-choice question answering setting. It specifically probes whether models can resist choice-induced prior bias where text-only answer options incorrectly steer predictions away from visually supported ground truth. Use when the user wants to benchmark on MAC, SciFIBench, MMSci, or asks about evaluating this task. Reports Accuracy.

qhjqhj00 9afd38a 3.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/scientific-figure-mcqa-eval commit 9afd38a7ec

Frequently asked questions

npx skillmds add qhjqhj00/scientific-figure-mcqa-eval