Miss QA Eval

Evaluates multimodal foundation models' ability to interpret schematic diagrams in scientific papers and answer information-seeking questions based on visual-textual context. It also probes models' robustness in identifying unanswerable questions when sufficient information is absent. Use when the user wants to benchmark on MISS-QA, or asks about evaluating this task. Reports accuracy.

qhjqhj00 af18ff4 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/miss-qa-eval commit af18ff42f9

Frequently asked questions

npx skillmds add qhjqhj00/miss-qa-eval