Chart Hqa Eval

Evaluates multimodal large language models' ability to perform counterfactual reasoning over chart visualizations. It probes whether models rely on parametric memory or truly understand the visual data when answering questions that contain hypothetical assumptions about the chart. Use when the user wants to benchmark on Chart-HQA, or asks about evaluating this task. Reports relaxed accuracy.

qhjqhj00 747ab1f 3.2 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/chart-hqa-eval commit 747ab1f2ce

Frequently asked questions

npx skillmds add qhjqhj00/chart-hqa-eval