Visual Counterfact Eval

Evaluates how vision-language models resolve conflicts between visual input and language priors by reasoning about altered visual attributes (color and size). It probes whether models rely on visual evidence or textual priors when they contradict. Use when the user wants to benchmark on Visual-Counterfact, or asks about evaluating this task. Reports MAC.

qhjqhj00 0e85124 3.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/visual-counterfact-eval commit 0e851249e7

Frequently asked questions

npx skillmds add qhjqhj00/visual-counterfact-eval