Imagenetvc Eval

Evaluates zero- and few-shot visual commonsense reasoning capabilities of language models and visually-augmented language models across 1,000 ImageNet categories using human-annotated QA pairs. Use when the user wants to benchmark on ImageNetVC, or asks about evaluating this task. Reports Top-1 accuracy.

qhjqhj00 c914a40 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/imagenetvc-eval commit c914a401e7

Frequently asked questions

npx skillmds add qhjqhj00/imagenetvc-eval