Clevr Ref Plus Eval

This benchmark evaluates a model's ability to comprehend referring expressions in synthetic visual scenes. It probes compositional visual reasoning by measuring how well models localize objects based on text descriptions that vary in attribute complexity, spatial relationships, and reasoning topology. Use when the user wants to benchmark on CLEVR-Ref+, or asks about evaluating this task. Reports IoU.

qhjqhj00 63a07e9 3.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/clevr-ref-plus-eval commit 63a07e91b5

Frequently asked questions

npx skillmds add qhjqhj00/clevr-ref-plus-eval