Refcoco Grounding Eval

Evaluates fine-grained visual grounding and spatial reasoning by measuring how accurately a multimodal model can localize regions in images corresponding to given referring expressions under varying linguistic and spatial conditions. Use when the user wants to benchmark on RefCOCO, RefCOCO+, RefCOCOg, or asks about evaluating this task. Reports IoU@50 accuracy.

qhjqhj00 93db7b7 3.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/refcoco-grounding-eval commit 93db7b799f

Frequently asked questions

npx skillmds add qhjqhj00/refcoco-grounding-eval