Omnilabel Eval

This benchmark evaluates a model's ability to perform language-based object detection using dynamic, open-vocabulary label spaces. It specifically probes handling of free-form text descriptions, negative examples (descriptions referring to zero objects), and multi-instance references within a single image. Use when the user wants to benchmark on OmniLabel, or asks about evaluating this task. Reports harmonic_mean_AP.

qhjqhj00 542cdca 4.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/omnilabel-eval commit 542cdca477

Frequently asked questions

npx skillmds add qhjqhj00/omnilabel-eval