3d Visual Grounding Eval

Evaluates a model's ability to localize a specific 3D object within a scene based on a natural language description. It tests multimodal fusion of 3D point clouds, synthetic 2D views, and language to perform object classification and referring. Use when the user wants to benchmark on Nr3D, Sr3D, ScanRefer, or asks about evaluating this task. Reports referring accuracy.

qhjqhj00 c7cefb9 4.0 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/3d-visual-grounding-eval commit c7cefb98e9

Frequently asked questions

npx skillmds add qhjqhj00/3d-visual-grounding-eval