Mmscan Visual Grounding Eval

Evaluates a model's ability to localize specific 3D objects or regions within a large-scale scene based on complex natural language prompts. It probes spatial reasoning, attribute understanding, and multi-target grounding capabilities in 3D point cloud environments. Use when the user wants to benchmark on MMScan (3D Visual Grounding), or asks about evaluating this task. Reports gTop-k.

qhjqhj00 981de81 3.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mmscan-visual-grounding-eval commit 981de81933

Frequently asked questions

npx skillmds add qhjqhj00/mmscan-visual-grounding-eval