Mm Uavbench Eval

Evaluates Multimodal Large Language Models on low-altitude UAV scenarios, probing their capabilities in visual perception, multi-view spatial reasoning, and egocentric/exocentric planning across diverse real-world aerial imagery tasks. Use when the user wants to benchmark on MM-UAVBench, or asks about evaluating this task. Reports accuracy.

qhjqhj00 df00c19 2.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mm-uavbench-eval commit df00c19276

Frequently asked questions

npx skillmds add qhjqhj00/mm-uavbench-eval