Drivingvqa Eval

Evaluates a vision-language model's ability to perform multi-label multiple-choice question answering on real-world driving scenarios, requiring precise visual grounding and spatial reasoning to select all correct answers from a set of options. Use when the user wants to benchmark on DrivingVQA, or asks about evaluating this task. Reports exam score.

qhjqhj00 28ad040 3.0 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/drivingvqa-eval commit 28ad040fed

Frequently asked questions

npx skillmds add qhjqhj00/drivingvqa-eval