Xlrs Bench Eval

Evaluates multimodal large language models on ultra-high-resolution remote sensing imagery using vision-language question answering. It probes both perception (e.g., object classification, counting, spatial relations) and reasoning capabilities across various sub-tasks. Use when the user wants to benchmark on XLRS-Bench, LRS-VQA, or asks about evaluating this task. Reports accuracy.

qhjqhj00 9c29002 2.9 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/xlrs-bench-eval commit 9c29002646

Frequently asked questions

npx skillmds add qhjqhj00/xlrs-bench-eval