Llava Onevision 1.5 Eval

This evaluation probes the multimodal reasoning, visual question answering, OCR, and chart understanding capabilities of large multimodal models. It tests the model's ability to process high-resolution images, extract fine-grained text, and perform complex reasoning across diverse visual domains. Use when the user wants to benchmark on MMStar, MMEBench, MME-RealWorld, SeedBench, CV-Bench, RealWorldQA, MathVista, WeMath, MathVision, MMMU, MMMU-Pro, ChartQA, CharXiv, DocVQA, OCRBench, AI2D, InfoVQA, PixmoCount, CountBench, VL-RewardBench, V*, or asks about evaluating this task. Reports accuracy / benchmark-specific score.

qhjqhj00 7a5b700 4.0 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/llava-onevision-1.5-eval commit 7a5b700897

Frequently asked questions

npx skillmds add qhjqhj00/llava-onevision-1-5-eval