Jmmmu Pro Eval

Evaluates multimodal language models' ability to perform integrated visual-textual reasoning on Japanese-language tasks where questions and reference images are combined into a single composite image. It specifically probes OCR capabilities, visual perception, and cross-modal alignment in a multilingual context. Use when the user wants to benchmark on JMMMU-Pro, or asks about evaluating this task. Reports accuracy.

qhjqhj00 3969ec4 3.1 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/jmmmu-pro-eval commit 3969ec46e7

Frequently asked questions

npx skillmds add qhjqhj00/jmmmu-pro-eval