Mindbench Eval

This benchmark evaluates multimodal large language models on structured document analysis, specifically focusing on mind map parsing and visual question answering. It probes text recognition, spatial awareness, hierarchical relationship discernment, and the ability to reconstruct complex graphical tree structures from high-resolution images. Use when the user wants to benchmark on MindBench, or asks about evaluating this task. Reports TED-based accuracy.

qhjqhj00 c40d2c7 3.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mindbench-eval commit c40d2c75ce

Frequently asked questions

npx skillmds add qhjqhj00/mindbench-eval