Mmdeepresearch Bench Eval

Evaluates multimodal deep research agents on iterative retrieval, citation-grounded reasoning, and long-form report synthesis. It probes how well models align textual claims with visual evidence, maintain citation discipline, and produce high-quality structured reports under multimodal constraints. Use when the user wants to benchmark on MMDeepResearch-Bench, or asks about evaluating this task. Reports Overall MMDR-Bench Score.

qhjqhj00 34781cc 3.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/mmdeepresearch-bench-eval commit 34781ccd77

Frequently asked questions

npx skillmds add qhjqhj00/mmdeepresearch-bench-eval