Med Mim Eval

Evaluates medical vision-language models on multi-image reasoning tasks, including temporal understanding, cross-modal comparison, multi-view diagnosis, and co-reference resolution across longitudinal and multi-modality medical imaging data. Use when the user wants to benchmark on Med-MIM Benchmark, or asks about evaluating this task. Reports closed-type accuracy.

qhjqhj00 03eda06 2.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/med-mim-eval commit 03eda0690f

Frequently asked questions

npx skillmds add qhjqhj00/med-mim-eval