Kormedmcqa V Eval

This benchmark evaluates vision-language models on multimodal medical reasoning using questions derived from the Korean Medical Licensing Examination. It probes the models' ability to integrate textual and visual evidence across diverse clinical imaging modalities, including cross-image reasoning when multiple scans are provided. Use when the user wants to benchmark on KorMedMCQA-V, or asks about evaluating this task. Reports accuracy.

qhjqhj00 94c618c 2.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/kormedmcqa-v-eval commit 94c618c9f5

Frequently asked questions

npx skillmds add qhjqhj00/kormedmcqa-v-eval