Kormedmcqa Eval

Probes large language models' ability to answer multiple-choice questions derived from South Korean healthcare professional licensing exams. It evaluates domain-specific medical knowledge, regional clinical guideline adherence, and reasoning capabilities in Korean. Use when the user wants to benchmark on KorMedMCQA, or asks about evaluating this task. Reports accuracy.

qhjqhj00 31f5c87 2.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/kormedmcqa-eval commit 31f5c87bc5

Frequently asked questions

npx skillmds add qhjqhj00/kormedmcqa-eval