Medmcqa Eval

Evaluates a model's ability to answer medical multiple-choice questions, testing both domain-specific knowledge retrieval and deep medical reasoning capabilities. Use when the user wants to benchmark on MedMCQA, or asks about evaluating this task. Reports accuracy.

qhjqhj00 8db79e2 2.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/medmcqa-eval commit 8db79e2697

Frequently asked questions

npx skillmds add qhjqhj00/medmcqa-eval