Medconceptsqaeval

Evaluates large language models' ability to reason about and identify medical concepts (diagnoses, procedures, drugs) across different semantic hierarchies and difficulty levels. Use when the user wants to benchmark on MedConceptsQA, or asks about evaluating this task. Reports accuracy.

qhjqhj00 628133f 2.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/medconceptsqaeval commit 628133f51d

Frequently asked questions

npx skillmds add qhjqhj00/medconceptsqaeval