Ming Moe Medical Eval

Evaluates large language models on a comprehensive suite of medical natural language processing tasks and medical licensing examinations. It probes the model's ability to process clinical text, perform information extraction, and demonstrate domain-specific knowledge and reasoning for medical exams. Use when the user wants to benchmark on CBLUE (via PromptCBLUE), MedQA, MedMCQA, CMB, CMExam, MMLU (medical subset), C-Eval (medical subset), CMMLU (medical subset), 2023 Chinese National Pharmacist Licensure Examination, or asks about evaluating this task. Reports score.

qhjqhj00 a171f8b 3.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/ming-moe-medical-eval commit a171f8b64e

Frequently asked questions

npx skillmds add qhjqhj00/ming-moe-medical-eval