Admedtagger Medical Eval

Evaluates the ability of lightweight BERT-based models to classify Polish medical texts into five clinical categories. The benchmark tests knowledge distillation from a large LLM teacher to smaller classifiers, with ground truth curated by medical experts. Use when the user wants to benchmark on ADMEDTAGGER Physician-Validated Test Sets, or asks about evaluating this task. Reports F1 score.

qhjqhj00 7353ece 2.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/admedtagger-medical-eval commit 7353eceb63

Frequently asked questions

npx skillmds add qhjqhj00/admedtagger-medical-eval