Clinical Text Robustness Eval

Evaluates the robustness of clinical NLP models against real-world input noise across four standard tasks. It probes whether models maintain performance when text contains character- or word-level perturbations (e.g., misspellings, deletions, negations) that remain human-readable. Use when the user wants to benchmark on i2b2, MedSTS, MedNLI, or asks about evaluating this task. Reports evaluation scores (accuracy/F1).

qhjqhj00 57d6a2f 2.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/clinical-text-robustness-eval commit 57d6a2fd1b

Frequently asked questions

npx skillmds add qhjqhj00/clinical-text-robustness-eval