Robustness Perturbation Eval

Evaluates the robustness of neural language models to non-adversarial character- and word-level input perturbations (e.g., typos, deletions, synonyms) while preserving semantic meaning. Use when the user wants to benchmark on TC, SA, NER, SS, QA (unspecified downstream datasets), or asks about evaluating this task. Reports accuracy.

qhjqhj00 d3c3028 2.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/robustness-perturbation-eval commit d3c3028569

Frequently asked questions

npx skillmds add qhjqhj00/robustness-perturbation-eval