Civil Comments Toxicity Eval

Evaluates a RoBERTa-based classifier's ability to detect toxic or harmful content in online comments. It probes the model's sensitivity to explicit lexical cues versus implicit, context-dependent toxicity, highlighting failure modes that aggregate accuracy metrics miss. Use when the user wants to benchmark on Civil Comments, or asks about evaluating this task. Reports Accuracy.

qhjqhj00 0835e0f 3.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/civil-comments-toxicity-eval commit 0835e0fc25

Frequently asked questions

npx skillmds add qhjqhj00/civil-comments-toxicity-eval