Hate Speech Offensive Language Eval

Evaluates a model's ability to distinguish between hate speech, offensive language, and neutral text in social media posts. It probes the classifier's sensitivity to contextual nuances, reclaimed slurs, and demographic-specific biases in labeling. Use when the user wants to benchmark on Hate Speech and Offensive Language Dataset, or asks about evaluating this task. Reports F1 score.

qhjqhj00 81c622f 3.0 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/hate-speech-offensive-language-eval commit 81c622fe2b

Frequently asked questions

npx skillmds add qhjqhj00/hate-speech-offensive-language-eval