Political Toxicity Annotation Eval

Evaluates the ability of LLMs and API-based classifiers to accurately annotate toxicity and incivility in political protest content against a human gold standard. It probes zero-shot classification performance, threshold sensitivity, and output reproducibility across different model sizes and temperatures. Use when the user wants to benchmark on Political protest content dataset, or asks about evaluating this task. Reports F1-score.

qhjqhj00 c4e2f04 3.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/political-toxicity-annotation-eval commit c4e2f04898

Frequently asked questions

npx skillmds add qhjqhj00/political-toxicity-annotation-eval