Xlist Hate Checklist Based Framework Interpretable

Decompose hate speech detection into a checklist of ten concept-level binary questions answered independently by an LLM, then aggregate results via a lightweight decision tree for interpretable, cross-dataset-robust classification. Use when asked to: 'build a hate speech detector', 'create an interpretable content moderation system', 'detect hateful content with explainability', 'classify toxic text with audit trails', 'implement checklist-based text classification', 'make a robust hate speech pipeline'.

ndpvt-web a6089f9 17.1 KB Updated

File contents

ndpvt-web/arxiv-claude-skills/tree/main/skills/xlist-hate-checklist-based-framework-interpretable commit a6089f9958

Frequently asked questions

npx skillmds@latest add ndpvt-web/xlist-hate-checklist-based-framework-interpretable