Unified Content Moderation Adjudication

Use this skill when an evaluation agent needs to moderate content, detect hate speech, identify offensive language, or spot implicit bias in human-ai interactions. Trigger it when users use casual language like 'clean up this chat', 'check for mean comments', 'is this post offensive?', 'spot the hidden racism', or 'scan for hateful slang across different cultures'.

dingxingdi Updated

File contents

dingxingdi/paper_fast_search_backup/tree/main/skill_bank_evolved/eval/skills/unified-content-moderation-adjudication commit fe96023192

Frequently asked questions

npx skillmds@latest add dingxingdi/unified-content-moderation-adjudication