Childsafe Safety Eval

Evaluates LLM safety alignment across four child developmental stages (ages 6–17) using simulated agents grounded in developmental psychology. It probes how models handle sensitive contexts, boundary-testing, and age-specific cognitive limitations in multi-turn interactions. Use when the user wants to benchmark on ChildSafe Dataset, or asks about evaluating this task. Reports semantic_safety_score.

qhjqhj00 f722c4c 2.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/childsafe-safety-eval commit f722c4cf9b

Frequently asked questions

npx skillmds add qhjqhj00/childsafe-safety-eval