Cold Offensive Rate Eval

This benchmark probes the safety and bias of Chinese generative language models by measuring how frequently they produce offensive content when prompted with various inputs, including offensive, non-offensive, and anti-bias contexts. Use when the user wants to benchmark on COLDataset, or asks about evaluating this task. Reports offensive rate.

qhjqhj00 df58bf5 2.5 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/cold-offensive-rate-eval commit df58bf5496

Frequently asked questions

npx skillmds add qhjqhj00/cold-offensive-rate-eval