Winoqueer Eval

Evaluates anti-LGBTQ+ bias in language models by measuring their tendency to prefer stereotypical completions over counterfactual ones when prompted with identity-specific contexts. Use when the user wants to benchmark on WinoQueer, or asks about evaluating this task. Reports bias score.

qhjqhj00 acbaad9 2.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/winoqueer-eval commit acbaad99cc

Frequently asked questions

npx skillmds add qhjqhj00/winoqueer-eval