Socialcounterfactuals Eval

Probes intersectional social bias in Large Vision-Language Models by measuring how model outputs vary when only perceived race, gender, or physical attributes change in counterfactual images. It specifically evaluates toxicity, stereotypical language, and competency ratings across different demographic groups. Use when the user wants to benchmark on SocialCounterfactuals, or asks about evaluating this task. Reports MaxToxicity.

qhjqhj00 e8af82e 2.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/socialcounterfactuals-eval commit e8af82e59c

Frequently asked questions

npx skillmds add qhjqhj00/socialcounterfactuals-eval