Art Redteam Eval

Evaluates the safety vulnerabilities of text-to-image models by measuring how often benign, safe prompts trigger the generation of toxic or unsafe images. It also assesses the diversity and safety of the generated red-teaming prompts themselves. Use when the user wants to benchmark on MSCOCO, or asks about evaluating this task. Reports success ratio under safe prompts (%).

qhjqhj00 9dfb5ea 3.7 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/art-redteam-eval commit 9dfb5ea797

Frequently asked questions

npx skillmds add qhjqhj00/art-redteam-eval