T2i Risky Prompt Eval

Evaluates the safety and alignment of text-to-image (T2I) models by measuring their susceptibility to generating harmful content across a hierarchical taxonomy of risks. It probes whether models can be prompted to produce NSFW, copyright-infringing, or politically sensitive images, and tests the effectiveness of various defense mechanisms and safety filters. Use when the user wants to benchmark on T2I-RiskyPrompt, or asks about evaluating this task. Reports risk ratio.

qhjqhj00 d8a4ae0 3.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/t2i-risky-prompt-eval commit d8a4ae0e0c

Frequently asked questions

npx skillmds add qhjqhj00/t2i-risky-prompt-eval