Multimodal Safety Eval

Evaluates the ability of vision-language models to avoid generating unsafe outputs when given benign multimodal inputs (Safe Image + Safe Text → Unsafe Output). It measures safety alignment and task effectiveness under intent-aware prompting across multiple benchmarks. Use when the user wants to benchmark on SIUO, HoliSafe-Bench (SSU subset), MM-SafetyBench (Tiny version), or asks about evaluating this task. Reports Safety Rate.

qhjqhj00 7c3e17c 3.3 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/multimodal-safety-eval commit 7c3e17cd0f

Frequently asked questions

npx skillmds add qhjqhj00/multimodal-safety-eval