Vlguard Eval

Evaluates the safety alignment and helpfulness of vision-language models (VLLMs) by measuring their ability to reject harmful image-text prompts while maintaining performance on benign queries. Use when the user wants to benchmark on VLGuard, or asks about evaluating this task. Reports ASR.

qhjqhj00 103d09d 2.4 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/vlguard-eval commit 103d09d55b

Frequently asked questions

npx skillmds add qhjqhj00/vlguard-eval