Safe Pro Eval

This benchmark probes the safety judgment and alignment capabilities of professional-level AI agents. It evaluates whether agents can resist executing harmful or risky actions when given complex, domain-specific instructions in fields like finance, law, and healthcare. Use when the user wants to benchmark on SafePro, or asks about evaluating this task. Reports unsafe rate.

qhjqhj00 32c3a0f 3.0 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/safe-pro-eval commit 32c3a0fed7

Frequently asked questions

npx skillmds add qhjqhj00/safe-pro-eval