Redteam

Use when the user wants to test their LLM/agent application for safety and security vulnerabilities — jailbreaks, prompt injection, PII extraction, harmful content generation, or evaluator gaming. Also use when the user mentions security testing, adversarial testing, red teaming, safety evaluation, ASR (Attack Success Rate), or "is my app safe to deploy." Outputs ASR paired with over-refusal rate and an audit document.

agentscope-ai d95b623 2 files · 22.3 KB Updated

File contents

agentscope-ai/OpenJudge commit d95b623efe

Frequently asked questions

npx skillmds@latest add agentscope-ai/redteam