Agent Safety Guard

Design and implement safety guardrails for AI agent systems. Use this skill when building production agents that need protection against prompt injection, jailbreaks, data leakage, uncontrolled tool use, and other adversarial attacks. Includes red-team testing checklists, defense-in-depth architectures, and monitoring strategies.

viliawang-pm 04d0bf4 13.4 KB Updated

File contents

viliawang-pm/ai-engineering-toolkit/tree/main/skills/agent-safety-guard commit 04d0bf4b62

Frequently asked questions

npx skillmds add viliawang-pm/agent-safety-guard