Scheduled upgrade in progress. Some pages may load slowly or ask you to retry.

AI Agent Security

Secure AI agents against prompt injection, tool abuse, and data exfiltration with defense-in-depth controls.

majiayu000 177b9fa 2 files · 1.9 KB Updated 567 repo stars

File contents

AI Agent Security

Protect agentic systems from adversarial input and unsafe tool execution.

Threats to Model

  • Prompt injection through untrusted content
  • Excessive permissions on tools and APIs
  • Data exfiltration via model responses
  • Cross-tenant context leakage

Security Controls

  1. Isolate tool execution with strict allowlists.
  2. Add policy checks before sensitive actions.
  3. Limit token scope and credential lifetimes.
  4. Apply output filtering for sensitive data.
  5. Log every privileged tool invocation.

Incident Readiness

  • Keep immutable audit trails for prompts and tool calls.
  • Build kill switches for high-risk tools.
  • Run regular red-team scenarios.

Related Skills

majiayu000/claude-skill-registry-data/tree/main/security/ai-agent-security commit 177b9fac5c

Frequently asked questions

npx skillmds add majiayu000/ai-agent-security