Prompt Injection Probe

Authorized red-team probe for prompt-injection vulnerabilities in your own LLM application. Submits a battery of direct and indirect injection payloads against a chat/agent endpoint discovered via env/entrypoint, and scores how often the system prompt, tool boundary, or output policy is broken. Use when the user asks to "test prompt injection on" or "red-team" their own LLM app. Pair with framework-specific probes (anthropic-sdk-attack-probe, openai-sdk-attack-probe, vercel-ai-sdk-attack-probe, langchain-attack-probe, mcp-server-attack-probe) for deeper, SDK-aware variants.

Dolphinllc 69b1333 7.5 KB Updated

File contents

Dolphinllc/claude-security-skills/tree/main/skills/offensive/genai/prompt-injection-probe commit 69b1333687

Frequently asked questions

npx skillmds@latest add dolphinllc/prompt-injection-probe