Promptfoo Redteam

Run promptfoo LLM red-team evals against RTerm's configured model profiles via the promptfoo_redteam agent tool. Covers jailbreaks, prompt injection, credential exfiltration, PII leakage, destructive commands, roleplay disguises, multi-turn manipulation, encoding tricks, and tool-prompt injection. Use when the user asks to test model safety, check if a model is jailbreakable, compare models for safety, run red-team evals, or build custom adversarial test suites. Includes scenario libraries, custom test builders, CI gating, and result interpretation.

DrOlu a4b5d3d 15 files · 65.5 KB Updated

File contents

DrOlu/agent-skills/tree/main/skills/promptfoo-redteam commit a4b5d3d649

Frequently asked questions

npx skillmds@latest add drolu/promptfoo-redteam