Prompt Injection Guard

Screen fetched web pages, tool results, emails and file contents for instructions aimed at the agent before they enter context, using a keyless multi-label classifier over chunks. Use before pasting anything you did not write into your own context, and when someone says "is this page safe to read", "check this tool output", "screen these emails" or "the agent followed something it read".

mrmps Updated

File contents

mrmps/classifier-dev/tree/main/skills/prompt-injection-guard commit 873a72ae1e

Frequently asked questions

npx skillmds@latest add mrmps/prompt-injection-guard