Mempot Defending Against Memory

Defend LLM agent memory systems against extraction attacks using optimized honeypot injection and sequential detection. Implements the MemPot framework: generates trap documents that lure attackers while staying invisible to legitimate users, then detects extraction attempts via Wald's Sequential Probability Ratio Test (SPRT). Trigger phrases: - "protect agent memory from extraction" - "add honeypots to my RAG memory" - "detect memory extraction attacks on my LLM agent" - "defend my vector database against adversarial queries" - "implement SPRT-based attack detection for my agent" - "harden my LLM agent's retrieval system"

ndpvt-web 6e43022 16.8 KB Updated

File contents

ndpvt-web/arxiv-claude-skills/tree/main/skills/mempot-defending-against-memory commit 6e43022886

Frequently asked questions

npx skillmds@latest add ndpvt-web/mempot-defending-against-memory