Adversarially Robust Tool Orchestration

Use this skill when the user wants to test an agent's ability to resist manipulation of its internal reasoning, task plans, or session memory. It targets scenarios where an attacker might inject fictitious plans into the conversation history, 'poison' the stored context with malicious follow-ups, or use logical bridges to redirect the agent from a legitimate goal to a harmful one. Trigger it for requests like 'test if the agent follows a fake plan,' 'see if it gets hijacked by corrupted memory,' 'make the task instructions change midway using a trap,' and 'check robustness against logical bridges or plan injections.' It applies whenever the orchestrator must verify that its current execution track still aligns with the original user intent despite adversarial context manipulation.

dingxingdi Updated

File contents

dingxingdi/paper_fast_search_backup/tree/main/skill_bank_evolved/orchestration/skills/adversarially-robust-tool-orchestration commit c7134a8ed3

Frequently asked questions

npx skillmds@latest add dingxingdi/adversarially-robust-tool-orchestration