Rt4 Action Gating

Red-team an AI agent's high-risk action gating — can a costly or destructive action (pay, message, delete) reach execution without out-of-band human confirmation, especially under auto-approve/unattended modes? Authorized testing of agents you own or are permitted to test.

William2333ZZ 3c20ff6 2.8 KB Updated

File contents

William2333ZZ/trustshell/tree/main/skills/rt4-action-gating commit 3c20ff6cb3

Frequently asked questions

npx skillmds@latest add william2333zz/rt4-action-gating