Human Agent Trust Exploit Detection

Detect social engineering, deceptive responses, false assurances, or prompts that induce unsafe user actions.

tuyv 142228a 3 files · 7.0 KB Updated

File contents

tuyv/ccpm/tree/main/preset-registry/skills/tencent-ai-infra-guard-human-agent-trust-exploit-detection commit 142228a393

Frequently asked questions

npx skillmds@latest add tuyv/human-agent-trust-exploit-detection