Triplay Rl Tri Role Self Play Reinforcement

Apply the TriPlay-RL tri-role adversarial self-play framework to systematically red-team, harden, and evaluate LLM-powered applications for safety. Trigger phrases: 'red-team my LLM app', 'adversarial safety testing', 'tri-role safety audit', 'harden my chatbot against jailbreaks', 'evaluate LLM safety alignment', 'self-play safety loop'

ndpvt-web Updated

File contents

ndpvt-web/arxiv-claude-skills/tree/main/skills/triplay-rl-tri-role-self-play-reinforcement commit 085248e42b

Frequently asked questions

npx skillmds@latest add ndpvt-web/triplay-rl-tri-role-self-play-reinforcement