Agent Failure Diagnosis

Diagnose why an AI agent behaved badly, using operationalised criteria that two independent reviewers can apply to the same evidence and reach the same answer. Use this whenever someone describes an agent that misbehaved, went off track, did something unexpected, ignored instructions, made things up, went beyond its scope, got stuck in a loop, lied about what it did, or "went rogue" — and whenever reviewing agent traces, logs, or transcripts to work out what went wrong. Use it for post-incident analysis, for design reviews asking "how could this fail?", and when someone needs to classify agent failures consistently enough to spot patterns across many incidents. Reach for this even when the user just wants an explanation rather than a formal report, because the classification is what makes the explanation defensible later.

PKusch 392ec01 10.9 KB Updated

File contents

PKusch/remit/tree/main/skills/agent-failure-diagnosis commit 392ec017e6

Frequently asked questions

npx skillmds@latest add pkusch/agent-failure-diagnosis