Chat Failure Audit

Audit a real user transcript or agent session for failure modes — extract what went wrong, map each symptom to its root cause and the system layer that owns it, and rank by frequency × severity. Distinguishes symptom from cause (e.g. "it worked on the next turn" → root cause is activation not happening in-turn, not "the model is bad") and classifies each failure (capability gap, friction, state/lifetime bug, grounding/ confabulation, recovery gap, UX dead-end). Triggers on "analyse this chat", "what went wrong here", "audit this session", "why did the agent fail", "find the failure modes", "look at this transcript", "where did this break", "review this conversation". Use when you have a real session/log and want grounded failure analysis, not guesses.

hiteshbandhu Updated

File contents

hiteshbandhu/skills-i-use/tree/main/skills/chat-failure-audit commit 46b2230277

Frequently asked questions

npx skillmds@latest add hiteshbandhu/chat-failure-audit