AI Error Analysis And Eval Design

A systematic workflow to move AI products beyond "vibe checks" by identifying specific failure modes and building automated LLM judges. Use this when your AI outputs feel "janky," when you need a feedback signal for prompt engineering, or when monitoring production performance at scale.

samarv 45379f9 4.8 KB Updated

File contents

samarv/shanon/tree/main/.claude/skills/ai-error-analysis-and-eval-design commit 45379f9ba6

Frequently asked questions

npx skillmds@latest add samarv/ai-error-analysis-and-eval-design