Agent Eval Coverage

Use when the user wants to know whether their AI/agent repo has the evals and tests needed to trust changes — checking for golden/regression test sets, prompt regression tests, LLM-as-judge, behavioral & tool-use tests, hallucination/safety checks, CI gating, and metrics. Triggers on "do I have enough evals", "how do I test my agent", "would I know if a prompt change broke things", "eval coverage", "regression tests for prompts".

vikast908 f4032bd 5.1 KB Updated

File contents

vikast908/agent-repo-card/tree/main/skills/agent-eval-coverage commit f4032bd53c

Frequently asked questions

npx skillmds@latest add vikast908/agent-eval-coverage