Harness First

Diagnose and fix an unreliable, expensive, or unsafe LLM agent by auditing its harness (golden set, judge, cost caps, data layer, action approvals, tracing) before blaming or swapping the model. Use when someone says an agent is "burning tokens", "hallucinating", "brittle", gives inconsistent answers, asks whether to switch to a cheaper/better model, or wants to ship an agent or prompt change to customers.

getedgehq Updated

File contents

getedgehq/skills/tree/main/harness-first commit 27111af350

Frequently asked questions

npx skillmds@latest add getedgehq/harness-first