Eval Spec Authoring

Use before building anything whose output needs a quality bar, or when the user says "write a rubric", "eval spec", "how do we score this", "define what good looks like", "quality bar", "grading criteria", "our AI output is inconsistent". Encodes judgment as hard gates plus weighted dimensions with 1-to-5 anchors and a frozen calibration set. Writes workspace/evals/specs/<type>.md. Writes the rubric. Not for running it against an artifact (`eval-loop`) or retesting it against its anchors (`eval-calibration`).

guerrilla2799 4851e65 5.0 KB Updated

File contents

guerrilla2799/ops-and-scale-os/tree/main/skills/eval-spec-authoring commit 4851e65f09

Frequently asked questions

npx skillmds@latest add guerrilla2799/eval-spec-authoring