Benchmark Author Skill Evals

Use when a reusable agent skill needs a small, scoreable eval pack with explicit success criteria, trigger prompts, deterministic checks, rubric-based qualitative grading, and extension hooks for regressions, thrashing, or permission drift.

LoogacyStudio 0cee854 3 files · 14.4 KB Updated

File contents

LoogacyStudio/skills/tree/main/.github/skills/benchmark-author-skill-evals commit 0cee854f91

Frequently asked questions

npx skillmds@latest add loogacystudio/benchmark-author-skill-evals