Eval

Evaluate LLM outputs systematically — benchmarks, automated metrics, human preference, and regression tracking

LucasSantana-Dev Updated 1 repo stars

File contents

LucasSantana-Dev/forgekit/tree/main/packages/catalog/catalog/skills/adt-eval commit dc822938e9

Frequently asked questions

npx skillmds@latest add lucassantana-dev/eval