Evals

Write and analyze evaluations for AI agents and LLM applications. Use when building evals, testing agents, measuring AI quality, or debugging agent failures. Use this skill when you need to test the performance of an LLM or Agent, or if the user mentions EZVals.

camronh Updated

File contents

camronh/evals-skill/tree/main/ commit 419df087c9

Frequently asked questions

npx skillmds@latest add camronh/evals