Evals

Write and analyze evaluations for AI agents and LLM applications. Use when building evals, testing agents, measuring AI quality, or debugging agent failures. Use this skill when you need to test the performance of an LLM or Agent, or if the user mentions EZVals. Use when this capability is needed.

tomevault-io Updated

File contents

tomevault-io/skills-registry/tree/main/camronh--ezvals--evals commit 833e00c458

Frequently asked questions

npx skillmds@latest add tomevault-io/evals