Agent Evals

Use when adding tests or evals for an LLM agent (LangGraph / langchain 1.x), or when "the tests pass but the agent regressed", "CI is paying for real model calls", "the event streamed but the panel was blank", or a tool-call trajectory must be asserted without real APIs. Pinned to langchain-core 1.6.3 / langchain 1.4.2 / langgraph 1.2.11 for the fake-model recipe.

andreasbloomquist Updated

File contents

andreasbloomquist/agent-skills/tree/main/skills/agent-evals commit 81c66468e3

Frequently asked questions

npx skillmds@latest add andreasbloomquist/agent-evals