Langsmith Evaluator

INVOKE THIS SKILL when building evaluation pipelines for LangSmith. Covers three core components: (1) Creating Evaluators - LLM-as-Judge, custom code; (2) Defining Run Functions - how to capture outputs and trajectories from your agent; (3) Running Evaluations - locally with evaluate() or auto-run via LangSmith. Uses the langsmith CLI tool.

langchain-ai Updated

File contents

langchain-ai/skills-benchmarks/tree/main/skills/main/langsmith-evaluator commit 9f70d1b03f

Frequently asked questions

npx skillmds@latest add langchain-ai/langsmith-evaluator-2