# Grade agent trajectories and tool-use decisions with AgentEvals

> Score whether an agent took a sensible intermediate path, called tools correctly, and reached the outcome without relying only on final-answer checks.

- Skill: `agentskillexchange/grade-agent-trajectories-and-tool-use-decisions-with-agentev` (Agent Skill)
- Install (CLI): `npx skillmds@latest add agentskillexchange/grade-agent-trajectories-and-tool-use-decisions-with-agentev`
- Raw SKILL.md: https://api.skillmd.com/api/skills/agentskillexchange/grade-agent-trajectories-and-tool-use-decisions-with-agentev/raw
- Safety review: pending (external: skill-scanner PASS, skillspector PASS)
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: AI & ML
- Author: agentskillexchange (https://skillmd.com/u/agentskillexchange)
- Updated: 2026-09-08
- Page: https://skillmd.com/skills/agentskillexchange/grade-agent-trajectories-and-tool-use-decisions-with-agentev

---


# Grade agent trajectories and tool-use decisions with AgentEvals

Score whether an agent took a sensible intermediate path, called tools correctly, and reached the outcome without relying only on final-answer checks.

## Prerequisites

Python or TypeScript runtime, agent run outputs or trajectories, optional LLM judge provider

## Installation

Use the upstream install or setup path that matches your environment:
- pip install agentevals
- npm install agentevals @langchain/core
- pip install openai
- npm install openai

Requirements and caveats from upstream:
- <summary>Python</summary>
- python
- [Python Async Support](#python-async-support)

Basic usage or getting-started notes:
- To get started, install agentevals:
- <details open>
- bash

- Source: https://github.com/langchain-ai/agentevals
- Extracted from upstream docs: https://raw.githubusercontent.com/langchain-ai/agentevals/HEAD/README.md

## Documentation

- https://github.com/langchain-ai/agentevals

## Source

- [Agent Skill Exchange](https://agentskillexchange.com/skills/grade-agent-trajectories-and-tool-use-decisions-with-agentevals/)

