Belt

Operate the belt CLI to evaluate headless coding agents (Claude Code, Cursor, Codex, Gemini, and others) end to end. Use when the user asks to write or run eval scenarios, compare agents, score outputs with rules or LLM judges, register a new agent adapter, interpret reports or benchmark cards, or set up evals in CI. Also use when the user mentions agent-belt, scenario JSON, BELT_ env vars, llm_scorer_instruction, llm_scorer_evidence_files, TurnExpectation, or benchmark cards.

jfrog 45e2ba4 16.0 KB Updated

File contents

jfrog/agent-belt/tree/main/src/belt/.agents/skills/belt commit 45e2ba42e2

Frequently asked questions

npx skillmds@latest add jfrog/belt