Eval

Evaluate and rank agent results by metric or LLM judge for an AgentHub session. Use when the user runs /hub:eval or asks to score, compare, or pick a winner among completed AgentHub agents.

bestagentkits Updated

File contents

bestagentkits/agency-skills/tree/main/skills/claude-skills/eval commit 93cb25eb86

Frequently asked questions

npx skillmds@latest add bestagentkits/eval