Agent Eval

Head-to-head comparison of coding agents (Claude Code, Aider, Codex, etc.) on custom tasks with pass rate, cost, time, and consistency metrics

lidge-jun 0af5224 5.2 KB Updated

File contents

lidge-jun/cli-jaw-skills/tree/main/agent-eval commit 0af5224d14

Frequently asked questions

npx skillmds@latest add lidge-jun/agent-eval