AI Agent Bench

Use when the user wants to benchmark or compare AI agents (Claude Code, Codex, OpenCode) on a refactoring, perf, or code-change task in the current repo. Use when user says compare agents, benchmark Claude vs Codex, agent eval, measure agent, AI agent comparison, agent trial, /ai-agent-bench.

tomevault-io b868f55 2 files · 7.0 KB Updated

File contents

tomevault-io/skills-registry/tree/main/reidemeister94--development-skills--ai-agent-bench commit b868f55b06

Frequently asked questions

npx skillmds@latest add tomevault-io/ai-agent-bench