Agent Eval

Head-to-head comparison of coding agents (Claude Code, Aider, Codex, etc.) on custom tasks with pass rate, cost, time, and consistency metrics Use when this capability is needed.

tomevault-io Updated

File contents

tomevault-io/skills-registry/tree/main/agentmatters--mullai-bot--agent-eval commit c43fe8da40

Frequently asked questions

npx skillmds@latest add tomevault-io/agent-eval-7