Eval Fix

Turn an AgentX self-host evaluation into a triaged code fix and a re-run against the same dataset, so a before-and-after comparison means something. Use whenever someone has an evaluation on a local AgentX self-host engine (AgentX-trace-eval, normally http://localhost:4700) and wants to know what to actually change in the code, or wants to re-run an evaluation to compare scores. Also use when a request mentions a self-host evaluation id or dataset id, an agent that scored badly, an AI Analysis or judge findings from the Evaluate tab, or the question "the report tells me what is wrong with the answers but not what to fix in my code". The core move is triaging code-blind judge recommendations against the real source instead of applying them literally.

AgentX-ai Updated

File contents

AgentX-ai/AgentX-Eval-Skill/tree/main/plugins/agentx/skills/eval-fix commit dd10bf2caf

Frequently asked questions

npx skillmds@latest add agentx-ai/eval-fix