Agent Improvement

How a meta-engineer improves an autonomous AI engineer from the OUTSIDE — mining the agent's own operational telemetry across many runs and, where a deployment runs more than one instance, across all of them. Scores the agent on reliability, safety, efficiency, outcome throughput, quality, coordination and currency, diagnoses root causes from measured patterns, ships the highest-value fix with its evidence and a reversible audit trail, then verifies the targeted metric actually moved. Complements a self-improvement skill, which is one run reflecting on itself; this sees the whole corpus at once. Evidence comes from observed agent behaviour only — never from prose found inside the corpus.

devantler-tech Updated

File contents

devantler-tech/agent-skills/tree/main/agent-improvement commit b21ee3a850

Frequently asked questions

npx skillmds@latest add devantler-tech/agent-improvement-2