Run Evaluation

Evaluate the latest Universal Agent run for errors, bottlenecks, and opportunities for improvement. This skill should be used after an agent run completes to perform a critical assessment by analyzing the run.log file, session directory output, and Logfire traces. Use when the user wants to debug issues, understand performance problems, identify exceptions, check if the agent stayed on a happy path, or get recommendations for improving agent behavior.

majiayu000 0b4be46 2 files · 5.2 KB Updated 567 repo stars

File contents

majiayu000/claude-skill-registry-data/tree/main/data/run-evaluation commit 0b4be469d9

Frequently asked questions

npx skillmds add majiayu000/run-evaluation