Eval Agent

Score any agent tool or platform against 3 structural questions (persistent memory, inspectable artifacts, compounding context) for a given use case, then generate a delegation spec that patches the weaknesses found. Use when evaluating agent tools, comparing platforms, or deciding whether to trust a tool with a specific workflow.

m2ai-portfolio f87fbbc 6.5 KB Updated

File contents

m2ai-portfolio/m2ai-skills-pack/tree/main/m2ai-build-tooling/skills/eval-agent commit f87fbbcece

Frequently asked questions

npx skillmds@latest add m2ai-portfolio/eval-agent