Evaluating Skills With Models

Evaluate skills by executing them across sonnet, opus, and haiku models using sub-agents. Use when testing if a skill works correctly, comparing model performance, or finding the cheapest compatible model. Returns numeric scores (0-100) to differentiate model capabilities. Use when this capability is needed.

tomevault-io Updated

File contents

tomevault-io/skills-registry/tree/main/taisukeoe--agentic-ai-skills-creator--evaluating-skills-with-models commit 0721dc0c30

Frequently asked questions

npx skillmds@latest add tomevault-io/evaluating-skills-with-models