Agentic Eval

Use when designing and implementing evaluation loops for AI agents, including reflection, evaluator-optimiser patterns, rubric scoring, LLM-as-judge review, test-driven refinement, convergence checks, and iteration logging.

MarieLynneBlock 8aaa0ca 10.6 KB Updated

File contents

MarieLynneBlock/arcanum-artifex/tree/main/skills/agentic/agentic-eval commit 8aaa0cab67

Frequently asked questions

npx skillmds@latest add marielynneblock/agentic-eval