Generate Eval

Generate a research-grounded LLM-as-a-Judge evaluator prompt for an AI system. Use when the user wants to build an eval, judge, scorer, or grader for their LLM app or agent — especially grounded in real production traces from the Progress Observability Platform. Triggers on "write an eval", "build a judge", "score my agent's outputs", "make a grader for these traces", "evaluate faithfulness/tool calls/tone".

telerik 0fae1cc 4 files · 25.3 KB Updated

File contents

telerik/observability-skills/tree/main/skills/generate-eval commit 0fae1cc4bb

Frequently asked questions

npx skillmds@latest add telerik/generate-eval