LLM Evaluator

LLM-as-a-Judge evaluator via Langfuse. Scores traces on relevance, accuracy, hallucination, and helpfulness using GPT-5-nano as judge. Supports single trace scoring, batch backfill, and test mode. Integrates with Langfuse dashboard for observability. Triggers: evaluate trace, score quality, check accuracy, backfill scores, test evaluator, LLM judge.

johnalbertini14-glitch Updated 1 repo stars

File contents

johnalbertini14-glitch/openclaw-skills/tree/main/skills/aiwithabidi/llm-evaluator-pro commit b949d88952

Frequently asked questions

npx skillmds@latest add johnalbertini14-glitch/llm-evaluator-2