Phoenix Evals

Build, run, validate, and operationalize Phoenix evaluators for AI and LLM applications in Python or TypeScript, including code evaluators, LLM judges, RAG evals, experiments, datasets, tracing, sampling, error analysis, and production guardrails. Use when the user asks for Phoenix evals, evaluator design, judge validation, experiments, or AI quality monitoring.

paulasilvatech 9e3b667 35 files · 94.3 KB Updated

File contents

paulasilvatech/awesome-harness-primitives/tree/main/harness/claude-code/skills/phoenix-evals commit 9e3b667b90

Frequently asked questions

npx skillmds@latest add paulasilvatech/phoenix-evals