Phoenix Evals

Build, run, validate, and operationalize Phoenix evaluators for AI and LLM applications in Python or TypeScript, including code evaluators, LLM judges, RAG evals, experiments, datasets, tracing, sampling, error analysis, and production guardrails. Use when the user asks for Phoenix evals, evaluator design, judge validation, experiments, or AI quality monitoring.

paulasilvatech 241bf78 35 files · 94.1 KB Updated

File contents

paulasilvatech/awesome-harness-primitives/tree/main/harness/github-copilot/skills/phoenix-evals commit 241bf78634

Frequently asked questions

npx skillmds@latest add paulasilvatech/phoenix-evals-2