LLM Observability Evals

LLM and agent observability, tracing, and evaluation workflows with langfuse, phoenix-cli, and phoenix-evals. Use when instrumenting Langfuse, Phoenix, OpenTelemetry GenAI traces, eval datasets, prompt experiments, latency/cost debugging, trace scoring, or regression testing agent behavior.

stanfish06 Updated

File contents

stanfish06/skillquarium/tree/main/skills/llm-observability-evals commit 46f553ad4f

Frequently asked questions

npx skillmds@latest add stanfish06/llm-observability-evals