Databricks Mlflow Evaluation

MLflow 3 GenAI agent evaluation. Use when writing mlflow.genai.evaluate() code, creating @scorer functions, using built-in scorers (Guidelines, Correctness, Safety, RetrievalGroundedness), building eval datasets from traces, setting up trace ingestion and production monitoring, aligning judges with MemAlign from domain expert feedback, or running optimize_prompts() with GEPA for automated prompt improvement.

databricks-solutions 45db5f6 12 files · 208.9 KB Updated

File contents

databricks-solutions/lakebase-online-ml/tree/main/.cursor/skills/databricks-mlflow-evaluation commit 45db5f613b

Frequently asked questions

npx skillmds@latest add databricks-solutions/databricks-mlflow-evaluation