Eng Langfuse Eval Runner

Use when setting up or operating automated quality evaluation of legal AI skill outputs using Langfuse. Covers the evaluation dataset structure for legal domains, judge-model configuration, scoring rubrics specific to legal drafting and analysis, how to run batch evals across skill versions, and how to connect eval results to feature-flag promotion decisions. Engineering skill for legal AI quality assurance.

sboghossian Updated

File contents

sboghossian/mini-claude-for-legal/tree/main/skills/eng/eng-langfuse-eval-runner commit 623e1d050d

Frequently asked questions

npx skillmds@latest add sboghossian-mini-claude-for-legal/eng-langfuse-eval-runner