LLM Eval Router

Shadow-test local Ollama models against a cloud baseline with a multi-judge ensemble. Automatically promotes models when statistically proven equivalent — reducing API costs with evidence, not hope.

modbender Updated 12 repo stars

File contents

modbender/skill-library-mcp/tree/main/data/llm-eval-router commit ff9dc84891

Frequently asked questions

npx skillmds@latest add modbender/llm-eval-router