Model Benchmarking

Use when selecting a model for any production feature, or evaluating whether to switch models. Requires task-specific benchmarking — not leaderboard lookup. Blocks "GPT-4 is the best model" decisions.

RBraga01 Updated

File contents

RBraga01/builder-ai/tree/main/skills/model-benchmarking commit 2fde4f1e54

Frequently asked questions

npx skillmds@latest add rbraga01/model-benchmarking