Evaluate Do

Expert guidance for ai-experiments — LLM benchmarking, parameter sweeps, model comparison, and pre-production evaluation of agents and functions.

dot-do 7e61f8b 2.3 KB Updated

File contents

dot-do/skills/tree/main/evaluate-do commit 7e61f8b98b

Frequently asked questions

npx skillmds@latest add dot-do/evaluate-do