Local Model Agent Evaluation

Evaluate local LLMs for agent/coding tasks — BenchLoop harnesses, coding agent selection, tool-calling format matching, and community-tested setups.

theheavenlyd3mon e4e73b8 2 files · 7.0 KB Updated 28 repo stars

File contents

theheavenlyd3mon/hermes-profiles/tree/main/profiles/mlops/skills/mlops/local-model-agent-evaluation commit e4e73b8422

Frequently asked questions

npx skillmds add theheavenlyd3mon/local-model-agent-evaluation