Evaluate Tool Calling Model

Evaluate tool selection, arguments, schemas, execution, observation use, recovery, and side effects. Use when comparing tool-capable checkpoints, prompts, parsers, agent loops, function schemas, or harnesses.

gaelic-ghost c76648a 3 files · 5.0 KB Updated

File contents

gaelic-ghost/socket/tree/main/skills/evaluate-tool-calling-model commit c76648a0b9

Frequently asked questions

npx skillmds@latest add gaelic-ghost/evaluate-tool-calling-model