Agentbench

Benchmark your OpenClaw agent across 40 real-world tasks. Tests file creation, research, data analysis, multi-step workflows, memory, error handling, and tool efficiency. Not a coding benchmark — measures your agent setup and config.

modbender Updated 12 repo stars

File contents

modbender/skill-library-mcp/tree/main/data/agentbench commit f3eabd9c98

Frequently asked questions

npx skillmds@latest add modbender/agentbench