Pinchbench

Run PinchBench benchmarks to evaluate OpenClaw agent performance across real-world tasks. Use when testing model capabilities, comparing models, submitting benchmark results to the leaderboard, or checking how well your OpenClaw setup handles calendar, email, research, coding, and multi-step workflows.

dixiyao e345437 252 files · 54.4 MB Updated

File contents

dixiyao/FoT/tree/main/experiment/pinchbench commit e34543795e

Frequently asked questions

npx skillmds@latest add dixiyao/pinchbench