Task Structure

SWE-bench is the gold-standard benchmark for evaluating AI agents on real-world software engineering tasks.

tools-only Updated 7 repo stars

File contents

tools-only/X-Skills/tree/main/development/1180-swe-bench_e7347b99 commit 7e8a7d7e62

Frequently asked questions

npx skillmds@latest add tools-only/task-structure-15