Code editing benchmarks for OpenAI's "1106" models

and there's a lot of interest about their ability to code compared to the previous versions. With that in mind, I've been benchmarking the new models.

tools-only Updated 7 repo stars

File contents

tools-only/X-Skills/tree/main/development/2493-benchmarks-1106_75ed7cba commit b563e8a252

Frequently asked questions

npx skillmds@latest add tools-only/code-editing-benchmarks-for-openai-s-1106-models