Opencodeinstruct Eval

Evaluates the code generation, algorithmic problem-solving, and complex function-calling capabilities of instruction-tuned LLMs across multiple standardized coding benchmarks. Use when the user wants to benchmark on HumanEval, MBPP, LiveCodeBench, BigIntCodeBench-Instruct, or asks about evaluating this task. Reports pass@1.

qhjqhj00 41d6eac 2.8 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/opencodeinstruct-eval commit 41d6eac3ed

Frequently asked questions

npx skillmds add qhjqhj00/opencodeinstruct-eval