Entropy Minimization Reasoning Eval

Evaluates the reasoning capabilities of LLMs on mathematical and coding benchmarks by measuring accuracy under various inference-time scaling and unsupervised entropy minimization techniques. It probes whether reducing output uncertainty improves correctness without labeled data or parameter updates. Use when the user wants to benchmark on AMC, AIME, Minerva, LeetCode Live Contest, or asks about evaluating this task. Reports accuracy.

qhjqhj00 0880892 2.6 KB Updated 3 repo stars

File contents

qhjqhj00/research-skills-pool/tree/main/skill-factory/output/entropy-minimization-reasoning-eval commit 0880892ad8

Frequently asked questions

npx skillmds add qhjqhj00/entropy-minimization-reasoning-eval