Colm Experiments

Use when designing or auditing the empirical core of a COLM paper — contamination analysis for evaluation data, fair baselines under matched prompting and compute, pinned model versions and decoding parameters, uncertainty over runs and samples, scaling coverage, and honest reporting of API-model comparisons.

brycewang-stanford Updated 1k repo stars

File contents

brycewang-stanford/Awesome-Journal-Skills/tree/main/COLM-Skills/skills/colm-experiments commit c6f113c78f

Frequently asked questions

npx skillmds@latest add brycewang-stanford/colm-experiments