Research

Conduct post-training research for LLMs using the Tinker API — replicate paper results, explore new training ideas, run and monitor experiments, and document findings. Use this skill whenever the user wants to do research, replicate experiments from a paper or repo, investigate training hypotheses, run experiment sweeps, explore post-training techniques (SFT, RL, DPO, distillation, etc.), set up training, write training code, choose a model, tune hyperparameters, manage checkpoints, export weights, or analyze training logs — even if they just say "try this idea" or "let's see what happens if...".

thinking-machines-lab Updated

File contents

thinking-machines-lab/tinker-cookbook/tree/main/skills/research commit 5de3b2e0e2

Frequently asked questions

npx skillmds@latest add thinking-machines-lab/research