Finetuning Method Selection

Decide whether to fine-tune at all, and route to the right method (SFT, DPO/ORPO/KTO, GRPO/RLVR, continued pretraining) and base model. Use when starting any fine-tuning effort, when unsure whether RAG or prompting would suffice, or when choosing between preference-optimization and reinforcement methods.

thedixitjain 6731c10 3 files · 15.8 KB Updated 2 repo stars

File contents

thedixitjain/the-mega-skill-library/tree/main/library/data-science-and-ml/finetuning-method-selection commit 6731c1032b

Frequently asked questions

npx skillmds add thedixitjain/finetuning-method-selection