Finetuning Method Selection

Decide whether to fine-tune at all, and route to the right method (SFT, DPO/ORPO/KTO, GRPO/RLVR, continued pretraining) and base model. Use when starting any fine-tuning effort, when unsure whether RAG or prompting would suffice, or when choosing between preference-optimization and reinforcement methods.

gabrielmoreira Updated 17 repo stars

File contents

gabrielmoreira/agent-skills-mirror/tree/main/mirrors/repos/wshobson@agents/plugins/llm-finetuning/skills/finetuning-method-selection commit f8939e057a

Frequently asked questions

npx skillmds@latest add gabrielmoreira/finetuning-method-selection