Fine Tuning Guide

Model fine-tuning covering dataset preparation, LoRA and QLoRA, instruction tuning, RLHF and DPO, benchmarking, overfitting prevention, compute requirements, Hugging Face Trainer, and the fine-tuning vs prompt engineering decision. Use when the user asks about fine tuning guide, fine tuning guide best practices, or needs guidance on fine tuning guide implementation. Do NOT use when the user needs a different specialized skill or is asking about an unrelated technology domain.

FerroxLabs ccd53f4 2 files · 15.8 KB Updated 37 repo stars

File contents

ferroxlabs/murage/tree/main/skills-library/fine-tuning-guide commit ccd53f459d

Frequently asked questions

npx skillmds@latest add ferroxlabs/fine-tuning-guide