Unsloth

Use when fine-tuning an open-weight LLM fast on ONE GPU with low VRAM — Unsloth's fast model loaders with 4-bit QLoRA and the trl trainer, response-only loss masking so the prompt is not trained on, GRPO reasoning fine-tunes, and export to merged 16-bit, GGUF or the Hub. NOT whether, why or which method to fine-tune (that is `finetuning`), NOT running the exported GGUF locally (that is `ollama`), NOT serving-engine flags and throughput (that is `vllm`), NOT building the JSONL dataset (that is `training-data`).

ericrisco b29e783 6 files · 29.6 KB Updated

File contents

ericrisco/rsc-harness/tree/main/skills/unsloth commit b29e783f0e

Frequently asked questions

npx skillmds@latest add ericrisco/unsloth