Qlora

Memory-efficient fine-tuning with 4-bit quantization and LoRA adapters. Use when fine-tuning large models (7B+) on consumer GPUs, when VRAM is limited, or when standard LoRA still exceeds memory. Builds on the lora skill.

Mostafa Ahmed Updated 1.1k repo stars

File contents

itsmostafa/llm-engineering-skills/tree/main/skills/qlora commit 47c4c0f66b

Frequently asked questions

npx skillmds@latest add itsmostafa/qlora