Quantized Export

Export a promoted fine-tuned model in the right deployment format — merged safetensors, LoRA-only, GGUF with imatrix, or FP8. Use after a checkpoint passes promotion, when choosing a quantization format for a target device, or when an exported model fails its smoke test.

wshobson ecf8c0c 2 files · 20.4 KB Updated

File contents

wshobson/agents commit ecf8c0c082

Frequently asked questions

npx skillmds add wshobson/quantized-export