Qwen Mtp Gguf

Complete agent-ready workflow for Qwen-family MTP or nextn GGUF conversion and release. Use when Codex or another coding agent needs to inspect a user's machine, estimate disk/RAM requirements from Hugging Face model sizes and requested quant formats, bootstrap llama.cpp and Python dependencies, extract MTP heads from a matching official/base Qwen model, merge them into a fine-tuned or target safetensors model, run local HF/GGUF smoke tests with Qwen chat formatting, quantize to GGUF, and optionally upload or resume Hugging Face releases.

r6410418 Updated

File contents

r6410418/jackrong-llm-finetuning-guide/tree/main/qwen-mtp-gguf commit d9f393955c

Frequently asked questions

npx skillmds@latest add r6410418/qwen-mtp-gguf