Huggingface Local Models

Use to select models to run locally with llama.cpp and GGUF on CPU, Mac Metal, CUDA, or ROCm. Covers finding GGUFs, quant selection, running servers, exact GGUF file lookup, conversion, and OpenAI-compatible local serving.

practicalswan 72491ab 6 files · 32.2 KB Updated

File contents

practicalswan/agent-skills/tree/main/huggingface-local-models commit 72491abb01

Frequently asked questions

npx skillmds add practicalswan/huggingface-local-models