Ollama

Use when running open-weight LLMs locally with Ollama — pulling and tagging models, calling the local API, picking a quantization or GGUF, writing Modelfiles, and sizing VRAM and RAM for the machine at hand. NOT remote or managed GPU serving and autoscaling (that is `runpod`), NOT downloading raw weights or datasets (that is `huggingface`), NOT retrieval pipeline design (that is `rag`).

ericrisco bfa8b0b 6 files · 28.7 KB Updated

File contents

ericrisco/rsc-harness/tree/main/skills/ollama commit bfa8b0b01d

Frequently asked questions

npx skillmds@latest add ericrisco/ollama