Local LLM

Manage local Ollama LLM models for development and testing. Use when: running local models, configuring Ollama, switching between fast/quality models, optimizing VRAM usage, troubleshooting model performance, creating Modelfiles, or integrating local LLMs with applications via OpenAI-compatible API. Triggers: ollama, local model, local LLM, VRAM, GPU memory, model speed, inference, Modelfile, llama.cpp, quantization.

jimmc414 77dd700 9 files · 36.7 KB Updated

File contents

jimmc414/claude-code-plugin-marketplace/tree/main/plugins/local-llm/skills/local-llm commit 77dd7008ee

Frequently asked questions

npx skillmds@latest add jimmc414/local-llm