Ollama

Runs large language models locally with Ollama, including model management, custom Modelfiles, and API integration. Use for private, offline LLM inference.

ssrjkk Updated 2 repo stars

File contents

Ollama

Run LLMs locally with simple commands and a REST API.

Quick Start

ollama pull llama3.2
ollama run llama3.2 "What is the capital of France?"

When to Use

  • Private/local LLM inference
  • Offline AI applications
  • Testing models without API costs
  • Custom fine-tuned models

Step-by-Step

  1. Install Ollama from ollama.com
  2. Pull a model: ollama pull llama3.2
  3. Run interactively or via API
  4. Create custom Modelfiles

Dependencies

# Install from https://ollama.com
ollama pull llama3.2:3b

Examples

import requests
response = requests.post("http://localhost:11434/api/generate", json={
  "model": "llama3.2",
  "prompt": "Why is the sky blue?",
  "stream": False
})
print(response.json()["response"])

Resources

Validation

  1. Ollama service is running
  2. Model pulls and runs successfully
  3. API returns responses at localhost:11434

ssrjkk/claude-skills/tree/main/.claude/skills/ai/ollama commit d5535f5db4

Frequently asked questions

npx skillmds@latest add ssrjkk/ollama