# Local LLM Tool

> Local LLM execution tool for text generation and chat through Ollama or vLLM endpoints. Use when: running on-prem inference, calling a local GPU model, or summarizing with a self-hosted LLM.

- Skill: `xuiltul/local-llm-tool` (Agent Skill)
- Install (CLI): `npx skillmds@latest add xuiltul/local-llm-tool`
- Raw SKILL.md: https://api.skillmd.com/api/skills/xuiltul/local-llm-tool/raw
- Safety review: pending
- Works with: Claude Code, Claude.ai, OpenAI Codex
- Category: AI & ML
- Author: xuiltul (https://skillmd.com/u/xuiltul)
- Updated: 2026-09-22
- Page: https://skillmd.com/skills/xuiltul/local-llm-tool

---


# Local LLM Tool

External tool for text generation and chat via local LLM (Ollama/vLLM).

## Invocation via Bash

Use **Bash** with `animaworks-tool local_llm <subcommand> [args]`. See Actions below for syntax.

## Actions

### generate — Text generation
```json
{"tool_name": "local_llm", "action": "generate", "args": {"prompt": "prompt text", "system": "system prompt (optional)", "temperature": 0.7, "max_tokens": 2048}}
```

### chat — Multi-turn chat
```json
{"tool_name": "local_llm", "action": "chat", "args": {"messages": [{"role": "user", "content": "question"}], "system": "system prompt (optional)"}}
```

### models — List available models
```json
{"tool_name": "local_llm", "action": "models", "args": {}}
```

### status — Server status
```json
{"tool_name": "local_llm", "action": "status", "args": {}}
```

## CLI Usage (S/C/D/G-mode)

```bash
animaworks-tool local_llm generate "prompt" [-S "system prompt"]
animaworks-tool local_llm list
animaworks-tool local_llm status
```

## Notes

- Ollama or vLLM server must be running
- Use -s/--server to specify server URL
- Use -m/--model to specify model

