Local LLM Router

Local LLM model router for Llama, Qwen, DeepSeek, Phi, Mistral, and Gemma across multiple devices. Self-hosted local LLM inference routing on macOS, Linux, and Windows. Local LLM 7-signal scoring engine picks the optimal machine for every local LLM request. OpenAI-compatible local LLM API with context protection, VRAM-aware fallback, and auto-retry. 本地LLM路由 inference router | LLM local enrutador de inferencia. Use when the user wants to optimize local LLM routing, reduce local LLM latency, or load balance local LLM across machines.

geeks-accelerator 5da926a 9.0 KB Updated

File contents

geeks-accelerator/ollama-herd/tree/main/skills/local-llm-router commit 5da926aedd

Frequently asked questions

npx skillmds@latest add geeks-accelerator/local-llm-router