Distributed Inference

Distributed inference for Llama, Qwen, DeepSeek across heterogeneous hardware. Self-hosted distributed inference — scatter requests across macOS, Linux, Windows, and any machine running Ollama. Thermal-aware distributed inference scheduling, 7-signal distributed inference scoring, adaptive capacity learning, context-aware model placement. No orchestration layer, no container runtime — just HTTP and mDNS. 分布式推理 across local devices | Inferencia distribuida en hardware local.

geeks-accelerator 5b7cf3d 9.3 KB Updated

File contents

geeks-accelerator/ollama-herd/tree/main/skills/distributed-inference commit 5b7cf3dba8

Frequently asked questions

npx skillmds@latest add geeks-accelerator/distributed-inference