Litellm Docs

LiteLLM documentation — unified Python SDK and proxy server for 100+ LLM providers (OpenAI, Anthropic, Google, Azure, AWS Bedrock, Vertex AI, Cohere, Mistral, Ollama, vLLM, etc). Covers completion, embedding, image generation, audio, reranking, fine-tuning API; LiteLLM Proxy server (load balancing, rate limiting, virtual keys, spend tracking, SSO, RBAC, guardrails); caching (Redis, S3, in-memory); observability (Langfuse, Datadog, Prometheus, OpenTelemetry); secret managers (AWS, Azure, GCP, HashiCorp Vault); and provider-specific configuration. USE THIS SKILL WHEN the user asks about LiteLLM proxy setup, multi-provider LLM routing, model fallbacks, spend tracking, or calling any LLM API through LiteLLM.

wenerme 3e715d7 2.8 KB Updated

File contents

wenerme/ai/tree/main/skills/litellm-docs commit 3e715d7e89

Frequently asked questions

npx skillmds@latest add wenerme/litellm-docs