import Image from '@theme/IdealImage';
import Tabs from '@theme/Tabs';
import TabItem from '@theme/TabItem';
Deploy this version
docker run \
-e STORE_MODEL_IN_DB=True \
-p 4000:4000 \
docker.litellm.ai/berriai/litellm:v1.80.11-stable
pip install litellm==1.80.11
Key Highlights
Cloudzero Integration on UI
<Image
img={require('../../img/ui_cloudzero.png')}
style={{width: '100%', display: 'block', margin: '2rem auto'}}
/>
Users can now configure their Cloudzero Integration directly on the UI.
Performance: 50% Reduction in Memory Usage and Import Latency for the LiteLLM SDK
We've completely restructured litellm.__init__.py to defer heavy imports until they're actually needed, implementing lazy loading for 109 components.
This refactoring includes 41 provider config classes, 40 utility functions, cache implementations (Redis, DualCache, InMemoryCache), HTTP handlers, logging, types, and other heavy dependencies. Heavy libraries like tiktoken and boto3 are now loaded on-demand rather than eagerly at import time.
This makes LiteLLM especially beneficial for serverless functions, Lambda deployments, and containerized environments where cold start times and memory footprint matter.
New Providers and Endpoints
New Providers (5 new providers)
| Provider |
Supported LiteLLM Endpoints |
Description |
| Stability AI |
/images/generations, /images/edits |
Stable Diffusion 3, SD3.5, image editing and generation |
| Venice.ai |
/chat/completions, /messages, /responses |
Venice.ai API integration via providers.json |
| Pydantic AI Agents |
/a2a |
Pydantic AI agents for A2A protocol workflows |
| VertexAI Agent Engine |
/a2a |
Google Vertex AI Agent Engine for agentic workflows |
| LinkUp Search |
/search |
LinkUp web search API integration |
New LLM API Endpoints (2 new endpoints)
| Endpoint |
Method |
Description |
Documentation |
/interactions |
POST |
Google Interactions API for conversational AI |
Docs |
/search |
POST |
RAG Search API with rerankers |
Docs |
New Models / Updated Models
New Model Support (55+ new models)
| Provider |
Model |
Context Window |
Input ($/1M tokens) |
Output ($/1M tokens) |
Features |
| Gemini |
gemini/gemini-3-flash-preview |
1M |
$0.50 |
$3.00 |
Reasoning, vision, audio, video, PDF |
| Vertex AI |
vertex_ai/gemini-3-flash-preview |
1M |
$0.50 |
$3.00 |
Reasoning, vision, audio, video, PDF |
| Azure AI |
azure_ai/deepseek-v3.2 |
164K |
$0.58 |
$1.68 |
Reasoning, function calling, caching |
| Azure AI |
azure_ai/cohere-rerank-v4.0-pro |
32K |
$0.0025/query |
- |
Rerank |
| Azure AI |
azure_ai/cohere-rerank-v4.0-fast |
32K |
$0.002/query |
- |
Rerank |
| OpenRouter |
openrouter/openai/gpt-5.2 |
400K |
$1.75 |
$14.00 |
Reasoning, vision, caching |
| OpenRouter |
openrouter/openai/gpt-5.2-pro |
400K |
$21.00 |
$168.00 |
Reasoning, vision |
| OpenRouter |
openrouter/mistralai/devstral-2512 |
262K |
$0.15 |
$0.60 |
Function calling |
| OpenRouter |
openrouter/mistralai/ministral-3b-2512 |
131K |
$0.10 |
$0.10 |
Function calling, vision |
| OpenRouter |
openrouter/mistralai/ministral-8b-2512 |
262K |
$0.15 |
$0.15 |
Function calling, vision |
| OpenRouter |
openrouter/mistralai/ministral-14b-2512 |
262K |
$0.20 |
$0.20 |
Function calling, vision |
| OpenRouter |
openrouter/mistralai/mistral-large-2512 |
262K |
$0.50 |
$1.50 |
Function calling, vision |
| OpenAI |
gpt-4o-transcribe-diarize |
16K |
$6.00/audio |
- |
Audio transcription with diarization |
| OpenAI |
gpt-image-1.5-2025-12-16 |
- |
Various |
Various |
Image generation |
| Stability |
stability/sd3-large |
- |
- |
$0.065/image |
Image generation |
| Stability |
stability/sd3.5-large |
- |
- |
$0.065/image |
Image generation |
| Stability |
stability/stable-image-ultra |
- |
- |
$0.08/image |
Image generation |
| Stability |
stability/inpaint |
- |
- |
$0.005/image |
Image editing |
| Stability |
stability/outpaint |
- |
- |
$0.004/image |
Image editing |
| Bedrock |
stability.stable-conservative-upscale-v1:0 |
- |
- |
$0.40/image |
Image upscaling |
| Bedrock |
stability.stable-creative-upscale-v1:0 |
- |
- |
$0.60/image |
Image upscaling |
| Vertex AI |
vertex_ai/deepseek-ai/deepseek-ocr-maas |
- |
$0.30 |
$1.20 |
OCR |
| LinkUp |
linkup/search |
- |
$5.87/1K queries |
- |
Web search |
| LinkUp |
linkup/search-deep |
- |
$58.67/1K queries |
- |
Deep web search |
| GitHub Copilot |
20+ models |
Various |
- |
- |
Chat completions |
Features
Bug Fixes
- Gemini
- Fix pricing for Gemini 3 Flash on Vertex AI - PR #18202
- Add output_cost_per_image_token for gemini-2.5-flash-image models - PR #18156
- Fix properties should be non-empty for OBJECT type - PR #18237
- Qwen
- Add qwen3-embedding-8b input per token price - PR #18018
- General
- Fix image URL handling - PR #18139
- Support Signed URLs with Query Parameters in Image Processing - PR #17976
- Add none to encoding_format instead of omitting it - PR #18042
LLM API Endpoints
Features
Bugs
- General
- Fix basemodel import in guardrail translation - PR #17977
- Fix No module named 'fastapi' error - PR #18239
Management Endpoints / UI
Features
- Virtual Keys
- Add master key rotation for credentials table - PR #17952
- Fix tag management to preserve encrypted fields in litellm_params - PR #17484
- Fix key delete and regenerate permissions - PR #18214
- Models + Endpoints
- Add Models Conditional Rendering in UI - PR #18071
- Add Health Check Model for Wildcard Model in UI - PR #18269
- Auto Resolve Vector Store Embedding Model Config - PR #18167
- Vector Stores
- Add Milvus Vector Store UI support - PR #18030
- Persist Vector Store Settings in Team Update - PR #18274
- Logs & Spend
- Add LiteLLM Overhead to Logs - PR #18033
- Show LiteLLM Overhead in Logs UI - PR #18034
- Resolve Team ID to Team Alias in Usage Page - PR #18275
- Fix Usage Page Top Key View Button Visibility - PR #18203
- SSO & Health
- Add SSO Readiness Health Check - PR #18078
- Fix /health/test_connection to resolve env variables like /chat/completions - PR #17752
- CloudZero
- General
- Update UI path handling for non-root Docker - PR #17989
Bugs
- UI Fixes
- Fix Login Page Failed To Parse JSON Error - PR #18159
- Fix new user route user_id collision handling - PR #17559
- Fix Callback Environment Variables Casing - PR #17912
AI Integrations
Logging
Guardrails
Secret Managers
Spend Tracking, Budgets and Rate Limiting
- Email Budget Alerts - Send email notifications when budgets are reached - PR #17995
MCP Gateway
- Auth Header Propagation - Add MCP auth header propagation - PR #17963
- Fix deepcopy error - Fix MCP tool call deepcopy error when processing requests - PR #18010
- Fix list tool - Fix MCP list_tools not working without database connection - PR #18161
Agent Gateway (A2A)
- New Provider: Agent Gateway - Add pydantic ai agents support - PR #18013
- VertexAI Agent Engine - Add Vertex AI Agent Engine provider - PR #18014
- Fix model extraction - Fix get_model_from_request() to extract model ID from Vertex AI passthrough URLs - PR #18097
Performance / Loadbalancing / Reliability improvements
- Lazy Imports - Use per-attribute lazy imports and extract shared constants - PR #17994
- Lazy Load HTTP Handlers - Lazy load http handlers - PR #17997
- Lazy Load Caches - Lazy load caches - PR #18001
- Lazy Load Types - Lazy load bedrock types, .types.utils, GuardrailItem - PR #18053, PR #18054, PR #18072
- Lazy Load Configs - Lazy load 41 configuration classes - PR #18267
- Lazy Load Client Decorators - Lazy load heavy client decorator imports - PR #18064
- Prisma Build Time - Download Prisma binaries at build time instead of runtime for security restricted environments - PR #17695
- Docker Alpine - Add libsndfile to Alpine image for ARM64 audio processing - PR #18092
- Security - Prevent LiteLLM API key leakage on /health endpoint failures - PR #18133
Documentation Updates
- SAP Docs - Update SAP documentation - PR #17974
- Pydantic AI Agents - Add docs on using pydantic ai agents with LiteLLM A2A gateway - PR #18026
- Vertex AI Agent Engine - Add Vertex AI Agent Engine documentation - PR #18027
- Router Order - Add router order parameter documentation - PR #18045
- Secret Manager Settings - Improve secret manager settings documentation - PR #18235
- Gemini 3 Flash - Add version requirement in Gemini 3 Flash blog - PR #18227
- README - Expand Responses API section and update endpoints - PR #17354
- Amazon Nova - Add Amazon Nova to sidebar and supported models - PR #18220
- Benchmarks - Add infrastructure recommendations to benchmarks documentation - PR #18264
- Broken Links - Fix broken link corrections - PR #18104
- README Fixes - Various README improvements - PR #18206
Infrastructure / CI/CD
- PR Templates - Add LiteLLM team PR template and CI/CD rules - PR #17983, PR #17985
- Issue Labeling - Improve issue labeling with component dropdown and more provider keywords - PR #17957
- PR Template Cleanup - Remove redundant fields from PR template - PR #17956
- Dependencies - Bump altcha-lib from 1.3.0 to 1.4.1 - PR #18017
New Contributors
- @dongbin-lunark made their first contribution in PR #17757
- @qdrddr made their first contribution in PR #18004
- @donicrosby made their first contribution in PR #17962
- @NicolaivdSmagt made their first contribution in PR #17992
- @Reapor-Yurnero made their first contribution in PR #18085
- @jk-f5 made their first contribution in PR #18086
- @castrapel made their first contribution in PR #18077
- @dtikhonov made their first contribution in PR #17484
- @opleonnn made their first contribution in PR #18175
- @eurogig made their first contribution in PR #18084
Full Changelog
View complete changelog on GitHub
1---2name: deploy-this-version-123description: import Image from '@theme/IdealImage'; import Tabs from '@theme/Tabs'; import TabItem from '@theme/TabItem';4---56import Image from '@theme/IdealImage';7import Tabs from '@theme/Tabs';8import TabItem from '@theme/TabItem';910## Deploy this version1112<Tabs>13<TabItem value="docker" label="Docker">1415``` showLineNumbers title="docker run litellm"16docker run \17-e STORE_MODEL_IN_DB=True \18-p 4000:4000 \19docker.litellm.ai/berriai/litellm:v1.80.11-stable20```2122</TabItem>2324<TabItem value="pip" label="Pip">2526``` showLineNumbers title="pip install litellm"27pip install litellm==1.80.1128```2930</TabItem>31</Tabs>3233---3435## Key Highlights3637- **Gemini 3 Flash Preview** - [Day 0 support for Google's Gemini 3 Flash Preview with reasoning capabilities](../../docs/providers/gemini)38- **Stability AI Image Generation** - [New provider for Stability AI image generation and editing](../../docs/providers/stability)39- **LiteLLM Content Filter** - [Built-in guardrails for harmful content, bias, and PII detection with image support](../../docs/proxy/guardrails/litellm_content_filter)40- **New Provider: Venice.ai** - Support for Venice.ai API via providers.json41- **Unified Skills API** - [Skills API works across Anthropic, Vertex, Azure, and Bedrock](../../docs/skills)42- **Azure Sentinel Logging** - [New logging integration for Azure Sentinel](../../docs/observability/azure_sentinel)43- **Guardrails Load Balancing** - [Load balance between multiple guardrail providers](../../docs/proxy/guardrails)44- **Email Budget Alerts** - [Send email notifications when budgets are reached](../../docs/proxy/email)45- **Cloudzero Integration on UI** - Setup your Cloudzero Integration Directly on the UI4647---4849### Cloudzero Integration on UI5051<Image52img={require('../../img/ui_cloudzero.png')}53style={{width: '100%', display: 'block', margin: '2rem auto'}}54/>5556Users can now configure their Cloudzero Integration directly on the UI.5758---59### Performance: 50% Reduction in Memory Usage and Import Latency for the LiteLLM SDK6061We've completely restructured `litellm.__init__.py` to defer heavy imports until they're actually needed, implementing lazy loading for **109 components**.6263This refactoring includes **41 provider config classes**, **40 utility functions**, cache implementations (Redis, DualCache, InMemoryCache), HTTP handlers, logging, types, and other heavy dependencies. Heavy libraries like tiktoken and boto3 are now loaded on-demand rather than eagerly at import time.6465This makes LiteLLM especially beneficial for serverless functions, Lambda deployments, and containerized environments where cold start times and memory footprint matter.6667---6869## New Providers and Endpoints7071### New Providers (5 new providers)7273| Provider | Supported LiteLLM Endpoints | Description |74| -------- | ------------------- | ----------- |75| [Stability AI](../../docs/providers/stability) | `/images/generations`, `/images/edits` | Stable Diffusion 3, SD3.5, image editing and generation |76| Venice.ai | `/chat/completions`, `/messages`, `/responses` | Venice.ai API integration via providers.json |77| [Pydantic AI Agents](../../docs/providers/pydantic_ai_agent) | `/a2a` | Pydantic AI agents for A2A protocol workflows |78| [VertexAI Agent Engine](../../docs/providers/vertex_ai_agent_engine) | `/a2a` | Google Vertex AI Agent Engine for agentic workflows |79| [LinkUp Search](../../docs/search/linkup) | `/search` | LinkUp web search API integration |8081### New LLM API Endpoints (2 new endpoints)8283| Endpoint | Method | Description | Documentation |84| -------- | ------ | ----------- | ------------- |85| `/interactions` | POST | Google Interactions API for conversational AI | [Docs](../../docs/interactions) |86| `/search` | POST | RAG Search API with rerankers | [Docs](../../docs/search/index) |8788---8990## New Models / Updated Models9192#### New Model Support (55+ new models)9394| Provider | Model | Context Window | Input ($/1M tokens) | Output ($/1M tokens) | Features |95| -------- | ----- | -------------- | ------------------- | -------------------- | -------- |96| Gemini | `gemini/gemini-3-flash-preview` | 1M | $0.50 | $3.00 | Reasoning, vision, audio, video, PDF |97| Vertex AI | `vertex_ai/gemini-3-flash-preview` | 1M | $0.50 | $3.00 | Reasoning, vision, audio, video, PDF |98| Azure AI | `azure_ai/deepseek-v3.2` | 164K | $0.58 | $1.68 | Reasoning, function calling, caching |99| Azure AI | `azure_ai/cohere-rerank-v4.0-pro` | 32K | $0.0025/query | - | Rerank |100| Azure AI | `azure_ai/cohere-rerank-v4.0-fast` | 32K | $0.002/query | - | Rerank |101| OpenRouter | `openrouter/openai/gpt-5.2` | 400K | $1.75 | $14.00 | Reasoning, vision, caching |102| OpenRouter | `openrouter/openai/gpt-5.2-pro` | 400K | $21.00 | $168.00 | Reasoning, vision |103| OpenRouter | `openrouter/mistralai/devstral-2512` | 262K | $0.15 | $0.60 | Function calling |104| OpenRouter | `openrouter/mistralai/ministral-3b-2512` | 131K | $0.10 | $0.10 | Function calling, vision |105| OpenRouter | `openrouter/mistralai/ministral-8b-2512` | 262K | $0.15 | $0.15 | Function calling, vision |106| OpenRouter | `openrouter/mistralai/ministral-14b-2512` | 262K | $0.20 | $0.20 | Function calling, vision |107| OpenRouter | `openrouter/mistralai/mistral-large-2512` | 262K | $0.50 | $1.50 | Function calling, vision |108| OpenAI | `gpt-4o-transcribe-diarize` | 16K | $6.00/audio | - | Audio transcription with diarization |109| OpenAI | `gpt-image-1.5-2025-12-16` | - | Various | Various | Image generation |110| Stability | `stability/sd3-large` | - | - | $0.065/image | Image generation |111| Stability | `stability/sd3.5-large` | - | - | $0.065/image | Image generation |112| Stability | `stability/stable-image-ultra` | - | - | $0.08/image | Image generation |113| Stability | `stability/inpaint` | - | - | $0.005/image | Image editing |114| Stability | `stability/outpaint` | - | - | $0.004/image | Image editing |115| Bedrock | `stability.stable-conservative-upscale-v1:0` | - | - | $0.40/image | Image upscaling |116| Bedrock | `stability.stable-creative-upscale-v1:0` | - | - | $0.60/image | Image upscaling |117| Vertex AI | `vertex_ai/deepseek-ai/deepseek-ocr-maas` | - | $0.30 | $1.20 | OCR |118| LinkUp | `linkup/search` | - | $5.87/1K queries | - | Web search |119| LinkUp | `linkup/search-deep` | - | $58.67/1K queries | - | Deep web search |120| GitHub Copilot | 20+ models | Various | - | - | Chat completions |121122#### Features123124- **[Gemini](../../docs/providers/gemini)**125 - Add Gemini 3 Flash Preview day 0 support with reasoning - [PR #18135](https://github.com/BerriAI/litellm/pull/18135)126 - Support extra_headers in batch embeddings - [PR #18004](https://github.com/BerriAI/litellm/pull/18004)127 - Propagate token usage when generating images - [PR #17987](https://github.com/BerriAI/litellm/pull/17987)128 - Use JSON instead of form-data for image edit requests - [PR #18012](https://github.com/BerriAI/litellm/pull/18012)129 - Fix web search requests count - [PR #17921](https://github.com/BerriAI/litellm/pull/17921)130- **[Anthropic](../../docs/providers/anthropic)**131 - Use dynamic max_tokens based on model - [PR #17900](https://github.com/BerriAI/litellm/pull/17900)132 - Fix claude-3-7-sonnet max_tokens to 64K default - [PR #17979](https://github.com/BerriAI/litellm/pull/17979)133 - Add OpenAI-compatible API with modify_params=True - [PR #17106](https://github.com/BerriAI/litellm/pull/17106)134- **[Vertex AI](../../docs/providers/vertex)**135 - Add Gemini 3 Flash Preview support - [PR #18164](https://github.com/BerriAI/litellm/pull/18164)136 - Add reasoning support for gemini-3-flash-preview - [PR #18175](https://github.com/BerriAI/litellm/pull/18175)137 - Fix image edit credential source - [PR #18121](https://github.com/BerriAI/litellm/pull/18121)138 - Pass credentials to PredictionServiceClient for custom endpoints - [PR #17757](https://github.com/BerriAI/litellm/pull/17757)139 - Fix multimodal embeddings for text + base64 image combinations - [PR #18172](https://github.com/BerriAI/litellm/pull/18172)140 - Add OCR support for DeepSeek model - [PR #17971](https://github.com/BerriAI/litellm/pull/17971)141- **[Azure AI](../../docs/providers/azure_ai)**142 - Add Azure Cohere 4 reranking models - [PR #17961](https://github.com/BerriAI/litellm/pull/17961)143 - Add Azure DeepSeek V3.2 versions - [PR #18019](https://github.com/BerriAI/litellm/pull/18019)144 - Return AzureAnthropicConfig for Claude models in get_provider_chat_config - [PR #18086](https://github.com/BerriAI/litellm/pull/18086)145- **[Fireworks AI](../../docs/providers/fireworks_ai)**146 - Add reasoning param support for Fireworks AI models - [PR #17967](https://github.com/BerriAI/litellm/pull/17967)147- **[Bedrock](../../docs/providers/bedrock)**148 - Add Qwen 2 and Qwen 3 to get_bedrock_model_id - [PR #18100](https://github.com/BerriAI/litellm/pull/18100)149 - Remove ttl field when routing to bedrock - [PR #18049](https://github.com/BerriAI/litellm/pull/18049)150 - Add Bedrock Stability image edit models - [PR #18254](https://github.com/BerriAI/litellm/pull/18254)151- **[Perplexity](../../docs/providers/perplexity)**152 - Use API-provided cost instead of manual calculation - [PR #17887](https://github.com/BerriAI/litellm/pull/17887)153- **[OpenAI](../../docs/providers/openai)**154 - Add diarize model for audio transcription - [PR #18117](https://github.com/BerriAI/litellm/pull/18117)155 - Add gpt-image-1.5-2025-12-16 in model cost map - [PR #18107](https://github.com/BerriAI/litellm/pull/18107)156 - Fix cost calculation of gpt-image-1 model - [PR #17966](https://github.com/BerriAI/litellm/pull/17966)157- **[GitHub Copilot](../../docs/providers/github_copilot)**158 - Add github_copilot model info - [PR #17858](https://github.com/BerriAI/litellm/pull/17858)159- **[Custom LLM](../../docs/providers/custom_llm_server)**160 - Add image_edit and aimage_edit support - [PR #17999](https://github.com/BerriAI/litellm/pull/17999)161162### Bug Fixes163164- **[Gemini](../../docs/providers/gemini)**165 - Fix pricing for Gemini 3 Flash on Vertex AI - [PR #18202](https://github.com/BerriAI/litellm/pull/18202)166 - Add output_cost_per_image_token for gemini-2.5-flash-image models - [PR #18156](https://github.com/BerriAI/litellm/pull/18156)167 - Fix properties should be non-empty for OBJECT type - [PR #18237](https://github.com/BerriAI/litellm/pull/18237)168- **[Qwen](../../docs/providers/fireworks_ai)**169 - Add qwen3-embedding-8b input per token price - [PR #18018](https://github.com/BerriAI/litellm/pull/18018)170- **General**171 - Fix image URL handling - [PR #18139](https://github.com/BerriAI/litellm/pull/18139)172 - Support Signed URLs with Query Parameters in Image Processing - [PR #17976](https://github.com/BerriAI/litellm/pull/17976)173 - Add none to encoding_format instead of omitting it - [PR #18042](https://github.com/BerriAI/litellm/pull/18042)174175---176177## LLM API Endpoints178179#### Features180181- **[Responses API](../../docs/response_api)**182 - Add provider specific tools support - [PR #17980](https://github.com/BerriAI/litellm/pull/17980)183 - Add custom headers support - [PR #18036](https://github.com/BerriAI/litellm/pull/18036)184 - Fix tool calls transformation in completion bridge - [PR #18226](https://github.com/BerriAI/litellm/pull/18226)185 - Use list format with input_text for tool results - [PR #18257](https://github.com/BerriAI/litellm/pull/18257)186 - Add cost tracking in background mode - [PR #18236](https://github.com/BerriAI/litellm/pull/18236)187 - Fix Claude code responses API bridge errors - [PR #18194](https://github.com/BerriAI/litellm/pull/18194)188- **[Chat Completions API](../../docs/completion/input)**189 - Add support for agent skills - [PR #18031](https://github.com/BerriAI/litellm/pull/18031)190- **[Skills API](../../docs/skills)**191 - Unified Skills API works across Anthropic, Vertex, Azure, Bedrock - [PR #18232](https://github.com/BerriAI/litellm/pull/18232)192- **[Search API](../../docs/search/index)**193 - Add new RAG Search API with rerankers - [PR #18217](https://github.com/BerriAI/litellm/pull/18217)194- **[Interactions API](../../docs/interactions)**195 - Add Google Interactions API on SDK and AI Gateway - [PR #18079](https://github.com/BerriAI/litellm/pull/18079), [PR #18081](https://github.com/BerriAI/litellm/pull/18081)196- **[Image Edit API](../../docs/image_edits)**197 - Add drop_params support and fix Vertex AI config - [PR #18077](https://github.com/BerriAI/litellm/pull/18077)198- **General**199 - Skip adding beta headers for Vertex AI as it is not supported - [PR #18037](https://github.com/BerriAI/litellm/pull/18037)200 - Fix managed files endpoint - [PR #18046](https://github.com/BerriAI/litellm/pull/18046)201 - Allow base_model for non-Azure providers in proxy - [PR #18038](https://github.com/BerriAI/litellm/pull/18038)202203#### Bugs204205- **General**206 - Fix basemodel import in guardrail translation - [PR #17977](https://github.com/BerriAI/litellm/pull/17977)207 - Fix No module named 'fastapi' error - [PR #18239](https://github.com/BerriAI/litellm/pull/18239)208209---210211## Management Endpoints / UI212213#### Features214215- **Virtual Keys**216 - Add master key rotation for credentials table - [PR #17952](https://github.com/BerriAI/litellm/pull/17952)217 - Fix tag management to preserve encrypted fields in litellm_params - [PR #17484](https://github.com/BerriAI/litellm/pull/17484)218 - Fix key delete and regenerate permissions - [PR #18214](https://github.com/BerriAI/litellm/pull/18214)219- **Models + Endpoints**220 - Add Models Conditional Rendering in UI - [PR #18071](https://github.com/BerriAI/litellm/pull/18071)221 - Add Health Check Model for Wildcard Model in UI - [PR #18269](https://github.com/BerriAI/litellm/pull/18269)222 - Auto Resolve Vector Store Embedding Model Config - [PR #18167](https://github.com/BerriAI/litellm/pull/18167)223- **Vector Stores**224 - Add Milvus Vector Store UI support - [PR #18030](https://github.com/BerriAI/litellm/pull/18030)225 - Persist Vector Store Settings in Team Update - [PR #18274](https://github.com/BerriAI/litellm/pull/18274)226- **Logs & Spend**227 - Add LiteLLM Overhead to Logs - [PR #18033](https://github.com/BerriAI/litellm/pull/18033)228 - Show LiteLLM Overhead in Logs UI - [PR #18034](https://github.com/BerriAI/litellm/pull/18034)229 - Resolve Team ID to Team Alias in Usage Page - [PR #18275](https://github.com/BerriAI/litellm/pull/18275)230 - Fix Usage Page Top Key View Button Visibility - [PR #18203](https://github.com/BerriAI/litellm/pull/18203)231- **SSO & Health**232 - Add SSO Readiness Health Check - [PR #18078](https://github.com/BerriAI/litellm/pull/18078)233 - Fix /health/test_connection to resolve env variables like /chat/completions - [PR #17752](https://github.com/BerriAI/litellm/pull/17752)234- **CloudZero**235 - Add CloudZero Cost Tracking UI - [PR #18163](https://github.com/BerriAI/litellm/pull/18163)236 - Add Delete CloudZero Settings Route and UI - [PR #18168](https://github.com/BerriAI/litellm/pull/18168), [PR #18170](https://github.com/BerriAI/litellm/pull/18170)237- **General**238 - Update UI path handling for non-root Docker - [PR #17989](https://github.com/BerriAI/litellm/pull/17989)239240#### Bugs241242- **UI Fixes**243 - Fix Login Page Failed To Parse JSON Error - [PR #18159](https://github.com/BerriAI/litellm/pull/18159)244 - Fix new user route user_id collision handling - [PR #17559](https://github.com/BerriAI/litellm/pull/17559)245 - Fix Callback Environment Variables Casing - [PR #17912](https://github.com/BerriAI/litellm/pull/17912)246247---248249## AI Integrations250251### Logging252253- **[Azure Sentinel](../../docs/observability/azure_sentinel)**254 - Add new Azure Sentinel Logger integration - [PR #18146](https://github.com/BerriAI/litellm/pull/18146)255- **[Prometheus](../../docs/proxy/logging#prometheus)**256 - Add extraction of top level metadata for custom labels - [PR #18087](https://github.com/BerriAI/litellm/pull/18087)257- **[Langfuse](../../docs/proxy/logging#langfuse)**258 - Fix not working log_failure_event - [PR #18234](https://github.com/BerriAI/litellm/pull/18234)259- **[Arize Phoenix](../../docs/observability/phoenix_integration)**260 - Fix nested spans - [PR #18102](https://github.com/BerriAI/litellm/pull/18102)261- **General**262 - Change extra_headers to additional_headers - [PR #17950](https://github.com/BerriAI/litellm/pull/17950)263264### Guardrails265266- **[LiteLLM Content Filter](../../docs/proxy/guardrails/litellm_content_filter)**267 - Add built-in guardrails for harmful content, bias, etc. - [PR #18029](https://github.com/BerriAI/litellm/pull/18029)268 - Add support for running content filters on images - [PR #18044](https://github.com/BerriAI/litellm/pull/18044)269 - Add support for Brazil PII field - [PR #18076](https://github.com/BerriAI/litellm/pull/18076)270 - Add configurable guardrail options for content filtering - [PR #18007](https://github.com/BerriAI/litellm/pull/18007)271- **[Guardrails API](../../docs/adding_provider/generic_guardrail_api)**272 - Support LLM tool call response checks on `/chat/completions`, `/v1/responses`, `/v1/messages` - [PR #17619](https://github.com/BerriAI/litellm/pull/17619)273 - Add guardrails load balancing - [PR #18181](https://github.com/BerriAI/litellm/pull/18181)274 - Fix guardrails for passthrough endpoint - [PR #18109](https://github.com/BerriAI/litellm/pull/18109)275 - Add headers to metadata for guardrails on pass-through endpoints - [PR #17992](https://github.com/BerriAI/litellm/pull/17992)276 - Various fixes for guardrail on OpenRouter models - [PR #18085](https://github.com/BerriAI/litellm/pull/18085)277- **[Lakera](../../docs/proxy/guardrails/lakera_ai)**278 - Add monitor mode for Lakera - [PR #18084](https://github.com/BerriAI/litellm/pull/18084)279- **[Pillar Security](../../docs/proxy/guardrails/pillar_security)**280 - Add masking support and MCP call support - [PR #17959](https://github.com/BerriAI/litellm/pull/17959)281- **[Bedrock Guardrails](../../docs/proxy/guardrails/bedrock)**282 - Add support for Bedrock image guardrails - [PR #18115](https://github.com/BerriAI/litellm/pull/18115)283 - Guardrails block action takes precedence over masking - [PR #17968](https://github.com/BerriAI/litellm/pull/17968)284285### Secret Managers286287- **[HashiCorp Vault](../../docs/secret_managers/hashicorp_vault)**288 - Add documentation for configurable Vault mount - [PR #18082](https://github.com/BerriAI/litellm/pull/18082)289 - Add per-team Vault configuration - [PR #18150](https://github.com/BerriAI/litellm/pull/18150)290- **UI**291 - Add secret manager settings controls to team management UI - [PR #18149](https://github.com/BerriAI/litellm/pull/18149)292293---294295## Spend Tracking, Budgets and Rate Limiting296297- **Email Budget Alerts** - Send email notifications when budgets are reached - [PR #17995](https://github.com/BerriAI/litellm/pull/17995)298299---300301## MCP Gateway302303- **Auth Header Propagation** - Add MCP auth header propagation - [PR #17963](https://github.com/BerriAI/litellm/pull/17963)304- **Fix deepcopy error** - Fix MCP tool call deepcopy error when processing requests - [PR #18010](https://github.com/BerriAI/litellm/pull/18010)305- **Fix list tool** - Fix MCP list_tools not working without database connection - [PR #18161](https://github.com/BerriAI/litellm/pull/18161)306307---308309## Agent Gateway (A2A)310311- **New Provider: Agent Gateway** - Add pydantic ai agents support - [PR #18013](https://github.com/BerriAI/litellm/pull/18013)312- **VertexAI Agent Engine** - Add Vertex AI Agent Engine provider - [PR #18014](https://github.com/BerriAI/litellm/pull/18014)313- **Fix model extraction** - Fix get_model_from_request() to extract model ID from Vertex AI passthrough URLs - [PR #18097](https://github.com/BerriAI/litellm/pull/18097)314315---316317## Performance / Loadbalancing / Reliability improvements318319- **Lazy Imports** - Use per-attribute lazy imports and extract shared constants - [PR #17994](https://github.com/BerriAI/litellm/pull/17994)320- **Lazy Load HTTP Handlers** - Lazy load http handlers - [PR #17997](https://github.com/BerriAI/litellm/pull/17997)321- **Lazy Load Caches** - Lazy load caches - [PR #18001](https://github.com/BerriAI/litellm/pull/18001)322- **Lazy Load Types** - Lazy load bedrock types, .types.utils, GuardrailItem - [PR #18053](https://github.com/BerriAI/litellm/pull/18053), [PR #18054](https://github.com/BerriAI/litellm/pull/18054), [PR #18072](https://github.com/BerriAI/litellm/pull/18072)323- **Lazy Load Configs** - Lazy load 41 configuration classes - [PR #18267](https://github.com/BerriAI/litellm/pull/18267)324- **Lazy Load Client Decorators** - Lazy load heavy client decorator imports - [PR #18064](https://github.com/BerriAI/litellm/pull/18064)325- **Prisma Build Time** - Download Prisma binaries at build time instead of runtime for security restricted environments - [PR #17695](https://github.com/BerriAI/litellm/pull/17695)326- **Docker Alpine** - Add libsndfile to Alpine image for ARM64 audio processing - [PR #18092](https://github.com/BerriAI/litellm/pull/18092)327- **Security** - Prevent LiteLLM API key leakage on /health endpoint failures - [PR #18133](https://github.com/BerriAI/litellm/pull/18133)328329---330331## Documentation Updates332333- **SAP Docs** - Update SAP documentation - [PR #17974](https://github.com/BerriAI/litellm/pull/17974)334- **Pydantic AI Agents** - Add docs on using pydantic ai agents with LiteLLM A2A gateway - [PR #18026](https://github.com/BerriAI/litellm/pull/18026)335- **Vertex AI Agent Engine** - Add Vertex AI Agent Engine documentation - [PR #18027](https://github.com/BerriAI/litellm/pull/18027)336- **Router Order** - Add router order parameter documentation - [PR #18045](https://github.com/BerriAI/litellm/pull/18045)337- **Secret Manager Settings** - Improve secret manager settings documentation - [PR #18235](https://github.com/BerriAI/litellm/pull/18235)338- **Gemini 3 Flash** - Add version requirement in Gemini 3 Flash blog - [PR #18227](https://github.com/BerriAI/litellm/pull/18227)339- **README** - Expand Responses API section and update endpoints - [PR #17354](https://github.com/BerriAI/litellm/pull/17354)340- **Amazon Nova** - Add Amazon Nova to sidebar and supported models - [PR #18220](https://github.com/BerriAI/litellm/pull/18220)341- **Benchmarks** - Add infrastructure recommendations to benchmarks documentation - [PR #18264](https://github.com/BerriAI/litellm/pull/18264)342- **Broken Links** - Fix broken link corrections - [PR #18104](https://github.com/BerriAI/litellm/pull/18104)343- **README Fixes** - Various README improvements - [PR #18206](https://github.com/BerriAI/litellm/pull/18206)344345---346347## Infrastructure / CI/CD348349- **PR Templates** - Add LiteLLM team PR template and CI/CD rules - [PR #17983](https://github.com/BerriAI/litellm/pull/17983), [PR #17985](https://github.com/BerriAI/litellm/pull/17985)350- **Issue Labeling** - Improve issue labeling with component dropdown and more provider keywords - [PR #17957](https://github.com/BerriAI/litellm/pull/17957)351- **PR Template Cleanup** - Remove redundant fields from PR template - [PR #17956](https://github.com/BerriAI/litellm/pull/17956)352- **Dependencies** - Bump altcha-lib from 1.3.0 to 1.4.1 - [PR #18017](https://github.com/BerriAI/litellm/pull/18017)353354---355356## New Contributors357358* @dongbin-lunark made their first contribution in [PR #17757](https://github.com/BerriAI/litellm/pull/17757)359* @qdrddr made their first contribution in [PR #18004](https://github.com/BerriAI/litellm/pull/18004)360* @donicrosby made their first contribution in [PR #17962](https://github.com/BerriAI/litellm/pull/17962)361* @NicolaivdSmagt made their first contribution in [PR #17992](https://github.com/BerriAI/litellm/pull/17992)362* @Reapor-Yurnero made their first contribution in [PR #18085](https://github.com/BerriAI/litellm/pull/18085)363* @jk-f5 made their first contribution in [PR #18086](https://github.com/BerriAI/litellm/pull/18086)364* @castrapel made their first contribution in [PR #18077](https://github.com/BerriAI/litellm/pull/18077)365* @dtikhonov made their first contribution in [PR #17484](https://github.com/BerriAI/litellm/pull/17484)366* @opleonnn made their first contribution in [PR #18175](https://github.com/BerriAI/litellm/pull/18175)367* @eurogig made their first contribution in [PR #18084](https://github.com/BerriAI/litellm/pull/18084)368369---370371## Full Changelog372373**[View complete changelog on GitHub](https://github.com/BerriAI/litellm/compare/v1.80.10-nightly...v1.80.11)**374