import Image from '@theme/IdealImage';
import Tabs from '@theme/Tabs';
import TabItem from '@theme/TabItem';
Deploy this version
docker run \
-e STORE_MODEL_IN_DB=True \
-p 4000:4000 \
docker.litellm.ai/berriai/litellm:v1.77.5-stable
pip install litellm==1.77.5
Key Highlights
- MCP OAuth 2.0 Support - Enhanced authentication for Model Context Protocol integrations
- Scheduled Key Rotations - Automated key rotation capabilities for enhanced security
- New Gemini 2.5 Flash & Flash-lite Models - Latest September 2025 preview models with improved pricing and features
- Performance Improvements - 54% RPS improvement
Performance Improvements - 54% RPS Improvement
<Image img={require('../../img/release_notes/perf_77_5.png')} style={{ width: '800px', height: 'auto' }} />
This release brings a 54% RPS improvement (1,040 → 1,602 RPS, aggregated) per instance.
The improvement comes from fixing O(n²) inefficiencies in the LiteLLM Router, primarily caused by repeated use of in statements inside loops over large arrays.
Tests were run with a database-only setup (no cache hits).
Test Setup
All benchmarks were executed using Locust with 1,000 concurrent users and a ramp-up of 500. The environment was configured to stress the routing layer and eliminate caching as a variable.
System Specs
- CPU: 8 vCPUs
- Memory: 32 GB RAM
Configuration (config.yaml)
View the complete configuration: gist.github.com/AlexsanderHamir/config.yaml
Load Script (no_cache_hits.py)
View the complete load testing script: gist.github.com/AlexsanderHamir/no_cache_hits.py
New Models / Updated Models
New Model Support
| Provider |
Model |
Context Window |
Input ($/1M tokens) |
Output ($/1M tokens) |
Features |
| Gemini |
gemini-2.5-flash-preview-09-2025 |
1M |
$0.30 |
$2.50 |
Chat, reasoning, vision, audio |
| Gemini |
gemini-2.5-flash-lite-preview-09-2025 |
1M |
$0.10 |
$0.40 |
Chat, reasoning, vision, audio |
| Gemini |
gemini-flash-latest |
1M |
$0.30 |
$2.50 |
Chat, reasoning, vision, audio |
| Gemini |
gemini-flash-lite-latest |
1M |
$0.10 |
$0.40 |
Chat, reasoning, vision, audio |
| DeepSeek |
deepseek-chat |
131K |
$0.60 |
$1.70 |
Chat, function calling, caching |
| DeepSeek |
deepseek-reasoner |
131K |
$0.60 |
$1.70 |
Chat, reasoning |
| Bedrock |
deepseek.v3-v1:0 |
164K |
$0.58 |
$1.68 |
Chat, reasoning, function calling |
| Azure |
azure/gpt-5-codex |
272K |
$1.25 |
$10.00 |
Responses API, reasoning, vision |
| OpenAI |
gpt-5-codex |
272K |
$1.25 |
$10.00 |
Responses API, reasoning, vision |
| SambaNova |
sambanova/DeepSeek-V3.1 |
33K |
$3.00 |
$4.50 |
Chat, reasoning, function calling |
| SambaNova |
sambanova/gpt-oss-120b |
131K |
$3.00 |
$4.50 |
Chat, reasoning, function calling |
| Bedrock |
qwen.qwen3-coder-480b-a35b-v1:0 |
262K |
$0.22 |
$1.80 |
Chat, reasoning, function calling |
| Bedrock |
qwen.qwen3-235b-a22b-2507-v1:0 |
262K |
$0.22 |
$0.88 |
Chat, reasoning, function calling |
| Bedrock |
qwen.qwen3-coder-30b-a3b-v1:0 |
262K |
$0.15 |
$0.60 |
Chat, reasoning, function calling |
| Bedrock |
qwen.qwen3-32b-v1:0 |
131K |
$0.15 |
$0.60 |
Chat, reasoning, function calling |
| Vertex AI |
vertex_ai/qwen/qwen3-next-80b-a3b-instruct-maas |
262K |
$0.15 |
$1.20 |
Chat, function calling |
| Vertex AI |
vertex_ai/qwen/qwen3-next-80b-a3b-thinking-maas |
262K |
$0.15 |
$1.20 |
Chat, function calling |
| Vertex AI |
vertex_ai/deepseek-ai/deepseek-v3.1-maas |
164K |
$1.35 |
$5.40 |
Chat, reasoning, function calling |
| OpenRouter |
openrouter/x-ai/grok-4-fast:free |
2M |
$0.00 |
$0.00 |
Chat, reasoning, function calling |
| XAI |
xai/grok-4-fast-reasoning |
2M |
$0.20 |
$0.50 |
Chat, reasoning, function calling |
| XAI |
xai/grok-4-fast-non-reasoning |
2M |
$0.20 |
$0.50 |
Chat, function calling |
Features
- Gemini
- Added Gemini 2.5 Flash and Flash-lite preview models (September 2025 release) with improved pricing - PR #14948
- Added new Anthropic web fetch tool support - PR #14951
- XAI
- Anthropic
- Updated Claude Sonnet 4 configs to reflect million-token context window pricing - PR #14639
- Added supported text field to anthropic citation response - PR #14164
- Bedrock
- Added support for Qwen models family & Deepseek 3.1 to Amazon Bedrock - PR #14845
- Support requestMetadata in Bedrock Converse API - PR #14570
- Vertex AI
- Added vertex_ai/qwen models and azure/gpt-5-codex - PR #14844
- Update vertex ai qwen model pricing - PR #14828
- Vertex AI Context Caching: use Vertex ai API v1 instead of v1beta1 and accept 'cachedContent' param - PR #14831
- SambaNova
- Add sambanova deepseek v3.1 and gpt-oss-120b - PR #14866
- OpenAI
- Fix inconsistent token configs for gpt-5 models - PR #14942
- GPT-3.5-Turbo price updated - PR #14858
- OpenRouter
- Add gpt-5 and gpt-5-codex to OpenRouter cost map - PR #14879
- VLLM
- Flux
Bug Fixes
- Anthropic
- Fix: Support claude code auth via subscription (anthropic) - PR #14821
- Fix Anthropic streaming IDs - PR #14965
- Revert incorrect changes to sonnet-4 max output tokens - PR #14933
- OpenAI
- Fix a bug where openai image edit silently ignores multiple images - PR #14893
- VLLM
- Fix: vLLM provider's rerank endpoint from /v1/rerank to /rerank - PR #14938
New Provider Support
LLM API Endpoints
Features
- General
- Add SDK support for additional headers - PR #14761
- Add shared_session parameter for aiohttp ClientSession reuse - PR #14721
Bugs
- General
- Fix: Streaming tool call index assignment for multiple tool calls - PR #14587
- Fix load credentials in token counter proxy - PR #14808
Management Endpoints / UI
Features
- Proxy CLI Auth
- Allow re-using cli auth token - PR #14780
- Create a python method to login using litellm proxy - PR #14782
- Fixes for LiteLLM Proxy CLI to Auth to Gateway - PR #14836
Virtual Keys
- Initial support for scheduled key rotations - PR #14877
- Allow scheduling key rotations when creating virtual keys - PR #14960
Models + Endpoints
- Fix: added Oracle to provider's list - PR #14835
Bugs
- SSO - Fix: SSO "Clear" button writes empty values instead of removing SSO config - PR #14826
- Admin Settings - Remove useful links from admin settings - PR #14918
- Management Routes - Add /user/list to management routes - PR #14868
Logging / Guardrail / Prompt Management Integrations
Features
Guardrails
- LakeraAI v2 Guardrail - Ensure exception is raised correctly - PR #14867
- Presidio Guardrail - Support custom entity types in Presidio guardrail with Union[PiiEntityType, str] - PR #14899
- Noma Guardrail - Add noma guardrail provider to ui - PR #14415
Prompt Management
- BitBucket Integration - Add BitBucket Integration for Prompt Management - PR #14882
Spend Tracking, Budgets and Rate Limiting
- Service Tier Pricing - Add service_tier based pricing support for openai (BOTH Service & Priority Support) - PR #14796
- Cost Tracking - Show input, output, tool call cost breakdown in StandardLoggingPayload - PR #14921
- Parallel Request Limiter v3
- Ensure Lua scripts can execute on redis cluster - PR #14968
- Fix: get metadata info from both metadata and litellm_metadata fields - PR #14783
- Priority Reservation - Fix: Priority Reservation: keys without priority metadata receive higher priority than keys with explicit priority configurations - PR #14832
MCP Gateway
- MCP Configuration - Enable custom fields in mcp_info configuration - PR #14794
- MCP Tools - Remove server_name prefix from list_tools - PR #14720
- OAuth Flow - Initial commit for v2 oauth flow - PR #14964
Performance / Loadbalancing / Reliability improvements
- Memory Leak Fix - Fix InMemoryCache unbounded growth when TTLs are set - PR #14869
- Cache Performance - Fix: cache root cause - PR #14827
- Concurrency Fix - Fix concurrency/scaling when many Python threads do streaming using sync completions - PR #14816
- Performance Optimization - Fix: reduce get_deployment cost to O(1) - PR #14967
- Performance Optimization - Fix: remove slow string operation - PR #14955
- DB Connection Management - Fix: DB connection state retries - PR #14925
Documentation Updates
- Provider Documentation - Fix docs for provider_specific_params.md - PR #14787
- Model References - Update model references from gemini-pro to gemini-2.5-pro - PR #14775
- Letta Guide - Add Letta Guide documentation - PR #14798
- README - Make the README document clearer - PR #14860
- Session Management - Update docs for session management availability - PR #14914
- Cost Documentation - Add documentation for additional cost-related keys in custom pricing - PR #14949
- Azure Passthrough - Add azure passthrough documentation - PR #14958
- General Documentation - Doc updates sept 2025 - PR #14769
- Clarified bridging between endpoints and mode in docs.
- Added Vertex AI Gemini API configuration as an alternative in relevant guides.
Linked AWS authentication info in the Bedrock guardrails documentation.
- Added Cancel Response API usage with code snippets
- Clarified that SSO (Single Sign-On) is free for up to 5 users:
- Alphabetized sidebar, leaving quick start / intros at top of categories
- Documented max_connections under cache_params.
- Clarified IAM AssumeRole Policy requirements.
- Added transform utilities example to Getting Started (showing request transformation).
- Added references to models.litellm.ai as the full models list in various docs.
- Added a code snippet for async_post_call_success_hook.
- Removed broken links to callbacks management guide. - Reformatted and linked cookbooks + other relevant docs
- Documentation Corrections - Corrected docs updates sept 2025 - PR #14916
New Contributors
- @uzaxirr made their first contribution in PR #14761
- @xprilion made their first contribution in PR #14416
- @CH-GAGANRAJ made their first contribution in PR #14779
- @otaviofbrito made their first contribution in PR #14778
- @danielmklein made their first contribution in PR #14639
- @Jetemple made their first contribution in PR #14826
- @akshoop made their first contribution in PR #14818
- @hazyone made their first contribution in PR #14821
- @leventov made their first contribution in PR #14816
- @fabriciojoc made their first contribution in PR #10955
- @onlylonly made their first contribution in PR #14845
- @Copilot made their first contribution in PR #14869
- @arsh72 made their first contribution in PR #14899
- @berri-teddy made their first contribution in PR #14914
- @vpbill made their first contribution in PR #14415
- @kgritesh made their first contribution in PR #14893
- @oytunkutrup1 made their first contribution in PR #14858
- @nherment made their first contribution in PR #14933
- @deepanshululla made their first contribution in PR #14974
- @TeddyAmkie made their first contribution in PR #14758
- @SmartManoj made their first contribution in PR #14775
- @uc4w6c made their first contribution in PR #14720
- @luizrennocosta made their first contribution in PR #14783
- @AlexsanderHamir made their first contribution in PR #14827
- @dharamendrak made their first contribution in PR #14721
- @TomeHirata made their first contribution in PR #14164
- @mrFranklin made their first contribution in PR #14860
- @luisfucros made their first contribution in PR #14866
- @huangyafei made their first contribution in PR #14879
- @thiswillbeyourgithub made their first contribution in PR #14949
- @Maximgitman made their first contribution in PR #14965
- @subnet-dev made their first contribution in PR #14938
- @22mSqRi made their first contribution in PR #14972
1---2name: deploy-this-version-243description: import Image from '@theme/IdealImage'; import Tabs from '@theme/Tabs'; import TabItem from '@theme/TabItem';4---56import Image from '@theme/IdealImage';7import Tabs from '@theme/Tabs';8import TabItem from '@theme/TabItem';910## Deploy this version1112<Tabs>13<TabItem value="docker" label="Docker">1415``` showLineNumbers title="docker run litellm"16docker run \17-e STORE_MODEL_IN_DB=True \18-p 4000:4000 \19docker.litellm.ai/berriai/litellm:v1.77.5-stable20```2122</TabItem>2324<TabItem value="pip" label="Pip">2526``` showLineNumbers title="pip install litellm"27pip install litellm==1.77.528```2930</TabItem>31</Tabs>3233---3435## Key Highlights3637- **MCP OAuth 2.0 Support** - Enhanced authentication for Model Context Protocol integrations38- **Scheduled Key Rotations** - Automated key rotation capabilities for enhanced security39- **New Gemini 2.5 Flash & Flash-lite Models** - Latest September 2025 preview models with improved pricing and features40- **Performance Improvements** - 54% RPS improvement4142---4344### Performance Improvements - 54% RPS Improvement4546<Image img={require('../../img/release_notes/perf_77_5.png')} style={{ width: '800px', height: 'auto' }} />4748<br/>4950This release brings a 54% RPS improvement (1,040 → 1,602 RPS, aggregated) per instance. 5152The improvement comes from fixing O(n²) inefficiencies in the LiteLLM Router, primarily caused by repeated use of `in` statements inside loops over large arrays. 5354Tests were run with a database-only setup (no cache hits).5556#### Test Setup5758All benchmarks were executed using Locust with 1,000 concurrent users and a ramp-up of 500. The environment was configured to stress the routing layer and eliminate caching as a variable.5960**System Specs**6162- **CPU:** 8 vCPUs63- **Memory:** 32 GB RAM6465**Configuration (config.yaml)**6667View the complete configuration: [gist.github.com/AlexsanderHamir/config.yaml](https://gist.github.com/AlexsanderHamir/53f7d554a5d2afcf2c4edb5b6be68ff4)6869**Load Script (no_cache_hits.py)**7071View the complete load testing script: [gist.github.com/AlexsanderHamir/no_cache_hits.py](https://gist.github.com/AlexsanderHamir/42c33d7a4dc7a57f56a78b560dee3a42)7273---747576## New Models / Updated Models7778#### New Model Support7980| Provider | Model | Context Window | Input ($/1M tokens) | Output ($/1M tokens) | Features |81| -------- | ----- | -------------- | ------------------- | -------------------- | -------- |82| Gemini | `gemini-2.5-flash-preview-09-2025` | 1M | $0.30 | $2.50 | Chat, reasoning, vision, audio |83| Gemini | `gemini-2.5-flash-lite-preview-09-2025` | 1M | $0.10 | $0.40 | Chat, reasoning, vision, audio |84| Gemini | `gemini-flash-latest` | 1M | $0.30 | $2.50 | Chat, reasoning, vision, audio |85| Gemini | `gemini-flash-lite-latest` | 1M | $0.10 | $0.40 | Chat, reasoning, vision, audio |86| DeepSeek | `deepseek-chat` | 131K | $0.60 | $1.70 | Chat, function calling, caching |87| DeepSeek | `deepseek-reasoner` | 131K | $0.60 | $1.70 | Chat, reasoning |88| Bedrock | `deepseek.v3-v1:0` | 164K | $0.58 | $1.68 | Chat, reasoning, function calling |89| Azure | `azure/gpt-5-codex` | 272K | $1.25 | $10.00 | Responses API, reasoning, vision |90| OpenAI | `gpt-5-codex` | 272K | $1.25 | $10.00 | Responses API, reasoning, vision |91| SambaNova | `sambanova/DeepSeek-V3.1` | 33K | $3.00 | $4.50 | Chat, reasoning, function calling |92| SambaNova | `sambanova/gpt-oss-120b` | 131K | $3.00 | $4.50 | Chat, reasoning, function calling |93| Bedrock | `qwen.qwen3-coder-480b-a35b-v1:0` | 262K | $0.22 | $1.80 | Chat, reasoning, function calling |94| Bedrock | `qwen.qwen3-235b-a22b-2507-v1:0` | 262K | $0.22 | $0.88 | Chat, reasoning, function calling |95| Bedrock | `qwen.qwen3-coder-30b-a3b-v1:0` | 262K | $0.15 | $0.60 | Chat, reasoning, function calling |96| Bedrock | `qwen.qwen3-32b-v1:0` | 131K | $0.15 | $0.60 | Chat, reasoning, function calling |97| Vertex AI | `vertex_ai/qwen/qwen3-next-80b-a3b-instruct-maas` | 262K | $0.15 | $1.20 | Chat, function calling |98| Vertex AI | `vertex_ai/qwen/qwen3-next-80b-a3b-thinking-maas` | 262K | $0.15 | $1.20 | Chat, function calling |99| Vertex AI | `vertex_ai/deepseek-ai/deepseek-v3.1-maas` | 164K | $1.35 | $5.40 | Chat, reasoning, function calling |100| OpenRouter | `openrouter/x-ai/grok-4-fast:free` | 2M | $0.00 | $0.00 | Chat, reasoning, function calling |101| XAI | `xai/grok-4-fast-reasoning` | 2M | $0.20 | $0.50 | Chat, reasoning, function calling |102| XAI | `xai/grok-4-fast-non-reasoning` | 2M | $0.20 | $0.50 | Chat, function calling |103104#### Features105106- **[Gemini](../../docs/providers/gemini)**107 - Added Gemini 2.5 Flash and Flash-lite preview models (September 2025 release) with improved pricing - [PR #14948](https://github.com/BerriAI/litellm/pull/14948)108 - Added new Anthropic web fetch tool support - [PR #14951](https://github.com/BerriAI/litellm/pull/14951)109- **[XAI](../../docs/providers/xai)**110 - Add xai/grok-4-fast models - [PR #14833](https://github.com/BerriAI/litellm/pull/14833)111- **[Anthropic](../../docs/providers/anthropic)**112 - Updated Claude Sonnet 4 configs to reflect million-token context window pricing - [PR #14639](https://github.com/BerriAI/litellm/pull/14639)113 - Added supported text field to anthropic citation response - [PR #14164](https://github.com/BerriAI/litellm/pull/14164)114- **[Bedrock](../../docs/providers/bedrock)**115 - Added support for Qwen models family & Deepseek 3.1 to Amazon Bedrock - [PR #14845](https://github.com/BerriAI/litellm/pull/14845)116 - Support requestMetadata in Bedrock Converse API - [PR #14570](https://github.com/BerriAI/litellm/pull/14570)117- **[Vertex AI](../../docs/providers/vertex)**118 - Added vertex_ai/qwen models and azure/gpt-5-codex - [PR #14844](https://github.com/BerriAI/litellm/pull/14844)119 - Update vertex ai qwen model pricing - [PR #14828](https://github.com/BerriAI/litellm/pull/14828)120 - Vertex AI Context Caching: use Vertex ai API v1 instead of v1beta1 and accept 'cachedContent' param - [PR #14831](https://github.com/BerriAI/litellm/pull/14831)121- **[SambaNova](../../docs/providers/sambanova)**122 - Add sambanova deepseek v3.1 and gpt-oss-120b - [PR #14866](https://github.com/BerriAI/litellm/pull/14866)123- **[OpenAI](../../docs/providers/openai)**124 - Fix inconsistent token configs for gpt-5 models - [PR #14942](https://github.com/BerriAI/litellm/pull/14942)125 - GPT-3.5-Turbo price updated - [PR #14858](https://github.com/BerriAI/litellm/pull/14858)126- **[OpenRouter](../../docs/providers/openrouter)**127 - Add gpt-5 and gpt-5-codex to OpenRouter cost map - [PR #14879](https://github.com/BerriAI/litellm/pull/14879)128- **[VLLM](../../docs/providers/vllm)**129 - Fix vllm passthrough - [PR #14778](https://github.com/BerriAI/litellm/pull/14778)130- **[Flux](../../docs/image_generation)**131 - Support flux image edit - [PR #14790](https://github.com/BerriAI/litellm/pull/14790)132133### Bug Fixes134135- **[Anthropic](../../docs/providers/anthropic)**136 - Fix: Support claude code auth via subscription (anthropic) - [PR #14821](https://github.com/BerriAI/litellm/pull/14821)137 - Fix Anthropic streaming IDs - [PR #14965](https://github.com/BerriAI/litellm/pull/14965)138 - Revert incorrect changes to sonnet-4 max output tokens - [PR #14933](https://github.com/BerriAI/litellm/pull/14933)139- **[OpenAI](../../docs/providers/openai)**140 - Fix a bug where openai image edit silently ignores multiple images - [PR #14893](https://github.com/BerriAI/litellm/pull/14893)141- **[VLLM](../../docs/providers/vllm)**142 - Fix: vLLM provider's rerank endpoint from /v1/rerank to /rerank - [PR #14938](https://github.com/BerriAI/litellm/pull/14938)143144#### New Provider Support145146- **[W&B Inference](../../docs/providers/wandb)**147 - Add W&B Inference to LiteLLM - [PR #14416](https://github.com/BerriAI/litellm/pull/14416)148149---150151## LLM API Endpoints152153#### Features154155- **General**156 - Add SDK support for additional headers - [PR #14761](https://github.com/BerriAI/litellm/pull/14761)157 - Add shared_session parameter for aiohttp ClientSession reuse - [PR #14721](https://github.com/BerriAI/litellm/pull/14721)158159#### Bugs160161- **General**162 - Fix: Streaming tool call index assignment for multiple tool calls - [PR #14587](https://github.com/BerriAI/litellm/pull/14587)163 - Fix load credentials in token counter proxy - [PR #14808](https://github.com/BerriAI/litellm/pull/14808)164165---166167## Management Endpoints / UI168169#### Features170171- **Proxy CLI Auth** 172 - Allow re-using cli auth token - [PR #14780](https://github.com/BerriAI/litellm/pull/14780)173 - Create a python method to login using litellm proxy - [PR #14782](https://github.com/BerriAI/litellm/pull/14782)174 - Fixes for LiteLLM Proxy CLI to Auth to Gateway - [PR #14836](https://github.com/BerriAI/litellm/pull/14836)175 176**Virtual Keys** 177 - Initial support for scheduled key rotations - [PR #14877](https://github.com/BerriAI/litellm/pull/14877)178 - Allow scheduling key rotations when creating virtual keys - [PR #14960](https://github.com/BerriAI/litellm/pull/14960)179180**Models + Endpoints**181 - Fix: added Oracle to provider's list - [PR #14835](https://github.com/BerriAI/litellm/pull/14835)182183184#### Bugs185186- **SSO** - Fix: SSO "Clear" button writes empty values instead of removing SSO config - [PR #14826](https://github.com/BerriAI/litellm/pull/14826)187- **Admin Settings** - Remove useful links from admin settings - [PR #14918](https://github.com/BerriAI/litellm/pull/14918)188- **Management Routes** - Add /user/list to management routes - [PR #14868](https://github.com/BerriAI/litellm/pull/14868)189---190191## Logging / Guardrail / Prompt Management Integrations192193#### Features194195- **[DataDog](../../docs/proxy/logging#datadog)**196 - Logging - `datadog` callback Log message content w/o sending to datadog - [PR #14909](https://github.com/BerriAI/litellm/pull/14909)197- **[Langfuse](../../docs/proxy/logging#langfuse)**198 - Adding langfuse usage details for cached tokens - [PR #10955](https://github.com/BerriAI/litellm/pull/10955)199- **[Opik](../../docs/proxy/logging#opik)**200 - Improve opik integration code - [PR #14888](https://github.com/BerriAI/litellm/pull/14888)201- **[SQS](../../docs/proxy/logging#sqs)**202 - Error logging support for SQS Logger - [PR #14974](https://github.com/BerriAI/litellm/pull/14974)203204#### Guardrails205206- **LakeraAI v2 Guardrail** - Ensure exception is raised correctly - [PR #14867](https://github.com/BerriAI/litellm/pull/14867)207- **Presidio Guardrail** - Support custom entity types in Presidio guardrail with Union[PiiEntityType, str] - [PR #14899](https://github.com/BerriAI/litellm/pull/14899)208- **Noma Guardrail** - Add noma guardrail provider to ui - [PR #14415](https://github.com/BerriAI/litellm/pull/14415)209210#### Prompt Management211212- **BitBucket Integration** - Add BitBucket Integration for Prompt Management - [PR #14882](https://github.com/BerriAI/litellm/pull/14882)213214---215216## Spend Tracking, Budgets and Rate Limiting217218- **Service Tier Pricing** - Add service_tier based pricing support for openai (BOTH Service & Priority Support) - [PR #14796](https://github.com/BerriAI/litellm/pull/14796)219- **Cost Tracking** - Show input, output, tool call cost breakdown in StandardLoggingPayload - [PR #14921](https://github.com/BerriAI/litellm/pull/14921)220- **Parallel Request Limiter v3** 221 - Ensure Lua scripts can execute on redis cluster - [PR #14968](https://github.com/BerriAI/litellm/pull/14968)222 - Fix: get metadata info from both metadata and litellm_metadata fields - [PR #14783](https://github.com/BerriAI/litellm/pull/14783)223- **Priority Reservation** - Fix: Priority Reservation: keys without priority metadata receive higher priority than keys with explicit priority configurations - [PR #14832](https://github.com/BerriAI/litellm/pull/14832)224225---226227## MCP Gateway228229- **MCP Configuration** - Enable custom fields in mcp_info configuration - [PR #14794](https://github.com/BerriAI/litellm/pull/14794)230- **MCP Tools** - Remove server_name prefix from list_tools - [PR #14720](https://github.com/BerriAI/litellm/pull/14720)231- **OAuth Flow** - Initial commit for v2 oauth flow - [PR #14964](https://github.com/BerriAI/litellm/pull/14964)232233---234235## Performance / Loadbalancing / Reliability improvements236237- **Memory Leak Fix** - Fix InMemoryCache unbounded growth when TTLs are set - [PR #14869](https://github.com/BerriAI/litellm/pull/14869)238- **Cache Performance** - Fix: cache root cause - [PR #14827](https://github.com/BerriAI/litellm/pull/14827)239- **Concurrency Fix** - Fix concurrency/scaling when many Python threads do streaming using *sync* completions - [PR #14816](https://github.com/BerriAI/litellm/pull/14816)240- **Performance Optimization** - Fix: reduce get_deployment cost to O(1) - [PR #14967](https://github.com/BerriAI/litellm/pull/14967)241- **Performance Optimization** - Fix: remove slow string operation - [PR #14955](https://github.com/BerriAI/litellm/pull/14955)242- **DB Connection Management** - Fix: DB connection state retries - [PR #14925](https://github.com/BerriAI/litellm/pull/14925)243244245246---247248## Documentation Updates249250- **Provider Documentation** - Fix docs for provider_specific_params.md - [PR #14787](https://github.com/BerriAI/litellm/pull/14787)251- **Model References** - Update model references from gemini-pro to gemini-2.5-pro - [PR #14775](https://github.com/BerriAI/litellm/pull/14775)252- **Letta Guide** - Add Letta Guide documentation - [PR #14798](https://github.com/BerriAI/litellm/pull/14798)253- **README** - Make the README document clearer - [PR #14860](https://github.com/BerriAI/litellm/pull/14860)254- **Session Management** - Update docs for session management availability - [PR #14914](https://github.com/BerriAI/litellm/pull/14914)255- **Cost Documentation** - Add documentation for additional cost-related keys in custom pricing - [PR #14949](https://github.com/BerriAI/litellm/pull/14949)256- **Azure Passthrough** - Add azure passthrough documentation - [PR #14958](https://github.com/BerriAI/litellm/pull/14958)257- **General Documentation** - Doc updates sept 2025 - [PR #14769](https://github.com/BerriAI/litellm/pull/14769)258 - Clarified bridging between endpoints and mode in docs.259 - Added Vertex AI Gemini API configuration as an alternative in relevant guides.260 Linked AWS authentication info in the Bedrock guardrails documentation.261 - Added Cancel Response API usage with code snippets262 - Clarified that SSO (Single Sign-On) is free for up to 5 users:263 - Alphabetized sidebar, leaving quick start / intros at top of categories264 - Documented max_connections under cache_params.265 - Clarified IAM AssumeRole Policy requirements.266 - Added transform utilities example to Getting Started (showing request transformation).267 - Added references to models.litellm.ai as the full models list in various docs.268 - Added a code snippet for async_post_call_success_hook.269 - Removed broken links to callbacks management guide. - Reformatted and linked cookbooks + other relevant docs270- **Documentation Corrections** - Corrected docs updates sept 2025 - [PR #14916](https://github.com/BerriAI/litellm/pull/14916)271272---273274## New Contributors275276* @uzaxirr made their first contribution in [PR #14761](https://github.com/BerriAI/litellm/pull/14761)277* @xprilion made their first contribution in [PR #14416](https://github.com/BerriAI/litellm/pull/14416)278* @CH-GAGANRAJ made their first contribution in [PR #14779](https://github.com/BerriAI/litellm/pull/14779)279* @otaviofbrito made their first contribution in [PR #14778](https://github.com/BerriAI/litellm/pull/14778)280* @danielmklein made their first contribution in [PR #14639](https://github.com/BerriAI/litellm/pull/14639)281* @Jetemple made their first contribution in [PR #14826](https://github.com/BerriAI/litellm/pull/14826)282* @akshoop made their first contribution in [PR #14818](https://github.com/BerriAI/litellm/pull/14818)283* @hazyone made their first contribution in [PR #14821](https://github.com/BerriAI/litellm/pull/14821)284* @leventov made their first contribution in [PR #14816](https://github.com/BerriAI/litellm/pull/14816)285* @fabriciojoc made their first contribution in [PR #10955](https://github.com/BerriAI/litellm/pull/10955)286* @onlylonly made their first contribution in [PR #14845](https://github.com/BerriAI/litellm/pull/14845)287* @Copilot made their first contribution in [PR #14869](https://github.com/BerriAI/litellm/pull/14869)288* @arsh72 made their first contribution in [PR #14899](https://github.com/BerriAI/litellm/pull/14899)289* @berri-teddy made their first contribution in [PR #14914](https://github.com/BerriAI/litellm/pull/14914)290* @vpbill made their first contribution in [PR #14415](https://github.com/BerriAI/litellm/pull/14415)291* @kgritesh made their first contribution in [PR #14893](https://github.com/BerriAI/litellm/pull/14893)292* @oytunkutrup1 made their first contribution in [PR #14858](https://github.com/BerriAI/litellm/pull/14858)293* @nherment made their first contribution in [PR #14933](https://github.com/BerriAI/litellm/pull/14933)294* @deepanshululla made their first contribution in [PR #14974](https://github.com/BerriAI/litellm/pull/14974)295* @TeddyAmkie made their first contribution in [PR #14758](https://github.com/BerriAI/litellm/pull/14758)296* @SmartManoj made their first contribution in [PR #14775](https://github.com/BerriAI/litellm/pull/14775)297* @uc4w6c made their first contribution in [PR #14720](https://github.com/BerriAI/litellm/pull/14720)298* @luizrennocosta made their first contribution in [PR #14783](https://github.com/BerriAI/litellm/pull/14783)299* @AlexsanderHamir made their first contribution in [PR #14827](https://github.com/BerriAI/litellm/pull/14827)300* @dharamendrak made their first contribution in [PR #14721](https://github.com/BerriAI/litellm/pull/14721)301* @TomeHirata made their first contribution in [PR #14164](https://github.com/BerriAI/litellm/pull/14164)302* @mrFranklin made their first contribution in [PR #14860](https://github.com/BerriAI/litellm/pull/14860)303* @luisfucros made their first contribution in [PR #14866](https://github.com/BerriAI/litellm/pull/14866)304* @huangyafei made their first contribution in [PR #14879](https://github.com/BerriAI/litellm/pull/14879)305* @thiswillbeyourgithub made their first contribution in [PR #14949](https://github.com/BerriAI/litellm/pull/14949)306* @Maximgitman made their first contribution in [PR #14965](https://github.com/BerriAI/litellm/pull/14965)307* @subnet-dev made their first contribution in [PR #14938](https://github.com/BerriAI/litellm/pull/14938)308* @22mSqRi made their first contribution in [PR #14972](https://github.com/BerriAI/litellm/pull/14972)309310---311312## **[Full Changelog](https://github.com/BerriAI/litellm/compare/v1.77.3.rc.1...v1.77.5.rc.1)**