1---2name: deploy-this-version-333description: import Image from '@theme/IdealImage'; import Tabs from '@theme/Tabs'; import TabItem from '@theme/TabItem';4---56import Image from '@theme/IdealImage';7import Tabs from '@theme/Tabs';8import TabItem from '@theme/TabItem';910## Deploy this version1112<Tabs>13<TabItem value="docker" label="Docker">1415``` showLineNumbers title="docker run litellm"16docker run \17-e STORE_MODEL_IN_DB=True \18-p 4000:4000 \19docker.litellm.ai/berriai/litellm:v1.79.0-stable20```2122</TabItem>2324<TabItem value="pip" label="Pip">2526``` showLineNumbers title="pip install litellm"27pip install litellm==1.79.028```2930</TabItem>31</Tabs>3233---3435## Major Changes3637- **Cohere models will now be routed to Cohere v2 API by default** - [PR #15722](https://github.com/BerriAI/litellm/pull/15722)3839---4041## Key Highlights4243- **Search APIs** - Native `/v1/search` endpoint with support for Perplexity, Tavily, Parallel AI, Exa AI, DataforSEO, and Google PSE with cost tracking44- **Vector Stores** - Vertex AI Search API integration as vector store through LiteLLM with passthrough endpoint support45- **Guardrails Expansion** - Apply guardrails across Responses API, Image Gen, Text completions, Audio transcriptions, Audio Speech, Rerank, and Anthropic Messages API via unified `apply_guardrails` function46- **New Guardrail Providers** - Gray Swan, Dynamo AI, IBM Guardrails, Lasso Security v3, and Bedrock Guardrail apply_guardrail endpoint support47- **Video Generation API** - Native support for OpenAI Sora-2 and Azure Sora-2 (Pro, Pro-High-Res) with cost tracking and logging support48- **Azure AI Speech (TTS)** - Native Azure AI Speech integration with cost tracking for standard and HD voices4950---5152## New Models / Updated Models5354#### New Model Support5556| Provider | Model | Context Window | Input ($/1M tokens) | Output ($/1M tokens) | Features |57| -------- | ----- | -------------- | ------------------- | -------------------- | -------- |58| Bedrock | `anthropic.claude-3-7-sonnet-20240620-v1:0` | 200K | $3.60 | $18.00 | Chat, reasoning, vision, function calling, prompt caching, computer use |59| Bedrock GovCloud | `us-gov-west-1/anthropic.claude-3-7-sonnet-20250219-v1:0` | 200K | $3.60 | $18.00 | Chat, reasoning, vision, function calling, prompt caching, computer use |60| Vertex AI | `mistral-medium-3` | 128K | $0.40 | $2.00 | Chat, function calling, tool choice |61| Vertex AI | `codestral-2` | 128K | $0.30 | $0.90 | Chat, function calling, tool choice |62| Bedrock | `amazon.titan-image-generator-v1` | - | - | - | Image generation - $0.008/image, $0.01/premium image |63| Bedrock | `amazon.titan-image-generator-v2` | - | - | - | Image generation - $0.008/image, $0.01/premium image |64| OpenAI | `sora-2` | - | - | - | Video generation - $0.10/video/second |65| Azure | `sora-2` | - | - | - | Video generation - $0.10/video/second |66| Azure | `sora-2-pro` | - | - | - | Video generation - $0.30/video/second |67| Azure | `sora-2-pro-high-res` | - | - | - | Video generation - $0.50/video/second |6869#### Features7071- **[Anthropic](../../docs/providers/anthropic)**72 - Fix cache_control incorrectly applied to all content items instead of last item only - [PR #15699](https://github.com/BerriAI/litellm/pull/15699)73 - Forward anthropic-beta headers to Bedrock, VertexAI - [PR #15700](https://github.com/BerriAI/litellm/pull/15700)74 - Change max_tokens value to match max_output_tokens for claude sonnet - [PR #15715](https://github.com/BerriAI/litellm/pull/15715)7576- **[Bedrock](../../docs/providers/bedrock)**77 - Add AWS us-gov-west-1 Claude 3.7 Sonnet costs - [PR #15775](https://github.com/BerriAI/litellm/pull/15775)78 - Fix the date for sonnet 3.7 in govcloud - [PR #15800](https://github.com/BerriAI/litellm/pull/15800)79 - Use proper bedrock model name in health check - [PR #15808](https://github.com/BerriAI/litellm/pull/15808)80 - Support for embeddings_by_type Response Format in Bedrock Cohere Embed v1 - [PR #15707](https://github.com/BerriAI/litellm/pull/15707)81 - Add titan image generations with cost tracking - [PR #15916](https://github.com/BerriAI/litellm/pull/15916)8283- **[Gemini](../../docs/providers/gemini)**84 - Add imageConfig parameter for gemini-2.5-flash-image - [PR #15530](https://github.com/BerriAI/litellm/pull/15530)85 - Replace deprecated gemini-1.5-pro-preview-0514 - [PR #15852](https://github.com/BerriAI/litellm/pull/15852)86 - Update vertex ai gemini costs - [PR #15911](https://github.com/BerriAI/litellm/pull/15911)8788- **[Ollama](../../docs/providers/ollama)**89 - Set 'think' to False when reasoning effort is minimal/none/disable - [PR #15763](https://github.com/BerriAI/litellm/pull/15763)90 - Handle parsing ollama chunk error - [PR #15717](https://github.com/BerriAI/litellm/pull/15717)9192- **[Vertex AI](../../docs/providers/vertex)**93 - Add mistral medium 3 and Codestral 2 on vertex - [PR #15887](https://github.com/BerriAI/litellm/pull/15887)9495- **[Databricks](../../docs/providers/databricks)**96 - Allow prompt caching to be used for Anthropic Claude on Databricks - [PR #15801](https://github.com/BerriAI/litellm/pull/15801)9798- **[Azure](../../docs/providers/azure)**99 - Add Azure AVA TTS integration - [PR #15749](https://github.com/BerriAI/litellm/pull/15749)100 - Add Azure AVA (Speech AI) Cost Tracking - [PR #15754](https://github.com/BerriAI/litellm/pull/15754)101 - Azure AI Speech - Ensure `voice` is mapped from request body to SSML body, allow sending `role` and `style` - [PR #15810](https://github.com/BerriAI/litellm/pull/15810)102 - Add Azure support for video generation functionality (Sora-2) - [PR #15901](https://github.com/BerriAI/litellm/pull/15901)103104- **[OpenAI](../../docs/providers/openai)**105 - OpenAI videos refactoring - [PR #15900](https://github.com/BerriAI/litellm/pull/15900)106107- **General**108 - Read from custom-llm-provider header - [PR #15528](https://github.com/BerriAI/litellm/pull/15528)109110---111112## LLM API Endpoints113114#### Features115116- **[Responses API](../../docs/response_api)**117 - Add gpt 4.1 pricing for response endpoint - [PR #15593](https://github.com/BerriAI/litellm/pull/15593)118 - Fix Incorrect status value in responses api with gemini - [PR #15753](https://github.com/BerriAI/litellm/pull/15753)119 - Simplify reasoning item handling for gpt-5-codex - [PR #15815](https://github.com/BerriAI/litellm/pull/15815)120 - ErrorEvent ValidationError when OpenAI Responses API returns nested error structure - [PR #15804](https://github.com/BerriAI/litellm/pull/15804)121 - Fix reasoning item ID auto-generation causing encrypted content verification errors - [PR #15782](https://github.com/BerriAI/litellm/pull/15782)122 - Support tags in metadata - [PR #15867](https://github.com/BerriAI/litellm/pull/15867)123 - Security: prevent User A from retrieving User B's response, if response.id is leaked - [PR #15757](https://github.com/BerriAI/litellm/pull/15757)124125- **[Batch API](../../docs/batch_api)**126 - Add pre and post call for list batches - [PR #15673](https://github.com/BerriAI/litellm/pull/15673)127 - Add function responsible to call precall - [PR #15636](https://github.com/BerriAI/litellm/pull/15636)128 - Fix "User default_user_id does not have access to the object" when object not in db - [PR #15873](https://github.com/BerriAI/litellm/pull/15873)129130- **[OCR API](../../docs/ocr)**131 - Add Azure AI - OCR to docs - [PR #15768](https://github.com/BerriAI/litellm/pull/15768)132 - Add mode + Health check support for OCR models - [PR #15767](https://github.com/BerriAI/litellm/pull/15767)133134- **[Search API](../../docs/search_api)**135 - Add def search() APIs for Web Search - Perplexity API - [PR #15769](https://github.com/BerriAI/litellm/pull/15769)136 - Add Tavily Search API - [PR #15770](https://github.com/BerriAI/litellm/pull/15770)137 - Add Parallel AI - Search API - [PR #15772](https://github.com/BerriAI/litellm/pull/15772)138 - Add EXA AI Search API to LiteLLM - [PR #15774](https://github.com/BerriAI/litellm/pull/15774)139 - Add /search endpoint on LiteLLM Gateway - [PR #15780](https://github.com/BerriAI/litellm/pull/15780)140 - Add DataforSEO Search API - [PR #15817](https://github.com/BerriAI/litellm/pull/15817)141 - Add Google PSE Search Provider - [PR #15816](https://github.com/BerriAI/litellm/pull/15816)142 - Add cost tracking for Search API requests - Google PSE, Tavily, Parallel AI, Exa AI - [PR #15821](https://github.com/BerriAI/litellm/pull/15821)143 - Backend: Allow storing configured Search APIs in DB - [PR #15862](https://github.com/BerriAI/litellm/pull/15862)144 - Exa Search API - ensure request params are sent to Exa AI - [PR #15855](https://github.com/BerriAI/litellm/pull/15855)145146- **[Vector Stores](../../docs/vector_stores)**147 - Support Vertex AI Search API as vector store through LiteLLM - [PR #15781](https://github.com/BerriAI/litellm/pull/15781)148 - Azure AI - Search Vector Stores - [PR #15873](https://github.com/BerriAI/litellm/pull/15873)149 - VertexAI Search Vector Store - Passthrough endpoint support + Vector store search Cost tracking support - [PR #15824](https://github.com/BerriAI/litellm/pull/15824)150 - Don't raise error if managed object is not found - [PR #15873](https://github.com/BerriAI/litellm/pull/15873)151 - Show config.yaml vector stores on UI - [PR #15873](https://github.com/BerriAI/litellm/pull/15873)152 - Cost tracking for search spend - [PR #15859](https://github.com/BerriAI/litellm/pull/15859)153154- **[Images API](../../docs/image_generation)**155 - Pass user-defined headers and extra_headers to image-edit calls - [PR #15811](https://github.com/BerriAI/litellm/pull/15811)156157- **[Video Generation API](../../docs/video_generation)**158 - Add Azure support for video generation functionality (Sora-2, Sora-2-Pro, Sora-2-Pro-High-Res) - [PR #15901](https://github.com/BerriAI/litellm/pull/15901)159 - OpenAI video generation refactoring (Sora-2) - [PR #15900](https://github.com/BerriAI/litellm/pull/15900)160161- **[Bedrock /invoke](../../docs/bedrock_invoke)**162 - Fix: Hooks broken on /bedrock passthrough due to missing metadata - [PR #15849](https://github.com/BerriAI/litellm/pull/15849)163164- **[Realtime API](../../docs/realtime_api)**165 - Fix: OpenAI Realtime API integration fails due to websockets.exceptions.PayloadTooBig error - [PR #15751](https://github.com/BerriAI/litellm/pull/15751)166167---168169## Management Endpoints / UI170171#### Features172173- **Passthrough**174 - Set auth on passthrough endpoints, on the UI - [PR #15778](https://github.com/BerriAI/litellm/pull/15778)175 - Fix pass-through endpoint budget enforcement bug - [PR #15805](https://github.com/BerriAI/litellm/pull/15805)176177- **Organizations**178 - Allow org admins to create teams on UI - [PR #15924](https://github.com/BerriAI/litellm/pull/15924)179180- **Search Tools**181 - UI - Search Tools, allow adding search tools on UI + testing search - [PR #15871](https://github.com/BerriAI/litellm/pull/15871)182 - UI - Add logos for search providers - [PR #15872](https://github.com/BerriAI/litellm/pull/15872)183184- **General**185 - Fix routing for custom server root path - [PR #15701](https://github.com/BerriAI/litellm/pull/15701)186187---188189## Logging / Guardrail / Prompt Management Integrations190191#### Features192193- **[OpenTelemetry](../../docs/proxy/logging#opentelemetry)**194 - Fix OpenTelemetry Logging functionality - [PR #15645](https://github.com/BerriAI/litellm/pull/15645)195 - Fix issue where headers were not being split correctly - [PR #15916](https://github.com/BerriAI/litellm/pull/15916)196197- **[Sentry](../../docs/proxy/logging#sentry)**198 - Add SENTRY_ENVIRONMENT configuration for Sentry integration - [PR #15760](https://github.com/BerriAI/litellm/pull/15760)199200- **[Helicone](../../docs/proxy/logging#helicone)**201 - Fix JSON serialization error in Helicone logging by removing OpenTelemetry span from metadata - [PR #15728](https://github.com/BerriAI/litellm/pull/15728)202203- **[MLFlow](../../docs/proxy/logging#mlflow)**204 - Fix MLFlow tags - split request_tags into (key, val) if request_tag has colon - [PR #15914](https://github.com/BerriAI/litellm/pull/15914)205206- **General**207 - Rename configured_cold_storage_logger to cold_storage_custom_logger - [PR #15798](https://github.com/BerriAI/litellm/pull/15798)208209#### Guardrails210211- **[Gray Swan](../../docs/proxy/guardrails)**212 - Add GraySwan Guardrails support - [PR #15756](https://github.com/BerriAI/litellm/pull/15756)213 - Rename GraySwan to Gray Swan - [PR #15771](https://github.com/BerriAI/litellm/pull/15771)214215- **[Dynamo AI](../../docs/proxy/guardrails)**216 - New Guardrail - Dynamo AI Guardrail - [PR #15920](https://github.com/BerriAI/litellm/pull/15920)217218- **[IBM Guardrails](../../docs/proxy/guardrails)**219 - IBM Guardrails integration - [PR #15924](https://github.com/BerriAI/litellm/pull/15924)220221- **[Lasso Security](../../docs/proxy/guardrails)**222 - Add v3 API Support - [PR #12452](https://github.com/BerriAI/litellm/pull/12452)223 - Fixed lasso import config, redis cluster hash tags for test keys - [PR #15917](https://github.com/BerriAI/litellm/pull/15917)224225- **[Bedrock Guardrails](../../docs/proxy/guardrails)**226 - Implement Bedrock Guardrail apply_guardrail endpoint support - [PR #15892](https://github.com/BerriAI/litellm/pull/15892)227228- **General**229 - Guardrails - Responses API, Image Gen, Text completions, Audio transcriptions, Audio Speech, Rerank, Anthropic Messages API support via the unified `apply_guardrails` function - [PR #15706](https://github.com/BerriAI/litellm/pull/15706)230231---232233## Spend Tracking, Budgets and Rate Limiting234235- **Rate Limiting**236 - Support absolute RPM/TPM in priority_reservation - [PR #15813](https://github.com/BerriAI/litellm/pull/15813)237 - Org level tpm/rpm limits + Team tpm/rpm validation when assigned to org - [PR #15549](https://github.com/BerriAI/litellm/pull/15549)238239---240241## MCP Gateway242243- **OAuth**244 - Auth Header Fix for MCP Tool Call - [PR #15736](https://github.com/BerriAI/litellm/pull/15736)245 - Add response_type + PKCE parameters to OAuth authorization endpoint - [PR #15720](https://github.com/BerriAI/litellm/pull/15720)246247---248249## Performance / Loadbalancing / Reliability improvements250251- **Database**252 - Minimize the occurrence of deadlocks - [PR #15281](https://github.com/BerriAI/litellm/pull/15281)253254- **Redis**255 - Apply max_connections configuration to Redis async client - [PR #15797](https://github.com/BerriAI/litellm/pull/15797)256257- **Caching**258 - Add documentation for `enable_caching_on_provider_specific_optional_params` setting - [PR #15885](https://github.com/BerriAI/litellm/pull/15885)259260---261262## Documentation Updates263264- **Provider Documentation**265 - Update worker recommendation - [PR #15702](https://github.com/BerriAI/litellm/pull/15702)266 - Fix the wrong request body in json mode doc - [PR #15729](https://github.com/BerriAI/litellm/pull/15729)267 - Add details in docs - [PR #15721](https://github.com/BerriAI/litellm/pull/15721)268 - Add responses api on openai docs - [PR #15866](https://github.com/BerriAI/litellm/pull/15866)269 - Add OpenAI responses api - [PR #15868](https://github.com/BerriAI/litellm/pull/15868)270271---272273## New Contributors274275* @tlecomte made their first contribution in [PR #15528](https://github.com/BerriAI/litellm/pull/15528)276* @tomhaynes made their first contribution in [PR #15645](https://github.com/BerriAI/litellm/pull/15645)277* @talalryz made their first contribution in [PR #15720](https://github.com/BerriAI/litellm/pull/15720)278* @1vinodsingh1 made their first contribution in [PR #15736](https://github.com/BerriAI/litellm/pull/15736)279* @nuernber made their first contribution in [PR #15775](https://github.com/BerriAI/litellm/pull/15775)280* @Thomas-Mildner made their first contribution in [PR #15760](https://github.com/BerriAI/litellm/pull/15760)281* @javiergarciapleo made their first contribution in [PR #15721](https://github.com/BerriAI/litellm/pull/15721)282* @lshgdut made their first contribution in [PR #15717](https://github.com/BerriAI/litellm/pull/15717)283* @kk-wangjifeng made their first contribution in [PR #15530](https://github.com/BerriAI/litellm/pull/15530)284* @anthonyivn2 made their first contribution in [PR #15801](https://github.com/BerriAI/litellm/pull/15801)285* @romanglo made their first contribution in [PR #15707](https://github.com/BerriAI/litellm/pull/15707)286* @mythral made their first contribution in [PR #15859](https://github.com/BerriAI/litellm/pull/15859)287* @mubashirosmani made their first contribution in [PR #15866](https://github.com/BerriAI/litellm/pull/15866)288* @CAFxX made their first contribution in [PR #15281](https://github.com/BerriAI/litellm/pull/15281)289* @reflection made their first contribution in [PR #15914](https://github.com/BerriAI/litellm/pull/15914)290* @shadielfares made their first contribution in [PR #15917](https://github.com/BerriAI/litellm/pull/15917)291292---293294## PR Count Summary295296### 10/26/2025297* New Models / Updated Models: 20298* LLM API Endpoints: 29299* Management Endpoints / UI: 5300* Logging / Guardrail / Prompt Management Integrations: 10301* Spend Tracking, Budgets and Rate Limiting: 2302* MCP Gateway: 2303* Performance / Loadbalancing / Reliability improvements: 3304* Documentation Updates: 5305306---307308## Full Changelog309310**[View complete changelog on GitHub](https://github.com/BerriAI/litellm/compare/v1.78.5-stable...v1.79.0-stable)**311