Commit graph

31415 commits

Author SHA1 Message Date
Sameer Kankute
8565a9f5a2
Merge pull request #19847 from BerriAI/litellm_image_streaming_download
Fix: Stream the download in chunks for image handling
2026-01-27 17:47:35 +05:30
Sameer Kankute
29fc4f8f61
Merge pull request #19850 from BerriAI/litellm_grok_reasonnig_support
Add grok reasoning content
2026-01-27 17:46:07 +05:30
Sameer Kankute
c834d7d1fe
Merge branch 'main' into litellm_oss_staging_01_27_2026 2026-01-27 17:11:15 +05:30
Sameer Kankute
ea0a264a3c
Merge pull request #19753 from BerriAI/litellm_oss_staging_01_26_2026
fix(proxy): support slashes in google generateContent model names (#1…
2026-01-27 17:09:12 +05:30
Sameer Kankute
0214cb04cd
Merge branch 'main' into litellm_oss_staging_01_26_2026 2026-01-27 17:00:58 +05:30
Sameer Kankute
adf6d7e1db
Merge pull request #19692 from BerriAI/litellm_oss_staging_01_24_2026
Litellm oss staging 01 24 2026
2026-01-27 16:59:28 +05:30
Sameer Kankute
9a2750f8ec
Merge pull request #19617 from BerriAI/litellm_oss_staging_01_23_2026
Litellm oss staging 01 23 2026
2026-01-27 16:55:32 +05:30
Sameer Kankute
f30742fe6e Fix mypy and code quality issues 2026-01-27 16:49:39 +05:30
Sameer Kankute
154ad179af Revert poetry lock 2026-01-27 16:39:56 +05:30
Sameer Kankute
e695cb5367 Add grok reasoning content 2026-01-27 16:34:57 +05:30
Sameer Kankute
bd95712a22
Merge pull request #19832 from BerriAI/litellm_fix_a2a_package
Fix: A2A Python SDK URL
2026-01-27 15:16:52 +05:30
Sameer Kankute
988dd2a911 Fix: Stream the download in chunks 2026-01-27 14:35:54 +05:30
Sameer Kankute
e273c85848 Add gemini-robotics-er-1.5-preview model documentation 2026-01-27 13:58:13 +05:30
Sameer Kankute
cf012a2f65 Add gemini-robotics-er-1.5-preview model in model map 2026-01-27 13:58:03 +05:30
Sameer Kankute
faf9c9ba76 Add sarvam doc 2026-01-27 13:11:29 +05:30
Sameer Kankute
13313ac2be
Merge pull request #19232 from natimofeev/fix-gigachat-function-output-format
Fix: ensure function content is valid JSON for GigaChat
2026-01-27 13:02:38 +05:30
Sameer Kankute
f98eba24d4
Merge pull request #19040 from Point72/ephrimstanley/batch-list
Fix /batches to return encoded ids (from managed objects table)
2026-01-27 13:02:05 +05:30
Sameer Kankute
9883c2fd64 Fix: timeout exception raised eror 2026-01-27 12:32:37 +05:30
Harshit Jain
fd2f148161
fix: resolve 'does not exist' migration errors as applied in setup_database (#19281) 2026-01-26 22:11:36 -08:00
Harshit Jain
a5bc98a18a
fix(prometheus): safely handle None metadata in logging to prevent At… (#19691)
* fix(prometheus): safely handle None metadata in logging to prevent AttributeError

* fix: lint issues
2026-01-26 22:10:54 -08:00
Harshit Jain
885a02e6c8
fix: token calculations and refactor (#19696) 2026-01-26 22:08:17 -08:00
Sameer Kankute
3f32562587 Translate advanced-tool-use to Bedrock-specific headers for Claude Opus 4.5 2026-01-27 11:20:16 +05:30
Sameer Kankute
1bf33c4e41 Fix:Support both JSON array format and comma-separated values from user headers 2026-01-27 11:19:31 +05:30
Cesar Garcia
e4a557d95f
fix(xai): correct cached token cost calculation for xAI models (#19772)
* fix(azure): use generic cost calculator for audio token pricing

Azure audio models were charging audio output tokens at the text token
rate instead of the correct audio token rate. This resulted in costs
being ~6.65x lower than expected.

The fix replaces Azure's custom cost calculation logic with the generic
cost calculator that properly handles text, audio, cached, reasoning,
and image tokens.

Fixes #19764

* fix(xai): correct cached token cost calculation for xAI models

- Fix double-counting issue where xAI reports text_tokens = prompt_tokens
  (including cached), causing tokens to be charged twice
- Add cache_read_input_token_cost to xAI grok-3 and grok-3-mini model variants
- Detection: when text_tokens + cached_tokens > prompt_tokens, recalculate
  text_tokens = prompt_tokens - cached_tokens

xAI pricing (25% of input for cached):
- grok-3 variants: $0.75/M cached (input $3/M)
- grok-3-mini variants: $0.075/M cached (input $0.30/M)
2026-01-26 21:00:35 -08:00
Cesar Garcia
16f456ad82
fix(azure): use generic cost calculator for audio token pricing (#19771)
Azure audio models were charging audio output tokens at the text token
rate instead of the correct audio token rate. This resulted in costs
being ~6.65x lower than expected.

The fix replaces Azure's custom cost calculation logic with the generic
cost calculator that properly handles text, audio, cached, reasoning,
and image tokens.

Fixes #19764
2026-01-26 21:00:03 -08:00
Cesar Garcia
b1968a8e33
fix(responses): update local_vars with detected provider (#19782) (#19798)
When using the responses API with provider-specific params (aws_*, vertex_*)
without explicitly passing custom_llm_provider, the code crashed with:
AttributeError: 'NoneType' object has no attribute 'startswith'

Root cause: local_vars was captured via locals() before get_llm_provider()
detected the provider from the model string (e.g., "bedrock/..."), so
custom_llm_provider remained None when processing provider-specific params.

Fix: Update local_vars["custom_llm_provider"] after get_llm_provider() call
so the detected provider is available for param processing.

Affected provider-specific params:
- aws_* (aws_region_name, aws_access_key_id, etc.) for Bedrock/SageMaker
- vertex_* (vertex_project, vertex_location, etc.) for Vertex AI
2026-01-26 20:47:35 -08:00
Cesar Garcia
0d45b01069
fix(models): set gpt-5.2-codex mode to responses for Azure and OpenRouter (#19770)
Fixes #19754

The gpt-5.2-codex model only supports the responses API, not chat completions.
Updated azure/gpt-5.2-codex and openrouter/openai/gpt-5.2-codex entries to use
mode: "responses" and supported_endpoints: ["/v1/responses"].
2026-01-26 20:36:10 -08:00
Krish Dholakia
6a54dcfa93
feat: Add model_id label to Prometheus metrics (#18048) (#19678)
Co-authored-by: Cursor Agent <cursoragent@cursor.com>
2026-01-26 20:32:08 -08:00
Krish Dholakia
7ba0235a50
Litellm release notes 01 26 2026 (#19838)
* docs: document new models/endpoints

* docs: cleanup

* feat: update model table

* fix: cleanup
2026-01-26 20:20:20 -08:00
yuneng-jiang
7307992cea fixing tests 2026-01-26 20:17:53 -08:00
Krish Dholakia
664715b68e
Litellm release notes 01 26 2026 (#19836)
* docs: document new models/endpoints

* docs: cleanup

* feat: update model table
2026-01-26 20:11:56 -08:00
yuneng-jiang
0dea6eeea3
Merge pull request #19721 from BerriAI/litellm_ui_bulk_fix
[Fix] UI - Internal User: Bulk Add
2026-01-26 20:10:37 -08:00
yuneng-jiang
c875cfb8cb
Merge pull request #19804 from BerriAI/litellm_ui_dark_mode_slider
[Feature] UI - Add Light/Dark Mode Switch for Development
2026-01-26 20:10:12 -08:00
yuneng-jiang
a97bf452f1
Merge pull request #19831 from BerriAI/litellm_ui_hide_send_feedback
[Feature] UI - Feedback Prompts: Option To Hide Prompts
2026-01-26 20:10:04 -08:00
yuneng-jiang
a7712ff8ef
Merge pull request #19807 from BerriAI/litellm_ui_regen_expiry
[Fix] UI - Create Key: Expire Key Input Duration
2026-01-26 20:09:39 -08:00
yuneng-jiang
ef7261d0eb Fixing tests 2026-01-26 20:07:30 -08:00
Cesar Garcia
014f783cc9
docs(readme): add OpenAI Agents SDK to OSS Adopters (#19820)
* docs(readme): add OpenAI Agents SDK to OSS Adopters

* docs(readme): add OpenAI Agents SDK logo
2026-01-26 19:24:46 -08:00
Ishaan Jaff
52d73c2c16
[Feat] Add UI for /rag/ingest API - upload docs, pdfs etc to create vector stores (#19822)
* feat: _save_vector_store_to_db_from_rag_ingest

* UI features for RAG ingest

* fix: Endpoints

* ragIngestCall

* _save_vector_store_to_db_from_rag_ingest

* fix: rag_ingest Code QA CHECK

* UI fixes unit tests
2026-01-26 19:23:43 -08:00
Sameer Kankute
c70d04d85f Fix: A2A Python SDK URL 2026-01-27 08:07:36 +05:30
yuneng-jiang
b10f71d583 fixing breaking change: just user_id provided should upsert still 2026-01-26 18:10:51 -08:00
yuneng-jiang
aa4b3ba87a Adding tests: 2026-01-26 17:53:01 -08:00
Alexsander Hamir
f95572e3ed
Fix broken mocks in 6 flaky tests to prevent real API calls (#19829)
* Fix broken mocks in 6 flaky tests to prevent real API calls

Added network-level HTTP blocking using respx to prevent tests from making real API calls when Python-level mocks fail. This makes tests more reliable and retryable in CI.

Changes:

- Azure OIDC test: Added Azure Identity SDK mock to prevent real Azure calls

- Vector store test: Added @respx.mock decorator to block HTTP requests

- Resend email tests (3): Added @respx.mock decorator for all 3 test functions

- SendGrid email test: Added @respx.mock decorator

All test assertions and verification logic remain unchanged - only added safety nets to catch leaked API calls.

* Fix failing OIDC secret manager tests

Fixed two test failures in test_secret_managers_main.py:

1. test_oidc_azure_ad_token_success: Corrected the patch path for get_bearer_token_provider from 'litellm.secret_managers.get_azure_ad_token_provider.get_bearer_token_provider' to 'azure.identity.get_bearer_token_provider' since the function is imported from azure.identity.

2. test_oidc_google_success: Added @patch('httpx.Client') decorator to prevent any real HTTP connections during test execution, resolving httpx.ConnectError issues.

Both tests now pass successfully.
2026-01-26 17:39:40 -08:00
Alexsander Hamir
c442fcd922
CI/CD: Increase retries and stabilize litellm_mapped_tests_core (#19826)
* Fix PLR0915: Extract system message handling to reduce statement count

* fix mypy

* fix: add host_progress_callback parameter to mock_call_tool in test

The test_call_tool_without_broken_pipe_error was failing because the mock function did not accept the host_progress_callback keyword argument that the actual implementation passes to client.call_tool(). Updated the mock to accept this parameter to match the real implementation signature.

* fixing flaky tests around oidc and email

* Add documentation comment to test file

* add retry

* add dependency

* increase retry

---------

Co-authored-by: yuneng-jiang <yuneng.jiang@gmail.com>
2026-01-26 17:00:18 -08:00
yuneng-jiang
9ff0daa24d Add dont ask me again option in nudges 2026-01-26 16:58:38 -08:00
mubashir1osmani
8908eff7b1
Fix(#19781): Unable to reset user max budget to unlimited
Fix(#19781): Unable to reset user max budget to unlimited
2026-01-26 18:37:49 -05:00
yuneng-jiang
801e0a6ce6
Merge pull request #19819 from BerriAI/litellm_cd_fix_yj_10
[Infra] CI/CD - Fixing Flaky Tests in OIDC and Email
2026-01-26 15:29:55 -08:00
mubashir1osmani
5d973477e6
fix(ui): prevent clearing content filter patterns when editing guardrail
fix(ui): prevent clearing content filter patterns when editing guardrail
2026-01-26 18:29:48 -05:00
yuneng-jiang
33b4b444ac fixing flaky tests around oidc and email 2026-01-26 15:23:17 -08:00
Ishaan Jaff
cec1a3c858
[Feat] CLI Auth - Add configurable CLI JWT expiration via environment variable (#19780)
* fix: add CLI_JWT_EXPIRATION_HOURS

* docs: CLI_JWT_EXPIRATION_HOURS

* fix: get_cli_jwt_auth_token

* test_get_cli_jwt_auth_token_custom_expiration
2026-01-26 14:56:17 -08:00
houdataali
29ee5aab5c
[Feat] enable progress notifications for MCP tool calls (#19809)
* enable progress notifications for MCP tool calls

* adjust mcp test
2026-01-26 14:48:22 -08:00