Luis Gallego Ledesma
52372dcbe9
fix(langfuse_otel): prevent empty proxy request spans from being sent to Langfuse
...
When using langfuse_otel callback, empty traces were being sent to Langfuse
for requests that didn't result in actual LLM calls (e.g., auth operations,
health checks, failed requests). These traces contained only internal proxy
operations (auth, postgres, proxy_pre_call) with no useful LLM data.
Root cause: LangfuseOtelLogger extends OpenTelemetry, which sets itself as
the proxy's open_telemetry_logger. This caused create_litellm_proxy_request_started_span
to be called for every request, creating a parent span that was sent to Langfuse
even when no LLM call occurred.
Fix: Override create_litellm_proxy_request_started_span in LangfuseOtelLogger
to return None, preventing the creation of empty parent spans. This is consistent
with the existing overrides for async_service_success_hook and async_service_failure_hook
which already prevent service-level logs from being sent to Langfuse.
Fixes: Empty traces in Langfuse v3 when using langfuse_otel callback
2026-01-28 15:35:35 +01:00
Sameer Kankute
169c9dae79
Merge pull request #19914 from BerriAI/litellm_responses_api_bridge_usage
...
Fix: output_tokens_details.reasoning_tokens None
2026-01-28 18:35:30 +05:30
Sameer Kankute
9fe8b12f44
Merge pull request #19924 from BerriAI/litellm_minimax_reasoning_caching_1
...
Add Prompt caching and reasoning support for MiniMax, GLM, Xiaomi
2026-01-28 18:04:13 +05:30
Sameer Kankute
b6c769880e
Merge pull request #19842 from BerriAI/litellm_fix_timeout_test_fix
...
Fixes Timeouts during chat completion calls no longer reported as timeout in failure callback
2026-01-28 18:03:21 +05:30
Sameer Kankute
f5e5569e40
Merge pull request #19636 from BerriAI/litellm_langfuse_callback
...
Add litellm_callback_logging_failures_metric for Langfuse, Langfuse Otel and other Otel providers
2026-01-28 18:02:17 +05:30
Sameer Kankute
c5c1fbc5a2
Fix test_calculate_usage_completion_tokens_details_always_populated and logging object test
2026-01-28 18:00:42 +05:30
Sameer Kankute
0fadcbb21f
Merge pull request #19915 from BerriAI/litellm_x_ai_responses_web
...
Add xai websearch params support fo Responses API
2026-01-28 17:34:28 +05:30
Sameer Kankute
7386621d04
Merge pull request #19839 from BerriAI/litellm_oss_staging_01_27_2026
...
Litellm oss staging 01 27 2026
2026-01-28 17:33:27 +05:30
Sameer Kankute
f6ead49afe
Add Prompt caching and reasoning support for MiniMax, GLM, Xiaomi
2026-01-28 17:25:26 +05:30
Sameer Kankute
4f7425df0c
Merge pull request #19661 from Chesars/fix/oci-image-url-format
...
fix(oci): serialize imageUrl as object for OCI GenAI API
2026-01-28 15:26:50 +05:30
Sameer Kankute
f58305b747
Merge pull request #19882 from milan-berri/fix/tiktoken-offline-import-order
...
initialize tiktoken environment at import time to support offline usage
2026-01-28 15:13:47 +05:30
Sameer Kankute
156751e8fc
Merge pull request #19919 from lizhen921/fix/anthropic-cache-control-null-issue
...
fix(anthropic): remove explicit cache_control null in tool_result content
2026-01-28 13:24:56 +05:30
lizhen
e4cb28aa07
fix(anthropic): remove explicit cache_control null in tool_result content
...
Fixes issue where tool_result content blocks include explicit
'cache_control': null which breaks some Anthropic API channels.
Changes:
- Only include cache_control field when explicitly set and not None
- Prevents serialization of null values in tool_result text content
- Maintains backward compatibility with existing cache_control usage
Related issue: Anthropic tool_result conversion adds explicit null values
that cause compatibility issues with certain API implementations.
Co-Authored-By: Claude (claude-4.5-sonnet) <noreply@anthropic.com>
2026-01-28 15:39:01 +08:00
Sameer Kankute
9b44984510
Merge pull request #19899 from xianzongxie-stripe/add_native_background_mode_override
...
Add native_background_mode to override polling_via_cache for specific models
2026-01-28 12:27:19 +05:30
Sameer Kankute
bd15ebba84
fix: Pydantic will fail to parse it because cached_tokens is required but not provided
2026-01-28 11:51:26 +05:30
yuneng-jiang
ab655ef296
Merge pull request #19913 from BerriAI/litellm_store_prompt_backend
...
[Feature] Allow Dynamic Setting of store_prompts_in_spend_logs
2026-01-27 22:07:57 -08:00
Sameer Kankute
6fb2a0d11f
Fix: output_tokens_details.reasoning_tokens None
2026-01-28 11:35:18 +05:30
yuneng-jiang
28ca991296
Allow dynamic setting of store_prompts_in_spend_logs
2026-01-27 20:52:07 -08:00
Sameer Kankute
d76fb5932a
Add xai websearch params support
2026-01-28 09:54:43 +05:30
Ishaan Jaff
3080e04180
[Feat] UI: Allow Admins to control what pages are visible on LeftNav ( #19907 )
...
* feat: enabled_ui_pages_internal_users
* init ui for internal user controsl
* fix ui settings
* fix build
* fix leftnav
* fix leftnav
* test fixes
* fix leftnav
* isPageAccessibleToInternalUsers
* docs fix
* docs ui viz
2026-01-27 19:31:24 -08:00
Sameer Kankute
74b308bc2d
Merge pull request #19911 from BerriAI/litellm_merge_timeout_issue
...
merge main in timeout
2026-01-28 08:56:49 +05:30
Sameer Kankute
5276085f3c
Merge branch 'litellm_fix_timeout_test_fix' into litellm_merge_timeout_issue
2026-01-28 08:56:29 +05:30
mubashir1osmani
9a245031bd
feat(hosted_vllm): support thinking parameter in anthropic_messages() and .completion()
...
feat(hosted_vllm): support `thinking` parameter in `anthropic_messages()` and `.completion()`
2026-01-27 22:13:53 -05:00
yuneng-jiang
b76286e11c
Merge pull request #19908 from BerriAI/litellm_ui_model_table_server_sort
...
[Feature] UI - Model Page: Server Sort
2026-01-27 19:02:54 -08:00
Sameer Kankute
42a0d576f3
Merge pull request #19910 from BerriAI/main
...
merge 01 27
2026-01-28 08:30:47 +05:30
yuneng-jiang
29c841d7f9
Fixing build and tests
2026-01-27 18:14:30 -08:00
Ryan Wilson
70eb732b41
docs: fix guardrail logging docs ( #19833 )
2026-01-27 18:13:04 -08:00
Pragya Sardana
b4a27712a1
Add Init Containers in the community helm chart ( #19816 )
2026-01-27 18:10:47 -08:00
yuneng-jiang
37bfab929d
All Models Page server side sorting
2026-01-27 18:08:17 -08:00
yuneng-jiang
936a4fc884
Merge pull request #19905 from BerriAI/litellm_ci_ui_fix
...
[Infra] CI/CD - Fixing UI Tests
2026-01-27 17:21:39 -08:00
yuneng-jiang
01e22fd5ab
Fixing UI tests
2026-01-27 17:19:02 -08:00
yuneng-jiang
7109aafe4c
Merge pull request #19903 from BerriAI/litellm_ui_model_table_adjustable_col
...
[Feature] Add sortBy and sortOrder params for /v2/model/info
2026-01-27 17:16:22 -08:00
yuneng-jiang
a7ead71499
ruff check
2026-01-27 17:00:14 -08:00
yuneng-jiang
1581bcf985
add sortBy and sortOrder params for /v2/model/info
2026-01-27 16:54:52 -08:00
Xianzong Xie
f9eea06a37
Add tests for native_background_mode feature
...
Added 8 new unit tests for the native_background_mode feature:
- test_polling_disabled_when_model_in_native_background_mode
- test_polling_disabled_for_native_background_mode_with_provider_list
- test_polling_enabled_when_model_not_in_native_background_mode
- test_polling_enabled_when_native_background_mode_is_none
- test_polling_enabled_when_native_background_mode_is_empty_list
- test_native_background_mode_exact_match_required
- test_native_background_mode_with_provider_prefix_in_request
- test_native_background_mode_with_router_lookup
Committed-By-Agent: cursor
2026-01-27 16:48:22 -08:00
Ishaan Jaff
51339f5ef1
[Feat] RAG API - Add s3_vectors as provider on /vector_store/search API + UI for creating + PDF support for /rag/ingest ( #19895 )
...
* init S3VectorsRAGIngestion as a supported ingestion provider for RAG API
* test: TestRAGS3Vectors
* init S3VectorsVectorStoreOptions
* init s3 vectors
* code clean up + QA
* fix: get_credentials
* S3VectorsRAGIngestion
* TestRAGS3Vectors
* docs: AWS S3 Vectors
* add asyncio QA checks
* fix: S3_VECTORS_DEFAULT_DIMENSION
* init ui for bedrock s3 vectors
* fix add /search support for s3_vectors
* init atransform_search_vector_store_request
* feat: S3VectorsVectorStoreConfig
* TestS3VectorsVectorStoreConfig
* atransform_search_vector_store_request
* fix: S3VectorsVectorStoreConfig
* add validation for bucket name etd
* fix UI validation for s3 vector store
* init extract_text_from_pdf
* add pypdf
* fix code QA checks
* fix navbar
* init s3_vector.png
* fix QA code
2026-01-27 16:30:59 -08:00
Xianzong Xie
0b8c10c488
Add native_background_mode to override polling_via_cache for specific models
...
This follow-up to PR #16862 allows users to specify models that should use
the native provider's background mode instead of polling via cache.
Config example:
litellm_settings:
responses:
background_mode:
polling_via_cache: ["openai"]
native_background_mode: ["o4-mini-deep-research"]
ttl: 3600
When a model is in native_background_mode list, should_use_polling_for_request
returns False, allowing the request to fall through to native provider handling.
Committed-By-Agent: cursor
2026-01-27 15:54:46 -08:00
Ishaan Jaff
fe444f3ed5
[Feat] RAG API - Add support for using s3 Vectors as Vector Store Provider for /rag/ingest ( #19888 )
...
* init S3VectorsRAGIngestion as a supported ingestion provider for RAG API
* test: TestRAGS3Vectors
* init S3VectorsVectorStoreOptions
* init s3 vectors
* code clean up + QA
* fix: get_credentials
* S3VectorsRAGIngestion
* TestRAGS3Vectors
* docs: AWS S3 Vectors
* add asyncio QA checks
* fix: S3_VECTORS_DEFAULT_DIMENSION
2026-01-27 14:45:26 -08:00
michelligabriele
7d5439adda
fix(bedrock): support tool search header translation for Sonnet 4.5 ( #19871 )
...
Extend advanced-tool-use header translation to include Claude Sonnet 4.5
in addition to Opus 4.5 on Bedrock Invoke API.
When Claude Code sends the advanced-tool-use-2025-11-20 header, it now
gets correctly translated to Bedrock-specific headers for both:
- Claude Opus 4.5
- Claude Sonnet 4.5
Headers translated:
- tool-search-tool-2025-10-19
- tool-examples-2025-10-29
Fixes defer_loading validation error on Bedrock with Sonnet 4.5.
Ref: https://platform.claude.com/docs/en/agents-and-tools/tool-use/tool-search-tool
2026-01-27 12:17:09 -08:00
Milan
3f99a91a47
initialize tiktoken environment at import time to support offline usage
2026-01-27 22:05:04 +02:00
yuneng-jiang
45954155d7
Merge pull request #19799 from BerriAI/litellm_sso_email_casing
...
[Fix] SSO Email Case Sensitivity
2026-01-27 09:52:03 -08:00
yuneng-jiang
50612715a5
Merge pull request #19814 from BerriAI/litellm_team_member_add_fix
...
[Fix] /team/member_add User Email and ID Verifications
2026-01-27 09:49:01 -08:00
michelligabriele
388b4c90b6
fix(proxy): handle agent parameter in /interactions endpoint ( #19866 )
2026-01-27 09:34:58 -08:00
michelligabriele
fc7a9b4cb0
fix(enterprise): correct error message for DISABLE_ADMIN_ENDPOINTS ( #19861 )
...
The error message for DISABLE_ADMIN_ENDPOINTS incorrectly said
"DISABLING LLM API ENDPOINTS is an Enterprise feature" instead of
"DISABLING ADMIN ENDPOINTS is an Enterprise feature".
This was a copy-paste bug from the is_llm_api_route_disabled() function.
Added regression tests to verify both error messages are correct.
2026-01-27 09:34:30 -08:00
Harshit Jain
0f0b71e6d9
feat: add feature to make silent calls ( #19544 )
...
* feat: add feature to make silent calls
* add test or silent feat
* add docs for silent feat
* fix lint issues and UI logs
* add docs of ab testing and deep copy
2026-01-27 09:16:53 -08:00
Sameer Kankute
5c1588e3b7
Merge pull request #19841 from BerriAI/litellm_bedrock_tool_search_header
...
Translate advanced-tool-use to Bedrock-specific headers for Claude Opus 4.5
2026-01-27 17:48:51 +05:30
Sameer Kankute
81020d05d2
Merge pull request #19844 from BerriAI/litellm_sarvam_doc
...
[Docs]Add sarvam usage documentation
2026-01-27 17:47:59 +05:30
Sameer Kankute
c8a93c7d81
Merge pull request #19845 from BerriAI/litellm_gemini-robotics-er-1.5-preview2
...
Add Gemini Robotics-ER 1.5 preview support
2026-01-27 17:47:44 +05:30
Sameer Kankute
8565a9f5a2
Merge pull request #19847 from BerriAI/litellm_image_streaming_download
...
Fix: Stream the download in chunks for image handling
2026-01-27 17:47:35 +05:30
Sameer Kankute
29fc4f8f61
Merge pull request #19850 from BerriAI/litellm_grok_reasonnig_support
...
Add grok reasoning content
2026-01-27 17:46:07 +05:30