Ishaan Jaffer
57a2ec3beb
fix: _extract_fields_recursive
2025-10-22 09:37:29 -07:00
Roman G
eac3cba44f
Support for embeddings_by_type Response Format in Bedrock Cohere Embed v1 ( #15707 )
...
* feat(cohere): Enhance embedding transformation to support Bedrock's embeddings by type
* test(cohere): Add unit tests for embedding transformation responses
2025-10-22 09:32:11 -07:00
Ishaan Jaffer
e80bba83e3
test fix
2025-10-22 09:29:04 -07:00
soo-jin.kim
03e1d93199
fix: Apply max_connections configuration to Redis async client ( #15797 )
...
* fix: Apply max_connections configuration to Redis async client
- Add max_connections to available Redis cluster kwargs
- Add connection_pool parameter to get_redis_async_client()
- Pass connection_pool to Redis client if provided
- Prevents Redis connection exhaustion under high load
* test: Add tests for Redis max_connections feature
- Test max_connections is included in cluster kwargs
- Test connection_pool parameter is properly passed to async client
- Test async client works without connection_pool parameter
All 3 tests pass successfully
2025-10-22 09:19:08 -07:00
soo-jin.kim
8050995dbb
fix: Rename configured_cold_storage_logger to cold_storage_custom_logger ( #15798 )
...
- Change variable name in litellm/__init__.py from configured_cold_storage_logger to cold_storage_custom_logger
- Update all references across the codebase to use the new variable name
- This fixes silent failure of cold storage logging due to variable name mismatch
- Configuration files use cold_storage_custom_logger, code should match
Files updated:
- litellm/__init__.py
- litellm/litellm_core_utils/litellm_logging.py
- litellm/proxy/spend_tracking/cold_storage_handler.py
- litellm/responses/litellm_completion_transformation/session_handler.py
- tests/test_litellm/litellm_core_utils/test_litellm_logging.py
- tests/test_litellm/responses/litellm_completion_transformation/test_session_handler.py
2025-10-22 09:17:08 -07:00
nuernber
69946bb35b
fix the date for sonnet 3.7 in govcloud ( #15800 )
2025-10-22 09:14:07 -07:00
Anthony Ivan
5f7a6b49eb
Feat: Allow prompt caching to be used for Anthropic Claude on Databricks ( #15801 )
2025-10-22 09:11:04 -07:00
Sameer Kankute
44495c0117
fix encrypted content error ( #15782 )
2025-10-21 23:29:48 -07:00
Krish Dholakia
29a97784e7
(feat) Passthrough - set auth on passthrough endpoints, on the UI ( #15778 )
...
* fix(add_pass_through.tsx): allow setting 'auth' to true for passthrough endpoints on the UI
* fix: working update auth on passthrough endpoints + show auth on passthrough table
2025-10-21 23:26:43 -07:00
Ishaan Jaffer
02e34a57d6
anthropic.claude-3-7-sonnet-20240620-v1:0
2025-10-21 19:19:52 -07:00
Ishaan Jaffer
bd0a8a047a
docs search_tools
2025-10-21 19:15:18 -07:00
Ishaan Jaffer
adc503c091
bump: version 1.78.6 → 1.78.7
2025-10-21 19:06:05 -07:00
Ishaan Jaff
f5a80110c1
[Feat] Add /search endpoint on LiteLLM Gateway ( #15780 )
...
* add SearchProvider
* add SearchToolTypedDict
* add search
* add SearchAPIRouter
* working router level search
* add search to allowed llm / ocr routes
* feat: add search_router
* add routing + proxy for search APIs
* /v1/search/{search_tool_name}
* fix search routing
* feat: parse_search_tools
* clean up sidebar
* docs fix
* router tests for search tools
* docs fix
2025-10-21 19:05:20 -07:00
Ishaan Jaffer
d9b85ab276
fix: rename search_provider
2025-10-21 17:42:18 -07:00
Ishaan Jaffer
bea8e13a94
fix: GuardrailConfigModel
2025-10-21 17:11:56 -07:00
wangjifeng
8cbaec0310
feat: Add imageConfig parameter for gemini-2.5-flash-image ( #15530 )
...
* Add imageConfig parameter support for Vertex AI to enable gemini-2.5-flash-image model requirements
* Add test for imageConfig parameter support in Vertex AI Gemini transformation
2025-10-21 17:08:30 -07:00
Ishaan Jaff
7b939b4558
[Feat] Add EXA AI Search API to LiteLLM ( #15774 )
...
* add BaseSearchConfig
* add BaseSearchConfig
* validate_environment
* fix handlers
* add PerplexitySearchConfig
* add PerplexitySearchConfig
* add LiteLLM Search API module.
* add BaseSearchConfig
* add _build_search_optional_params
* add search_testing
* add BaseSearchTest
* add TestPerplexitySearch
* fix BASE
* fix handler
* add search API
* add to init
* fix: working perplexity search API
* add _hidden_params to search
* add TAVILY to LlmProviders
* add TavilySearchConfig
* add TavilySearchConfig
* TestTavilySearch
* add tavily transform
* TestParallelAISearch
* add LlmProviders
* add ParallelAISearchConfig
* add ParallelAISearchConfig
* ParallelAISearchConfig
* add EXA AI Search API
* add ExaAISearchConfig
* TestExaAISearch
* add get_supported_perplexity_optional_params
* add Exa AI Search API
* add transform_search_request
* add ExaAISearchConfig
* fix linting errors
* transform_search_request
2025-10-21 17:06:23 -07:00
Ishaan Jaff
208f76f8ad
[Feat] Add Parallel AI - Search API ( #15772 )
...
* add BaseSearchConfig
* add BaseSearchConfig
* validate_environment
* fix handlers
* add PerplexitySearchConfig
* add PerplexitySearchConfig
* add LiteLLM Search API module.
* add BaseSearchConfig
* add _build_search_optional_params
* add search_testing
* add BaseSearchTest
* add TestPerplexitySearch
* fix BASE
* fix handler
* add search API
* add to init
* fix: working perplexity search API
* add _hidden_params to search
* add TAVILY to LlmProviders
* add TavilySearchConfig
* add TavilySearchConfig
* TestTavilySearch
* add tavily transform
* TestParallelAISearch
* add LlmProviders
* add ParallelAISearchConfig
* add ParallelAISearchConfig
* ParallelAISearchConfig
2025-10-21 17:00:05 -07:00
Ishaan Jaff
b9f3f9fb79
[Feat] Add Tavily Search API ( #15770 )
...
* add BaseSearchConfig
* add BaseSearchConfig
* validate_environment
* fix handlers
* add PerplexitySearchConfig
* add PerplexitySearchConfig
* add LiteLLM Search API module.
* add BaseSearchConfig
* add _build_search_optional_params
* add search_testing
* add BaseSearchTest
* add TestPerplexitySearch
* fix BASE
* fix handler
* add search API
* add to init
* fix: working perplexity search API
* add _hidden_params to search
* add TAVILY to LlmProviders
* add TavilySearchConfig
* add TavilySearchConfig
* TestTavilySearch
* add tavily transform
2025-10-21 16:59:29 -07:00
wenhua
b0ccc35a9c
fix(ollama): Enhance chunk parsing for empty responses without 'thinking' and improve error logging ( #13333 ) ( #15717 )
2025-10-21 16:59:01 -07:00
Ishaan Jaff
e1cb92862e
[Feat] Add def search() APIs for Web Search - Perplexity API ( #15769 )
...
* add BaseSearchConfig
* add BaseSearchConfig
* validate_environment
* fix handlers
* add PerplexitySearchConfig
* add PerplexitySearchConfig
* add LiteLLM Search API module.
* add BaseSearchConfig
* add _build_search_optional_params
* add search_testing
* add BaseSearchTest
* add TestPerplexitySearch
* fix BASE
* fix handler
* add search API
* add to init
* fix: working perplexity search API
* add _hidden_params to search
2025-10-21 16:58:51 -07:00
Ishaan Jaff
9135e748a0
[Feat ] /ocr - Add mode + Health check support for OCR models ( #15767 )
...
* get_mode_handlers
* use get_mode_handlers
* test_ahealth_check_ocr
* Add OCR mode to test models
* docs OCR Health Checks
* fix connection endpoint
2025-10-21 16:58:37 -07:00
Javier Garcia
b0a3a7c4fb
Add details in docs ( #15721 )
...
* Add details in docs
* add logic to set span attributes and unit tests
* Restore html files
* Remove html files
* Remove html files
2025-10-21 16:57:51 -07:00
Thomas Mildner
1cfc4624c3
[Feat] Add SENTRY_ENVIRONMENT configuration for Sentry integration ( #15760 )
...
* [Feat] Add SENTRY_ENVIRONMENT configuration for Sentry integration and corresponding tests
* [Refactor] Enhance test_sentry_environment by mocking sentry_sdk and improving environment handling
* [Fix] Update default SENTRY_ENVIRONMENT to 'production' and enhance test for Sentry integration
* [Fix] Update test_sentry_environment to verify correct handling of SENTRY_ENVIRONMENT values
* [Fix] Update test_sentry_environment to assert correct handling of production environment
2025-10-21 16:40:55 -07:00
Kowyo
1fcadd6c05
feat(ollama): set 'think' to False when reasoning effort is not high/medium/low ( #15763 )
2025-10-21 16:39:08 -07:00
Krrish Dholakia
2a1dbb5b9e
docs(creating_adapters.md): document how to write an adapter
2025-10-21 16:20:31 -07:00
nuernber
353dfb1238
Add AWS us-gov-west-1 Claude 3.7 Sonnet costs ( #15775 )
...
* add us-gov-west-1 claude 3.7 sonnet to prices
* add to _backup file as well
2025-10-21 16:17:07 -07:00
YutaSaito
39641e7e68
chore: rename GraySwan to Gray Swan ( #15771 )
2025-10-21 15:18:55 -07:00
Vinod Singh
d4aadda692
Auth Header Fix for MCP Tool Call ( #15736 )
...
* fixed the Auth header for MCP Tool Call
* Final fix for Auth header
* testcase for mcp_auth_header_extraction, insensitive_alias_matching, insensitive_servername_matching added
2025-10-21 13:58:03 -07:00
Krrish Dholakia
1e0368521e
refactor: cleanup
2025-10-21 13:46:19 -07:00
Ishaan Jaffer
3741c43396
docs fix
2025-10-21 13:21:59 -07:00
Ishaan Jaff
8ad9bbbd02
[Docs] Add Azure AI - OCR to docs ( #15768 )
...
* add Azure OCR to docs
* docs fix
* docs fix
* docs fix
* docs OCR
2025-10-21 13:10:45 -07:00
Ishaan Jaffer
185182bebc
Revert "add Azure OCR to docs"
...
This reverts commit a3699e28a4 .
2025-10-21 13:02:31 -07:00
Ishaan Jaffer
a3699e28a4
add Azure OCR to docs
2025-10-21 13:02:21 -07:00
Ishaan Jaffer
6605aba307
docs grayswan
2025-10-21 11:16:25 -07:00
YutaSaito
d79bdd491f
feat: add GraySwan Guardrails support ( #15756 )
2025-10-21 11:13:50 -07:00
Talal
46d55bd92a
fix: Add response_type + PKCE parameters to OAuth authorization endpoint ( #15720 )
...
* fix: Add response_type parameter to OAuth authorization endpoint
Fixes #15684
OAuth providers like Google require the response_type parameter during
the authorization flow. This commit adds response_type=code to the
authorization redirect parameters, which is required by the OAuth 2.0
specification (RFC 6749 Section 4.1.1).
Changes:
- Added response_type=code to authorization params in discoverable_endpoints.py
- Added test coverage for the response_type parameter
🤖 Generated with [Claude Code](https://claude.com/claude-code )
Co-Authored-By: Claude <noreply@anthropic.com>
* fix oauth flow by forwarding code_challenge and forwarding code_verifier
---------
Co-authored-by: Claude <noreply@anthropic.com>
2025-10-21 09:43:19 -07:00
Tom Haynes
98f1d63508
use correct otel logger, and normalise otel paths ( #15645 )
2025-10-21 09:16:03 -07:00
Ishaan Jaffer
8b522d88a2
is_llm_api_route
2025-10-20 18:05:35 -07:00
Ishaan Jaff
92335d991c
[Feat] Add Azure AVA (Speech AI) Cost Tracking ( #15754 )
...
* add azure/speech/ cost tracking
* test_azure_ava_tts_async
* add azure/speech to model cost map
* docs cost tracking
* docs tts AVA
* add azure/speech/azure-tts
2025-10-20 18:01:51 -07:00
Ishaan Jaffer
60fab591db
rename test files
2025-10-20 18:00:17 -07:00
Ishaan Jaff
157739da01
[Bug]: Fix Incorrect status value in responses api with gemini ( #15753 )
...
* _map_chat_completion_finish_reason_to_responses_status
* test_transform_chat_completion_response_with_reasoning_content
* test_transform_chat_completion_response_output_item_status
2025-10-20 17:58:56 -07:00
Ishaan Jaffer
5ce2be732e
get_provider_text_to_speech_config
2025-10-20 17:10:09 -07:00
Ishaan Jaffer
9a25eeccb2
docs fix
2025-10-20 17:02:38 -07:00
Ishaan Jaffer
c9152003bd
bump V
2025-10-20 16:55:03 -07:00
Ishaan Jaff
73a23a6c78
[Feat] Add Azure AVA TTS integration ( #15749 )
...
* add AzureBaseIssueTokenHandler
* add BaseTextToSpeechConfig
* async_text_to_speech_handler
* add AzureAVATextToSpeechConfig
* add get_provider_text_to_speech_config
* add AzureAVATextToSpeechConfig
* fixes for base_llm_http_handler
* fix transform_text_to_speech_request
* test_azure_ava_tts_async
* test_azure_ava_tts_async
* fix TextToSpeechRequestData
* fix transform_text_to_speech_request
* add text_to_speech_handler in LLMHttpHandler
* remove old file
* fix transform_text_to_speech_request
* fix dispatch_text_to_speech
* fix azure TTS
* fix AVA TTS
* fix transform
* fix linting
* ci/cd - use one job for audio testing
* fix tests
* fix llm http handler debugging
* unit tests azure tts
* docs Azure speech
* docs fix
* docs azure AVA
* docs azure AVA
* fix handlers
* test_async_realtime_uses_max_size_parameter
2025-10-20 16:52:23 -07:00
akraines
41a6ecd5b6
Change max_tokens value to match max_output_tokens for claude sonnet 4.5: 64000 ( #15715 )
...
See https://github.com/RooCodeInc/Roo-Code/issues/8454
2025-10-20 16:11:36 -07:00
Ishaan Jaff
0c25b1a256
[Fix] OpenAI Realtime API integration fails due to websockets.exceptions.PayloadTooBig error ( #15751 )
...
* fix REALTIME_WEBSOCKET_MAX_MESSAGE_SIZE_BYTES
* edit max_size for websockets
* fix AzureOpenAIRealtime
2025-10-20 15:54:14 -07:00
Sameer Kankute
1fb798f81d
(Bug) Fix JSON serialization error in Helicone logging by removing OpenTelemetry span from metadata ( #15728 )
...
* remove span object from helicon metadata
* Add test
2025-10-20 08:53:22 -07:00
Sameer Kankute
3955a3de5d
fix the wrong request body in json mode doc ( #15729 )
2025-10-20 08:44:14 -07:00