Ishaan Jaff
ec32a0a7d7
fix: get_base_completion_call_args
2025-09-13 09:59:08 -07:00
Ishaan Jaff
69e451458c
test_logging_opentelemetry_context_propagation
2025-09-13 09:44:36 -07:00
Krish Dholakia
38efd358eb
Merge pull request #14401 from Noma-Security/noma_non_blocking_monitor_mode
...
Noma non blocking monitor mode & anonymize input support
2025-09-13 09:41:41 -07:00
Krish Dholakia
550feffeb9
Merge pull request #14512 from timelfrink/fix/lm-studio-bearer-header-14502
...
fix(lm_studio): resolve illegal Bearer header value issue
2025-09-13 09:40:30 -07:00
Krish Dholakia
ad9515a81e
Merge branch 'main' into litellm_contributor_prs_09_12_2025_p1
2025-09-13 09:38:43 -07:00
Krish Dholakia
663dbc6080
Merge pull request #14477 from BerriAI/litellm_dev_09_11_2025_p2
...
`/v1/messages` - don't send content block after message w/ finish reason + usage block + `/key/unblock` - support hashed tokens
2025-09-12 19:51:44 -07:00
Krish Dholakia
e254d9013b
Merge pull request #14469 from sashank5644/litellm_log_key_alias_filtering
...
Fixed Log Tab Key Alias filtering inaccurately for failed logs
2025-09-12 19:46:17 -07:00
Ishaan Jaff
93af8fd6ba
[QA] E2E - Testing for bedrock batches api ( #14525 )
...
* add bedrock/batch-anthropic.claude-3-5-sonnet-20240620-v1:0
* test_bedrock_batches_api
* fix
* fix import
* test_bedrock_batches_api
2025-09-12 19:31:19 -07:00
Ishaan Jaff
075a089d82
[Feat] Bedrock Batches - Ensure correct transformation applied to incoming requests ( #14522 )
...
* use is_batch_jsonl_file
* fix valid_content_type
* fix transform_create_file_request
* fix _transform_openai_jsonl_content_to_bedrock_jsonl_content
* test_transform_openai_jsonl_content_to_bedrock_jsonl_content
* fix mypy linting errors
* fix BEDROCK_BATCH_MODEL
* fix working sample
* fix comment
* fix model list
* fix: use with managed batches
* refactor
2025-09-12 18:32:57 -07:00
Arseny Boykov
f4318bccd3
[Performance] Use _PROXY_MaxParallelRequestsHandler_v3 by default again ( #14450 )
...
* Use _PROXY_MaxParallelRequestsHandler_v3 by default (#14352 )
(cherry picked from commit f3fa45cf8fbd5f5cce2f45a7312776d5005fb08e)
(cherry picked from commit 5b680bb4a3 )
* Use random api_key for parallel requests test
* Fix off-by-one error in parallel request rate limit
The rate limiter was incorrectly rejecting requests when the limit was met, but not exceeded. The check in `is_cache_list_over_limit` was `int(counter_value) + 1 > current_limit`, which caused the first request to be rejected if the limit was 1.
This commit removes the `+ 1`, changing the logic to `int(counter_value) > current_limit`. The check now correctly allows requests up to the specified parallel limit.
* Test actual parallel requests
* Ensure rate limiting works correctly for multiple users
* Add sequential rate-limit test
* Revert random key usage
2025-09-12 17:33:55 -07:00
Ishaan Jaff
e87e50328e
[Feat] Bedrock Batches - Working e2e flow to upload file + create batch ( #14518 )
...
* fix: bedrock batches transform
* fix: upload_url
* fixes for model name
* fix upload_url
* fix bedrock batch test
* test_mock_bedrock_file_url_mapping
2025-09-12 15:37:09 -07:00
Tim Elfrink
b84785b5b7
fix(lm_studio): resolve illegal Bearer header value issue
...
- Change default API key from space ' ' to 'fake-api-key'
- Fixes httpcore.LocalProtocolError: Illegal header value b'Bearer '
- Maintains compatibility with explicit API keys and environment variables
- Add comprehensive tests for provider info retrieval
Fixes #14502
2025-09-12 22:41:30 +02:00
Ishaan Jaff
18372f9ebe
Revert "fix vertex ai file upload" ( #14501 )
2025-09-12 12:02:24 -07:00
Sameerlite
fa175e8d90
Fix gemini cli error ( #14417 )
...
* Fix gemini cli error
* Added better handling
---------
Co-authored-by: Ishaan Jaff <ishaanjaffer0324@gmail.com>
2025-09-12 11:56:51 -07:00
Sameer Kankute
1a123b2cd5
Litellm gemini cli bug fix ( #14451 )
...
* Fix gemini cli error
* Add reasoning request support
* Added better handling
* remove other PR code
* refactored code for better structure following
---------
Co-authored-by: sameer@berri.ai <sameer@berri.ai>
2025-09-12 11:55:26 -07:00
Boopesh Shanmugam
8b338a4d8c
User Headers X LiteLLM Users Mapping feature ( #14485 )
...
* Draft commit.
* user header mapping feature with backward compatibility with user_header_name field.
* user header mapping feature with backward compatibility with user_header_name field optimizations.
* Added unit tests.
2025-09-12 11:49:37 -07:00
Krish Dholakia
f8036a25a2
Merge pull request #14455 from lmnr-ai/fix/async-logging-tasks-context
...
propagate execution context into logging tasks
2025-09-12 00:36:08 -07:00
Krish Dholakia
113d2a8c5a
Merge pull request #14459 from holzman/fix-provider-budget
...
Fix provider budgets
2025-09-12 00:34:52 -07:00
Krish Dholakia
0413a701a2
Merge pull request #14460 from Sameerlite/litellm_vertex_gcs_bucket_issue
...
fix vertex ai file upload
2025-09-12 00:04:21 -07:00
Ishaan Jaff
32d87c242b
[Fixes] Using Qwen API Tiered Pricing ( #14479 )
...
* fix: use dashscope cost calc
* add qwen logo
2025-09-11 20:07:41 -07:00
Ishaan Jaff
69ef062f55
fix tiered_pricing test
2025-09-11 19:56:44 -07:00
Ishaan Jaff
51d5255452
[Bug]: Azure OpenAI & AI Foundry Reject Image Generation Payload Due to extra_body Injection in LiteLLM v1.76.3 ( #14475 )
...
* add request body azure img gen
* fix test_get_optional_params_image_gen_filters_empty_values
* test_azure_image_generation_request_body
* test_azure_image_generation_request_body
2025-09-11 19:39:06 -07:00
Krrish Dholakia
0c8b311155
test: add unit testing for both flows on key unblock
2025-09-11 19:15:15 -07:00
Krrish Dholakia
805069c287
fix(adapters/streaming_iterator.py): Don't send content block after message delta block is sent
...
Fixes https://github.com/BerriAI/litellm/issues/14315
2025-09-11 18:52:02 -07:00
Ishaan Jaff
dda115cc6d
[Feat] Cost Tracking - Add support for Tiered Cost Tracking for Qwen API (Dashscope) ( #14471 )
...
* add dashscope logo
* docs fix
* docs fix
* fix supports_batch_calling
* fix naming
* fix input_cost_per_audio_token
* use output_cost_per_reasoning_token
* add tiered_pricing in get_model_info
* test fixes
* fix cost calc
* ruff fix
2025-09-11 18:14:39 -07:00
Sashanken
c6626559a2
Fixed Log Tab Key Alias filtering inaccurately for failed logs
2025-09-11 13:05:48 -07:00
Burt Holzman
e9e548d797
Fix provider budgets
2025-09-11 12:08:43 -05:00
Sameer Kankute
090e0fddf4
fix vertex ai file upload
2025-09-11 22:20:54 +05:30
Din
ee5a9d0aa0
propagate execution context into logging tasks
2025-09-11 15:54:40 +01:00
Tom Alon
b83b497d38
PR fixes
2025-09-11 13:38:06 +03:00
Tom Alon
b473344f70
Implement anonymization logic
2025-09-11 11:47:19 +03:00
Ishaan Jaff
258b674dbb
fix deepinfra test
2025-09-10 19:39:23 -07:00
Ishaan Jaff
a13aa4740a
[Fixes] Bug fixes to using LiteLLM MCP Gateway ( #14392 )
...
* fix: use _get_mcp_servers_in_path
* fix checks for using litellm_proxy as MCP tool provider
* fix: fix mcp_tools_with_litellm_proxy
* fix: fix aresponses_api_with_mcp
* aresponses_api_with_mcp
* test_mcp_allowed_tools_filtering
* fix: _filter_mcp_tools_by_allowed_tools
* fix: _filter_mcp_tools_by_allowed_tools
* test_streaming_responses_api_with_mcp_tools
* fixes: test tools transfrom MCP->OpenaI spec
* test_streaming_responses_api_with_mcp_tools
* fix: chat ui allow multi select with allowed tools
* fix: use correct MCP events with litellm proxy response API
* fix get_event_model_class
* fix litellm proxy MCP handler
* fix MCPEnhancedStreamingIterator
* chat ui show list tools result
* UI: show MCP events
* fix stream iterator
* fixes: litellm proxy mcp handler
* test responses + mcp
* fix: update responses api with mcp handling
* ruff check fix
* central: _process_mcp_tools_to_openai_format
* fix: refactor code
* test_mcp_allowed_tools_filtering
* test mcp with litellm proxy
* fix mcp call
* demo: video using MCP ui
* fixes for using stream iterator
* test_no_duplicate_mcp_tools_in_streaming_e2e
* docs fix
* fix code snippet
2025-09-10 19:12:11 -07:00
Ishaan Jaff
1f42e41c8d
[Bug]: Fix Authorization header not being sent to configured MCP servers ( #14422 )
...
* test: test_mcp_server_config_auth_value_header_used
* fix: authentication_token
* docs: fix instructions on using responses api with MCPs
* mcp fixes
2025-09-10 16:41:08 -07:00
Ishaan Jaff
dc5650eeda
Revert "fix: remove anthropic-beta header for Vertex AI requests with prompt caching" ( #14421 )
2025-09-10 15:47:18 -07:00
Tom Alon
f6bc4d0bf9
Noma non blocking on monitor mode
2025-09-10 13:35:17 +03:00
Krish Dholakia
a80cfe5fb5
Merge branch 'main' into feature/databricks-function-call-missing-pass-description
2025-09-09 22:42:32 -07:00
Krish Dholakia
53bf023876
Merge pull request #14092 from gotsysdba/main
...
OCI Provder: Update OCIPromptTokensDetails
2025-09-09 22:39:16 -07:00
Krish Dholakia
d72081113e
Merge pull request #14111 from dharamendrak/feature/aiohttp-dependency-injection
...
feat: Add dependency injection support to BaseLLMAIOHTTPHandler
2025-09-09 22:35:30 -07:00
Krish Dholakia
03a457842d
Merge pull request #14310 from swarna1101/fix-anthropic-prompt-caching-vertex-ai
...
fix: remove anthropic-beta header for Vertex AI requests with prompt caching
2025-09-09 22:32:40 -07:00
Krrish Dholakia
a504c7dae3
test: update tests
2025-09-09 21:43:37 -07:00
Krrish Dholakia
bdd7255bab
test: remove redundant tests - moved to parallel_request_limiter_v3.py
2025-09-09 21:35:50 -07:00
Krrish Dholakia
c45ede7187
test: update test
2025-09-09 21:31:34 -07:00
Krrish Dholakia
f0de7d1dfd
fix: remove EOL model name
2025-09-09 21:15:12 -07:00
Krrish Dholakia
d05f58721e
test: remove end of life model from tests
2025-09-09 21:01:45 -07:00
Krrish Dholakia
e443d01925
test: remove redundant test
2025-09-09 20:37:09 -07:00
Krrish Dholakia
bff76715a5
test: skip test with invalid arn
2025-09-09 20:35:44 -07:00
Krrish Dholakia
07a83056a6
test: update test
2025-09-09 20:30:50 -07:00
Krrish Dholakia
44566977f1
test: update test
2025-09-09 20:18:03 -07:00
Krrish Dholakia
0854c35d3e
test: remove eol bedrock model from tests
2025-09-09 19:48:35 -07:00