Commit graph

41449 commits

Author SHA1 Message Date
R.Sicart
608979c7e9
feat: add support for keda in helm chart (#19337)
* feat: add support for keda in helm chart

Signed-off-by: R.Sicart <roger.sicart@gmail.com>

* chore: bump chart version

---------

Signed-off-by: R.Sicart <roger.sicart@gmail.com>
2026-01-19 10:38:41 -08:00
Cesar Garcia
4ad5de10cb
fix(realtime): disable SSL for ws:// WebSocket connections (#19345)
When using http:// api_base (converted to ws://), the websockets library
throws "ssl argument is incompatible with a ws:// URI". Only pass SSL
context for secure wss:// connections.

Co-authored-by: Krish Dholakia <krrishdholakia@gmail.com>
2026-01-19 10:37:41 -08:00
Harshit Jain
99c4ba7adf
docs: fix bad examples from sdk (#19322) 2026-01-19 10:27:25 -08:00
Harshit Jain
1678f621db
feat: add retry_delay, exponential_backoff, and jitter to completion() (#19371) 2026-01-19 10:27:01 -08:00
Ishaan Jaff
e817aa713e
[Fix] Claude Code x Bedrock Invoke fails with advanced-tool-use-2025-11-20 (#19373)
* _filter_unsupported_beta_headers_for_bedrock

* test_bedrock_sonnet_4_5_with_advanced_tool_use_beta_header
2026-01-19 10:16:18 -08:00
Lucky Lodhi
74d3b11290 undid changes 2026-01-19 17:38:29 +00:00
Lucky Lodhi
c4013a34b8 fix tool call for ollama - #19357 2026-01-19 17:20:18 +00:00
Alexsander Hamir
0cd7763d5f
Add health check scripts and parallel execution support (#19295)
- Add health_check_client.py for monitoring model availability
- Add health_check_client_README.md with usage documentation
- Add health_check_requirements.txt for dependencies
- Add run_parallel_health_checks.ps1 (PowerShell version)
- Add run_parallel_health_checks.sh (Bash version)
- Organize all scripts under scripts/health_check/ directory
2026-01-19 08:38:38 -08:00
Krish Dholakia
0862373b38
docs: add note about no limits on users/keys/teams in LiteLLM OSS (#19367)
Co-authored-by: Cursor Agent <cursoragent@cursor.com>
2026-01-19 08:22:18 -08:00
Sameer Kankute
bb67f9267a
Merge pull request #19354 from BerriAI/litellm_fix_streaming_test
Fix : test_responses_streaming_failure_triggers_failure_handlers
2026-01-19 21:22:53 +05:30
Benedikt Óskarsson
406cdbe321
Merge branch 'litellm_staging_01_19_2026' into fix/bedrock-thinking-tool-call-2 2026-01-19 15:18:47 +00:00
Benedikt Óskarsson
0367e9c9f1
fix: pr ammends. 2026-01-19 14:50:57 +00:00
Chesars
f76719750b feat(helicone): add Vertex AI support for non-Gemini models
Extends HeliconeLogger to properly log Vertex AI partner models (GLM, DeepSeek, etc.)
that don't contain "gemini" in their name. Uses custom_llm_provider to detect vertex_ai.
2026-01-19 11:35:34 -03:00
Lucky Lodhi
15e8eea0b5 Merge remote-tracking branch 'upstream/main' into fix-litellm-params 2026-01-19 14:21:44 +00:00
superpoussin22
1be7e87783
Fix HTML entity in survey description text (#19307) 2026-01-19 06:20:08 -08:00
Cesar Garcia
4d6a430adc
docs: update UI contributing guide (#19353)
* docs: update UI contributing guide with correct commands

- Replace outdated proxy_cli.py command with poetry run litellm
- Add config.yaml example with required settings
- Clarify that UI comes pre-built in the repo
- Add two development options: Build Mode and Dev Mode (hot reload)
- Note about redirect issues in Dev Mode

* docs: add hot reload login flow and PR submission section

- Document the 3000 -> 4000 -> 3000 login flow for hot reload
- Reorder: Hot Reload as Option A, Build Mode as Option B
- Add section 4 on submitting PRs
- Add note that UI changes don't require tests

* Update login flow navigation URL in contributing.md
2026-01-19 06:18:45 -08:00
Harshit Jain
fc9988b686
fix/bedrock-inconsistent-postcall-hook (#19151)
* fix/bedrock-inconsistent-postcall-hook

* Add condition check to avoid multiple validation
2026-01-19 06:18:02 -08:00
Sameer Kankute
9e405ce6cc
Merge pull request #19323 from BerriAI/litellm_fix_stability_issues12
[Fix] Bedrock stability model usage issues
2026-01-19 19:43:32 +05:30
Sameer Kankute
ff7bb59824
Merge branch 'main' into litellm_fix_streaming_test 2026-01-19 19:43:16 +05:30
Sameer Kankute
c9de4776bc Fix test_process_chunk_exception_calls_handle_failure_once 2026-01-19 19:39:12 +05:30
Sameer Kankute
d6baa9a4ba
Merge pull request #19234 from BerriAI/litellm_staging_01_16_2026
Litellm staging 01 16 2026
2026-01-19 19:34:53 +05:30
Harshit Jain
1dc2d2ddac
fix(utils.py): correctly extract messages from google genai contents (#19156)
* fix(utils.py): correctly extract messages from google genai contents

* refactor use shared utilities
2026-01-19 06:00:23 -08:00
Harshit Jain
98e87c3e67
feat: Add Redis-based migration lock with bug fixes (#19261) 2026-01-19 05:57:24 -08:00
Harshit Jain
fe92f4af9c
fix(langfuse_otel): ignore service logs and fix callback shadowing (#19298)
* fix(langfuse_otel): ignore service logs and fix callback shadowing

* add test cases for service logger
2026-01-19 05:53:47 -08:00
Harshit Jain
07fbd77c91
fix(logging): prevent duplicate StandardLoggingPayload logs (#19325) 2026-01-19 05:53:23 -08:00
Sameer Kankute
ac24696556 Remove runtime error 2026-01-19 19:19:16 +05:30
Cesar Garcia
b49f0a91e4
fix(responses): resolve deepcopy error with tool_choice ValidatorIterator (#17192) (#17205)
Replace copy.deepcopy with model_dump + model_validate in streaming
iterator logging to handle Pydantic ValidatorIterator objects that
cannot be pickled when tool_choice uses allowed_tools mode.

Co-authored-by: Krish Dholakia <krrishdholakia@gmail.com>
2026-01-19 05:44:20 -08:00
Sameer Kankute
daf70f7221
Merge pull request #19329 from BerriAI/litellm_vector_store_sync
Fix: vector store sync issues
2026-01-19 19:11:48 +05:30
Sameer Kankute
e126e98a7f Fix azure image mypy issues 2026-01-19 19:10:43 +05:30
Manuel Schweigert
29adf34313
Add ChatGPT subscription support and responses bridge (#19030)
* Add ChatGPT subscription support and responses bridge

* Fix typing import for responses bridge

* Guard device code timestamp parsing

* add /v1/messages endpoint to chatgpt model
2026-01-19 05:37:45 -08:00
Harshit Jain
32562708ee
fix(proxy): add /a2a/{agent_id}/.well-known/agent-card.json to agent_routes allowlist (#19277) 2026-01-19 05:36:54 -08:00
Sameer Kankute
42175a411d Fix streaming tests 2026-01-19 19:04:52 +05:30
Sameer Kankute
443b888d84
Merge pull request #19352 from BerriAI/revert-19158-main
Revert "Fix audio cost per second override"
2026-01-19 18:52:08 +05:30
Sameer Kankute
574391c118
Revert "Fix audio cost per second override (#19158)"
This reverts commit 2a0f87bde0.
2026-01-19 18:51:08 +05:30
Jón Levy
5db0e3289a
fix(agentcore): simplify agentcore streaming (#17141)
* fix(agentcore): simplify agentcore streaming

* fix(agentcore): move CustomStreamWrapper import to module level

The deferred imports inside streaming methods caused initialization delays
during health check requests, leading to timeouts in ECS deployments.

- Move CustomStreamWrapper import to module-level (line 19)
- Remove deferred imports from get_sync_custom_stream_wrapper (line 588)
- Remove deferred import from get_async_custom_stream_wrapper (line 747)
- Remove from TYPE_CHECKING block to use actual import

This ensures the import happens at module load time rather than during
first request processing, preventing health check endpoint blocking.

🤖 Generated with [Claude Code](https://claude.com/claude-code)

Co-Authored-By: Claude <noreply@anthropic.com>

* test(agentcore): ensure sync response

* chore: upgrade boto3 to 1.40.76 in pyproject.toml

* chore: added taplo.toml

* fix(types): correct annotation type hint for MyPy compatibility

Update _convert_annotations_to_chat_format return type from
Dict[str, Any] to ChatCompletionAnnotation TypedDict to match
the Message class's expected type signature.

Co-Authored-By: Claude <noreply@anthropic.com>

---------

Co-authored-by: Claude <noreply@anthropic.com>
Co-authored-by: Benedikt Óskarsson <bensi94@hotmail.com>
2026-01-19 05:20:24 -08:00
Sameer Kankute
6eb3f579d7 Fix module import for opentelemetry 2026-01-19 18:50:18 +05:30
Harshit Jain
6cd4b3603f
fix(router): prevent retrying 4xx client errors (#19275) 2026-01-19 05:18:35 -08:00
Sameer Kankute
2e6ce1469e Fix module import for opentelemetry 2026-01-19 18:47:28 +05:30
Benedikt Óskarsson
f09cae2107
Merge branch 'main' into fix/bedrock-thinking-tool-call-2 2026-01-19 13:08:17 +00:00
Sameer Kankute
75932e08da
Merge pull request #19349 from BerriAI/litellm_doc_update_05
Fix: update the doc
2026-01-19 18:25:43 +05:30
Sameer Kankute
351f921561
Merge pull request #19347 from BerriAI/litellm_fix_replicate_output
Fix Output None for replicate handler
2026-01-19 18:25:25 +05:30
Sameer Kankute
480fa13b1d
Merge pull request #19343 from BerriAI/litellm_anthropic_header_fix_19_jan
Fix: anthropic-beta is getting overriden and set to anthropic-beta
2026-01-19 18:24:46 +05:30
Sameer Kankute
a9475be06d
Merge pull request #19338 from BerriAI/litellm_fix_managed_load_balancing_batches
Add managed files support when load_balancing is True
2026-01-19 18:24:21 +05:30
Sameer Kankute
a8883a45bf
Merge pull request #19327 from BerriAI/litellm_vertex_ai_file_upload
Fix: upload pdfs for file endpoint
2026-01-19 18:23:41 +05:30
Sameer Kankute
68294228c2
Merge pull request #19326 from BerriAI/litellm_handle_failer_2_times
Fix: _handle_failure method getting called 2 times
2026-01-19 18:22:44 +05:30
Sameer Kankute
896d1a7dad Fix Error: Found packages that need verification: 2026-01-19 18:18:24 +05:30
Sameer Kankute
c5a8d4e34e
Merge branch 'main' into litellm_staging_01_16_2026 2026-01-19 18:11:21 +05:30
Sameer Kankute
d5293af053 Fix: update the doc 2026-01-19 17:57:00 +05:30
Sameer Kankute
58daf3eabf Fix Output None for replicate handler 2026-01-19 17:22:06 +05:30
Chesars
45eb35938b fix: drop_params not dropping prompt_cache_key for non-OpenAI providers
Fixes #19225

Add prompt_cache_key and other missing OpenAI Chat Completions params
to DEFAULT_CHAT_COMPLETION_PARAM_VALUES so drop_params: true works.

Also fix additional_drop_params to filter extra params for all providers,
not just OpenAI/Azure.
2026-01-19 08:49:03 -03:00