Commit graph

30699 commits

Author SHA1 Message Date
houdataali
ebcc37dfdf
add redisvl dependency to the root requiremnts.tx (#19417) 2026-01-21 20:12:07 -08:00
Harshit Jain
746414eb9b
Fix/per service ssl override v2 (#19538)
* refactor(ssl): support per-service SSL verification overrides

* add test cases for ssl
2026-01-21 20:10:04 -08:00
davida-ps
7777aeb695
fixing prompt-security's guardrail implementation (#19374)
* Consolidated change

* fix(prompt_security): update message processing to persist sanitized files and filter for API calls

* fix per krrishdholakia suggestion
2026-01-21 20:09:40 -08:00
jay prajapati
363b0cc132
fix(azure): preserve content_policy_violation details for images (#19328) (#19372)
Azure OpenAI Images (DALL·E 3) returns policy violations as a structured payload under body["error"], including inner_error.content_filter_results and revised_prompt.

LiteLLM previously:
- Failed to extract nested error messages (get_error_message only handled body["message"])
- Missed policy violation detection when error strings were generic
- Dropped inner_error details when raising ContentPolicyViolationError

This change:
- Extracts nested Azure error fields (code/type/message + inner_error)
- Detects policy violations via structured error codes
- Passes an OpenAI-style error body + provider_specific_fields to preserve details

Tests:
- python3 -m pytest tests/test_litellm/llms/azure/test_azure_exception_mapping.py
- python3 -m pytest tests/test_litellm/litellm_core_utils/test_exception_mapping_utils.py

Fixes #19328
2026-01-21 20:06:51 -08:00
Cesar Garcia
4106d24215
feat: add GMI Cloud provider support (#19376)
* feat: add GMI Cloud provider support

Add GMI Cloud as an OpenAI-compatible provider with:
- Provider configuration in providers.json
- Documentation page with usage examples
- Model pricing for 16 models (Claude, GPT, DeepSeek, Gemini, etc.)
- Sidebar entry for docs navigation

* Add gmi_cloud to provider_endpoints_support.json

Add provider entry to pass CI validation check that ensures all
providers in openai_like/providers.json are documented.

* Fix provider key: gmi_cloud -> gmi

Match the provider key with providers.json

---------

Co-authored-by: Krish Dholakia <krrishdholakia@gmail.com>
2026-01-21 15:48:15 -08:00
Ryne Carbone
15013cec4b
feat(gemini): add file content support in tool results (#19416)
Add support for 'file' and 'input_file' content types in
convert_to_gemini_tool_call_result(). File content in tool
results was previously silently dropped.

Supports base64 data URIs and HTTP URLs, matching the existing
image handling pattern. Enables PDF, audio, video, and other
file types as inline_data for Gemini.
2026-01-20 19:54:12 -08:00
Cesar Garcia
12c6dde1c3
fix(ui): increase model selector width in playground for larger screens (#19423)
Make the model/agent selector responsive with wider widths on tablet (256px) and desktop (288px) for better usability.
2026-01-20 19:47:52 -08:00
Lucky-Lodhi2004
efbdf9d60d
fix #19414 - [Bug]: get_model_info on Bedrock suffixed models doesn't return proper information (#19421)
* fixed get_model_info for bedrock models

* fixed import
2026-01-20 19:45:55 -08:00
Harshit Jain
76433f9f04
fix: add better error handling for misconfig on health check (#19441) 2026-01-20 19:24:28 -08:00
Sampson
09941dd1d1
add search provider for brave search api (#19433)
* add search provider for brave search api

Introduces a minimal implementation of the Brave Search API as a search provider. Additionally, this PR introduces a test file to ensure the provider works properly, and numerous other smaller changes (e.g., changes to docs to mention the new option).

* Update transformation.py
2026-01-20 19:23:29 -08:00
Kamil Jopek
ce722ab763
Make grpc dependency optional (#19447)
* Make grpc optional and document gRPC OTEL setup

* Add tests for missing OTLP gRPC imports
2026-01-20 19:03:52 -08:00
Harshit Jain
3c3148789a
fix: resolve Read-only file system error in non-root images (#19449) 2026-01-20 19:00:52 -08:00
yuneng-jiang
498ee5a662
Merge pull request #19452 from BerriAI/litellm_cicd_fix_yj_04
[Infra] Increase Time to Wait for Spend Accuracy Tests
2026-01-20 16:34:05 -08:00
yuneng-jiang
65829411e0 increasing time for spend tracking 2026-01-20 16:31:25 -08:00
yuneng-jiang
d5305c6a61
Merge pull request #19450 from BerriAI/litellm_cicd_fix_yj_02
[Infra] Fix test_route_checks
2026-01-20 16:22:20 -08:00
yuneng-jiang
b37e42ac0f
Merge pull request #19451 from BerriAI/litellm_cicd_fix_yj_03
[Infra] Use mock db for claude code marketplace tests
2026-01-20 16:19:15 -08:00
yuneng-jiang
1a9a7df437 use mock db for cluade code marketplace 2026-01-20 16:18:21 -08:00
yuneng-jiang
232ae52b94 attempt test_route_checks fix 2026-01-20 15:55:44 -08:00
yuneng-jiang
e0811ad848
Merge pull request #19448 from BerriAI/litellm_cicd_fix_yj_01
[Infra[ Fixing dynamic_router_retry_policy CI
2026-01-20 15:52:40 -08:00
yuneng-jiang
dd9e8833db Fixing dynamic_router_retry_policy 2026-01-20 15:51:57 -08:00
yuneng-jiang
231023c422
Merge pull request #19446 from BerriAI/migration_fix_yj
[Infra] Fixing LiteLLM Proxy Extras
2026-01-20 15:24:15 -08:00
Cesar Garcia
94055741d4
docs: clarify Gemini vs Vertex AI model prefix behavior (#19443)
Add documentation explaining the difference between model formats:
- `gemini/model` → Gemini API (simple API key)
- `vertex_ai/model` → Vertex AI (GCP credentials)
- `model` (no prefix) → defaults to Vertex AI

This addresses user confusion when models without prefix require
GCP authentication instead of simple API key auth.

Ref #8424
2026-01-20 15:22:52 -08:00
yuneng-jiang
4b25ae6693 Adding build artifacts 2026-01-20 15:22:47 -08:00
yuneng-jiang
2c2e0649d9 bump: version 0.4.24 → 0.4.25 2026-01-20 15:22:18 -08:00
Cesar Garcia
2b44d02682
fix: add google-cloud-aiplatform as optional dependency with clear error message (#19437)
- Add google-cloud-aiplatform as optional dependency in pyproject.toml
- Add 'google' extra for easy installation: pip install litellm[google]
- Improve error messages when Google SDK is not installed to guide users

Fixes #5483
2026-01-20 15:22:13 -08:00
yuneng-jiang
0bcf7097d2 bump: version 0.4.23 → 0.4.24 2026-01-20 15:22:11 -08:00
yuneng-jiang
2b62e9fedf
Merge pull request #19440 from BerriAI/litellm_ui_chat-autofill
[Feature] UI - Playground: Button to Fill Custom API Base
2026-01-20 14:23:39 -08:00
yuneng-jiang
71a2fc5331 fix classnames 2026-01-20 14:16:03 -08:00
Cesar Garcia
7515f179e7
fix: sync Helm chart version with LiteLLM release version (#19438)
Replace independent auto-incrementing chart versioning with 1-1 sync
to LiteLLM version. This allows users to easily map Helm chart versions
to LiteLLM versions without needing to inspect appVersion.

Changes:
- Remove auto-increment logic that read from OCI registry
- Chart version now equals LiteLLM tag without 'v' prefix (v1.81.0 -> 1.81.0)
- appVersion equals full Docker tag (v1.81.0)
- Update both ghcr_deploy.yml and ghcr_helm_deploy.yml workflows

Before: helm chart 0.1.837 -> user has to guess LiteLLM version
After:  helm chart 1.81.0  -> matches LiteLLM v1.81.0

References:
- https://codefresh.io/docs/docs/ci-cd-guides/helm-best-practices/
2026-01-20 14:13:26 -08:00
yuneng-jiang
f434d1c847 Option to pre fill custom proxy base URL 2026-01-20 14:07:54 -08:00
Alexsander Hamir
7f81dea8b3
Add custom auth header support and increase default prompt size to 100k chars (#19436) 2026-01-20 13:25:12 -08:00
yuneng-jiang
e142474e0b
Merge pull request #19431 from BerriAI/litellm_ui_fix_build_002
[Infra] UI - Fixing UI Build
2026-01-20 13:17:25 -08:00
yuneng-jiang
3ae71bf49e fixing ui build 2026-01-20 13:03:27 -08:00
yuneng-jiang
bfb94f56b7
Merge pull request #19276 from stiyyagura0901/litellm_fix_ui_auth_header_override
fix: UI dashboard respects custom authentication header override
2026-01-20 12:24:14 -08:00
Alexsander Hamir
5a06868652
Fix in-flight request termination on SIGTERM when health-check runs in a separate process (#19427) 2026-01-20 12:17:06 -08:00
Krrish Dholakia
f95f5563ea docs: document input/output/total tokens behaviour
Closes https://github.com/BerriAI/litellm/issues/17480
2026-01-20 10:45:47 -08:00
Alexsander Hamir
1377721715
Fix: Handle PostgreSQL cached plan errors during rolling deployments (#19424) 2026-01-20 10:44:31 -08:00
Otavio Brito
ce37729da4
remove count tokens optional param before request is sent to vertex (#19359) 2026-01-20 09:03:27 -08:00
Ishaan Jaffer
f6d6455cbc fix rc 2026-01-20 08:39:17 -08:00
Sameer Kankute
11dbae85d1
Merge pull request #19390 from BerriAI/litellm_consistent_id_streaming_responses
Fix: ID mismatch between text-start and text-delta
2026-01-20 20:46:34 +05:30
Sameer Kankute
961a424069
Merge pull request #19355 from BerriAI/litellm_staging_01_19_2026
Litellm staging 01 19 2026
2026-01-20 20:45:12 +05:30
Sameer Kankute
9e1275b76c
Merge branch 'main' into litellm_staging_01_19_2026 2026-01-20 19:19:36 +05:30
Sameer Kankute
a3c1f4758d
Merge branch 'main' into litellm_consistent_id_streaming_responses 2026-01-20 19:02:23 +05:30
Sameer Kankute
e69c12b6db
Merge pull request #19396 from BerriAI/litellm_responses_route_fix
Fix for Prometheus Metric Cardinality Issue with /responses Endpoint
2026-01-20 19:01:18 +05:30
Sameer Kankute
bd6f7bae21
Merge pull request #19397 from BerriAI/litellm_google_computer_use_cost_tracking
Add gemini-2.5-computer-use-preview-10-2025 model for vertex ai provider
2026-01-20 19:00:34 +05:30
Sameer Kankute
172ad17fbc
Merge pull request #19398 from BerriAI/litellm_add_multimodal_cost_tracking
Add input_cost_per_video_per_second in ModelInfoBase
2026-01-20 19:00:05 +05:30
Sameer Kankute
37ce6957ab
Merge pull request #19386 from BerriAI/litellm_staging_01_20_2026
Litellm staging 01 20 2026
2026-01-20 18:53:59 +05:30
Sameer Kankute
12c556e485
Merge pull request #19409 from BerriAI/revert-19261-feat/redis-migration-lock-safe
Revert "feat: Add Redis-based migration lock with bug fixes"
2026-01-20 18:46:38 +05:30
Sameer Kankute
3cc19c56ba
Revert "feat: Add Redis-based migration lock with bug fixes (#19261)"
This reverts commit 98e87c3e67.
2026-01-20 18:46:26 +05:30
Sameer Kankute
dc3ee63359 fix: test_env_keys 2026-01-20 18:37:56 +05:30