yucheng
0a88658227
chore: retrigger ci after docs merge
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:25:56 +00:00
Yuneng Jiang
ddf6565970
Merge remote-tracking branch 'origin/main' into litellm_config_read_source
2026-09-21 13:25:45 -07:00
Yuneng Jiang
be2f0d081b
fix(proxy): report sources only on the read endpoints main does not cover
...
/config/field/info and /config/list already report per-key source on main,
so this drops the branch's versions of those and keeps /alerting/settings,
/get/ui_settings and /router/settings.
Read endpoints no longer write the freshly read database row back into the
shared settings store; the reload path already keeps it current, and a GET
that mutates global state leaks across callers.
Regenerates the lazy OpenAPI snapshot on Python 3.12, matching CI, and the
dashboard API types for the two new response fields.
2026-09-21 13:25:39 -07:00
kerry-berri
89e13ee959
chore: merge litellm_cost_shard_proxy_behaviour into litellm_cost_shard_batches_realtime
2026-09-21 20:25:25 +00:00
kerry-berri
1f71e4c8f5
chore: merge litellm_cost_shard_provider_wires into litellm_cost_shard_proxy_behaviour
2026-09-21 20:25:21 +00:00
kerry-berri
5d3dfe9b70
chore: merge litellm_cost_shard_pricing_dimensions into litellm_cost_shard_provider_wires
2026-09-21 20:25:17 +00:00
kerry-berri
5b51be83b8
chore: merge litellm_cost_shard_passthrough into litellm_cost_shard_pricing_dimensions
2026-09-21 20:25:13 +00:00
kerry-berri
f63a85e9cc
chore: merge litellm_cost_shard_audio_images into litellm_cost_shard_passthrough
2026-09-21 20:25:10 +00:00
Yujong Lee
fe6804ea74
refactor(rust): reject negative CyberArk refresh intervals
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:25:07 +00:00
mateo-berri
61fcfd986d
fix(anthropic): map the dangerous-tool-use beta for Bedrock Mantle so safeguards never reach it without the beta
2026-09-21 13:24:52 -07:00
kerry-berri
134c7111a5
chore: merge litellm_cost_shard_harness_extensions into litellm_cost_shard_audio_images
2026-09-21 20:24:37 +00:00
Yujong Lee
1f86bb8e46
feat(cache-valkey-semantic): add native Valkey semantic cache backend
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:24:34 +00:00
kerry-berri
7b1283ba18
chore: merge litellm_cost_shard_harness_extensions into litellm_cost_shard_embeddings_rerank
2026-09-21 20:24:32 +00:00
Yujong Lee
8d9ab9eeaa
feat(cache-redis): expose the pooled connection handling for reuse
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:24:23 +00:00
kerry
1ae2c0938a
chore: merge main into litellm_cost_shard_harness_extensions
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:23:58 +00:00
Yujong Lee
4e2d4b5ff9
feat(rust): add CyberArk Conjur secret manager backend
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:23:54 +00:00
Yujong Lee
1320eeeb41
refactor(cache-response): generalize ResponseCache over the backend context
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:23:29 +00:00
Yujong Lee
8dc960c928
feat(cache): add SemanticCacheContext
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:22:13 +00:00
Yujong Lee
df97b274fc
refactor(cache-response): generalize ResponseCache over the backend context
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:21:19 +00:00
Yujong Lee
d2f457f144
feat(cache): add semantic cache context and unsupported operation error
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:20:31 +00:00
mateo-berri
bdbb4cf527
test(e2e): send no-cache on rerank bodies like the other request models
2026-09-21 13:20:06 -07:00
Mateo Wang
662e5b6e32
Merge pull request #42284 from BerriAI/litellm_qianwen_ai_platform_rename
...
fix: rename the mainland China brand to Qianwen AI Platform
2026-09-21 13:18:20 -07:00
kerry
ffa1cceb11
refactor(fal_ai): share result request derivation between status fetchers
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:17:40 +00:00
kerry
a909a7908e
fix(fal_ai): surface fal errors in video status and content
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:17:29 +00:00
mateo-berri
56d1f042ef
fix(types): read upstream headers through a typed helper
2026-09-21 13:16:58 -07:00
Mateo Wang
13c604c124
Merge pull request #42296 from BerriAI/litellm_ci_smoke_test_master_key
...
fix(ci): let the install smoke test boot its key-less proxy config
2026-09-21 13:16:44 -07:00
kerry-berri
5216844c40
Merge pull request #42286 from BerriAI/litellm_fal_ai_minimax_h3
...
feat(fal_ai): add MiniMax H3 text-to-video and reference-to-video
2026-09-21 13:16:08 -07:00
Yuneng Jiang
dc85812971
test(migrations): close the gaps the upgrade assertions left open
...
Three holes in the new suite, all of which let a test pass without proving
what its name claims:
- A migration recorded twice, once per replica, each with
applied_steps_count = 1, slipped past both the step-count check and
migration_names(), which collapses the history into a set. Reject
duplicate migration_name rows outright.
- auth_traffic only asserted the failures it had seen by the time
keep_serving hit its target. A request failing after that, or on the
other replica while the test waited on one stream, was recorded and
never read. Assert the recorded failures once the thread has joined.
- The rolling test warmed the baseline replica's virtual-key cache before
the upgrade, and that cache holds for 60 seconds by default
(UserAPIKeyCacheTTLEnum.in_memory_cache_ttl). The candidate migrates
well inside that window, so the post-upgrade requests could be served
from cache without ever repeating the whole-row token lookup that the
stale prepared statement breaks. Drive the baseline replica with a key
minted after the schema moved, which it has never seen and must resolve
from the database.
Re-ran against v1.101.0 -> v1.102.0: 6 passed.
2026-09-21 13:15:46 -07:00
mateo-berri
b0651d52ec
Merge remote-tracking branch 'origin/main' into litellm_safeguards_bedrock_vertex_messages
2026-09-21 13:15:33 -07:00
Joshua Valluru
ef67412e50
fix(mcp): keep OAuth prefetch failure logs free of caller data
2026-09-21 13:13:20 -07:00
kerry-berri
d79eadfe8c
Merge pull request #42297 from BerriAI/litellm-providers/price-sync-openrouter
...
chore(prices): sync OpenRouter prices: 1 model
2026-09-21 13:11:01 -07:00
kerry-berri
dd73c9fe3d
Merge pull request #42298 from BerriAI/litellm-providers/price-sync-aws-bedrock
...
chore(prices): sync AWS Bedrock prices: 1 model [enrichment failed: AWS Bedrock, 16 held]
2026-09-21 13:10:57 -07:00
mateo-berri
f78fdc4ce7
Merge remote-tracking branch 'origin/main' into claude/e2e-tests-custom-endpoints-qxoi1o
...
# Conflicts:
# litellm/model_prices_and_context_window_backup.json
# model_prices_and_context_window.json
2026-09-21 13:10:29 -07:00
kerry-berri
246a6ea54a
Merge pull request #42282 from BerriAI/litellm_fal_price_from_response_dims
...
fix(fal_ai): price images from the dimensions fal returns
2026-09-21 13:08:49 -07:00
kerry
9611af7817
Merge remote-tracking branch 'origin/main' into litellm_upgrade_banner_changelog_stats
2026-09-21 20:06:41 +00:00
yujonglee
18f77e96b5
Merge pull request #42196 from BerriAI/litellm_cache_static_dispatch
...
feat(rust): scaffold cache foundation for Python parity
2026-09-21 13:05:41 -07:00
shivam
db7d52eeda
test(router): track attempted fallback groups via the mock call log
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:01:50 +00:00
berriai-litellm-provider-info-sync[bot]
4a825a7259
chore(prices): sync AWS Bedrock prices: 1 model [enrichment failed: AWS Bedrock, 16 held]
...
zai.glm-5: supports_vision, supports_audio_input, supports_response_schema
2026-09-21 20:01:35 +00:00
berriai-litellm-provider-info-sync[bot]
a2bb1f7e57
chore(prices): sync OpenRouter prices: 1 model
...
openrouter/deepseek/deepseek-v4-pro: input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
2026-09-21 20:01:21 +00:00
yuneng
4117d9f786
style(ui): format complexity router files
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 20:01:08 +00:00
yucheng-berri
42519a7680
Merge pull request #42262 from BerriAI/litellm_bedrock_batch_s3_bucket_owner
...
* fix(bedrock): send s3BucketOwner on batch input and output data config
Resolve s3_bucket_owner from litellm_params, then optional_params, then
AWS_S3_BUCKET_OWNER and emit it on both S3 data configs so cross-account
batch buckets pass Bedrock ownership validation. Omitted when unset
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
* refactor(bedrock): build batch output config with explicit returns
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
---------
Co-authored-by: yucheng <yucheng@berri.ai>
Co-authored-by: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 12:59:08 -07:00
kerry
b53f9ad658
fix(fal_ai): validate returned image dimensions
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 19:58:59 +00:00
yassin
9dee1d86e7
fix(edenai): advertise reasoning_effort only for models the price map flags as reasoning
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 19:58:58 +00:00
Joshua Valluru
9002749e29
fix(mcp): preserve Python 3.10 imports and integration test seams
2026-09-21 12:58:06 -07:00
Tin Chi Lo
7f997a420d
chore(ui): resolve routing forecast merge conflict
2026-09-21 12:57:57 -07:00
mateo-berri
596783c257
Merge remote-tracking branch 'origin/main' into litellm_decrease_anys_opus5_r5
2026-09-21 12:57:47 -07:00
mateo-berri
c62c187054
Merge remote-tracking branch 'origin/main' into litellm_decrease_anys_opus5_r5
...
# Conflicts:
# litellm/integrations/otel/model/metadata.py
# litellm/litellm_core_utils/coroutine_checker.py
# litellm/litellm_core_utils/exception_mapping_utils.py
# litellm/litellm_core_utils/llm_response_utils/response_metadata.py
# litellm/llms/bedrock_mantle/responses/transformation.py
# litellm/llms/openai/chat/guardrail_translation/handler.py
# litellm/proxy/common_utils/reset_budget_job.py
# litellm/proxy/db/db_spend_update_writer.py
# litellm/proxy/guardrails/guardrail_hooks/llm_as_a_judge/__init__.py
# litellm/proxy/openai_files_endpoints/storage_backend_service.py
# litellm/proxy/pass_through_endpoints/llm_provider_handlers/vertex_passthrough_logging_handler.py
# litellm/realtime_api/main.py
2026-09-21 12:57:38 -07:00
joshua-berri
3549143bcd
Merge pull request #34919 from BerriAI/litellm_fix_mcp_peek_utf8_boundary
...
fix(mcp): handle split UTF-8 routing previews
2026-09-21 19:56:50 +00:00
yuneng
f70683ae92
fix(ui): satisfy complexity router CI lint budgets
...
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 19:56:30 +00:00
mateo-berri
a8745f24a2
fix(ci): let the install smoke test boot its key-less proxy config
...
The install smoke test starts the proxy on test_config_no_auth.yaml, which has no master key on purpose, and #42019 's boot check now refuses that, so the three installing_litellm_on_python jobs have been red on main since 2026-09-20. Pass the documented local-dev override to the proxy child so the test keeps its no-auth config and the boot check stays as it is
2026-09-21 12:56:21 -07:00