Commit graph

51909 commits

Author SHA1 Message Date
MHammett
b2ef9d5b64 test(responses): drop sys.path.insert from the response_format tests
main's test-quality gate counts sys.path.insert under TQ003, and main is
already over that rule's ceiling (63 against 62), so any new occurrence
fails lint. The line was copied from the sibling text_format test, which
main has since cleaned up the same way; pytest resolves litellm from the
repo root without it.

Also corrects a stale docstring: the helper raises litellm.BadRequestError
directly, not ValueError.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
2026-09-21 05:12:49 -05:00
MHammett
bb304c8f35 Merge main into fix/responses-response-format-alias
The PR was retargeted from litellm_internal_staging to main. test-linting.yml checks out the PR head and runs .github/actions/detect-changes, which was added to main after this branch was cut, so the lint job could not start. Bringing main in resolves it; no change to this PR's code.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
2026-09-21 04:59:39 -05:00
MHammett
c7daa77fc2 style(responses): build text.format without mutable-collection literals
litellm main tightened the LIT002 (mutable-collection construction)
ceiling after this branch was cut and is currently over it, so the
type-discipline gate now rejects any net-new dict literal. Build the two
ResponseText values as TypedDict-annotated literals and freeze the
intermediate mappings with MappingProxyType -- the forms the checker
documents as one-shot builds rather than seed-then-mutate accumulators.

No behaviour change; the tests are untouched.

Co-Authored-By: Claude Opus 5 <noreply@anthropic.com>
2026-09-21 04:36:21 -05:00
Mateo Wang
1cac8bd9ab
Merge pull request #42193 from BerriAI/litellm_responses_bridge_safety_identifier
fix(responses): forward safety_identifier through the chat completion bridge
2026-09-21 01:36:26 -07:00
Devin AI
b38504b5a6 test(e2e): assert the forwarded Converse body without requiring the model to accept safety_identifier
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 07:37:28 +00:00
Devin AI
0317903a44 ci(e2e-changed): surface failed test ids from the pytest log
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 07:16:02 +00:00
Devin AI
7e0fa40fe3 test(e2e): gate the Bedrock edge capture behind a provider_edge_host opt-in
The Buildkite ephemeral stack runs the gateway in another pod, so it cannot reach the pytest host's provider edge. The GitHub changed-e2e lane runs gateways on the runner and sets E2E_PROVIDER_EDGE_HOST_REACHABLE

Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 06:17:03 +00:00
Devin AI
ba93c7402a test(e2e): tolerate provider retries in safety_identifier capture assertion
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 04:53:10 +00:00
Devin AI
83223885e6 fix(responses): forward safety_identifier through the chat completion bridge
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-21 04:09:59 +00:00
kerry-berri
e484a7c89c
Merge pull request #42192 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 2 models
2026-09-20 20:39:41 -07:00
berriai-litellm-provider-info-sync[bot]
d5921713a8
chore(prices): sync OpenRouter prices: 2 models
openrouter/~moonshotai/kimi-latest: input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
openrouter/moonshotai/kimi-k3: input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
2026-09-21 03:30:42 +00:00
kerry-berri
0ece1cd426
Merge pull request #42187 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 2 models
2026-09-20 20:09:22 -07:00
berriai-litellm-provider-info-sync[bot]
4258bd366c
chore(prices): sync OpenRouter prices: 2 models
openrouter/deepseek/deepseek-v4-flash: input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
openrouter/deepseek/deepseek-v4-pro: input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
2026-09-21 03:00:39 +00:00
kerry-berri
946da34260
Merge pull request #42184 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 2 models
2026-09-20 19:38:29 -07:00
berriai-litellm-provider-info-sync[bot]
2bacd6caa0
chore(prices): sync OpenRouter prices: 2 models
openrouter/~moonshotai/kimi-latest: input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
openrouter/moonshotai/kimi-k3: input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
2026-09-21 02:30:52 +00:00
yujonglee
211f4a6851
Merge pull request #42173 from BerriAI/litellm_scaffold_secrets
feat(rust): add typed secret managers and shared auth adapters
2026-09-20 19:15:36 -07:00
Yujong Lee
e3f69fc4d8 refactor(rust): define consistent secret lookup contracts 2026-09-20 19:08:25 -07:00
kerry-berri
9456f48427
Merge pull request #42182 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 1 model
2026-09-20 18:40:00 -07:00
berriai-litellm-provider-info-sync[bot]
acd303043b
chore(prices): sync OpenRouter prices: 1 model
openrouter/qwen/qwen3.8-27b: output_cost_per_token, cache_read_input_token_cost
2026-09-21 01:31:07 +00:00
kerry-berri
6f37808d44
Merge pull request #42179 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 3 models
2026-09-20 17:39:05 -07:00
berriai-litellm-provider-info-sync[bot]
a3ffc395b9
chore(prices): sync OpenRouter prices: 3 models
openrouter/ibm-granite/granite-4.2-8b: input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
openrouter/meta-llama/llama-3.1-70b-instruct: max_tokens, max_output_tokens, input_cost_per_token, output_cost_per_token
openrouter/meta-llama/llama-4-maverick: input_cost_per_token, output_cost_per_token
2026-09-21 00:31:06 +00:00
kerry-berri
5e58b76f4e
Merge pull request #42178 from BerriAI/litellm-providers/price-sync-openrouter
Some checks are pending
Unit Tests: Documentation Validation / documentation (push) Waiting to run
Unit Tests: Proxy DB Operations / assert-shard-coverage (push) Waiting to run
Unit Tests: Proxy DB Operations / auth-checks (push) Blocked by required conditions
Unit Tests: Proxy DB Operations / budgets (push) Blocked by required conditions
Unit Tests: Proxy DB Operations / custom-logging (push) Blocked by required conditions
Unit Tests: Proxy DB Operations / db-and-spend (push) Blocked by required conditions
Unit Tests: Proxy DB Operations / endpoints-and-responses (push) Blocked by required conditions
Unit Tests: Proxy DB Operations / proxy-server-core (push) Blocked by required conditions
Unit Tests: Proxy DB Operations / guardrails-hooks (push) Blocked by required conditions
Unit Tests: Proxy DB Operations / jwt-and-keys (push) Blocked by required conditions
Unit Tests: Proxy DB Operations / key-generation (push) Blocked by required conditions
Unit Tests: Proxy DB Operations / logging-misc (push) Blocked by required conditions
Unit Tests: Proxy DB Operations / proxy-runtime (push) Blocked by required conditions
Unit Tests: Proxy DB Operations / proxy-utils (push) Blocked by required conditions
Unit Tests / caching-local (push) Waiting to run
Unit Tests / core-utils (push) Waiting to run
Unit Tests / enterprise-package (push) Waiting to run
Unit Tests / enterprise-routing (push) Waiting to run
Unit Tests / integrations (push) Waiting to run
Unit Tests / All Other Providers (push) Waiting to run
Unit Tests / Vertex AI (push) Waiting to run
Unit Tests / mcp-integration (push) Waiting to run
Unit Tests / misc (push) Waiting to run
Unit Tests / proxy-auth (push) Waiting to run
Unit Tests / proxy-endpoints (push) Waiting to run
Unit Tests / proxy-extras (push) Waiting to run
Unit Tests / proxy-infra (push) Waiting to run
Unit Tests / proxy-server (push) Waiting to run
Unit Tests / responses-caching-types (push) Waiting to run
GitHub Actions Security Analysis / zizmor (push) Waiting to run
chore(prices): sync OpenRouter prices: 4 models
2026-09-20 17:09:14 -07:00
berriai-litellm-provider-info-sync[bot]
7bcb01a40c
chore(prices): sync OpenRouter prices: 4 models
openrouter/~deepseek/deepseek-flash-latest: max_tokens, max_output_tokens
openrouter/~z-ai/glm-flash-latest: max_tokens, max_output_tokens
openrouter/deepseek/deepseek-v4-flash: input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
openrouter/deepseek/deepseek-v4-pro: input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
2026-09-21 00:00:57 +00:00
kerry-berri
9ab609948d
Merge pull request #42175 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 3 models
2026-09-20 16:43:57 -07:00
kerry
5ee75af908 fix(openrouter): keep glm-latest output limit at alias target value
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-20 23:34:40 +00:00
berriai-litellm-provider-info-sync[bot]
15f368060f
chore(prices): sync OpenRouter prices: 3 models
openrouter/~deepseek/deepseek-pro-latest: off_peak_pricing, input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
openrouter/~z-ai/glm-latest: max_tokens, max_output_tokens
openrouter/deepseek/deepseek-v4-pro-0813: off_peak_pricing, input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
2026-09-20 23:30:54 +00:00
Yujong Lee
82bc67b122 feat(rust): add typed secret managers and shared auth adapters 2026-09-20 16:09:20 -07:00
Yujong Lee
3fd2dd635b refactor(rust): split auth facade from shared types 2026-09-20 15:45:38 -07:00
kerry-berri
6814373a48
Merge pull request #42169 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 2 models
2026-09-20 15:39:53 -07:00
berriai-litellm-provider-info-sync[bot]
668a89fbfd
chore(prices): sync OpenRouter prices: 2 models
openrouter/~deepseek/deepseek-pro-latest: max_tokens, max_output_tokens, input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
openrouter/deepseek/deepseek-v4-pro-0813: max_tokens, max_output_tokens, input_cost_per_token, output_cost_per_token, cache_read_input_token_cost, off_peak_pricing
2026-09-20 22:30:41 +00:00
yujonglee
65c6495616
Merge pull request #42165 from BerriAI/refactor-tokenizer-rs
refactor(rust): split token counter backends
2026-09-20 15:22:33 -07:00
Yujong Lee
661da87c91 fix(rust): validate tokenizer ranks and cover backend features 2026-09-20 15:13:59 -07:00
kerry-berri
e7c8928459
Merge pull request #42168 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 3 models
2026-09-20 15:13:08 -07:00
berriai-litellm-provider-info-sync[bot]
10715b1b46
chore(prices): sync OpenRouter prices: 3 models
openrouter/~deepseek/deepseek-pro-latest: max_tokens, max_output_tokens, input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
openrouter/~z-ai/glm-latest: max_tokens, max_output_tokens
openrouter/deepseek/deepseek-v4-pro-0813: max_tokens, max_output_tokens, input_cost_per_token, output_cost_per_token, cache_read_input_token_cost, off_peak_pricing
2026-09-20 22:00:46 +00:00
Yujong Lee
56c5d31e73 test(rust): run fast token counter parity tests by default
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-20 21:52:48 +00:00
Yujong Lee
831810248a refactor(rust): keep the Python token counter on the fast backend
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-20 21:49:39 +00:00
kerry-berri
5483f49a4d
Merge pull request #42166 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 1 model
2026-09-20 14:40:14 -07:00
berriai-litellm-provider-info-sync[bot]
be7b7d2ade
chore(prices): sync OpenRouter prices: 1 model
openrouter/~z-ai/glm-latest: max_tokens, max_output_tokens
2026-09-20 21:30:55 +00:00
Yujong Lee
a619b765fc refactor(rust): simplify tokenizer bridge features
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-20 21:25:19 +00:00
Yujong Lee
9b2b3d0b90 refactor(rust): align tokenizer docs and lint
Co-Authored-By: Devin AI <158243242+devin-ai-integration[bot]@users.noreply.github.com>
2026-09-20 21:21:52 +00:00
Yujong Lee
84c26978a3 refactor(rust): split token counter backends 2026-09-20 14:10:31 -07:00
kerry-berri
3c69ad983a
Merge pull request #42164 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 2 models
2026-09-20 14:08:42 -07:00
berriai-litellm-provider-info-sync[bot]
8a8a15af93
chore(prices): sync OpenRouter prices: 2 models
openrouter/~deepseek/deepseek-v4-flash-latest: output_cost_per_token
openrouter/deepseek/deepseek-v4-flash-0731: output_cost_per_token
2026-09-20 21:00:50 +00:00
kerry-berri
0e61ed51c0
Merge pull request #42163 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 3 models
2026-09-20 13:39:16 -07:00
berriai-litellm-provider-info-sync[bot]
83e1f0834e
chore(prices): sync OpenRouter prices: 3 models
openrouter/~deepseek/deepseek-v4-flash-latest: output_cost_per_token
openrouter/deepseek/deepseek-v4-flash-0731: output_cost_per_token
openrouter/deepseek/deepseek-v4-flash-vision-exp: max_tokens, max_output_tokens, input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
2026-09-20 20:30:54 +00:00
kerry-berri
c20135621f
Merge pull request #42162 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 2 models
2026-09-20 12:39:29 -07:00
joshua-berri
b4d9231b31
Merge pull request #42148 from BerriAI/litellm_mcp_public_client_7740
fix(mcp): explain missing public client dependencies
2026-09-20 19:36:17 +00:00
joshua-berri
c25601fb5f
Merge pull request #42064 from BerriAI/litellm_fix_jwt_deactivated_users_8217
fix(auth): reject deactivated JWT users and refresh cached status
2026-09-20 19:35:53 +00:00
berriai-litellm-provider-info-sync[bot]
9e84b8b048
chore(prices): sync OpenRouter prices: 2 models
openrouter/~deepseek/deepseek-pro-latest: input_cost_per_token, output_cost_per_token, cache_read_input_token_cost
openrouter/deepseek/deepseek-v4-pro-0813: input_cost_per_token, output_cost_per_token, cache_read_input_token_cost, off_peak_pricing
2026-09-20 19:30:56 +00:00
kerry-berri
f8cc79bbcd
Merge pull request #42157 from BerriAI/litellm-providers/price-sync-openrouter
chore(prices): sync OpenRouter prices: 2 models
2026-09-20 12:12:14 -07:00