Commit graph

31034 commits

Author SHA1 Message Date
Sameer Kankute
bfd410704b Fix anthropic.claude-opus-4-6-v1 for bedrock 2026-02-11 13:15:21 +05:30
Sameer Kankute
f2d1c0eb22 Add adaptive thinking support for anthropic opus 4.6 2026-02-11 13:14:33 +05:30
Cesar Garcia
a0707bcb03 [Feat] add ElevenLabs eleven_v3 and eleven_multilingual_v2 to model cost map (#20522)
* [Feat] add ElevenLabs `eleven_v3` and `eleven_multilingual_v2` to model cost map

Register ElevenLabs TTS models for cost tracking:
- elevenlabs/eleven_v3: most expressive model, 70+ languages, audio tags
- elevenlabs/eleven_multilingual_v2: default TTS model, 29 languages

Also update ElevenLabs docs with supported models table and eleven_v3 audio tags example.

* docs: remove model-agnostic tip from ElevenLabs docs
2026-02-11 13:14:33 +05:30
Cesar Garcia
0baaf9125c feat(web_search): add gpt-5-search-api model and docs clarifications (#20512)
* docs(web_search): clarify OpenAI search model requirements

- Add gpt-5-search-api to supported OpenAI search models
- Add warning that regular models (gpt-5, gpt-4.1) do NOT support web_search_options
- Add tip that web_search_options is optional for search models

* feat(models): add gpt-5-search-api pricing for OpenAI and Azure
2026-02-11 13:14:33 +05:30
Swayambhu
af41713aa1 refactor: Add error handling for network calls and apply consistent formatting across networking functions. 2026-02-11 13:14:30 +05:30
Peter Dave Hello
ee634375e2 Align Claude Opus 4.6 Bedrock metadata and model IDs
Unify follow-up fixes for Opus 4.6 pricing and routing metadata into
a single changeset.

Set long-context-capable Opus 4.6 entries to 1M input tokens where
>200K pricing is defined, align alias and dated capability metadata,
and add Bedrock Converse v1 IDs with and without :0 suffixes.

Keep regional endpoint pricing at a 10% premium over global entries
and mirror all cost-map changes in the backup file used for local
loading and offline fallback behavior.

Extend Opus 4.6 regression tests to verify metadata parity, Bedrock
regional pricing parity across :0 and non-:0 IDs, and converse model
registration in constants and runtime model sets.
2026-02-11 13:13:50 +05:30
Swayambhu
bb739de689 fix: Add array type checks for model, agent, and MCP hub data to prevent crashes from non-array API responses and include a regression test.
This PR:Fixes a frontend regression where the
PublicModelHub
 page would crash with TypeError: e.filter is not a function when the API returned an error object (e.g. { "detail": "..." }) instead of the expected data array.
Changes:

Added defensive Array.isArray() checks in
src/components/public_model_hub.tsx
 for:
modelHubData
agentHubData
mcpHubData
Updated useMemo hooks and helper functions to handle invalid data gracefully.
Added a regression test in
src/components/public_model_hub.test.tsx
 that mocks a non-array API response to ensure the component renders without crashing.
2026-02-11 13:12:52 +05:30
yuneng-jiang
9da55d7c4e small fixes 2026-02-11 13:12:28 +05:30
yuneng-jiang
2bbd88703d refactor admin page 2026-02-11 13:11:45 +05:30
Ishaan Jaffer
61ed8f9e03 fix MYPY lint 2026-02-07 15:55:54 -08:00
Ishaan Jaffer
fac094da7e Revert "Merge pull request #18790 from BerriAI/litellm_key_team_routing_3"
This reverts commit ae26d8e68a, reversing
changes made to 864e8c6543.
2026-02-07 15:51:33 -08:00
michelligabriele
3a10de7d31 fix: revert httpx client caching that caused closed client errors
AsyncHTTPHandler.__del__ was closing httpx clients still in use by
AsyncOpenAI/AsyncAzureOpenAI due to independent cache lifecycles.
Restores standalone httpx client creation for OpenAI/Azure providers.
2026-02-07 15:45:46 -08:00
Alexsander Hamir
461fff1a87 perf(prometheus): parallelize budget metrics, fix caching bug, reduce CPU by ~40% (#20544) 2026-02-07 15:43:35 -08:00
Ishaan Jaff
c0fb179a32 [Feat] - Search API add /list endpoint to list what search tools exist in router (#19969)
* feat: List all available search tools configured in the router.

* add debugging search API

* add debugging search API
2026-01-28 17:59:04 -08:00
Ishaan Jaff
4d4acbecde [Fix] VertexAI Pass through - fix regression that caused vertex ai passthroughs to stop working for router models (#19967)
* fix(vertex_ai): replace custom model names with actual Vertex AI model names in passthrough URLs (#19948)

When the passthrough URL already contains project and location, the code
was skipping the deployment lookup and forwarding the URL as-is to Vertex AI.
For custom model names like gcp/google/gemini-2.5-flash, Vertex AI returned
404 because it only knows the actual model name (gemini-2.5-flash).

The fix makes the deployment lookup always run, so the custom model name
gets replaced with the actual Vertex AI model name before forwarding.

* add _resolve_vertex_model_from_router

* fix: get_llm_provider

* Potential fix for code scanning alert no. 4020: Clear-text logging of sensitive information

Co-authored-by: Copilot Autofix powered by AI <62310815+github-advanced-security[bot]@users.noreply.github.com>

---------

Co-authored-by: michelligabriele <gabriele.michelli@icloud.com>
Co-authored-by: Copilot Autofix powered by AI <62310815+github-advanced-security[bot]@users.noreply.github.com>
2026-01-28 16:55:03 -08:00
michelligabriele
0003e647ea fix(vertex_ai): support model names with slashes in passthrough URLs (#19944)
The regex in get_vertex_model_id_from_url() was using [^/:]+
which stopped at the first slash, truncating model names like
'gcp/google/gemini-2.5-flash' to just 'gcp'. This caused
access_groups checks to fail for custom model names.

Changed the pattern to [^:]+ to allow slashes in model names,
only stopping at the colon before the action (e.g., :generateContent).
2026-01-28 09:34:21 -08:00
Jay Prajapati
43642391c5 fix(proxy): support slashes in google generateContent model names (#19737)
* fix(proxy): support slashes in google route params

* fix(proxy): extract google model ids with slashes

* test(proxy): cover google model ids with slashes
2026-01-27 19:33:19 -08:00
Harshit Jain
51d744c9d0 feat: tpm-rpm limit in prometheus metrics (#19725)
Co-authored-by: Krish Dholakia <krrishdholakia@gmail.com>
2026-01-27 19:32:47 -08:00
Harshit Jain
5426b3c940 fix: server rooth path (#19790) 2026-01-26 09:49:41 -08:00
Harshit Jain
4c4d4f6d42
Resolved merge conflict in proxy_config.yaml, changes from main into litellm_rc_branch 2026-01-25 10:58:36 +05:30
Ishaan Jaffer
e41b9c29a8 fix MCP tests 2026-01-24 18:00:22 -08:00
Ishaan Jaffer
f148207b11 fix patch reliability mock tests 2026-01-24 17:37:17 -08:00
Ishaan Jaffer
73dd1bd97a test_stream_transformation_error_sync 2026-01-24 17:19:21 -08:00
yuneng-jiang
06f7a7ec69
Merge pull request #19716 from BerriAI/ui_build_yj_0124_2
[Infra] Rebuilding UI
2026-01-24 16:28:22 -08:00
yuneng-jiang
3c94458f35 chore: update Next.js build artifacts (2026-01-25 00:27 UTC, node v22.16.0) 2026-01-24 16:27:30 -08:00
yuneng-jiang
8094aff8c5
Merge pull request #19715 from BerriAI/key_teams_fallback_docs
[Docs] UI Keys Teams Router Settings docs
2026-01-24 16:26:28 -08:00
yuneng-jiang
937ccf1977 UI Keys Teams Router Settings docs 2026-01-24 16:23:46 -08:00
Harshit Jain
05fdd099ba
fix(presidio): resolve runtime error by handling asyncio loops in bac… (#19714)
* fix(presidio): resolve runtime error by handling asyncio loops in background threads

* add test case for thread safety
2026-01-24 15:36:49 -08:00
Ishaan Jaffer
53d3868ff2 TestBedrockInvokeToolSearch 2026-01-24 15:36:30 -08:00
yuneng-jiang
99e9462ec9
Merge pull request #19713 from BerriAI/litellm_model_search_id_team
[Feature] UI - Model Page: Filter by Model ID and Team ID
2026-01-24 15:04:20 -08:00
yuneng-jiang
47810f1523 Model and Team filtering 2026-01-24 14:45:14 -08:00
Ishaan Jaffer
9f99e8231c test_retrieve_container_basic 2026-01-24 14:31:54 -08:00
Ishaan Jaffer
f2fd54ffcf test fixes 2026-01-24 14:27:56 -08:00
Ishaan Jaffer
d8e0c43a21 test fixes 2026-01-24 14:08:48 -08:00
Ishaan Jaffer
df3c54ec6a BUMP extras 2026-01-24 14:04:30 -08:00
Ishaan Jaffer
e4e76a4963 test_team_update_sc_2 2026-01-24 13:17:32 -08:00
Ishaan Jaffer
489c986cac test_hanging_request_azure 2026-01-24 13:14:48 -08:00
Ishaan Jaffer
a62ccd582d fix flaky tests 2026-01-24 13:04:20 -08:00
Ishaan Jaffer
31a4cb65bf test_get_default_unvicorn_init_args 2026-01-24 12:59:51 -08:00
Ishaan Jaffer
241c0c6d2a docs fix 2026-01-24 12:32:18 -08:00
Ishaan Jaffer
6c6ed0dad2 docs fix 2026-01-24 12:11:39 -08:00
Ishaan Jaffer
28a9003103 docs fix 2026-01-24 12:10:42 -08:00
Ishaan Jaffer
bbeb007f4e docs fix 2026-01-24 12:09:16 -08:00
Ishaan Jaffer
a081dc2ee2 docs fix 2026-01-24 11:41:16 -08:00
Ishaan Jaffer
a7e26460d0 fix unstable tests 2026-01-24 11:15:16 -08:00
Ishaan Jaffer
bd38374a45 fix: FLAKY tests 2026-01-24 11:13:44 -08:00
Alexsander Hamir
9cdd7a8fd2
Fix: log duplication when json_logs is enabled (#19705) 2026-01-24 11:09:04 -08:00
Ishaan Jaffer
6710abd1fa test_web_search 2026-01-24 11:06:23 -08:00
Ishaan Jaffer
6587cd228b test_partner_models_httpx_streaming 2026-01-24 10:58:35 -08:00
Ishaan Jaffer
1a7274aa4e fix: _apply_search_filter_to_models mypy linting 2026-01-24 09:24:50 -08:00