Commit graph

22419 commits

Author SHA1 Message Date
Ishaan Jaff
eb02cf1a2d
Revert "Nebius model pricing info updted (#11445)" (#11493)
This reverts commit 32281de91f.
2025-06-06 11:04:21 -07:00
Fadil Rahman
fb5f2c5441
Add batch polling to python code in batches docs (#11286) 2025-06-06 10:49:23 -07:00
AyrennC
900c6e5d7b
[Docs] Add audio / tts section for gemini and vertex (#11306)
* added audio and tts doc for gemini

* updated gemini and vertex audio gen doc to be more concise
2025-06-06 10:48:50 -07:00
Krrish Dholakia
05125a9691 fix: fix linting 2025-06-06 10:43:57 -07:00
Akim Tsvigun
32281de91f
Nebius model pricing info updted (#11445) 2025-06-06 10:43:04 -07:00
Ishaan Jaff
2aa75e1403
add codex-mini-latest (#11492) 2025-06-06 10:39:09 -07:00
Ishaan Jaff
fdaad51015
Feat: add add azure endpoint for image endpoints (#11482)
* feat: add add azure endpoint for image endpoints

* test: azure image routes working as expected

* test azure routes
2025-06-06 10:38:37 -07:00
Krrish Dholakia
1af0b62578 refactor: cleanup huggingface rerank transformation 2025-06-06 10:30:44 -07:00
Krrish Dholakia
b70574017c docs: document new env vars 2025-06-06 09:59:13 -07:00
Peter Dave Hello
b452f82045
Add Google Gemini 2.5 Pro Preview 06-05 (#11447) 2025-06-06 09:28:53 -07:00
Krrish Dholakia
e5f228abd5 fix(utils.py): handle litellm proxy case for checking model info 2025-06-06 09:24:41 -07:00
Krrish Dholakia
e4d1d88d15 fix: remove redundant f-string 2025-06-06 09:18:20 -07:00
Krrish Dholakia
e2da29c54d test: update test 2025-06-06 09:15:08 -07:00
Krrish Dholakia
3608db5ffe fix(prometheus.py): update tests 2025-06-06 09:12:54 -07:00
Cole McIntosh
1b7056f281
fix(vertex_and_google_ai_studio_gemini.py): remove redundant initialization of url_context_metadata, linting error (#11486) 2025-06-06 09:01:21 -07:00
Krrish Dholakia
398fef8391 fix(bedrock/): add generic support for tool calling on bedrock models
Closes https://github.com/BerriAI/litellm/issues/11430
2025-06-05 23:37:02 -07:00
Krish Dholakia
603bd73a17
Gemini - web search cost tracking + Update max output tokens for nova models
* fix(vertex_and_google_ai_studio_gemini.py): add web search request tracking

Enables cost calculation for google web search

* fix(vertex_and_gemini): use common processing logic across stream / non-stream calls

* fix(vertex_And_google_ai_studio_Gemini.py): fix initial choice

* fix: fix linting error

* fix: add initial support for google search cost tracking

* fix(tool_call_cost_tracking.py): working tool cost tracking for gemini

* fix(vertex_ai/gemini/cost_calculator.py): add google web search tool cost tracking for vertex ai

Closes LIT-210

* fix: fix check

* build(model_prices_and_context_window.json): fix amazon nova max output tokens

Closes https://github.com/BerriAI/litellm/issues/11441

* fix: fix ruff check
2025-06-05 23:25:18 -07:00
cainiaoit
be12416863
feat: add HuggingFace rerank provider support (#11438)
++ feat: add HuggingFace rerank provider support

feat: add HuggingFace rerank provider support

feat: add HuggingFace rerank provider support

feat: add HuggingFace rerank provider support

feat: add HuggingFace rerank provider support
2025-06-05 23:23:01 -07:00
Pedro Azevedo
5d516aace1
fix: supports_function_calling works with llm_proxy models (#11381)
* Add tests for function calling support in LiteLLM proxy models

- Introduced a new test script `test_proxy_function_calling.py` to validate function calling capabilities for both direct and proxied models.
- Created a comprehensive test suite in `tests/litellm_utils_tests/test_proxy_function_calling.py` using pytest, covering various model configurations and edge cases.
- Implemented parameterized tests to ensure consistency between direct and proxied model function calling support.
- Added tests for specific proxy models, edge cases, and import verification for the `supports_function_calling` function.
- Included a demonstration test to highlight the current issue with proxy model resolution.

* feat: add fallback handling for litellm_proxy models in model info retrieval

* feat: enhance proxy function calling tests with custom model name handling and documentation

* fix: add type ignore comments for custom logger callback initialization

* fix: remove styling diff

* fix: style

* fix(utils.py): remove outdated comment regarding litellm_proxy models

* feat(utils.py): add proxy model handling for underlying model extraction

* feat(utils.py): enhance model name handling for litellm_proxy integration

* refactor(utils.py): remove unused _handle_proxy_model_names function
2025-06-05 23:15:33 -07:00
Krrish Dholakia
1ca85161a5 docs(users.md): clarify how budgets are applied 2025-06-05 23:12:24 -07:00
Ishaan Jaff
c99daef689
[Fix]: /v1/messages - return streaming usage statistics when using litellm with bedrock models (#11469)
* fix: using litellm with claude code bedrock

* fix: usage for bedrock with /messages

* fix: bedrock_sse_wrapper

* tests: test for test_chunk_parser_usage_transformation

* test fix
2025-06-05 21:18:19 -07:00
Ishaan Jaff
f0cb80ec50
[Feat] Return response_id == upstream response ID for VertexAI + Google AI studio (Stream+Non stream) (#11456)
* fix: vertexAI return responseID

* fix: vertexAI return responseID

* test_vertex_ai_response_id

* test: test_vertex_ai_streaming_response_id

* test_vertex_ai_streaming_response_id
2025-06-05 20:18:55 -07:00
Ishaan Jaff
23627d6a26
[Fix] [Bug]: Knowledge Base Call returning error (#11467)
* fix:get_and_pop_recognised_vector_store_tools

* test: tools wwith vector stores

* test - bedrock kb tools

* fix: add clear comment

* fix: vector store tools
2025-06-05 18:24:36 -07:00
RMeans
742405f6cf
Add pangea to guardrails sidebar (#11464) 2025-06-05 18:11:52 -07:00
Ishaan Jaff
18ea65218b
[Feat] Make batch size for maximum retention in spend logs a controllable parameter (#11459)
* feat: add SPEND_LOG_CLEANUP_BATCH_SIZE

* docs update

* test: test_cleanup_batch_size_env_var
2025-06-05 17:11:51 -07:00
Krish Dholakia
d05eda0311
Custom Root Path Improvements: don't require reserving /litellm route (#11460)
* fix(proxy_server.py): initial commit with asset prefix rewriting for custom base path

Closes https://github.com/BerriAI/litellm/issues/11451

* docs(litellm_proxy.md): clarify version requirement

* fix(proxy_server.py): replace litellm well known route with custom server root path

Ensures UI calls correct endpoint

* build(ui/): update ui build
2025-06-05 16:36:47 -07:00
Cole McIntosh
a3da7f1876
Add AGENTS.md (#11461) 2025-06-05 16:29:28 -07:00
Sean Walker
29dc4e51f9
Fix HuggingFace embeddings using non-default input_type (#11452)
* fix(huggingface): use get() instead of pop() for input_type parameter

Fixes embedding generation for HuggingFace models where input_type override
is required (e.g. BAAI/bge-m3). The pop() method was mutating optional_params
and removing input_type before downstream functions could access it.

* Add unit tests to catch regression

* Move tests around
2025-06-05 15:48:55 -07:00
Krrish Dholakia
ab9d09a464 fix(ui/): fix linting errors 2025-06-05 15:17:30 -07:00
Krrish Dholakia
30f1c5e852 docs: clarify pre-release 2025-06-05 15:00:04 -07:00
Cole McIntosh
7c513856dc
Fix None values in usage field for gpt-image-1 model responses (#11448)
* fix(convert_dict_to_response.py): handle None values in usage field for gpt-image-1

* test: add tests for handling None and partial values in usage fields for gpt-image-1 responses
2025-06-05 13:19:18 -07:00
Krish Dholakia
69c9d75f20
fix(prometheus.py): pass custom metadata labels in litellm_total_toke… (#11414)
* fix(prometheus.py): pass custom metadata labels in litellm_total_tokens metric

* fix(handler.py): handle /v1 for openai realtime translation

Closes https://github.com/BerriAI/litellm/pull/11398

* fix(prometheus.py): fix incrementing total tokens metric
2025-06-05 00:15:23 -07:00
Low Jian Sheng
a3e5bc4856
Support no reasoning option for gemini models (#11393)
* support no reasoning for gemini models

* change none to disable

* remove print statements

* update docs
2025-06-05 00:11:45 -07:00
Krrish Dholakia
505d2fe0c7 build: bump 2025-06-05 00:08:53 -07:00
Krish Dholakia
db23016536
fix(redis_cache.py): support pipeline redis lpop for older redis vers… (#11425)
* fix(redis_cache.py): support pipeline redis lpop for older redis versions

Fixes https://github.com/BerriAI/litellm/issues/10379

* test: add mock host
2025-06-05 00:05:54 -07:00
Sha
a301ef873e
added gemini url context support (#11351)
* added gemini url context support

* lint issue fix
2025-06-04 23:56:21 -07:00
Tom Bocklisch
d7982bb0af
Use proper attribute for sagemaker request (#11362) 2025-06-04 22:47:43 -07:00
Jimmy Tsai
4019f79808
feat: add deepseek-r1 family model configuration to pricing JSON (#11394) 2025-06-04 22:39:06 -07:00
Ishaan Jaff
02a34d319a
bump to ddtrace==3.8.0 (#11426) 2025-06-04 22:18:07 -07:00
Ishaan Jaff
f0e0007eaf fix: gemini-2.0-flash-preview-image-generation test 2025-06-04 21:21:28 -07:00
Cole McIntosh
049e65a84e
Merge pull request #11417 from colesmcintosh/sso-config-ui 2025-06-04 20:51:59 -06:00
Ishaan Jaff
de306cfcb3
[Performance] Performance improvements for /v1/messages route (#11421)
* fix: perf anthropic /v1/messages

* fix: perf anthropic /v1/messages

* fix: linting checks

* fix: linting checks
2025-06-04 18:47:53 -07:00
raz-alon
fada9c79be
Add User ID validation to ensure it is not an email or phone number (#10102) 2025-06-04 18:38:02 -07:00
Cole McIntosh
c1a324c2fb Merge remote-tracking branch 'origin/main' into sso-config-ui 2025-06-04 18:42:43 -06:00
Cole McIntosh
65b28826a6 Add uiAuditLogsCall function 2025-06-04 18:19:53 -06:00
Krish Dholakia
9da32d9e14
Litellm audit log staging (#11418)
* Audit logs added (#11226)

* audit logs added

* audit logs populated

* adding json response

* collapsible json columns

* add created at column

* added changed field

* added premiumUser description

* added paginated filtered logs

* convert table names

* remove test file

* added new ui for audit logs

* only show the difference in before value and updated value

* fix: add lucide-react to package json

---------

Co-authored-by: tanjiro <56165694+NANDINI-star@users.noreply.github.com>
2025-06-04 14:34:17 -07:00
Cole McIntosh
3a946933ee Refactor settings response models in proxy_setting_endpoints.py
- Renamed SSOSettingsResponse to inherit from a new base class SettingsResponse for better structure.
- Introduced InternalUserSettingsResponse and DefaultTeamSettingsResponse models for internal user and default team settings.
- Updated endpoint responses to use field_schema instead of schema for consistency.
- Enhanced test cases to validate the new response structure and ensure proper functionality of SSO settings.
2025-06-04 15:08:05 -06:00
Cole McIntosh
3cc9460922 Add SSO settings response model in proxy_setting_endpoints.py
- Introduced SSOSettingsResponse model to encapsulate SSO configuration values and schema information.
- Updated the get_sso_settings endpoint to utilize the new response model, enhancing API clarity and usability.
2025-06-04 14:55:29 -06:00
Cole McIntosh
c41b14e27b Add clear SSO settings functionality in SSOModals component
- Introduced a confirmation modal for clearing SSO settings.
- Implemented handleClearSSO function to reset SSO settings and provide user feedback.
- Updated UI to include a 'Clear' button for SSO settings, enhancing user experience.
- Added state management for the confirmation modal visibility.
2025-06-04 14:39:26 -06:00
Cole McIntosh
72c7fd63bf Implement SSO configuration check in AdminPanel and update SSOModals to reflect SSO status
- Added logic to check SSO configuration and set state in AdminPanel.
- Introduced a new function to handle SSO configuration checks.
- Updated UI to conditionally render SSO button text based on configuration status.
- Passed SSO configuration status as a prop to SSOModals for better integration.
2025-06-04 14:34:18 -06:00