Timothy Lowrimore
586c0a7be5
Update docs/my-website/docs/providers/heroku.md
...
Co-authored-by: Claire Riley <claire.riley@salesforce.com>
2025-08-12 08:45:47 -06:00
Timothy Lowrimore
7b0ecb1847
Update docs/my-website/docs/providers/heroku.md
...
Co-authored-by: Claire Riley <claire.riley@salesforce.com>
2025-08-12 08:45:36 -06:00
Timothy Lowrimore
d4e320a4ce
Update docs/my-website/docs/providers/heroku.md
...
Co-authored-by: Claire Riley <claire.riley@salesforce.com>
2025-08-12 08:45:21 -06:00
Timothy Lowrimore
5857d17fad
Update docs/my-website/docs/providers/heroku.md
...
Co-authored-by: Claire Riley <claire.riley@salesforce.com>
2025-08-12 08:45:06 -06:00
Timothy Lowrimore
3cfb85228f
Update docs/my-website/docs/providers/heroku.md
...
Co-authored-by: Claire Riley <claire.riley@salesforce.com>
2025-08-12 08:44:52 -06:00
Timothy Lowrimore
417b18a727
Update docs/my-website/docs/providers/heroku.md
...
Co-authored-by: Claire Riley <claire.riley@salesforce.com>
2025-08-12 08:42:46 -06:00
Timothy Lowrimore
fa587c2eba
Update docs/my-website/docs/providers/heroku.md
...
Co-authored-by: Claire Riley <claire.riley@salesforce.com>
2025-08-12 08:42:27 -06:00
Timothy Lowrimore
1be33ec47f
Update docs/my-website/docs/providers/heroku.md
...
Co-authored-by: Claire Riley <claire.riley@salesforce.com>
2025-08-12 08:41:38 -06:00
Timothy Lowrimore
00c9f1fe0e
Update docs/my-website/docs/providers/heroku.md
...
Co-authored-by: Claire Riley <claire.riley@salesforce.com>
2025-08-12 08:41:22 -06:00
Timothy Lowrimore
e97fa833b3
Update docs/my-website/docs/providers/heroku.md
...
Co-authored-by: Claire Riley <claire.riley@salesforce.com>
2025-08-12 08:41:08 -06:00
Timothy Lowrimore
b20b28a912
Update docs/my-website/docs/providers/heroku.md
...
Co-authored-by: Claire Riley <claire.riley@salesforce.com>
2025-08-12 08:40:54 -06:00
Timothy Lowrimore
e2165ff76e
Update docs/my-website/docs/providers/heroku.md
...
Co-authored-by: Claire Riley <claire.riley@salesforce.com>
2025-08-12 08:40:43 -06:00
Timothy Lowrimore
cc3ec33fe9
removes unused imports that are causing linting failures
2025-08-11 09:52:20 -06:00
Timothy Lowrimore
95d9e30448
Merge branch 'main' into heroku-llms
2025-08-11 09:45:52 -06:00
Timothy Lowrimore
51b52534fb
fixes misconfigured OCI models by setting supports_tool_choice: true
2025-08-11 09:40:40 -06:00
Ishaan Jaff
1cd827874f
[Bug Fix] - Allow using reasoning_effort for gpt-5 model family and reasoning for Responses API ( #13475 )
...
* test_openai_gpt5_reasoning
* test_openai_gpt5_reasoning_effort_parameter
* add OpenAIGPT5ResponsesAPIConfig
* test_openai_gpt5_reasoning_effort_parameter
* fixes
2025-08-10 09:55:36 -07:00
Krrish Dholakia
bd8a0ae0d0
docs: fix order
2025-08-10 09:42:55 -07:00
Krrish Dholakia
1dbac75675
docs(index.md): update release with deployment information
2025-08-10 09:31:28 -07:00
Krish Dholakia
0aeb4f1653
fix(health_check_helpers.py): set max tokens for wildcard call to 10, fixes calling gpt-5-nano via wildcard on openai ( #13482 )
...
gpt-5-nano raises errors for max_tokens=1
2025-08-10 09:23:36 -07:00
Krish Dholakia
184687157e
Litellm model cost map fixes ( #13480 )
...
* build(model_prices_and_context_window.json): fix max token values
* build(model_prices_and_context_window.json): fix max token values
* build(model_prices_and_context_window.json): fix azure gpt-5-chat pricing
2025-08-10 07:38:35 -07:00
Krish Dholakia
c742c76288
Litellm release notes 08 10 2025 ( #13479 )
...
* docs(index.md): initial doc
* build(index.md): initial notes
* docs(index.md): add llm translation tickets
* docs(index.md): document new model support
* docs(index.md): document all pricing changes
* docs(index.md): add llm api endpoints
* docs(index.md): add doc on mcp gateway
* docs(index.md): add all remaining rc notes
* docs(index.md): cleanup
2025-08-10 07:32:11 -07:00
Krrish Dholakia
ece2c9c65d
bump: version 1.75.4 → 1.75.5
2025-08-09 16:31:51 -07:00
Krrish Dholakia
0eedf7c447
build: update local model cost map
2025-08-09 16:31:41 -07:00
Krish Dholakia
9f6f96d76c
Litellm dev 08 07 2025 p1 ( #13418 )
...
* fix(router.py): support base model for model group usage
allows model group info to show accurate cost information for azure models
* fix(router.py): fix changes
* test: add unit tests
* build(pyproject.toml): bump openai version requirements
support custom tool from responses api
Closes https://github.com/BerriAI/litellm/issues/13391
* docs(responses_api.md): add verbosity + free-form function calling parameters
* docs(responses_api.md): add cfg + minimal reasoning to docs
Closes https://github.com/BerriAI/litellm/issues/13391
* docs(responses_api.md): add proxy examples to docs
* refactor: fix ruff error
2025-08-09 16:30:04 -07:00
Sannan Nasir
0e53b1feab
Add digitalocean provider ( #12169 )
...
* Add digitalocean provider
* Add digitalocean provider
* Revert "Add digitalocean provider"
This reverts commit 96dda40f45 .
* changes
* fixes
* Update transformation
* refactoring
* rename provider to Gradient AI
* fixes
* Incorporte review comments
* revert changes
* fix typo
* revert change
* incorporated review comments
* Revert "Incorporte review comments"
This reverts commit 37bd51bd54 .
* changes
* Revert "Revert "Incorporte review comments"
This reverts commit 37bd51bd54ef4fd52ccc12866e47f8de9476d597."
This reverts commit 68c8a198ee .
* changes
* fixes
* Update provider_specific_fields.tsx
2025-08-09 16:26:33 -07:00
Ishaan Jaff
f60a9cf908
[Bug]: Fix JWTs access not working with model groups ( #13474 )
...
* fix can_team_access_model
* test_find_team_with_model_access_model_group
2025-08-09 16:14:51 -07:00
Jugal D. Bhatt
95fbe59c46
Add local storage auth ( #13473 )
2025-08-09 16:13:56 -07:00
Jugal D. Bhatt
67833590d6
[Proxy changes] Litellm add model price reload schedule for multi-pod ( #13470 )
...
* added mcp guardrails doc in mcp.md
* add button to reload models
* Added button changes
* added button for scheduling reload
* add multi pod support to reloading the model price json
* fix ruff
2025-08-09 16:12:13 -07:00
Krish Dholakia
1c8761111f
Router - reduce p99 latency w/ redis enabled by 50% + OTEL - track pre_call hook latency ( #13362 )
...
* feat(proxy/utils.py): track pre-call hooks in OTEL
some pre call hooks can cause latency in high traffic - make sure this is tracked
* fix(router.py): move redis call on deployment_callback_on_success to pipeline operation
reduces p99 latency by half when redis is enabled
* fix(parallel_request_limiter_v3.py): only run check if any item has rate limits set
Prevents unnecessary latency added by rate limit checks
* test: add unit tests
* Latency Improvements: only track tpm/rpm usage when set on deployment+ LLM Caching - use an in-memory cache to reduce redis calls + OTEL - track time spent on LLM caching (#13472 )
* fix(router.py): only track usage for deployments with tpm/rpm set
ensures additional latency avoided for non-tpm/rpm models
* fix(caching_handler.py): log time spent on request get cache to OTEL
enables easy debugging of call latency
* fix(caching_handler.py): use dual cache object for in-memory caching + trace redis call within caching handler
* fix(caching_handler.py): working in-memory cache for redis calls
ensures dual cache works when redis cache setup for llm calls
makes calls quicker by only checking redis when in-memory cache missed for llm api call
* test: remove redundant test
* test: add unit tests
2025-08-09 16:09:51 -07:00
Ishaan Jaff
60306d34a0
[Bug Fix] Allow using Swagger for /chat/completions ( #13469 )
...
* fix get_openapi_schema
* fixes for ProxyChatCompletionRequest
* TestSwaggerChatCompletions
* fix working request body
* fix - add "messages"
* fix messages
* TestSwaggerChatCompletions
* test_messages_field_has_example
* ruff check fix
2025-08-09 15:35:45 -07:00
Jugal D. Bhatt
1270df08a4
[Proxy + UI] Litellm add reload model api and button ( #13464 )
...
* added mcp guardrails doc in mcp.md
* add button to reload models
* Added button changes
* remove the model_reload
2025-08-09 13:52:56 -07:00
Jugal D. Bhatt
10a1fe21c5
[LLM Translation] Litellm azure o series drop params ( #13353 )
...
* added route check
* fix ruff
* Added support for dropping o_series params
* Added ruff fix
* fix tests
2025-08-09 13:52:45 -07:00
Ishaan Jaff
6184e898b7
Generate unique IDs for litellm_call_id and function_id using UUID ( #13468 )
...
Co-authored-by: Cursor Agent <cursoragent@cursor.com>
Co-authored-by: ishaan <ishaan@berri.ai>
2025-08-09 12:59:09 -07:00
Ishaan Jaff
eb4bd26f24
[Bug Fix] - Get Routes ( #13466 )
...
* fixes get_routes_for_mounted_app
* fix - use _safe_get_endpoint_name
* fix code QA check
* test_get_routes_for_mounted_app_with_static_files
* test fixes
2025-08-09 12:52:23 -07:00
Ishaan Jaff
825ea65b96
[Bug Fix] Responses API - Responses API failed if input containing ResponseReasoningItem ( #13465 )
...
* add test_responses_api_multi_turn_with_reasoning_and_structured_output
* fix transform_responses_api_request
2025-08-09 11:20:34 -07:00
Ishaan Jaff
ee40db7b31
docs native litellm prompts
2025-08-09 09:46:31 -07:00
Ishaan Jaff
94c33200a4
docs - native prompt mgmt ( #13463 )
2025-08-09 09:39:16 -07:00
Ishaan Jaff
3999e65a97
docs update
2025-08-09 09:24:41 -07:00
Cole McIntosh
d874bec480
feat(models): add OpenRouter and Cerebras GPT-OSS models (20b, 120b) with pricing and context windows; update backup; refs #13428 ( #13442 )
2025-08-08 22:47:51 -07:00
Jugal D. Bhatt
035e5497e0
added mcp guardrails doc in mcp.md ( #13452 )
2025-08-08 22:47:31 -07:00
Ishaan Jaff
3905cee579
test fixes
2025-08-08 18:50:09 -07:00
Ishaan Jaff
05b48eba62
fix security issue
2025-08-08 18:32:50 -07:00
Ishaan Jaff
32db7f1508
bump: version 1.75.3 → 1.75.4
2025-08-08 18:30:27 -07:00
Ishaan Jaff
edc38b73f9
UI new build
2025-08-08 18:30:15 -07:00
Ishaan Jaff
a843e876a8
[Feat] Working e2e flow for Responses API session management with media ( #13456 )
...
* add MultimodalContent on chat UI
* add multi modal img on chat ui
* utils for responses API imgs
* add code snippet with imgs
* chat UI add imgs
* add imge upload
* chat ui allow adding images
* fix chat send button
* fix button styles
* fix clear chat
* fixes session management
* fixes for session management
* QA fix _should_check_cold_storage_for_full_payload
* test_should_check_cold_storage_for_full_payload
2025-08-08 18:28:10 -07:00
Cole McIntosh
1d514cc68b
feat(reasoning): support 'minimal' effort type for OpenAI ( #13447 )
...
* feat(reasoning): support 'minimal' effort type for OpenAI
* fix(reasoning): correctly map 'minimal' effort to Reasoning object
* chore(dependencies): update OpenAI package version to 1.99.5 in pyproject.toml and requirements.txt
* chore(dependencies): update poetry.lock for OpenAI package version 1.99.5 and Poetry version 2.1.3
2025-08-08 17:56:23 -07:00
tanjiro
4571002e19
disable logging settings for non-enterprise users ( #13431 )
2025-08-08 17:35:29 -07:00
Ishaan Jaff
840db3fe48
LiteLLM UI - Test Key Page - allow uploading images for /chat/completions and /responses ( #13445 )
...
* add MultimodalContent on chat UI
* add multi modal img on chat ui
* utils for responses API imgs
* add code snippet with imgs
* chat UI add imgs
* add imge upload
* chat ui allow adding images
* fix chat send button
* fix button styles
* fix clear chat
2025-08-08 16:57:08 -07:00
Ishaan Jaff
7e2a00c848
[Docs] Add docs on how router / cooldowns work ( #13444 )
...
* add theme-mermaid
* docs cool down
* docs cooldown
2025-08-08 15:13:37 -07:00
Ishaan Jaff
d0aa12f3bf
Enhance team member permission error message with guidance for key creation ( #13443 )
...
Co-authored-by: Cursor Agent <cursoragent@cursor.com>
Co-authored-by: ishaan <ishaan@berri.ai>
2025-08-08 14:35:06 -07:00