Commit graph

24591 commits

Author SHA1 Message Date
Krish Dholakia
be30bc68ae
Merge pull request #13759 from kankute-sameer/litellm_feat_correct_cost_calculations
Add long context support for claude-4-sonnet
2025-08-19 22:30:25 -07:00
Krrish Dholakia
7d09375d52 fix: fix gpt-5-chat mappings 2025-08-19 22:21:00 -07:00
Ishaan Jaff
1832c09d6b
[Feat] - UI Allow using Key/Team Based Logging for Langfuse OTEL (#13791)
* add langfuse_otel

* ui allow setting langfuse OTEL

* fixes - using langfuse OTEL

* add description for logging integrations

* add description
2025-08-19 19:06:26 -07:00
Ishaan Jaff
a328ad56e3
[Bug Fix] Fixes for using Auto Router with LiteLLM Docker Image (#13788)
* fix install auto router.sh

* fixes for Docker IMG
2025-08-19 18:36:30 -07:00
Ishaan Jaff
56ac778316
[Bug Fix] Bedrock KB - Using LiteLLM Managed Credentials for Query (#13787)
* fix: add get_credentials_for_vector_store

* test_search_uses_registry_credentials

* test_bedrock_search_with_credentials_managed_registry
2025-08-19 15:39:36 -07:00
mubashir1osmani
ef9e50458d removed faq 2025-08-19 17:01:31 -04:00
tanjiro
b0e3469202
Models page row UI restructure (#13771)
* remove unused model dashboard

* move model_dashboard to templates folder

* moving  columns to molecules directory

* group model names and provider icon

* increase table column name font size. cleanup extra title and description inside the tab

* fix column width and truncate string

* moved credentials column

* combined created by and created at

* combine in and out costs

* remove edit button

* remove 2nd extra tooltip
2025-08-19 14:00:58 -07:00
Philip Kiely
b8167cf304 rip out some old stuff 2025-08-19 13:55:33 -07:00
mubashir1osmani
2edcc51e57 updated claude-code docs and deployment faq 2025-08-19 16:52:46 -04:00
Philip Kiely
6fd1705183 lint 2025-08-19 13:36:41 -07:00
Sameer Kankute
48622a4ee7 add cache above 200k keys in INTENDED_SCHEMA 2025-08-20 01:47:34 +05:30
Philip Kiely
7c3d522435 Update Baseten LiteLLM integration 2025-08-19 12:21:05 -07:00
Sameer Kankute
fa81b6c639
Update test_cost_calculator.py 2025-08-19 23:14:22 +05:30
Ishaan Jaff
195ea6515e
[Feat] Datadog LLM Observability - Add support for tracing guardrail input/output (#13767)
* add guardrail information on DD LLM Obs

* test_guardrail_information_in_metadata
2025-08-19 10:26:25 -07:00
Sameer Kankute
b5f0a7b49b
Merge branch 'main' into litellm_feat_correct_cost_calculations 2025-08-19 18:55:36 +05:30
Sameer Kankute
d39b2e8888 Add test for long context cost calculation 2025-08-19 16:56:13 +05:30
Sameer Kankute
4f38a63152 Add long context support for claude-4-sonnet 2025-08-19 16:40:12 +05:30
drorbaron
7fedcf1ea9 migrate stream ws 2025-08-19 12:36:29 +03:00
drorbaron
6b78ade918 migrate to use new aim FW API 2025-08-19 12:20:19 +03:00
openhands
7530f13e05 Fix IRSA role assumption logic for AWS Bedrock
- Fixed condition to properly detect IRSA environments using AWS_WEB_IDENTITY_TOKEN_FILE
- Skip role assumption when already running as target role in IRSA environment
- Prevents unnecessary AWS API calls that cause 'root account cannot assume role' errors
- Resolves failing test test_auth_with_aws_role_same_role_irsa

The original issue was that aws_access_key_id and aws_secret_access_key were being
populated from environment variables even when passed as None, causing the IRSA
detection condition to fail. The fix checks for IRSA-specific environment variables
instead of relying on the absence of explicit credentials.

Fixes #13417
2025-08-19 08:27:24 +00:00
openhands
93651c9da7 Revert "Revert "fix: role chaining and session name with webauthentication for aws be…" (#13230)"
This reverts commit 342fd2d8b6.
2025-08-19 08:19:18 +00:00
Tim Elfrink
b5fa2ee73f Merge remote-tracking branch 'origin/main' into feat/github-copilot-thinking-reasoning-support 2025-08-19 10:11:59 +02:00
Krish Dholakia
de59691c4b Enable update/delete org members on UI (#8560)
* feat(organization_endpoints.py): expose new `/organization/delete` endpoint. Cascade org deletion to member, teams and keys

Ensures any org deletion is handled correctly

* test(test_organizations.py): add simple test to ensure org deletion works

* feat(organization_endpoints.py): expose /organization/update endpoint, and define response models for org delete + update

* fix(organizations.tsx): support org delete on UI + move org/delete endpoint to use DELETE

* feat(organization_endpoints.py): support `/organization/member_update` endpoint

Allow admin to update member's role within org

* feat(organization_endpoints.py): support deleting member from org

* test(test_organizations.py): add e2e test to ensure org member flow works

* fix(organization_endpoints.py): fix code qa check

* fix(schema.prisma): don't introduce ondelete:cascade - breaking change

* docs(organization_endpoints.py): document missing params
2025-08-19 10:53:24 +03:00
Tim Elfrink
9b0fda7b14 fix: resolve case sensitivity and test failures for extended thinking support
- Fix supports_reasoning() call to use lowercase model names for proper lookup
- Remove custom_llm_provider parameter as model registry entries are provider-agnostic
- Update tests to use full model names with date stamps (required for supports_reasoning)
- Add test coverage for models without extended thinking support
2025-08-19 08:40:10 +02:00
Tim Elfrink
9f82b89051 fix: remove redundant github_copilot check in get_supported_openai_params
The provider_config_manager already handles github_copilot provider
through LlmProviders.GITHUB_COPILOT mapping, making the explicit
check unnecessary.
2025-08-19 08:21:18 +02:00
Tim Elfrink
8b66b50c31 fix: restrict thinking/reasoning_effort parameters to models with extended thinking support
Only models in the 4 family and 3-7 family support extended thinking features.
Previously all models would incorrectly receive these parameters.

Now uses supports_reasoning() to check model registry for actual capability.
2025-08-19 08:15:50 +02:00
Krrish Dholakia
137a98a5af bump: version 1.75.8 → 1.75.9 2025-08-18 23:11:08 -07:00
Krish Dholakia
435995ba5c
Merge pull request #13617 from moandersson/fix/migratejob-resources
Add possibility to configure resources for migrations-job in Helm chart
2025-08-18 23:03:35 -07:00
Krish Dholakia
88e52c55d0
Merge pull request #13675 from colesmcintosh/fix/groq-streaming-encoding
Fix Groq streaming ASCII encoding issue
2025-08-18 23:00:00 -07:00
Krish Dholakia
9dadd279a4
Merge pull request #13741 from BerriAI/litellm_dev_08_18_2025_p1
Refactor - forward model group headers - reuse same logic as global header forwarding
2025-08-18 22:58:39 -07:00
Krrish Dholakia
2c0520635d test: cleanup old tests 2025-08-18 22:58:29 -07:00
Krish Dholakia
048f22b7ec
Merge pull request #13742 from BerriAI/litellm_dev_08_18_2025_p2
Fix - gemini prompt caching cost calculation
2025-08-18 22:54:28 -07:00
Krish Dholakia
b5f06e2bd3
Merge pull request #13685 from BerriAI/azure-deployment-name-preset
Add Azure Deployment Name Support in UI
2025-08-18 22:48:33 -07:00
Krish Dholakia
78fcf7afa9
Merge pull request #13687 from BerriAI/model-filter-on-models
Add Search Functionality for Public Model Names in Model Dashboard
2025-08-18 22:44:36 -07:00
Krish Dholakia
3713c926c0
Merge pull request #13704 from michal-otmianowski/use-namespace-as-prefix-for-s3-cache
Use namespace as prefix for s3 cache
2025-08-18 22:37:18 -07:00
Krrish Dholakia
f7f1a0d0b7 test: add unit test 2025-08-18 22:32:36 -07:00
Krrish Dholakia
241f32b2a3 fix(vertex_and_google_ai_studio_gemini.py): adjust 'text token' value to be the non-cached tokens - for accurate cost tracking 2025-08-18 22:09:07 -07:00
Krrish Dholakia
2e16f2cb13 test: add unit tests 2025-08-18 21:19:43 -07:00
Krrish Dholakia
36f93444b2 refactor: cleanup 2025-08-18 21:14:38 -07:00
Krrish Dholakia
06d05c691d fix(litellm_pre_call_utils.py): forward headers by model group at litellm pre call utils level
do it at the proxy level instead of router - allows reusing same forwarding logic as global forwarding
2025-08-18 21:13:52 -07:00
Krrish Dholakia
549acbc822 fix: fix test 2025-08-18 19:16:53 -07:00
Krish Dholakia
422447b7f1
Responses API - add default api version for openai responses api calls + Openrouter - fix claude-sonnet-4 on openrouter + Azure - Handle openai/v1/responses
Responses API - add default api version for openai responses api calls + Openrouter - fix claude-sonnet-4 on openrouter + Azure - Handle `openai/v1/responses`
2025-08-18 18:59:28 -07:00
Krrish Dholakia
0459604721 docs: document new param 2025-08-18 18:56:39 -07:00
Krish Dholakia
3b52545db3
Merge pull request #13529 from BerriAI/litellm_dev_08_11_2025_p1
[Fix] Cooldowns - don't return raw Azure Exceptions to client
2025-08-18 18:54:19 -07:00
Ishaan Jaff
0945483721 fix mypy linting errors 2025-08-18 18:51:14 -07:00
Ishaan Jaff
aac7cdf2c8 ruff check fix 2025-08-18 18:28:25 -07:00
Ishaan Jaff
ba0881d728
[Bug Fix] image_edit() function returns APIConnectionError with litellm_proxy - Support for both image edits and image generations (#13735)
* add image edits litellm proxy on SDK

* add image gen provider

* add IMG Gen support for litellm_proxy provider
2025-08-18 18:26:32 -07:00
Ishaan Jaff
76f1064229
[Bug Fix] litellm incompatible with newest release of openAI v1.100.0 (#13728)
* fix imports OpenAI SDK

* ResponseText fixes

* fixes ResponseText

* fix imports

* catch AttributeError

* fix import

* use openai==1.100.1

* fix build from PIP

* fix lint test

* Print OpenAI version

* fix Install dependencies
2025-08-18 18:26:17 -07:00
Ishaan Jaff
ba1d2e8749
[Feat] DD LLM Observability - Add time to first token, litellm overhead, guardrail overhead latency metrics (#13734)
* fixes for DDLLMObsLatencyMetrics

* use _get_latency_metrics

* DD LLM Obs - track latency metrics

* fixes for bedrock guardrails

* DD unit tests

* test DD
2025-08-18 17:38:04 -07:00
Ishaan Jaff
ef08e18c66
[Feat] Datadog LLM Observability - Add support for Failure Logging (#13726)
* add async_log_failure_event for DD LLM Obs

* update types

* DataDogLLMObsLogger  add failure logging support

* test_async_log_failure_event

* dd test failure
2025-08-18 15:19:48 -07:00