yuneng-jiang
111593397a
fixing core proxy tests
2026-02-12 17:54:32 -08:00
yuneng-jiang
8d10311b4b
content filter test fix
2026-02-12 17:54:16 -08:00
yuneng-jiang
e49d094606
fix openai tests
2026-02-12 17:53:47 -08:00
yuneng-jiang
c37e3be933
fixing mcp tests
2026-02-12 17:53:40 -08:00
yuneng-jiang
c41459c8e3
fixing mistral model deprecation, cloud zero transform bug
2026-02-12 17:53:04 -08:00
yuneng-jiang
514645777b
adding key envs to docs
2026-02-12 17:52:35 -08:00
yuneng-jiang
0efe3ea825
bumping pillow and cryptography for security fixes
2026-02-12 17:52:20 -08:00
yuneng-jiang
2a4066646a
fix mypy linting
2026-02-12 17:52:12 -08:00
yuneng-jiang
2c8225b465
fix ruff check
2026-02-12 17:52:00 -08:00
yuneng-jiang
918376ddaf
Merge pull request #20124 from naaa760/fix/management-key-access
...
fix: allow Management keys to access user/daily/activity and team
2026-02-12 17:11:18 -08:00
Harshit Jain
f77fbefc22
fix: resolve conflicts with verification e2e
2026-02-13 06:11:03 +05:30
milan-berri
a2e9e73b64
fix(proxy): change model mismatch logs from WARNING to DEBUG ( #20994 )
...
Fixes #20990
PR #19943 added logging when the proxy overrides model names to prevent
internal provider prefixes from leaking to clients. The behavior works
correctly but logs a WARNING on every request with model mismatch.
For high-traffic customers using model aliases or provider prefixes,
this creates millions of warnings per day, flooding logs and causing
disk space issues.
Changed log level from WARNING to DEBUG since:
- The model mismatch is expected behavior when using aliases
- The override happens correctly regardless of log level
- Operators can still enable with LITELLM_LOG=DEBUG for debugging
Changes:
- common_request_processing.py: 2 warnings -> debug (non-streaming)
- proxy_server.py: 1 warning -> debug (streaming)
2026-02-12 16:40:58 -08:00
yuneng-jiang
ce3bb97d40
Merge pull request #21076 from BerriAI/litellm_ui_model_table_cred
...
[Feature] UI - Model Page: Improve Credentials Messaging
2026-02-12 16:28:45 -08:00
Emerson Gomes
d9606773ea
feat(vertex_ai): add zai-org/glm-5-maas model pricing ( #21053 )
...
Add Vertex AI ZAI GLM-5 model map entry with reasoning + prompt caching metadata and cache-read pricing.\n\nRefs #21052
Co-authored-by: Codex <codex@example.com>
2026-02-12 16:11:59 -08:00
yuneng-jiang
168a8731ce
improve credentials messaging
2026-02-12 16:02:54 -08:00
yuneng-jiang
3cbb12b9c8
Merge pull request #21074 from milan-berri/fix/mcp-server-name-validation-spaces
...
fix(ui): Block spaces and hyphens in MCP server names and aliases
2026-02-12 15:35:18 -08:00
yuneng-jiang
2864ce73da
Merge pull request #21022 from BerriAI/litellm_unified_ag
...
[Feature] Access Groups
2026-02-12 15:34:38 -08:00
Milan
de42f733df
fix: Update alias tooltip - remove outdated space replacement text
...
Since spaces are now blocked in server names, the tooltip text about
'spaces replaced by underscores' is no longer accurate.
2026-02-13 01:20:57 +02:00
Harshit Jain
c56bbb9067
Update tests/test_litellm/proxy/management_endpoints/test_ui_sso.py
...
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
2026-02-13 04:48:36 +05:30
Milan
b769fa08d2
chore: Remove accidentally committed image file
2026-02-13 01:18:19 +02:00
Milan
bab5173500
fix: Make validation message generic and restore alias tooltip text
...
- Change error message to be generic (works for both server_name and alias)
- Restore 'Defaults to server name with spaces replaced' text in alias tooltip
2026-02-13 01:17:01 +02:00
Milan
1d67476ed0
refactor: Simplify validateMCPServerName to match original ternary style
2026-02-13 01:14:37 +02:00
Ishaan Jaff
5f40f93846
fix: MCP - inject NPM_CONFIG_CACHE into STDIO MCP subprocess env ( #21069 )
...
* fix: inject NPM_CONFIG_CACHE into STDIO MCP subprocess env for Docker
npm/npx needs a writable cache directory. In containers the default
(~/.npm) may not exist or be read-only, causing STDIO MCP servers
launched via npx to fail with ENOENT. Inject NPM_CONFIG_CACHE=/tmp/.npm_mcp_cache
into the subprocess env when not already set.
* test: add unit test for NPM_CONFIG_CACHE injection in STDIO MCP
Verifies that NPM_CONFIG_CACHE is auto-injected when not set, and
preserved when explicitly provided. Also moves the import to module
level per code style rules.
* Update litellm/constants.py
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
* Apply suggestion from @greptile-apps[bot]
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
---------
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
2026-02-12 15:11:37 -08:00
Milan
8fa2734830
fix(ui): Block spaces and hyphens in MCP server names and aliases
...
- Update validateMCPServerName to reject both spaces and hyphens
- Apply shared validation to alias field in create form (was inline)
- Update tooltips to mention space restriction
- Ensures consistency across create/edit forms for server_name and alias fields
2026-02-13 01:11:06 +02:00
Harshit Jain
a2b4728e74
Update tests/test_litellm/proxy/management_endpoints/test_ui_sso.py
...
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
2026-02-13 04:40:29 +05:30
Alejandro Tapia
82f6d0fe43
healthcheck-model_id-fix: There was quite a bit of code that needed to be changed since health checks were entirely keyed by model name. This includes, proxy logic, the dashboard, and even networking, because model name was the identifier everywhere. All changed files had tests added to them, which are passing with no regressions.
2026-02-12 14:47:09 -08:00
yuneng-jiang
a37623945d
migration and build
2026-02-12 14:34:10 -08:00
Harshit Jain
e0846389e9
add modify test to perform async run
2026-02-13 04:04:03 +05:30
yuneng-jiang
ed59c7c84d
bump: version 0.4.35 → 0.4.36
2026-02-12 14:33:38 -08:00
yuneng-jiang
a45028f623
Merge remote-tracking branch 'origin' into litellm_unified_ag
2026-02-12 14:32:52 -08:00
Ryan Crabbe
2065e5b88b
perf: cache model_fields.keys() as frozensets in convert_to_model_response_object (15% faster)
...
Replace per-call .model_fields.keys() allocations and linear-scan membership
checks with module-level frozenset constants and dict.keys() set difference.
Defer locals() from hot path to except block. 617µs → 524µs/call.
2026-02-12 14:25:09 -08:00
Harshit Jain
847402b68d
Merge branch 'fix/sso_PKCE_deployments' of https://github.com/Harshit28j/litellm into fix/sso_PKCE_deployments
2026-02-13 03:50:16 +05:30
Harshit Jain
eb249b2f06
fix: add await in tests
2026-02-13 03:46:31 +05:30
Harshit Jain
bc5543cfdc
Update tests/test_litellm/proxy/management_endpoints/test_ui_sso.py
...
Co-authored-by: greptile-apps[bot] <165735046+greptile-apps[bot]@users.noreply.github.com>
2026-02-13 03:39:57 +05:30
yuneng-jiang
ea8c89ea3f
rename file and add tests
2026-02-12 14:05:51 -08:00
Harshit Jain
1792b3c8e5
fix: add async call to avoid server pauses
2026-02-13 03:28:51 +05:30
yuneng-jiang
c34cdb29cd
remove double auth and add alias
2026-02-12 13:40:58 -08:00
Ishaan Jaff
736daf0a7d
[Feat] Adds Shell tool support for the OpenAI Responses API ( #21063 )
...
* test_responses_api_context_management_server_side_compaction
* Server-side compaction
* docs fix
* test_responses_api_shell_tool
* add SHELL tool
* test_responses_api_shell_tool
* add SHELL_CALL_IN_PROGRESS
* add SHELL_CALL_IN_PROGRESS events
* TestOpenAIResponsesAPITest
* transform_streaming_response
* test_responses_api_shell_tool_streaming_sees_shell_output
* test_responses_api_shell_tool_streaming_sees_shell_output
* test_responses_api_shell_tool
* docs fix
2026-02-12 13:04:29 -08:00
yuneng-jiang
e6df587bfb
adding tests and fixing prisma lookup table
2026-02-12 12:48:05 -08:00
yuneng-jiang
fbfaa6c8af
rename unified access group to access group
2026-02-12 12:30:22 -08:00
yuneng-jiang
5a0db00d8a
Merge pull request #21061 from BerriAI/migration_yj_feb12
...
[Infra] Add Mmigration for Tags Adjustment on Policy Table
2026-02-12 10:36:10 -08:00
yuneng-jiang
5147515d78
add migration + build files
2026-02-12 10:34:59 -08:00
yuneng-jiang
5152bf4f4e
bump: version 0.4.34 → 0.4.35
2026-02-12 10:34:17 -08:00
yuneng-jiang
df15456bcc
Merge pull request #20598 from muraliavarma/fix/team-update-empty-premium-fields-403
...
fix(proxy): skip premium check for empty metadata fields on team/key update
2026-02-12 10:10:27 -08:00
Ishaan Jaff
89565c97cc
[Feat] AI Gateway - Add Tracing for MCP Calls running through AI Gateway ( #21018 )
...
* commit new expansion
* fix MCP
* fix: LiteLLMProxyRequestSetup
* _process_mcp_tools_without_openai_transform
* UI fixes
* UI refactor view logs/sessions
* index
* _add_mcp_tool_metadata_to_final_chunk
* add badges
* add getEventDisplayName
* ui fixes
* backend fix
* fix
* UI fix
* UI fix
* fix row
* fix: address Greptile review feedback on PR #21018 (#21057 )
- Fix session time range calculation: use Math.min/Math.max across all
entries instead of relying on array order (sessionLogs is sorted by
type, not time).
Other Greptile comments were already addressed in the branch:
- LogDetailContent.tsx exists
- Clipboard call already wrapped in try/catch
- Dedup already uses O(1) Map lookup
- model_dump() serialization is documented
- GROUP BY performance comment already present
---------
Co-authored-by: shin-bot-litellm <shin-bot-litellm@berri.ai>
2026-02-12 10:02:03 -08:00
Ishaan Jaff
3d9b145b04
[Feat] Adds support for server-side compaction on the OpenAI Responses API context_management ( #21058 )
...
* test_responses_api_context_management_server_side_compaction
* Server-side compaction
* docs fix
* test_responses_api_shell_tool
2026-02-12 10:00:30 -08:00
Krrish Dholakia
f5382ebac9
docs: fix docs
2026-02-12 08:45:57 -08:00
Sameer Kankute
556bcd7203
Merge pull request #21055 from BerriAI/litellm_day_0_MiniMax-M2.1
...
fix docs
2026-02-12 22:04:23 +05:30
Sameer Kankute
9f15eca6b6
fix docs
2026-02-12 22:03:21 +05:30
Sameer Kankute
4e62386c65
Merge pull request #21054 from BerriAI/litellm_day_0_MiniMax-M2.1
...
Add support for MiniMax-M2.1 and MiniMax-M2.1-lightining
2026-02-12 21:51:46 +05:30