Krish Dholakia
d1adc7896d
Merge pull request #14262 from TobiMayr/feature/gemini-reasoning-effort-minimal
...
feat: Add support for reasoning_effort='minimal' for Gemini models
2025-09-06 09:17:23 -07:00
Krish Dholakia
026fc2553d
Merge pull request #14287 from alpha-pet/fix/gemini-25-flash-image-price
...
fix: correct output pricing for gemini-2.5-flash-image-preview
2025-09-06 09:14:26 -07:00
Krish Dholakia
943f6a38ed
Merge pull request #14300 from eycjur/gpt-oss-reasoning-effort-mapping
...
Bug fix for openai.gpt-oss when using reasoning_effort parameter
2025-09-06 09:13:03 -07:00
Ishaan Jaff
29e410b04c
docs Video Walkthrough claude code
2025-09-06 09:12:30 -07:00
Mubashir Osmani
cb117647fc
[docs]: added loom for claude code ( #14223 )
...
* added loom for claude code
* docs: add web search models
* added new loom
2025-09-06 09:09:22 -07:00
Ishaan Jaff
0cb01d6027
Fix: Include model name in Azure base_model error ( #14294 )
...
Co-authored-by: Cursor Agent <cursoragent@cursor.com>
Co-authored-by: ishaan <ishaan@berri.ai>
2025-09-06 09:06:43 -07:00
Duc Tran
3478c53c60
Update constants.py ( #14242 )
2025-09-06 09:06:04 -07:00
katsuhiro muto
51de2ebb64
[Feat]Cancel upstream on client disconnect ( #14295 )
...
* cancel upstream on client disconnect
* add comments
* add test
* set timeout in constraints.py
* Guard against missing 'type' key
* update dependency to fix uvicorn bugs
2025-09-06 08:58:51 -07:00
eycjur
67315d8727
fix ci
2025-09-06 21:43:30 +09:00
eycjur
6eb1b40336
refactor
2025-09-06 21:30:08 +09:00
eycjur
b472bf6aef
update docs
2025-09-06 21:15:59 +09:00
eycjur
58cf72ef5e
add test
2025-09-06 21:12:18 +09:00
eycjur
e27c4c98c0
Added conditional branch for gpt-oss
2025-09-06 21:11:53 +09:00
Ishaan Jaff
5310bba35b
[Feat] Litellm x CloudZero Integration - Cost Tracking ( #14296 )
...
* fix: just pull LiteLLM_DailyUserSpend
* get the team_id from user daily spend table
* cloudzero_dry_run_export
* fix CZ endpoints
* trace entity_id
* fix: get_usage_data
* fix get_usage_data
* fix _create_cbf_record
* fix get_usage_data
* ensure start and end time is used for exporting data
* fix init_cloudzero_background_job
* fix CloudZeroExportRequest
* fix initialize_cloudzero_export_job
* fix initialize_cloudzero_export_job
* allow init with env + config.yaml for cloudzero
* fix: init CZ through config.yaml
* fix DRY run on CZ
* TestCloudZeroDryRunEndpoint
* fix: CLOUDZERO_EXPORT_INTERVAL_MINUTES
* fix init_cloudzero_background_job
* fix exporting data
* fix transform
* stash cloudzero docs
* docs: CloudZero
* ruff fix
* fix rendering key alias
* fix polars
2025-09-05 21:29:41 -07:00
Sameer Kankute
07ba3ff036
[Feat] Add pass through image gen and image editing on OpenAI ( #14292 )
...
* add pass through image gen and image editing on OpenAI
* fix lint
2025-09-05 12:25:49 -07:00
Pierre-Emmanuel MERCIER
31f806f7d0
feat: add redis ssl and username support ( #11319 )
2025-09-05 10:35:11 -07:00
Krrish Dholakia
c051ab5b5a
refactor: remove unused function
2025-09-05 10:21:16 -07:00
Ishaan Jaff
0a60390521
Revert "[Feat] LiteLLM CloudZero Integration updates - using LiteLLM_SpendLogs Table ( #12922 )"
...
This reverts commit e3b752d3dc .
2025-09-05 10:04:07 -07:00
Ishaan Jaff
982800069c
[Bug Fix] x-litellm-tags not routing with Responses API ( #14289 )
...
* fix: get_deployments_for_tag
* fix get_deployments_for_tag
* test_router_tag_routing.py
* test_get_metadata_variable_name_from_kwargs
* fix mapped tests
* docs fix
2025-09-05 09:40:37 -07:00
Thomas Rehn
d88771ca49
fix: correct output pricing for gemini-2.5-flash-image-preview https://ai.google.dev/gemini-api/docs/pricing#gemini-2.5-flash-image-preview
2025-09-05 16:22:07 +02:00
Krish Dholakia
63a2a47ce0
Merge pull request #14003 from BerriAI/teams-on-users-page
...
Team name badge added on the User Details
2025-09-04 21:03:02 -07:00
Krish Dholakia
f67339a86c
Merge pull request #14028 from onlylhf/volcengine-embedding-support
...
Add Volcengine embedding module with handler and transformation logic
2025-09-04 21:01:25 -07:00
Krish Dholakia
62d4623d1c
Merge pull request #14107 from uc4w6c/feat/add_guardrails_anthropic
...
feat: Add guardrail to the Anthropic API endpoint
2025-09-04 20:50:43 -07:00
Krish Dholakia
5d5c19a302
Merge pull request #14077 from TomeHirata/codex/add-support-for-anthropic-citation-api
...
Add support for anthropic citation api in Databricks
2025-09-04 20:46:11 -07:00
Ishaan Jaff
379b0dbf14
[Fix] Ensure team_id is a required field for generating service account keys ( #14270 )
...
* generate_service_account_key_fn
* fix validate_team_id_used_in_service_account_request
* fix types
* test_validate_team_id_used_in_service_account_request_requires_team_id
2025-09-04 18:14:38 -07:00
Ishaan Jaff
5847037b3a
Add validation for STORE_MODEL_IN_DB when updating public model groups ( #14269 )
...
Co-authored-by: Cursor Agent <cursoragent@cursor.com>
Co-authored-by: ishaan <ishaan@berri.ai>
2025-09-04 14:40:28 -07:00
Ishaan Jaff
1352545617
[Fix] DD LLM Observability - Ensure apm_id is set on traces ( #14272 )
...
* add apm_id for DD LLM
* feat: add _get_apm_trace_id
2025-09-04 14:35:41 -07:00
tobias-mayr
29bbde5257
fix condition ordering and test
2025-09-04 22:18:50 +01:00
Krish Dholakia
796192dabe
Merge pull request #14130 from moshemorad/bedrock_fix_structure_output
...
Bedrock fix structure output
2025-09-04 12:22:35 -07:00
Krish Dholakia
1c75139cd3
Merge pull request #14136 from BerriAI/add-client-side-pagination
...
Add client side pagination on All Models table
2025-09-04 12:21:48 -07:00
Krish Dholakia
e7b4892124
Merge pull request #14156 from byrongrogan/byron--bedrock-passthrough-override
...
fix: Support AWS_BEDROCK_RUNTIME_ENDPOINT on bedrock passthrough, make work for URLs with a base path
2025-09-04 12:20:17 -07:00
tobias-mayr
d9304b74bd
fix accidental changes
2025-09-04 19:03:25 +01:00
tobias-mayr
2d30b55964
distinguish between gemini models
2025-09-04 18:59:37 +01:00
tobias-mayr
0ade6cceff
remove old code
2025-09-04 18:40:46 +01:00
Sameer Kankute
fc9560573b
[BUG] Fix response api for reasoning item in input for litellm proxy ( #14200 )
...
* fix response api for litellm proxy
* Add test for checking if status is getting removed
* add test in correct file
* remove hardcoded fields
* Make the handling simpler
* fix lint error:
2025-09-04 10:36:48 -07:00
tobias-mayr
99eceb8835
feat: Add support for reasoning_effort='minimal' for Gemini models
...
- Add DEFAULT_REASONING_EFFORT_MINIMAL_THINKING_BUDGET constant (128 tokens)
- Update Gemini transformation to handle 'minimal' reasoning_effort
- Maps 'minimal' to 128 tokens (Gemini's minimum thinking budget)
- Maintains backward compatibility with existing reasoning_effort values
- Fixes issue where Gemini API rejected 0 token thinking budget
2025-09-04 18:29:06 +01:00
Sameer Kankute
5f79e8aac6
Litellm passthrough cost tracking chat completion ( #14256 )
...
* feat: add structured output for sdk
* Add support for cost tracking for chat completion in passthrough
* remove not required changes
2025-09-04 09:57:48 -07:00
Ishaan Jaff
1237be04a5
test_aaamodel_prices_and_context_window_json_is_valid
2025-09-04 07:58:20 -07:00
Krish Dholakia
572ac0b88f
Merge pull request #14211 from zhxlp/main
...
fix: image_generation supports extra_body parameter
2025-09-04 07:35:53 -07:00
Krish Dholakia
cb8e3936ae
Merge pull request #14216 from mubashir1osmani/fix_all_docs
...
Fix custom callbacks doc
2025-09-04 07:35:38 -07:00
Krish Dholakia
8738f630ac
Merge pull request #14232 from mubashir1osmani/litellm_docs
...
[docs]: added more info to load balancing & pass through endpoints
2025-09-04 07:31:54 -07:00
Ishaan Jaff
8878951f16
docs fix: disable_add_user_agent_to_request_tags
2025-09-04 07:30:18 -07:00
Krish Dholakia
a8cce5fbf3
Merge pull request #14237 from yeahyung/fix/tpm_limit_bug
...
Fixes #14204 TPM Rate Limit Bug
2025-09-04 07:29:47 -07:00
Krish Dholakia
a74feb221f
Merge pull request #14241 from 22mSqRi/fix-key-reset-and-expires
...
fix: Key Budget not resets at expectable times
2025-09-04 07:29:06 -07:00
22mSqRi
25afa67616
fix: Key Budget not resets at expectable times
2025-09-04 09:04:01 +00:00
yeahyung
7829de2948
( #14204 ) add test code
2025-09-04 16:41:29 +09:00
yeahyung
9e3010daa4
( #14204 ) increase token usage with TTL preservation
2025-09-04 16:41:23 +09:00
mubashir1osmani
b0450d2ddf
docs: added more info to load balancing & passthrough endpoints
2025-09-03 23:58:48 -04:00
Ishaan Jaff
69d5a91e02
bump: version 1.76.2 → 1.76.3
2025-09-03 18:28:52 -07:00
Ishaan Jaff
ab3cd5e96e
fix memory_usage_in_mem_cache cache endpoint vulnerability ( #14229 )
2025-09-03 18:28:11 -07:00