Krrish Dholakia
|
071a70c5fc
|
test: fix watsonx api error
|
2024-05-13 19:01:19 -07:00 |
|
Krrish Dholakia
|
724d880a45
|
test(test_completion.py): handle async watsonx call fail
|
2024-05-13 18:40:51 -07:00 |
|
Krrish Dholakia
|
d4123951d9
|
test: handle watsonx rate limit error
|
2024-05-13 18:27:39 -07:00 |
|
Krrish Dholakia
|
155f1f164f
|
refactor(utils.py): trigger local_testing
|
2024-05-13 18:18:22 -07:00 |
|
Krrish Dholakia
|
29449aa5c1
|
fix(utils.py): fix watsonx exception mapping
|
2024-05-13 18:13:13 -07:00 |
|
Krrish Dholakia
|
3694b5e7c0
|
refactor(main.py): trigger new build
|
2024-05-13 18:12:01 -07:00 |
|
Krrish Dholakia
|
38988f030a
|
fix(router.py): fix typing
|
2024-05-13 18:06:10 -07:00 |
|
Krrish Dholakia
|
5488bf4921
|
feat(router.py): enable default fallbacks
allow user to define a generic list of fallbacks, in case a new deployment is bad
Closes https://github.com/BerriAI/litellm/issues/3623
|
2024-05-13 17:49:56 -07:00 |
|
Krrish Dholakia
|
d7c28509d7
|
fix(utils.py): watsonx ai exception mapping fix
|
2024-05-13 17:11:33 -07:00 |
|
Krrish Dholakia
|
af9489bbfd
|
fix(router.py): overloads fix
|
2024-05-13 17:04:04 -07:00 |
|
Krrish Dholakia
|
240c9550f0
|
fix(utils.py): handle api assistant returning 'null' role
Fixes https://github.com/BerriAI/litellm/issues/3621
|
2024-05-13 16:46:07 -07:00 |
|
Krish Dholakia
|
e92f433566
|
Merge pull request #3554 from paneru-rajan/Issue-3544-fix-message
Fixes #3544 based on the data-type of message
|
2024-05-13 16:23:58 -07:00 |
|
Krrish Dholakia
|
c7b3193944
|
fix(proxy/_types.py): allow jwt admin to access /team/list route
|
2024-05-13 16:07:31 -07:00 |
|
Ishaan Jaff
|
ea9b4dc439
|
Merge pull request #3619 from BerriAI/litellm_show_spend_reports
[Feat] - `/global/spend/report`
|
2024-05-13 16:06:02 -07:00 |
|
Ishaan Jaff
|
eb2d6ba20a
|
ui - new build
|
2024-05-13 15:56:59 -07:00 |
|
Krrish Dholakia
|
b4a8665d11
|
fix(utils.py): fix custom pricing when litellm model != response obj model name
|
2024-05-13 15:25:35 -07:00 |
|
Ishaan Jaff
|
1be6ea0c0d
|
Merge pull request #3603 from alexanderepstein/langfuse_turn_off_messaging
feat(langfuse.py): Allow for individual call message/response redaction
|
2024-05-13 15:21:41 -07:00 |
|
Ishaan Jaff
|
12cf9d71c7
|
feat - /spend/report endpoint
|
2024-05-13 15:01:02 -07:00 |
|
Krrish Dholakia
|
1312eece6d
|
fix(router.py): overloads for better router.acompletion typing
|
2024-05-13 14:27:16 -07:00 |
|
Krrish Dholakia
|
bd2f46fd75
|
fix(slack_alerting.py): if 'turn_off_message_logging' enabled, do not log the message to logging integration
|
2024-05-13 14:02:43 -07:00 |
|
Ishaan Jaff
|
9d4b727913
|
fix - show team based spend reports
|
2024-05-13 13:56:48 -07:00 |
|
Krrish Dholakia
|
20456968e9
|
fix(openai.py): creat MistralConfig with response_format mapping for mistral api
|
2024-05-13 13:29:58 -07:00 |
|
Ishaan Jaff
|
21845bc061
|
Merge pull request #3609 from BerriAI/litellm_send_daily_spend_report
[Feat] send weekly spend reports by Team/Tag
|
2024-05-13 12:45:37 -07:00 |
|
Krrish Dholakia
|
39e4927752
|
fix(utils.py): fix vertex ai function calling + streaming
Completes https://github.com/BerriAI/litellm/issues/3147
|
2024-05-13 12:32:39 -07:00 |
|
Ishaan Jaff
|
40b2f33a80
|
fix - only schedule spend alerting when db is not none
|
2024-05-13 12:30:54 -07:00 |
|
Krrish Dholakia
|
5dc3f157a6
|
build(model_prices_and_context_window.json): fix gpt-4o max tokens
|
2024-05-13 12:11:15 -07:00 |
|
Ishaan Jaff
|
3fbebe16aa
|
fix - spend reports on alerts
|
2024-05-13 10:51:59 -07:00 |
|
Ishaan Jaff
|
aac81c59b5
|
test - weekly / monthly spend report alerts on /health/services
|
2024-05-13 10:50:26 -07:00 |
|
Ishaan Jaff
|
197eb44832
|
fix scheduling spend reports
|
2024-05-13 10:45:22 -07:00 |
|
Ishaan Jaff
|
4a679bb640
|
schedule weekly/monthly spend reports
|
2024-05-13 10:44:19 -07:00 |
|
Krrish Dholakia
|
04ae285001
|
fix(vertex_ai.py): support tool call list response async completion
|
2024-05-13 10:42:31 -07:00 |
|
Ishaan Jaff
|
2cd584a3ad
|
feat - send_monthly_spend_report
|
2024-05-13 10:17:40 -07:00 |
|
Krrish Dholakia
|
7f6e933372
|
fix(router.py): give an 'info' log when fallbacks work successfully
|
2024-05-13 10:17:32 -07:00 |
|
Ishaan Jaff
|
55f747fb1d
|
fix - show monthly spend in slack reports
|
2024-05-13 10:17:09 -07:00 |
|
Ishaan Jaff
|
07247452c5
|
feat - show monthly spend reports
|
2024-05-13 10:10:44 -07:00 |
|
Krrish Dholakia
|
13e1577753
|
fix(slack_alerting.py): don't fire spam alerts when backend api call fails
|
2024-05-13 10:04:43 -07:00 |
|
Ishaan Jaff
|
50f3677989
|
feat - _get_weekly_spend_reports
|
2024-05-13 09:26:51 -07:00 |
|
Ishaan Jaff
|
b7bbaf1a68
|
feat - send daily spend reports
|
2024-05-13 09:25:31 -07:00 |
|
Krrish Dholakia
|
5342b3dc05
|
fix(router.py): fix error message to return if pre-call-checks + allowed model region
|
2024-05-13 09:04:38 -07:00 |
|
Krrish Dholakia
|
c3293474dd
|
fix(proxy_server.py): return 'allowed-model-region' in headers
|
2024-05-13 08:48:16 -07:00 |
|
Ishaan Jaff
|
514c5737f8
|
Merge pull request #3587 from BerriAI/litellm_proxy_use_batch_completions_model_csv
[Feat] Use csv values for proxy batch completions (OpenAI Python compatible)
|
2024-05-13 07:55:12 -07:00 |
|
Alex Epstein
|
3bf2ccc856
|
feat(langfuse.py): Allow for individual call message/response redaction
|
2024-05-12 22:38:29 -04:00 |
|
Krrish Dholakia
|
61143c8b45
|
refactor(main.py): trigger new build
|
2024-05-11 22:53:09 -07:00 |
|
Krrish Dholakia
|
b4684d5132
|
fix(proxy_server.py): linting fix
|
2024-05-11 22:05:01 -07:00 |
|
Krrish Dholakia
|
094f20121a
|
build(model_prices_and_context_window.json): add bedrock cohere command r pricing
|
2024-05-11 21:38:53 -07:00 |
|
Krrish Dholakia
|
15a6e59431
|
fix(proxy/_types.py): allow jwt admin to access spend routes
|
2024-05-11 21:31:34 -07:00 |
|
Krish Dholakia
|
1d651c6049
|
Merge branch 'main' into litellm_bedrock_command_r_support
|
2024-05-11 21:24:42 -07:00 |
|
Krrish Dholakia
|
e8437e52fa
|
test(test_rules.py): fix test
|
2024-05-11 21:22:37 -07:00 |
|
Krish Dholakia
|
7566a2fc78
|
Merge pull request #3589 from msabramo/msabramo/make_test_load_router_config_pass
Make `test_load_router_config` pass
|
2024-05-11 21:15:07 -07:00 |
|
Krrish Dholakia
|
d142478b75
|
fix(langfuse.py): fix handling of dict object for langfuse prompt management
|
2024-05-11 20:42:55 -07:00 |
|