ishaan-jaff
|
8175fb4deb
|
(fix) mark semantic caching as beta test
|
2024-02-06 11:04:19 -08:00 |
|
ishaan-jaff
|
1afdf5cf36
|
(fix) semantic caching
|
2024-02-06 10:55:15 -08:00 |
|
ishaan-jaff
|
54c920c299
|
(docs) litellm semantic caching
|
2024-02-06 10:54:55 -08:00 |
|
ishaan-jaff
|
93504915d7
|
(docs) redis cache
|
2024-02-06 10:53:28 -08:00 |
|
ishaan-jaff
|
c8a83bb745
|
(fix) test-semantic caching
|
2024-02-06 10:39:44 -08:00 |
|
ishaan-jaff
|
2732c47b70
|
(feat) redis-semantic cache on proxy
|
2024-02-06 10:35:21 -08:00 |
|
ishaan-jaff
|
bdc2091838
|
(docs) using semantic caching on proxy
|
2024-02-06 10:32:07 -08:00 |
|
ishaan-jaff
|
a1fc1e49c7
|
(fix) use semantic cache on proxy
|
2024-02-06 10:27:33 -08:00 |
|
ishaan-jaff
|
05f379234d
|
allow setting redis_semantic cache_embedding model
|
2024-02-06 10:22:02 -08:00 |
|
ishaan-jaff
|
751fb1af89
|
(feat) log semantic_sim to langfuse
|
2024-02-06 09:31:57 -08:00 |
|
ishaan-jaff
|
c4e73768cf
|
(fix) add redisvl==0.0.7
|
2024-02-06 09:30:45 -08:00 |
|
ishaan-jaff
|
70a895329e
|
(feat) working semantic cache on proxy
|
2024-02-06 08:55:25 -08:00 |
|
ishaan-jaff
|
a3b1e3bc84
|
(feat) redis-semantic cache
|
2024-02-06 08:54:36 -08:00 |
|
ishaan-jaff
|
6249a97098
|
(feat) working semantic-cache on litellm proxy
|
2024-02-06 08:52:57 -08:00 |
|
ishaan-jaff
|
a125ffe190
|
(test) async semantic cache
|
2024-02-06 08:14:54 -08:00 |
|
ishaan-jaff
|
76def20ffe
|
(feat) RedisSemanticCache - async
|
2024-02-06 08:13:12 -08:00 |
|
ishaan-jaff
|
ccc94128d3
|
(fix) semantic cache
|
2024-02-05 18:25:22 -08:00 |
|
ishaan-jaff
|
81f8ac00b2
|
(test) semantic caching
|
2024-02-05 18:22:50 -08:00 |
|
ishaan-jaff
|
cf4bd1cf4e
|
(test) semantic cache
|
2024-02-05 17:58:32 -08:00 |
|
ishaan-jaff
|
1b39454a08
|
(feat) working - sync semantic caching
|
2024-02-05 17:58:12 -08:00 |
|
ishaan-jaff
|
d4a799a3ca
|
(feat )add semantic cache
|
2024-02-05 12:28:21 -08:00 |
|
Krish Dholakia
|
646764f1d4
|
Merge pull request #1826 from promptmetheus/update-perplexity-model-info
Update Perplexity models in model_prices_and_context_window.json
|
2024-02-05 08:53:16 -08:00 |
|
Krrish Dholakia
|
031bc96f88
|
bump: version 1.22.3 → 1.22.4
|
2024-02-05 08:47:10 -08:00 |
|
Krrish Dholakia
|
1bdb332454
|
fix(utils.py): handle count response tokens false case token counting
|
2024-02-05 08:47:10 -08:00 |
|
Toni Engelhardt
|
e832492423
|
Update Perplexity models in model_prices_and_context_window.json
According to https://docs.perplexity.ai/docs/model-cards
|
2024-02-05 16:46:10 +00:00 |
|
Ishaan Jaff
|
109ccf4cef
|
Merge pull request #1602 from ShaunMaher/add_helm_chart
Add a Helm chart for deploying LiteLLM Proxy
|
2024-02-05 08:12:58 -08:00 |
|
Ishaan Jaff
|
14c9e239a1
|
Merge pull request #1750 from vanpelt/patch-2
Re-raise exception in async ollama streaming
|
2024-02-05 08:12:17 -08:00 |
|
Krrish Dholakia
|
d1f9165af1
|
bump: version 1.22.2 → 1.22.3
|
2024-02-03 22:23:43 -08:00 |
|
Krish Dholakia
|
640572647a
|
Merge pull request #1805 from BerriAI/litellm_cost_tracking_image_gen
feat(utils.py): support cost tracking for openai/azure image gen models
|
2024-02-03 22:23:22 -08:00 |
|
Krrish Dholakia
|
49b2dc4180
|
test(test_completion_cost.py): fix test
|
2024-02-03 22:00:49 -08:00 |
|
Krrish Dholakia
|
66565f96b1
|
test(test_completion.py): fix test
|
2024-02-03 21:44:57 -08:00 |
|
Krrish Dholakia
|
d2d57ecf1c
|
test(test_parallel_request_limiter.py): fix test
|
2024-02-03 21:31:29 -08:00 |
|
Krrish Dholakia
|
3a19c8b600
|
test(test_completion.py): fix test
|
2024-02-03 21:30:45 -08:00 |
|
Krish Dholakia
|
7ee5d2e0d9
|
Update model_prices_and_context_window.json
|
2024-02-03 21:12:26 -08:00 |
|
Krish Dholakia
|
6eed51607a
|
Update model_prices_and_context_window.json
|
2024-02-03 21:09:00 -08:00 |
|
Krish Dholakia
|
28df60b609
|
Merge pull request #1809 from BerriAI/litellm_embedding_caching_updates
Support caching individual items in embedding list (Async embedding only)
|
2024-02-03 21:04:23 -08:00 |
|
ishaan-jaff
|
c353161456
|
(fix) test_parallel limiter fix
|
2024-02-03 21:03:15 -08:00 |
|
Ishaan Jaff
|
24aa304061
|
Update README.md
|
2024-02-03 20:54:14 -08:00 |
|
Ishaan Jaff
|
254a5e8533
|
Update README.md
|
2024-02-03 20:51:43 -08:00 |
|
Ishaan Jaff
|
fc8811a5e7
|
Update ghcr_deploy.yml
|
2024-02-03 20:50:18 -08:00 |
|
ishaan-jaff
|
7e43d22da4
|
(docs) update ui
|
2024-02-03 20:43:43 -08:00 |
|
Ishaan Jaff
|
f3ec2efa21
|
Update README.md
|
2024-02-03 20:40:10 -08:00 |
|
Krrish Dholakia
|
3e35041758
|
test(test_parallel_request_limiter.py): fix test to handle minute changes
|
2024-02-03 20:39:31 -08:00 |
|
ishaan-jaff
|
1155025e6a
|
(ci/cd) run again
|
2024-02-03 20:36:35 -08:00 |
|
Krrish Dholakia
|
b47b2837eb
|
test(test_parallel_request_limiter.py): fix test
|
2024-02-03 20:34:05 -08:00 |
|
Krrish Dholakia
|
10822ded76
|
bump: version 1.22.1 → 1.22.2
|
2024-02-03 20:23:55 -08:00 |
|
Krrish Dholakia
|
312c7462c8
|
refactor(ollama.py): trigger rebuild
|
2024-02-03 20:23:43 -08:00 |
|
Krrish Dholakia
|
01cef1fe9e
|
fix(ollama.py): fix api connection error
https://github.com/BerriAI/litellm/issues/1735
|
2024-02-03 20:22:33 -08:00 |
|
ishaan-jaff
|
dc506e3df1
|
bump: version 1.22.0 → 1.22.1
|
2024-02-03 20:01:05 -08:00 |
|
ishaan-jaff
|
774cbbde52
|
(test) tgai is unstable
|
2024-02-03 20:00:40 -08:00 |
|