litellm/litellm
Ishaan Jaff cef27241e1
Merge pull request #4427 from BerriAI/litellm_fix_cost_tracking_whisper
[Fix-Proxy] Store SpendLogs when using Whisper, Moderations etc
2024-06-26 15:27:50 -07:00
..
assistants feat(assistants/main.py): support arun_thread_stream 2024-06-04 16:47:51 -07:00
batches docs(customers.md): add customer cost tracking to docs 2024-05-29 14:55:33 -07:00
deprecated_litellm_server refactor: add black formatting 2023-12-25 14:11:20 +05:30
integrations fix(router.py): set cooldown_time: per model 2024-06-25 16:51:55 -07:00
litellm_core_utils fix(router.py): set cooldown_time: per model 2024-06-25 16:51:55 -07:00
llms Merge pull request #4418 from BerriAI/litellm_fireworks_ai_tool_calling 2024-06-26 08:30:06 -07:00
proxy fix cost tracking for whisper 2024-06-26 14:21:57 -07:00
router_strategy refactor: replace 'traceback.print_exc()' with logging library 2024-06-06 13:47:43 -07:00
router_utils fix use safe access for router alerting 2024-06-14 15:17:32 -07:00
tests test_spend_logs_payload_whisper 2024-06-26 15:21:49 -07:00
types Add return type annotations to util types 2024-06-26 12:46:59 -04:00
__init__.py add fireworks ai param mapping 2024-06-26 06:43:18 -07:00
_logging.py fix(_logging.py): fix timestamp format for json logs 2024-06-20 15:20:21 -07:00
_redis.py feat(proxy_server.py): return litellm version in response headers 2024-05-08 16:00:08 -07:00
_service_logger.py feat(dynamic_rate_limiter.py): update cache with active project 2024-06-21 20:25:40 -07:00
_version.py (fix) ci/cd don't let importing litellm._version block starting proxy 2024-02-01 16:23:16 -08:00
budget_manager.py feat(proxy_server.py): return litellm version in response headers 2024-05-08 16:00:08 -07:00
caching.py feat(dynamic_rate_limiter.py): update cache with active project 2024-06-21 20:25:40 -07:00
cost.json store llm costs in budget manager 2023-09-09 19:11:35 -07:00
cost_calculator.py fix: use per-token costs for claude via vertex_ai 2024-06-21 11:21:36 -05:00
exceptions.py fix(utils.py): fix exception_mapping check for errors 2024-06-24 16:55:19 -07:00
main.py fix(router.py): set cooldown_time: per model 2024-06-25 16:51:55 -07:00
model_prices_and_context_window_backup.json fix add ollama codegemma 2024-06-26 12:57:09 -07:00
py.typed feature - Types for mypy - #360 2024-05-30 14:14:41 -04:00
requirements.txt Add symlink and only copy in source dir to stay under 50MB compressed limit for Lambdas. 2023-11-22 23:07:33 -05:00
router.py fix(router.py): set cooldown_time: per model 2024-06-25 16:51:55 -07:00
scheduler.py feat(scheduler.py): support redis caching for req. prioritization 2024-06-06 14:19:21 -07:00
timeout.py refactor: add black formatting 2023-12-25 14:11:20 +05:30
utils.py add fireworks ai param mapping 2024-06-26 06:43:18 -07:00