litellm/litellm
Krrish Dholakia 16da21e839 feat(llm_cost_calc/google.py): do character based cost calculation for vertex ai
Calculate cost for vertex ai responses using characters in query/response

 Closes https://github.com/BerriAI/litellm/issues/4165
2024-06-19 17:18:42 -07:00
..
assistants feat(assistants/main.py): support arun_thread_stream 2024-06-04 16:47:51 -07:00
batches docs(customers.md): add customer cost tracking to docs 2024-05-29 14:55:33 -07:00
deprecated_litellm_server refactor: add black formatting 2023-12-25 14:11:20 +05:30
integrations Merge pull request #4275 from BerriAI/litellm_fix_langfuse_log_prompts 2024-06-18 20:09:13 -07:00
litellm_core_utils feat(llm_cost_calc/google.py): do character based cost calculation for vertex ai 2024-06-19 17:18:42 -07:00
llms fix(vertex_httpx.py): fix supports system message check for vertex_ai_beta 2024-06-19 13:17:22 -07:00
proxy Merge pull request #4286 from BerriAI/litellm_support_options_health_endpoints 2024-06-19 12:25:32 -07:00
router_strategy refactor: replace 'traceback.print_exc()' with logging library 2024-06-06 13:47:43 -07:00
router_utils fix use safe access for router alerting 2024-06-14 15:17:32 -07:00
tests feat(llm_cost_calc/google.py): do character based cost calculation for vertex ai 2024-06-19 17:18:42 -07:00
types feat(llm_cost_calc/google.py): do character based cost calculation for vertex ai 2024-06-19 17:18:42 -07:00
__init__.py Merge branch 'main' into litellm_gemini_refactoring 2024-06-17 19:50:56 -07:00
_logging.py Merge branch 'main' into litellm_gemini_refactoring 2024-06-17 19:50:56 -07:00
_redis.py feat(proxy_server.py): return litellm version in response headers 2024-05-08 16:00:08 -07:00
_service_logger.py feat - working exception logs for Redis errors 2024-06-07 16:30:29 -07:00
_version.py (fix) ci/cd don't let importing litellm._version block starting proxy 2024-02-01 16:23:16 -08:00
budget_manager.py feat(proxy_server.py): return litellm version in response headers 2024-05-08 16:00:08 -07:00
caching.py fix(caching.py): Stop throwing constant spam errors on every single S3 cache miss. Fixes #4146. 2024-06-13 21:13:29 -07:00
cost.json store llm costs in budget manager 2023-09-09 19:11:35 -07:00
cost_calculator.py feat(llm_cost_calc/google.py): do character based cost calculation for vertex ai 2024-06-19 17:18:42 -07:00
exceptions.py feat(router.py): support content policy fallbacks 2024-06-14 17:15:44 -07:00
main.py fix(main.py): route openai calls to /completion when text_completion is True 2024-06-19 12:37:05 -07:00
model_prices_and_context_window_backup.json build(model_prices_and_context_window.json): fix gemini pricing 2024-06-19 14:14:44 -07:00
py.typed feature - Types for mypy - #360 2024-05-30 14:14:41 -04:00
requirements.txt Add symlink and only copy in source dir to stay under 50MB compressed limit for Lambdas. 2023-11-22 23:07:33 -05:00
router.py fix(router.py): support multiple orgs in 1 model definition 2024-06-18 19:36:58 -07:00
scheduler.py feat(scheduler.py): support redis caching for req. prioritization 2024-06-06 14:19:21 -07:00
timeout.py refactor: add black formatting 2023-12-25 14:11:20 +05:30
utils.py feat(llm_cost_calc/google.py): do character based cost calculation for vertex ai 2024-06-19 17:18:42 -07:00