mirror of
https://github.com/BerriAI/litellm.git
synced 2026-09-20 00:11:50 +00:00
* feat(pass_through_endpoints/): support logging anthropic/gemini pass through calls to langfuse/s3/etc. * fix(utils.py): allow disabling end user cost tracking with new param Allows proxy admin to disable cost tracking for end user - keeps prometheus metrics small * docs(configs.md): add disable_end_user_cost_tracking reference to docs * feat(key_management_endpoints.py): add support for restricting access to `/key/generate` by team/proxy level role Enables admin to restrict key creation, and assign team admins to handle distributing keys * test(test_key_management.py): add unit testing for personal / team key restriction checks * docs: add docs on restricting key creation * docs(finetuned_models.md): add new guide on calling finetuned models * docs(input.md): cleanup anthropic supported params Closes https://github.com/BerriAI/litellm/issues/6856 * test(test_embedding.py): add test for passing extra headers via embedding * feat(cohere/embed): pass client to async embedding * feat(rerank.py): add `/v1/rerank` if missing for cohere base url Closes https://github.com/BerriAI/litellm/issues/6844 * fix(main.py): pass extra_headers param to openai Fixes https://github.com/BerriAI/litellm/issues/6836 * fix(litellm_logging.py): don't disable global callbacks when dynamic callbacks are set Fixes issue where global callbacks - e.g. prometheus were overriden when langfuse was set dynamically * fix(handler.py): fix linting error * fix: fix typing * build: add conftest to proxy_admin_ui_tests/ * test: fix test * fix: fix linting errors * test: fix test * fix: fix pass through testing |
||
|---|---|---|
| .. | ||
| AI21 | ||
| anthropic | ||
| azure_ai | ||
| AzureOpenAI | ||
| bedrock | ||
| cerebras | ||
| cohere | ||
| custom_httpx | ||
| databricks | ||
| deepseek/chat | ||
| files_apis | ||
| fine_tuning_apis | ||
| fireworks_ai | ||
| groq | ||
| hosted_vllm/chat | ||
| huggingface_llms_metadata | ||
| jina_ai | ||
| lm_studio | ||
| mistral | ||
| nvidia_nim | ||
| OpenAI | ||
| openai_like | ||
| perplexity/chat | ||
| prompt_templates | ||
| sagemaker | ||
| sambanova | ||
| together_ai | ||
| tokenizers | ||
| vertex_ai_and_google_ai_studio | ||
| watsonx | ||
| xai/chat | ||
| __init__.py | ||
| aleph_alpha.py | ||
| azure_text.py | ||
| base.py | ||
| base_aws_llm.py | ||
| baseten.py | ||
| clarifai.py | ||
| cloudflare.py | ||
| custom_llm.py | ||
| gemini.py | ||
| huggingface_restapi.py | ||
| maritalk.py | ||
| nlp_cloud.py | ||
| ollama.py | ||
| ollama_chat.py | ||
| oobabooga.py | ||
| openrouter.py | ||
| palm.py | ||
| petals.py | ||
| predibase.py | ||
| README.md | ||
| replicate.py | ||
| text_completion_codestral.py | ||
| triton.py | ||
| vllm.py | ||
| volcengine.py | ||
File Structure
August 27th, 2024
To make it easy to see how calls are transformed for each model/provider:
we are working on moving all supported litellm providers to a folder structure, where folder name is the supported litellm provider name.
Each folder will contain a *_transformation.py file, which has all the request/response transformation logic, making it easy to see how calls are modified.
E.g. cohere/, bedrock/.