litellm/litellm/llms
Ishaan Jaff ca79ebe7d2 UI - Fix regression where Guardrail Entity Could not be selected and entity was not displayed (#16165)
* fix PiiEntityCategoryMap

* fix OpenAIChatCompletionsHandler

* fix lint
2025-11-01 18:02:13 -07:00
..
ai21/chat (code quality) run ruff rule to ban unused imports (#7313) 2024-12-19 12:33:42 -08:00
aiml fixes: AI/ML API 2025-10-16 17:52:37 -07:00
aiohttp_openai/chat VertexAI non-jsonl file storage support (#9781) 2025-04-09 14:01:48 -07:00
anthropic get_computer_tool_beta_header 2025-10-31 20:38:30 -07:00
azure [Fix] Azure OpenAI - Add handling for v1 under azure api versions (#15984) 2025-10-27 13:45:44 -07:00
azure_ai fix(managed_files.py): don't raise error if managed object is not found + (Feat) Azure AI - Search Vector Stores + (Fix) Batches - “User default_user_id does not have access to the object” when object not in db + (fix) Vector Stores - show config.yaml vector stores on UI (#15873) 2025-10-25 12:06:24 -07:00
base_llm Guardrails - Responses API, Image Gen, Text completions, Audio transcriptions, Audio Speech, Rerank, Anthropic Messages API support via the unified apply_guardrails function (#15706) 2025-10-25 13:38:57 -07:00
baseten lint 2025-08-19 13:36:41 -07:00
bedrock Allow using ARNs when generation images via Bedrock (#15789) 2025-10-28 19:41:35 -07:00
bytez ruff check ./litellm --fix 2025-07-12 11:04:02 -07:00
cerebras (code quality) run ruff rule to ban unused imports (#7313) 2024-12-19 12:33:42 -08:00
clarifai/chat UI - add arize on ui, LLMs - clarifai refactor to openai compatible route, added azure ai/grok-4 model family 2025-10-16 20:39:15 -07:00
cloudflare/chat VertexAI non-jsonl file storage support (#9781) 2025-04-09 14:01:48 -07:00
codestral/completion Codestral - return litellm latency overhead on /v1/completions + Add '__contains__' support for ChatCompletionDeltaToolCall (#10879) 2025-05-27 16:13:44 -07:00
cohere Add OpenAI-compatible annotations support for Cohere v2 citations 2025-10-29 19:12:17 -07:00
cometapi feat(cometapi): Add CometAPI provider support (embeddings, image generation, docs) 2025-10-16 13:08:14 +08:00
compactifai Fix CompactifAI provider tests and implementation 2025-09-15 22:03:42 +02:00
custom_httpx Add Add per model group header forwarding for Bedrock Invoke API (#16042) 2025-10-30 20:10:17 -07:00
dashscope [Fixes] Using Qwen API Tiered Pricing (#14479) 2025-09-11 20:07:41 -07:00
databricks Feat: Allow prompt caching to be used for Anthropic Claude on Databricks (#15801) 2025-10-22 09:11:04 -07:00
dataforseo/search [Feat] UI - Search Tools, allow adding search tools on UI + testing search (#15871) 2025-10-23 17:59:29 -07:00
datarobot/chat Updated URL handling for DataRobot provider base 2025-08-21 19:42:46 -06:00
deepgram feat: handle Deepgram detected language when available (#16093) 2025-10-30 19:19:34 -07:00
deepinfra [Feat] Add Nvidia NIM Rerank Support (#15152) 2025-10-02 18:58:52 -07:00
deepseek Litellm staging 05 10 2025 - openai pdf url support + sagemaker chat content length error fix (#10724) 2025-05-10 17:41:57 -07:00
deprecated_providers build(pyproject.toml): add new dev dependencies - for type checking (#9631) 2025-03-29 11:02:13 -07:00
elevenlabs [Feat] Add Eleven Labs - Speech To Text Support on LiteLLM (#12119) 2025-06-27 17:50:49 -07:00
empower/chat LiteLLM Common Base LLM Config (pt.3): Move all OAI compatible providers to base llm config (#7148) 2024-12-10 17:12:42 -08:00
exa_ai/search [Feat] UI - Search Tools, allow adding search tools on UI + testing search (#15871) 2025-10-23 17:59:29 -07:00
fal_ai [Feat] Add FAL AI Image Generations on LiteLLM (#16067) 2025-10-29 13:10:51 -07:00
featherless_ai/chat fix: fix model param mapping 2025-05-19 20:24:47 -07:00
fireworks_ai Don't add "accounts/fireworks/models" prefix for Fireworks Provider (#15938) 2025-10-30 20:36:42 -07:00
friendliai/chat (code quality) run ruff rule to ban unused imports (#7313) 2024-12-19 12:33:42 -08:00
galadriel/chat (code quality) run ruff rule to ban unused imports (#7313) 2024-12-19 12:33:42 -08:00
gemini [Oct Staging Branch] (#15460) 2025-10-17 17:52:25 -07:00
github/chat (code quality) run ruff rule to ban unused imports (#7313) 2024-12-19 12:33:42 -08:00
github_copilot Merge branch 'main' into feat/github-copilot-thinking-reasoning-support 2025-08-27 22:09:15 -07:00
google_pse/search [Feat] UI - Search Tools, allow adding search tools on UI + testing search (#15871) 2025-10-23 17:59:29 -07:00
gradient_ai/chat Add digitalocean provider (#12169) 2025-08-09 16:26:33 -07:00
groq [Feat] Support reasoning_effort in Groq (#14207) 2025-09-03 10:43:47 -07:00
heroku/chat fixes linter error 2025-07-28 09:44:02 -06:00
hosted_vllm [Feat] Add Nvidia NIM Rerank Support (#15152) 2025-10-02 18:58:52 -07:00
huggingface [Feat] Add Nvidia NIM Rerank Support (#15152) 2025-10-02 18:58:52 -07:00
hyperbolic Revert "Litellm dev 07 21 2025 p1 (#12848)" 2025-07-22 18:28:36 -07:00
infinity Merge pull request #14764 from daily-kim/litellm_fix_bearer_capitalization 2025-10-03 22:02:46 -07:00
jina_ai fix: transform_rerank_response 2025-10-04 08:55:48 -07:00
lambda_ai Revert "Litellm dev 07 21 2025 p1 (#12848)" 2025-07-22 18:28:36 -07:00
lemonade Removing unecessary import 2025-09-30 12:12:24 -06:00
litellm_proxy Add native Responses API support for litellm_proxy provider (#15347) 2025-10-08 18:31:26 -07:00
llamafile/chat Add llamafile as a provider (#10203) (#10482) 2025-05-01 18:36:55 -07:00
lm_studio fix(lm_studio): resolve illegal Bearer header value issue 2025-09-12 22:41:30 +02:00
meta_llama/chat [Feat] Enable Tool Calling for meta_llama (#11895) 2025-06-19 13:44:22 -07:00
mistral [Feat] Native /ocr endpoint support (#15573) 2025-10-15 17:20:01 -07:00
moonshot/chat Fix MoonshotChatConfig to address limitations of kimi-thinking-preview model by excluding additional parameters (#12772) 2025-07-19 15:10:57 -07:00
morph fix morph api tests 2025-07-22 18:44:44 -07:00
nebius Integration with Nebius AI Studio added (#11143) 2025-05-27 11:05:22 -07:00
nlp_cloud VertexAI non-jsonl file storage support (#9781) 2025-04-09 14:01:48 -07:00
novita/chat Add new model provider Novita AI (#7582) (#9527) 2025-05-12 21:49:30 -07:00
nscale/chat Add nscale support for streaming (#10698) 2025-05-09 11:39:23 -07:00
nvidia_nim [Feat] Add Nvidia NIM Rerank Support (#15152) 2025-10-02 18:58:52 -07:00
oci Add OCI Signer Authentication. Closes #16048, Closes #15654 (#16064) 2025-10-30 19:59:01 -07:00
ollama fix(ollama): Enhance chunk parsing for empty responses without 'thinking' and improve error logging (#13333) (#15717) 2025-10-21 16:59:01 -07:00
oobabooga VertexAI non-jsonl file storage support (#9781) 2025-04-09 14:01:48 -07:00
openai UI - Fix regression where Guardrail Entity Could not be selected and entity was not displayed (#16165) 2025-11-01 18:02:13 -07:00
openai_like [Bug] Fix: Vertex Mistral not working for streaming (#13952) 2025-08-25 17:39:40 -07:00
openrouter direct cost calculation from openrouter 2025-10-14 13:57:39 -07:00
ovhcloud feat: Add OVHCloud AI Endpoints as a provider 2025-09-12 13:37:03 +02:00
parallel_ai/search [Feat] UI - Search Tools, allow adding search tools on UI + testing search (#15871) 2025-10-23 17:59:29 -07:00
perplexity [Feat] UI - Search Tools, allow adding search tools on UI + testing search (#15871) 2025-10-23 17:59:29 -07:00
petals VertexAI non-jsonl file storage support (#9781) 2025-04-09 14:01:48 -07:00
pg_vector/vector_stores [Feat] UI Vector Stores - Allow adding Vertex RAG Engine, OpenAI, Azure (#12752) 2025-07-18 18:25:26 -07:00
predibase VertexAI non-jsonl file storage support (#9781) 2025-04-09 14:01:48 -07:00
recraft [Feat] Add Google AI Studio Imagen4 model family (#13065) 2025-07-28 21:25:40 -07:00
replicate VertexAI non-jsonl file storage support (#9781) 2025-04-09 14:01:48 -07:00
sagemaker [Oct Staging Branch] (#15460) 2025-10-17 17:52:25 -07:00
sambanova Feat/sambanova embeddings (#13308) 2025-08-12 17:15:26 -07:00
snowflake feat(snowflake): add function calling support for Snowflake Cortex REST API 2025-10-05 13:08:33 +02:00
tavily/search [Feat] UI - Search Tools, allow adding search tools on UI + testing search (#15871) 2025-10-23 17:59:29 -07:00
together_ai fix: use fastuuid helper (#14903) 2025-09-25 15:47:01 -07:00
topaz Add /vllm/* and /mistral/* passthrough endpoints (adds support for Mistral OCR via passthrough) 2025-04-14 22:06:33 -07:00
triton fix(triton/completion/transformation.py): remove bad_words / stop wor… (#10163) 2025-04-19 11:23:37 -07:00
v0 feat: add v0 provider support (#12751) 2025-07-18 18:26:44 -07:00
vercel_ai_gateway Merge branch 'main' into add-vercel-ai-gateway-provider 2025-07-31 18:49:03 -07:00
vertex_ai Changes to fix frequency_penalty and presence_penalty issue for gemini-2.5-pro model (#16041) 2025-10-30 20:02:58 -07:00
vllm fix vllm passthrough 2025-09-22 10:09:02 -03:00
volcengine fix bug 2025-09-15 17:20:30 +08:00
voyage/embedding Add support for voyage-context-3 embedding model 2025-08-22 00:15:12 +05:30
wandb (feat): Add W&B Inference to LiteLLM 2025-09-11 00:07:30 +05:30
watsonx [Fix] Watsonx - Apply correct prompt templates for openai/gpt-oss model family (#15341) 2025-10-08 15:39:36 -07:00
xai Add Xai websearch cost (#16001) 2025-10-30 20:35:34 -07:00
xinference/image_generation [Feat] Add XInference Image Generation API Provider (#12439) 2025-07-08 21:17:38 -07:00
__init__.py Add Xai websearch cost (#16001) 2025-10-30 20:35:34 -07:00
base.py Gemini - web search cost tracking + Update max output tokens for nova models 2025-06-05 23:25:18 -07:00
custom_llm.py Add bridge for /chat/completion -> /responses API (#11632) 2025-06-11 22:20:18 -07:00
maritalk.py build(pyproject.toml): add new dev dependencies - for type checking (#9631) 2025-03-29 11:02:13 -07:00
ollama_chat.py fix: add 'think' parameter handling in ollama_chat.py 2025-10-14 13:57:39 -07:00
README.md LiteLLM Minor Fixes and Improvements (09/13/2024) (#5689) 2024-09-14 10:02:55 -07:00

File Structure

August 27th, 2024

To make it easy to see how calls are transformed for each model/provider:

we are working on moving all supported litellm providers to a folder structure, where folder name is the supported litellm provider name.

Each folder will contain a *_transformation.py file, which has all the request/response transformation logic, making it easy to see how calls are modified.

E.g. cohere/, bedrock/.