litellm/docs/my-website/docs/completion
Michael-RZ-Berri 4f823cedac
Add supported providers to prompt caching doc (#26124)
* Add supported providers to prompt caching doc

* Move Z.ai / GLM to cache_control marker list

* Mark xAI models as supporting prompt caching

* Narrow xAI prompt caching flag to models with documented cache pricing

* Add prompt caching flag to grok-4, grok-4-0709, grok-4-latest

---------

Co-authored-by: Michael Riad Zaky <michaelr@Michaels-MacBook-Air.local>
2026-04-20 15:25:21 -07:00
..
anthropic_advisor_tool.md docs(advisor): move supported providers to top, focus how it works on litellm native loop 2026-04-11 18:27:18 -07:00
audio.md Fix typos (#10232) 2025-04-23 20:59:25 -07:00
batching.md docs(batching.md): add batch completion fastest response on proxy to docs 2024-05-28 22:14:22 -07:00
computer_use.md Litellm fix update bedrock models (#24947) 2026-04-01 19:22:54 -07:00
document_understanding.md Litellm fix update bedrock models (#24947) 2026-04-01 19:22:54 -07:00
drop_params.md feat: Replace jsonpath-ng with custom minimal parser for additional_drop_params 2025-12-09 17:30:30 +05:30
function_call.md Added compatibility guidance, etc. for xAI Grok model (#8282) 2025-02-05 17:21:47 -08:00
http_handler_config.md docs for custom aiohttp session 2025-09-07 17:42:26 -07:00
image_generation_chat.md feat: Add gemini-3-pro-image-preview model support for imageSize parameters (#17019) 2025-11-25 19:38:29 -08:00
input.md feat: Limit stop sequence as per openai spec 2026-01-22 17:52:13 +05:30
json_mode.md feat(gemini): use responseJsonSchema for Gemini 2.0+ models (#19314) 2026-01-19 10:45:37 -08:00
knowledgebase.md Add Azure AI Search to supported vector stores (#17726) 2025-12-09 09:04:04 -08:00
message_sanitization.md build: migrate packaging, CI, and Docker from Poetry to uv (#25007) 2026-04-09 11:46:23 -07:00
message_trimming.md (docs) update token trimming 2023-10-30 13:56:23 -07:00
mock_requests.md docs update 2023-09-16 08:55:08 -07:00
model_alias.md (docs) update model alias 2023-10-30 13:58:07 -07:00
multiple_deployments.md docs(multiple_deployments.md): docs on how to route between multiple deployments 2023-10-20 14:30:29 -07:00
output.md feat(types): expose native_finish_reason in provider_specific_fields 2026-03-10 18:43:51 -03:00
predict_outputs.md (feat) add Predicted Outputs for OpenAI (#6594) 2024-11-04 21:16:57 -08:00
prefix.md Update prefix.md (#6734) 2024-11-14 11:18:35 +05:30
prompt_caching.md Add supported providers to prompt caching doc (#26124) 2026-04-20 15:25:21 -07:00
prompt_compression.md Prompt Compression - add it to the proxy (#25729) 2026-04-20 15:08:00 -07:00
prompt_formatting.md initial 2024-04-04 16:58:51 -03:00
provider_specific_params.md Litellm fix update bedrock models (#24947) 2026-04-01 19:22:54 -07:00
reliable_completions.md Revert "feat: add retry_delay, exponential_backoff, and jitter to completion(…" 2026-01-20 17:07:00 +05:30
shared_session.md feat: Add shared_session parameter for aiohttp ClientSession reuse 2025-09-19 01:46:53 -07:00
stream.md doc on streaming usage litellm proxy 2024-12-30 21:06:34 -08:00
token_usage.md docs: fix bad examples from sdk (#19322) 2026-01-19 10:27:25 -08:00
usage.md Add Default usage data configuration 2026-02-19 14:04:07 +05:30
vision.md fix img URL for tests 2025-11-22 09:41:15 -08:00
web_fetch.md docs(web_fetch): add newer Claude models to supported models list (#23251) 2026-03-11 19:09:28 +05:30
web_search.md docs(web_search): add gpt-5-search-api usage examples for SDK and AI Gateway (#20616) 2026-02-12 19:58:12 +05:30