litellm/litellm
Alexsander Hamir b4def8899f
add: shared_session support to responses API (#16260)
This change enables HTTP client session reuse by:
- Adding shared_session parameter to all responses API methods (responses, delete_responses, get_responses, list_input_items, cancel_responses)
- Passing shared_session to get_async_httpx_client() for connection pooling
- Adding debug logging to track shared session usage

This helps reduce memory overhead by reusing HTTP connections instead of creating new clients for each request, which is particularly important for high-throughput proxy scenarios.
2025-11-04 18:53:54 -08:00
..
anthropic_interface [Feat] Allow using litellm.completion with /v1/messages API Spec (use gpt-4, gemini etc with claude code) (#11502) 2025-06-06 20:35:53 -07:00
assistants Contributor PR - Support OPENAI_BASE_URL in addition to OPENAI_API_BASE (#9995) (#10423) 2025-04-29 21:27:37 -07:00
batch_completion (code quality) run ruff rule to ban unused imports (#7313) 2024-12-19 12:33:42 -08:00
batches [Feat] Add support for Batch API Rate limiting - PR1 adds support for input based rate limits (#16075) 2025-10-29 18:28:52 -07:00
caching [Feat] UI - Allow setting cache settings on UI (#16143) 2025-10-31 17:43:59 -07:00
completion_extras fix: fix transformation.py 2025-10-11 13:45:41 -07:00
containers Add E2E Container API Support (#16136) 2025-11-01 14:03:51 -07:00
endpoints/speech/speech_to_completion_bridge fix: fix import errors 2025-09-14 09:32:21 -07:00
experimental_mcp_client fix(client.py): fix rest api tool call 2025-10-09 14:42:04 -07:00
files fix: Convert object to a correct type (#15634) 2025-10-16 21:47:32 -07:00
fine_tuning Litellm managed file updates combined (#11040) 2025-05-22 17:20:41 -07:00
google_genai fix test /generateContent route 2025-10-04 10:49:46 -07:00
images [Feat] Add FAL AI Image Generations on LiteLLM (#16067) 2025-10-29 13:10:51 -07:00
integrations Add Prometheus metric to track callback logging failures in S3 (#16209) 2025-11-03 18:46:52 -08:00
litellm_core_utils fix _handle_callback_failure 2025-11-04 18:00:12 -08:00
llms add: shared_session support to responses API (#16260) 2025-11-04 18:53:54 -08:00
ocr [Feat] /ocr - Add VertexAI OCR provider support + cost tracking (#16216) 2025-11-03 15:56:49 -08:00
passthrough working - errors from bedrock through pass throughs 2025-10-16 15:40:42 -07:00
proxy [Feat] Add Bedrock Agentcore as a provider on LiteLLM Python SDK and LiteLLM AI Gateway (#16252) 2025-11-04 16:35:12 -08:00
realtime_api [Fix] OpenAI Realtime API integration fails due to websockets.exceptions.PayloadTooBig error (#15751) 2025-10-20 15:54:14 -07:00
rerank_api [Feat] Add Nvidia NIM Rerank Support (#15152) 2025-10-02 18:58:52 -07:00
responses add: shared_session support to responses API (#16260) 2025-11-04 18:53:54 -08:00
router_strategy fix: remove router inefficiencies (from O(M*N) to O(1)) - 62.5% faster P99 latency (#15046) 2025-09-29 15:49:46 -07:00
router_utils [Feat] UI - Search Tools, allow adding search tools on UI + testing search (#15871) 2025-10-23 17:59:29 -07:00
search [Bug Fix] Exa Search API - ensure request params are sent to Exa AI (#15855) 2025-10-23 11:56:30 -07:00
secret_managers Add tags and descriptions support to aws secrets manager (#16224) 2025-11-04 16:11:51 -08:00
types [Feat] add serxng search API provider (#16259) 2025-11-04 17:56:07 -08:00
vector_stores (feat) Azure AI Vector Stores - support "virtual" indexes + create vector store on passthrough API (#16160) 2025-11-01 12:01:32 -07:00
videos Add custom_llm_provider support for video endpoints (non-generation) (#16121) 2025-11-01 12:09:11 -07:00
__init__.py Add E2E Container API Support (#16136) 2025-11-01 14:03:51 -07:00
_logging.py [Feat] Add support for returning images with gemini/gemini-2.5-flash-image-preview with /chat/completions (#13983) 2025-08-27 16:16:19 -07:00
_redis.py fix: Apply max_connections configuration to Redis async client (#15797) 2025-10-22 09:19:08 -07:00
_service_logger.py [⚡️ Python SDK import] - reduce python sdk import time by .3s (#12140) 2025-06-28 14:57:10 -07:00
_uuid.py Fix: revert fastuuid optional dependency, always use fastuuid in .__uid helper (#14941) 2025-09-26 09:14:20 -07:00
_version.py Virtual key based policies in Aim Guardrails (#9499) 2025-04-01 21:57:23 -07:00
budget_manager.py Squashed commit of the following: (#9709) 2025-04-02 21:24:54 -07:00
constants.py [Feat] Add Azure AI Doc Intelligence OCR (#16219) 2025-11-03 17:22:19 -08:00
cost.json
cost_calculator.py fix: resolve memory leak caused by Pydantic 2.11+ deprecation warnings (#16110) 2025-11-01 13:16:32 -07:00
exceptions.py build: squash merge litellm_dev_10_10_2025_p1 2025-10-25 12:21:12 -07:00
main.py Update perplexity cost tracking (#15743) 2025-11-03 08:45:34 -08:00
model_prices_and_context_window_backup.json [Feat] add serxng search API provider (#16259) 2025-11-04 17:56:07 -08:00
mypy.ini fix mypy 2025-09-27 12:21:32 -07:00
py.typed
router.py refactor 2025-11-03 17:31:55 -08:00
scheduler.py Squashed commit of the following: (#9709) 2025-04-02 21:24:54 -07:00
timeout.py Litellm ruff linting enforcement (#5992) 2024-10-01 19:44:20 -04:00
utils.py [Feat] add serxng search API provider (#16259) 2025-11-04 17:56:07 -08:00