litellm/tests/test_litellm/llms/sagemaker
mateo-berri 2ea633d223 fix(sagemaker_chat): send the inference component header and honor hf_model_name
sagemaker_chat never put X-Amzn-SageMaker-Inference-Component on the request, so any endpoint
backed by inference components answered 400 INFERENCE_COMPONENT_NAME_MISSING and the call never
reached the container. The legacy sagemaker provider has built that header from model_id since
#8889, and this brings the chat provider in line. It goes on in validate_environment, which runs
before the request is SigV4-signed, so the signature covers it

The request body also always named the endpoint rather than the served model, which containers
that validate the body's model answer with a 404. hf_model_name now becomes the body's model
when it is set, and endpoints that do not set it keep sending exactly what they send today
2026-08-20 19:57:46 -07:00
..
test_sagemaker_chat_transformation.py fix(sagemaker_chat): send the inference component header and honor hf_model_name 2026-08-20 19:57:46 -07:00
test_sagemaker_common_utils.py test: add six ruff rules that catch tests which cannot fail (#37709) 2026-08-20 14:21:26 -07:00
test_sagemaker_completion_handler.py test(sagemaker): assert make_sync_call maps non-200 to SagemakerError 2026-07-23 04:31:11 +00:00
test_sagemaker_embedding_role_assumption.py style: run black formatter on files from main merge 2026-04-17 13:02:59 -07:00
test_sagemaker_embedding_voyage.py fix(sagemaker): send native Cohere embed payload to Cohere SageMaker endpoints (#28613) 2026-05-22 12:00:42 -07:00
test_sagemaker_nova_transformation.py style: run black formatter on files from main merge 2026-04-17 13:02:59 -07:00