mirror of
https://github.com/BerriAI/litellm.git
synced 2026-10-08 03:08:45 +00:00
Every /azure_ai/<router model>/<native path> relay failed with HTTP 500 because azure_ai had no passthrough config. The new AzureAIPassthroughConfig strips the router-model prefix from the relayed path, forwards to the deployment's api_base with its own credential (api-key on Foundry and Azure OpenAI hosts, Bearer elsewhere, Entra as the fallback), and delegates chat/completions cost logging to the Azure passthrough config. The router's provider inference now receives the deployment's api_base so an OpenAI-family model on a Foundry resource stays azure_ai instead of flipping to azure through the AZURE_AI_API_BASE env var.
38 lines
437 B
JSON
38 lines
437 B
JSON
{
|
|
"LIT001": {
|
|
"limit": 22325
|
|
},
|
|
"LIT002": {
|
|
"limit": 26748
|
|
},
|
|
"LIT003": {
|
|
"limit": 261
|
|
},
|
|
"LIT004": {
|
|
"limit": 40
|
|
},
|
|
"LIT005": {
|
|
"limit": 0
|
|
},
|
|
"LIT006": {
|
|
"limit": 1038
|
|
},
|
|
"LIT007": {
|
|
"limit": 0
|
|
},
|
|
"LIT008": {
|
|
"limit": 945
|
|
},
|
|
"LIT009": {
|
|
"limit": 0
|
|
},
|
|
"LIT010": {
|
|
"limit": 16468
|
|
},
|
|
"LIT011": {
|
|
"limit": 5512
|
|
},
|
|
"LIT012": {
|
|
"limit": 4487
|
|
}
|
|
}
|