litellm/type-discipline-budget.json
mateo-berri 512e730f2f fix(azure_ai): add passthrough config so router-model relays reach the deployment's own endpoint
Every /azure_ai/<router model>/<native path> relay failed with HTTP 500 because
azure_ai had no passthrough config. The new AzureAIPassthroughConfig strips the
router-model prefix from the relayed path, forwards to the deployment's api_base
with its own credential (api-key on Foundry and Azure OpenAI hosts, Bearer
elsewhere, Entra as the fallback), and delegates chat/completions cost logging
to the Azure passthrough config.

The router's provider inference now receives the deployment's api_base so an
OpenAI-family model on a Foundry resource stays azure_ai instead of flipping to
azure through the AZURE_AI_API_BASE env var.
2026-09-04 21:30:48 -07:00

38 lines
437 B
JSON

{
"LIT001": {
"limit": 22325
},
"LIT002": {
"limit": 26748
},
"LIT003": {
"limit": 261
},
"LIT004": {
"limit": 40
},
"LIT005": {
"limit": 0
},
"LIT006": {
"limit": 1038
},
"LIT007": {
"limit": 0
},
"LIT008": {
"limit": 945
},
"LIT009": {
"limit": 0
},
"LIT010": {
"limit": 16468
},
"LIT011": {
"limit": 5512
},
"LIT012": {
"limit": 4487
}
}