An Azure service that provides access to OpenAI’s GPT-3 models with enterprise capabilities.
Hello Ashika Hussain
I see that slider jumping from 1K directly to approximately 2.01M is not documented by Microsoft as expected behavior. I would treat it as a Foundry portal/UI issue unless the backend APIs show otherwise.
For a Standard Azure OpenAI deployment, Microsoft documents that:
-
sku.capacity = 1means 1,000 TPM -
sku.capacity = 100means 100K TPM -
sku.capacity = 200means 200K TPM
So you can bypass the slider with Azure CLI:
az cognitiveservices account deployment create -g <RG> -n <resource> --deployment-name <deployment> --model-name gpt-4.1-mini --model-version "2025-04-14" --model-format OpenAI --sku-name Standard --sku-capacity 100
Microsoft documentation: Automate Azure OpenAI deployments with quota
Having 200K quota available does not necessarily mean 200K deployment capacity is currently available in that region. Microsoft now exposes these separately:
- Usages API → assigned quota and current consumption.
- Model Capacities API → how much capacity can actually be deployed for that model, version, region and SKU.
You can check capacity with:
GET https://management.azure.com/subscriptions/<subscription-id>/providers/Microsoft.CognitiveServices/modelCapacities?api-version=2024-10-01&modelFormat=OpenAI&modelName=gpt-4.1-mini&modelVersion=2025-04-14
Official documentation: Manage Azure OpenAI quota and check capacity
Microsoft introduced quota tiers and subscription-level quota changes in 2026, so generic quota values in documentation might differ from the allocation shown for a specific subscription.
So using CLI/REST with sku-capacity 100 or 200 is the supported workaround. If the APIs show at least that much quota and deployment capacity but creation still fails or the portal continues jumping from 1K to 2.01M capture the API response/error and open an Azure Support request, as that would indicate a portal or backend quota/capacity inconsistency rather than an expected slider restriction.