Tag not monitored by Microsoft.
To get quota so model deployments work in a new Foundry project, use both the Foundry portal quota tools and the quota request form.
- Verify roles on the subscription
- Ensure the account has the required roles on the Azure subscription:
- Cognitive Services Usages Reader to view quota.
- Owner or Contributor to request quota increases.
- Cognitive Services Contributor + Cognitive Services Usages Reader to edit allocations in Foundry.
- Without these, the Quota page may show no allocations or the Request quota button may be disabled.
- Ensure the account has the required roles on the Azure subscription:
- Check and manage quota in Foundry (new)
- Go to Microsoft Foundry and ensure New Foundry is enabled.
- Select the correct project (upper-left corner).
- Select Operate (top navigation).
- Select Quota in the left pane.
- Use the Token per minute tab to view TPM quota for standard deployments.
- If any deployments exist that are using quota, select them, open the details pane, and in Affiliated deployments using shared quota use the pencil icon to reduce or reallocate TPM so that quota is freed for the new gpt‑4.1 / gpt‑4.1‑mini deployments.
- Request more quota from within Foundry
- On the Quota page for the project, select Request quota (upper-right corner).
- Submit a request for the models and regions where gpt‑4.1 and gpt‑4.1‑mini are needed.
- Allow up to ~15 minutes after approval for changes to propagate, then refresh the Quota page and retry deployment.
- Request quota via the central form (if no quota at all)
- If the subscription has 0 quota for these models in all regions (not just unallocated), submit the global quota request form: quota increase request form.
- Use this to request quota for Azure OpenAI models (including gpt‑4.1 and gpt‑4.1‑mini) in the desired regions.
- Quota requests are processed in order and are more likely to be approved if existing quota (if any) is actively used.
- If quota UI shows issues
- If the Quota page is empty or shows no allocations even after roles are confirmed, verify the correct subscription is selected.
- If the Request quota button is disabled, confirm Owner or Contributor role and that the model/region supports quota increases.
- If quota changes are approved but not visible after 24 hours, open an Azure support ticket as advised in the troubleshooting guidance.
Once quota is granted and visible on the Quota page, retry deploying gpt‑4.1 and gpt‑4.1‑mini in a region where TPM is available.
References:
- Manage and increase quotas for resources with Microsoft Foundry (Foundry projects)
- View and request quotas in Foundry portal
- Manage Azure OpenAI in Microsoft Foundry Models quota
- Deploy Microsoft Foundry Models in the Foundry portal
- Azure OpenAI in Microsoft Foundry Models quotas and limits
- How to troubleshoot your deployments and monitors in Microsoft Foundry portal (classic)
- Quota shows as fully utilised for Australia East region for Standard deployment - Microsoft Q&A