An Azure service that provides access to OpenAI’s GPT-3 models with enterprise capabilities.
Hi pipi,
To request a significantly higher quota such as 150M TPM for gpt-4.1-mini, the correct process is to submit a support request through the Azure Portal by selecting Service and subscription limits (quotas) → Azure OpenAI Service, and then choosing the specific deployment region. In the details of your request, please include your current usage limits with OpenAI, the intended use case for Azure as a backup environment, and any supporting information on expected sustained throughput requirements. Requests of this scale require review by our Capacity and Quota Management team and may also require coordination with your Microsoft account team. If you do not already have a dedicated Microsoft representative, our support engineers can help facilitate this escalation.
Please note that quotas of this size are typically approved only for customers on Enterprise-level agreements (Enterprise Agreement or Microsoft Customer Agreement – Enterprise tier) and may only be available in regions with sufficient capacity, such as East US, South Central US, West Europe, or Sweden Central. Capacity availability can vary by region, so part of the review process will involve confirming that the requested throughput can be supported in your chosen deployment location.
For reference:
Azure OpenAI in Azure AI Foundry Models quotas and limits
Manage Azure OpenAI in Azure AI Foundry Models quota
Create an Azure support request