-<description><h3>Azure OpenAI</h3></p><p>- <strong>GPT-5.4</strong> meters were added for <code>gpt-5.4</code>, <code>gpt-5.4-mini</code>, <code>gpt-5.4-nano</code>, and <code>gpt-5.4-pro</code> in Flex Global token billing, covering Input, Cached Input, Output, and for some variants LongCo pricing. Prices start from $0.0375/1M tokens for cached input on <code>gpt-5.4-mini</code>, $0.1/1M tokens for <code>gpt-5.4-nano</code> input, $0.375/1M tokens for <code>gpt-5.4-mini</code> input, $1.25/1M tokens for <code>gpt-5.4</code> input, $15/1M tokens for <code>gpt-5.4-pro</code> input, and up to $135/1M tokens for <code>gpt-5.4-pro</code> LongCo output.<br>- <strong>GPT-5.5</strong> meters were added for <code>gpt-5.5</code> with ShortCo and LongCo Flex Global Input and Cached Input pricing. Prices start from $0.25/1M tokens for cached input, $2.5/1M tokens for ShortCo input, and $5/1M tokens for LongCo input.<br>- <strong>GPT-5.6 luna</strong> meters were added in both Flex Global and Std DZ forms, with ShortCo and LongCo variants across Input, Cached Input, Cached Write, and Output. Prices start from $0.0275/1M tokens for ShortCo Cached Input in Std DZ, $0.1/1M tokens for ShortCo input in Flex Global, $0.2/1M tokens for LongCo input in Flex Global, $0.25/1M tokens for LongCo Cached Write in Flex Global, $0.55/1M tokens for LongCo input in Std DZ, $0.9/1M tokens for LongCo output in Flex Global, and $2.475/1M tokens for LongCo output in Std DZ.<br>- <strong>GPT-5.6 terra</strong> meters were added in both Flex Global and Std DZ forms, with ShortCo and LongCo variants across Input, Cached Input, Cached Write, and Output. Prices start from $0.1/1M tokens for ShortCo Cached Input in Flex Global, $0.2/1M tokens for LongCo Cached Input in Flex Global, $1/1M tokens for ShortCo input in Flex Global, $2/1M tokens for LongCo input in Flex Global, $2.5/1M tokens for LongCo Cached Write in Flex Global, $2.75/1M tokens for ShortCo input in Std DZ, $6/1M tokens for ShortCo output in Flex Global, $9/1M tokens for LongCo output in Flex Global, $16.5/1M tokens for ShortCo output in Std DZ, and $24.75/1M tokens for LongCo output in Std DZ.<br>- <strong>GPT-5.6 sol</strong> meters were added in both Flex Global and Std DZ forms, with ShortCo and LongCo variants across Input, Cached Input, Cached Write, and Output. Prices start from $0.25/1M tokens for ShortCo Cached Input in Flex Global, $0.5/1M tokens for LongCo Cached Input in Flex Global, $2.5/1M tokens for ShortCo input in Flex Global, $3.125/1M tokens for ShortCo Cached Write in Flex Global, $5/1M tokens for LongCo input in Flex Global, $6.875/1M tokens for ShortCo input in Std DZ, $15/1M tokens for ShortCo output in Flex Global, $22.5/1M tokens for LongCo output in Flex Global, $41.25/1M tokens for ShortCo output in Std DZ, and $61.875/1M tokens for LongCo output in Std DZ.<br>- <strong>Provisioned Managed</strong> Azure OpenAI meters were added for both Data Zone and Regional deployment types, with prices from $1.2/hour for Data Zone and from $2.28/hour for Regional.</p><p>Documentation links:<br>- <a href="https://learn.microsoft.com/azure/foundry/foundry-models/concepts/models-sold-directly-by-azure#azure-openai-in-microsoft-foundry-models">Azure OpenAI models in Foundry, including GPT-5.4, GPT-5.5, and GPT-5.6 families</a><br>- <a href="https://learn.microsoft.com/azure/foundry/openai/concepts/provisioned-throughput">Provisioned throughput for Foundry Models</a></p><p><h3>Azure DeepSeek</h3></p><p>- <strong>Provisioned Managed</strong> Azure DeepSeek meters were added for Data Zone and Regional deployment types. Prices are $1.2/hour for Data Zone and from $2.28/hour for Regional.</p><p>Documentation links:<br>- <a href="https://learn.microsoft.com/azure/foundry/foundry-models/concepts/models-sold-directly-by-azure?pivots=azure-direct-others">Foundry Models sold by Azure - other model collections, including DeepSeek</a><br>- <a href="https://learn.microsoft.com/azure/foundry/openai/concepts/provisioned-throughput">Provisioned throughput for Foundry Models</a></p><p><h3>Azure Grok</h3></p><p>- <strong>Provisioned Managed</strong> Azure Grok meters were added for Data Zone and Regional deployment types. Prices are $1.2/hour for Data Zone and from $2.28/hour for Regional.</p><p>Documentation links:<br>- <a href="https://learn.microsoft.com/azure/foundry/foundry-models/concepts/models-sold-directly-by-azure?pivots=azure-direct-others">Foundry Models sold by Azure - other model collections, including xAI Grok</a><br>- <a href="https://learn.microsoft.com/azure/foundry/openai/concepts/provisioned-throughput">Provisioned throughput for Foundry Models</a></p><p><h3>Azure Llama</h3></p><p>- <strong>Provisioned Managed</strong> Azure Llama meters were added for Data Zone and Regional deployment types. Prices are $1.2/hour for Data Zone and from $2.28/hour for Regional.</p><p>Documentation links:<br>- <a href="https://learn.microsoft.com/azure/foundry/foundry-models/concepts/models-sold-directly-by-azure?pivots=azure-direct-others">Foundry Models sold by Azure - other model collections, including Meta Llama</a><br>- <a href="https://learn.microsoft.com/azure/foundry/openai/concepts/provisioned-throughput">Provisioned throughput for Foundry Models</a></p><p><h3>Azure Mistral</h3></p><p>- <strong>Provisioned Managed</strong> Azure Mistral meters were added for Data Zone and Regional deployment types. Prices are $1.2/hour for Data Zone and from $2.28/hour for Regional.</p><p>Documentation links:<br>- <a href="https://learn.microsoft.com/azure/foundry/foundry-models/concepts/models-sold-directly-by-azure?pivots=azure-direct-others">Foundry Models sold by Azure - other model collections, including Mistral AI</a><br>- <a href="https://learn.microsoft.com/azure/foundry/openai/concepts/provisioned-throughput">Provisioned throughput for Foundry Models</a></p><p><h3>Azure BFL Flux</h3></p><p>- <strong>Provisioned Managed</strong> Azure BFL Flux meters were added for Data Zone and Regional deployment types. Prices are $1.2/hour for Data Zone and from $2.28/hour for Regional.</p><p>Documentation links:<br>- <a href="https://learn.microsoft.com/azure/foundry/foundry-models/how-to/use-foundry-models-flux">Deploy and use FLUX models in Microsoft Foundry</a><br>- <a href="https://learn.microsoft.com/azure/foundry/foundry-models/concepts/models-sold-directly-by-azure?pivots=azure-direct-others">Foundry Models sold by Azure - other model collections, including Black Forest Labs</a></p><p><h3>Azure Fireworks</h3></p><p>- <strong>Provisioned Managed</strong> Azure Fireworks meters were added for Data Zone deployment at $1.2/hour. No Regional Fireworks meter appears in this file.</p><p>Documentation links:<br>- <a href="https://learn.microsoft.com/azure/foundry/how-to/fireworks/enable-fireworks-models">Fireworks models on Microsoft Foundry</a><br>- <a href="https://learn.microsoft.com/azure/foundry/openai/concepts/provisioned-throughput">Provisioned throughput for Foundry Models</a></p><p><h3>Azure AI Foundry Provisioned Throughput Reservation</h3></p><p>- <strong>Provisioned Throughput Reservation</strong> meters were added for Data Zone and Regional reservation billing. Data Zone prices are from $312/hour and $3180/hour. Regional prices are from $326/hour, with additional listed reservation meters up to $3972/hour.</p><p>Documentation links:<br>- <a href="https://learn.microsoft.com/azure/foundry/openai/concepts/provisioned-throughput">Provisioned throughput for Foundry Models</a><br>- <a href="https://learn.microsoft.com/en-us/azure/cost-management-billing/reservations/azure-openai">Save costs with Microsoft Azure AI Foundry Provisioned Throughput Reservations</a></p><p>_This summary was AI-generated in 28 seconds on 1 September 2026 using gpt-5.4.<br>It may contain mistakes, or outdated pricing data. Always use the Azure Retail Prices API for live pricing. The existence of a price meter does not always imply model/service availability. Prices vary depending on different factors, including region.<br>Input tokens: 24,453 | Output tokens: 2,037 | Total tokens: 26,490._</description>
0 commit comments