Azure Speed Test

Prices by deployment type

All rates are USD per 1 million tokens. Each row uses its own deployment type and published price regions. A range reflects regional differences within that type. Missing prices mean no published meter in this snapshot; they do not mean free usage. A retail price does not guarantee deployment availability or quota.

codex mini OpenAI prices in USD per 1 million tokens.
Deployment typeInput / 1MOutput / 1MCached input / 1MCache write / 1MPrice regions
Data Zone Standard $1.65 $6.60 $0.413 Not published 12
Global Standard $1.50 $6.00 $0.375 Not published 23
Regional Standard $1.65 - $1.815 $6.60 - $7.26 $0.413 - $0.454 Not published 8

The tables contain only this OpenAI listing. Other providers, versions and similarly named Azure products can have different meters. Some token categories may be missing in individual regions; expand the regional tables for exact coverage.

How token costs are calculated

For each region and deployment type, multiply each token category by its matching rate, then divide the sum by 1,000,000. Monthly estimates multiply that request cost by the number of requests. Ranges use complete workloads priced in each region, so input and output prices from different regions are never combined.

Request cost = (uncached input tokens × input rate + cached input tokens × cached input rate + cache-write input tokens × cache-write rate + output tokens × output rate) / 1,000,000.

The examples below use zero cache writes. The cached example assumes half of the input tokens are eligible cache reads on every request; actual cache hit rates and any cache creation charges depend on the deployment. Missing cache rates are not replaced with ordinary input rates. Estimates cover these token meters only, before taxes, negotiated discounts and any separate hosting, tools, media or provisioned-capacity charges.

Some model token limits are not published in the metadata. These are arithmetic cost examples; verify that your deployment supports the workload before using it.

Data Zone Standard prices and cost examples

Data Zone processing
Inference data is processed within the Microsoft-specified US, EU, or APAC data zone.
Standard pay-per-token
Input and output usage is billed by token. Standard deployments provide best-effort throughput; reserved PTU pricing is separate.

Example workload: 3,000 input and 1,000 output tokens per request, with 10,000 requests per month.

No cached input

Per request
$0.01155
Per month
$115.50

Across 12 published price regions.

Customize workload or price region for Data Zone Standard, No cached input

50% cached input

Per request
$0.0096945
Per month
$96.945

Across 12 published price regions.

Customize workload or price region for Data Zone Standard, 50% cached input
View 12 region prices for Data Zone Standard

Price region identifies the Azure retail meter; processing location follows the deployment type described above. Dates show the latest effective meter in each row, not when prices were last checked. All token rates are USD per 1 million tokens.

codex mini Data Zone Standard exact regional rates in USD per 1 million tokens.
Azure Public price regionInput / 1MOutput / 1MCached input / 1MCache write / 1MMeter effective dates (UTC)
East USeastus$1.65 $6.60 $0.413 Not published
East US 2eastus2$1.65 $6.60 $0.413 Not published
France Centralfrancecentral$1.65 $6.60 $0.413 Not published
Germany West Centralgermanywestcentral$1.65 $6.60 $0.413 Not published
North Central USnorthcentralus$1.65 $6.60 $0.413 Not published
Poland Centralpolandcentral$1.65 $6.60 $0.413 Not published
South Central USsouthcentralus$1.65 $6.60 $0.413 Not published
Spain Centralspaincentral$1.65 $6.60 $0.413 Not published
Sweden Centralswedencentral$1.65 $6.60 $0.413 Not published
West Europewesteurope$1.65 $6.60 $0.413 Not published
West USwestus$1.65 $6.60 $0.413 Not published
West US 3westus3$1.65 $6.60 $0.413 Not published

Global Standard prices and cost examples

Global processing
Inference data may be processed in any Azure region.
Standard pay-per-token
Input and output usage is billed by token. Standard deployments provide best-effort throughput; reserved PTU pricing is separate.

Example workload: 3,000 input and 1,000 output tokens per request, with 10,000 requests per month.

No cached input

Per request
$0.0105
Per month
$105.00

Across 23 published price regions.

Customize workload or price region for Global Standard, No cached input

50% cached input

Per request
$0.0088125
Per month
$88.125

Across 23 published price regions.

Customize workload or price region for Global Standard, 50% cached input
View 23 region prices for Global Standard

Price region identifies the Azure retail meter; processing location follows the deployment type described above. Dates show the latest effective meter in each row, not when prices were last checked. All token rates are USD per 1 million tokens.

codex mini Global Standard exact regional rates in USD per 1 million tokens.
Azure Public price regionInput / 1MOutput / 1MCached input / 1MCache write / 1MMeter effective dates (UTC)
Australia Eastaustraliaeast$1.50 $6.00 $0.375 Not published
Brazil Southbrazilsouth$1.50 $6.00 $0.375 Not published
Canada Eastcanadaeast$1.50 $6.00 $0.375 Not published
East USeastus$1.50 $6.00 $0.375 Not published
East US 2eastus2$1.50 $6.00 $0.375 Not published
France Centralfrancecentral$1.50 $6.00 $0.375 Not published
Germany West Centralgermanywestcentral$1.50 $6.00 $0.375 Not published
Japan Eastjapaneast$1.50 $6.00 $0.375 Not published
Korea Centralkoreacentral$1.50 $6.00 $0.375 Not published
North Central USnorthcentralus$1.50 $6.00 $0.375 Not published
Norway Eastnorwayeast$1.50 $6.00 $0.375 Not published
Poland Centralpolandcentral$1.50 $6.00 $0.375 Not published
South Africa Northsouthafricanorth$1.50 $6.00 $0.375 Not published
South Central USsouthcentralus$1.50 $6.00 $0.375 Not published
South Indiasouthindia$1.50 $6.00 $0.375 Not published
Spain Centralspaincentral$1.50 $6.00 $0.375 Not published
Sweden Centralswedencentral$1.50 $6.00 $0.375 Not published
Switzerland Northswitzerlandnorth$1.50 $6.00 $0.375 Not published
UAE Northuaenorth$1.50 $6.00 $0.375 Not published
UK Southuksouth$1.50 $6.00 $0.375 Not published
West Europewesteurope$1.50 $6.00 $0.375 Not published
West USwestus$1.50 $6.00 $0.375 Not published
West US 3westus3$1.50 $6.00 $0.375 Not published

Regional Standard prices and cost examples

Single-region processing
Inference data is processed in the deployment region. Microsoft names the pay-per-token deployment type Standard; this directory uses Regional Standard to make its scope explicit.
Standard pay-per-token
Input and output usage is billed by token. Standard deployments provide best-effort throughput; reserved PTU pricing is separate.

Example workload: 3,000 input and 1,000 output tokens per request, with 10,000 requests per month.

No cached input

Per request
$0.01155 - $0.012705
Per month
$115.50 - $127.05

Across 8 published price regions.

Customize workload or price region for Regional Standard, No cached input

50% cached input

Per request
$0.0096945 - $0.0106635
Per month
$96.945 - $106.635

Across 8 published price regions.

Customize workload or price region for Regional Standard, 50% cached input
View 8 region prices for Regional Standard

Price region identifies the Azure retail meter; processing location follows the deployment type described above. Dates show the latest effective meter in each row, not when prices were last checked. All token rates are USD per 1 million tokens.

codex mini Regional Standard exact regional rates in USD per 1 million tokens.
Azure Public price regionInput / 1MOutput / 1MCached input / 1MCache write / 1MMeter effective dates (UTC)
Central UScentralus$1.65 $6.60 $0.413 Not published
East USeastus$1.65 $6.60 $0.413 Not published
East US 2eastus2$1.65 $6.60 $0.413 Not published
North Central USnorthcentralus$1.65 $6.60 $0.413 Not published
South Central USsouthcentralus$1.65 $6.60 $0.413 Not published
Sweden Centralswedencentral$1.815 $7.26 $0.454 Not published
West USwestus$1.65 $6.60 $0.413 Not published
West US 3westus3$1.65 $6.60 $0.413 Not published

Model information and limits

Coding-optimized GPT model for repository edits, reviews, and agentic software work

Metadata model ID
codex-mini
Metadata match
model
Input modalities
text
Output modalities
text
Context window (tokens)
200,000
Maximum input tokens
Not published
Maximum output tokens
100,000
Reasoning
Yes
Tool calling
Yes
Structured output
Not published

Supplemental metadata comes from Models.dev, with verified Azure version corrections where available. Family matches do not establish version-specific limits. Confirm capabilities and lifecycle in the official model catalog.

Price sources and deployment guidance

Rates come from the Azure Retail Prices API. This page covers consumption token meters in Azure Public. Provisioned throughput, training, non-token media meters and sovereign-cloud pricing are separate billing scopes.

Browse all Azure AI model prices