OpenAI Pricing
OpenAI charges from $0.140 per 1 million output tokens, depending on the model. gpt-oss-20b is the cheapest at $0.140/1M output and $0.030/1M input; its current flagship GPT-5.5 Pro costs $180.00/1M output. The largest context window is 2.0M tokens (gpt-5.4 (>272K context length)). Prices are USD, current as of August 11, 2026.
OpenAI publishes separate input and output rates per model, with cached input billed at a large discount. Reasoning models bill their internal thinking as output tokens, so a reasoning call can cost several times a standard completion on the same visible response length.
Last updated · synced weekly from the upstream model catalog
Cheapest input
$0.030/1M
gpt-oss-20b
Cheapest output
$0.140/1M
gpt-oss-20b
Longest context
2.0M
gpt-5.4 (>272K context length)
Models priced
95
Avg $29.47/1M output
Pricing by model
All prices in USD per 1 million tokens. Showing 80 of 95 priced models, cheapest first. The batch column is the asynchronous batch-API rate, 50% below the standard output rate on the 35 models that publish one.
| Model | Input / 1M | Output / 1M | Batch out / 1M | Context |
|---|---|---|---|---|
| gpt-oss-20b | $0.030 | $0.140 | — | 131K |
| gpt-oss-120b (exacto) | $0.050 | $0.240 | — | 131K |
| gpt-oss-safeguard-20b | $0.075 | $0.300 | — | 131K |
| GPT-4.1 Nano | $0.100 | $0.400 | $0.200−50% | 1.0M |
| GPT-5 Nano | $0.050 | $0.400 | — | 400K |
| GPT 4o Mini Audio Preview | $0.150 | $0.600 | — | 128K |
| GPT-4o-mini | $0.150 | $0.600 | $0.300−50% | 128K |
| GPT-4o-mini Search Preview | $0.150 | $0.600 | $0.300−50% | 128K |
| Ft:gpt 4.1 Nano 2025 04.14 | $0.200 | $0.800 | $0.400−50% | 1.0M |
| Ft:gpt 4o Mini 2024 07.18 | $0.300 | $1.20 | $0.600−50% | 128K |
| GPT-5.6 Luna VisionReasoningToolsCache | $0.200 | $1.20 | $0.600−50% | 1.1M |
| gpt-5.4-nano | $0.200 | $1.25 | $0.625−50% | 272K |
| GPT 3.5 Turbo 0125 | $0.500 | $1.50 | — | 16K |
| GPT-3.5 Turbo | $0.500 | $1.50 | — | 16K |
| GPT-4.1 Mini | $0.400 | $1.60 | $0.800−50% | 1.0M |
| GPT 3.5 Turbo 1106 | $1.00 | $2.00 | — | 16K |
| GPT-3.5 Turbo (older v0613) | $1.00 | $2.00 | — | 4K |
| GPT-3.5 Turbo Instruct | $1.50 | $2.00 | — | 4K |
| GPT-5 Image Mini | $2.50 | $2.00 | — | 400K |
| GPT-5 Mini | $0.250 | $2.00 | — | 400K |
| GPT-5.1-Codex-Mini | $0.250 | $2.00 | — | 400K |
| GPT Audio Mini | $0.600 | $2.40 | — | 128K |
| Ft:gpt 4.1 Mini 2025 04.14 | $0.800 | $3.20 | $1.60−50% | 1.0M |
| GPT-3.5 Turbo 16k | $3.00 | $4.00 | — | 16K |
| o3 Mini | $1.10 | $4.40 | — | 200K |
| o3 Mini High | $1.10 | $4.40 | — | 200K |
| o4 Mini | $1.10 | $4.40 | — | 200K |
| o4 Mini High | $1.10 | $4.40 | — | 200K |
| gpt-5.4-mini | $0.750 | $4.50 | $2.25−50% | 272K |
| Codex Mini | $1.50 | $6.00 | — | 200K |
| Ft:gpt 3.5 Turbo | $3.00 | $6.00 | $3.00−50% | 16K |
| Ft:gpt 3.5 Turbo 0125 | $3.00 | $6.00 | — | 16K |
| Ft:gpt 3.5 Turbo 0613 | $3.00 | $6.00 | — | 4K |
| Ft:gpt 3.5 Turbo 1106 | $3.00 | $6.00 | — | 16K |
| GPT-4.1 | $2.00 | $8.00 | $4.00−50% | 1.0M |
| o3 | $2.00 | $8.00 | — | 200K |
| o4 Mini Deep Research | $2.00 | $8.00 | — | 200K |
| GPT 4o Audio Preview | $2.50 | $10.00 | — | 128K |
| GPT 5 Chat Latest | $1.25 | $10.00 | — | 128K |
| GPT 5 Search Api | $1.25 | $10.00 | — | 272K |
| GPT 5.1 Chat Latest | $1.25 | $10.00 | — | 128K |
| GPT Audio | $2.50 | $10.00 | — | 128K |
| GPT Audio 1.5 | $2.50 | $10.00 | — | 128K |
| GPT-4o | $2.50 | $10.00 | $5.00−50% | 128K |
| GPT-4o Audio | $2.50 | $10.00 | — | 128K |
| GPT-4o Search Preview | $2.50 | $10.00 | $5.00−50% | 128K |
| GPT-5 | $1.25 | $10.00 | — | 400K |
| GPT-5 Chat | $1.25 | $10.00 | — | 128K |
| GPT-5 Codex | $1.25 | $10.00 | — | 400K |
| GPT-5 Image | $10.00 | $10.00 | — | 400K |
| GPT-5.1 | $1.25 | $10.00 | — | 400K |
| GPT-5.1 Chat | $1.25 | $10.00 | — | 128K |
| GPT-5.1-Codex | $1.25 | $10.00 | — | 400K |
| Ft:gpt 4.1 2025 04.14 | $3.00 | $12.00 | $6.00−50% | 1.0M |
| GPT-5.6 Terra VisionReasoningToolsCache | $2.00 | $12.00 | $6.00−50% | 1.1M |
| GPT 5.2 Chat Latest | $1.75 | $14.00 | — | 128K |
| GPT 5.3 Chat Latest | $1.75 | $14.00 | — | 128K |
| GPT-5.2 | $1.75 | $14.00 | — | 400K |
| GPT-5.2 Chat VisionReasoningToolsCache | $1.75 | $14.00 | — | 128K |
| GPT-5.3 Chat VisionToolsCache | $1.75 | $14.00 | — | 128K |
| GPT-5.3 Codex VisionReasoningToolsCache | $1.75 | $14.00 | — | 400K |
| GPT-5.3 Codex Spark VisionReasoningToolsCache | $1.75 | $14.00 | — | 128K |
| Chatgpt 4o Latest | $5.00 | $15.00 | — | 128K |
| ChatGPT-4o | $5.00 | $15.00 | — | 128K |
| Ft:gpt 4o 2024 08.06 | $3.75 | $15.00 | $7.50−50% | 128K |
| GPT-4o (2024-05-13) | $5.00 | $15.00 | $7.50−50% | 128K |
| GPT-5.4 VisionReasoningToolsCache | $2.50 | $15.00 | $7.50−50% | 1.1M |
| Ft:o4 Mini 2025 04.16 | $4.00 | $16.00 | $8.00−50% | 200K |
| GPT-4o (extended) | $6.00 | $18.00 | — | 128K |
| gpt-5.4 (>272K context length) | $5.00 | $22.50 | — | 2.0M |
| GPT-Realtime-2.1 VisionReasoningToolsCache | $4.00 | $24.00 | — | 128K |
| GPT 4 0125 Preview | $10.00 | $30.00 | — | 128K |
| GPT 4 1106 Preview | $10.00 | $30.00 | — | 128K |
| GPT-4 Turbo | $10.00 | $30.00 | — | 128K |
| GPT-4 Turbo (older v1106) | $10.00 | $30.00 | — | 128K |
| GPT-4 Turbo Preview | $10.00 | $30.00 | — | 128K |
| GPT-5.5 VisionReasoningToolsCache | $5.00 | $30.00 | $15.00−50% | 1.1M |
| GPT-5.6 VisionReasoningToolsCache | $5.00 | $30.00 | $15.00−50% | 1.1M |
| GPT-5.6 Sol VisionReasoningToolsCache | $5.00 | $30.00 | $15.00−50% | 1.1M |
| o1-pro | $150.00 | $600.00 | — | 200K |
A dash in the batch column means our catalog carries no published batch rate for that model — not that the model has no batch tier. Batch coverage in the upstream pricing data is uneven.
Frequently asked questions
How much does OpenAI cost per 1M tokens?
OpenAI pricing starts at $0.140 per 1 million output tokens (gpt-oss-20b) across 95 models, with its current flagship GPT-5.5 Pro at $180.00. Older premium models in the catalog list higher. Rates current as of August 11, 2026.
What is the cheapest OpenAI model?
gpt-oss-20b is the cheapest OpenAI model on both axes — $0.030 per 1M input tokens and $0.140 per 1M output. It is the right default for high-volume, low-complexity work such as classification, extraction and routing.
What is the largest OpenAI context window?
gpt-5.4 (>272K context length) has the largest context window in the OpenAI catalog at 2.0M input tokens, with up to 16K output tokens per response.
Does OpenAI support prompt caching?
Yes. GPT-5.6 Luna reads cached input at $0.020 per 1M tokens, against $0.200 per 1M uncached. Caching pays off when a long system prompt or document is reused across many requests.
Does OpenAI offer batch pricing?
Yes — 35 OpenAI models in this catalog publish a discounted rate for asynchronous batch jobs, at 50% off output tokens. GPT-5.6 bills $15.00 per 1M output in batch against $30.00 synchronously. Batch trades real-time responses for the lower rate, so it suits backfills, evaluations and bulk enrichment rather than user-facing calls.
Which OpenAI models support reasoning or vision?
The OpenAI catalog includes 12 reasoning models, 13 vision models, 13 with tool calling. Reasoning models bill their internal thinking as output tokens, so they cost more per visible response than the headline rate suggests.
How do I calculate my actual OpenAI bill?
Multiply your input tokens by the input rate and your output tokens by the output rate, both per million, then add retries and any conversation history resent on each turn. Calcaas does this against your real usage pattern and shows the margin left at your price point.
Head-to-head comparisons
Other providers
What will this actually cost you?
Per-token rates only tell you half the story. Calcaas multiplies them by your real usage — prompt size, response length, retries and conversation history — then shows the margin left at your price point.
Open the free cost calculator