OpenAI Pricing
OpenAI charges between $0.140 and $180.00 per 1 million output tokens, depending on the model. gpt-oss-20b is the cheapest at $0.140/1M output and $0.030/1M input; GPT-5.5 Pro is the most expensive at $180.00/1M output. The largest context window is 2.0M tokens (gpt-5.4 (>272K context length)). Prices are USD, current as of July 20, 2026.
OpenAI publishes separate input and output rates per model, with cached input billed at a large discount. Reasoning models bill their internal thinking as output tokens, so a reasoning call can cost several times a standard completion on the same visible response length.
Last updated · synced weekly from the upstream model catalog
Cheapest input
$0.030/1M
gpt-oss-20b
Cheapest output
$0.140/1M
gpt-oss-20b
Longest context
2.0M
gpt-5.4 (>272K context length)
Models priced
97
Avg $29.26/1M output
Pricing by model
All prices in USD per 1 million tokens. Showing 80 of 97 priced models, cheapest first.
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
| gpt-oss-20b | $0.030 | $0.140 | 131K |
| gpt-oss-120b (exacto) | $0.050 | $0.240 | 131K |
| gpt-oss-safeguard-20b | $0.075 | $0.300 | 131K |
| GPT-4.1 Nano | $0.100 | $0.400 | 1.0M |
| GPT-5 Nano | $0.050 | $0.400 | 400K |
| GPT 4o Mini Audio Preview | $0.150 | $0.600 | 128K |
| GPT-4o-mini | $0.150 | $0.600 | 128K |
| GPT-4o-mini Search Preview | $0.150 | $0.600 | 128K |
| Ft:gpt 4.1 Nano 2025 04.14 | $0.200 | $0.800 | 1.0M |
| Ft:gpt 4o Mini 2024 07.18 | $0.300 | $1.20 | 128K |
| gpt-5.4-nano | $0.200 | $1.25 | 272K |
| GPT 3.5 Turbo 0125 | $0.500 | $1.50 | 16K |
| GPT-3.5 Turbo | $0.500 | $1.50 | 16K |
| GPT-4.1 Mini | $0.400 | $1.60 | 1.0M |
| GPT 3.5 Turbo 1106 | $1.00 | $2.00 | 16K |
| GPT-3.5 Turbo (older v0613) | $1.00 | $2.00 | 4K |
| GPT-3.5 Turbo Instruct | $1.50 | $2.00 | 4K |
| GPT-5 Image Mini | $2.50 | $2.00 | 400K |
| GPT-5 Mini | $0.250 | $2.00 | 400K |
| GPT-5.1-Codex-Mini | $0.250 | $2.00 | 400K |
| GPT Audio Mini | $0.600 | $2.40 | 128K |
| Ft:gpt 4.1 Mini 2025 04.14 | $0.800 | $3.20 | 1.0M |
| GPT-3.5 Turbo 16k | $3.00 | $4.00 | 16K |
| o3 Mini | $1.10 | $4.40 | 200K |
| o3 Mini High | $1.10 | $4.40 | 200K |
| o4 Mini | $1.10 | $4.40 | 200K |
| o4 Mini High | $1.10 | $4.40 | 200K |
| gpt-5.4-mini | $0.750 | $4.50 | 272K |
| Codex Mini | $1.50 | $6.00 | 200K |
| Ft:gpt 3.5 Turbo | $3.00 | $6.00 | 16K |
| Ft:gpt 3.5 Turbo 0125 | $3.00 | $6.00 | 16K |
| Ft:gpt 3.5 Turbo 0613 | $3.00 | $6.00 | 4K |
| Ft:gpt 3.5 Turbo 1106 | $3.00 | $6.00 | 16K |
| GPT-5.6 Luna VisionReasoningToolsCache | $1.00 | $6.00 | 1.1M |
| GPT-4.1 | $2.00 | $8.00 | 1.0M |
| o3 | $2.00 | $8.00 | 200K |
| o4 Mini Deep Research | $2.00 | $8.00 | 200K |
| GPT 4o Audio Preview | $2.50 | $10.00 | 128K |
| GPT 5 Chat Latest | $1.25 | $10.00 | 128K |
| GPT 5 Search Api | $1.25 | $10.00 | 272K |
| GPT 5.1 Chat Latest | $1.25 | $10.00 | 128K |
| GPT Audio | $2.50 | $10.00 | 128K |
| GPT Audio 1.5 | $2.50 | $10.00 | 128K |
| GPT-4o | $2.50 | $10.00 | 128K |
| GPT-4o Audio | $2.50 | $10.00 | 128K |
| GPT-4o Search Preview | $2.50 | $10.00 | 128K |
| GPT-5 | $1.25 | $10.00 | 400K |
| GPT-5 Chat | $1.25 | $10.00 | 128K |
| GPT-5 Codex | $1.25 | $10.00 | 400K |
| GPT-5 Image | $10.00 | $10.00 | 400K |
| GPT-5.1 | $1.25 | $10.00 | 400K |
| GPT-5.1 Chat | $1.25 | $10.00 | 128K |
| GPT-5.1 Codex Max VisionReasoningToolsCache | $1.25 | $10.00 | 400K |
| GPT-5.1-Codex | $1.25 | $10.00 | 400K |
| Ft:gpt 4.1 2025 04.14 | $3.00 | $12.00 | 1.0M |
| GPT 5.2 Chat Latest | $1.75 | $14.00 | 128K |
| GPT 5.3 Chat Latest | $1.75 | $14.00 | 128K |
| GPT-5.2 | $1.75 | $14.00 | 400K |
| GPT-5.2 Chat VisionReasoningToolsCache | $1.75 | $14.00 | 128K |
| GPT-5.2 Codex VisionReasoningToolsCache | $1.75 | $14.00 | 400K |
| GPT-5.3 Chat VisionToolsCache | $1.75 | $14.00 | 128K |
| GPT-5.3 Codex VisionReasoningToolsCache | $1.75 | $14.00 | 400K |
| GPT-5.3 Codex Spark VisionReasoningToolsCache | $1.75 | $14.00 | 128K |
| Chatgpt 4o Latest | $5.00 | $15.00 | 128K |
| ChatGPT-4o | $5.00 | $15.00 | 128K |
| Ft:gpt 4o 2024 08.06 | $3.75 | $15.00 | 128K |
| GPT-4o (2024-05-13) | $5.00 | $15.00 | 128K |
| GPT-5.4 VisionReasoningToolsCache | $2.50 | $15.00 | 1.1M |
| GPT-5.6 Terra VisionReasoningToolsCache | $2.50 | $15.00 | 1.1M |
| Ft:o4 Mini 2025 04.16 | $4.00 | $16.00 | 200K |
| GPT-4o (extended) | $6.00 | $18.00 | 128K |
| gpt-5.4 (>272K context length) | $5.00 | $22.50 | 2.0M |
| GPT-Realtime-2.1 VisionReasoningToolsCache | $4.00 | $24.00 | 128K |
| GPT 4 0125 Preview | $10.00 | $30.00 | 128K |
| GPT 4 1106 Preview | $10.00 | $30.00 | 128K |
| GPT-4 Turbo | $10.00 | $30.00 | 128K |
| GPT-4 Turbo (older v1106) | $10.00 | $30.00 | 128K |
| GPT-4 Turbo Preview | $10.00 | $30.00 | 128K |
| GPT-5.5 VisionReasoningToolsCache | $5.00 | $30.00 | 1.1M |
| o1-pro | $150.00 | $600.00 | 200K |
Frequently asked questions
How much does OpenAI cost per 1M tokens?
OpenAI pricing ranges from $0.140 to $180.00 per 1 million output tokens across 97 models. Input tokens are cheaper than output on every model. Rates current as of July 20, 2026.
What is the cheapest OpenAI model?
gpt-oss-20b is the cheapest OpenAI model on both axes — $0.030 per 1M input tokens and $0.140 per 1M output. It is the right default for high-volume, low-complexity work such as classification, extraction and routing.
What is the largest OpenAI context window?
gpt-5.4 (>272K context length) has the largest context window in the OpenAI catalog at 2.0M input tokens, with up to 16K output tokens per response.
Does OpenAI support prompt caching?
Yes. GPT-5.6 Luna reads cached input at $0.100 per 1M tokens, against $1.00 per 1M uncached. Caching pays off when a long system prompt or document is reused across many requests.
Which OpenAI models support reasoning or vision?
The OpenAI catalog includes 14 reasoning models, 15 vision models, 15 with tool calling. Reasoning models bill their internal thinking as output tokens, so they cost more per visible response than the headline rate suggests.
How do I calculate my actual OpenAI bill?
Multiply your input tokens by the input rate and your output tokens by the output rate, both per million, then add retries and any conversation history resent on each turn. Calcaas does this against your real usage pattern and shows the margin left at your price point.
Head-to-head comparisons
Other providers
What will this actually cost you?
Per-token rates only tell you half the story. Calcaas multiplies them by your real usage — prompt size, response length, retries and conversation history — then shows the margin left at your price point.
Open the free cost calculator