API pricing

OpenAI Pricing

OpenAI charges between $0.140 and $180.00 per 1 million output tokens, depending on the model. gpt-oss-20b is the cheapest at $0.140/1M output and $0.030/1M input; GPT-5.5 Pro is the most expensive at $180.00/1M output. The largest context window is 2.0M tokens (gpt-5.4 (>272K context length)). Prices are USD, current as of July 20, 2026.

OpenAI publishes separate input and output rates per model, with cached input billed at a large discount. Reasoning models bill their internal thinking as output tokens, so a reasoning call can cost several times a standard completion on the same visible response length.

Last updated · synced weekly from the upstream model catalog

Cheapest input

$0.030/1M

gpt-oss-20b

Cheapest output

$0.140/1M

gpt-oss-20b

Longest context

2.0M

gpt-5.4 (>272K context length)

Models priced

97

Avg $29.26/1M output

Pricing by model

All prices in USD per 1 million tokens. Showing 80 of 97 priced models, cheapest first.

ModelInput / 1MOutput / 1MContext
gpt-oss-20b$0.030$0.140131K
gpt-oss-120b (exacto)$0.050$0.240131K
gpt-oss-safeguard-20b$0.075$0.300131K
GPT-4.1 Nano$0.100$0.4001.0M
GPT-5 Nano$0.050$0.400400K
GPT 4o Mini Audio Preview$0.150$0.600128K
GPT-4o-mini$0.150$0.600128K
GPT-4o-mini Search Preview$0.150$0.600128K
Ft:gpt 4.1 Nano 2025 04.14$0.200$0.8001.0M
Ft:gpt 4o Mini 2024 07.18$0.300$1.20128K
gpt-5.4-nano$0.200$1.25272K
GPT 3.5 Turbo 0125$0.500$1.5016K
GPT-3.5 Turbo$0.500$1.5016K
GPT-4.1 Mini$0.400$1.601.0M
GPT 3.5 Turbo 1106$1.00$2.0016K
GPT-3.5 Turbo (older v0613)$1.00$2.004K
GPT-3.5 Turbo Instruct$1.50$2.004K
GPT-5 Image Mini$2.50$2.00400K
GPT-5 Mini$0.250$2.00400K
GPT-5.1-Codex-Mini$0.250$2.00400K
GPT Audio Mini$0.600$2.40128K
Ft:gpt 4.1 Mini 2025 04.14$0.800$3.201.0M
GPT-3.5 Turbo 16k$3.00$4.0016K
o3 Mini$1.10$4.40200K
o3 Mini High$1.10$4.40200K
o4 Mini$1.10$4.40200K
o4 Mini High$1.10$4.40200K
gpt-5.4-mini$0.750$4.50272K
Codex Mini$1.50$6.00200K
Ft:gpt 3.5 Turbo$3.00$6.0016K
Ft:gpt 3.5 Turbo 0125$3.00$6.0016K
Ft:gpt 3.5 Turbo 0613$3.00$6.004K
Ft:gpt 3.5 Turbo 1106$3.00$6.0016K
GPT-5.6 Luna
VisionReasoningToolsCache
$1.00$6.001.1M
GPT-4.1$2.00$8.001.0M
o3$2.00$8.00200K
o4 Mini Deep Research$2.00$8.00200K
GPT 4o Audio Preview$2.50$10.00128K
GPT 5 Chat Latest$1.25$10.00128K
GPT 5 Search Api$1.25$10.00272K
GPT 5.1 Chat Latest$1.25$10.00128K
GPT Audio$2.50$10.00128K
GPT Audio 1.5$2.50$10.00128K
GPT-4o$2.50$10.00128K
GPT-4o Audio$2.50$10.00128K
GPT-4o Search Preview$2.50$10.00128K
GPT-5$1.25$10.00400K
GPT-5 Chat$1.25$10.00128K
GPT-5 Codex$1.25$10.00400K
GPT-5 Image$10.00$10.00400K
GPT-5.1$1.25$10.00400K
GPT-5.1 Chat$1.25$10.00128K
GPT-5.1 Codex Max
VisionReasoningToolsCache
$1.25$10.00400K
GPT-5.1-Codex$1.25$10.00400K
Ft:gpt 4.1 2025 04.14$3.00$12.001.0M
GPT 5.2 Chat Latest$1.75$14.00128K
GPT 5.3 Chat Latest$1.75$14.00128K
GPT-5.2$1.75$14.00400K
GPT-5.2 Chat
VisionReasoningToolsCache
$1.75$14.00128K
GPT-5.2 Codex
VisionReasoningToolsCache
$1.75$14.00400K
GPT-5.3 Chat
VisionToolsCache
$1.75$14.00128K
GPT-5.3 Codex
VisionReasoningToolsCache
$1.75$14.00400K
GPT-5.3 Codex Spark
VisionReasoningToolsCache
$1.75$14.00128K
Chatgpt 4o Latest$5.00$15.00128K
ChatGPT-4o$5.00$15.00128K
Ft:gpt 4o 2024 08.06$3.75$15.00128K
GPT-4o (2024-05-13)$5.00$15.00128K
GPT-5.4
VisionReasoningToolsCache
$2.50$15.001.1M
GPT-5.6 Terra
VisionReasoningToolsCache
$2.50$15.001.1M
Ft:o4 Mini 2025 04.16$4.00$16.00200K
GPT-4o (extended)$6.00$18.00128K
gpt-5.4 (>272K context length)$5.00$22.502.0M
GPT-Realtime-2.1
VisionReasoningToolsCache
$4.00$24.00128K
GPT 4 0125 Preview$10.00$30.00128K
GPT 4 1106 Preview$10.00$30.00128K
GPT-4 Turbo$10.00$30.00128K
GPT-4 Turbo (older v1106)$10.00$30.00128K
GPT-4 Turbo Preview$10.00$30.00128K
GPT-5.5
VisionReasoningToolsCache
$5.00$30.001.1M
o1-pro$150.00$600.00200K

Frequently asked questions

How much does OpenAI cost per 1M tokens?

OpenAI pricing ranges from $0.140 to $180.00 per 1 million output tokens across 97 models. Input tokens are cheaper than output on every model. Rates current as of July 20, 2026.

What is the cheapest OpenAI model?

gpt-oss-20b is the cheapest OpenAI model on both axes — $0.030 per 1M input tokens and $0.140 per 1M output. It is the right default for high-volume, low-complexity work such as classification, extraction and routing.

What is the largest OpenAI context window?

gpt-5.4 (>272K context length) has the largest context window in the OpenAI catalog at 2.0M input tokens, with up to 16K output tokens per response.

Does OpenAI support prompt caching?

Yes. GPT-5.6 Luna reads cached input at $0.100 per 1M tokens, against $1.00 per 1M uncached. Caching pays off when a long system prompt or document is reused across many requests.

Which OpenAI models support reasoning or vision?

The OpenAI catalog includes 14 reasoning models, 15 vision models, 15 with tool calling. Reasoning models bill their internal thinking as output tokens, so they cost more per visible response than the headline rate suggests.

How do I calculate my actual OpenAI bill?

Multiply your input tokens by the input rate and your output tokens by the output rate, both per million, then add retries and any conversation history resent on each turn. Calcaas does this against your real usage pattern and shows the margin left at your price point.

Head-to-head comparisons

Other providers

What will this actually cost you?

Per-token rates only tell you half the story. Calcaas multiplies them by your real usage — prompt size, response length, retries and conversation history — then shows the margin left at your price point.

Open the free cost calculator