API pricing

OpenAI Pricing

OpenAI charges from $0.140 per 1 million output tokens, depending on the model. gpt-oss-20b is the cheapest at $0.140/1M output and $0.030/1M input; its current flagship GPT-5.5 Pro costs $180.00/1M output. The largest context window is 2.0M tokens (gpt-5.4 (>272K context length)). Prices are USD, current as of August 11, 2026.

OpenAI publishes separate input and output rates per model, with cached input billed at a large discount. Reasoning models bill their internal thinking as output tokens, so a reasoning call can cost several times a standard completion on the same visible response length.

Last updated · synced weekly from the upstream model catalog

Cheapest input

$0.030/1M

gpt-oss-20b

Cheapest output

$0.140/1M

gpt-oss-20b

Longest context

2.0M

gpt-5.4 (>272K context length)

Models priced

95

Avg $29.47/1M output

Pricing by model

All prices in USD per 1 million tokens. Showing 80 of 95 priced models, cheapest first. The batch column is the asynchronous batch-API rate, 50% below the standard output rate on the 35 models that publish one.

ModelInput / 1MOutput / 1MBatch out / 1MContext
gpt-oss-20b$0.030$0.140131K
gpt-oss-120b (exacto)$0.050$0.240131K
gpt-oss-safeguard-20b$0.075$0.300131K
GPT-4.1 Nano$0.100$0.400$0.20050%1.0M
GPT-5 Nano$0.050$0.400400K
GPT 4o Mini Audio Preview$0.150$0.600128K
GPT-4o-mini$0.150$0.600$0.30050%128K
GPT-4o-mini Search Preview$0.150$0.600$0.30050%128K
Ft:gpt 4.1 Nano 2025 04.14$0.200$0.800$0.40050%1.0M
Ft:gpt 4o Mini 2024 07.18$0.300$1.20$0.60050%128K
GPT-5.6 Luna
VisionReasoningToolsCache
$0.200$1.20$0.60050%1.1M
gpt-5.4-nano$0.200$1.25$0.62550%272K
GPT 3.5 Turbo 0125$0.500$1.5016K
GPT-3.5 Turbo$0.500$1.5016K
GPT-4.1 Mini$0.400$1.60$0.80050%1.0M
GPT 3.5 Turbo 1106$1.00$2.0016K
GPT-3.5 Turbo (older v0613)$1.00$2.004K
GPT-3.5 Turbo Instruct$1.50$2.004K
GPT-5 Image Mini$2.50$2.00400K
GPT-5 Mini$0.250$2.00400K
GPT-5.1-Codex-Mini$0.250$2.00400K
GPT Audio Mini$0.600$2.40128K
Ft:gpt 4.1 Mini 2025 04.14$0.800$3.20$1.6050%1.0M
GPT-3.5 Turbo 16k$3.00$4.0016K
o3 Mini$1.10$4.40200K
o3 Mini High$1.10$4.40200K
o4 Mini$1.10$4.40200K
o4 Mini High$1.10$4.40200K
gpt-5.4-mini$0.750$4.50$2.2550%272K
Codex Mini$1.50$6.00200K
Ft:gpt 3.5 Turbo$3.00$6.00$3.0050%16K
Ft:gpt 3.5 Turbo 0125$3.00$6.0016K
Ft:gpt 3.5 Turbo 0613$3.00$6.004K
Ft:gpt 3.5 Turbo 1106$3.00$6.0016K
GPT-4.1$2.00$8.00$4.0050%1.0M
o3$2.00$8.00200K
o4 Mini Deep Research$2.00$8.00200K
GPT 4o Audio Preview$2.50$10.00128K
GPT 5 Chat Latest$1.25$10.00128K
GPT 5 Search Api$1.25$10.00272K
GPT 5.1 Chat Latest$1.25$10.00128K
GPT Audio$2.50$10.00128K
GPT Audio 1.5$2.50$10.00128K
GPT-4o$2.50$10.00$5.0050%128K
GPT-4o Audio$2.50$10.00128K
GPT-4o Search Preview$2.50$10.00$5.0050%128K
GPT-5$1.25$10.00400K
GPT-5 Chat$1.25$10.00128K
GPT-5 Codex$1.25$10.00400K
GPT-5 Image$10.00$10.00400K
GPT-5.1$1.25$10.00400K
GPT-5.1 Chat$1.25$10.00128K
GPT-5.1-Codex$1.25$10.00400K
Ft:gpt 4.1 2025 04.14$3.00$12.00$6.0050%1.0M
GPT-5.6 Terra
VisionReasoningToolsCache
$2.00$12.00$6.0050%1.1M
GPT 5.2 Chat Latest$1.75$14.00128K
GPT 5.3 Chat Latest$1.75$14.00128K
GPT-5.2$1.75$14.00400K
GPT-5.2 Chat
VisionReasoningToolsCache
$1.75$14.00128K
GPT-5.3 Chat
VisionToolsCache
$1.75$14.00128K
GPT-5.3 Codex
VisionReasoningToolsCache
$1.75$14.00400K
GPT-5.3 Codex Spark
VisionReasoningToolsCache
$1.75$14.00128K
Chatgpt 4o Latest$5.00$15.00128K
ChatGPT-4o$5.00$15.00128K
Ft:gpt 4o 2024 08.06$3.75$15.00$7.5050%128K
GPT-4o (2024-05-13)$5.00$15.00$7.5050%128K
GPT-5.4
VisionReasoningToolsCache
$2.50$15.00$7.5050%1.1M
Ft:o4 Mini 2025 04.16$4.00$16.00$8.0050%200K
GPT-4o (extended)$6.00$18.00128K
gpt-5.4 (>272K context length)$5.00$22.502.0M
GPT-Realtime-2.1
VisionReasoningToolsCache
$4.00$24.00128K
GPT 4 0125 Preview$10.00$30.00128K
GPT 4 1106 Preview$10.00$30.00128K
GPT-4 Turbo$10.00$30.00128K
GPT-4 Turbo (older v1106)$10.00$30.00128K
GPT-4 Turbo Preview$10.00$30.00128K
GPT-5.5
VisionReasoningToolsCache
$5.00$30.00$15.0050%1.1M
GPT-5.6
VisionReasoningToolsCache
$5.00$30.00$15.0050%1.1M
GPT-5.6 Sol
VisionReasoningToolsCache
$5.00$30.00$15.0050%1.1M
o1-pro$150.00$600.00200K

A dash in the batch column means our catalog carries no published batch rate for that model — not that the model has no batch tier. Batch coverage in the upstream pricing data is uneven.

Frequently asked questions

How much does OpenAI cost per 1M tokens?

OpenAI pricing starts at $0.140 per 1 million output tokens (gpt-oss-20b) across 95 models, with its current flagship GPT-5.5 Pro at $180.00. Older premium models in the catalog list higher. Rates current as of August 11, 2026.

What is the cheapest OpenAI model?

gpt-oss-20b is the cheapest OpenAI model on both axes — $0.030 per 1M input tokens and $0.140 per 1M output. It is the right default for high-volume, low-complexity work such as classification, extraction and routing.

What is the largest OpenAI context window?

gpt-5.4 (>272K context length) has the largest context window in the OpenAI catalog at 2.0M input tokens, with up to 16K output tokens per response.

Does OpenAI support prompt caching?

Yes. GPT-5.6 Luna reads cached input at $0.020 per 1M tokens, against $0.200 per 1M uncached. Caching pays off when a long system prompt or document is reused across many requests.

Does OpenAI offer batch pricing?

Yes — 35 OpenAI models in this catalog publish a discounted rate for asynchronous batch jobs, at 50% off output tokens. GPT-5.6 bills $15.00 per 1M output in batch against $30.00 synchronously. Batch trades real-time responses for the lower rate, so it suits backfills, evaluations and bulk enrichment rather than user-facing calls.

Which OpenAI models support reasoning or vision?

The OpenAI catalog includes 12 reasoning models, 13 vision models, 13 with tool calling. Reasoning models bill their internal thinking as output tokens, so they cost more per visible response than the headline rate suggests.

How do I calculate my actual OpenAI bill?

Multiply your input tokens by the input rate and your output tokens by the output rate, both per million, then add retries and any conversation history resent on each turn. Calcaas does this against your real usage pattern and shows the margin left at your price point.

Head-to-head comparisons

Other providers

What will this actually cost you?

Per-token rates only tell you half the story. Calcaas multiplies them by your real usage — prompt size, response length, retries and conversation history — then shows the margin left at your price point.

Open the free cost calculator