Pricing comparison

OpenAI vs Mistral AI Pricing

Mistral AI is cheaper than OpenAI at the entry level: Ministral 3B costs $0.040 per 1M output tokens against $0.140 for gpt-oss-20b — 3.5× the price. On current flagships, OpenAI's GPT-5.6 costs $30.00/1M output against $7.50/1M for Mistral AI's Mistral Medium. Prices are USD, current as of August 11, 2026.

Per-million-token pricing for OpenAI and Mistral AI, with side-by-side flagship models, cheapest tiers, and context windows. Pricing data syncs weekly from a continuously-updated model catalog — last updated August 11, 2026.

Who wins on what

Cheapest input tokens

$0.02/1M

Mistral AI

Mistral Nemo — $0.02/1M input

Cheapest output tokens

$0.04/1M

Mistral AI

Ministral 3B — $0.04/1M output

Longest context window

2.0M

OpenAI

gpt-5.4 (>272K context length) — 2.0M input tokens

Lowest average output cost

$2.23/1M

Mistral AI

Provider-wide average across 87 models

Largest model catalog

129 models

OpenAI

More options to match cost vs capability

Most reasoning models

12 models

OpenAI

Models with dedicated reasoning / thinking support

Most vision models

13 models

OpenAI

Models that accept image input

Open-weights available

Yes

Mistral AI

Offers open-weight models you can self-host

Side-by-side

129 models

OpenAI

Full OpenAI pricing →

Cheapest input

$0.030

gpt-oss-20b

Cheapest output

$0.140

gpt-oss-20b

Longest context

2.0M

gpt-5.4 (>272K context length)

Avg output / 1M

$29.47

Across catalog

Cheapest cached input

$0.020

GPT-5.6 Luna

ModelIn/1MOut/1MCtx
GPT-5.6
VisionReasoningToolsCache
$5.00$30.001.1M
GPT-5.6 Sol
VisionReasoningToolsCache
$5.00$30.001.1M
GPT-5.6 Terra
VisionReasoningToolsCache
$2.00$12.001.1M
GPT-5.6 Luna
VisionReasoningToolsCache
$0.200$1.201.1M
GPT-Realtime-2.1
VisionReasoningToolsCache
$4.00$24.00128K
gpt-oss-20b$0.030$0.140131K
87 models

Mistral AI

Full Mistral AI pricing →

Cheapest input

$0.020

Mistral Nemo

Cheapest output

$0.040

Ministral 3B

Longest context

262K

Devstral 2

Avg output / 1M

$2.23

Across catalog

ModelIn/1MOut/1MCtx
Mistral Medium
VisionReasoningTools
$1.50$7.50262K
Mistral Medium 3.5
VisionReasoningTools
$1.50$7.50262K
Mistral Small 4
VisionReasoningTools
$0.150$0.600256K
Devstral 2
Tools
$0.400$2.00262K
Mistral Large 3
VisionTools
$0.500$1.50262K
Ministral 3B$0.040$0.040131K

All prices in USD per 1 million tokens. Showing top 6 models per provider, sorted by output cost.

Frequently asked questions

Is OpenAI or Mistral AI cheaper?

Mistral AI has the cheaper entry point at $0.04/1M output (Ministral 3B — $0.04/1M output). Provider-wide, OpenAI averages $29.47/1M output against $2.23/1M for Mistral AI. Which is cheaper for you depends on which model tier your workload actually needs.

How much do OpenAI and Mistral AI cost per 1M tokens?

OpenAI starts at $0.140 per 1M output tokens (gpt-oss-20b), with its current flagship GPT-5.6 at $30.00. Mistral AI starts at $0.040 (Ministral 3B), with Mistral Medium at $7.50. Input tokens cost less than output on both.

Which has the larger context window, OpenAI or Mistral AI?

OpenAI — gpt-5.4 (>272K context length) — 2.0M input tokens. For comparison, OpenAI's largest is 2.0M tokens (gpt-5.4 (>272K context length)) and Mistral AI's is 262K tokens (Devstral 2).

Which has more reasoning models, OpenAI or Mistral AI?

OpenAI lists 12 reasoning models and Mistral AI lists 5. Reasoning models bill their internal thinking as output tokens, so a reasoning call costs several times a standard completion of the same visible length — compare them on total tokens billed, not headline rate.

Can I self-host OpenAI or Mistral AI models?

Mistral AI publishes open-weight models you can run on your own hardware; OpenAI does not. Self-hosting swaps per-token pricing for GPU-hour cost, which usually only wins above sustained high utilisation — below that, hosted inference is cheaper.

Should I switch from OpenAI to Mistral AI to save money?

Only if the cheaper model still meets your quality bar. Token price is one input; the ones that decide your bill are prompt size, response length, retries and how much conversation history you resend each turn. Model the switch against your real traffic before committing. Prices here are current as of August 11, 2026.

Related comparisons

See the full AI model leaderboard — every model ranked by intelligence, real-world usage, and value.

Run the numbers for your workload

Calcaas multiplies per-token costs by your real usage patterns — inputs, outputs, retries, and conversation history — across both providers in one model.