API pricing

OpenRouter Pricing

OpenRouter charges between $0.030 and $30.00 per 1 million output tokens, depending on the model. Ling-2.6-flash is the cheapest at $0.030/1M output and $0.010/1M input; GPT-5.6 Sol is the most expensive at $30.00/1M output. The largest context window is 10.0M tokens (Llama 4 Scout). Prices are USD, current as of July 20, 2026.

Last updated · synced weekly from the upstream model catalog

Cheapest input

$0.010/1M

Ling-2.6-flash

Cheapest output

$0.030/1M

Ling-2.6-flash

Longest context

10.0M

Llama 4 Scout

Models priced

403

Avg $10.21/1M output

Pricing by model

All prices in USD per 1 million tokens. Showing 80 of 403 priced models, cheapest first.

ModelInput / 1MOutput / 1MContext
Ling-2.6-flash
ToolsCache
$0.010$0.030262K
Mistral Nemo
Tools
$0.019$0.030131K
Llama 3 8B Lunaris$0.040$0.0508K
MythoMax 13B$0.060$0.0604K
Llama 3.1 8B Instruct
ToolsCache
$0.050$0.080131K
Mistral Small 3$0.050$0.08033K
Gemma 3 4B
Vision
$0.050$0.100131K
Granite 4.1 8B
ToolsCache
$0.050$0.100131K
Ministral 3 3B 2512
VisionToolsCache
$0.100$0.100131K
Mistralai Ministral 3B 2512$0.100$0.100131K
Nex-N2-Mini
VisionReasoningToolsCache
$0.025$0.100262K
OpenAI GPT Oss 20B$0.020$0.100131K
Qwen Qwen3 235B A22b 2507$0.071$0.100262K
Qwen2.5 7B Instruct
Tools
$0.040$0.100131K
Reka Edge
VisionTools
$0.100$0.10016K
Granite 4.0 Micro$0.017$0.112131K
Gemma 3n 4B$0.060$0.12033K
Laguna XS 2.1
ReasoningToolsCache
$0.060$0.120262K
GPT OSS 20B
ReasoningToolsCache
$0.030$0.130131K
Mistralai Mistral 7B Instruct$0.130$0.13033K
Nova Micro 1.0
Tools
$0.035$0.140128K
Phi 4$0.070$0.14016K
Command R7B$0.037$0.150128K
Gemma 3 12B
VisionTools
$0.050$0.150131K
Ministral 3 8B 2512
VisionToolsCache
$0.150$0.150262K
Mistralai Ministral 8B 2512$0.150$0.150262K
Qwen3.5 9B
VisionReasoningTools
$0.100$0.150262K
GPT OSS 120B
ReasoningTools
$0.037$0.170131K
Llama Guard 4 12B
Vision
$0.180$0.180164K
Qwen Qwen 2.5 Coder 32B Instruct$0.180$0.18034K
DeepSeek V4 Flash
ReasoningToolsCache
$0.098$0.1961.0M
Bytedance Ui Tars 1.5 7B$0.100$0.200131K
Ministral 3 14B 2512
VisionToolsCache
$0.200$0.200262K
Mistralai Ministral 14B 2512$0.200$0.200262K
Nemotron 3 Nano 30B A3B
ReasoningTools
$0.050$0.200262K
Reka Flash 3
Reasoning
$0.100$0.20066K
UI-TARS 7B
VisionCache
$0.100$0.200128K
Llama 3.2 1B Instruct$0.027$0.201131K
Hy3 preview
ReasoningToolsCache
$0.063$0.210262K
Nova Lite 1.0
VisionTools
$0.060$0.240300K
Qwen3 14B
ReasoningTools
$0.120$0.240132K
Qwen3.5-Flash
VisionReasoningTools
$0.065$0.2601.0M
Qwen3-Coder 30B-A3B Instruct
Tools
$0.070$0.270160K
DeepSeek DeepSeek Chat$0.140$0.28066K
DeepSeek DeepSeek Chat V3 0324$0.140$0.28066K
MiMo-V2.5
VisionReasoningToolsCache
$0.140$0.2801.0M
Qwen3 32B
ReasoningTools
$0.080$0.280131K
Gemma 3 27B
VisionTools
$0.100$0.300131K
gpt-oss-safeguard-20b
ReasoningToolsCache
$0.075$0.300131K
Llama 4 Scout
VisionTools
$0.100$0.30010.0M
Mistral Small 3.2 24B
VisionToolsCache
$0.100$0.300131K
Mistralai Mistral Small 3.1 24B Instruct$0.100$0.300131K
Mistralai Mistral Small 3.2 24B Instruct$0.100$0.300128K
Qwen3 30B A3B Instruct 2507
Tools
$0.100$0.300262K
Seed 1.6 Flash
VisionReasoningTools
$0.075$0.300262K
Step 3.5 Flash
ReasoningTools
$0.100$0.300262K
Voxtral Small 24B 2507
ToolsCache
$0.100$0.30032K
Xiaomi Mimo V2 Flash$0.100$0.300262K
Llama 3.2 3B Instruct$0.051$0.335131K
Gemma 4 26B A4B IT
VisionReasoningTools
$0.070$0.340262K
DeepSeek DeepSeek V3.2$0.280$0.400164K
DeepSeek DeepSeek V3.2 Exp$0.200$0.400164K
DeepSeek V3.2
ReasoningToolsCache
$0.269$0.400164K
Gemini 2.5 Flash-Lite
VisionReasoningToolsCache
$0.100$0.4001.0M
GLM-4.7-Flash
ReasoningTools
$0.060$0.400200K
Google Gemini 2.0 Flash 001$0.100$0.4001.0M
GPT-4.1 nano
VisionToolsCache
$0.100$0.4001.0M
GPT-5 Nano
VisionReasoningToolsCache
$0.050$0.400400K
Hermes 4 70B
Reasoning
$0.130$0.400131K
Laguna M.1
ReasoningToolsCache
$0.200$0.400262K
Llama 3.1 70B Instruct
Tools
$0.400$0.400131K
Llama-3.3-70B-Instruct
Tools
$0.130$0.400131K
Nemotron 3 Super 120B A12B
ReasoningTools
$0.085$0.4001.0M
OpenAI GPT 4.1 Nano$0.100$0.4001.0M
OpenAI GPT 5 Nano$0.050$0.400272K
Qwen Qwen3.5 Flash 02.23$0.100$0.4001.0M
Qwen2.5 72B Instruct
Tools
$0.360$0.400131K
Seed-2.0-Mini
VisionReasoningTools
$0.100$0.400262K
UnslopNemo 12B
Tools
$0.400$0.40033K
o1-pro
VisionReasoning
$150.00$600.00200K

Frequently asked questions

How much does OpenRouter cost per 1M tokens?

OpenRouter pricing ranges from $0.030 to $30.00 per 1 million output tokens across 403 models. Input tokens are cheaper than output on every model. Rates current as of July 20, 2026.

What is the cheapest OpenRouter model?

Ling-2.6-flash is the cheapest OpenRouter model on both axes — $0.010 per 1M input tokens and $0.030 per 1M output. It is the right default for high-volume, low-complexity work such as classification, extraction and routing.

What is the largest OpenRouter context window?

Llama 4 Scout has the largest context window in the OpenRouter catalog at 10.0M input tokens, with up to 16K output tokens per response.

Does OpenRouter support prompt caching?

Yes. Ling-2.6-flash reads cached input at $0.0020 per 1M tokens, against $0.010 per 1M uncached. Caching pays off when a long system prompt or document is reused across many requests.

Which OpenRouter models support reasoning or vision?

The OpenRouter catalog includes 187 reasoning models, 167 vision models, 251 with tool calling. Reasoning models bill their internal thinking as output tokens, so they cost more per visible response than the headline rate suggests.

How do I calculate my actual OpenRouter bill?

Multiply your input tokens by the input rate and your output tokens by the output rate, both per million, then add retries and any conversation history resent on each turn. Calcaas does this against your real usage pattern and shows the margin left at your price point.

Other providers

What will this actually cost you?

Per-token rates only tell you half the story. Calcaas multiplies them by your real usage — prompt size, response length, retries and conversation history — then shows the margin left at your price point.

Open the free cost calculator