API pricing

xAI (Grok) Pricing

xAI (Grok) charges between $0.500 and $25.00 per 1 million output tokens, depending on the model. Grok 3 Mini is the cheapest at $0.500/1M output and $0.300/1M input; Grok 3 Fast Beta is the most expensive at $25.00/1M output. The largest context window is 2.0M tokens (Grok 4 Fast). Prices are USD, current as of July 20, 2026.

xAI prices Grok per million tokens, with the fast and mini variants sitting well below the flagship. Context windows are large across the lineup, which makes prompt size less of a cost lever here than output length.

Last updated · synced weekly from the upstream model catalog

Cheapest input

$0.200/1M

Grok 4 Fast

Cheapest output

$0.500/1M

Grok 3 Mini

Longest context

2.0M

Grok 4 Fast

Models priced

45

Avg $6.72/1M output

Pricing by model

All prices in USD per 1 million tokens. All 45 priced models, cheapest first.

ModelInput / 1MOutput / 1MContext
Grok 3 Mini$0.300$0.500131K
Grok 3 Mini Beta$0.300$0.500131K
Grok 3 Mini Latest$0.300$0.500131K
Grok 4 Fast$0.200$0.5002.0M
Grok 4 Fast Non Reasoning$0.200$0.5002.0M
Grok 4 Fast Reasoning$0.200$0.5002.0M
Grok 4.1 Fast$0.200$0.5002.0M
Grok 4.1 Fast Non Reasoning$0.200$0.5002.0M
Grok 4.1 Fast Non Reasoning Latest$0.200$0.5002.0M
Grok 4.1 Fast Reasoning$0.200$0.5002.0M
Grok 4.1 Fast Reasoning Latest$0.200$0.5002.0M
Grok Code Fast$0.200$1.50256K
Grok Code Fast 1$0.200$1.50256K
Grok Code Fast 1 0825$0.200$1.50256K
Grok Build 0.1
VisionReasoningToolsCache
$1.00$2.00256K
Grok 4.20 (Non-Reasoning)
VisionToolsCache
$1.25$2.501.0M
Grok 4.20 (Reasoning)
VisionReasoningToolsCache
$1.25$2.501.0M
Grok 4.20 Multi-Agent
VisionReasoningCache
$1.25$2.501.0M
Grok 4.3
VisionReasoningToolsCache
$1.25$2.501.0M
Grok 4.3 Latest$1.25$2.501.0M
Grok 3 Mini Fast$0.600$4.00131K
Grok 3 Mini Fast Beta$0.600$4.00131K
Grok 3 Mini Fast Latest$0.600$4.00131K
Grok 4.20 0309 Reasoning$2.00$6.002.0M
Grok 4.20 Beta 0309 Non Reasoning$2.00$6.002.0M
Grok 4.20 Beta 0309 Reasoning$2.00$6.002.0M
Grok 4.20 Multi Agent Beta 0309$2.00$6.002.0M
Grok 4.5
VisionReasoningToolsCache
$2.00$6.00500K
Grok 4.5 Latest$2.00$6.00500K
Grok 2$2.00$10.00131K
Grok 2 1212$2.00$10.00131K
Grok 2 Latest$2.00$10.00131K
Grok 2 Vision$2.00$10.0033K
Grok 2 Vision 1212$2.00$10.0033K
Grok 2 Vision Latest$2.00$10.0033K
Grok 3$3.00$15.00131K
Grok 3 Beta$3.00$15.00131K
Grok 3 Latest$3.00$15.00131K
Grok 4$3.00$15.00256K
Grok 4 0709$3.00$15.00256K
Grok 4 Latest$3.00$15.00256K
Grok Beta$5.00$15.00131K
Grok Vision Beta$5.00$15.008K
Grok 3 Fast Beta$5.00$25.00131K
Grok 3 Fast Latest$5.00$25.00131K

Frequently asked questions

How much does xAI (Grok) cost per 1M tokens?

xAI (Grok) pricing ranges from $0.500 to $25.00 per 1 million output tokens across 45 models. Input tokens are cheaper than output on every model. Rates current as of July 20, 2026.

What is the cheapest xAI (Grok) model?

Grok 3 Mini is the cheapest xAI (Grok) model on output tokens at $0.500 per 1M, while Grok 4 Fast is cheapest on input at $0.200 per 1M. Which wins for you depends on your input-to-output ratio.

What is the largest xAI (Grok) context window?

Grok 4 Fast has the largest context window in the xAI (Grok) catalog at 2.0M input tokens, with up to 16K output tokens per response.

Does xAI (Grok) support prompt caching?

Yes. Grok 4.20 (Non-Reasoning) reads cached input at $0.200 per 1M tokens, against $1.25 per 1M uncached. Caching pays off when a long system prompt or document is reused across many requests.

Which xAI (Grok) models support reasoning or vision?

The xAI (Grok) catalog includes 5 reasoning models, 6 vision models, 5 with tool calling. Reasoning models bill their internal thinking as output tokens, so they cost more per visible response than the headline rate suggests.

How do I calculate my actual xAI (Grok) bill?

Multiply your input tokens by the input rate and your output tokens by the output rate, both per million, then add retries and any conversation history resent on each turn. Calcaas does this against your real usage pattern and shows the margin left at your price point.

Head-to-head comparisons

Other providers

What will this actually cost you?

Per-token rates only tell you half the story. Calcaas multiplies them by your real usage — prompt size, response length, retries and conversation history — then shows the margin left at your price point.

Open the free cost calculator