API pricing

xAI (Grok) Pricing

xAI (Grok) charges from $0.500 per 1 million output tokens, depending on the model. Grok 3 Mini is the cheapest at $0.500/1M output and $0.300/1M input; its current flagship Grok 4.6 costs $6.00/1M output. The largest context window is 2.0M tokens (Grok 4 Fast). Prices are USD, current as of August 31, 2026.

xAI prices Grok per million tokens, with the fast and mini variants sitting well below the flagship. Context windows are large across the lineup, which makes prompt size less of a cost lever here than output length.

Last updated · synced weekly from the upstream model catalog

Cheapest input

$0.200/1M

Grok 4 Fast

Cheapest output

$0.500/1M

Grok 3 Mini

Longest context

2.0M

Grok 4 Fast

Models priced

41

Avg $3.46/1M output

Pricing by model

All prices in USD per 1 million tokens. All 41 priced models, cheapest first.

ModelInput / 1MOutput / 1MContext
Grok 3 Mini$0.300$0.500131K
Grok 3 Mini Beta$0.300$0.500131K
Grok 4 Fast$0.200$0.5002.0M
Grok Code Fast 1$0.200$1.50256K
Grok Build 0.1
VisionReasoningToolsCache
$1.00$2.00256K
Grok Code Fast$1.00$2.00256K
Grok Code Fast 1 0825$1.00$2.00256K
Grok 3 Fast Beta$1.25$2.50131K
Grok 3 Fast Latest$1.25$2.50131K
Grok 3 Latest$1.25$2.50131K
Grok 3 Mini Fast$1.25$2.50131K
Grok 3 Mini Fast Beta$1.25$2.50131K
Grok 3 Mini Fast Latest$1.25$2.50131K
Grok 3 Mini Latest$1.25$2.50131K
Grok 4 0709$1.25$2.50256K
Grok 4 Fast Non Reasoning$1.25$2.502.0M
Grok 4 Fast Reasoning$1.25$2.502.0M
Grok 4 Latest$1.25$2.50256K
Grok 4.1 Fast$1.25$2.502.0M
Grok 4.1 Fast Non Reasoning$1.25$2.502.0M
Grok 4.1 Fast Non Reasoning Latest$1.25$2.502.0M
Grok 4.1 Fast Reasoning$1.25$2.502.0M
Grok 4.1 Fast Reasoning Latest$1.25$2.502.0M
Grok 4.20$1.25$2.501.0M
Grok 4.20 (Non-Reasoning)
VisionToolsCache
$1.25$2.501.0M
Grok 4.20 (Reasoning)
VisionReasoningToolsCache
$1.25$2.501.0M
Grok 4.20 0309 Non Reasoning$1.25$2.501.0M
Grok 4.20 0309 Reasoning$1.25$2.501.0M
Grok 4.20 Beta 0309 Non Reasoning$1.25$2.501.0M
Grok 4.20 Beta 0309 Reasoning$1.25$2.501.0M
Grok 4.20 Multi-Agent
VisionReasoningCache
$1.25$2.501.0M
Grok 4.20 Non Reasoning Latest$1.25$2.501.0M
Grok 4.20 Reasoning Latest$1.25$2.501.0M
Grok 4.3
VisionReasoningToolsCache
$1.25$2.501.0M
Grok 4.3 Latest$1.25$2.501.0M
Grok 4.5
VisionReasoningToolsCache
$2.00$6.00500K
Grok 4.5 Latest$2.00$6.00500K
Grok 4.6
VisionReasoningToolsCache
$2.00$6.00500K
Grok 3$3.00$15.00131K
Grok 3 Beta$3.00$15.00131K
Grok 4$3.00$15.00256K

Frequently asked questions

How much does xAI (Grok) cost per 1M tokens?

xAI (Grok) pricing starts at $0.500 per 1 million output tokens (Grok 3 Mini) across 41 models, with its current flagship Grok 4.6 at $6.00. Older premium models in the catalog list higher. Rates current as of August 31, 2026.

What is the cheapest xAI (Grok) model?

Grok 3 Mini is the cheapest xAI (Grok) model on output tokens at $0.500 per 1M, while Grok 4 Fast is cheapest on input at $0.200 per 1M. Which wins for you depends on your input-to-output ratio.

What is the largest xAI (Grok) context window?

Grok 4 Fast has the largest context window in the xAI (Grok) catalog at 2.0M input tokens, with up to 16K output tokens per response.

Does xAI (Grok) support prompt caching?

Yes. Grok 4.20 (Non-Reasoning) reads cached input at $0.200 per 1M tokens, against $1.25 per 1M uncached. Caching pays off when a long system prompt or document is reused across many requests.

Which xAI (Grok) models support reasoning or vision?

The xAI (Grok) catalog includes 6 reasoning models, 7 vision models, 6 with tool calling. Reasoning models bill their internal thinking as output tokens, so they cost more per visible response than the headline rate suggests.

How do I calculate my actual xAI (Grok) bill?

Multiply your input tokens by the input rate and your output tokens by the output rate, both per million, then add retries and any conversation history resent on each turn. Calcaas does this against your real usage pattern and shows the margin left at your price point.

Head-to-head comparisons

Other providers

What will this actually cost you?

Per-token rates only tell you half the story. Calcaas multiplies them by your real usage — prompt size, response length, retries and conversation history — then shows the margin left at your price point.

Open the free cost calculator