Cohere Pricing
Cohere charges from $0.150 per 1 million output tokens, depending on the model. Command R7B is the cheapest at $0.150/1M output and $0.037/1M input; its current flagship Command A Plus costs $10.00/1M output. The largest context window is 256K tokens (Command A). Prices are USD, current as of August 11, 2026.
Last updated · synced weekly from the upstream model catalog
Cheapest input
$0.037/1M
Command R7B
Cheapest output
$0.150/1M
Command R7B
Longest context
256K
Command A
Models priced
16
Avg $6.39/1M output
Pricing by model
All prices in USD per 1 million tokens. All 16 priced models, cheapest first.
| Model | Input / 1M | Output / 1M | Context |
|---|---|---|---|
| Command R7B Tools | $0.037 | $0.150 | 128K |
| Command R7B (12-2024) | $0.037 | $0.150 | 128K |
| Command R7B Arabic Tools | $0.037 | $0.150 | 128K |
| Command Light | $0.300 | $0.600 | 4K |
| Command R Tools | $0.150 | $0.600 | 128K |
| Command R (08-2024) | $0.150 | $0.600 | 128K |
| Command A | $2.50 | $10.00 | 256K |
| Command A 03 2025 | $2.50 | $10.00 | 256K |
| Command A Plus VisionReasoningTools | $2.50 | $10.00 | 128K |
| Command A Reasoning ReasoningTools | $2.50 | $10.00 | 256K |
| Command A Translate Tools | $2.50 | $10.00 | 8K |
| Command A Vision Vision | $2.50 | $10.00 | 128K |
| Command R Plus | $2.50 | $10.00 | 128K |
| Command R Plus 08 2024 | $2.50 | $10.00 | 128K |
| Command R+ Tools | $2.50 | $10.00 | 128K |
| Command R+ (08-2024) | $2.50 | $10.00 | 128K |
Frequently asked questions
How much does Cohere cost per 1M tokens?
Cohere pricing starts at $0.150 per 1 million output tokens (Command R7B) across 16 models, with its current flagship Command A Plus at $10.00. Older premium models in the catalog list higher. Rates current as of August 11, 2026.
What is the cheapest Cohere model?
Command R7B is the cheapest Cohere model on both axes — $0.037 per 1M input tokens and $0.150 per 1M output. It is the right default for high-volume, low-complexity work such as classification, extraction and routing.
What is the largest Cohere context window?
Command A has the largest context window in the Cohere catalog at 256K input tokens, with up to 16K output tokens per response.
Which Cohere models support reasoning or vision?
The Cohere catalog includes 2 reasoning models, 2 vision models, 7 with tool calling. Reasoning models bill their internal thinking as output tokens, so they cost more per visible response than the headline rate suggests.
How do I calculate my actual Cohere bill?
Multiply your input tokens by the input rate and your output tokens by the output rate, both per million, then add retries and any conversation history resent on each turn. Calcaas does this against your real usage pattern and shows the margin left at your price point.
Other providers
What will this actually cost you?
Per-token rates only tell you half the story. Calcaas multiplies them by your real usage — prompt size, response length, retries and conversation history — then shows the margin left at your price point.
Open the free cost calculator