API pricing

Qwen (Alibaba) Pricing

Qwen (Alibaba) charges between $0.090 and $6.40 per 1 million output tokens, depending on the model. Qwen2.5 Coder 7B Instruct is the cheapest at $0.090/1M output and $0.030/1M input; Qwen-Max is the most expensive at $6.40/1M output. The largest context window is 1.0M tokens (Qwen Plus 0728). Prices are USD, current as of July 20, 2026.

Alibaba prices Qwen well below frontier Western models, and much of the lineup is open-weight. Context windows run long across the family, which makes Qwen a common pick for document-heavy workloads on a budget.

Last updated · synced weekly from the upstream model catalog

Cheapest input

$0.030/1M

Qwen2.5 Coder 7B Instruct

Cheapest output

$0.090/1M

Qwen2.5 Coder 7B Instruct

Longest context

1.0M

Qwen Plus 0728

Models priced

39

Avg $1.17/1M output

Pricing by model

All prices in USD per 1 million tokens. All 39 priced models, cheapest first.

ModelInput / 1MOutput / 1MContext
Qwen2.5 Coder 7B Instruct$0.030$0.09033K
Qwen2.5 7B Instruct$0.040$0.10033K
Qwen3 8B$0.035$0.138128K
Qwen2.5 Coder 32B Instruct$0.040$0.16033K
Qwen-Turbo$0.050$0.2001.0M
Qwen2.5-VL 7B Instruct$0.200$0.20033K
Qwen3 32B$0.050$0.20041K
Qwen2.5 VL 32B Instruct$0.050$0.22016K
Qwen3 14B$0.050$0.22041K
Qwen3 30B A3B$0.060$0.22041K
Qwen3 Coder 30B A3B Instruct$0.060$0.250262K
Qwen2.5 72B Instruct$0.070$0.26033K
Qwen3 30B A3B Thinking 2507$0.090$0.300262K
Qwen2.5 VL 72B Instruct$0.080$0.33033K
Qwen3 30B A3B Instruct 2507$0.080$0.330262K
QwQ 32B$0.150$0.40033K
Tongyi DeepResearch 30B A3B$0.090$0.400131K
Qwen3 VL 8B Instruct$0.080$0.500131K
Qwen3 235B A22B$0.180$0.54041K
Qwen3 235B A22B Instruct 2507$0.080$0.550262K
Qwen3 235B A22B Thinking 2507$0.110$0.600262K
Qwen3 VL 30B A3B Instruct$0.150$0.600262K
Qwen VL Plus$0.210$0.6308K
Qwen3 Next 80B A3B Instruct$0.100$0.800262K
Qwen3 VL 235B A22B Instruct$0.220$0.880262K
Qwen3 Coder 480B A35B$0.220$0.950262K
Qwen3 VL 30B A3B Thinking$0.200$1.00131K
Qwen Plus 0728$0.400$1.201.0M
Qwen-Plus$0.400$1.20131K
Qwen3 Next 80B A3B Thinking$0.150$1.20262K
Qwen3 VL 235B A22B Thinking$0.300$1.20262K
Qwen3 Coder Flash$0.300$1.50128K
Qwen3 Coder 480B A35B (exacto)$0.380$1.53262K
Qwen3 VL 8B Thinking$0.180$2.10256K
Qwen VL Max$0.800$3.20131K
Qwen Plus 0728 (thinking)$0.400$4.001.0M
Qwen3 Coder Plus$1.00$5.00128K
Qwen3 Max$1.20$6.00256K
Qwen-Max$1.60$6.4033K

Frequently asked questions

How much does Qwen (Alibaba) cost per 1M tokens?

Qwen (Alibaba) pricing ranges from $0.090 to $6.40 per 1 million output tokens across 39 models. Input tokens are cheaper than output on every model. Rates current as of July 20, 2026.

What is the cheapest Qwen (Alibaba) model?

Qwen2.5 Coder 7B Instruct is the cheapest Qwen (Alibaba) model on both axes — $0.030 per 1M input tokens and $0.090 per 1M output. It is the right default for high-volume, low-complexity work such as classification, extraction and routing.

What is the largest Qwen (Alibaba) context window?

Qwen Plus 0728 has the largest context window in the Qwen (Alibaba) catalog at 1.0M input tokens, with up to 16K output tokens per response.

How do I calculate my actual Qwen (Alibaba) bill?

Multiply your input tokens by the input rate and your output tokens by the output rate, both per million, then add retries and any conversation history resent on each turn. Calcaas does this against your real usage pattern and shows the margin left at your price point.

Head-to-head comparisons

Other providers

What will this actually cost you?

Per-token rates only tell you half the story. Calcaas multiplies them by your real usage — prompt size, response length, retries and conversation history — then shows the margin left at your price point.

Open the free cost calculator