// pricing · September 2026
Qwen API Pricing
Every Qwen model's list price per 1M tokens — input, output and cached input — with context windows, LiveBench scores and what a real workload costs per month.
52 Qwen models are listed, with a median list price of $0.205 per 1M input tokens and $1.11 per 1M output tokens. Paid prices run from $0.055 (Qwen3.7 Flash) to $3.00 (Qwen3.8 2.4T A95B) per 1M tokens blended. 1 has a free listing.
- Models
- 52
- 1 with a free listing
- Median input / 1M
- $0.205
- paid models
- Median output / 1M
- $1.11
- paid models
- Cheapest paid
- $0.055
- Qwen3.7 Flash
- Highest LiveBench
- 75.3
- Qwen3.8 27B
Qwen model prices
List prices in USD per 1M tokens. “Blended” is a 3:1 input:output mix; the per-month column prices the selected workload using each model's cached-input rate where it publishes one.
Showing 52 of 52 models · per-month column: 8K in / 600 out × 100K requests, 50% of input cached
| qwen/qwen3.7-flash | $0.030 | $0.130 | $0.0060 | $0.055 | 1M | — | $22.20 |
|---|---|---|---|---|---|---|---|
| qwen/qwen3-30b-a3b-instruct-2507 | $0.048 | $0.193 | — | $0.084 | 262K | — | $50.10 |
| qwen/qwen3.5-9b | $0.100 | $0.150 | — | $0.112 | 262K | — | $89.00 |
| qwen/qwen3.5-flash-02-23 | $0.065 | $0.260 | — | $0.114 | 1M | — | $67.60 |
| qwen/qwen3-coder-30b-a3b-instruct | $0.070 | $0.280 | — | $0.123 | 262K | — | $72.80 |
| qwen/qwen-2.5-7b-instruct | $0.100 | $0.200 | — | $0.125 | 33K | — | $92.00 |
| qwen/qwen3-32b | $0.080 | $0.280 | — | $0.130 | 131K | — | $80.80 |
| qwen/qwen3-14b | $0.120 | $0.240 | — | $0.150 | 131K | — | $110.40 |
| qwen/qwen3-235b-a22b-2507 | $0.087 | $0.350 | $0.018 | $0.153 | 262K | — | $63.00 |
| qwen/qwen3-vl-32b-instruct | $0.104 | $0.416 | — | $0.182 | 131K | — | $108.16 |
| qwen/qwen3-vl-8b-instruct | $0.117 | $0.455 | — | $0.202 | 262K | — | $120.90 |
| qwen/qwen3-8b | $0.117 | $0.455 | — | $0.202 | 131K | — | $120.90 |
| qwen/qwen3-30b-a3b | $0.120 | $0.500 | — | $0.215 | 131K | — | $126.00 |
| qwen/qwen3-vl-30b-a3b-instruct | $0.130 | $0.520 | — | $0.228 | 262K | — | $135.20 |
| qwen/qwen3.8-omni-flash | $0.150 | $0.470 | $0.016 | $0.230 | 1M | — | $94.60 |
| qwen/qwen3.8-flash | $0.150 | $0.470 | $0.016 | $0.230 | 1M | — | $94.60 |
| qwen/qwen3-coder-next | $0.120 | $0.800 | $0.070 | $0.290 | 262K | — | $124.00 |
| qwen/qwen3-next-80b-a3b-instruct | $0.090 | $1.10 | — | $0.343 | 262K | — | $138.00 |
| qwen/qwen3.6-35b-a3b | $0.150 | $1.00 | $0.050 | $0.362 | 262K | — | $140.00 |
| qwen/qwen-2.5-72b-instruct | $0.360 | $0.400 | — | $0.370 | 33K | — | $312.00 |
| qwen/qwen3-coder-flash | $0.195 | $0.975 | $0.039 | $0.390 | 1M | — | $152.10 |
| qwen/qwen-plus-2025-07-28 | $0.260 | $0.780 | — | $0.390 | 1M | — | $254.80 |
| qwen/qwen-plus | $0.260 | $0.780 | $0.052 | $0.390 | 1M | — | $171.60 |
| qwen/qwen3-next-80b-a3b-thinking | $0.150 | $1.20 | — | $0.412 | 262K | — | $192.00 |
| qwen/qwen3.6-flash | $0.188 | $1.13 | — | $0.422 | 1M | — | $217.50 |
| qwen/qwen3-coder | $0.300 | $1.00 | $0.100 | $0.475 | 262K | — | $220.00 |
| qwen/qwen3.5-27b | $0.195 | $1.56 | — | $0.536 | 262K | — | $249.60 |
| qwen/qwen3.5-35b-a3b | $0.313 | $1.25 | $0.156 | $0.547 | 262K | — | $262.50 |
| qwen/qwen3.7-plus | $0.320 | $1.28 | $0.064 | $0.560 | 1M | — | $230.40 |
| qwen/qwen3.5-plus-02-15 | $0.260 | $1.56 | — | $0.585 | 1M | — | $301.60 |
| qwen/qwen3-vl-235b-a22b-instruct | $0.210 | $1.90 | $0.100 | $0.632 | 262K | — | $238.00 |
| qwen/qwen3-vl-8b-thinking | $0.180 | $2.10 | — | $0.660 | 131K | — | $270.00 |
| qwen/qwen3.5-plus-20260420 | $0.300 | $1.80 | — | $0.675 | 1M | — | $348.00 |
| qwen/qwen3.5-122b-a10b | $0.260 | $2.08 | — | $0.715 | 262K | — | $332.80 |
| qwen/qwen3.6-plus | $0.325 | $1.95 | — | $0.731 | 1M | 68.9 | $377.00 |
| qwen/qwen-2.5-coder-32b-instruct | $0.660 | $1.00 | — | $0.745 | 33K | — | $588.00 |
| qwen/qwen3-235b-a22b-thinking-2507 | $0.230 | $2.30 | — | $0.747 | 131K | — | $322.00 |
| qwen/qwen3-vl-30b-a3b-thinking | $0.200 | $2.40 | — | $0.750 | 262K | — | $304.00 |
| qwen/qwen3-30b-a3b-thinking-2507 | $0.200 | $2.40 | — | $0.750 | 82K | — | $304.00 |
| qwen/qwen3-235b-a22b | $0.455 | $1.82 | — | $0.796 | 131K | — | $473.20 |
| qwen/qwen2.5-vl-72b-instruct | $0.800 | $1.00 | $0.400 | $0.850 | 128K | — | $540.00 |
| qwen/qwen3.6-27b | $0.320 | $2.70 | $0.150 | $0.915 | 262K | 64.0 | $350.00 |
| qwen/qwen3.8-27b | $0.420 | $3.00 | $0.085 | $1.06 | 1M | 75.3 | $382.00 |
| qwen/qwen3.5-397b-a17b | $0.550 | $3.50 | $0.225 | $1.29 | 262K | — | $520.00 |
| qwen/qwen3-vl-235b-a22b-thinking | $0.400 | $4.00 | — | $1.30 | 131K | — | $560.00 |
| qwen/qwen3-coder-plus | $0.650 | $3.25 | $0.130 | $1.30 | 1M | — | $507.00 |
| qwen/qwen3-max-thinking | $0.780 | $3.90 | — | $1.56 | 262K | — | $858.00 |
| qwen/qwen3-max | $0.780 | $3.90 | $0.156 | $1.56 | 262K | — | $608.40 |
| qwen/qwen3.7-max | $1.48 | $4.42 | $0.295 | $2.21 | 1M | 73.1 | $973.50 |
| qwen/qwen3.6-max-preview | $1.03 | $6.16 | — | $2.31 | 262K | — | $1,191 |
| qwen/qwen3.8-max-0902 | $2.00 | $6.00 | $0.250 | $3.00 | 1M | — | $1,260 |
| qwen/qwen3.8-2.4t-a95b | $2.00 | $6.00 | $0.250 | $3.00 | 1.0M | — | $1,260 |
Qwen pricing FAQ
What is the cheapest Qwen model?
Qwen3.7 Flash is the cheapest paid Qwen model, at $0.030 per 1M input tokens and $0.130 per 1M output tokens. 1 Qwen models also have a free listing, usually with tighter rate limits.
How much does the most capable Qwen model cost?
The highest LiveBench overall score among Qwen models belongs to Qwen3.8 27B (75.3), priced at $0.420 input and $3.00 output per 1M tokens. A RAG assistant at 100K requests a month would list at about $382.00.
Why do output tokens cost more than input tokens?
Generating a token takes a full forward pass per token, while input is processed in parallel. Across the paid Qwen models listed here the median output price is $1.11 per 1M against $0.205 for input — reasoning tokens bill as output too.
Which Qwen models support prompt caching?
21 of 52 publish a separate cached-input price, which bills repeated prompt prefixes at a fraction of the normal input rate. Filter the table by "Caching" to see them.
How current are these prices?
List prices are pulled from OpenRouter’s public models API and refreshed every 15 minutes. Batch discounts and enterprise pricing are not included.
Other providers
What these prices include
- List prices from OpenRouter's public models API, which mirrors provider pricing. Batch discounts and negotiated enterprise rates are not included.
- Variants —
:freeand:batchlistings are folded into their base model; its model page lists them. - Per-month cost is a floor: tokens you describe times list price. Cache writes and reasoning-token volume depend on your workload.