Qwen API pricing
How Qwen API charges, what to know before you commit, and what you'd pay at your usage — next to what the alternatives would cost for the same thing.
How it charges
Checked 2026-10-07 on www.alibabacloud.com ↗ · USD, list pricesIn short
Pay per token, from $0.15 / $0.47 per million input/output tokens on Qwen3.8-Flash to $2 / $6 on Qwen3.8-Max (international region).
- Free plan
- Yes
- Side project
- $1.22/mo Qwen3.8-Flash
- Growing
- $12/mo Qwen3.8-Flash
- Scaling
- $244/mo Qwen3.8-Flash
What the model leaves out
International (Singapore) list prices for the lowest input-length tier. Qwen3.7-Plus is marked "limited-time 20% off" on the page; the list price is shown. Batch calls are half price.
Before you commit
- Prices are tiered by request size: a request's total input tokens pick the tier, and every token in that request is billed at it (Qwen3.7-Plus: $0.40 / $1.60 up to 256K input, $1.20 / $4.80 above).
- Explicit context-cache writes cost 125% of the input price; cache hits cost 10%.
- The free quota (1 million tokens per model for 90 days) applies to the Singapore region only.
- Mainland-China (Beijing) region prices are in yuan and differ from the international ones.
| Plan | Monthly fee | Input tokens per month | Output tokens per month |
|---|---|---|---|
| Qwen3.8-Max | None | $2 per million tokens | $6 per million tokens |
| Qwen3.7-Plus | None | $0.4 per million tokens | $1.6 per million tokens |
| Qwen3.8-Flash | None | $0.15 per million tokens | $0.47 per million tokens |
What you'd pay
Qwen3.8-Flash50 million tokens × $0.15 per million tokens + 10 million tokens × $0.47 per million tokens$12/mo
Cost as you grow
All LLM API pricing, compared →Which one fits your situation →