Qwen API pricing

How Qwen API charges, what to know before you commit, and what you'd pay at your usage — next to what the alternatives would cost for the same thing.

How it charges

Checked 2026-10-07 on www.alibabacloud.com ↗ · USD, list prices
In short

Pay per token, from $0.15 / $0.47 per million input/output tokens on Qwen3.8-Flash to $2 / $6 on Qwen3.8-Max (international region).

Free plan
Yes
Side project
$1.22/mo Qwen3.8-Flash
Growing
$12/mo Qwen3.8-Flash
Scaling
$244/mo Qwen3.8-Flash
What the model leaves out

International (Singapore) list prices for the lowest input-length tier. Qwen3.7-Plus is marked "limited-time 20% off" on the page; the list price is shown. Batch calls are half price.

Before you commit
  • Prices are tiered by request size: a request's total input tokens pick the tier, and every token in that request is billed at it (Qwen3.7-Plus: $0.40 / $1.60 up to 256K input, $1.20 / $4.80 above).
  • Explicit context-cache writes cost 125% of the input price; cache hits cost 10%.
  • The free quota (1 million tokens per model for 90 days) applies to the Singapore region only.
  • Mainland-China (Beijing) region prices are in yuan and differ from the international ones.
PlanMonthly feeInput tokens per monthOutput tokens per month
Qwen3.8-MaxNone$2 per million tokens$6 per million tokens
Qwen3.7-PlusNone$0.4 per million tokens$1.6 per million tokens
Qwen3.8-FlashNone$0.15 per million tokens$0.47 per million tokens

What you'd pay

Cost as you grow

$0$100$500$1,000$2,0001510501005001k5k
Qwen APIDeepSeekMistral AIGroqGemini APIClaudeOpenAI APIx: input tokens per month (million tokens), other usage scaled with it · cheapest usable plan at each point, list prices