Qwen API Cost Calculator
How much does the Qwen API cost per month?
For Qwen 3.7 Plus at 200K requests/month, 2,400 input tokens and 350 output tokens per request costs about $524 per month after 30% cache use and 100% batch share. Across Qwen's 3 priced models, the cheapest default ranking is Qwen 3.7 Plus at $524 per month.
Pricing data as of June 2026. Sources: Qwen pricing and model documentation. Parameters are shareable in this URL.
Qwen 3.7 Plus: estimated monthly cost
Formula: provider-specific cached input multiplier × cache rate, plus uncached input and verbosity-adjusted output; cache writes are amortised over the provider TTL, then batch savings are applied.
| Per request | $0.003 |
|---|---|
| Per day | $17.47 |
| Per month | $524 |
| Per year | $6,288 |
OpenAI scenario sensitivity
The default is a server-rendered estimate. Change the cache and batch shares to see when a cheaper qualified model overtakes the selected model; the URL is shareable and preserves the inputs.
| Default | $524/month |
|---|---|
| No caching or batch | $524/month |
| 50% cache, 50% batch | $524/month |
| 100% cache, 100% batch | $524/month |
Batch 11 Qwen workload decision depth
1. Plus/Max context eligibility and cost ladder
| Input tokens | Model | Cost | Window result |
|---|---|---|---|
| 32K | Qwen 3.7 Plus | $0.03 | Fit only when model context window covers full request; otherwise excluded |
| 32K | Qwen 3.8 Max | $0.05 | Fit only when model context window covers full request; otherwise excluded |
| 32K | Qwen 3.7 Max | $0.05 | Fit only when model context window covers full request; otherwise excluded |
| 128K | Qwen 3.7 Plus | $0.10 | Fit only when model context window covers full request; otherwise excluded |
| 128K | Qwen 3.8 Max | $0.21 | Fit only when model context window covers full request; otherwise excluded |
| 128K | Qwen 3.7 Max | $0.21 | Fit only when model context window covers full request; otherwise excluded |
| 500K | Qwen 3.7 Plus | $0.40 | Fit only when model context window covers full request; otherwise excluded |
| 500K | Qwen 3.8 Max | $0.80 | Fit only when model context window covers full request; otherwise excluded |
| 500K | Qwen 3.7 Max | $0.80 | Fit only when model context window covers full request; otherwise excluded |
| 1M | Qwen 3.7 Plus | $0.80 | Fit only when model context window covers full request; otherwise excluded |
| 1M | Qwen 3.8 Max | $1.60 | Fit only when model context window covers full request; otherwise excluded |
| 1M | Qwen 3.7 Max | $1.60 | Fit only when model context window covers full request; otherwise excluded |
2. Output-heavy coding/agent sensitivity
| Output multiplier | Model | Bill | Uplift required |
|---|---|---|---|
| 1× | Qwen 3.7 Plus | $524.00 | Accepted-result uplift for premium tier: User-supplied |
| 1× | Qwen 3.8 Max | $1216.00 | Accepted-result uplift for premium tier: User-supplied |
| 1× | Qwen 3.7 Max | $1216.00 | Accepted-result uplift for premium tier: User-supplied |
| 2× | Qwen 3.7 Plus | $664.00 | Accepted-result uplift for premium tier: User-supplied |
| 2× | Qwen 3.8 Max | $1664.00 | Accepted-result uplift for premium tier: User-supplied |
| 2× | Qwen 3.7 Max | $1664.00 | Accepted-result uplift for premium tier: User-supplied |
| 4× | Qwen 3.7 Plus | $944.00 | Accepted-result uplift for premium tier: User-supplied |
| 4× | Qwen 3.8 Max | $2560.00 | Accepted-result uplift for premium tier: User-supplied |
| 4× | Qwen 3.7 Max | $2560.00 | Accepted-result uplift for premium tier: User-supplied |
3. Direct-versus-hosted-alias break-even ceiling
| Model | Dated Alibaba spend | External premium ceiling | Region/currency |
|---|---|---|---|
| Qwen 3.7 Plus | $524.00 | External monthly bill ≤ direct spend; exact alias rate: Unavailable | Region and currency: Unavailable |
| Qwen 3.8 Max | $1216.00 | External monthly bill ≤ direct spend; exact alias rate: Unavailable | Region and currency: Unavailable |
| Qwen 3.7 Max | $1216.00 | External monthly bill ≤ direct spend; exact alias rate: Unavailable | Region and currency: Unavailable |
Provenance: Batch 11 Qwen direct/hosted context module; selected model Qwen 3.7 Plus; inputs are 2,400 input tokens, 350 output tokens, 200,000 calls/month, cache rate 30%, batch share 100%. Verified 2026-06-21. Provider source: https://www.alibabacloud.com/help/en/model-studio/models · Run this scenario →
Ranked cost — 200K requests/month
| Model | Provider | List monthly | Effective monthly | Verbosity | Rank Δ |
|---|---|---|---|---|---|
| Qwen 3.7 Plusmid | Qwen | $524 | $524 | — | — |
| Qwen 3.8 Maxmid | Qwen | $1,216 | $1,216 | — | — |
| Qwen 3.7 Maxmid | Qwen | $1,216 | $1,216 | — | — |
Levers live on Qwen
The billing gotcha
Qwen's direct Model Studio API and its hosted aliases are separate billing surfaces: this page prices the Qwen models sold through Alibaba Cloud, not similarly named Qwen weights on Groq or another host. Region matters too — the international DashScope endpoint and mainland China service can expose different quotas and currency treatment, so the monthly estimate below is a token-rate projection, not an AWS-style all-in account bill.
