Qwen API Cost Calculator

How much does the Qwen API cost per month?

For Qwen 3.7 Plus at 200K requests/month, 2,400 input tokens and 350 output tokens per request costs about $524 per month after 30% cache use and 100% batch share. Across Qwen's 3 priced models, the cheapest default ranking is Qwen 3.7 Plus at $524 per month.

Verified 2026-06-21

Pricing data as of June 2026. Sources: Qwen pricing and model documentation. Parameters are shareable in this URL.

Qwen 3.7 Plus: estimated monthly cost

$524

Formula: provider-specific cached input multiplier × cache rate, plus uncached input and verbosity-adjusted output; cache writes are amortised over the provider TTL, then batch savings are applied.

Per request$0.003
Per day$17.47
Per month$524
Per year$6,288

OpenAI scenario sensitivity

The default is a server-rendered estimate. Change the cache and batch shares to see when a cheaper qualified model overtakes the selected model; the URL is shareable and preserves the inputs.

Default$524/month
No caching or batch$524/month
50% cache, 50% batch$524/month
100% cache, 100% batch$524/month

Batch 11 Qwen workload decision depth

1. Plus/Max context eligibility and cost ladder

Input tokensModelCostWindow result
32KQwen 3.7 Plus$0.03Fit only when model context window covers full request; otherwise excluded
32KQwen 3.8 Max$0.05Fit only when model context window covers full request; otherwise excluded
32KQwen 3.7 Max$0.05Fit only when model context window covers full request; otherwise excluded
128KQwen 3.7 Plus$0.10Fit only when model context window covers full request; otherwise excluded
128KQwen 3.8 Max$0.21Fit only when model context window covers full request; otherwise excluded
128KQwen 3.7 Max$0.21Fit only when model context window covers full request; otherwise excluded
500KQwen 3.7 Plus$0.40Fit only when model context window covers full request; otherwise excluded
500KQwen 3.8 Max$0.80Fit only when model context window covers full request; otherwise excluded
500KQwen 3.7 Max$0.80Fit only when model context window covers full request; otherwise excluded
1MQwen 3.7 Plus$0.80Fit only when model context window covers full request; otherwise excluded
1MQwen 3.8 Max$1.60Fit only when model context window covers full request; otherwise excluded
1MQwen 3.7 Max$1.60Fit only when model context window covers full request; otherwise excluded

2. Output-heavy coding/agent sensitivity

Output multiplierModelBillUplift required
Qwen 3.7 Plus$524.00Accepted-result uplift for premium tier: User-supplied
Qwen 3.8 Max$1216.00Accepted-result uplift for premium tier: User-supplied
Qwen 3.7 Max$1216.00Accepted-result uplift for premium tier: User-supplied
Qwen 3.7 Plus$664.00Accepted-result uplift for premium tier: User-supplied
Qwen 3.8 Max$1664.00Accepted-result uplift for premium tier: User-supplied
Qwen 3.7 Max$1664.00Accepted-result uplift for premium tier: User-supplied
Qwen 3.7 Plus$944.00Accepted-result uplift for premium tier: User-supplied
Qwen 3.8 Max$2560.00Accepted-result uplift for premium tier: User-supplied
Qwen 3.7 Max$2560.00Accepted-result uplift for premium tier: User-supplied

3. Direct-versus-hosted-alias break-even ceiling

ModelDated Alibaba spendExternal premium ceilingRegion/currency
Qwen 3.7 Plus$524.00External monthly bill ≤ direct spend; exact alias rate: UnavailableRegion and currency: Unavailable
Qwen 3.8 Max$1216.00External monthly bill ≤ direct spend; exact alias rate: UnavailableRegion and currency: Unavailable
Qwen 3.7 Max$1216.00External monthly bill ≤ direct spend; exact alias rate: UnavailableRegion and currency: Unavailable

Provenance: Batch 11 Qwen direct/hosted context module; selected model Qwen 3.7 Plus; inputs are 2,400 input tokens, 350 output tokens, 200,000 calls/month, cache rate 30%, batch share 100%. Verified 2026-06-21. Provider source: https://www.alibabacloud.com/help/en/model-studio/models · Run this scenario →

Ranked cost — 200K requests/month

ModelProviderList monthlyEffective monthlyVerbosityRank Δ
Qwen 3.7 PlusmidQwen$524$524
Qwen 3.8 MaxmidQwen$1,216$1,216
Qwen 3.7 MaxmidQwen$1,216$1,216

Levers live on Qwen

Context trimming: up to 47%Batch API: up to 50%

The billing gotcha

Qwen's direct Model Studio API and its hosted aliases are separate billing surfaces: this page prices the Qwen models sold through Alibaba Cloud, not similarly named Qwen weights on Groq or another host. Region matters too — the international DashScope endpoint and mainland China service can expose different quotas and currency treatment, so the monthly estimate below is a token-rate projection, not an AWS-style all-in account bill.

Related

Qwen provider profile →Compare against another provider →Global cost calculator →How to reduce LLM API costs →