← Back to all pricing

GLM-5.3 Flash API Pricing — $0.15/M input, $0.50/M output

Native multimodal Flash: beats GLM-5.2 at a fraction of the cost with vision built into the coding loop.

Full specs, context window and API limits →

How much does GLM-5.3 Flash cost per million tokens?

GLM-5.3 Flash, from Z.ai, charges $0.15 per million input tokens and $0.50 per million output tokens — $0.24 per million blended at a 3:1 input:output ratio. That makes it pricier than Qwen 3.8 Flash on a blended-token basis.

Verified 2026-10-01 — source
Input
$0.15/M
Output
$0.50/M
Blended
$0.24/M
Provider
Verified 2026-10-01 — source

How much does GLM-5.3 Flash cost per 1,000 requests?

Computed from generated token pricing. Each row assumes the listed input and output tokens per request; this model has no measured verbosity factor, so the unadjusted output estimate is shown.

Request shapeInput tokensOutput tokensCost / 1,000 requests
Short10050$0.0400
Medium1,000500$0.4000
Long4,0002,000$1.6000

Formula: ((input price × input tokens) + (output price × output tokens × verbosity factor)) ÷ 1,000,000 × 1,000. Assumptions: short 100/50, medium 1,000/500, long 4,000/2,000 input/output tokens per request. Unadjusted — no measured verbosity factor is available.

How fast is GLM-5.3 Flash?

Not yet measured — see the speed benchmark leaderboard for models we do track.

How much does GLM-5.3 Flash cost at scale?

Tokens / monthEst. cost (blended 3:1)
100,000$0.02
1,000,000$0.24
10,000,000$2.38
100,000,000$23.75

How does GLM-5.3 Flash compare with other models?

GLM-5.1 — $1.00/MGLM-5.3 — $2.15/MGLM-5.2 — $2.15/MQwen 3.8 Flash — $0.23/MGPT-4o Mini — $0.26/MGrok-3 Mini — $0.26/M
See all Z.ai models →

What is GLM-5.3 Flash best for?

#4 for Math & Reasoning#4 for Agents & Tool Use#4 for Image Understanding
Looking for a cheaper option?
Qwen 3.8 Flash is 3.2% cheaper — a code-change migration. See all 8 alternatives to GLM-5.3 Flash →

What are common questions about GLM-5.3 Flash?

Is GLM-5.3 Flash cheaper than Qwen 3.8 Flash?

GLM-5.3 Flash costs $0.24/M blended tokens, Qwen 3.8 Flash costs $0.23/M — Qwen 3.8 Flash is cheaper.

How much does 1 million tokens cost with GLM-5.3 Flash?

At a 3:1 input:output ratio, 1 million blended tokens costs approximately $0.24. Pure input costs $0.15/M; pure output costs $0.50/M.

What does GLM-5.3 Flash cost at high volume?

At 100 million blended tokens a month, GLM-5.3 Flash costs approximately $23.75. See the cost-at-scale table below for other volumes.

Try GLM-5.3 Flash for free

Run real prompts against GLM-5.3 Flash and every other model on this page in one workspace.

Try GLM-5.3 Flash Free