GLM 5.2 vs GLM 5.3 Flash

Price, context window & benchmark comparison

Z

GLM 5.2

Z.ai (Zhipu)
$0.65 / $2.04
Input / Output · per 1M tokens
VS
Z

GLM 5.3 Flash

Z.ai (Zhipu)
$0.15 / $0.50
Input / Output · per 1M tokens
Compare different models
Z
Z
Verdict: GLM 5.3 Flash is 4.2× cheaper than GLM 5.2 on a typical 3:1 input/output mix. GLM 5.3 Flash scores higher on the Artificial Analysis Intelligence Index (41.8 vs 33.7). GLM 5.3 Flash has the larger context window (1.3M vs 1M).
GLM 5.2GLM 5.3 Flash
Input price / 1M$0.65$0.15
Output price / 1M$2.04$0.50
Blended price (3:1)$1.00$0.24
Cached input / 1M$0.12$0.05
Context window1M1.3M
Max output131K944K
Intelligence index33.741.8
Coding index68.871.5
Providers31
Best provider uptime (30m)100.0%
Vision inputNoYes
ReasoningYesYes
Tool callingYesYes
Released2026-06-162026-08-26
Retirement

Cost calculator

FAQ

Which is cheaper, GLM 5.2 or GLM 5.3 Flash?

GLM 5.3 Flash is 4.2× cheaper than GLM 5.2 on a typical 3:1 input/output mix.

GLM 5.2 vs GLM 5.3 Flash: which has a bigger context window?

GLM 5.2: 1,048,576 tokens. GLM 5.3 Flash: 1,310,720 tokens.

How much would 1 million requests cost?

At 2,000 input + 500 output tokens per request: GLM 5.2 ≈ $2,320, GLM 5.3 Flash ≈ $550.00.

More comparisons