GLM 5.3 pricing & price history
GLM 5.3 by Z.ai (Zhipu) costs $0.84 per 1M input tokens and $2.64 per 1M output tokens, with a 1.3M token context window. The cheapest provider currently charges $0.56 input / $2.28 output.
Price history
No price changes since tracking began on 2026-09-24. We check hourly.
Providers serving GLM 5.3 (30)
| Provider | Input | Output | Context | Uptime 30m | Uptime 24h |
|---|---|---|---|---|---|
| DeepInfra fp4 | $0.56 | $2.50 | 1M | 92.9% | 95.3% |
| InferenceNet fp4 | $0.68 | $2.28 | 1M | 99.1% | 98.8% |
| Reka fp8 | $0.76 | $2.57 | 262K | 99.6% | 98.8% |
| Io Net fp8 | $0.77 | $2.60 | 262K | 99.9% | 99.9% |
| Sail Research fp8 | $0.77 | $4.00 | 1M | 100.0% | 99.4% |
| Morph | $0.77 | $2.43 | 1M | 100.0% | 99.7% |
| Novita fp8 | $0.78 | $2.46 | 1M | 99.2% | 99.7% |
| Phala | $0.84 | $2.64 | 1M | 98.9% | 99.8% |
| Inceptron fp4 | $0.90 | $3.53 | 1M | 99.9% | 99.4% |
| DigitalOcean | $0.91 | $2.86 | 1M | 99.8% | 99.3% |
| GMICloud fp8 | $0.98 | $3.08 | 1M | 98.3% | 99.5% |
| SiliconFlow fp8 | $1.12 | $3.52 | 1M | 99.7% | 96.9% |
| Alibaba | $1.19 | $3.74 | 1M | 99.8% | 98.6% |
| Decart fp4 | $1.19 | $3.74 | 1M | 99.8% | 99.5% |
| Sail Research fp8 | $1.21 | $3.87 | 1M | — | 99.6% |
| Friendli | $1.26 | $3.96 | 1M | 100.0% | 100.0% |
| AkashML fp8 | $1.30 | $4.40 | 1M | 100.0% | 99.9% |
| Baidu fp8 | $1.40 | $4.40 | 1M | 100.0% | 100.0% |
| BaseTen fp4 | $1.40 | $4.40 | 1M | 100.0% | 99.7% |
| Mistral nvfp4 | $1.40 | $4.40 | 1M | 100.0% | 99.9% |
| Crusoe fp4 | $1.40 | $4.40 | 1M | — | 99.6% |
| PrimeIntellect | $1.40 | $4.40 | 1M | 100.0% | 100.0% |
| Wafer | $1.40 | $4.40 | 1M | 100.0% | 99.9% |
| Venice | $1.40 | $4.40 | 1M | 99.6% | 98.5% |
| Together | $1.40 | $4.40 | 1M | 99.9% | 99.4% |
| Parasail fp8 | $1.40 | $4.40 | 1M | 94.5% | 99.0% |
| Modal | $1.40 | $4.40 | 1M | 99.9% | 99.3% |
| BaseTen fp4 | $1.40 | $4.40 | 1M | 100.0% | 99.5% |
| Fireworks | $1.40 | $4.40 | 1M | 99.8% | 99.7% |
| Cloudflare | $1.40 | $4.40 | 1.3M | 100.0% | 97.2% |
Provider availability updated 53m ago.
Cost calculator
Cheaper alternatives to GLM 5.3
| Model | Input | Output | Context | Intel. | |
|---|---|---|---|---|---|
O GPT-6 Luna OpenAI | $0.10 | $0.50 | 1.1M | 37.3 | compare |
D DeepSeek V4.1 Flash DeepSeek | $0.14 | $0.42 | 1M | 39.5 | compare |
Z GLM 5.3 Flash Z.ai (Zhipu) | $0.15 | $0.50 | 1.3M | 41.8 | compare |
O GPT-5.6 Luna OpenAI | $0.20 | $1.20 | 1.1M | 37.3 | compare |
X MiMo-V2.6-Pro Xiaomi | $0.43 | $0.87 | 1M | 46.3 | compare |
D DeepSeek V4 Pro 0813 DeepSeek | $0.46 | $1.39 | 1M | 36 | compare |
About GLM 5.3
GLM-5.3 is a large-scale reasoning model from Z.ai, built for complex software engineering and long-horizon agent tasks. It supports text input and output with a 1M-token context window, and improves...
FAQ
How much does GLM 5.3 cost?
GLM 5.3 by Z.ai (Zhipu) costs $0.84 per 1M input tokens and $2.64 per 1M output tokens, with a 1.3M token context window. The cheapest provider currently charges $0.56 input / $2.28 output. A typical request with 2,000 input and 500 output tokens costs about $0.003.
What is the context window of GLM 5.3?
GLM 5.3 supports up to 1,310,720 tokens of context and up to 131,072 output tokens per request.
Has GLM 5.3’s price changed?
No price changes recorded since ModelBank started tracking it on 2026-09-24.
Is GLM 5.3 being deprecated?
No retirement date has been announced for GLM 5.3 as of 2026-09-24.
What are cheaper alternatives to GLM 5.3?
Cheaper options include GPT-6 Luna ($0.10/$0.50), DeepSeek V4.1 Flash ($0.14/$0.42), GLM 5.3 Flash ($0.15/$0.50), GPT-5.6 Luna ($0.20/$1.20).