GLM 4.5 pricing & price history
GLM 4.5 by Z.ai (Zhipu) costs $0.60 per 1M input tokens and $2.20 per 1M output tokens, with a 131K token context window.
Price history
No price changes since tracking began on 2026-09-24. We check hourly.
Cost calculator
Cheaper alternatives to GLM 4.5
| Model | Input | Output | Context | Intel. | |
|---|---|---|---|---|---|
N Nemotron 3.5 Lightning NVIDIA | $0.08 | $0.20 | 262K | 12.9 | compare |
D DeepSeek V4 Flash 0731 DeepSeek | $0.04 | $0.32 | 1.3M | 34.3 | compare |
M Ministral 3 8B 2512 Mistral AI | $0.15 | $0.15 | 262K | 5.5 | compare |
M Llama 4 Scout Meta Llama | $0.10 | $0.30 | 1.3M | — | compare |
X MiMo-V2.6-Flash Xiaomi | $0.14 | $0.28 | 1M | — | compare |
M Llama Guard 4 12B Meta Llama | $0.18 | $0.18 | 164K | — | compare |
O GPT-6 Luna Pro OpenAI | $0.10 | $0.50 | 1.1M | — | compare |
O GPT-6 Luna OpenAI | $0.10 | $0.50 | 1.1M | 37.3 | compare |
About GLM 4.5
GLM-4.5 is our latest flagship foundation model, purpose-built for agent-based applications. It leverages a Mixture-of-Experts (MoE) architecture and supports a context length of up to 128k tokens. GLM-4.5 delivers significantly...
FAQ
How much does GLM 4.5 cost?
GLM 4.5 by Z.ai (Zhipu) costs $0.60 per 1M input tokens and $2.20 per 1M output tokens, with a 131K token context window. A typical request with 2,000 input and 500 output tokens costs about $0.0023.
What is the context window of GLM 4.5?
GLM 4.5 supports up to 131,072 tokens of context and up to 98,304 output tokens per request.
Has GLM 4.5’s price changed?
No price changes recorded since ModelBank started tracking it on 2026-09-24.
Is GLM 4.5 being deprecated?
Yes — GLM 4.5 is scheduled to be retired on 2026-12-31. Consider migrating to one of the alternatives on this page.
What are cheaper alternatives to GLM 4.5?
Cheaper options include Nemotron 3.5 Lightning ($0.08/$0.20), DeepSeek V4 Flash 0731 ($0.04/$0.32), Ministral 3 8B 2512 ($0.15/$0.15), Llama 4 Scout ($0.10/$0.30).