GLM 4.7 Flash pricing & price history
GLM 4.7 Flash by Z.ai (Zhipu) costs $0.0605 per 1M input tokens and $0.40 per 1M output tokens, with a 200K token context window.
Price history
No price changes since tracking began on 2026-09-24. We check hourly.
Cost calculator
Cheaper alternatives to GLM 4.7 Flash
| Model | Input | Output | Context | Intel. | |
|---|---|---|---|---|---|
N Nemotron 3.5 Lightning NVIDIA | $0.08 | $0.20 | 262K | 12.9 | compare |
D DeepSeek V4 Flash 0731 DeepSeek | $0.04 | $0.32 | 1.3M | 34.3 | compare |
About GLM 4.7 Flash
As a 30B-class SOTA model, GLM-4.7-Flash offers a new option that balances performance and efficiency. It is further optimized for agentic coding use cases, strengthening coding capabilities, long-horizon task planning,...
FAQ
How much does GLM 4.7 Flash cost?
GLM 4.7 Flash by Z.ai (Zhipu) costs $0.0605 per 1M input tokens and $0.40 per 1M output tokens, with a 200K token context window. A typical request with 2,000 input and 500 output tokens costs about $0.0003.
What is the context window of GLM 4.7 Flash?
GLM 4.7 Flash supports up to 200,000 tokens of context and up to 117,964 output tokens per request.
Has GLM 4.7 Flash’s price changed?
No price changes recorded since ModelBank started tracking it on 2026-09-24.
Is GLM 4.7 Flash being deprecated?
No retirement date has been announced for GLM 4.7 Flash as of 2026-09-24.
What are cheaper alternatives to GLM 4.7 Flash?
Cheaper options include Nemotron 3.5 Lightning ($0.08/$0.20), DeepSeek V4 Flash 0731 ($0.04/$0.32).