Qwen3 32B pricing & price history
Qwen3 32B by Qwen (Alibaba) costs $0.08 per 1M input tokens and $0.28 per 1M output tokens, with a 131K token context window.
Price history
No price changes since tracking began on 2026-09-24. We check hourly.
Cost calculator
Cheaper alternatives to Qwen3 32B
| Model | Input | Output | Context | Intel. | |
|---|---|---|---|---|---|
N Nemotron 3.5 Lightning NVIDIA | $0.08 | $0.20 | 262K | 12.9 | compare |
D DeepSeek V4 Flash 0731 DeepSeek | $0.04 | $0.32 | 1.3M | 34.3 | compare |
About Qwen3 32B
Qwen3-32B is a dense 32.8B parameter causal language model from the Qwen3 series, optimized for both complex reasoning and efficient dialogue. It supports seamless switching between a "thinking" mode for...
FAQ
How much does Qwen3 32B cost?
Qwen3 32B by Qwen (Alibaba) costs $0.08 per 1M input tokens and $0.28 per 1M output tokens, with a 131K token context window. A typical request with 2,000 input and 500 output tokens costs about $0.0003.
What is the context window of Qwen3 32B?
Qwen3 32B supports up to 131,072 tokens of context and up to 16,384 output tokens per request.
Has Qwen3 32B’s price changed?
No price changes recorded since ModelBank started tracking it on 2026-09-24.
Is Qwen3 32B being deprecated?
No retirement date has been announced for Qwen3 32B as of 2026-09-24.
What are cheaper alternatives to Qwen3 32B?
Cheaper options include Nemotron 3.5 Lightning ($0.08/$0.20), DeepSeek V4 Flash 0731 ($0.04/$0.32).