Q
Qwen (Alibaba)

Qwen3 8B pricing & price history

reasoningtoolsqwen/qwen3-8b

Qwen3 8B by Qwen (Alibaba) costs $0.12 per 1M input tokens and $0.46 per 1M output tokens, with a 131K token context window.

$0.12Input / 1M tokens
$0.46Output / 1M tokens
131KContext window
Intelligence index · coding 9

Price history

$0.000$0.262$0.523 2026-09-24today input /1M output /1M

No price changes since tracking began on 2026-09-24. We check hourly.

Cost calculator

Cheaper alternatives to Qwen3 8B

ModelInputOutputContextIntel.
$0.08$0.20262K12.9compare
$0.04$0.321.3M34.3compare
$0.15$0.15262K5.5compare
M
Llama 4 Scout
Meta Llama
$0.10$0.301.3Mcompare
$0.14$0.281Mcompare
$0.18$0.18164Kcompare
$0.10$0.501.1Mcompare
O
$0.10$0.501.1M37.3compare

About Qwen3 8B

Qwen3-8B is a dense 8.2B parameter causal language model from the Qwen3 series, designed for both reasoning-heavy tasks and efficient dialogue. It supports seamless switching between "thinking" mode for math,...

FAQ

How much does Qwen3 8B cost?

Qwen3 8B by Qwen (Alibaba) costs $0.12 per 1M input tokens and $0.46 per 1M output tokens, with a 131K token context window. A typical request with 2,000 input and 500 output tokens costs about $0.0005.

What is the context window of Qwen3 8B?

Qwen3 8B supports up to 131,072 tokens of context and up to 8,192 output tokens per request.

Has Qwen3 8B’s price changed?

No price changes recorded since ModelBank started tracking it on 2026-09-24.

Is Qwen3 8B being deprecated?

No retirement date has been announced for Qwen3 8B as of 2026-09-24.

What are cheaper alternatives to Qwen3 8B?

Cheaper options include Nemotron 3.5 Lightning ($0.08/$0.20), DeepSeek V4 Flash 0731 ($0.04/$0.32), Ministral 3 8B 2512 ($0.15/$0.15), Llama 4 Scout ($0.10/$0.30).