Q
Qwen (Alibaba)

Qwen2.5 72B Instruct pricing & price history

toolsqwen/qwen-2.5-72b-instruct

Qwen2.5 72B Instruct by Qwen (Alibaba) costs $0.36 per 1M input tokens and $0.40 per 1M output tokens, with a 33K token context window.

$0.36Input / 1M tokens
$0.40Output / 1M tokens
33KContext window
Intelligence index

Price history

$0.000$0.230$0.460 2026-09-24today input /1M output /1M

No price changes since tracking began on 2026-09-24. We check hourly.

Cost calculator

Cheaper alternatives to Qwen2.5 72B Instruct

ModelInputOutputContextIntel.
$0.08$0.20262K12.9compare
$0.04$0.321.3M34.3compare
$0.15$0.15262K5.5compare
M
Llama 4 Scout
Meta Llama
$0.10$0.301.3Mcompare
$0.14$0.281Mcompare
$0.18$0.18164Kcompare
$0.10$0.501.1Mcompare
O
$0.10$0.501.1M37.3compare

About Qwen2.5 72B Instruct

Qwen2.5 72B is the latest series of Qwen large language models. Qwen2.5 brings the following improvements upon Qwen2: - Significantly more knowledge and has greatly improved capabilities in coding and...

FAQ

How much does Qwen2.5 72B Instruct cost?

Qwen2.5 72B Instruct by Qwen (Alibaba) costs $0.36 per 1M input tokens and $0.40 per 1M output tokens, with a 33K token context window. A typical request with 2,000 input and 500 output tokens costs about $0.0009.

What is the context window of Qwen2.5 72B Instruct?

Qwen2.5 72B Instruct supports up to 32,768 tokens of context and up to 16,384 output tokens per request.

Has Qwen2.5 72B Instruct’s price changed?

No price changes recorded since ModelBank started tracking it on 2026-09-24.

Is Qwen2.5 72B Instruct being deprecated?

No retirement date has been announced for Qwen2.5 72B Instruct as of 2026-09-24.

What are cheaper alternatives to Qwen2.5 72B Instruct?

Cheaper options include Nemotron 3.5 Lightning ($0.08/$0.20), DeepSeek V4 Flash 0731 ($0.04/$0.32), Ministral 3 8B 2512 ($0.15/$0.15), Llama 4 Scout ($0.10/$0.30).