M
Meta Llama

Llama 3.1 70B Instruct pricing & price history

toolsmeta-llama/llama-3.1-70b-instruct

Llama 3.1 70B Instruct by Meta Llama costs $0.40 per 1M input tokens and $0.40 per 1M output tokens, with a 131K token context window.

$0.40Input / 1M tokens
$0.40Output / 1M tokens
131KContext window
Intelligence index

Price history

$0.000$0.230$0.460 2026-09-24today input /1M output /1M

No price changes since tracking began on 2026-09-24. We check hourly.

Cost calculator

Cheaper alternatives to Llama 3.1 70B Instruct

ModelInputOutputContextIntel.
$0.08$0.20262K12.9compare
$0.04$0.321.3M34.3compare
$0.15$0.15262K5.5compare
M
Llama 4 Scout
Meta Llama
$0.10$0.301.3Mcompare
$0.14$0.281Mcompare
$0.18$0.18164Kcompare
$0.10$0.501.1Mcompare
O
$0.10$0.501.1M37.3compare

About Llama 3.1 70B Instruct

Meta's latest class of model (Llama 3.1) launched with a variety of sizes & flavors. This 70B instruct-tuned version is optimized for high quality dialogue usecases. It has demonstrated strong...

FAQ

How much does Llama 3.1 70B Instruct cost?

Llama 3.1 70B Instruct by Meta Llama costs $0.40 per 1M input tokens and $0.40 per 1M output tokens, with a 131K token context window. A typical request with 2,000 input and 500 output tokens costs about $0.001.

What is the context window of Llama 3.1 70B Instruct?

Llama 3.1 70B Instruct supports up to 131,072 tokens of context and up to 16,384 output tokens per request.

Has Llama 3.1 70B Instruct’s price changed?

No price changes recorded since ModelBank started tracking it on 2026-09-24.

Is Llama 3.1 70B Instruct being deprecated?

No retirement date has been announced for Llama 3.1 70B Instruct as of 2026-09-24.

What are cheaper alternatives to Llama 3.1 70B Instruct?

Cheaper options include Nemotron 3.5 Lightning ($0.08/$0.20), DeepSeek V4 Flash 0731 ($0.04/$0.32), Ministral 3 8B 2512 ($0.15/$0.15), Llama 4 Scout ($0.10/$0.30).