Llama 3.2 3B Instruct pricing & price history
Llama 3.2 3B Instruct by Meta Llama costs $0.05 per 1M input tokens and $0.33 per 1M output tokens, with a 131K token context window.
Price history
No price changes since tracking began on 2026-09-24. We check hourly.
Cost calculator
Cheaper alternatives to Llama 3.2 3B Instruct
| Model | Input | Output | Context | Intel. | |
|---|---|---|---|---|---|
N Nemotron 3.5 Lightning NVIDIA | $0.08 | $0.20 | 262K | 12.9 | compare |
D DeepSeek V4 Flash 0731 DeepSeek | $0.04 | $0.32 | 1.3M | 34.3 | compare |
About Llama 3.2 3B Instruct
Llama 3.2 3B is a 3-billion-parameter multilingual large language model, optimized for advanced natural language processing tasks like dialogue generation, reasoning, and summarization. Designed with the latest transformer architecture, it...
FAQ
How much does Llama 3.2 3B Instruct cost?
Llama 3.2 3B Instruct by Meta Llama costs $0.05 per 1M input tokens and $0.33 per 1M output tokens, with a 131K token context window. A typical request with 2,000 input and 500 output tokens costs about $0.0003.
What is the context window of Llama 3.2 3B Instruct?
Llama 3.2 3B Instruct supports up to 131,072 tokens of context and up to 117,964 output tokens per request.
Has Llama 3.2 3B Instruct’s price changed?
No price changes recorded since ModelBank started tracking it on 2026-09-24.
Is Llama 3.2 3B Instruct being deprecated?
No retirement date has been announced for Llama 3.2 3B Instruct as of 2026-09-24.
What are cheaper alternatives to Llama 3.2 3B Instruct?
Cheaper options include Nemotron 3.5 Lightning ($0.08/$0.20), DeepSeek V4 Flash 0731 ($0.04/$0.32).