Gemma 3 4B vs Llama 4 Scout

Price, context window & benchmark comparison

G

Gemma 3 4B

Google
$0.05 / $0.10
Input / Output · per 1M tokens
VS
M

Llama 4 Scout

Meta Llama
$0.10 / $0.30
Input / Output · per 1M tokens
Compare different models
G
M
Verdict: Gemma 3 4B is 2.4× cheaper than Llama 4 Scout on a typical 3:1 input/output mix. Llama 4 Scout has the larger context window (1.3M vs 131K).
Gemma 3 4BLlama 4 Scout
Input price / 1M$0.05$0.10
Output price / 1M$0.10$0.30
Blended price (3:1)$0.0625$0.15
Cached input / 1M
Context window131K1.3M
Max output16K16K
Intelligence index
Coding index2.78.2
Providers3
Best provider uptime (30m)99.9%
Vision inputYesYes
ReasoningNoNo
Tool callingNoYes
Released2025-03-132025-04-05
Retirement

Cost calculator

FAQ

Which is cheaper, Gemma 3 4B or Llama 4 Scout?

Gemma 3 4B is 2.4× cheaper than Llama 4 Scout on a typical 3:1 input/output mix.

Gemma 3 4B vs Llama 4 Scout: which has a bigger context window?

Gemma 3 4B: 131,072 tokens. Llama 4 Scout: 1,310,720 tokens.

How much would 1 million requests cost?

At 2,000 input + 500 output tokens per request: Gemma 3 4B ≈ $150.00, Llama 4 Scout ≈ $350.00.

More comparisons