DeepSeek V4 Flash 0731 vs Nemotron 3.5 Lightning

Price, context window & benchmark comparison

D

DeepSeek V4 Flash 0731

DeepSeek
$0.03 / $0.32
Input / Output · per 1M tokens
VS
N

Nemotron 3.5 Lightning

NVIDIA
$0.08 / $0.20
Input / Output · per 1M tokens
Compare different models
D
N
Verdict: DeepSeek V4 Flash 0731 is 7% cheaper than Nemotron 3.5 Lightning on a typical 3:1 input/output mix. DeepSeek V4 Flash 0731 scores higher on the Artificial Analysis Intelligence Index (34.3 vs 12.9). DeepSeek V4 Flash 0731 has the larger context window (1.3M vs 262K).
DeepSeek V4 Flash 0731Nemotron 3.5 Lightning
Input price / 1M$0.03$0.08
Output price / 1M$0.32$0.20
Blended price (3:1)$0.10$0.11
Cached input / 1M$0.016$0.04
Context window1.3M262K
Max output944K131K
Intelligence index34.312.9
Coding index69.126.8
Providers304
Best provider uptime (30m)100.0%100.0%
Vision inputNoNo
ReasoningYesYes
Tool callingYesYes
Released2026-07-312026-08-11
Retirement

Cost calculator

FAQ

Which is cheaper, DeepSeek V4 Flash 0731 or Nemotron 3.5 Lightning?

DeepSeek V4 Flash 0731 is 7% cheaper than Nemotron 3.5 Lightning on a typical 3:1 input/output mix.

DeepSeek V4 Flash 0731 vs Nemotron 3.5 Lightning: which has a bigger context window?

DeepSeek V4 Flash 0731: 1,310,720 tokens. Nemotron 3.5 Lightning: 262,144 tokens.

How much would 1 million requests cost?

At 2,000 input + 500 output tokens per request: DeepSeek V4 Flash 0731 ≈ $220.00, Nemotron 3.5 Lightning ≈ $260.00.

More comparisons