Ling 3.0 Flash VL vs Nemotron 3 Ultra

Price, context window & benchmark comparison

I

Ling 3.0 Flash VL

inclusionAI
$0.06 / $0.18
Input / Output · per 1M tokens
VS
N

Nemotron 3 Ultra

NVIDIA
$0.60 / $2.40
Input / Output · per 1M tokens
Compare different models
I
N
Verdict: Ling 3.0 Flash VL is 11.7× cheaper than Nemotron 3 Ultra on a typical 3:1 input/output mix. Ling 3.0 Flash VL scores higher on the Artificial Analysis Intelligence Index (24.6 vs 22.9).
Ling 3.0 Flash VLNemotron 3 Ultra
Input price / 1M$0.06$0.60
Output price / 1M$0.18$2.40
Blended price (3:1)$0.09$1.05
Cached input / 1M$0.012$0.12
Context window262K262K
Max output33K183K
Intelligence index24.622.9
Coding index5749.3
Providers
Best provider uptime (30m)
Vision inputYesNo
ReasoningYesYes
Tool callingYesYes
Released2026-09-102026-06-04
Retirement

Cost calculator

FAQ

Which is cheaper, Ling 3.0 Flash VL or Nemotron 3 Ultra?

Ling 3.0 Flash VL is 11.7× cheaper than Nemotron 3 Ultra on a typical 3:1 input/output mix.

Ling 3.0 Flash VL vs Nemotron 3 Ultra: which has a bigger context window?

Ling 3.0 Flash VL: 262,144 tokens. Nemotron 3 Ultra: 262,144 tokens.

How much would 1 million requests cost?

At 2,000 input + 500 output tokens per request: Ling 3.0 Flash VL ≈ $210.00, Nemotron 3 Ultra ≈ $2,400.

More comparisons