gpt-oss-120b vs Ling 3.0 Flash VL

Price, context window & benchmark comparison

O

gpt-oss-120b

OpenAI
$0.15 / $0.60
Input / Output · per 1M tokens
VS
I

Ling 3.0 Flash VL

inclusionAI
$0.06 / $0.18
Input / Output · per 1M tokens
Compare different models
O
I
Verdict: Ling 3.0 Flash VL is 2.9× cheaper than gpt-oss-120b on a typical 3:1 input/output mix. Ling 3.0 Flash VL scores higher on the Artificial Analysis Intelligence Index (24.6 vs 11.6). Ling 3.0 Flash VL has the larger context window (262K vs 131K).
gpt-oss-120bLing 3.0 Flash VL
Input price / 1M$0.15$0.06
Output price / 1M$0.60$0.18
Blended price (3:1)$0.26$0.09
Cached input / 1M$0.075$0.012
Context window131K262K
Max output66K33K
Intelligence index11.624.6
Coding index30.457
Providers
Best provider uptime (30m)
Vision inputNoYes
ReasoningYesYes
Tool callingYesYes
Released2025-08-052026-09-10
Retirement

Cost calculator

FAQ

Which is cheaper, gpt-oss-120b or Ling 3.0 Flash VL?

Ling 3.0 Flash VL is 2.9× cheaper than gpt-oss-120b on a typical 3:1 input/output mix.

gpt-oss-120b vs Ling 3.0 Flash VL: which has a bigger context window?

gpt-oss-120b: 131,072 tokens. Ling 3.0 Flash VL: 262,144 tokens.

How much would 1 million requests cost?

At 2,000 input + 500 output tokens per request: gpt-oss-120b ≈ $600.00, Ling 3.0 Flash VL ≈ $210.00.

More comparisons