DeepSeek V4.1 Flash vs Qwen3 30B A3B Thinking 2507

Price, context window & benchmark comparison

D

DeepSeek V4.1 Flash

DeepSeek
$0.14 / $0.42
Input / Output · per 1M tokens
VS
Q

Qwen3 30B A3B Thinking 2507

Qwen (Alibaba)
$0.20 / $2.40
Input / Output · per 1M tokens
Compare different models
D
Q
Verdict: DeepSeek V4.1 Flash is 3.6× cheaper than Qwen3 30B A3B Thinking 2507 on a typical 3:1 input/output mix. DeepSeek V4.1 Flash scores higher on the Artificial Analysis Intelligence Index (39.5 vs 9.8). DeepSeek V4.1 Flash has the larger context window (1M vs 82K).
DeepSeek V4.1 FlashQwen3 30B A3B Thinking 2507
Input price / 1M$0.14$0.20
Output price / 1M$0.42$2.40
Blended price (3:1)$0.21$0.75
Cached input / 1M$0.0042
Context window1M82K
Max output131K33K
Intelligence index39.59.8
Coding index12.1
Providers26
Best provider uptime (30m)100.0%
Vision inputYesNo
ReasoningYesYes
Tool callingYesYes
Released2026-09-102025-08-28
Retirement

Cost calculator

FAQ

Which is cheaper, DeepSeek V4.1 Flash or Qwen3 30B A3B Thinking 2507?

DeepSeek V4.1 Flash is 3.6× cheaper than Qwen3 30B A3B Thinking 2507 on a typical 3:1 input/output mix.

DeepSeek V4.1 Flash vs Qwen3 30B A3B Thinking 2507: which has a bigger context window?

DeepSeek V4.1 Flash: 1,048,576 tokens. Qwen3 30B A3B Thinking 2507: 81,920 tokens.

How much would 1 million requests cost?

At 2,000 input + 500 output tokens per request: DeepSeek V4.1 Flash ≈ $490.00, Qwen3 30B A3B Thinking 2507 ≈ $1,600.

More comparisons