DeepSeek V4 Flash 0731 vs Qwen3 235B A22B Thinking 2507

Price, context window & benchmark comparison

DeepSeek

DeepSeek V4 Flash 0731

DeepSeek
$0.03 / $0.32
Input / Output · per 1M tokens
VS
Qwen (Alibaba)

Qwen3 235B A22B Thinking 2507

Qwen (Alibaba)
$0.23 / $2.30
Input / Output · per 1M tokens
Compare different models
DeepSeek
Qwen (Alibaba)
Verdict: DeepSeek V4 Flash 0731 is 7.3× cheaper than Qwen3 235B A22B Thinking 2507 on a typical 3:1 input/output mix. DeepSeek V4 Flash 0731 scores higher on the Artificial Analysis Intelligence Index (34.3 vs 12.7). DeepSeek V4 Flash 0731 has the larger context window (1.3M vs 131K).
DeepSeek V4 Flash 0731Qwen3 235B A22B Thinking 2507
Input price / 1M$0.03$0.23
Output price / 1M$0.32$2.30
Blended price (3:1)$0.10$0.75
Cached input / 1M$0.016
Context window1.3M131K
Max output944K118K
Intelligence index34.312.7
Coding index69.122.1
Providers30
Best provider uptime (30m)100.0%
Vision inputNoNo
ReasoningYesYes
Tool callingYesYes
Released2026-07-312025-07-25
Retirement

Cost calculator

FAQ

Which is cheaper, DeepSeek V4 Flash 0731 or Qwen3 235B A22B Thinking 2507?

DeepSeek V4 Flash 0731 is 7.3× cheaper than Qwen3 235B A22B Thinking 2507 on a typical 3:1 input/output mix.

DeepSeek V4 Flash 0731 vs Qwen3 235B A22B Thinking 2507: which has a bigger context window?

DeepSeek V4 Flash 0731: 1,310,720 tokens. Qwen3 235B A22B Thinking 2507: 131,072 tokens.

How much would 1 million requests cost?

At 2,000 input + 500 output tokens per request: DeepSeek V4 Flash 0731 ≈ $220.00, Qwen3 235B A22B Thinking 2507 ≈ $1,610.

More comparisons