MiniMax M2.5 vs Mistral Small 4
Price, context window & benchmark comparison
VS
Compare different models
Verdict: Mistral Small 4 is 44% cheaper than MiniMax M2.5 on a typical 3:1 input/output mix. Mistral Small 4 has the larger context window (262K vs 205K).
| MiniMax M2.5 | Mistral Small 4 | |
|---|---|---|
| Input price / 1M | $0.27 | $0.15 |
| Output price / 1M | $1.08 | $0.60 |
| Blended price (3:1) | $0.47 | $0.26 |
| Cached input / 1M | $0.027 | $0.015 |
| Context window | 205K | 262K |
| Max output | 128K | 210K |
| Intelligence index | — | 11.3 |
| Coding index | — | 26.6 |
| Providers | 8 | 3 |
| Best provider uptime (30m) | 100.0% | 100.0% |
| Vision input | No | Yes |
| Reasoning | Yes | Yes |
| Tool calling | Yes | Yes |
| Released | 2026-02-12 | 2026-03-16 |
| Retirement | — | — |
Cost calculator
FAQ
Which is cheaper, MiniMax M2.5 or Mistral Small 4?
Mistral Small 4 is 44% cheaper than MiniMax M2.5 on a typical 3:1 input/output mix.
MiniMax M2.5 vs Mistral Small 4: which has a bigger context window?
MiniMax M2.5: 204,800 tokens. Mistral Small 4: 262,144 tokens.
How much would 1 million requests cost?
At 2,000 input + 500 output tokens per request: MiniMax M2.5 ≈ $1,080, Mistral Small 4 ≈ $600.00.
More comparisons
MiniMax M2.5 vs GLM 5.3 PrimeMiniMax M2.5 vs Qwen3.8 Max PrimeMiniMax M2.5 vs Command A+MiniMax M2.5 vs GPT-6 Luna ProMiniMax M2.5 vs GPT-6 LunaMiniMax M2.5 vs GPT-6 Sol ProMiniMax M2.5 vs GPT-6 SolMiniMax M2.5 vs Claude Opus 5.5MiniMax M2.5 vs MiMo-V2.6-Pro-UltraSpeedMiniMax M2.5 vs MiMo-V2.6-FlashMiniMax M2.5 vs Grok 4.7MiniMax M2.5 vs Qwen3.8 Omni FlashMiniMax M2.5 vs GLM 5.3 FlashXMiniMax M2.5 vs DeepSeek V4.1 FlashMiniMax M2.5 vs GPT-6 AstraMiniMax M2.5 vs GPT-6 Astra ProMistral Small 4 vs GLM 5.3 PrimeMistral Small 4 vs Qwen3.8 Max PrimeMistral Small 4 vs Command A+Mistral Small 4 vs GPT-6 Luna ProMistral Small 4 vs GPT-6 LunaMistral Small 4 vs GPT-6 Sol ProMistral Small 4 vs GPT-6 SolMistral Small 4 vs Claude Opus 5.5