D
DeepSeek

DeepSeek V4.1 Flash pricing & price history

visionreasoningtoolsdeepseek/deepseek-v4.1-flash

DeepSeek V4.1 Flash by DeepSeek costs $0.14 per 1M input tokens and $0.42 per 1M output tokens, with a 1M token context window. The cheapest provider currently charges $0.05 input / $0.25 output.

$0.14Input / 1M tokens
$0.42Output / 1M tokens
1MContext window
39.5Intelligence index

Price history

$0.000$0.241$0.483 2026-09-24today input /1M output /1M

No price changes since tracking began on 2026-09-24. We check hourly.

Providers serving DeepSeek V4.1 Flash (26)

ProviderInputOutputContextUptime 30mUptime 24h
Relace$0.05$0.251M99.9%100.0%
Morph$0.072$0.291M100.0%99.8%
OpenInference fp4$0.10$0.501M100.0%99.2%
DekaLLM$0.10$1.001M99.9%95.4%
Sail Research fp4$0.13$0.751M99.9%99.8%
DeepInfra fp8$0.14$0.421M99.8%99.7%
DeepSeek$0.15$0.601M100.0%100.0%
StreamLake fp8$0.20$0.791M99.7%99.8%
Wafer$0.20$0.601M100.0%100.0%
CoreWeave fp8$0.20$0.651M99.8%99.5%
Fireworks$0.22$0.661M99.8%99.9%
Krea fp8$0.23$0.901M97.2%96.3%
GMICloud fp8$0.23$0.901M98.8%99.3%
Phala$0.28$1.101M100.0%99.1%
Novita fp8$0.28$1.141M100.0%99.9%
AtlasCloud fp8$0.30$1.201M99.7%99.8%
BaseTen fp8$0.30$1.201M100.0%99.2%
Makora fp8$0.30$1.201M98.0%98.6%
DigitalOcean$0.30$1.201M99.8%98.6%
Alibaba$0.30$1.201M100.0%99.9%
Together$0.30$1.201M99.9%99.9%
SiliconFlow fp8$0.30$1.201M99.9%99.2%
Modal$0.30$1.201M99.7%99.8%
BaseTen fp8$0.30$1.201M99.5%99.0%
Parasail fp8$0.30$1.201M93.9%95.6%
Venice fp8$0.38$1.501M99.6%98.6%

Provider availability updated 54m ago.

Cost calculator

Cheaper alternatives to DeepSeek V4.1 Flash

ModelInputOutputContextIntel.
O
$0.10$0.501.1M37.3compare

About DeepSeek V4.1 Flash

DeepSeek V4.1 Flash is a sparse mixture-of-experts model from DeepSeek, and the first built on the company's Causal Encoder-Decoder (CED) architecture. It activates 8B parameters on input and 16B on...

FAQ

How much does DeepSeek V4.1 Flash cost?

DeepSeek V4.1 Flash by DeepSeek costs $0.14 per 1M input tokens and $0.42 per 1M output tokens, with a 1M token context window. The cheapest provider currently charges $0.05 input / $0.25 output. A typical request with 2,000 input and 500 output tokens costs about $0.0005.

What is the context window of DeepSeek V4.1 Flash?

DeepSeek V4.1 Flash supports up to 1,048,576 tokens of context and up to 131,072 output tokens per request.

Has DeepSeek V4.1 Flash’s price changed?

No price changes recorded since ModelBank started tracking it on 2026-09-24.

Is DeepSeek V4.1 Flash being deprecated?

No retirement date has been announced for DeepSeek V4.1 Flash as of 2026-09-24.

What are cheaper alternatives to DeepSeek V4.1 Flash?

Cheaper options include GPT-6 Luna ($0.10/$0.50).