DeepSeek V4.1 Flash (Reasoning, Max Effort) API PricingNEW

DeepSeek · released 2026-09-10

Pricing

$0.30
input / 1M tokens
0.030¢ / 1K
$1.20
output / 1M tokens
0.120¢ / 1K
$0.525
blended / 1M tokens
0.053¢ / 1K

What it actually costs

WorkloadTokensCost
Short chat reply500 in / 300 out0.051¢
RAG query4,000 in / 500 out0.18¢
Long document summary50,000 in / 1,000 out$0.0162
Code review of a 2k-line file30,000 in / 2,000 out$0.0114
1,000 chat replies/day for 30 days30,000 × (500 in / 300 out)$15.30

Cost = (input tokens ÷ 1,000,000 × $0.30) + (output tokens ÷ 1,000,000 × $1.20).

Benchmarks

39.5
Intelligence
#24 of 533
Coding
Math
231
Tokens/sec
0.81s
Time to first token

Similar intelligence, lower price

ModelIntelligenceBlended $/1M
Qwen3.8-Flash-Next39.8$0.23
GLM 5.3 Flash41.8$0.237
GPT-5.6 Luna (max)37.3$0.45

Leaderboard neighbours

ModelIntelligenceBlended $/1M
Qwen3.8-Flash-Next39.8$0.23
Gemini 3.7 Flash (medium)39.6$1.50
Muse Spark 1.2 (xhigh)39.6$2.00
DeepSeek V4.1 Flash (Reasoning, Max Effort)39.5$0.525
GPT-5.4 (xhigh)39.0$5.625
Grok 4.5 (high)38.8$3.00
GPT-5.5 (xhigh)38.4$11.25

More from DeepSeek

ModelIntelligenceBlended $/1M
DeepSeek V4 Pro 0813 (Reasoning, Max Effort)36.0$1.98
DeepSeek V4 Flash Vision (Reasoning, Max Effort)34.8$0.66
DeepSeek V4 Flash 0731 (Reasoning, Max Effort)34.3$0.66
DeepSeek V4 Pro 0424 (Reasoning, Max Effort)30.4$0.544
DeepSeek V4 Flash 0420 (Reasoning, High Effort)26.0$0.168
DeepSeek V3.2 (Reasoning)21.5$0.315
DeepSeek V3.2 Exp (Reasoning)16.6$0.315
DeepSeek V3.2 (Non-reasoning)16.0$0.315

Count tokens & estimate cost for DeepSeek V4.1 Flash (Reasoning, Max Effort) See DeepSeek V4.1 Flash (Reasoning, Max Effort) on the leaderboard

Source: Artificial Analysis · last synced 2026-09-22

FAQ

Who develops DeepSeek V4.1 Flash (Reasoning, Max Effort)?

DeepSeek V4.1 Flash (Reasoning, Max Effort) is developed by DeepSeek, released 2026-09-10.

How much does DeepSeek V4.1 Flash (Reasoning, Max Effort) cost per million tokens?

DeepSeek V4.1 Flash (Reasoning, Max Effort) costs $0.30 per 1M input tokens and $1.20 per 1M output tokens ($0.525 blended), per Artificial Analysis's official pricing data.

How does DeepSeek V4.1 Flash (Reasoning, Max Effort) rank on benchmarks?

DeepSeek V4.1 Flash (Reasoning, Max Effort) scores 39.5 on the Artificial Analysis Intelligence Index, ranking #24 of 533 models we track.

Is there a cheaper model with similar intelligence to DeepSeek V4.1 Flash (Reasoning, Max Effort)?

Yes — Qwen3.8-Flash-Next scores similar intelligence (39.8) at $0.23/1M blended, versus $0.525/1M for DeepSeek V4.1 Flash (Reasoning, Max Effort).

Where does this data come from?

Pricing and benchmark data for DeepSeek V4.1 Flash (Reasoning, Max Effort) comes from Artificial Analysis (artificialanalysis.ai), synced daily. Last synced 2026-09-22.