ZMIME
DeepSeek · Released Sep 10, 2026

DeepSeek V4.1 Flash

DeepSeek V4.1 Flash model for reasoning and agentic coding.

  • Open weights
  • Reasoning
  • Tool calling
  • Vision
Capability
155.0ECI · #29 of 148
Input
$0.15per 1M tokens
Output
$0.60per 1M tokens
Context
1M384K max output
Pricing

What it costs

Input
$0.15per million tokens
Output
$0.60per million tokens
Cached input
$0.003per million tokens
Blended (3:1)
$0.263cheaper than 73% of priced models

Typical monthly bills

Monthly usageEstimated cost
Side project2M input + 0.5M output tokens$0.60
Team assistant25M input + 5M output tokens$6.75
Production app250M input + 50M output tokens$67.50

Official DeepSeek API price, as listed on models.dev.

Capability

Independent benchmarks

155.0Epoch Capabilities Index
#29 of 148 scored models
88Median 146167

The shaded band is Epoch AI’s confidence range (148.8–157.6); the tick marks the median scored model.

Scores from Epoch AI, run independently of DeepSeek.

Specifications

The details

Lab
DeepSeek
Released
Sep 10, 2026
Knowledge cutoff
May 2025
Context window
1,000,000 tokens
Max output
384,000 tokens
Inputs
Text, Images
Output
Text
Reasoning
Adjustable effort low · high · max
Tool calling
Yes
Structured output
Yes
API model ID
deepseek-flashon DeepSeek
Availability
50 API providerslisted on models.dev

DeepSeek model documentation

Lab-reported

What DeepSeek claims

Published by the lab at launch. Settings vary, so compare these only with care.

BenchmarkScoreSettingSource
GPQA Diamond90.9reasoning_effort=100Source
MathArena Apex65.6reasoning_effort=100Source
Humanity's Last Exam36.8reasoning_effort=100Source
Humanity's Last Exam39.1reasoning_effort=100Source
Codeforces3471 ratingreasoning_effort=100Source
Terminal-Bench v2.190.6reasoning_effort=100Source
Terminal-Bench v3.030reasoning_effort=100Source
Terminal-Bench v4.031.2reasoning_effort=100Source
DeepSWE v1.174.2 resolvedreasoning_effort=100Source
ProgramBench20.3 almost@1reasoning_effort=100Source
NL2Repo64reasoning_effort=100Source
CyberGym88.1reasoning_effort=100Source
Same lab

More from DeepSeek

Questions

About DeepSeek V4.1 Flash

How much does DeepSeek V4.1 Flash cost?

DeepSeek V4.1 Flash costs $0.15 input / $0.60 output per million tokens (official DeepSeek API price). Cached input is $0.003 per million tokens. At a 3:1 input-to-output mix that is $0.263 per million tokens, cheaper than 73% of the 360 priced models we track.

What is the context window of DeepSeek V4.1 Flash?

DeepSeek V4.1 Flash accepts up to 1,000,000 tokens per request and can write up to 384,000 tokens in one response.

How good is DeepSeek V4.1 Flash?

Epoch AI gives DeepSeek V4.1 Flash a Capabilities Index score of 155.0 (likely range 148.8–157.6), ranking it #29 of 148 models Epoch has scored.

Is DeepSeek V4.1 Flash open source?

Yes. DeepSeek publishes the weights under the MIT licence, so it can be downloaded and self-hosted.

What inputs does DeepSeek V4.1 Flash support?

DeepSeek V4.1 Flash accepts text and images and replies in text. It is a reasoning model with low, high and max effort settings, supports tool calling and can return structured JSON output.

When was DeepSeek V4.1 Flash released?

DeepSeek released DeepSeek V4.1 Flash on Sep 10, 2026. Its training data runs to May 2025.

What are the best alternatives to DeepSeek V4.1 Flash?

The closest current models from other labs on capability, price and release date are GLM-5.3-Flash (Z.ai (Zhipu), ECI 151.9, $0.15 / $0.50), GPT-5.6 Luna (OpenAI, ECI 156.5, $0.20 / $1.20), Gemini 3.6 Flash (Google, ECI 154.3, $0.75 / $3.75) and Inkling Small (Thinking Machines, ECI 150.2, $0.50 / $1.20).