DeepSeek V4.1 Flash
DeepSeek V4.1 Flash model for reasoning and agentic coding.
- Capability
- 155.0ECI · #29 of 148
- Input
- $0.15per 1M tokens
- Output
- $0.60per 1M tokens
- Context
- 1M384K max output
What it costs
- Input
- $0.15per million tokens
- Output
- $0.60per million tokens
- Cached input
- $0.003per million tokens
- Blended (3:1)
- $0.263cheaper than 73% of priced models
Typical monthly bills
| Monthly usage | Estimated cost |
|---|---|
| Side project2M input + 0.5M output tokens | $0.60 |
| Team assistant25M input + 5M output tokens | $6.75 |
| Production app250M input + 50M output tokens | $67.50 |
Official DeepSeek API price, as listed on models.dev.
Independent benchmarks
#29 of 148 scored models
The shaded band is Epoch AI’s confidence range (148.8–157.6); the tick marks the median scored model.
Scores from Epoch AI, run independently of DeepSeek.
The details
- Lab
- DeepSeek
- Released
- Sep 10, 2026
- Knowledge cutoff
- May 2025
- Context window
- 1,000,000 tokens
- Max output
- 384,000 tokens
- Inputs
- Text, Images
- Output
- Text
- Reasoning
- Adjustable effort low · high · max
- Tool calling
- Yes
- Structured output
- Yes
- Weights
- Open weights · MIT
- API model ID
deepseek-flashon DeepSeek- Availability
- 50 API providerslisted on models.dev
What DeepSeek claims
Published by the lab at launch. Settings vary, so compare these only with care.
| Benchmark | Score | Setting | Source |
|---|---|---|---|
| GPQA Diamond | 90.9 | reasoning_effort=100 | Source |
| MathArena Apex | 65.6 | reasoning_effort=100 | Source |
| Humanity's Last Exam | 36.8 | reasoning_effort=100 | Source |
| Humanity's Last Exam | 39.1 | reasoning_effort=100 | Source |
| Codeforces | 3471 rating | reasoning_effort=100 | Source |
| Terminal-Bench v2.1 | 90.6 | reasoning_effort=100 | Source |
| Terminal-Bench v3.0 | 30 | reasoning_effort=100 | Source |
| Terminal-Bench v4.0 | 31.2 | reasoning_effort=100 | Source |
| DeepSWE v1.1 | 74.2 resolved | reasoning_effort=100 | Source |
| ProgramBench | 20.3 almost@1 | reasoning_effort=100 | Source |
| NL2Repo | 64 | reasoning_effort=100 | Source |
| CyberGym | 88.1 | reasoning_effort=100 | Source |
More from DeepSeek
About DeepSeek V4.1 Flash
How much does DeepSeek V4.1 Flash cost?
DeepSeek V4.1 Flash costs $0.15 input / $0.60 output per million tokens (official DeepSeek API price). Cached input is $0.003 per million tokens. At a 3:1 input-to-output mix that is $0.263 per million tokens, cheaper than 73% of the 360 priced models we track.
What is the context window of DeepSeek V4.1 Flash?
DeepSeek V4.1 Flash accepts up to 1,000,000 tokens per request and can write up to 384,000 tokens in one response.
How good is DeepSeek V4.1 Flash?
Epoch AI gives DeepSeek V4.1 Flash a Capabilities Index score of 155.0 (likely range 148.8–157.6), ranking it #29 of 148 models Epoch has scored.
Is DeepSeek V4.1 Flash open source?
Yes. DeepSeek publishes the weights under the MIT licence, so it can be downloaded and self-hosted.
What inputs does DeepSeek V4.1 Flash support?
DeepSeek V4.1 Flash accepts text and images and replies in text. It is a reasoning model with low, high and max effort settings, supports tool calling and can return structured JSON output.
When was DeepSeek V4.1 Flash released?
DeepSeek released DeepSeek V4.1 Flash on Sep 10, 2026. Its training data runs to May 2025.
What are the best alternatives to DeepSeek V4.1 Flash?
The closest current models from other labs on capability, price and release date are GLM-5.3-Flash (Z.ai (Zhipu), ECI 151.9, $0.15 / $0.50), GPT-5.6 Luna (OpenAI, ECI 156.5, $0.20 / $1.20), Gemini 3.6 Flash (Google, ECI 154.3, $0.75 / $3.75) and Inkling Small (Thinking Machines, ECI 150.2, $0.50 / $1.20).