ZMIME
DeepSeek · Released Jul 31, 2026

DeepSeek V4 Flash 0731

Official DeepSeek V4 Flash release with enhanced agentic capabilities and integrated DSpark speculative decoding.

  • Open weights
  • Reasoning
  • Tool calling
Capability
154.5ECI · #32 of 148
Input
$0.14per 1M tokens
Output
$0.28per 1M tokens
Context
1M384K max output
Pricing

What it costs

Input
$0.14per million tokens
Output
$0.28per million tokens
Blended (3:1)
$0.175cheaper than 79% of priced models

Typical monthly bills

Monthly usageEstimated cost
Side project2M input + 0.5M output tokens$0.42
Team assistant25M input + 5M output tokens$4.90
Production app250M input + 50M output tokens$49.00

Median of 48 paid API providers on models.dev.

Capability

Independent benchmarks

154.5Epoch Capabilities Index
#32 of 148 scored models
88Median 146167

The shaded band is Epoch AI’s confidence range (152.0–156.6); the tick marks the median scored model.

  • GPQA DiamondGraduate-level science questions · max setting91.0%
  • FrontierMath Tiers 1–3Research-level mathematics · max setting57.5%
  • OTIS Mock AIME 2024–2025Competition mathematics · max setting94.4%
  • SimpleQA VerifiedShort factual questions · max setting33.6%

Scores from Epoch AI, run independently of DeepSeek.

Specifications

The details

Lab
DeepSeek
Released
Jul 31, 2026
Knowledge cutoff
May 2025
Context window
1,000,000 tokens
Max output
384,000 tokens
Inputs
Text
Output
Text
Reasoning
Always on
Tool calling
Yes
Structured output
Yes
Availability
49 API providerslisted on models.dev
Lab-reported

What DeepSeek claims

Published by the lab at launch. Settings vary, so compare these only with care.

BenchmarkScoreSettingSource
Terminal-Bench v2.182.7maxSource
NL2Repo54.2 resolve ratemax effortSource
CyberGym76.7max effortSource
DeepSWE54.4 resolve ratemax effortSource
Toolathlon-Verified70.3max effortSource
Agents' Last Exam25.2max effortSource
AutomationBench25.1 success ratemax effortSource
DSBench-FullStack68.7max effortSource
DSBench-Hard59.6max effortSource
Same lab

More from DeepSeek

Questions

About DeepSeek V4 Flash 0731

How much does DeepSeek V4 Flash 0731 cost?

DeepSeek V4 Flash 0731 costs $0.14 input / $0.28 output per million tokens (median across 48 API providers). At a 3:1 input-to-output mix that is $0.175 per million tokens, cheaper than 79% of the 360 priced models we track.

What is the context window of DeepSeek V4 Flash 0731?

DeepSeek V4 Flash 0731 accepts up to 1,000,000 tokens per request and can write up to 384,000 tokens in one response.

How good is DeepSeek V4 Flash 0731?

Epoch AI gives DeepSeek V4 Flash 0731 a Capabilities Index score of 154.5 (likely range 152.0–156.6), ranking it #32 of 148 models Epoch has scored. Epoch AI benchmark results: GPQA Diamond 91.0%, FrontierMath Tiers 1–3 57.5%, OTIS Mock AIME 2024–2025 94.4%, SimpleQA Verified 33.6%.

Is DeepSeek V4 Flash 0731 open source?

Yes. DeepSeek publishes the weights under the MIT licence, so it can be downloaded and self-hosted.

What inputs does DeepSeek V4 Flash 0731 support?

DeepSeek V4 Flash 0731 accepts text and replies in text. It is a reasoning model, supports tool calling and can return structured JSON output.

When was DeepSeek V4 Flash 0731 released?

DeepSeek released DeepSeek V4 Flash 0731 on Jul 31, 2026. Its training data runs to May 2025.

What are the best alternatives to DeepSeek V4 Flash 0731?

The closest current models from other labs on capability, price and release date are GPT-6 Luna (OpenAI, $0.10 / $0.50), GLM-5.3-Flash (Z.ai (Zhipu), ECI 151.9, $0.15 / $0.50), Gemini 3.6 Flash (Google, ECI 154.3, $0.75 / $3.75) and Inkling Small (Thinking Machines, ECI 150.2, $0.50 / $1.20).