Model record

DeepSeek V4 Flash 0731

Open weights

DeepSeek · released Jul 31, 2026

Overall standing

Overall
Insufficient data
0 of 8 measured
Category standingsReasoning Coding Math Agentic
Reasoning
Insufficient data
0 of 4 measured
Coding
Insufficient data
0 of 2 measured
Math
Insufficient data
0 of 1 measured
Agentic
Insufficient data
0 of 1 measured
Model factsOpen weights · $0.14 / $0.28 per 1M

Copied from the linked provider page; no separate retrieval date is stored.

Scores and sources

Measured scores retain their source and retrieval date. Missing scores are not treated as zero.

Benchmark scores, evaluation settings, sources, and retrieval dates for DeepSeek V4 Flash 0731
BenchmarkCategoryScoreEvaluation settingsSource
Terminal-Bench v2.1CodingNot measured in this datasetNo source
τ³-BankingAgenticNot measured in this datasetNo source
AA-LCRReasoningNot measured in this datasetNo source
Humanity's Last ExamReasoningNot measured in this datasetNo source
GPQA DiamondReasoningNot measured in this datasetNo source
SciCodeCodingNot measured in this datasetNo source
IFBenchReasoningNot measured in this datasetNo source
CritPtMathNot measured in this datasetNo source