DataLearner logo

DeepSeek-V4.1-FlashvsDeepSeek-V4-Flash

Across 12 shared benchmarks, DeepSeek-V4.1-Flash leads overall: DeepSeek-V4.1-Flash wins 12, DeepSeek-V4-Flash wins 0, with 0 ties and an average score difference of +29.58.

DeepSeek-AI
DeepSeek-V4.1-Flash

DeepSeek-AI · 2026-09-10 · Multimodal model

DeepSeek-AI
DeepSeek-V4-Flash

DeepSeek-AI · 2026-04-24 · Reasoning model

DeepSeek-V4.1-Flash12 wins(100%)(0%)0 winsDeepSeek-V4-Flash

Benchmark scores

Grouped by capability, sorted by largest gap within each. 12 shared benchmarks.

AI Agent - Tool Usage

DeepSeek-V4.1-Flash 4/4
BenchmarkDeepSeek-V4.1-FlashDeepSeek-V4-FlashDiff
Terminal-Bench 4.031.207 / 16Max (With Tools)716 / 16Max (With Tools)+24.20
Terminal-Bench 3.0304 / 11Max (With Tools)7.6011 / 11Max (With Tools)+22.40
CyberGym88.101 / 8Max (With Tools)76.707 / 8Max (With Tools)+11.40
Terminal-Bench 2.190.601 / 53Max (With Tools)82.7024 / 53Max (With Tools)+7.90

Coding and Software Engineer

DeepSeek-V4.1-Flash 4/4
BenchmarkDeepSeek-V4.1-FlashDeepSeek-V4-FlashDiff
CodeForces3,4711 / 21Max (No Tools)3,2893 / 21Max (No Tools)+182
SEC-Bench Pro62.804 / 5Max (With Tools)30.905 / 5Max (With Tools)+31.90
DeepSWE74.202 / 38Max (With Tools)54.4026 / 38Max (With Tools)+19.80
NL2Repo-Bench65.402 / 16Max (With Tools)54.2012 / 16Max (With Tools)+11.20

Agent Capability

DeepSeek-V4.1-Flash 1/1
BenchmarkDeepSeek-V4.1-FlashDeepSeek-V4-FlashDiff
ExploitGym (budget unspecified)15.302 / 4Max (With Tools)1.804 / 4Max (With Tools)+13.50

Agent Level Benchmark

DeepSeek-V4.1-Flash 1/1
BenchmarkDeepSeek-V4.1-FlashDeepSeek-V4-FlashDiff
Agents' Last Exam31.806 / 19Max (With Tools)25.2016 / 19Max (With Tools)+6.60

Math and Reasoning

DeepSeek-V4.1-Flash 1/1
BenchmarkDeepSeek-V4.1-FlashDeepSeek-V4-FlashDiff
MathArena Apex65.601 / 3Max (No Tools)58.603 / 3Max (No Tools)+7

Productivity Knowledge

DeepSeek-V4.1-Flash 1/1
BenchmarkDeepSeek-V4.1-FlashDeepSeek-V4-FlashDiff
AutomationBench54.801 / 17Max (With Tools)37.7010 / 17Max (With Tools)+17.10

Specs

FieldDeepSeek-V4.1-FlashDeepSeek-V4-Flash
PublisherDeepSeek-AIDeepSeek-AI
Release date2026-09-102026-04-24
Model typeMultimodal modelReasoning model
ArchitectureMoEMoE
Parameters552B284B
Context length1M1M
Max output384K384K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemDeepSeek-V4.1-FlashDeepSeek-V4-Flash
Text input¥1 / 1M tokens$0.14 / 1M tokens
Text output¥4 / 1M tokens$0.28 / 1M tokens
Cache read¥0.02 / 1M tokens$0.0028 / 1M tokens

Summary

  • DeepSeek-V4.1-Flashleads in:AI Agent - Tool Usage (4/4), Coding and Software Engineer (4/4), Agent Capability (1/1), Agent Level Benchmark (1/1), Math and Reasoning (1/1), Productivity Knowledge (1/1)

On average across the 12 shared benchmarks, DeepSeek-V4.1-Flash scores 29.58 higher.

Largest single-benchmark gap: CodeForces — DeepSeek-V4.1-Flash 3,471 vs DeepSeek-V4-Flash 3,289 (+182).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.