DataLearner logo

DeepSeek-V4.1-FlashvsDeepSeek-V4-Flash-Vision-Exp

Across 8 shared benchmarks, DeepSeek-V4.1-Flash leads overall: DeepSeek-V4.1-Flash wins 8, DeepSeek-V4-Flash-Vision-Exp wins 0, with 0 ties and an average score difference of +13.04.

DeepSeek-AI
DeepSeek-V4.1-Flash

DeepSeek-AI · 2026-09-10 · Multimodal model

DeepSeek-AI
DeepSeek-V4-Flash-Vision-Exp

DeepSeek-AI · 2026-08-21 · Multimodal model

DeepSeek-V4.1-Flash8 wins(100%)(0%)0 winsDeepSeek-V4-Flash-Vision-Exp

Benchmark scores

Grouped by capability, sorted by largest gap within each. 8 shared benchmarks.

AI Agent - Tool Usage

DeepSeek-V4.1-Flash 2/2
BenchmarkDeepSeek-V4.1-FlashDeepSeek-V4-Flash-Vision-ExpDiff
CyberGym88.101 / 8Max (With Tools)75.308 / 8Max (With Tools)+12.80
Terminal-Bench 2.190.601 / 53Max (With Tools)83.9018 / 53Max (With Tools)+6.70

Coding and Software Engineer

DeepSeek-V4.1-Flash 2/2
BenchmarkDeepSeek-V4.1-FlashDeepSeek-V4-Flash-Vision-ExpDiff
DeepSWE74.202 / 38Max (With Tools)59.3021 / 38Max (With Tools)+14.90
NL2Repo-Bench65.402 / 16Max (With Tools)57.708 / 16Max (With Tools)+7.70

Multimodal Understanding

DeepSeek-V4.1-Flash 2/2
BenchmarkDeepSeek-V4.1-FlashDeepSeek-V4-Flash-Vision-ExpDiff
Chartography78.904 / 8Max (With Tools)64.307 / 8Max (With Tools)+14.60
ZeroBench Main493 / 6Max (With Tools)355 / 6Max (With Tools)+14

Agent Level Benchmark

DeepSeek-V4.1-Flash 1/1
BenchmarkDeepSeek-V4.1-FlashDeepSeek-V4-Flash-Vision-ExpDiff
Agents' Last Exam31.806 / 19Max (With Tools)27.3010 / 19Max (With Tools)+4.50

Productivity Knowledge

DeepSeek-V4.1-Flash 1/1
BenchmarkDeepSeek-V4.1-FlashDeepSeek-V4-Flash-Vision-ExpDiff
AutomationBench54.801 / 17Max (With Tools)25.7016 / 17Max (With Tools)+29.10

Specs

FieldDeepSeek-V4.1-FlashDeepSeek-V4-Flash-Vision-Exp
PublisherDeepSeek-AIDeepSeek-AI
Release date2026-09-102026-08-21
Model typeMultimodal modelMultimodal model
ArchitectureMoEMoE
Parameters552B305B
Context length1M1M
Max output384K384K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemDeepSeek-V4.1-FlashDeepSeek-V4-Flash-Vision-Exp
Text input¥1 / 1M tokens$0.22 / 1M tokens
Text output¥4 / 1M tokens$0.66 / 1M tokens
Cache read¥0.02 / 1M tokens$0.007 / 1M tokens

Summary

  • DeepSeek-V4.1-Flashleads in:AI Agent - Tool Usage (2/2), Coding and Software Engineer (2/2), Multimodal Understanding (2/2), Agent Level Benchmark (1/1), Productivity Knowledge (1/1)

On average across the 8 shared benchmarks, DeepSeek-V4.1-Flash scores 13.04 higher.

Largest single-benchmark gap: AutomationBench — DeepSeek-V4.1-Flash 54.80 vs DeepSeek-V4-Flash-Vision-Exp 25.70 (+29.10).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.