DeepSeek-V4.1-Flash Benchmark Details
DeepSeek-V4.1-Flash currently shows benchmark results led by HLE (3 / 197, score 63.90), Terminal-Bench 2.1 (1 / 53, score 90.60), CodeForces (1 / 21, score 3471). This page also compares it with 3 competitor models and 2 predecessor or same-series models, including performance and pricing views when available.
Benchmark Results
Benchmark Results
General Knowledge
3 evaluationsCoding and Software Engineer
5 evaluationsAI Agent - Tool Usage
4 evaluationsMultimodal Understanding
3 evaluationsCompetitor Comparison
Benchmark scores for DeepSeek-V4.1-Flash compared against top models in its class
12 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | DeepSeek-V4.1-FlashCurrent | Kimi K3 | GLM-5.3 | Claude Opus 5 |
|---|---|---|---|---|
63.90Thinking Level · High | Tools | 59.80Thinking Level · High | Tools | 62.50Thinking Level · High | Tools | 63.60Thinking Level · High | Tools | |
90.90Thinking Level · High | 92.90Thinking Level · High | 88.10Thinking Level · High | 93.40Thinking Level · High | |
74.20Thinking Level · High | Tools | 67.50Thinking Level · High | Tools | 66.90Thinking Level · High | Tools | 74.00Thinking Level · High | Tools | |
65.40Thinking Level · High | Tools | 58.00Thinking Level · High | Tools | 58.00Thinking Level · High | Tools | 75.30Thinking Level · High | Tools | |
20.30Thinking Level · High | Tools | 17.50Thinking Level · High | Tools | 19.00Thinking Level · High | Tools | 37.00Thinking Level · High | Tools | |
88.10Thinking Level · High | Tools | 80.00Thinking Level · High | Tools | 84.50Thinking Level · High | Tools | -- | |
90.60Thinking Level · High | Tools | 88.30Thinking Level · High | Tools | 88.20Thinking Level · High | Tools | 89.10Thinking Level · High | Tools | |
30.00Thinking Level · High | Tools | 17.70Thinking Level · High | Tools | 28.30Thinking Level · High | Tools | 43.30Thinking Level · High | Tools | |
31.20Thinking Level · High | Tools | 12.60Thinking Level · High | Tools | 37.90Thinking Level · High | Tools | 51.80Thinking Level · High | Tools | |
31.80Thinking Level · High | Tools | 27.60Thinking Level · High | Tools | 28.50Thinking Level · High | Tools | 28.60Thinking Level · High | Tools | |
54.80Thinking Level · High | Tools | 46.70Thinking Level · High | Tools | 48.80Thinking Level · High | Tools | 50.30Thinking Level · High | Tools | |
89.60Thinking Level · High | Tools | 85.70Thinking Level · High | Tools | -- | 94.10Thinking Level · High | Tools |
Standard API Pricing: DeepSeek-V4.1-Flash vs. Peer Models
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier.
These models use different currencies or billing units, so the page falls back to raw price values instead of a shared bar chart.
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
Kimi K3 | Moonshot AI | ¥20 / 1M tokens | ¥100 / 1M tokens | — |
GLM-5.3 | 智谱AI | $1.4 / 1M tokens | $4.4 / 1M tokens | — |
Claude Opus 5 | Anthropic | $5 / 1M tokens | $25 / 1M tokens | — |
Version History
How each version of the DeepSeek-V4.1-Flash series stacks up on benchmark tests
12 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.
| Benchmark | DeepSeek-V4.1-FlashCurrent | DeepSeek-V4-Flash-Vision-Exp | DeepSeek-V4-Flash |
|---|---|---|---|
63.90Thinking Level · High | Tools | -- | 51.50Thinking Level · High | Tools | |
90.90Thinking Level · High | -- | 89.90Thinking Level · High | |
3471.00Thinking Level · High | -- | 3289.00Thinking Level · High | |
74.20Thinking Level · High | Tools | 59.30Thinking Level · High | Tools | 54.40Thinking Level · High | Tools | |
65.40Thinking Level · High | Tools | 57.70Thinking Level · High | Tools | 54.20Thinking Level · High | Tools | |
62.80Thinking Level · High | Tools | -- | 30.90Thinking Level · High | Tools | |
88.10Thinking Level · High | Tools | 75.30Thinking Level · High | Tools | 76.70Thinking Level · High | Tools | |
90.60Thinking Level · High | Tools | 83.90Thinking Level · High | Tools | 82.70Thinking Level · High | Tools | |
30.00Thinking Level · High | Tools | -- | 7.60Thinking Level · High | Tools | |
31.20Thinking Level · High | Tools | -- | 7.00Thinking Level · High | Tools | |
31.80Thinking Level · High | Tools | 27.30Thinking Level · High | Tools | 25.20Thinking Level · High | Tools | |
54.80Thinking Level · High | Tools | 25.70Thinking Level · High | Tools | 37.70Thinking Level · High | Tools |
Single-Benchmark Version Trend
Viewing: HLE · 综合评估
Standard API Pricing Across the DeepSeek-V4.1-Flash Series
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
DeepSeek-V4-Flash | DeepSeek-AI | $0.14 / 1M tokens | $0.28 / 1M tokens | — |