DataLearner logo
DE

DeepSeek-V3-0324

Chat modelDeepSeek VDeepSeek V3

DeepSeek-V3-0324

Release date: 2025-03-24Updated: 2025-08-23Views: 1,829
Parameters
671B
Context length
128K
Multilingual
Supported
Reasoning ability
3/5

DeepSeek-V3-0324 is an AI model published by DeepSeek-AI, released on 2025-03-24, for Chat model, with 671B parameters, and 128K context length, requiring about 1442GB storage, with a 1470.20 score on Creative Writing.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

DeepSeek-V3-0324

Model basics

Reasoning traces
No data
Thinking modes
Standard Mode (Default)Thinking Mode
Context length
128K tokens
Max output length
No data
Model type
Chat model
Modality (in / out)
Text → Text
Release date
2025-03-24
Model file size
1442GB
MoE architecture
Yes
Total params / Active params
671B / 37B
Knowledge cutoff
No data
DeepSeek-V3-0324

Open source & experience

Code license
Weights license
MIT License- Commercial use permitted
DeepSeek-V3-0324

Official resources

Paper
N/A
DataLearnerAI blog
DeepSeek-V3-0324

API details

API speed
3/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-$0.200/ 1M tokens$0.880/ 1M tokens
Cache PricingPrompt Cache
TypeTTLWriteRead
Text-$0.200/ 1M tokens$0.106/ 1M tokens
Text-$0.200/ 1M tokens
Cache = write
Text-$0.106/ 1M tokens
Cache = hit

“—” means the modality is not billed in that direction, or the vendor has not published a price for it.

DeepSeek-V3-0324

Benchmark Results

DeepSeek-V3-0324 currently shows benchmark results led by GSM8K (3 / 70, score 96.30), MMLU (29 / 124, score 86.50), GPQA (4 / 17, score 68.40). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage

General Knowledge

5 evaluations
Benchmark / mode
Score
Rank/total
MMLU
Standard Mode
86.50
29 / 124
MMLU Pro
Standard Mode
81.20
56 / 134
GPQA
Standard Mode
68.40
4 / 17
ARC-AGI-1
Standard Mode
9
86 / 92
HLE
Standard Mode
5.20
190 / 197

Math and Reasoning

7 evaluations
Benchmark / mode
Score
Rank/total
GSM8K
Standard Mode
96.30
3 / 70
MATH-500
Standard Mode
94
29 / 45
AIME 2024
Standard Mode
59.40
43 / 62
AIME2025
Standard Mode
47.70
89 / 107
IMO-ProofBench
Standard Mode
4.30
15 / 16
IMO 2024
Standard Mode
1.70
9 / 10
IMO 2025
Standard Mode
1.70
9 / 9

Reading Comprehension

1 evaluations
Benchmark / mode
Score
Rank/total
DROP
Standard Mode
89.70
3 / 9

General Evaluation

1 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Standard Mode
68.40
205 / 274

Common Sense

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleQA
Standard Mode
27.20
28 / 47

Coding and Software Engineer

2 evaluations
Benchmark / mode
Score
Rank/total
LiveCodeBench
Standard Mode
49.20
100 / 128
SWE-bench Verified
Standard Mode
38.80
107 / 116

Writing and Creative Capabilities

1 evaluations
Benchmark / mode
Score
Rank/total
Creative Writing
Standard Mode
1470.20
55 / 106

AI Agent - Tool Usage

1 evaluations
Benchmark / mode
Score
Rank/total
Terminal-Bench
Standard Mode
13.30
34 / 35

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleBench
Standard Mode
27.20
79 / 94

Agent Level Benchmark

2 evaluations
Benchmark / mode
Score
Rank/total
Aider-Polyglot
Standard Mode
55.10
27 / 59
τ²-Bench
Standard ModeTools
38.80
39 / 44
DeepSeek-V3-0324

Publisher

DeepSeek-V3-0324

Model Overview

DeepSeek-V3-0324 is an AI model published by DeepSeek-AI, released on 2025-03-24, for Chat model, with 671B parameters, and 128K context length, requiring about 1442GB storage, with a 1470.20 score on Creative Writing.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code