DataLearner logo
DE

DeepSeek-V3

Chat modelDeepSeek VDeepSeek V3

DeepSeek-V3

Release date: 2024-12-26Updated: 2025-03-21 11:14:411,583
Parameters
681B
Context length
128K
Chinese support
Supported
Reasoning ability

DeepSeek-V3 is an AI model published by DeepSeek-AI, released on 2024-12-26, for Chat model, with 681B parameters, and 128K context length, requiring about 687.9 GB storage, with a 92.30 score on BBH.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

DeepSeek-V3

Model basics

Reasoning traces
Not supported
Thinking modes
Thinking modes not supported
Context length
128K tokens
Max output length
No data
Model type
Chat model
Modality (in / out)
No data
Release date
2024-12-26
Model file size
687.9 GB
MoE architecture
No
Total params / Active params
681B / N/A
Knowledge cutoff
No data
DeepSeek-V3

Open source & experience

Code license
Weights license
DeepSeek License Agreement- Commercial use permitted
Live demo
No live demo
DeepSeek-V3

Official resources

Paper
DataLearnerAI blog
DeepSeek-V3

API details

API speed
No data
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-$0.270/ 1M$1.10/ 1M
Cache PricingPrompt Cache
TypeTTLWriteRead
Text-$0.270/ 1M$0.070/ 1M
DeepSeek-V3

Benchmark Results

DeepSeek-V3 currently shows benchmark results led by BBH (3 / 21, score 92.30), MATH (7 / 42, score 87.80), HumanEval (9 / 39, score 89). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking

General Knowledge

5 evaluations
Benchmark / mode
Score
Rank/total
92.30
3 / 21
88.50
17 / 66
75.90
83 / 132
59.10
148 / 187
59.10
6 / 15

Coding and Software Engineer

2 evaluations
Benchmark / mode
Score
Rank/total
89
9 / 39
34.60
110 / 123

Math and Reasoning

4 evaluations
Benchmark / mode
Score
Rank/total
87.80
7 / 42
87.80
39 / 44
39
52 / 62
1.70
49 / 60

Common Sense

1 evaluations
Benchmark / mode
Score
Rank/total
24.90
31 / 47

Writing and Creative Capabilities

1 evaluations
Benchmark / mode
Score
Rank/total
81.60
15 / 23

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
Simple Bench
Standard Mode
18.90
59 / 63

Agent Level Benchmark

1 evaluations
Benchmark / mode
Score
Rank/total
Aider-Polyglot
Standard Mode
48.40
34 / 59

Compare with other models

No curated comparisons for this model yet.

Want a custom combination? Open the compare tool

DeepSeek-V3

Publisher

DeepSeek-V3

Model Overview

DeepSeek-V3 is an AI model published by DeepSeek-AI, released on 2024-12-26, for Chat model, with 681B parameters, and 128K context length, requiring about 687.9 GB storage, with a 92.30 score on BBH.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code