DataLearner logo
DE

DeepSeek-R1

Reasoning modelDeepSeek R1DeepSeek R1

DeepSeek-R1

Release date: 2025-01-20Updated: 2025-03-21Views: 1,864
Live demoGitHubHugging FaceCompare
Parameters
671B
Context length
128K
Multilingual
Supported
Reasoning ability
No data

DeepSeek-R1 is an AI model published by DeepSeek-AI, released on 2025-01-20, for Reasoning model, with 671B parameters, and 128K context length, requiring about 134GB storage, with a 1500.00 score on Creative Writing.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

DeepSeek-R1

Model basics

Reasoning traces
Supported
Thinking modes
Thinking modes not supported
Context length
128K tokens
Max output length
No data
Model type
Reasoning model
Modality (in / out)
No data
Release date
2025-01-20
Model file size
134GB
MoE architecture
No
Total params / Active params
671B / Not applicable
Knowledge cutoff
No data
DeepSeek-R1

Open source & experience

Code license
Weights license
MIT License- Commercial use permitted
GitHub repo
N/A
Live demo
N/A
DeepSeek-R1

Official resources

Paper
DataLearnerAI blog
DeepSeek-R1

API details

API speed
No data
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-$0.550/ 1M$2.19/ 1M
Cache PricingPrompt Cache
TypeTTLWriteRead
Text-$0.550/ 1M$0.140/ 1M
Text-$0.550/ 1M
Cache = write
Text-$0.140/ 1M
Cache = hit

“—” means the modality is not billed in that direction, or the vendor has not published a price for it.

DeepSeek-R1

Benchmark Results

DeepSeek-R1 currently shows benchmark results led by MMLU (8 / 124, score 90.80), MATH-500 (13 / 45, score 97.30), MMLU Pro (39 / 134, score 84). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage

General Knowledge

3 evaluations
Benchmark / mode
Score
Rank/total
MMLU
Standard Mode
90.80
8 / 124
MMLU Pro
Standard Mode
84
39 / 134
ARC-AGI-1
Standard Mode
15.80
82 / 92

General Evaluation

1 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Standard Mode
71.50
187 / 274

Common Sense

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleQA
Standard Mode
30.10
24 / 47

Coding and Software Engineer

3 evaluations
Benchmark / mode
Score
Rank/total
LiveCodeBench
Standard Mode
65.90
64 / 128
SWE-bench Verified
Standard Mode
49.20
98 / 116
WeirdML v2
Standard ModeTools
36.49
49 / 52

Math and Reasoning

4 evaluations
Benchmark / mode
Score
Rank/total
MATH-500
Standard Mode
97.30
13 / 45
AIME 2024
Standard Mode
79.80
28 / 62
AIME2025
Standard Mode
70
74 / 107
3.80
17 / 19

Writing and Creative Capabilities

1 evaluations
Benchmark / mode
Score
Rank/total
Creative Writing
Standard Mode
1500
51 / 106

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleBench
Standard Mode
30.90
76 / 94

Agent Level Benchmark

3 evaluations
Benchmark / mode
Score
Rank/total
Aider-Polyglot
Thinking Mode
56.90
25 / 59
BALROG
Standard ModeTools
34.90
8 / 12
METR Time Horizons v1.1
Standard ModeTools
26.93
21 / 22

Long Context

1 evaluations
Benchmark / mode
Score
Rank/total
Fiction.liveBench
Standard Mode
69.40
11 / 16
DeepSeek-R1

Publisher

DeepSeek-R1

Model Overview

DeepSeek-R1 is an AI model published by DeepSeek-AI, released on 2025-01-20, for Reasoning model, with 671B parameters, and 128K context length, requiring about 134GB storage, with a 1500.00 score on Creative Writing.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code