DataLearner logo
DE

DeepSeek-R1

Reasoning modelDeepSeek R1DeepSeek R1

DeepSeek-R1

Release date: 2025-01-20Updated: 2025-03-21 11:14:181,823
Live demoGitHubHugging FaceCompare
Parameters
671B
Context length
128K
Chinese support
Supported
Reasoning ability

DeepSeek-R1 is an AI model published by DeepSeek-AI, released on 2025-01-20, for Reasoning model, with 671B parameters, and 128K context length, requiring about 134GB storage, with a 97.30 score on MATH-500.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

DeepSeek-R1

Model basics

Reasoning traces
Supported
Thinking modes
Thinking modes not supported
Context length
128K tokens
Max output length
No data
Model type
Reasoning model
Modality (in / out)
No data
Release date
2025-01-20
Model file size
134GB
MoE architecture
No
Total params / Active params
671B / N/A
Knowledge cutoff
No data
DeepSeek-R1

Open source & experience

Code license
Weights license
MIT License- Commercial use permitted
GitHub repo
GitHub link unavailable
Live demo
No live demo
DeepSeek-R1

Official resources

Paper
DataLearnerAI blog
DeepSeek-R1

API details

API speed
No data
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-$0.550/ 1M$2.19/ 1M
Cache PricingPrompt Cache
TypeTTLWriteRead
Text-$0.550/ 1M$0.140/ 1M
DeepSeek-R1

Benchmark Results

DeepSeek-R1 currently shows benchmark results led by MMLU (8 / 66, score 90.80), MMLU Pro (38 / 132, score 84), MATH-500 (13 / 44, score 97.30). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking

General Knowledge

4 evaluations
Benchmark / mode
Score
Rank/total
90.80
8 / 66
84
38 / 132
71.50
109 / 187
15.80
58 / 68

Common Sense

1 evaluations
Benchmark / mode
Score
Rank/total
30.10
24 / 47

Coding and Software Engineer

2 evaluations
Benchmark / mode
Score
Rank/total
65.90
60 / 123
49.20
96 / 112

Math and Reasoning

3 evaluations
Benchmark / mode
Score
Rank/total
97.30
13 / 44
79.80
28 / 62
70
74 / 107

Writing and Creative Capabilities

1 evaluations
Benchmark / mode
Score
Rank/total
84.60
11 / 23

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
Simple Bench
Standard Mode
30.90
48 / 63

Agent Level Benchmark

1 evaluations
Benchmark / mode
Score
Rank/total
Aider-Polyglot
Thinking Mode
56.90
25 / 59

Compare with other models

No curated comparisons for this model yet.

Want a custom combination? Open the compare tool

DeepSeek-R1

Publisher

DeepSeek-R1

Model Overview

DeepSeek-R1 is an AI model published by DeepSeek-AI, released on 2025-01-20, for Reasoning model, with 671B parameters, and 128K context length, requiring about 134GB storage, with a 97.30 score on MATH-500.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code