DeepSeek-R1-Distill-Llama-70B
DeepSeek-R1-Distill-Llama-70B is an AI model published by DeepSeek-AI, released on 2025-01-20, for Reasoning model, with 70B parameters, and 128K context length, requiring about 140GB storage, with a 94.50 score on MATH-500.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
Model basics
Open source & experience
Official resources
API details
| Type | Condition | Input | Output |
|---|---|---|---|
| Text | - | $0.600/ 1M tokens | $1.20/ 1M tokens |
“—” means the modality is not billed in that direction, or the vendor has not published a price for it.
Benchmark Results
DeepSeek-R1-Distill-Llama-70B currently shows benchmark results led by MATH-500 (28 / 45, score 94.50), AIME2025 (151 / 215, score 53.70), GPQA Diamond (355 / 462, score 65.20). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.
General Evaluation
2 evaluationsMath and Reasoning
2 evaluationsAgent Level Benchmark
2 evaluationsPublisher
Model Overview
DeepSeek-R1-Distill-Llama-70B is an AI model published by DeepSeek-AI, released on 2025-01-20, for Reasoning model, with 70B parameters, and 128K context length, requiring about 140GB storage, with a 94.50 score on MATH-500.
DataLearner on WeChat
Follow DataLearner on WeChat for AI model updates and research notes.
