DeepSeek-V2-MoE-236B-Chat
DeepSeek-V2-MoE-236B-Chat is an AI model published by DeepSeek-AI, released on 2024-05-06, for Chat model, with 236B parameters, and 128K context length, requiring about 472GB storage, with a 54.81 score on MMLU Pro.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
Model basics
Open source & experience
Official resources
API details
Benchmark Results
DeepSeek-V2-236B-Chat currently shows benchmark results led by MMLU Pro (133 / 176, score 54.81), Aider-Polyglot (52 / 59, score 17.80). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.
Publisher
Model Overview
DeepSeek-V2-MoE-236B-Chat is an AI model published by DeepSeek-AI, released on 2024-05-06, for Chat model, with 236B parameters, and 128K context length, requiring about 472GB storage, with a 54.81 score on MMLU Pro.
DataLearner on WeChat
Follow DataLearner on WeChat for AI model updates and research notes.
