DE
DeepSeek-V2-MoE-236B
Release date: 2024-05-06Updated: 2024-05-07 08:37:17705
Parameters
236B
Context length
128K
Chinese support
Supported
Reasoning ability
DeepSeek-V2-MoE-236B is an AI model published by DeepSeek-AI, released on 2024-05-06, for Foundation model, with 236B parameters, and 128K context length, requiring about 472GB storage.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
DeepSeek-V2-236B
Model basics
Reasoning traces
Not supported
Thinking modes
Thinking modes not supported
Context length
128K tokens
Max output length
No data
Model type
Foundation model
Modality (in / out)
No data
Release date
2024-05-06
Model file size
472GB
MoE architecture
No
Total params / Active params
236B / N/A
Knowledge cutoff
No data
DeepSeek-V2-236B
Open source & experience
Code license
Weights license
DeepSeek License Agreement- Commercial use permitted
GitHub repo
Hugging Face
Live demo
No live demo
DeepSeek-V2-236B
Official resources
Paper
No paper available
DataLearnerAI blog
No blog post yet
DeepSeek-V2-236B
API details
API speed
No data
No public API pricing yet.
DeepSeek-V2-236B
Benchmark Results
No benchmark data to show.
Compare with other models
No curated comparisons for this model yet.
Want a custom combination? Open the compare tool
DeepSeek-V2-236B
Publisher
DeepSeek-AI
View publisher details DeepSeek-V2-MoE-236B
Model Overview
DeepSeek-V2-MoE-236B is an AI model published by DeepSeek-AI, released on 2024-05-06, for Foundation model, with 236B parameters, and 128K context length, requiring about 472GB storage.
DataLearner on WeChat
Follow DataLearner on WeChat for AI model updates and research notes.
