Qwen3.8-Max-0902 leads on the shared-benchmark average
Ahead on 3 of 4 shared benchmarks, averaging 7.0 points higher
Summarised only from the 4 percentage-scale benchmarks scored by every selected model; details are below.
“Best available” takes each model’s highest recorded non-parallel mode per benchmark, so it may combine modes into a virtual configuration that does not exist. Read it with the mode breakdown.

Qwen3.8-Max-0902
阿里巴巴
Benchmark-by-benchmark comparison. Changing the thinking mode or tool filters updates the chart and table below.
“Best available” picks the highest non-parallel mode separately for each benchmark. The resulting series can combine several reasoning levels and is not one reproducible runtime configuration. Choose a mode filter to compare like-for-like runs.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
Every model and runtime mode, benchmark by benchmark. Values are comparable along a row, not between different benchmarks.
4 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | Qwen3.8-Max-0902 | DeepSeek-V4-Pro |
|---|---|---|
DeepSWE 编程与软件工程 | 69.30Thinking Level · Extra High | Tools | 62.70Thinking Level · Extra High | Tools |
NL2Repo-Bench 编程与软件工程 | 64.90Thinking Level · Extra High | Tools | 61.50Thinking Level · Extra High | Tools |
AutomationBench AI Agent - 工具使用 | 50.80Thinking Level · Extra High | Tools | 31.80Thinking Level · Extra High | Tools |
Toolathlon-Verified AI Agent - 工具使用 | 73.30Thinking Level · Extra High | Tools | 74.10Thinking Level · Extra High | Tools |
Official list prices per model API, split by input and output. Unit: USD per 1M tokens.
Architecture, licensing and API modalities. "Not provided" means the field is missing from our database.
| Features & specs | Qwen3.8-Max-0902阿里巴巴 | DeepSeek-V4-ProDeepSeek-AI |
|---|---|---|
Core specsRelease | 2026-09-02 | 2026-08-13 |
Context length | 1M | 1M |
Total parameters | 2.4T | 1.6T |
Active parameters | 95B | 49B |
Max output length | 131,072 tokens | 384,000 tokens |
Architecture | MoE (mixture of experts) | MoE (mixture of experts) |
Runtime modes | 低高极高 | 关闭高最高 |
Availability & licensingCode availability | Not public | Available · MIT License |
Weight availability | Not public | Available · MIT License |
Use & commercial terms | Official service only; subject to provider terms | 免费商用授权 |
Local deploymentWeights | Not provided | Hugging Face |
API modality supportText Input/Output | Input:YesOutput:Yes | Input:YesOutput:Yes |
Image Input/Output | Input:YesOutput:No | Input:NoOutput:No |
Video Input/Output | Input:YesOutput:No | Input:NoOutput:No |
ResourcesPaper / report | Qwen3.8-Max just got upgraded. Meet Qwen3.8-Max-0902! | DeepSeek-V4: Towards Highly Efficient Million-Token Context Intelligence |

DeepSeek-V4-Pro
DeepSeek-AI