Qwen3.8-Max Benchmark Details
Qwen3.8-Max currently shows benchmark results led by Context Arena (3 / 126, score 96.92), LongBench v2 (1 / 13, score 66.30), HLE (15 / 189, score 56.20). This page also compares it with 3 competitor models and 3 predecessor or same-series models, including performance and pricing views when available.
Benchmark Results
Benchmark Results
Writing and Creative Capabilities
1 evaluationsCoding and Software Engineer
5 evaluationsText Embedding
3 evaluationsAI Agent - Tool Usage
2 evaluationsMath and Reasoning
2 evaluationsCompetitor Comparison
Benchmark scores for Qwen3.8-Max compared against top models in its class
12 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | Qwen3.8-MaxCurrent | Claude Opus 5 | GPT-5.6 Sol | Kimi K3 |
|---|---|---|---|---|
56.20Thinking Level · Extra High | Tools | 64.70Thinking Level · High | Tools | 49.50Thinking Level · High | 56.00Thinking Level · High | Tools | |
92.60Thinking Level · Extra High | 93.88Thinking Level · High | 93.50Thinking Level · High | 93.50Thinking Level · High | |
1842.00Standard Mode | 2116.10Standard Mode | 1964.10Standard Mode | 2070.80Standard Mode | |
62.50Thinking Level · Extra High | 80.60Thinking Level · High | 64.80Thinking Level · Extra High | 60.70Thinking Level · High | |
56.60Thinking Level · Extra High | Tools | 68.80Thinking Level · High | Tools | 72.70Thinking Level · Extra High | Tools | 67.50Thinking Level · High | Tools | |
73.50Thinking Level · Extra High | Tools | -- | -- | 81.20Thinking Level · High | Tools | |
41.00Thinking Level · Extra High | Tools | -- | -- | 48.30Thinking Level · High | Tools | |
67.70Thinking Level · Extra High | Tools | 79.20Thinking Level · High | Tools | 64.60Thinking Level · Extra High | Tools | -- | |
96.92Thinking Level · Extra High | 97.72Thinking Level · High | 97.63Thinking Level · High | 71.75Thinking Level · High | |
86.60Thinking Level · Extra High | Tools | -- | 88.80Thinking Level · High | 88.30Thinking Level · High | Tools | |
72.50Thinking Level · Extra High | Tools | -- | -- | 76.50Thinking Level · High | Tools | |
27.00Thinking Level · Extra High | Tools | -- | 52.70Thinking Level · Extra High | Tools | 28.30Thinking Level · High | Tools |
Standard API Pricing: Qwen3.8-Max vs. Peer Models
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier.
These models use different currencies or billing units, so the page falls back to raw price values instead of a shared bar chart.
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
Qwen3.8-Max | 阿里巴巴 | ¥12 / 1M tokens | ¥36 / 1M tokens | — |
Claude Opus 5 | Anthropic | $5 / 1M tokens | $25 / 1M tokens | — |
GPT-5.6 Sol | OpenAI | $4 / 1M tokens | $20 / 1M tokens | — |
Kimi K3 | Moonshot AI | ¥20 / 1M tokens | ¥100 / 1M tokens | — |
Version History
How each version of the Qwen3.8-Max series stacks up on benchmark tests
5 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.
| Benchmark | Qwen3.8-MaxCurrent | Qwen3.7 Max | Qwen3.6-Max-Preview |
|---|---|---|---|
56.20Thinking Level · Extra High | Tools | 53.50Thinking Enabled | Tools | 50.20Thinking Enabled | Tools | |
92.60Thinking Level · Extra High | 92.40Thinking Level · High | 90.40Thinking Level · High | |
62.50Thinking Level · Extra High | 70.40Standard Mode | 63.00Standard Mode | |
67.70Thinking Level · Extra High | Tools | 60.60Thinking Enabled | Tools | 57.30Deep Thinking Mode | Tools | |
96.92Thinking Level · Extra High | 56.01Standard Mode | 88.14Thinking Enabled |
Single-Benchmark Version Trend
Viewing: HLE · 综合评估
Standard API Pricing Across the Qwen3.8-Max Series
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier.
These models use different currencies or billing units, so the page falls back to raw price values instead of a shared bar chart.
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
Qwen3.8-Max | 阿里巴巴 | ¥12 / 1M tokens | ¥36 / 1M tokens | — |
Qwen3.7 Max | 阿里巴巴 | ¥12 / 1M tokens | ¥36 / 1M tokens | — |
Qwen3.6-Max-Preview | 阿里巴巴 | $1.3 / 1M tokens | $7.8 / 1M tokens | <= 128 |