DataLearner logo

Qwen3.8-Max-0902 Benchmark Details

Qwen3.8-Max-0902 currently shows benchmark results led by NL2Repo-Bench (1 / 12, score 64.90), AutomationBench (1 / 12, score 50.80), DeepSWE (4 / 32, score 69.30). This page also compares it with 3 competitor models and 3 predecessor or same-series models, including performance and pricing views when available.

Benchmark Results

Qwen3.8-Max-0902

Benchmark Results

Thinking
Tool usage

Coding and Software Engineer

5 evaluations
Benchmark / mode
Score
Rank/total
DeepSWE
Extra-HighTools
69.30
4 / 32
NL2Repo-Bench
Extra-HighTools
64.90
1 / 12
MLS Bench
Extra-HighTools
50.10
1 / 5
SWE-Marathon
Extra-HighTools
44.80
1 / 6
Program Bench
Extra-HighTools
28
5 / 7

AI Agent - Tool Usage

4 evaluations
Benchmark / mode
Score
Rank/total
ClawEval-MM
Extra-HighTools
80.20
1 / 2
Toolathlon-Verified
Extra-HighTools
73.30
7 / 10
AutomationBench
Extra-HighTools
50.80
1 / 12
Terminal-Bench 3.0
Extra-HighTools
29
3 / 7

Agent Level Benchmark

2 evaluations
Benchmark / mode
Score
Rank/total
CoWorkBench
Extra-HighTools
76.10
1 / 2
Job Bench
Extra-HighTools
64
1 / 4

Multimodal Understanding

3 evaluations
Benchmark / mode
Score
Rank/total
BabyVision
Extra-HighTools
93.80
1 / 6
MMMU-Pro
Extra-High
82.70
2 / 8
ERQA
Extra-High
78.30
1 / 2

Competitor Comparison

Benchmark scores for Qwen3.8-Max-0902 compared against top models in its class

Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

11 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.

BenchmarkQwen3.8-Max-0902CurrentKimi K3GLM-5.3DeepSeek-V4-Pro
DeepSWE
编程与软件工程
69.30Thinking Level · Extra High | Tools
67.50Thinking Level · High | Tools
66.90Thinking Level · High | Tools
62.70Thinking Level · Extra High | Tools
MLS Bench
编程与软件工程
50.10Thinking Level · Extra High | Tools
48.30Thinking Level · High | Tools
--
--
NL2Repo-Bench
编程与软件工程
64.90Thinking Level · Extra High | Tools
--
58.00Thinking Level · High | Tools
61.50Thinking Level · Extra High | Tools
Program Bench
编程与软件工程
28.00Thinking Level · Extra High | Tools
77.80Thinking Level · High | Tools
19.00Thinking Level · High | Tools
--
SWE-Marathon
编程与软件工程
44.80Thinking Level · Extra High | Tools
42.00Thinking Level · High | Tools
42.50Thinking Level · High | Tools
--
AutomationBench
AI Agent - 工具使用
50.80Thinking Level · Extra High | Tools
30.80Thinking Level · High | Tools
48.20Thinking Level · High | Tools
31.80Thinking Level · Extra High | Tools
Terminal-Bench 3.0
AI Agent - 工具使用
29.00Thinking Level · Extra High | Tools
--
28.30Thinking Level · High | Tools
--
Toolathlon-Verified
AI Agent - 工具使用
73.30Thinking Level · Extra High | Tools
76.50Thinking Level · High | Tools
73.00Thinking Level · High | Tools
74.10Thinking Level · Extra High | Tools
Job Bench
Agent能力评测
64.00Thinking Level · Extra High | Tools
54.30Thinking Level · High | Tools
--
--
BabyVision
多模态理解
93.80Thinking Level · Extra High | Tools
85.70Thinking Level · High | Tools
--
--
MMMU-Pro
多模态理解
82.70Thinking Level · Extra High
83.40Thinking Level · High | Tools
--
--

Standard API Pricing: Qwen3.8-Max-0902 vs. Peer Models

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier.

These models use different currencies or billing units, so the page falls back to raw price values instead of a shared bar chart.

Qwen3.8-Max-0902
Supplier: 阿里巴巴
Standard input: $2 / 1M tokens
Standard output: $6 / 1M tokens
Kimi K3
Supplier: Moonshot AI
Standard input: ¥20 / 1M tokens
Standard output: ¥100 / 1M tokens
GLM-5.3
Supplier: 智谱AI
Standard input: $1.4 / 1M tokens
Standard output: $4.4 / 1M tokens
DeepSeek-V4-Pro
Supplier: DeepSeek-AI
Standard input: $0.435 / 1M tokens
Standard output: $0.87 / 1M tokens
ModelSupplierStandard inputStandard outputBase price applies to
Qwen3.8-Max-0902
阿里巴巴$2 / 1M tokens$6 / 1M tokens
Kimi K3
Moonshot AI¥20 / 1M tokens¥100 / 1M tokens
GLM-5.3
智谱AI$1.4 / 1M tokens$4.4 / 1M tokens
DeepSeek-V4-Pro
DeepSeek-AI$0.435 / 1M tokens$0.87 / 1M tokens

Version History

How each version of the Qwen3.8-Max-0902 series stacks up on benchmark tests

Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

5 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.

BenchmarkQwen3.8-Max-0902CurrentQwen3.8-Max
DeepSWE
编程与软件工程
69.30Thinking Level · Extra High | Tools
56.60Thinking Level · Extra High | Tools
MLS Bench
编程与软件工程
50.10Thinking Level · Extra High | Tools
41.00Thinking Level · Extra High | Tools
NL2Repo-Bench
编程与软件工程
64.90Thinking Level · Extra High | Tools
55.90Thinking Level · Extra High | Tools
AutomationBench
AI Agent - 工具使用
50.80Thinking Level · Extra High | Tools
27.30Thinking Level · Extra High | Tools
Toolathlon-Verified
AI Agent - 工具使用
73.30Thinking Level · Extra High | Tools
72.50Thinking Level · Extra High | Tools

Single-Benchmark Version Trend

Viewing: DeepSWE · 编程与软件工程

Benchmark
NormalNormal + ToolsThinkingThinking + ToolsDeepDeep + Tools

X-axis shows model and release date, Y-axis shows score; solid lines connect the same mode across versions, while dotted guides align modes within the same generation.

Standard API Pricing Across the Qwen3.8-Max-0902 Series

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier.

These models use different currencies or billing units, so the page falls back to raw price values instead of a shared bar chart.

Qwen3.8-Max-0902
Supplier: 阿里巴巴
Standard input: $2 / 1M tokens
Standard output: $6 / 1M tokens
Qwen3.8-Max
Supplier: 阿里巴巴
Standard input: ¥12 / 1M tokens
Standard output: ¥36 / 1M tokens
Qwen3.7 Max
Supplier: 阿里巴巴
Standard input: ¥12 / 1M tokens
Standard output: ¥36 / 1M tokens
Qwen3.6-Max-Preview
Supplier: 阿里巴巴
Standard input: $1.3 / 1M tokens
Standard output: $7.8 / 1M tokens
Base price applies to <= 128
ModelSupplierStandard inputStandard outputBase price applies to
Qwen3.8-Max-0902
阿里巴巴$2 / 1M tokens$6 / 1M tokens
Qwen3.8-Max
阿里巴巴¥12 / 1M tokens¥36 / 1M tokens
Qwen3.7 Max
阿里巴巴¥12 / 1M tokens¥36 / 1M tokens
Qwen3.6-Max-Preview
阿里巴巴$1.3 / 1M tokens$7.8 / 1M tokens<= 128