DataLearner logo

Qwen3-Coder-480B-A35B Benchmark Details

Qwen3-Coder-480B-A35B currently shows benchmark results led by Terminal-Bench (15 / 35, score 37.50), LiveCodeBench (137 / 250, score 58.50), Terminal Bench Hard (150 / 244, score 18.90).

Benchmark Results

Qwen3-Coder-480B-A35B

Benchmark Results

Thinking
Tool usage

General Knowledge

1 evaluations
Benchmark / mode
Score
Rank/total
HLE
Standard Mode
4.50
500 / 563

Other

1 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Standard Mode
61.80
370 / 462

Coding and Software Engineer

3 evaluations
Benchmark / mode
Score
Rank/total
CodeClash
Standard ModeTools
952
8 / 8
SWE-bench Verified
Standard Mode
67
75 / 116
LiveCodeBench
Standard Mode
58.50
137 / 250

Math and Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
AIME2025
Standard Mode
39.30
167 / 215

AI Agent - Tool Usage

1 evaluations
Benchmark / mode
Score
Rank/total
Terminal-Bench
Standard Mode
37.50
15 / 35

Agent Level Benchmark

2 evaluations
Benchmark / mode
Score
Rank/total
τ²-Bench - Telecom
Standard ModeTools
43.60
184 / 264
Terminal Bench Hard
Standard ModeTools
18.90
150 / 244

Instruction Following

1 evaluations
Benchmark / mode
Score
Rank/total
IF Bench
Standard Mode
40.50
214 / 282