DataLearner logo

GLM-4.6 Benchmark Details

GLM-4.6 currently shows benchmark results led by AIME2025 (15 / 107, score 98.60), LiveCodeBench (18 / 123, score 84.50), MMLU Pro (45 / 132, score 83).

Benchmark Results

GLM-4.6

Benchmark Results

Thinking

General Knowledge

9 evaluations
Benchmark / mode
Score
Rank/total
83
45 / 132
78
72 / 132
82.90
67 / 188
81
76 / 188
63
142 / 188
LiveBench
Standard Mode
55.19
81 / 115
30.40
87 / 173
17.20
132 / 173
5.20
166 / 173

Coding and Software Engineer

5 evaluations
Benchmark / mode
Score
Rank/total
84.50
18 / 123
82.80
24 / 123
56
80 / 123

Math and Reasoning

4 evaluations
Benchmark / mode
Score
Rank/total
98.60
15 / 107
98.60
15 / 107
44
93 / 107
2.10
56 / 80

AI Agent - Tool Usage

1 evaluations
Benchmark / mode
Score
Rank/total
40.50
12 / 35

Agent Level Benchmark

2 evaluations
Benchmark / mode
Score
Rank/total
75.90
21 / 43

Instruction Following

1 evaluations
Benchmark / mode
Score
Rank/total
43
31 / 31

AI Agent - Information Search

1 evaluations
Benchmark / mode
Score
Rank/total
45.10
45 / 53