DataLearner logo

GLM-4.7 Benchmark Details

GLM-4.7 currently shows benchmark results led by LiveCodeBench (17 / 123, score 84.90), τ²-Bench (6 / 43, score 87.40), AIME2025 (23 / 107, score 95.70).

Benchmark Results

GLM-4.7

Benchmark Results

Thinking
Tool usage

General Knowledge

5 evaluations
Benchmark / mode
Score
Rank/total
85.70
49 / 187
84.30
37 / 132
LiveBench
Standard Mode
58.09
78 / 115
42.80
52 / 172
24.80
103 / 172

Coding and Software Engineer

3 evaluations
Benchmark / mode
Score
Rank/total
84.90
17 / 123
73.80
43 / 112

Math and Reasoning

3 evaluations
Benchmark / mode
Score
Rank/total
95.70
23 / 107
92.90
8 / 18
2.10
56 / 80

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
Simple Bench
Thinking Enabled
47.70
29 / 63

Agent Level Benchmark

2 evaluations
Benchmark / mode
Score
Rank/total
87.40
6 / 43

AI Agent - Information Search

1 evaluations
Benchmark / mode
Score
Rank/total
52
41 / 53

AI Agent - Tool Usage

2 evaluations
Benchmark / mode
Score
Rank/total
MCP-Atlas
Standard ModeTools
58.10
22 / 27