DataLearner logo

Grok 4 Benchmark Details

Grok 4 currently shows benchmark results led by IMO 2024 (1 / 10, score 23.20), MMLU Pro (14 / 132, score 87), IMO 2025 (1 / 9, score 29.20).

Benchmark Results

Grok 4

Benchmark Results

Thinking

General Knowledge

8 evaluations
Benchmark / mode
Score
Rank/total
87
14 / 132
87
42 / 188
66.70
32 / 68
LiveBench
Standard Mode
62.02
59 / 115
38.60
65 / 173
38.60
65 / 173
25.40
101 / 173
15.90
37 / 62

Coding and Software Engineer

2 evaluations
Benchmark / mode
Score
Rank/total
82
25 / 123
58.60
84 / 113

Math and Reasoning

9 evaluations
Benchmark / mode
Score
Rank/total
98.80
13 / 107
91.70
36 / 107
46.70
4 / 16
23.30
10 / 16
29.20
1 / 9
23.20
1 / 10
12.10
22 / 60
2.10
56 / 80

AI Agent - Tool Usage

1 evaluations
Benchmark / mode
Score
Rank/total

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
Simple Bench
Thinking Enabled
60.50
15 / 63

Agent Level Benchmark

2 evaluations
Benchmark / mode
Score
Rank/total
Aider-Polyglot
Thinking Level · High
79.60
7 / 59
Grok 4 Benchmark Results & Rankings | DataLearnerAI