DataLearner logo

Gemini 2.5 Pro Experimental 03-25 Benchmark Details

Gemini 2.5 Pro Experimental 03-25 currently shows benchmark results led by AIME 2024 (9 / 62, score 92), Aider-Polyglot (12 / 59, score 72.90), SimpleQA (13 / 47, score 52.90).

Benchmark Results

Gemini 2.5 Pro Experimental 03-25

Benchmark Results

Thinking
Tool usage

General Knowledge

2 evaluations
Benchmark / mode
Score
Rank/total
84
59 / 188
18.80
124 / 173

Common Sense

1 evaluations
Benchmark / mode
Score
Rank/total
52.90
13 / 47

Coding and Software Engineer

2 evaluations
Benchmark / mode
Score
Rank/total
70.40
54 / 123
63.80
77 / 113

Math and Reasoning

3 evaluations
Benchmark / mode
Score
Rank/total
92
9 / 62
86.90
47 / 107
4.20
40 / 80

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
Simple Bench
Standard Mode
51.60
27 / 63

Agent Level Benchmark

1 evaluations
Benchmark / mode
Score
Rank/total
Aider-Polyglot
Standard Mode
72.90
12 / 59

Claw-style Agent Evaluation

2 evaluations
Benchmark / mode
Score
Rank/total
Claw Bench
Thinking EnabledTools
80.40
20 / 29
Pinch Bench
Thinking EnabledTools
71.90
29 / 37