DataLearner logo

OpenAI o3-mini Benchmark Details

OpenAI o3-mini currently shows benchmark results led by Aider-Polyglot (21 / 59, score 60.40), AIME2025 (48 / 107, score 86.50), FrontierMath - Tier 4 (40 / 80, score 4.20).

Benchmark Results

OpenAI o3-mini

Benchmark Results

Thinking

General Knowledge

3 evaluations
Benchmark / mode
Score
Rank/total
84.90
41 / 66
70.60
117 / 187
13.40
139 / 172

Coding and Software Engineer

2 evaluations
Benchmark / mode
Score
Rank/total
2073
13 / 16
40.80
102 / 112

Math and Reasoning

4 evaluations
Benchmark / mode
Score
Rank/total
95.80
24 / 44
86.50
48 / 107
60
42 / 62
FrontierMath - Tier 4
Thinking Level · High
4.20
40 / 80

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
Simple Bench
Thinking Level · High
22.80
55 / 63

Agent Level Benchmark

2 evaluations
Benchmark / mode
Score
Rank/total
Aider-Polyglot
Thinking Level · Medium
53.80
29 / 59
Aider-Polyglot
Thinking Level · High
60.40
21 / 59