DataLearner logo

Mistral-Small-3.2 Benchmark Details

Mistral-Small-3.2 currently shows benchmark results led by MMLU (53 / 124, score 80.50), MATH (20 / 42, score 69.42), GPQA (11 / 16, score 44.22).

Benchmark Results

Mistral-Small-3.2

Benchmark Results

Thinking

General Knowledge

3 evaluations
Benchmark / mode
Score
Rank/total
MMLU
Standard Mode
80.50
53 / 124
MMLU Pro
Standard Mode
69.06
98 / 133
GPQA
Standard Mode
44.22
11 / 16

Math and Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
MATH
Standard Mode
69.42
20 / 42

Other

1 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Standard Mode
46.13
247 / 270

Common Sense

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleQA
Standard Mode
12.10
39 / 47

Writing and Creative Capabilities

1 evaluations
Benchmark / mode
Score
Rank/total
Creative Writing
Standard Mode
1252.70
72 / 99