DataLearner logo

Qwen1.5-72B-Chat Benchmark Details

Qwen1.5-72B-Chat currently shows benchmark results led by C-Eval (10 / 48, score 84.10), GSM8K (24 / 68, score 79.50), MMLU (62 / 124, score 77.50).

Benchmark Results

Qwen1.5-72B-Chat

Benchmark Results

Thinking

Knowledge Exams

3 evaluations
Benchmark / mode
Score
Rank/total
C-Eval
Standard Mode
84.10
10 / 48
MMLU
Standard Mode
77.50
62 / 124
MMLU-Pro
unknown
52.64
137 / 175

Mathematics

1 evaluations
Benchmark / mode
Score
Rank/total
GSM8K
Standard Mode
79.50
24 / 68

Code Generation & Editing

2 evaluations
Benchmark / mode
Score
Rank/total
MBPP
Standard Mode
53.40
61 / 95
HumanEval
Standard Mode
41.50
95 / 140