DataLearner logo

GLM-4-9B-Chat Benchmark Details

GLM-4-9B-Chat currently shows benchmark results led by GPQA (9 / 17, score 58.50), MMLU Pro (95 / 176, score 72.40), AIME 2024 (34 / 62, score 76.40).

Benchmark Results

GLM-4-9B-Chat

Benchmark Results

Thinking

General Knowledge

3 evaluations
Benchmark / mode
Score
Rank/total
MMLU Pro
Standard Mode
72.40
95 / 176
MMLU Pro
unknown
48
143 / 176
GPQA
Standard Mode
58.50
9 / 17

Math and Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
AIME 2024
Standard Mode
76.40
34 / 62

Coding and Software Engineer

1 evaluations
Benchmark / mode
Score
Rank/total
LiveCodeBench
Standard Mode
51.80
160 / 250