DataLearner logo

GLM-5.2 Benchmark Details

GLM-5.2 currently shows benchmark results led by IMO-AnswerBench (1 / 21, score 91), AIME 2026 (1 / 18, score 99.20), HLE (13 / 172, score 54.70). This page also compares it with 4 competitor models and 3 predecessor or same-series models, including performance and pricing views when available.

Benchmark Results

GLM-5.2

Benchmark Results

Thinking
Tool usage

General Knowledge

4 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Thinking Enabled
91.20
16 / 187
LiveBench
Standard Mode
76.24
9 / 115
HLE
Thinking Enabled
40.50
61 / 172
HLE
Thinking EnabledTools
54.70
13 / 172

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
Simple Bench
Standard Mode
58.80
17 / 63

Coding and Software Engineer

2 evaluations
Benchmark / mode
Score
Rank/total
SWE-Bench Pro - Public
Thinking EnabledTools
62.10
8 / 54
DeepSWE
Deep Thinking ModeTools
44
15 / 20

Math and Reasoning

2 evaluations
Benchmark / mode
Score
Rank/total
AIME 2026
Thinking Enabled
99.20
1 / 18
IMO-AnswerBench
Thinking Enabled
91
1 / 21

AI Agent - Tool Usage

1 evaluations
Benchmark / mode
Score
Rank/total
81
10 / 28

Competitor Comparison

Benchmark scores for GLM-5.2 compared against top models in its class

Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

8 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.

BenchmarkGLM-5.2CurrentKimi K2.7 CodeMiniMax M3DeepSeek-V4-ProClaude Opus 4.8
GPQA Diamond
综合评估
91.20Thinking Enabled
--
--
90.10Thinking Level · High
93.60Thinking Level · High
HLE
综合评估
54.70Thinking Enabled | Tools
--
--
48.20Thinking Level · Extra High | Tools
57.90Extended Thinking | Tools
LiveBench
综合评估
76.24Standard Mode
71.89Standard Mode
70.02Deep Thinking Mode
73.58Standard Mode
78.79Deep Thinking Mode
Simple Bench
常识推理
58.80Standard Mode
--
--
50.90Standard Mode
64.80Standard Mode
DeepSWE
编程与软件工程
44.00Deep Thinking Mode | Tools
31.00Standard Mode | Tools
--
--
59.00Deep Thinking Mode | Tools
SWE-Bench Pro - Public
编程与软件工程
62.10Thinking Enabled | Tools
--
59.00Thinking Enabled | Tools
55.40Thinking Level · Extra High | Tools
69.20Extended Thinking | Tools
IMO-AnswerBench
数学推理
91.00Thinking Enabled
--
--
89.80Thinking Level · High
--
TerminalBench 2.1
AI Agent - 工具使用
81.00Thinking Level · High | Tools
67.04Thinking Enabled | Tools
66.00Thinking Enabled | Tools
--
78.90Thinking Level · High | Tools

Standard API Pricing: GLM-5.2 vs. Peer Models

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier.

These models use different currencies or billing units, so the page falls back to raw price values instead of a shared bar chart.

GLM-5.2
Supplier: 智谱AI
Standard input: $1.4 / 1M tokens
Standard output: $4.4 / 1M tokens
Kimi K2.7 Code
Supplier: Moonshot AI
Standard input: $0.95 / 1M tokens
Standard output: $4 / 1M tokens
MiniMax M3
Supplier: MiniMaxAI
Standard input: ¥2.1 / 1M tokens
Standard output: ¥8.4 / 1M tokens
DeepSeek-V4-Pro
Supplier: DeepSeek-AI
Standard input: $0.435 / 1M tokens
Standard output: $0.87 / 1M tokens
Claude Opus 4.8
Supplier: Anthropic
Standard input: $5 / 1M tokens
Standard output: $25 / 1M tokens
ModelSupplierStandard inputStandard outputBase price applies to
GLM-5.2
智谱AI$1.4 / 1M tokens$4.4 / 1M tokens
Kimi K2.7 Code
Moonshot AI$0.95 / 1M tokens$4 / 1M tokens
MiniMax M3
MiniMaxAI¥2.1 / 1M tokens¥8.4 / 1M tokens
DeepSeek-V4-Pro
DeepSeek-AI$0.435 / 1M tokens$0.87 / 1M tokens
Claude Opus 4.8
Anthropic$5 / 1M tokens$25 / 1M tokens

Version History

How each version of the GLM-5.2 series stacks up on benchmark tests

Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

8 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.

BenchmarkGLM-5.2CurrentGLM 5.1GLM-5GLM-4.7
GPQA Diamond
综合评估
91.20Thinking Enabled
86.20Thinking Enabled
86.00Thinking Enabled
85.70Thinking Enabled
HLE
综合评估
54.70Thinking Enabled | Tools
52.30Thinking Enabled | Tools
50.40Thinking Enabled | Tools
42.80Thinking Enabled | Tools
LiveBench
综合评估
76.24Standard Mode
70.18Standard Mode
68.85Standard Mode
58.09Standard Mode
Simple Bench
常识推理
58.80Standard Mode
--
53.20Standard Mode
47.70Thinking Enabled
SWE-Bench Pro - Public
编程与软件工程
62.10Thinking Enabled | Tools
58.40Thinking Enabled | Tools
--
40.60Thinking Enabled | Tools
AIME 2026
数学推理
99.20Thinking Enabled
95.30Thinking Enabled
92.70Thinking Enabled
92.90Thinking Enabled
IMO-AnswerBench
数学推理
91.00Thinking Enabled
83.80Thinking Enabled
82.50Thinking Enabled
--
TerminalBench 2.1
AI Agent - 工具使用
81.00Thinking Level · High | Tools
58.70Thinking Level · High | Tools
--
--

Single-Benchmark Version Trend

Viewing: GPQA Diamond · 综合评估

Benchmark
NormalNormal + ToolsThinkingThinking + ToolsDeepDeep + Tools

X-axis shows model and release date, Y-axis shows score; solid lines connect the same mode across versions, while dotted guides align modes within the same generation.

Standard API Pricing Across the GLM-5.2 Series

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier.

These models use different currencies or billing units, so the page falls back to raw price values instead of a shared bar chart.

GLM-5.2
Supplier: 智谱AI
Standard input: $1.4 / 1M tokens
Standard output: $4.4 / 1M tokens
GLM 5.1
Supplier: 智谱AI
Standard input: $1.4 / 1M tokens
Standard output: $4.4 / 1M tokens
GLM-5
Supplier: 智谱AI
Standard input: $1 / 1M tokens
Standard output: $3.2 / 1M tokens
GLM-4.7
Supplier: 智谱AI
Standard input: ¥4 / 1M tokens
Standard output: ¥16 / 1M tokens
ModelSupplierStandard inputStandard outputBase price applies to
GLM-5.2
智谱AI$1.4 / 1M tokens$4.4 / 1M tokens
GLM 5.1
智谱AI$1.4 / 1M tokens$4.4 / 1M tokens
GLM-5
智谱AI$1 / 1M tokens$3.2 / 1M tokens
GLM-4.7
智谱AI¥4 / 1M tokens¥16 / 1M tokens