DataLearner logo

Claude Opus 5vsGLM-5.2

Across 3 shared benchmarks, Claude Opus 5 leads overall: Claude Opus 5 wins 3, GLM-5.2 wins 0, with 0 ties and an average score difference of +17.30.

Anthropic
Claude Opus 5

Anthropic · 2026-07-24 · Reasoning model

智谱AI
GLM-5.2

智谱AI · 2026-06-13 · Reasoning model

Claude Opus 53 wins(100%)(0%)0 winsGLM-5.2

Benchmark scores

Grouped by capability, sorted by largest gap within each. 3 shared benchmarks.

Coding and Software Engineer

Claude Opus 5 2/2
BenchmarkClaude Opus 5GLM-5.2Diff
DeepSWE68.804 / 19Max (With Tools)4414 / 19Deep Thinking (With Tools)+24.80
SWE-Bench Pro - Public79.202 / 54Max (With Tools)62.108 / 54Thinking (With Tools)+17.10

General Knowledge

Claude Opus 5 1/1
BenchmarkClaude Opus 5GLM-5.2Diff
HLE64.701 / 172Max (With Tools)54.7013 / 172Thinking (With Tools)+10

Specs

FieldClaude Opus 5GLM-5.2
PublisherAnthropic智谱AI
Release date2026-07-242026-06-13
Model typeReasoning modelReasoning model
ArchitectureDenseMoE
ParametersNot available753.33B
Context length1M1M
Max output128K128K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemClaude Opus 5GLM-5.2
Text input$5 / 1M tokens$1.4 / 1M tokens
Text output$25 / 1M tokens$4.4 / 1M tokens
Cache read$0.5 / 1M tokens$0.26 / 1M tokens
Cache write$6.25 / 1M tokensNot public

Summary

  • Claude Opus 5leads in:Coding and Software Engineer (2/2), General Knowledge (1/1)

On average across the 3 shared benchmarks, Claude Opus 5 scores 17.30 higher.

Largest single-benchmark gap: DeepSWE — Claude Opus 5 68.80 vs GLM-5.2 44 (+24.80).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.