DataLearner logo

Qwen 3.6 Plus PreviewvsGLM 5.1

Across 9 shared benchmarks, GLM 5.1 leads overall: Qwen 3.6 Plus Preview wins 2, GLM 5.1 wins 5, with 2 ties and an average score difference of +1.51.

阿里巴巴
Qwen 3.6 Plus Preview

阿里巴巴 · 2026-03-31 · Chat model

智谱AI
GLM 5.1

智谱AI · 2026-03-27 · Reasoning model

Qwen 3.6 Plus Preview2 wins(22%)Ties2(56%)5 winsGLM 5.1

Benchmark scores

Grouped by capability, sorted by largest gap within each. 9 shared benchmarks.

AI Agent - Tool Usage

GLM 5.1 3/3
BenchmarkQwen 3.6 Plus PreviewGLM 5.1Diff
Terminal Bench 2.061.6016 / 48Thinking (With Tools)63.5013 / 48Thinking (With Tools)-1.90
Tool Decathlon39.807 / 10Thinking (With Tools)40.706 / 10Thinking (With Tools)-0.90
Terminal-Bench 2.161.40106 / 191Thinking (With Tools)61.80104 / 191Thinking (With Tools)-0.40

General Knowledge

GLM 5.1 2/2
BenchmarkQwen 3.6 Plus PreviewGLM 5.1Diff
CritPt2.90115 / 200Thinking (No Tools)4.60102 / 200Thinking (No Tools)-1.70
LiveBench68.9142 / 117Normal (No Tools)70.1837 / 117Normal (No Tools)-1.27

Math and Reasoning

Even 2/2
BenchmarkQwen 3.6 Plus PreviewGLM 5.1Diff
AIME 202695.3012 / 29Thinking (No Tools)95.3012 / 29Thinking (No Tools)
IMO-AnswerBench83.8014 / 24Thinking (No Tools)83.8014 / 24Thinking (No Tools)

Agent Level Benchmark

Qwen 3.6 Plus Preview 1/1
BenchmarkQwen 3.6 Plus PreviewGLM 5.1Diff
τ³-Banking20.8088 / 164Thinking (With Tools)13.60115 / 164Thinking (With Tools)+7.20

Claw-style Agent Evaluation

Qwen 3.6 Plus Preview 1/1
BenchmarkQwen 3.6 Plus PreviewGLM 5.1Diff
PinchBench v272.5023 / 45Reported best (effort unspecified)59.9533 / 45Reported best (effort unspecified)+12.55

Specs

FieldQwen 3.6 Plus PreviewGLM 5.1
Publisher阿里巴巴智谱AI
Release date2026-03-312026-03-27
Model typeChat modelReasoning model
ArchitectureDenseMoE
ParametersNot available754B
Context length1M200K
Max output64K125K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemQwen 3.6 Plus PreviewGLM 5.1
Text input$0.5 / 1M tokens$1.4 / 1M tokens
Text output$3 / 1M tokens$4.4 / 1M tokens
Cache read$0.05 / 1M tokens$4.4 / 1M tokens
Cache write$0.625 / 1M tokens$0.26 / 1M tokens

Summary

  • Qwen 3.6 Plus Previewleads in:Agent Level Benchmark (1/1), Claw-style Agent Evaluation (1/1)
  • GLM 5.1leads in:AI Agent - Tool Usage (3/3), General Knowledge (2/2)
  • Tied in:Math and Reasoning

On average across the 9 shared benchmarks, Qwen 3.6 Plus Preview scores 1.51 higher.

Largest single-benchmark gap: PinchBench v2 — Qwen 3.6 Plus Preview 72.50 vs GLM 5.1 59.95 (+12.55).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.