DataLearner logo

Qwen3.6-Max-PreviewvsGLM 5.1

Across 6 shared benchmarks, Qwen3.6-Max-Preview leads overall: Qwen3.6-Max-Preview wins 3, GLM 5.1 wins 2, with 1 ties and an average score difference of +1.20.

阿里巴巴
Qwen3.6-Max-Preview

阿里巴巴 · 2026-04-18 · Chat model

智谱AI
GLM 5.1

智谱AI · 2026-03-27 · Reasoning model

Qwen3.6-Max-Preview3 wins(50%)Ties1(33%)2 winsGLM 5.1

Benchmark scores

Grouped by capability, sorted by largest gap within each. 6 shared benchmarks.

AI Agent - Tool Usage

Qwen3.6-Max-Preview 1/1
BenchmarkQwen3.6-Max-PreviewGLM 5.1Diff
Terminal Bench 2.065.4011 / 48Deep Thinking (With Tools)63.5013 / 48Thinking (With Tools)+1.90

Coding and Software Engineer

GLM 5.1 1/1
BenchmarkQwen3.6-Max-PreviewGLM 5.1Diff
SWE-Bench Pro - Public57.3020 / 57Deep Thinking (With Tools)58.4017 / 57Thinking (With Tools)-1.10

Commonsense Reasoning

Qwen3.6-Max-Preview 1/1
BenchmarkQwen3.6-Max-PreviewGLM 5.1Diff
SimpleBench6311 / 67Normal (No Tools)58.7022 / 67Normal (No Tools)+4.30

General Evaluation

Qwen3.6-Max-Preview 1/1
BenchmarkQwen3.6-Max-PreviewGLM 5.1Diff
GPQA Diamond90.4036 / 226最高(无工具)86.2078 / 226Thinking (No Tools)+4.20

General Knowledge

GLM 5.1 1/1
BenchmarkQwen3.6-Max-PreviewGLM 5.1Diff
HLE50.2029 / 181Thinking (With Tools)52.3021 / 181Thinking (With Tools)-2.10

Math and Reasoning

Even 1/1
BenchmarkQwen3.6-Max-PreviewGLM 5.1Diff
IMO-AnswerBench83.8014 / 23最高(无工具)83.8014 / 23Thinking (No Tools)

Specs

FieldQwen3.6-Max-PreviewGLM 5.1
Publisher阿里巴巴智谱AI
Release date2026-04-182026-03-27
Model typeChat modelReasoning model
ArchitectureDenseMoE
ParametersNot available75.4B
Context length262K200K
Max output64K125K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemQwen3.6-Max-PreviewGLM 5.1
Text input$1.3 / 1M tokens$1.4 / 1M tokens
Text output$7.8 / 1M tokens$4.4 / 1M tokens
Cache readNot public$4.4 / 1M tokens
Cache writeNot public$0.26 / 1M tokens

Summary

  • Qwen3.6-Max-Previewleads in:AI Agent - Tool Usage (1/1), Commonsense Reasoning (1/1), General Evaluation (1/1)
  • GLM 5.1leads in:Coding and Software Engineer (1/1), General Knowledge (1/1)
  • Tied in:Math and Reasoning

On average across the 6 shared benchmarks, Qwen3.6-Max-Preview scores 1.20 higher.

Largest single-benchmark gap: SimpleBench — Qwen3.6-Max-Preview 63 vs GLM 5.1 58.70 (+4.30).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.