DataLearner logo

Qwen3.6-35B-A3BvsGLM-4.7-Flash

Across 3 shared benchmarks, Qwen3.6-35B-A3B leads overall: Qwen3.6-35B-A3B wins 3, GLM-4.7-Flash wins 0, with 0 ties and an average score difference of +16.12.

阿里巴巴
Qwen3.6-35B-A3B

阿里巴巴 · 2026-04-16 · Reasoning model

智谱AI
GLM-4.7-Flash

智谱AI · 2026-01-19 · Reasoning model

Qwen3.6-35B-A3B3 wins(100%)(0%)0 winsGLM-4.7-Flash

Benchmark scores

Grouped by capability, sorted by largest gap within each. 3 shared benchmarks.

Coding and Software Engineer

Qwen3.6-35B-A3B 1/1
BenchmarkQwen3.6-35B-A3BGLM-4.7-FlashDiff
SWE-bench Verified73.4048 / 114Thinking (No Tools)59.2084 / 114Thinking (No Tools)+14.20

General Evaluation

Qwen3.6-35B-A3B 1/1
BenchmarkQwen3.6-35B-A3BGLM-4.7-FlashDiff
GPQA Diamond84.85100 / 270Normal (No Tools)66209 / 270Normal (No Tools)+18.85

General Knowledge

Qwen3.6-35B-A3B 1/1
BenchmarkQwen3.6-35B-A3BGLM-4.7-FlashDiff
HLE21.40127 / 185Thinking (No Tools)6.10175 / 185Normal (No Tools)+15.30

Specs

FieldQwen3.6-35B-A3BGLM-4.7-Flash
Publisher阿里巴巴智谱AI
Release date2026-04-162026-01-19
Model typeReasoning modelReasoning model
ArchitectureMoEMoE
Parameters35B31B
Context length200K200K
Max output80K128K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemQwen3.6-35B-A3BGLM-4.7-Flash
Text inputNot public¥0 / 1M tokens
Text outputNot public¥0 / 1M tokens

One or both models have incomplete public pricing.

Summary

  • Qwen3.6-35B-A3Bleads in:Coding and Software Engineer (1/1), General Evaluation (1/1), General Knowledge (1/1)

On average across the 3 shared benchmarks, Qwen3.6-35B-A3B scores 16.12 higher.

Largest single-benchmark gap: GPQA Diamond — Qwen3.6-35B-A3B 84.85 vs GLM-4.7-Flash 66 (+18.85).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.