DataLearner logo

Gemma 4 26B A4BvsGLM-4.7-Flash

Across 3 shared benchmarks, Gemma 4 26B A4B leads overall: Gemma 4 26B A4B wins 2, GLM-4.7-Flash wins 1, with 0 ties and an average score difference of -0.47.

DeepMind
Gemma 4 26B A4B

DeepMind · 2026-04-02 · Chat model

智谱AI
GLM-4.7-Flash

智谱AI · 2026-01-19 · Reasoning model

Gemma 4 26B A4B2 wins(67%)(33%)1 winGLM-4.7-Flash

Benchmark scores

Grouped by capability, sorted by largest gap within each. 3 shared benchmarks.

Agent Level Benchmark

GLM-4.7-Flash 1/1
BenchmarkGemma 4 26B A4BGLM-4.7-FlashDiff
τ²-Bench68.2026 / 43Thinking (With Tools)79.5016 / 43-11.30

General Evaluation

Gemma 4 26B A4B 1/1
BenchmarkGemma 4 26B A4BGLM-4.7-FlashDiff
GPQA Diamond82.30108 / 226Thinking (No Tools)75.20137 / 226+7.10

General Knowledge

Gemma 4 26B A4B 1/1
BenchmarkGemma 4 26B A4BGLM-4.7-FlashDiff
HLE17.20140 / 181Thinking (With Tools + Internet)14.40145 / 181+2.80

Specs

FieldGemma 4 26B A4BGLM-4.7-Flash
PublisherDeepMind智谱AI
Release date2026-04-022026-01-19
Model typeChat modelReasoning model
ArchitectureMoEMoE
Parameters25.2B31B
Context length256K200K
Max output32K128K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemGemma 4 26B A4BGLM-4.7-Flash
Text inputNot public¥0 / 1M tokens
Text outputNot public¥0 / 1M tokens

One or both models have incomplete public pricing.

Summary

  • Gemma 4 26B A4Bleads in:General Evaluation (1/1), General Knowledge (1/1)
  • GLM-4.7-Flashleads in:Agent Level Benchmark (1/1)

On average across the 3 shared benchmarks, GLM-4.7-Flash scores 0.47 higher.

Largest single-benchmark gap: τ²-Bench — Gemma 4 26B A4B 68.20 vs GLM-4.7-Flash 79.50 (-11.30).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.