DataLearner logo

MiniCPM5-1BvsGemma 4 E4B

Across 3 shared benchmarks, Gemma 4 E4B leads overall: MiniCPM5-1B wins 0, Gemma 4 E4B wins 3, with 0 ties and an average score difference of -23.96.

OpenBMB
MiniCPM5-1B

OpenBMB · 2026-05-01 · Reasoning model

DeepMind
Gemma 4 E4B

DeepMind · 2026-04-02 · Multimodal model

MiniCPM5-1B0 wins(0%)(100%)3 winsGemma 4 E4B

Benchmark scores

Grouped by capability, sorted by largest gap within each. 3 shared benchmarks.

General Evaluation

Gemma 4 E4B 1/1
BenchmarkMiniCPM5-1BGemma 4 E4BDiff
GPQA Diamond26.26268 / 274Thinking (No Tools)58.60232 / 274Thinking (No Tools)-32.34

General Knowledge

Gemma 4 E4B 1/1
BenchmarkMiniCPM5-1BGemma 4 E4BDiff
MMLU Pro48.85124 / 134Thinking (No Tools)69.4096 / 134Thinking (No Tools)-20.55

Long Context

Gemma 4 E4B 1/1
BenchmarkMiniCPM5-1BGemma 4 E4BDiff
AA-LCR5170 / 170Normal (No Tools)24165 / 170Normal (No Tools)-19

Specs

FieldMiniCPM5-1BGemma 4 E4B
PublisherOpenBMBDeepMind
Release date2026-05-012026-04-02
Model typeReasoning modelMultimodal model
ArchitectureMoEDense
Parameters1.08B8B
Context length128K128K
Max outputNot available8K

Summary

  • Gemma 4 E4Bleads in:General Evaluation (1/1), General Knowledge (1/1), Long Context (1/1)

On average across the 3 shared benchmarks, Gemma 4 E4B scores 23.96 higher.

Largest single-benchmark gap: GPQA Diamond — MiniCPM5-1B 26.26 vs Gemma 4 E4B 58.60 (-32.34).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.