DataLearner logo

MiniCPM5-1B Benchmark Details

MiniCPM5-1B currently shows benchmark results led by τ²-Bench - Telecom (106 / 264, score 82.50), IF Bench (160 / 282, score 49.30), BBH (13 / 21, score 71.89). This page also compares it with 2 competitor models and 1 predecessor or same-series models, including performance and pricing views when available.

Benchmark Results

MiniCPM5-1B

Benchmark Results

Thinking
Tool usage

General Knowledge

4 evaluations
Benchmark / mode
Score
Rank/total
BBH
Thinking Mode
71.89
13 / 21
MMLU-Pro
Thinking Mode
48.85
142 / 176
HLE
Standard Mode
4.50
502 / 565
HLE
Thinking Mode
6.50
454 / 565

Other

2 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Standard Mode
26.90
451 / 463
GPQA Diamond
Thinking Mode
26.26
454 / 463

Agent Level Benchmark

2 evaluations
Benchmark / mode
Score
Rank/total
τ²-Bench - Telecom
Standard ModeTools
82.50
106 / 264
τ²-Bench - Telecom
Thinking ModeTools
81
109 / 264

Instruction Following

2 evaluations
Benchmark / mode
Score
Rank/total
IF Bench
Standard Mode
35.20
245 / 282
IF Bench
Thinking Mode
49.30
160 / 282

Long Context

1 evaluations
Benchmark / mode
Score
Rank/total
AA-LCR
Standard Mode
5
171 / 171

Competitor Comparison

Benchmark scores for MiniCPM5-1B compared against top models in its class

Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

6 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.

BenchmarkMiniCPM5-1BCurrentGemma 4 E2BGemma 4 E4B
HLE
Accuracy
综合评估
6.50Thinking Enabled
4.80Thinking Enabled
4.80Standard Mode
MMLU-Pro
Accuracy
综合评估
48.85Thinking Enabled
60.00Thinking Enabled
69.40Thinking Enabled
GPQA Diamond
Accuracy
科学与综合推理
26.90Standard Mode
43.30Thinking Enabled
58.60Thinking Enabled
τ²-Bench - Telecom
Accuracy
Agent能力评测
82.50Standard Mode | Tools
22.20Standard Mode | Tools
26.00Standard Mode | Tools
IF Bench
Accuracy
指令跟随
49.30Thinking Enabled
38.00Thinking Enabled
44.20Thinking Enabled
AA-LCR
Accuracy
长上下文能力
5.00Standard Mode
16.30Standard Mode
24.00Standard Mode

Standard API Pricing: MiniCPM5-1B vs. Peer Models

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier.

Comparable standard text pricing is not available for these models.

Version History

How each version of the MiniCPM5-1B series stacks up on benchmark tests

MiniCPM5-1BMiniCPM-1B-SFT
No benchmark data matches the selected filters.

Standard API Pricing Across the MiniCPM5-1B Series

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier.

Comparable standard text pricing is not available for these models.