DataLearner logo

Inkling Benchmark Details

Inkling currently shows benchmark results led by AIME 2026 (2 / 19, score 97.10), IF Bench (4 / 34, score 79.80), SWE-bench Verified (26 / 114, score 77.60). This page also compares it with 3 competitor models, including performance and pricing views when available.

Benchmark Results

Inkling

Benchmark Results

Thinking
Tool usage
Internet

General Knowledge

2 evaluations
Benchmark / mode
Score
Rank/total
HLE
Thinking Mode
29.70
100 / 185
HLE
Thinking ModeTools
46
43 / 185

Other

1 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Thinking Mode
87.20
71 / 226

Common Sense

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleQA
Thinking Mode
43.90
17 / 47

Coding and Software Engineer

2 evaluations
Benchmark / mode
Score
Rank/total
SWE-bench Verified
Thinking ModeTools
77.60
26 / 114
SWE-Bench Pro - Public
Thinking ModeTools
54.30
35 / 59

Writing and Creative Capabilities

1 evaluations
Benchmark / mode
Score
Rank/total
Creative Writing
Standard Mode
1607.90
33 / 99

Instruction Following

1 evaluations
Benchmark / mode
Score
Rank/total
IF Bench
Thinking Mode
79.80
4 / 34

AI Agent - Information Search

1 evaluations
Benchmark / mode
Score
Rank/total
BrowseComp
Thinking ModeToolsInternet
77.10
22 / 54

Math and Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
AIME 2026
Thinking Mode
97.10
2 / 19

AI Agent - Tool Usage

3 evaluations
Benchmark / mode
Score
Rank/total
MCP-Atlas
Thinking ModeTools
74.10
25 / 40
MCP-Atlas
Thinking Level · Extra HighTools
76
19 / 40
Terminal-Bench 2.1
Thinking ModeTools
63.80
39 / 47

Competitor Comparison

Benchmark scores for Inkling compared against top models in its class

Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

9 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.

BenchmarkInklingCurrentKimi K3GLM-5.2DeepSeek-V4-Pro
HLE
综合评估
46.00Thinking Enabled | Tools
56.00Thinking Level · High | Tools
54.70Thinking Enabled | Tools
48.20Thinking Level · Extra High | Tools
GPQA Diamond
科学与综合推理
87.20Thinking Enabled
93.50Thinking Level · High
91.86Thinking Level · High
90.10Thinking Level · High
SWE-Bench Pro - Public
编程与软件工程
54.30Thinking Enabled | Tools
--
62.10Thinking Enabled | Tools
55.40Thinking Level · Extra High | Tools
SWE-bench Verified
编程与软件工程
77.60Thinking Enabled | Tools
--
--
80.60Thinking Level · Extra High | Tools
Creative Writing
写作和创作
1607.90Standard Mode
2070.80Standard Mode
1750.90Standard Mode
1552.10Standard Mode
BrowseComp
AI Agent - 信息收集
77.10Thinking Enabled | Tools
91.20Thinking Level · High | Tools
--
83.40Thinking Level · Extra High | Tools
AIME 2026
数学推理
97.10Thinking Enabled
--
99.20Thinking Enabled
--
MCP-Atlas
AI Agent - 工具使用
76.00Thinking Level · Extra High | Tools
84.20Thinking Level · High | Tools
76.80Thinking Enabled | Tools
--
Terminal-Bench 2.1
AI Agent - 工具使用
63.80Thinking Enabled | Tools
88.30Thinking Level · High | Tools
81.00Thinking Level · High | Tools
87.90Thinking Level · Extra High | Tools

Standard API Pricing: Inkling vs. Peer Models

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier.

These models use different currencies or billing units, so the page falls back to raw price values instead of a shared bar chart.

Kimi K3
Supplier: Moonshot AI
Standard input: ¥20 / 1M tokens
Standard output: ¥100 / 1M tokens
GLM-5.2
Supplier: 智谱AI
Standard input: $1.4 / 1M tokens
Standard output: $4.4 / 1M tokens
DeepSeek-V4-Pro
Supplier: DeepSeek-AI
Standard input: $0.435 / 1M tokens
Standard output: $0.87 / 1M tokens
ModelSupplierStandard inputStandard outputBase price applies to
Kimi K3
Moonshot AI¥20 / 1M tokens¥100 / 1M tokens
GLM-5.2
智谱AI$1.4 / 1M tokens$4.4 / 1M tokens
DeepSeek-V4-Pro
DeepSeek-AI$0.435 / 1M tokens$0.87 / 1M tokens