DataLearner logo

Claude Fable 5.1 Benchmark Details

Claude Fable 5.1 currently shows benchmark results led by HLE (1 / 190, score 65), SimpleBench (1 / 92, score 86.60), AA Intelligence Index (1 / 28, score 66). This page also compares it with 3 competitor models and 1 predecessor or same-series models, including performance and pricing views when available.

Benchmark Results

Claude Fable 5.1

Benchmark Results

Thinking
Tool usage

General Knowledge

17 evaluations
Benchmark / mode
Score
Rank/total
ARC-AGI-1
Thinking Level · Low
90
35 / 91
ARC-AGI-1
Thinking Level · Medium
94.50
18 / 91
ARC-AGI-1
Thinking Level · High
96
12 / 91
ARC-AGI-1
Thinking Level · Max
97.50
4 / 91
ARC-AGI-1
Thinking Level · Extra High
96.50
8 / 91
ARC-AGI-2
Thinking Level · Low
78.30
22 / 85
ARC-AGI-2
Thinking Level · Medium
86.30
11 / 85
ARC-AGI-2
Thinking Level · High
88.80
8 / 85
ARC-AGI-2
Thinking Level · Max
90
4 / 85
ARC-AGI-2
Thinking Level · Extra High
90
4 / 85
AA Intelligence Index
Thinking Level · LowTools
58
14 / 28
AA Intelligence Index
Thinking Level · MediumTools
60
11 / 28
AA Intelligence Index
Thinking Level · HighTools
62
5 / 28
AA Intelligence Index
Thinking Level · MaxTools
66
1 / 28
AA Intelligence Index
Thinking Level · Extra HighTools
65
2 / 28
HLE
Thinking Level · Max
60.90
6 / 190
HLE
Thinking Level · MaxTools
65
1 / 190

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleBench
Thinking Enabled
86.60
1 / 92

AI Agent - Tool Usage

4 evaluations
Benchmark / mode
Score
Rank/total
OSWorld 2.0
Thinking Level · MaxTools
77.90
1 / 9
Terminal-Bench 4.0
Thinking Level · MaxTools
55.80
2 / 13
Terminal-Bench-Science 0.1
Thinking Level · MaxTools
52.60
2 / 11
AutomationBench
Thinking Level · MaxTools
31.40
8 / 14

Productivity Knowledge

1 evaluations
Benchmark / mode
Score
Rank/total
GDPval-AA v2
Thinking Level · MaxTools
1853
2 / 25

Coding and Software Engineer

1 evaluations
Benchmark / mode
Score
Rank/total
CursorBench 3.2
Thinking Level · MaxTools
73.40
1 / 5

Competitor Comparison

Benchmark scores for Claude Fable 5.1 compared against top models in its class

Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

10 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.

BenchmarkClaude Fable 5.1CurrentGPT-5.6 Sol ProGPT-5.6 Sol
66.00Thinking Level · High | Tools
--
61.00Thinking Level · High | Tools
ARC-AGI-1
综合评估
97.50Thinking Level · High
--
97.50Thinking Level · Extra High
ARC-AGI-2
综合评估
90.00Thinking Level · Extra High
--
92.50Thinking Level · High
HLE
综合评估
65.00Thinking Level · High | Tools
--
49.50Thinking Level · High
SimpleBench
常识推理
86.60Thinking Enabled
71.70Thinking Level · Extra High
64.80Thinking Level · Extra High
OSWorld 2.0
AI Agent - 工具使用
77.90Thinking Level · High | Tools
--
62.60Thinking Level · Extra High | Tools
Terminal-Bench 4.0
AI Agent - 工具使用
55.80Thinking Level · High | Tools
--
37.27Thinking Level · High | Tools
Terminal-Bench-Science 0.1
AI Agent - 工具使用
52.60Thinking Level · High | Tools
--
22.40Thinking Level · High | Tools
GDPval-AA v2
生产力知识
1853.00Thinking Level · High | Tools
--
1728.00Thinking Level · High | Tools
CursorBench 3.2
编程与软件工程
73.40Thinking Level · High | Tools
--
67.20Thinking Level · High | Tools

Standard API Pricing: Claude Fable 5.1 vs. Peer Models

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens

When a context threshold exists, the charted base price only applies within these limits:

GPT-5.6 Sol Pro: Base price applies to <= 272000
ModelSupplierStandard inputStandard outputBase price applies to
Claude Fable 5.1
Anthropic$10 / 1M tokens$50 / 1M tokens
GPT-5.6 Sol Pro
OpenAI$4 / 1M tokens$20 / 1M tokens<= 272000
GPT-5.6 Sol
OpenAI$4 / 1M tokens$20 / 1M tokens

Version History

How each version of the Claude Fable 5.1 series stacks up on benchmark tests

Claude Fable 5.1Claude Fable 5
Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

9 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.

BenchmarkClaude Fable 5.1CurrentClaude Fable 5
66.00Thinking Level · High | Tools
62.00Thinking Level · High | Tools
ARC-AGI-1
综合评估
97.50Thinking Level · High
98.50Thinking Level · Extra High
ARC-AGI-2
综合评估
90.00Thinking Level · Extra High
89.20Thinking Level · High
HLE
综合评估
65.00Thinking Level · High | Tools
59.00Deep Thinking Mode
SimpleBench
常识推理
86.60Thinking Enabled
81.90Standard Mode
Terminal-Bench 4.0
AI Agent - 工具使用
55.80Thinking Level · High | Tools
44.55Thinking Level · High | Tools
Terminal-Bench-Science 0.1
AI Agent - 工具使用
52.60Thinking Level · High | Tools
21.40Thinking Level · High | Tools
GDPval-AA v2
生产力知识
1853.00Thinking Level · High | Tools
1741.00Thinking Level · High | Tools
CursorBench 3.2
编程与软件工程
73.40Thinking Level · High | Tools
70.50Thinking Level · High | Tools

Single-Benchmark Version Trend

Viewing: AA Intelligence Index · 综合评估

Benchmark
NormalNormal + ToolsThinkingThinking + ToolsDeepDeep + Tools

X-axis shows model and release date, Y-axis shows score; solid lines connect the same mode across versions, while dotted guides align modes within the same generation.

Standard API Pricing Across the Claude Fable 5.1 Series

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens

ModelSupplierStandard inputStandard outputBase price applies to
Claude Fable 5.1
Anthropic$10 / 1M tokens$50 / 1M tokens
Claude Fable 5
Anthropic$10 / 1M tokens$50 / 1M tokens