DataLearner logo

Claude Fable 5.1 Benchmark Details

Claude Fable 5.1 currently shows benchmark results led by HLE (1 / 188, score 65), AA Intelligence Index (1 / 27, score 66), GDPval-AA v2 (2 / 23, score 1853). This page also compares it with 3 competitor models and 1 predecessor or same-series models, including performance and pricing views when available.

Benchmark Results

Claude Fable 5.1

Benchmark Results

Thinking
Tool usage

General Knowledge

7 evaluations
Benchmark / mode
Score
Rank/total
AA Intelligence Index
Thinking Level · LowTools
58
13 / 27
AA Intelligence Index
Thinking Level · MediumTools
60
10 / 27
AA Intelligence Index
Thinking Level · HighTools
62
5 / 27
AA Intelligence Index
Thinking Level · MaxTools
66
1 / 27
AA Intelligence Index
Thinking Level · Extra HighTools
65
2 / 27
HLE
Thinking Level · Max
60.90
6 / 188
HLE
Thinking Level · MaxTools
65
1 / 188

AI Agent - Tool Usage

4 evaluations
Benchmark / mode
Score
Rank/total
Terminal-Bench 4.0
Thinking Level · MaxTools
55.80
1 / 11
Terminal-Bench-Science 0.1
Thinking Level · MaxTools
52.60
1 / 10
OSWorld 2.0
Thinking Level · MaxTools
41.70
6 / 6
AutomationBench
Thinking Level · MaxTools
31.40
6 / 12

Productivity Knowledge

1 evaluations
Benchmark / mode
Score
Rank/total
GDPval-AA v2
Thinking Level · MaxTools
1853
2 / 23

Coding and Software Engineer

1 evaluations
Benchmark / mode
Score
Rank/total
CursorBench 3.2
Thinking Level · MaxTools
73.40
1 / 5

Competitor Comparison

Benchmark scores for Claude Fable 5.1 compared against top models in its class

Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

7 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.

BenchmarkClaude Fable 5.1CurrentGPT-5.6 Sol
66.00Thinking Level · High | Tools
61.00Thinking Level · High | Tools
HLE
综合评估
65.00Thinking Level · High | Tools
49.50Thinking Level · High
OSWorld 2.0
AI Agent - 工具使用
41.70Thinking Level · High | Tools
62.60Thinking Level · Extra High | Tools
Terminal-Bench 4.0
AI Agent - 工具使用
55.80Thinking Level · High | Tools
37.27Thinking Level · High | Tools
Terminal-Bench-Science 0.1
AI Agent - 工具使用
52.60Thinking Level · High | Tools
22.40Thinking Level · High | Tools
GDPval-AA v2
生产力知识
1853.00Thinking Level · High | Tools
1728.00Thinking Level · High | Tools
CursorBench 3.2
编程与软件工程
73.40Thinking Level · High | Tools
67.20Thinking Level · High | Tools

Standard API Pricing: Claude Fable 5.1 vs. Peer Models

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens

ModelSupplierStandard inputStandard outputBase price applies to
Claude Fable 5.1
Anthropic$10 / 1M tokens$50 / 1M tokens
GPT-5.6 Sol
OpenAI$4 / 1M tokens$20 / 1M tokens

Version History

How each version of the Claude Fable 5.1 series stacks up on benchmark tests

Claude Fable 5.1Claude Fable 5
Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

6 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.

BenchmarkClaude Fable 5.1CurrentClaude Fable 5
66.00Thinking Level · High | Tools
62.00Thinking Level · High | Tools
HLE
综合评估
65.00Thinking Level · High | Tools
59.00Deep Thinking Mode
Terminal-Bench 4.0
AI Agent - 工具使用
55.80Thinking Level · High | Tools
44.55Thinking Level · High | Tools
Terminal-Bench-Science 0.1
AI Agent - 工具使用
52.60Thinking Level · High | Tools
21.40Thinking Level · High | Tools
GDPval-AA v2
生产力知识
1853.00Thinking Level · High | Tools
1741.00Thinking Level · High | Tools
CursorBench 3.2
编程与软件工程
73.40Thinking Level · High | Tools
70.50Thinking Level · High | Tools

Single-Benchmark Version Trend

Viewing: AA Intelligence Index · 综合评估

Benchmark
NormalNormal + ToolsThinkingThinking + ToolsDeepDeep + Tools

X-axis shows model and release date, Y-axis shows score; solid lines connect the same mode across versions, while dotted guides align modes within the same generation.

Standard API Pricing Across the Claude Fable 5.1 Series

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens

ModelSupplierStandard inputStandard outputBase price applies to
Claude Fable 5.1
Anthropic$10 / 1M tokens$50 / 1M tokens
Claude Fable 5
Anthropic$10 / 1M tokens$50 / 1M tokens