Claude Fable 5.1 Benchmark Details
Claude Fable 5.1 currently shows benchmark results led by HLE (1 / 188, score 65), AA Intelligence Index (1 / 27, score 66), GDPval-AA v2 (2 / 23, score 1853). This page also compares it with 3 competitor models and 1 predecessor or same-series models, including performance and pricing views when available.
Benchmark Results
Benchmark Results
General Knowledge
7 evaluationsAI Agent - Tool Usage
4 evaluationsCompetitor Comparison
Benchmark scores for Claude Fable 5.1 compared against top models in its class
7 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | Claude Fable 5.1Current | GPT-5.6 Sol |
|---|---|---|
66.00Thinking Level · High | Tools | 61.00Thinking Level · High | Tools | |
HLE 综合评估 | 65.00Thinking Level · High | Tools | 49.50Thinking Level · High |
OSWorld 2.0 AI Agent - 工具使用 | 41.70Thinking Level · High | Tools | 62.60Thinking Level · Extra High | Tools |
Terminal-Bench 4.0 AI Agent - 工具使用 | 55.80Thinking Level · High | Tools | 37.27Thinking Level · High | Tools |
Terminal-Bench-Science 0.1 AI Agent - 工具使用 | 52.60Thinking Level · High | Tools | 22.40Thinking Level · High | Tools |
GDPval-AA v2 生产力知识 | 1853.00Thinking Level · High | Tools | 1728.00Thinking Level · High | Tools |
CursorBench 3.2 编程与软件工程 | 73.40Thinking Level · High | Tools | 67.20Thinking Level · High | Tools |
Standard API Pricing: Claude Fable 5.1 vs. Peer Models
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
Claude Fable 5.1 | Anthropic | $10 / 1M tokens | $50 / 1M tokens | — |
GPT-5.6 Sol | OpenAI | $4 / 1M tokens | $20 / 1M tokens | — |
Version History
How each version of the Claude Fable 5.1 series stacks up on benchmark tests
6 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.
| Benchmark | Claude Fable 5.1Current | Claude Fable 5 |
|---|---|---|
66.00Thinking Level · High | Tools | 62.00Thinking Level · High | Tools | |
HLE 综合评估 | 65.00Thinking Level · High | Tools | 59.00Deep Thinking Mode |
Terminal-Bench 4.0 AI Agent - 工具使用 | 55.80Thinking Level · High | Tools | 44.55Thinking Level · High | Tools |
Terminal-Bench-Science 0.1 AI Agent - 工具使用 | 52.60Thinking Level · High | Tools | 21.40Thinking Level · High | Tools |
GDPval-AA v2 生产力知识 | 1853.00Thinking Level · High | Tools | 1741.00Thinking Level · High | Tools |
CursorBench 3.2 编程与软件工程 | 73.40Thinking Level · High | Tools | 70.50Thinking Level · High | Tools |
Single-Benchmark Version Trend
Viewing: AA Intelligence Index · 综合评估
Standard API Pricing Across the Claude Fable 5.1 Series
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
Claude Fable 5.1 | Anthropic | $10 / 1M tokens | $50 / 1M tokens | — |
Claude Fable 5 | Anthropic | $10 / 1M tokens | $50 / 1M tokens | — |