DataLearner logo

Grok 4.5vsClaude Fable 5

Across 14 shared benchmarks, Claude Fable 5 leads overall: Grok 4.5 wins 1, Claude Fable 5 wins 13, with 0 ties and an average score difference of -46.96.

xAI
Grok 4.5

xAI · 2026-07-08 · Coding model

Anthropic
Claude Fable 5

Anthropic · 2026-06-09 · Reasoning model

Grok 4.51 win(7%)(93%)13 winsClaude Fable 5

Benchmark scores

Grouped by capability, sorted by largest gap within each. 14 shared benchmarks.

Coding and Software Engineer

Claude Fable 5 5/5
BenchmarkGrok 4.5Claude Fable 5Diff
DeepSWE5318 / 27Thinking High (With Tools)702 / 27Deep Thinking (With Tools)-17
SWE-Bench Pro - Public64.706 / 57Thinking High (With Tools)80.301 / 57Deep Thinking (With Tools)-15.60
FrontierCode 1.156.604 / 5Thinking High (With Tools)63.601 / 5Max (With Tools)-7
APEX-SWE53.603 / 3Thinking High (With Tools)58.801 / 3Max (With Tools)-5.20
CursorBench 3.266.704 / 4Thinking High (With Tools)70.501 / 4Max (With Tools)-3.80

Productivity Knowledge

Claude Fable 5 2/3
BenchmarkGrok 4.5Claude Fable 5Diff
AA-Briefcase1,3136 / 6Thinking High (With Tools)1,5743 / 6Max (With Tools)-261
GDPval-AA v21,5267 / 13Thinking High (With Tools)1,7414 / 13Max (With Tools)-215
Harvey Lab-AA12.904 / 6Thinking High (With Tools)11.305 / 6Max (With Tools)+1.60

AI Agent - Tool Usage

Claude Fable 5 2/2
BenchmarkGrok 4.5Claude Fable 5Diff
Terminal-Bench 3.015.705 / 6Thinking High (With Tools)34.102 / 6Max (With Tools)-18.40
Terminal-Bench 2.183.3014 / 44Thinking High (With Tools)884 / 44Deep Thinking (With Tools)-4.70

Math and Reasoning

Claude Fable 5 2/2
BenchmarkGrok 4.5Claude Fable 5Diff
FrontierMath Tier 4 v224.3920 / 34Thinking High (No Tools)87.801 / 34最高(无工具)-63.41
FrontierMath v257.1921 / 34Thinking High (No Tools)87.023 / 34最高(无工具)-29.82

Agent Level Benchmark

Claude Fable 5 1/1
BenchmarkGrok 4.5Claude Fable 5Diff
APEX-Agents47.104 / 5Thinking High (With Tools)59.201 / 5Max (With Tools)-12.10

General Knowledge

Claude Fable 5 1/1
BenchmarkGrok 4.5Claude Fable 5Diff
AA Intelligence Index564 / 8Thinking High (With Tools)621 / 8Max (With Tools)-6

Specs

FieldGrok 4.5Claude Fable 5
PublisherxAIAnthropic
Release date2026-07-082026-06-09
Model typeCoding modelReasoning model
ArchitectureDenseDense
ParametersNot availableNot available
Context length500K1M
Max outputNot available128K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemGrok 4.5Claude Fable 5
Text input$2 / 1M tokens$10 / 1M tokens
Text output$6 / 1M tokens$50 / 1M tokens
Cache read$0.5 / 1M tokens$1 / 1M tokens
Cache writeNot public$12.5 / 1M tokens

Summary

  • Claude Fable 5leads in:Coding and Software Engineer (5/5), Productivity Knowledge (2/3), AI Agent - Tool Usage (2/2), Math and Reasoning (2/2), Agent Level Benchmark (1/1), General Knowledge (1/1)

On average across the 14 shared benchmarks, Claude Fable 5 scores 46.96 higher.

Largest single-benchmark gap: AA-Briefcase — Grok 4.5 1,313 vs Claude Fable 5 1,574 (-261).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.