DataLearner logo

Claude Fable 5vsGemini 3.1 Pro Preview

Across 11 shared benchmarks, Claude Fable 5 leads overall: Claude Fable 5 wins 10, Gemini 3.1 Pro Preview wins 1, with 0 ties and an average score difference of +84.24.

Anthropic
Claude Fable 5

Anthropic · 2026-06-09 · Reasoning model

Google Deep Mind
Gemini 3.1 Pro Preview

Google Deep Mind · 2026-02-20 · Multimodal model

Claude Fable 510 wins(91%)(9%)1 winGemini 3.1 Pro Preview

Benchmark scores

Grouped by capability, sorted by largest gap within each. 11 shared benchmarks.

Coding and Software Engineer

Claude Fable 5 4/4
BenchmarkClaude Fable 5Gemini 3.1 Pro PreviewDiff
DeepSWE702 / 27Deep Thinking (With Tools)1227 / 27Thinking High (With Tools)+58
SWE-Bench Pro - Public80.301 / 57Deep Thinking (With Tools)54.2034 / 57Thinking High (With Tools)+26.10
WeirdML v287.853 / 52Thinking High (With Tools)72.1017 / 52Normal (With Tools)+15.75
SWE-bench Verified952 / 114Deep Thinking (With Tools)80.6011 / 114Thinking High (With Tools)+14.40

AI Agent - Tool Usage

Claude Fable 5 3/3
BenchmarkClaude Fable 5Gemini 3.1 Pro PreviewDiff
Terminal-Bench 2.1884 / 44Deep Thinking (With Tools)73.8028 / 44Thinking High (With Tools)+14.20
OSWorld-Verified851 / 26Thinking High (With Tools)76.2012 / 26Thinking (With Tools)+8.80
MCP-Atlas83.305 / 38Normal (With Tools)78.2012 / 38Thinking High (With Tools)+5.10

General Knowledge

Even 2/2
BenchmarkClaude Fable 5Gemini 3.1 Pro PreviewDiff
HLE595 / 181Deep Thinking (No Tools)51.4024 / 181Thinking High (With Tools)+7.60
LiveBench78.315 / 115Deep Thinking (No Tools)79.933 / 115Thinking High (No Tools)-1.62

Commonsense Reasoning

Claude Fable 5 1/1
BenchmarkClaude Fable 5Gemini 3.1 Pro PreviewDiff
SimpleBench81.901 / 67Normal (No Tools)79.602 / 67Normal (No Tools)+2.30

Productivity Knowledge

Claude Fable 5 1/1
BenchmarkClaude Fable 5Gemini 3.1 Pro PreviewDiff
GDPval-AA v21,7414 / 13Max (With Tools)96512 / 13Thinking (No Tools)+776

Specs

FieldClaude Fable 5Gemini 3.1 Pro Preview
PublisherAnthropicGoogle Deep Mind
Release date2026-06-092026-02-20
Model typeReasoning modelMultimodal model
ArchitectureDenseDense
ParametersNot availableNot available
Context length1M1M
Max output128K64K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemClaude Fable 5Gemini 3.1 Pro Preview
Text input$10 / 1M tokens$2 / 1M tokens
Text output$50 / 1M tokens$12 / 1M tokens
Cache read$1 / 1M tokensNot public
Cache write$12.5 / 1M tokensNot public

Summary

  • Claude Fable 5leads in:Coding and Software Engineer (4/4), AI Agent - Tool Usage (3/3), Commonsense Reasoning (1/1), Productivity Knowledge (1/1)
  • Tied in:General Knowledge

On average across the 11 shared benchmarks, Claude Fable 5 scores 84.24 higher.

Largest single-benchmark gap: GDPval-AA v2 — Claude Fable 5 1,741 vs Gemini 3.1 Pro Preview 965 (+776).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.