DataLearner logo

Claude Sonnet 5vsClaude Sonnet 4.6

Across 7 shared benchmarks, Claude Sonnet 5 leads overall: Claude Sonnet 5 wins 6, Claude Sonnet 4.6 wins 1, with 0 ties and an average score difference of +5.78.

Anthropic
Claude Sonnet 5

Anthropic · 2026-06-30 · Multimodal model

Anthropic
Claude Sonnet 4.6

Anthropic · 2026-02-17 · Chat model

Claude Sonnet 56 wins(86%)(14%)1 winClaude Sonnet 4.6

Benchmark scores

Grouped by capability, sorted by largest gap within each. 7 shared benchmarks.

Coding and Software Engineer

Claude Sonnet 5 2/2
BenchmarkClaude Sonnet 5Claude Sonnet 4.6Diff
DeepSWE5417 / 28Deep Thinking (With Tools)3026 / 28Thinking High (With Tools)+24
SWE-bench Verified85.207 / 114极高强度思考(工具)79.6018 / 114Thinking (No Tools)+5.60

AI Agent - Information Search

Claude Sonnet 5 1/1
BenchmarkClaude Sonnet 5Claude Sonnet 4.6Diff
BrowseComp84.707 / 54Thinking (With Tools + Internet)74.7027 / 54Thinking (With Tools)+10

AI Agent - Tool Usage

Claude Sonnet 5 1/1
BenchmarkClaude Sonnet 5Claude Sonnet 4.6Diff
OSWorld-Verified81.206 / 26极高强度思考(工具)72.5017 / 26Thinking (With Tools)+8.70

General Evaluation

Claude Sonnet 5 1/1
BenchmarkClaude Sonnet 5Claude Sonnet 4.6Diff
GPQA Diamond90.5333 / 224极高强度思考(无工具)89.9042 / 224Thinking (No Tools)+0.63

General Knowledge

Claude Sonnet 5 1/1
BenchmarkClaude Sonnet 5Claude Sonnet 4.6Diff
HLE57.409 / 181极高强度思考(工具)4934 / 181Thinking (With Tools)+8.40

Writing and Creative Capabilities

Claude Sonnet 4.6 1/1
BenchmarkClaude Sonnet 5Claude Sonnet 4.6Diff
Creative Writing1,78818 / 99Normal (No Tools)1,80515 / 99Normal (No Tools)-16.90

Specs

FieldClaude Sonnet 5Claude Sonnet 4.6
PublisherAnthropicAnthropic
Release date2026-06-302026-02-17
Model typeMultimodal modelChat model
ArchitectureDenseDense
ParametersNot availableNot available
Context length1M1M
Max output128K8K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemClaude Sonnet 5Claude Sonnet 4.6
Text input$2 / 1M tokens$3 / 1M tokens
Text output$10 / 1M tokens$15 / 1M tokens
Cache read$0.3 / 1M tokens$0.3 / 1M tokens
Cache write$3.75 / 1M tokens$3.75 / 1M tokens

Summary

  • Claude Sonnet 5leads in:Coding and Software Engineer (2/2), AI Agent - Information Search (1/1), AI Agent - Tool Usage (1/1), General Evaluation (1/1), General Knowledge (1/1)
  • Claude Sonnet 4.6leads in:Writing and Creative Capabilities (1/1)

On average across the 7 shared benchmarks, Claude Sonnet 5 scores 5.78 higher.

Largest single-benchmark gap: DeepSWE — Claude Sonnet 5 54 vs Claude Sonnet 4.6 30 (+24).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.