DataLearner logo
CL

Claude Sonnet 4

Reasoning modelSonnetClaude 4

Claude Sonnet 4

Release date: 2025-05-23Updated: 2025-10-19 12:24:142,128
Live demoGitHubHugging FaceCompare
Parameters
Not disclosed
Context length
200K
Chinese support
Supported
Reasoning ability

Claude Sonnet 4 is an AI model published by Anthropic, released on 2025-05-23, for Reasoning model, and 200K context length, with a 1223.00 score on CodeClash.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

Claude Sonnet 4

Model basics

Reasoning traces
Supported
Thinking modes
Thinking modes not supported
Context length
200K tokens
Max output length
64K tokens
Model type
Reasoning model
Modality (in / out)
Text, Image → Text
Release date
2025-05-23
Model file size
No data
MoE architecture
No
Total params / Active params
No data / N/A
Knowledge cutoff
No data
Claude Sonnet 4

Open source & experience

Code license
Proprietary
Weights license
Proprietary
GitHub repo
GitHub link unavailable
Hugging Face
Hugging Face link unavailable
Live demo
No live demo
Claude Sonnet 4

Official resources

Paper
DataLearnerAI blog
Claude Sonnet 4

API details

API speed
4/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-$3.00/ 1M$15.00/ 1M
TextContext > 200000$6.00/ 1M$22.50/ 1M
Batch
TypeConditionInputOutput
Text-$1.50/ 1M$7.50/ 1M
Cache PricingPrompt Cache
TypeTTLWriteRead
Text-$3.75/ 1M$0.300/ 1M
Claude Sonnet 4

Benchmark Results

Claude Sonnet 4 currently shows benchmark results led by SWE-bench Verified (14 / 113, score 80.20), Terminal-Bench (10 / 35, score 41.30), MMLU Pro (38 / 132, score 84). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage

General Knowledge

12 evaluations
Benchmark / mode
Score
Rank/total
84
38 / 132
83.80
62 / 188
75.40
98 / 188
68
129 / 188
LiveBench
Standard Mode
50.98
89 / 115
61.27
65 / 115
40
49 / 68
23.80
56 / 68
9.60
150 / 173
5.52
164 / 173
5.90
46 / 62
1.30
55 / 62

Coding and Software Engineer

6 evaluations
Benchmark / mode
Score
Rank/total
CodeClash
Standard ModeTools
1223
4 / 8
80.20
14 / 113
72.70
52 / 113
66
59 / 123
48.50
96 / 123

Math and Reasoning

12 evaluations
Benchmark / mode
Score
Rank/total
85
51 / 107
70.50
72 / 107
38
96 / 107
43.40
50 / 62
27.10
8 / 16
9.70
5 / 10
5.20
8 / 10
4.10
41 / 60
4
5 / 9
3.30
6 / 9
0
72 / 80

Writing and Creative Capabilities

1 evaluations
Benchmark / mode
Score
Rank/total
83.05
14 / 23

AI Agent - Tool Usage

4 evaluations
Benchmark / mode
Score
Rank/total
42.20
23 / 25
41.30
10 / 35
35.50
18 / 35

Multimodal Understanding

1 evaluations
Benchmark / mode
Score
Rank/total
76.50
17 / 29

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
Simple Bench
Thinking Mode
45.50
34 / 63

Agent Level Benchmark

4 evaluations
Benchmark / mode
Score
Rank/total
Aider-Polyglot
Standard Mode
56.40
26 / 59
61.30
20 / 59
52
34 / 43

Instruction Following

1 evaluations
Benchmark / mode
Score
Rank/total
55
24 / 31

Productivity Knowledge

1 evaluations
Benchmark / mode
Score
Rank/total
33
19 / 21

Long Context

1 evaluations
Benchmark / mode
Score
Rank/total
65
13 / 16

Claw-style Agent Evaluation

2 evaluations
Benchmark / mode
Score
Rank/total
Pinch Bench
Thinking ModeTools
80.50
22 / 37
Claw Bench
Thinking ModeTools
77.80
23 / 29

Compare with other models

No curated comparisons for this model yet.

Want a custom combination? Open the compare tool

Claude Sonnet 4

Publisher

Claude Sonnet 4

Model Overview

Claude Sonnet 4 is an AI model published by Anthropic, released on 2025-05-23, for Reasoning model, and 200K context length, with a 1223.00 score on CodeClash.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code