DataLearner logo
CL

Claude Opus 4.6

Reasoning modelOpusClaude 4.6

Anthropic Claude Opus 4.6

Release date: 2026-02-05Updated: 2026-07-17 21:57:33.324Knowledge cutoff: 2025-055,355
Live demoGitHubHugging FaceCompare
Parameters
Not disclosed
Context length
1000K
Chinese support
Supported
Reasoning ability

Claude Opus 4.6 is a reasoning model from Anthropic, released on 2026-02-05. It accepts text and image input and returns text output. The recorded context window is 1000K, and the recorded maximum output is 64K. Cataloged capabilities include Reasoning model and Multilingual. The model weights are proprietary and are not published for download. Use the linked references to confirm current access, licensing, and provider-specific limits.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

Claude Opus 4.6

Model basics

Reasoning traces
Supported
Thinking modes
Thinking Level · Extended (Default)Standard ModeThinking Level · LowThinking Level · MediumThinking Level · High
Context length
1000K tokens
Max output length
64K tokens
Model type
Reasoning model
Modality (in / out)
Text, Image → Text
Release date
2026-02-05
Model file size
0B
MoE architecture
No
Total params / Active params
No data / N/A
Knowledge cutoff
2025-05
Claude Opus 4.6

Open source & experience

Code license
Proprietary
Weights license
Proprietary
GitHub repo
GitHub link unavailable
Hugging Face
Hugging Face link unavailable
Claude Opus 4.6

Official resources

Paper
DataLearnerAI blog
No blog post yet
Claude Opus 4.6

API details

API speed
3/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
TextContext > 200K$10.00/ 1M$37.50/ 1M
TextContext <= 200K$5.00/ 1M$25.00/ 1M
Batch
TypeConditionInputOutput
Text-$2.50/ 1M$12.50/ 1M
Turbo
TypeConditionInputOutput
TextContext <= 200K$30.00/ 1M$150.00/ 1M
TextContext > 200K$60.00/ 1M$225.00/ 1M
Cache PricingPrompt Cache
TypeTTLWriteRead
Text5m$6.25/ 1M
Context <= 200K
$0.500/ 1M
Context <= 200K
Text5m$12.50/ 1M
Context > 200K
$1.00/ 1M
Context > 200K
Text10m$10.00/ 1M
Context <= 200K
$0.500/ 1M
Context <= 200K
Text10m$20.00/ 1M
Context > 200K
$1.00/ 1M
Context > 200K
Claude Opus 4.6

Benchmark Results

Claude Opus 4.6 currently shows benchmark results led by τ²-Bench (1 / 43, score 91.89), IF Bench (1 / 31, score 94), HumanEval (2 / 39, score 95). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage
Internet

General Knowledge

9 evaluations
Benchmark / mode
Score
Rank/total
86
23 / 68
ARC-AGI
Extended
92
13 / 68
GPQA Diamond
Extended
91.31
15 / 188
MMLU
Extended
91.05
7 / 66
76.33
8 / 115
64.60
18 / 62
ARC-AGI-2
Extended
66.30
17 / 62
HLE
ExtendedToolsInternet
53
18 / 173
0
4 / 9

Coding and Software Engineer

5 evaluations
Benchmark / mode
Score
Rank/total
HumanEval
Extended
95
2 / 39
SWE-bench Verified
ExtendedTools
80.84
10 / 113
SWE-bench
ExtendedTools
77.83
1 / 2
76
38 / 123
72
14 / 23

Common Sense

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleQA
Extended
72
7 / 47

Math and Reasoning

7 evaluations
Benchmark / mode
Score
Rank/total
AIME2025
Extended
99.79
7 / 107
MATH-500
Extended
97.60
10 / 44
40.70
7 / 60
20.80
14 / 80
20.80
14 / 80
14.60
23 / 80
22.90
12 / 80

Multimodal Understanding

2 evaluations
Benchmark / mode
Score
Rank/total
MMMU
Extended
73.90
19 / 29
MMMU
ExtendedTools
77.30
16 / 29

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
Simple Bench
Standard Mode
67.60
8 / 63

Agent Level Benchmark

2 evaluations
Benchmark / mode
Score
Rank/total
99.25
2 / 35
τ²-Bench
ExtendedTools
91.89
1 / 43

Instruction Following

1 evaluations
Benchmark / mode
Score
Rank/total
IF Bench
Extended
94
1 / 31

AI Agent - Information Search

1 evaluations
Benchmark / mode
Score
Rank/total
BrowseComp
Thinking ModeToolsInternet
84
11 / 53

AI Agent - Tool Usage

3 evaluations
Benchmark / mode
Score
Rank/total
MCP-Atlas
DeepTools
76.80
10 / 28
OSWorld-Verified
ExtendedTools
72.70
15 / 25
Terminal Bench 2.0
ExtendedTools
65.40
11 / 47

Productivity Knowledge

1 evaluations
Benchmark / mode
Score
Rank/total
GDPval-AA
ExtendedToolsInternet
1606
3 / 21

Claw-style Agent Evaluation

1 evaluations
Benchmark / mode
Score
Rank/total
Pinch Bench
Thinking ModeTools
87.40
7 / 37

Compare with other models

Claude Opus 4.6

Publisher

Anthropic Claude Opus 4.6

Model Overview

Claude Opus 4.6 is a reasoning model from Anthropic, released on 2026-02-05.

It accepts text and image input and produces text output. Its cataloged capabilities include Reasoning model and Multilingual. The recorded context window is 1000K, and the recorded maximum output is 64K.

The model weights are proprietary and are not published for download. The page records 18 API pricing rules from Anthropic; current provider pricing and conditions should be checked before deployment. The evaluation section contains 34 cataloged benchmark results with their recorded modes and scores. The page links 2 release, model-card, repository, or provider references for checking the underlying claims. Specifications, availability, and prices can change; undisclosed values are intentionally left unstated.

Claude Opus 4.6

FAQ

What is Claude Opus 4.6?

Claude Opus 4.6 is a reasoning model from Anthropic, released on 2026-02-05. It accepts text and image input and returns text output. The recorded context window is 1000K, and the recorded maximum output is 64K. Cataloged capabilities include Reasoning model and Multilingual. The model weights are proprietary and are not published for download. Use the linked references to confirm current access, licensing, and provider-specific limits.

What input and output modalities does Claude Opus 4.6 support?

The current model record lists text and image as input and text as output.

What are the main recorded specifications for Claude Opus 4.6?

The recorded context window is 1000K, and the recorded maximum output is 64K. Fields without a source-backed value remain undisclosed.

Does Claude Opus 4.6 have API pricing?

The page records 18 API pricing rules from Anthropic; current provider pricing and conditions should be checked before deployment.

Are benchmark results available for Claude Opus 4.6?

The evaluation section contains 34 cataloged benchmark results with their recorded modes and scores. Compare only results that use the same benchmark version and evaluation mode.

Is Claude Opus 4.6 open source?

The model weights are proprietary and are not published for download. Review the linked license text before commercial or derivative use.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code