DataLearner logo
GP

GPT-5.6 Sol

Reasoning modelCoding modelGPT-5.6

GPT-5.6 Sol

Release date: 2026-06-26Updated: 2026-09-05Knowledge cutoff: 2026-02-16Views: 5,758
Live demoGitHubHugging FaceCompare
Parameters
No data
Context length
1.05M
Multilingual
Supported
Reasoning ability
5/5

GPT-5.6 Sol is a reasoning model from OpenAI, released on 2026-06-26. It accepts text and image input and returns text output. The recorded context window is 1.05M, and the recorded maximum output is 128K. Cataloged capabilities include Reasoning model, Multilingual, and Coding model. The model weights are proprietary and are not published for download. Use the linked references to confirm current access, licensing, and provider-specific limits.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

GPT-5.6 Sol

Model basics

Reasoning traces
Supported
Thinking modes
Thinking Level · Medium (Default)Standard ModeThinking Level · LowThinking Level · HighThinking Level · Extra-HighThinking Level · Max
Context length
1.05M tokens
Max output length
128K tokens
Model type
Reasoning model
Modality (in / out)
Text, Image → Text
Release date
2026-06-26
Model file size
No data
MoE architecture
No
Total params / Active params
No data / Not applicable
Knowledge cutoff
2026-02-16
GPT-5.6 Sol

Open source & experience

Code license
Proprietary
Weights license
Proprietary
GitHub repo
N/A
Hugging Face
N/A
Live demo
N/A
GPT-5.6 Sol

Official resources

GPT-5.6 Sol

API details

API speed
3/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
TextContext <= 272000$4.00/ 1M tokens$20.00/ 1M tokens
long-context
TypeConditionInputOutput
TextContext > 272000$8.00/ 1M tokens$30.00/ 1M tokens
Cache PricingPrompt Cache
TypeTTLWriteRead
Text30m$5.00/ 1M tokens
Context <= 272000
$0.400/ 1M tokens
Context <= 272000

“—” means the modality is not billed in that direction, or the vendor has not published a price for it.

GPT-5.6 Sol

Benchmark Results

GPT-5.6 Sol currently shows benchmark results led by Terminal Bench Hard (1 / 244, score 65.90), CritPt (1 / 201, score 32.30), GPQA Diamond (6 / 463, score 94.60). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage
Internet

General Knowledge

38 evaluations
Benchmark / mode
Score
Rank/total
74.50
82 / 147
ARC-AGI-1
Medium
92.50
35 / 147
97
14 / 147
96.50
15 / 147
ARC-AGI-1
Extra-High
97.50
7 / 147
42.50
70 / 136
ARC-AGI-2
Medium
67.08
43 / 136
85.42
17 / 136
92.50
3 / 136
ARC-AGI-2
Extra-High
90
7 / 136
Vals Index
Extra-High
72.63
1 / 2
54.50
2 / 6
HLE
Standard Mode
16.70
325 / 565
HLE
Low
39.40
138 / 565
HLE
Medium
42.20
115 / 565
HLE
High
46
78 / 565
HLE
Max
49.50
56 / 565
HLE
Max
44.50
89 / 565
HLE
Extra-High
47.30
69 / 565
47.10
6 / 26
CritPt
Standard Mode
5.10
95 / 201
14.90
57 / 201
CritPt
Medium
22.90
31 / 201
CritPt
High
25.70
24 / 201
32.30
1 / 201
CritPt
Extra-High
28.60
14 / 201

General Evaluation

11 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Standard Mode
82.83
195 / 463
94.60
6 / 463
89.90
89 / 463
92.60
44 / 463
92.80
42 / 463
94.10
13 / 463
GPQA Diamond
Extra-High
93.10
38 / 463
59.90
2 / 2
47.40
2 / 2
32.30
2 / 2

Writing and Creative Capabilities

2 evaluations
Benchmark / mode
Score
Rank/total
Creative Writing
Standard Mode
1963.40
6 / 106

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleBench
Extra-High
64.80
21 / 93

Coding and Software Engineer

38 evaluations
Benchmark / mode
Score
Rank/total
1620.27
4 / 35
WeirdML v2
HighTools
88.76
2 / 52
79.10
2 / 6
74.30
3 / 6
DeepSWE
LowTools
45.35
66 / 86
DeepSWE
MediumTools
61.06
42 / 86
DeepSWE
HighTools
69.40
19 / 86
DeepSWE
MaxTools
72.67
13 / 86
DeepSWE
Extra-HighTools
72.70
12 / 86
SRE-Bench
unknown
68.70
3 / 4
SRE-Bench
unknown
55.90
4 / 4
67.20
4 / 5
SWE-Bench Pro - Public
Extra-HighTools
64.60
9 / 62
60.60
5 / 7
56.40
25 / 131
SciCode
Medium
57.40
16 / 131
57.80
15 / 131
57.10
19 / 131
SciCode
Extra-High
57.10
19 / 131
56.80
9 / 16
54.60
4 / 11
53.50
3 / 6
47.50
5 / 5
24.60
39 / 43
CursorBench 4.0
MediumTools
31.10
30 / 43
35.70
22 / 43
41.70
10 / 43
CursorBench 4.0
Extra-HighTools
37.70
17 / 43
23
8 / 12
WeirdML v3
Extra-HighTools
0.16
3 / 5

Agent Level Benchmark

20 evaluations
Benchmark / mode
Score
Rank/total
76
120 / 264
81
109 / 264
83.30
102 / 264
85.10
86 / 264
τ²-Bench - Telecom
Extra-HighTools
84.80
90 / 264
60.60
7 / 244
62.90
2 / 244
62.10
5 / 244
65.90
1 / 244
Terminal Bench Hard
Extra-HighTools
61.40
6 / 244
APEX-Agents
MaxTools
56.70
3 / 7
53.60
2 / 20
26.70
13 / 20
Agents' Last Exam
Extra-HighTools
52.70
3 / 20
τ³-Banking
Standard ModeTools
19.60
95 / 167
τ³-Banking
LowTools
29.10
65 / 167
τ³-Banking
MediumTools
36.50
45 / 167
τ³-Banking
HighTools
36.70
43 / 167
τ³-Banking
MaxTools
44.30
20 / 167
τ³-Banking
Extra-HighTools
46.91
13 / 167

Instruction Following

5 evaluations
Benchmark / mode
Score
Rank/total
66.50
93 / 282
IF Bench
Medium
69.60
77 / 282
69.20
78 / 282
72.70
49 / 282
IF Bench
Extra-High
71
62 / 282

AI Agent - Information Search

1 evaluations
Benchmark / mode
Score
Rank/total
BrowseComp
unknown
90.40
4 / 58

Text Embedding

1 evaluations
Benchmark / mode
Score
Rank/total
97.63
2 / 126

Long Context

6 evaluations
Benchmark / mode
Score
Rank/total
AA-LCR
Standard Mode
62.30
126 / 171
78
67 / 171
AA-LCR
Medium
80.30
40 / 171
AA-LCR
High
81.70
29 / 171
84
7 / 171
AA-LCR
Extra-High
82.30
21 / 171

AI Agent - Tool Usage

25 evaluations
Benchmark / mode
Score
Rank/total
92.91
2 / 8
Terminal-Bench 2.1
Standard ModeTools
74.20
82 / 194
76.80
73 / 194
86.10
26 / 194
87.30
21 / 194
88.80
9 / 194
88
11 / 194
Terminal-Bench 2.1
Extra-HighTools
89.50
6 / 194
Vals CyberBench
Extra-HighTools
88.14
1 / 1
CyberGym
MaxTools
84.50
3 / 9
MCP-Atlas
MaxTools
81.80
13 / 44
78.50
2 / 3
65.70
5 / 10
OSWorld 2.0
Extra-HighTools
62.60
6 / 10
60.10
6 / 14
1
76 / 88
14.60
37 / 88
20.70
28 / 88
37.27
16 / 88
Terminal-Bench 4.0
Extra-HighTools
24.70
25 / 88
34.40
2 / 11
30.30
2 / 2
22.40
4 / 11

Productivity Knowledge

19 evaluations
Benchmark / mode
Score
Rank/total
GDPval-AA v2
Standard ModeTools
1299
64 / 106
GDPval-AA v2
LowTools
1355
56 / 106
GDPval-AA v2
MediumTools
1456
38 / 106
GDPval-AA v2
HighTools
1524
26 / 106
GDPval-AA v2
MaxTools
1624
17 / 106
GDPval-AA v2
Extra-HighTools
1585
18 / 106
AA-Briefcase
Standard ModeTools
1013
62 / 84
AA-Briefcase
LowTools
1042
59 / 84
AA-Briefcase
MediumTools
1240
42 / 84
AA-Briefcase
HighTools
1361
28 / 84
AA-Briefcase
MaxTools
1475
20 / 84
AA-Briefcase
Extra-HighTools
1435
23 / 84
87.18
17 / 43
BenchCAD
unknown
83.30
2 / 2
AA-AnalystAgent
MaxToolsInternet
47.50
7 / 29
47.40
2 / 2
18.10
18 / 18
45.80
8 / 18

Multimodal Understanding

15 evaluations
Benchmark / mode
Score
Rank/total
BabyVision
MaxTools
88.90
4 / 9
MMMU-Pro
Standard Mode
71.90
118 / 229
81
35 / 229
MMMU-Pro
Medium
81.40
31 / 229
81.80
27 / 229
83.40
16 / 229
MMMU-Pro
Extra-High
82.70
19 / 229
Chartography
MaxTools
79.90
3 / 8
76.90
2 / 2
53
1 / 6
21
31 / 119
GDP.pdf
Medium
26.20
14 / 119
27.80
8 / 119
27.20
10 / 119
GDP.pdf
Extra-High
27.60
9 / 119

Math and Reasoning

3 evaluations
Benchmark / mode
Score
Rank/total
89.12
1 / 58
83
3 / 42

Truthfulness Evaluation

6 evaluations
Benchmark / mode
Score
Rank/total
48.20
2 / 2
22
2 / 2
22
15 / 44
12.20
2 / 2
0.29
2 / 2

Long Context

2 evaluations
Benchmark / mode
Score
Rank/total

Agent Capability

1 evaluations
Benchmark / mode
Score
Rank/total

Claw-style Agent Evaluation

1 evaluations
Benchmark / mode
Score
Rank/total
84.19
7 / 45

Compare with other models

GPT-5.6 Sol

Publisher

GPT-5.6 Sol

Model Overview

GPT-5.6 Sol is a reasoning model from OpenAI, released on 2026-06-26.

It accepts text and image input and produces text output. Its cataloged capabilities include Reasoning model, Multilingual, and Coding model. The recorded context window is 1.05M, and the recorded maximum output is 128K.

The model weights are proprietary and are not published for download. The page records 6 API pricing rules from OpenAI; current provider pricing and conditions should be checked before deployment. The evaluation section contains 9 cataloged benchmark results with their recorded modes and scores. The page links 1 release, model-card, repository, or provider references for checking the underlying claims. Specifications, availability, and prices can change; undisclosed values are intentionally left unstated.

GPT-5.6 Sol

FAQ

What is GPT-5.6 Sol?

GPT-5.6 Sol is a reasoning model from OpenAI, released on 2026-06-26. It accepts text and image input and returns text output. The recorded context window is 1.05M, and the recorded maximum output is 128K. Cataloged capabilities include Reasoning model, Multilingual, and Coding model. The model weights are proprietary and are not published for download. Use the linked references to confirm current access, licensing, and provider-specific limits.

What input and output modalities does GPT-5.6 Sol support?

The current model record lists text and image as input and text as output.

What are the main recorded specifications for GPT-5.6 Sol?

The recorded context window is 1.05M, and the recorded maximum output is 128K. Fields without a source-backed value remain undisclosed.

Does GPT-5.6 Sol have API pricing?

The page records 6 API pricing rules from OpenAI; current provider pricing and conditions should be checked before deployment.

Are benchmark results available for GPT-5.6 Sol?

The evaluation section contains 9 cataloged benchmark results with their recorded modes and scores. Compare only results that use the same benchmark version and evaluation mode.

Is GPT-5.6 Sol open source?

The model weights are proprietary and are not published for download. Review the linked license text before commercial or derivative use.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code