DataLearner logo
GP

GPT-5.4

Multimodal modelGPTGPT-5.4

GPT-5.4

Release date: 2026-03-05Updated: 2026-06-15 07:18:17.137Knowledge cutoff: 2025-083,674
Live demoGitHubHugging FaceCompare
Parameters
Not disclosed
Context length
1M
Chinese support
Supported
Reasoning ability

GPT-5.4 is a multimodal model from OpenAI, released on 2026-03-05. It accepts text and image input and returns text output. The recorded context window is 1M, and the recorded maximum output is 125K. Cataloged capabilities include Reasoning model and Multilingual. The model weights are proprietary and are not published for download. Use the linked references to confirm current access, licensing, and provider-specific limits.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

GPT-5.4

Model basics

Reasoning traces
Supported
Thinking modes
Thinking Level · Extra-High (Default)Thinking Level · LowThinking Level · MediumThinking Level · High
Context length
1M tokens
Max output length
125K tokens
Model type
Multimodal model
Modality (in / out)
Text, Image → Text
Release date
2026-03-05
Model file size
No data
MoE architecture
No
Total params / Active params
No data / N/A
Knowledge cutoff
2025-08
GPT-5.4

Open source & experience

Code license
Proprietary
Weights license
Proprietary
GitHub repo
GitHub link unavailable
Hugging Face
Hugging Face link unavailable
GPT-5.4

Official resources

Paper
DataLearnerAI blog
No blog post yet
GPT-5.4

API details

API speed
3/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
TextContext <= 272K$2.50/ 1M$15.00/ 1M
TextContext > 272K$5.00/ 1M$22.50/ 1M
Cache PricingPrompt Cache
TypeTTLWriteRead
Text5m$0.250/ 1M-
GPT-5.4

Benchmark Results

GPT-5.4 currently shows benchmark results led by LiveBench (2 / 115, score 80.28), Pinch Bench (1 / 37, score 90.50), GPQA Diamond (11 / 188, score 92.80). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage

General Knowledge

14 evaluations
Benchmark / mode
Score
Rank/total
ARC-AGI
Standard Mode
93.70
9 / 68
68.20
31 / 68
ARC-AGI
Medium
86.20
21 / 68
ARC-AGI
Extra-High
93.70
9 / 68
GPQA Diamond
Extra-High
92.80
11 / 188
75.07
16 / 115
80.28
2 / 115
ARC-AGI-2
Standard Mode
77.10
9 / 62
29.20
33 / 62
ARC-AGI-2
Medium
55.40
22 / 62
ARC-AGI-2
Extra-High
74
12 / 62
HLE
Extra-High
39.80
64 / 173
HLE
Extra-HighTools
52.10
21 / 173
0
7 / 9

Math and Reasoning

2 evaluations
Benchmark / mode
Score
Rank/total
FrontierMath
Extra-High
47.60
5 / 60
27.10
11 / 80

Coding and Software Engineer

2 evaluations
Benchmark / mode
Score
Rank/total
57.70
17 / 55
DeepSWE
Extra-HighTools
52
13 / 20

Agent Level Benchmark

2 evaluations
Benchmark / mode
Score
Rank/total
τ²-Bench - Telecom
Standard ModeTools
64.30
30 / 35
τ²-Bench - Telecom
Extra-HighTools
98.90
3 / 35

AI Agent - Information Search

1 evaluations
Benchmark / mode
Score
Rank/total
BrowseComp
Extra-HighTools
82.70
15 / 53

AI Agent - Tool Usage

3 evaluations
Benchmark / mode
Score
Rank/total
Terminal Bench 2.0
Extra-HighTools
75.10
4 / 47
OSWorld-Verified
Extra-HighTools
75
12 / 25
MCP-Atlas
Extra-HighTools
70.60
15 / 28

Claw-style Agent Evaluation

2 evaluations
Benchmark / mode
Score
Rank/total
Claw Bench
Thinking ModeTools
92.70
3 / 29
Pinch Bench
Thinking ModeTools
90.50
1 / 37

Compare with other models

GPT-5.4

Publisher

GPT-5.4

Model Overview

GPT-5.4 is a multimodal model from OpenAI, released on 2026-03-05.

It accepts text and image input and produces text output. Its cataloged capabilities include Reasoning model and Multilingual. The recorded context window is 1M, and the recorded maximum output is 125K.

The model weights are proprietary and are not published for download. The page records 5 API pricing rules from OpenAI; current provider pricing and conditions should be checked before deployment. The evaluation section contains 26 cataloged benchmark results with their recorded modes and scores. The page links 2 release, model-card, repository, or provider references for checking the underlying claims. Specifications, availability, and prices can change; undisclosed values are intentionally left unstated.

GPT-5.4

FAQ

What is GPT-5.4?

GPT-5.4 is a multimodal model from OpenAI, released on 2026-03-05. It accepts text and image input and returns text output. The recorded context window is 1M, and the recorded maximum output is 125K. Cataloged capabilities include Reasoning model and Multilingual. The model weights are proprietary and are not published for download. Use the linked references to confirm current access, licensing, and provider-specific limits.

What input and output modalities does GPT-5.4 support?

The current model record lists text and image as input and text as output.

What are the main recorded specifications for GPT-5.4?

The recorded context window is 1M, and the recorded maximum output is 125K. Fields without a source-backed value remain undisclosed.

Does GPT-5.4 have API pricing?

The page records 5 API pricing rules from OpenAI; current provider pricing and conditions should be checked before deployment.

Are benchmark results available for GPT-5.4?

The evaluation section contains 26 cataloged benchmark results with their recorded modes and scores. Compare only results that use the same benchmark version and evaluation mode.

Is GPT-5.4 open source?

The model weights are proprietary and are not published for download. Review the linked license text before commercial or derivative use.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code