DataLearner logo
GP

GPT-image-2

Multimodal modelImageGPT Image 2

GPT-image-2

Release date: 2026-04-21Updated: 2026-07-17 21:57:35.3932,290
Live demoGitHubHugging FaceCompare
Parameters
Not disclosed
Context length
32K
Chinese support
Supported
Reasoning ability

GPT-image-2 is a multimodal model from OpenAI, released on 2026-04-21. It accepts text and image input and returns image output. The recorded context window is 32K. Cataloged capabilities include Reasoning model and Multilingual. The model weights are proprietary and are not published for download. Use the linked references to confirm current access, licensing, and provider-specific limits.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

GPT-image-2

Model basics

Reasoning traces
Supported
Thinking modes
Standard ModeThinking Mode
Context length
32K tokens
Max output length
No data
Model type
Multimodal model
Modality (in / out)
Text, Image → Image
Release date
2026-04-21
Model file size
No data
MoE architecture
No
Total params / Active params
No data / N/A
Knowledge cutoff
No data
GPT-image-2

Open source & experience

Code license
Proprietary
Weights license
Proprietary
GitHub repo
GitHub link unavailable
Hugging Face
Hugging Face link unavailable
Live demo
No live demo
GPT-image-2

Official resources

Paper
DataLearnerAI blog
No blog post yet
GPT-image-2

API details

API speed
4/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-$5.00/ 1M$10.00/ 1M
Image-$8.00/ 1M$30.00/ 1M
Cache PricingPrompt Cache
TypeTTLWriteRead
Text--$1.25/ 1M
Image--$2.00/ 1M
GPT-image-2

Benchmark Results

No benchmark data to show.

Compare with other models

No curated comparisons for this model yet.

Want a custom combination? Open the compare tool

GPT-image-2

Publisher

GPT-image-2

Model Overview

GPT-image-2 is a multimodal model from OpenAI, released on 2026-04-21.

It accepts text and image input and produces image output. Its cataloged capabilities include Reasoning model and Multilingual. The recorded context window is 32K.

The model weights are proprietary and are not published for download. The page records 6 API pricing rules from the listed provider; current provider pricing and conditions should be checked before deployment. The page links 1 release, model-card, repository, or provider references for checking the underlying claims. Specifications, availability, and prices can change; undisclosed values are intentionally left unstated.

GPT-image-2

FAQ

What is GPT-image-2?

GPT-image-2 is a multimodal model from OpenAI, released on 2026-04-21. It accepts text and image input and returns image output. The recorded context window is 32K. Cataloged capabilities include Reasoning model and Multilingual. The model weights are proprietary and are not published for download. Use the linked references to confirm current access, licensing, and provider-specific limits.

What input and output modalities does GPT-image-2 support?

The current model record lists text and image as input and image as output.

What are the main recorded specifications for GPT-image-2?

The recorded context window is 32K. Fields without a source-backed value remain undisclosed.

Does GPT-image-2 have API pricing?

The page records 6 API pricing rules from the listed provider; current provider pricing and conditions should be checked before deployment.

Why are no benchmark results shown for GPT-image-2?

Image-generation model; the current text/agent benchmark catalog is not applicable.

Is GPT-image-2 open source?

The model weights are proprietary and are not published for download. Review the linked license text before commercial or derivative use.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code