DataLearner logo
GP

GPT-5.6 Luna

Reasoning modelTool useGPT-5.6

GPT-5.6 Luna

Release date: 2026-06-26Updated: 2026-08-23Knowledge cutoff: 2026-02-16Views: 875
Live demoGitHubHugging FaceCompare
Parameters
No data
Context length
1.05M
Multilingual
Supported
Reasoning ability
3/5

GPT-5.6 Luna is an AI model from OpenAI, released on 2026-06-26.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

GPT-5.6 Luna

Model basics

Reasoning traces
Supported
Thinking modes
Thinking Level · Medium (Default)Standard ModeThinking Level · LowThinking Level · High
Context length
1.05M tokens
Max output length
128K tokens
Model type
Reasoning model
Modality (in / out)
Text, Image → Text
Release date
2026-06-26
Model file size
No data
MoE architecture
No
Total params / Active params
No data / Not applicable
Knowledge cutoff
2026-02-16
GPT-5.6 Luna

Open source & experience

Code license
Proprietary
Weights license
Proprietary
GitHub repo
N/A
Hugging Face
N/A
Live demo
N/A
GPT-5.6 Luna

Official resources

GPT-5.6 Luna

API details

API speed
5/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-$0.200/ 1M$1.20/ 1M
long-context
TypeConditionInputOutput
Text-$0.400/ 1M$1.80/ 1M
Cache PricingPrompt Cache
TypeTTLWriteRead
Text30m$0.250/ 1M$0.020/ 1M

“—” means the modality is not billed in that direction, or the vendor has not published a price for it.

GPT-5.6 Luna

Benchmark Results

GPT-5.6 Luna currently shows benchmark results led by GPQA Diamond (34 / 274, score 91.60), FrontierMath v2 (8 / 58, score 82.11), Creative Writing (17 / 106, score 1825.80). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage

General Knowledge

4 evaluations
Benchmark / mode
Score
Rank/total

General Evaluation

3 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Standard Mode
63.64
222 / 274
82.32
129 / 274
91.60
34 / 274

Writing and Creative Capabilities

1 evaluations
Benchmark / mode
Score
Rank/total
Creative Writing
Standard Mode
1825.80
17 / 106

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleBench
Extra-High
46.80
54 / 94

Text Embedding

1 evaluations
Benchmark / mode
Score
Rank/total
81.80
34 / 126

AI Agent - Tool Usage

4 evaluations
Benchmark / mode
Score
Rank/total
84.70
16 / 53
75.70
33 / 53
17.27
16 / 20
3.30
11 / 11

Coding and Software Engineer

4 evaluations
Benchmark / mode
Score
Rank/total
1522.94
15 / 35
DeepSWE
Extra-HighTools
67.20
13 / 38
WeirdML v2
HighTools
60.86
27 / 52

Agent Level Benchmark

1 evaluations
Benchmark / mode
Score
Rank/total
Agents' Last Exam
Extra-HighTools
50.30
5 / 19

Productivity Knowledge

1 evaluations
Benchmark / mode
Score
Rank/total
AA-Briefcase
MaxTools
1359.76
11 / 20

Math and Reasoning

4 evaluations
Benchmark / mode
Score
Rank/total
FrontierMath v2
Standard Mode
39.65
36 / 58
41.40
35 / 58
82.11
8 / 58

Compare with other models

GPT-5.6 Luna

Publisher

GPT-5.6 Luna

Model Overview

GPT-5.6 Luna

GPT-5.6 Luna is OpenAI's most cost-efficient tier in the GPT-5.6 family, first previewed on June 26, 2026 and made generally available on July 9, 2026. It's built for high-throughput, cost-sensitive tasks — summarization, drafting, routine automation — and is the fastest and cheapest of the three GPT-5.6 models. The family also includes Sol (the flagship) and Terra (the balanced mid-tier).

On Artificial Analysis's Intelligence Index (max reasoning), Luna scores 51 at roughly one-fifth Sol's cost per task (~$0.21 vs ~$1.04), delivering about 24 benchmark points per estimated API dollar. It reaches 84.70 on Terminal-Bench 2.1, 67.20 on DeepSWE, 74.6 on the AA Coding Agent Index, and 50.3 on Agents' Last Exam. Long-context recall is comparatively weak (41.3% on MRCR), so it's best reserved for shorter, well-scoped jobs rather than large-document workflows.

Luna shares OpenAI's GPT-5.6 system card with Sol and Terra, including "High" Preparedness Framework ratings for cybersecurity and biological/chemical risk; see the GPT-5.6 Sol page for the detailed safety-evaluation breakdown.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code