GPT-5.6 Terra Benchmark Details
GPT-5.6 Terra currently shows benchmark results led by Terminal Bench Hard (2 / 244, score 62.90), CritPt (7 / 201, score 30), FrontierMath v2 (4 / 58, score 85.96). This page also compares it with 3 competitor models and 2 predecessor or same-series models, including performance and pricing views when available.
Benchmark Results
Benchmark Results
General Knowledge
31 evaluationsOther
6 evaluationsWriting and Creative Capabilities
1 evaluationsAgent Level Benchmark
16 evaluationsInstruction Following
5 evaluationsLong Context
6 evaluationsAI Agent - Tool Usage
14 evaluationsCoding and Software Engineer
18 evaluationsProductivity Knowledge
13 evaluationsMultimodal Understanding
11 evaluationsMath and Reasoning
2 evaluationsCompetitor Comparison
Benchmark scores for GPT-5.6 Terra compared against top models in its class
12 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | GPT-5.6 TerraCurrent | Claude Sonnet 5 | Gemini 3.5 Flash | GLM-5.2 |
|---|---|---|---|---|
42.30Standard Mode | Tools | 38.40Standard Mode | Tools | -- | 34.00Standard Mode | Tools | |
96.50Thinking Level · High | -- | 92.50Thinking Level · High | Tools | -- | |
83.90Thinking Level · High | -- | 72.08Thinking Level · High | Tools | -- | |
30.00Thinking Level · High | 16.90Thinking Level · High | 13.10Thinking Level · High | 20.90Thinking Level · High | |
42.90Thinking Level · High | 57.40Thinking Level · Extra High | Tools | 42.70Thinking Level · High | 54.70Thinking Enabled | Tools | |
51.10Thinking Level · High | 31.00Thinking Level · High | -- | -- | |
93.31Thinking Level · High | 90.53Thinking Level · Extra High | 92.80Thinking Level · High | 91.86Thinking Level · High | |
1850.30Standard Mode | 1790.50Standard Mode | -- | 1752.80Standard Mode | |
48.90Thinking Level · Extra High | 60.60Standard Mode | 76.70Standard Mode | 58.80Standard Mode | |
62.90Thinking Level · Extra High | Tools | -- | 46.20Standard Mode | Tools | 50.80Thinking Level · High | Tools | |
86.30Thinking Level · High | Tools | -- | 95.60Thinking Level · Medium | Tools | 99.10Thinking Level · High | Tools | |
40.20Thinking Level · High | Tools | 37.30Thinking Level · High | Tools | 32.20Thinking Level · High | Tools | 37.11Thinking Level · Extra High | Tools |
Standard API Pricing: GPT-5.6 Terra vs. Peer Models
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
GPT-5.6 Terra | OpenAI | $2 / 1M tokens | $12 / 1M tokens | — |
Claude Sonnet 5 | Anthropic | $2 / 1M tokens | $10 / 1M tokens | — |
Gemini 3.5 Flash | Google DeepMind | $1.5 / 1M tokens | $9 / 1M tokens | — |
GLM-5.2 | 智谱AI | $1.4 / 1M tokens | $4.4 / 1M tokens | — |
Version History
How each version of the GPT-5.6 Terra series stacks up on benchmark tests
12 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.
| Benchmark | GPT-5.6 TerraCurrent | GPT-5.5 | GPT-5.4 |
|---|---|---|---|
42.30Standard Mode | Tools | 38.60Thinking Level · Extra High | Tools | -- | |
96.50Thinking Level · High | 95.00Thinking Level · Extra High | 93.67Thinking Level · Extra High | |
83.90Thinking Level · High | 85.00Thinking Level · Extra High | 77.10Standard Mode | |
30.00Thinking Level · High | 27.10Thinking Level · Extra High | 23.40Thinking Level · Extra High | |
42.90Thinking Level · High | 52.20Thinking Level · High | Tools | 52.10Thinking Level · Extra High | Tools | |
93.31Thinking Level · High | 94.00Thinking Level · Extra High | 92.00Thinking Level · Extra High | |
1850.30Standard Mode | 1843.50Standard Mode | 1835.60Standard Mode | |
48.90Thinking Level · Extra High | 69.00Standard Mode | -- | |
62.90Thinking Level · Extra High | Tools | 60.60Thinking Level · Extra High | Tools | 57.60Thinking Level · Extra High | Tools | |
86.30Thinking Level · High | Tools | 93.90Thinking Level · Extra High | Tools | 87.10Thinking Level · Extra High | Tools | |
40.20Thinking Level · High | Tools | 44.59Thinking Level · Extra High | Tools | 39.43Thinking Level · Extra High | Tools | |
71.20Thinking Level · High | 75.90Thinking Level · Extra High | 73.90Thinking Level · Extra High |
Single-Benchmark Version Trend
Viewing: AA Intelligence Index v4.3 · 综合评估
Standard API Pricing Across the GPT-5.6 Terra Series
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
GPT-5.6 Terra | OpenAI | $2 / 1M tokens | $12 / 1M tokens | — |
GPT-5.5 | OpenAI | $5 / 1M tokens | $30 / 1M tokens | — |
GPT-5.4 | OpenAI | $2.5 / 1M tokens | $15 / 1M tokens | — |