GPT-5.3 Codex Benchmark Details
GPT-5.3 Codex currently shows benchmark results led by Terminal Bench 2.0 (3 / 48, score 77.30), Terminal Bench Hard (17 / 244, score 53), ECI (16 / 167, score 156.84). This page also compares it with 2 competitor models and 3 predecessor or same-series models, including performance and pricing views when available. 1 source link is attached for reference.
Benchmark Results
Benchmark Results
Repository Engineering
2 evaluationsCross-capability Suites
2 evaluationsAgentic Development
2 evaluationsML Engineering
2 evaluationsCapability Frontier Metrics
1 evaluationsCode Generation & Editing
1 evaluationsCompetitor Comparison
Benchmark scores for GPT-5.3 Codex compared against top models in its class
9 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | GPT-5.3 CodexCurrent | Claude Opus 4.6 | Gemini 3.0 Pro (Preview 11-2025) |
|---|---|---|---|
72.76Thinking Level · High | 74.52Thinking Level · High | 73.39Thinking Level · High | |
77.30Thinking Level · Extra High | Tools | 65.40Extended Thinking | Tools | 56.90Thinking Level · High | Tools | |
53.00Thinking Level · Extra High | Tools | 48.50Standard Mode | Tools | 41.70Thinking Level · High | Tools | |
78.50Thinking Level · Extra High | 77.30Thinking Enabled | Tools | 80.20Thinking Level · High | |
16.90Thinking Level · Extra High | 12.60Thinking Level · High | 9.10Thinking Level · High | |
79.30Thinking Level · Extra High | Tools | 77.95Thinking Level · High | Tools | -- | |
1406.65Standard Mode | 1555.35Standard Mode | -- | |
61.77Thinking Level · Extra High | Tools | 57.57Standard Mode | Tools | -- | |
156.84Thinking Level · High | 155.36Thinking Level · High | -- |
Standard API Pricing: GPT-5.3 Codex vs. Peer Models
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
When a context threshold exists, the charted base price only applies within these limits:
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
GPT-5.3 Codex | OpenAI | $1.75 / 1M tokens | $14 / 1M tokens | — |
Claude Opus 4.6 | Anthropic | $5 / 1M tokens | $25 / 1M tokens | — |
Gemini 3.0 Pro (Preview 11-2025) | Google DeepMind | $2 / 1M tokens | $12 / 1M tokens | <= 200000 |
Version History
How each version of the GPT-5.3 Codex series stacks up on benchmark tests
6 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.
| Benchmark | GPT-5.3 CodexCurrent | GPT-5.2-Codex | GPT-5.1-Codex-Max | GPT-5 Codex |
|---|---|---|---|---|
72.76Thinking Level · High | 73.98Standard Mode | 73.98Deep Thinking Mode | -- | |
53.00Thinking Level · Extra High | Tools | 37.10Thinking Level · Extra High | Tools | -- | 37.90Thinking Level · High | Tools | |
78.50Thinking Level · Extra High | 76.30Thinking Level · Extra High | -- | 73.80Thinking Level · High | |
16.90Thinking Level · Extra High | 8.70Thinking Level · Extra High | -- | 5.10Thinking Level · High | |
349.53Standard Mode | Tools | -- | 161.75Standard Mode | Tools | -- | |
61.77Thinking Level · Extra High | Tools | 37.91Thinking Level · High | Tools | 22.17Thinking Level · High | Tools | -- |
Single-Benchmark Version Trend
Viewing: LiveBench · Cross-capability Suites
Standard API Pricing Across the GPT-5.3 Codex Series
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
GPT-5.3 Codex | OpenAI | $1.75 / 1M tokens | $14 / 1M tokens | — |
GPT-5.2-Codex | OpenAI | $1.25 / 1M tokens | $10 / 1M tokens | — |
GPT-5.1-Codex-Max | OpenAI | $1.25 / 1M tokens | $10 / 1M tokens | — |
GPT-5 Codex | OpenAI | $1.25 / 1M tokens | $10 / 1M tokens | — |

