GPT-5.4 mini Benchmark Details
GPT-5.4 mini currently shows benchmark results led by GPQA Diamond (73 / 274, score 88), Creative Writing (34 / 106, score 1661.80), HLE (70 / 197, score 41.50). This page also compares it with 2 competitor models and 1 predecessor or same-series models, including performance and pricing views when available.
Benchmark Results
Benchmark Results
General Knowledge
7 evaluationsWriting and Creative Capabilities
1 evaluationsMath and Reasoning
5 evaluationsCoding and Software Engineer
1 evaluationsAI Agent - Tool Usage
4 evaluationsText Embedding
5 evaluationsCompetitor Comparison
Benchmark scores for GPT-5.4 mini compared against top models in its class
9 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | GPT-5.4 miniCurrent | Haiku 4.5 | Gemini 3.0 Flash |
|---|---|---|---|
41.50Thinking Level · Extra High | Tools | 9.70Extended Thinking | 43.50Thinking Enabled | Tools | |
66.37Deep Thinking Mode | 61.3264K | 72.40Thinking Level · High | |
88.00Thinking Level · Extra High | 73.30Extended Thinking | 90.40Thinking Enabled | |
2.10Thinking Level · High | 2.1032K | 4.20Standard Mode | |
54.40Thinking Level · Extra High | Tools | 39.45Extended Thinking | Tools | 49.60Thinking Level · High | Tools | |
56.70Thinking Level · Extra High | Tools | 40.20Standard Mode | Tools | 62.00Standard Mode | Tools | |
60.00Thinking Level · Extra High | Tools | -- | 47.60Thinking Enabled | Tools | |
50.92Thinking Level · Extra High | 36.27Thinking Enabled | -- | |
75.30Thinking Enabled | Tools | 89.40Thinking Enabled | Tools | 85.70Thinking Enabled | Tools |
Standard API Pricing: GPT-5.4 mini vs. Peer Models
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
GPT-5.4 mini | OpenAI | $0.75 / 1M tokens | $4.5 / 1M tokens | — |
Haiku 4.5 | Anthropic | $1 / 1M tokens | $5 / 1M tokens | — |
Gemini 3.0 Flash | Google Deep Mind | $0.5 / 1M tokens | $3 / 1M tokens | — |
Version History
How each version of the GPT-5.4 mini series stacks up on benchmark tests
7 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.
| Benchmark | GPT-5.4 miniCurrent | GPT-5-mini |
|---|---|---|
41.50Thinking Level · Extra High | Tools | 5.00Thinking Enabled | |
66.37Deep Thinking Mode | 65.91Thinking Level · High | |
88.00Thinking Level · Extra High | 69.00Thinking Enabled | |
1661.80Standard Mode | 1310.30Standard Mode | |
2.10Thinking Level · High | 6.30Thinking Level · High | |
9.76Thinking Level · Extra High | 12.20Thinking Level · High | |
51.23Thinking Level · Extra High | 46.67Thinking Level · High |
Single-Benchmark Version Trend
Viewing: HLE · 综合评估
Standard API Pricing Across the GPT-5.4 mini Series
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
GPT-5.4 mini | OpenAI | $0.75 / 1M tokens | $4.5 / 1M tokens | — |
GPT-5-mini | OpenAI | $0.25 / 1M tokens | $2 / 1M tokens | — |