DataLearner logo
GP

GPT-4.1

Chat modelGPTGPT-4.1

GPT-4.1

Release date: 2025-04-14Updated: 2025-04-15Views: 1,290
Live demoGitHubHugging FaceCompare
Parameters
No data
Context length
1024K
Multilingual
Supported
Reasoning ability
4/5

GPT-4.1 is an AI model published by OpenAI, released on 2025-04-14, for Chat model, and 1024K context length, with a 1417.00 score on Creative Writing.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

GPT-4.1

Model basics

Reasoning traces
No data
Thinking modes
Thinking modes not supported
Context length
1024K tokens
Max output length
32K tokens
Model type
Chat model
Modality (in / out)
Text, Image → Text
Release date
2025-04-14
Model file size
No data
MoE architecture
No
Total params / Active params
No data / Not applicable
Knowledge cutoff
No data
GPT-4.1

Open source & experience

Code license
Proprietary
Weights license
Proprietary
GitHub repo
N/A
Hugging Face
N/A
Live demo
N/A
GPT-4.1

Official resources

Paper
DataLearnerAI blog
N/A
GPT-4.1

API details

API speed
3/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-$2.00/ 1M$8.00/ 1M
Image-$2.00/ 1M
Cache PricingPrompt Cache
TypeTTLWriteRead
Text-$0.500/ 1M
Text-$0.500/ 1M
Cache = hit
Image-$0.500/ 1M
Cache = hit

“—” means the modality is not billed in that direction, or the vendor has not published a price for it.

GPT-4.1

Benchmark Results

GPT-4.1 currently shows benchmark results led by GSM8K (5 / 70, score 95.90), MMLU (9 / 124, score 90.20), DROP (4 / 9, score 89.20). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage

General Knowledge

3 evaluations
Benchmark / mode
Score
Rank/total
MMLU
Standard Mode
90.20
9 / 124
MMLU Pro
Standard Mode
80.50
61 / 133
HLE
Standard Mode
3.70
184 / 185

Math and Reasoning

6 evaluations
Benchmark / mode
Score
Rank/total
GSM8K
Standard Mode
95.90
5 / 70
MATH-500
Standard Mode
92.80
30 / 44
AIME 2024
Standard Mode
48.10
49 / 62
AIME2025
Standard Mode
36.70
97 / 106
FrontierMath
Standard Mode
5.50
37 / 60
0
72 / 80

Reading Comprehension

1 evaluations
Benchmark / mode
Score
Rank/total
DROP
Standard Mode
89.20
4 / 9

General Evaluation

1 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Standard Mode
66.30
208 / 270

Coding and Software Engineer

5 evaluations
Benchmark / mode
Score
Rank/total
SWE-bench Verified
Standard Mode
54.60
90 / 114
LiveCodeBench
Standard Mode
40.50
106 / 127
WeirdML v2
Standard ModeTools
39.04
48 / 52
35.10
1 / 1
14.40
8 / 8

Writing and Creative Capabilities

1 evaluations
Benchmark / mode
Score
Rank/total
Creative Writing
Standard Mode
1417
56 / 99

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleBench
Standard Mode
27
56 / 67

Agent Level Benchmark

2 evaluations
Benchmark / mode
Score
Rank/total
τ²-Bench
Standard ModeTools
54.70
32 / 43
Aider-Polyglot
Standard Mode
52.40
31 / 59

Long Context

1 evaluations
Benchmark / mode
Score
Rank/total
Fiction.liveBench
Standard Mode
63.90
15 / 16

Multimodal Understanding

1 evaluations
Benchmark / mode
Score
Rank/total
GeoBench ACW
Standard Mode
72
12 / 20
GPT-4.1

Publisher

GPT-4.1

Model Overview

GPT-4.1 is an AI model published by OpenAI, released on 2025-04-14, for Chat model, and 1024K context length, with a 1417.00 score on Creative Writing.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code