DataLearner logo
GR

Grok 4

Reasoning modelGrok 4

Grok 4

Release date: 2025-07-10Updated: 2026-07-17 23:18:25.5634,058
Live demoGitHubHugging FaceCompare
Parameters
Not disclosed
Context length
256K
Chinese support
Supported
Reasoning ability

Grok 4 is a reasoning model from xAI, released on 2025-07-10. It accepts text and image input and returns text output. The recorded context window is 256K, and the recorded maximum output is 256K. Cataloged capabilities include Reasoning model and Multilingual. The model weights are proprietary and are not published for download. Use the linked references to confirm current access, licensing, and provider-specific limits.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

Grok 4

Model basics

Reasoning traces
Supported
Thinking modes
Thinking modes not supported
Context length
256K tokens
Max output length
256K tokens
Model type
Reasoning model
Modality (in / out)
Text, Image → Text
Release date
2025-07-10
Model file size
No data
MoE architecture
No
Total params / Active params
No data / N/A
Knowledge cutoff
No data
Grok 4

Open source & experience

Code license
Proprietary
Weights license
Proprietary
GitHub repo
GitHub link unavailable
Hugging Face
Hugging Face link unavailable
Grok 4

Official resources

Paper
DataLearnerAI blog
Grok 4

API details

API speed
3/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-$3.00/ 1M$15.00/ 1M
Grok 4

Benchmark Results

Grok 4 currently shows benchmark results led by IMO 2024 (1 / 10, score 23.20), MMLU Pro (14 / 132, score 87), IMO 2025 (1 / 9, score 29.20). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking

General Knowledge

8 evaluations
Benchmark / mode
Score
Rank/total
87
14 / 132
87
42 / 188
66.70
32 / 68
LiveBench
Standard Mode
62.02
59 / 115
38.60
65 / 173
38.60
65 / 173
25.40
101 / 173
15.90
37 / 62

Coding and Software Engineer

2 evaluations
Benchmark / mode
Score
Rank/total
82
25 / 123
58.60
84 / 113

Math and Reasoning

9 evaluations
Benchmark / mode
Score
Rank/total
98.80
13 / 107
91.70
36 / 107
46.70
4 / 16
23.30
10 / 16
29.20
1 / 9
23.20
1 / 10
12.10
22 / 60
2.10
56 / 80

AI Agent - Tool Usage

1 evaluations
Benchmark / mode
Score
Rank/total

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
Simple Bench
Thinking Mode
60.50
15 / 63

Agent Level Benchmark

2 evaluations
Benchmark / mode
Score
Rank/total
79.60
7 / 59

Compare with other models

No curated comparisons for this model yet.

Want a custom combination? Open the compare tool

Grok 4

Publisher

Grok 4

Model Overview

Grok 4 is a reasoning model from xAI, released on 2025-07-10.

It accepts text and image input and produces text output. Its cataloged capabilities include Reasoning model and Multilingual. The recorded context window is 256K, and the recorded maximum output is 256K.

The model weights are proprietary and are not published for download. The page records 2 API pricing rules from xAI; current provider pricing and conditions should be checked before deployment. The evaluation section contains 23 cataloged benchmark results with their recorded modes and scores. The page links 3 release, model-card, repository, or provider references for checking the underlying claims. Specifications, availability, and prices can change; undisclosed values are intentionally left unstated.

Grok 4

FAQ

What is Grok 4?

Grok 4 is a reasoning model from xAI, released on 2025-07-10. It accepts text and image input and returns text output. The recorded context window is 256K, and the recorded maximum output is 256K. Cataloged capabilities include Reasoning model and Multilingual. The model weights are proprietary and are not published for download. Use the linked references to confirm current access, licensing, and provider-specific limits.

What input and output modalities does Grok 4 support?

The current model record lists text and image as input and text as output.

What are the main recorded specifications for Grok 4?

The recorded context window is 256K, and the recorded maximum output is 256K. Fields without a source-backed value remain undisclosed.

Does Grok 4 have API pricing?

The page records 2 API pricing rules from xAI; current provider pricing and conditions should be checked before deployment.

Are benchmark results available for Grok 4?

The evaluation section contains 23 cataloged benchmark results with their recorded modes and scores. Compare only results that use the same benchmark version and evaluation mode.

Is Grok 4 open source?

The model weights are proprietary and are not published for download. Review the linked license text before commercial or derivative use.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code