Grok 3 is an AI model published by xAI, released on 2025-02-17, for Chat model, and 128K context length, with a 84.20 score on AIME 2024.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
Model basics
Open source & experience
Official resources
API details
| Type | Condition | Input | Output |
|---|---|---|---|
| Text | - | $3.00/ 1M | $15.00/ 1M |
Benchmark Results
Grok 3 currently shows benchmark results led by AIME 2024 (22 / 62, score 84.20), SimpleQA (18 / 47, score 43.40), GPQA Diamond (81 / 188, score 80.40). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.
Math and Reasoning
4 evaluationsCompare with other models
No curated comparisons for this model yet.
Want a custom combination? Open the compare tool
Publisher
Model Overview
Grok 3 is an AI model published by xAI, released on 2025-02-17, for Chat model, and 128K context length, with a 84.20 score on AIME 2024.
DataLearner on WeChat
Follow DataLearner on WeChat for AI model updates and research notes.
