DataLearner logo
GE

Gemini 3.1 Pro Preview

Multimodal modelGemini ProGemini 3.1

Gemini 3.1 Pro Preview

Release date: 2026-02-20Updated: 2026-07-17 21:57:28.408Knowledge cutoff: 2025-017,187
Live demoGitHubHugging FaceCompare
Parameters
Not disclosed
Context length
1M
Chinese support
Supported
Reasoning ability

Gemini 3.1 Pro Preview is a multimodal model from Google Deep Mind, released on 2026-02-20. It accepts text, image, audio, and video input and returns text output. The recorded context window is 1M, and the recorded maximum output is 64K. Cataloged capabilities include Reasoning model and Multilingual. The model weights are proprietary and are not published for download. Use the linked references to confirm current access, licensing, and provider-specific limits.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

Gemini 3.1 Pro Preview

Model basics

Reasoning traces
Supported
Thinking modes
Thinking Level · High (Default)Thinking Level · Low
Context length
1M tokens
Max output length
64K tokens
Model type
Multimodal model
Modality (in / out)
Text, Image, Audio, Video → Text
Release date
2026-02-20
Model file size
No data
MoE architecture
No
Total params / Active params
No data / N/A
Knowledge cutoff
2025-01
Gemini 3.1 Pro Preview

Open source & experience

Code license
Proprietary
Weights license
Proprietary
GitHub repo
GitHub link unavailable
Hugging Face
Hugging Face link unavailable
Gemini 3.1 Pro Preview

Official resources

Paper
DataLearnerAI blog
No blog post yet
Gemini 3.1 Pro Preview

API details

API speed
3/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
TextContext <= 200K$2.00/ 1M$12.00/ 1M
TextContext > 200K$4.00/ 1M$18.00/ 1M
Gemini 3.1 Pro Preview

Benchmark Results

Gemini 3.1 Pro Preview currently shows benchmark results led by GPQA Diamond (4 / 226, score 94.30), LiveCodeBench (3 / 126, score 91.70), LiveBench (3 / 115, score 79.93). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage
Internet

General Knowledge

6 evaluations
Benchmark / mode
Score
Rank/total
MMLU
High
92.60
3 / 66
79.93
3 / 115
77.10
9 / 62
HLE
High
44.40
47 / 181
HLE
HighTools
51.40
24 / 181
0
6 / 9

General Evaluation

1 evaluations
Benchmark / mode
Score
Rank/total
94.30
4 / 226

Coding and Software Engineer

8 evaluations
Benchmark / mode
Score
Rank/total
Text Arena (Coding)
Standard Mode
1461.49
23 / 35
LiveCodeBench
HighTools
91.70
3 / 126
80.60
11 / 114
WeirdML v2
Standard ModeTools
72.10
17 / 52
54.20
34 / 57
SWE-Bench Pro - Commercial
Thinking ModeTools
32.20
3 / 3
GSO
Standard ModeTools
22.55
10 / 21
DeepSWE
HighTools
12
27 / 27

Multimodal Understanding

3 evaluations
Benchmark / mode
Score
Rank/total
CharXiv RQ
Thinking Mode
83.30
13 / 15
CharXiv RQ
Thinking ModeTools
83.20
14 / 15
MMMU
High
80.50
12 / 29

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleBench
Standard Mode
79.60
2 / 67

Agent Level Benchmark

4 evaluations
Benchmark / mode
Score
Rank/total
METR Time Horizons v1.1
Standard ModeTools
384.15
2 / 22
99.30
1 / 35
τ²-Bench
HighTools
90.80
2 / 43
BALROG
Standard ModeTools
57
2 / 12

Math and Reasoning

3 evaluations
Benchmark / mode
Score
Rank/total
36.90
11 / 60
16.70
20 / 80
16.70
20 / 80

AI Agent - Information Search

1 evaluations
Benchmark / mode
Score
Rank/total
BrowseComp
HighToolsInternet
85.90
5 / 54

AI Agent - Tool Usage

5 evaluations
Benchmark / mode
Score
Rank/total
MCP-Atlas
HighTools
78.20
12 / 38
OSWorld-Verified
Thinking ModeTools
76.20
12 / 26
73.80
28 / 44
68.50
8 / 48
MLE-Bench
Thinking ModeTools
42.60
3 / 3

Claw-style Agent Evaluation

1 evaluations
Benchmark / mode
Score
Rank/total
Pinch Bench
Thinking ModeTools
86.70
11 / 38

Productivity Knowledge

1 evaluations
Benchmark / mode
Score
Rank/total
GDPval-AA v2
Thinking Mode
965
12 / 13

Long Context

2 evaluations
Benchmark / mode
Score
Rank/total
84.90
3 / 8
26.30
3 / 3

Compare with other models

Gemini 3.1 Pro Preview

Publisher

Google Deep Mind
Google Deep Mind
View publisher details
Gemini 3.1 Pro Preview

Model Overview

Gemini 3.1 Pro Preview is a multimodal model from Google Deep Mind, released on 2026-02-20.

It accepts text, image, audio, and video input and produces text output. Its cataloged capabilities include Reasoning model and Multilingual. The recorded context window is 1M, and the recorded maximum output is 64K.

The model weights are proprietary and are not published for download. The page records 4 API pricing rules from Google Deep Mind; current provider pricing and conditions should be checked before deployment. The evaluation section contains 23 cataloged benchmark results with their recorded modes and scores. The page links 2 release, model-card, repository, or provider references for checking the underlying claims. Specifications, availability, and prices can change; undisclosed values are intentionally left unstated.

Gemini 3.1 Pro Preview

FAQ

What is Gemini 3.1 Pro Preview?

Gemini 3.1 Pro Preview is a multimodal model from Google Deep Mind, released on 2026-02-20. It accepts text, image, audio, and video input and returns text output. The recorded context window is 1M, and the recorded maximum output is 64K. Cataloged capabilities include Reasoning model and Multilingual. The model weights are proprietary and are not published for download. Use the linked references to confirm current access, licensing, and provider-specific limits.

What input and output modalities does Gemini 3.1 Pro Preview support?

The current model record lists text, image, audio, and video as input and text as output.

What are the main recorded specifications for Gemini 3.1 Pro Preview?

The recorded context window is 1M, and the recorded maximum output is 64K. Fields without a source-backed value remain undisclosed.

Does Gemini 3.1 Pro Preview have API pricing?

The page records 4 API pricing rules from Google Deep Mind; current provider pricing and conditions should be checked before deployment.

Are benchmark results available for Gemini 3.1 Pro Preview?

The evaluation section contains 23 cataloged benchmark results with their recorded modes and scores. Compare only results that use the same benchmark version and evaluation mode.

Is Gemini 3.1 Pro Preview open source?

The model weights are proprietary and are not published for download. Review the linked license text before commercial or derivative use.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code