DataLearner logo

Gemini 3.6 Flash Benchmark Details

Gemini 3.6 Flash currently shows benchmark results led by OSWorld-Verified (3 / 23, score 83), SWE-Bench Pro - Public (11 / 53, score 58.70), MLE-Bench (1 / 3, score 63.90). This page also compares it with 3 competitor models and 3 predecessor or same-series models, including performance and pricing views when available.

Benchmark Results

Gemini 3.6 Flash

Benchmark Results

Thinking
Tool usage

Coding and Software Engineer

2 evaluations
Benchmark / mode
Score
Rank/total
SWE-Bench Pro - Public
Thinking EnabledTools
58.70
11 / 53
DeepSWE
Thinking EnabledTools
49
12 / 18

AI Agent - Tool Usage

3 evaluations
Benchmark / mode
Score
Rank/total
OSWorld-Verified
Thinking EnabledTools
83
3 / 23
TerminalBench 2.1
Thinking EnabledTools
78
13 / 27
MLE-Bench
Thinking EnabledTools
63.90
1 / 3

Productivity Knowledge

1 evaluations
Benchmark / mode
Score
Rank/total
GDPval-AA v2
Thinking Enabled
1421
2 / 4

Other

2 evaluations
Benchmark / mode
Score
Rank/total
91.80
1 / 3
54
1 / 3

Competitor Comparison

Benchmark scores for Gemini 3.6 Flash compared against top models in its class

Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

4 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.

BenchmarkGemini 3.6 FlashCurrentGPT-5.6 TerraClaude Sonnet 5Grok 4.5
DeepSWE
编程与软件工程
49.00Thinking Enabled | Tools
69.60Thinking Level · Extra High | Tools
54.00Deep Thinking Mode | Tools
53.00Thinking Level · High | Tools
SWE-Bench Pro - Public
编程与软件工程
58.70Thinking Enabled | Tools
--
--
64.70Thinking Level · High | Tools
OSWorld-Verified
AI Agent - 工具使用
83.00Thinking Enabled | Tools
--
81.20Thinking Level · Extra High | Tools
--
TerminalBench 2.1
AI Agent - 工具使用
78.00Thinking Enabled | Tools
87.40Thinking Level · High
80.40Thinking Level · Extra High | Tools
83.30Thinking Level · High | Tools

Standard API Pricing: Gemini 3.6 Flash vs. Peer Models

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens

ModelSupplierStandard inputStandard outputBase price applies to
Gemini 3.6 Flash
DeepMind$1.5 / 1M tokens$7.5 / 1M tokens
GPT-5.6 Terra
OpenAI$2.5 / 1M tokens$15 / 1M tokens
Claude Sonnet 5
Anthropic$2 / 1M tokens$10 / 1M tokens
Grok 4.5
xAI$2 / 1M tokens$6 / 1M tokens

Version History

How each version of the Gemini 3.6 Flash series stacks up on benchmark tests

Benchmark categories:
The chart shows each model’s highest score per benchmark within the current filter. Out-of-100 benchmarks use raw heights; out-of-range benchmarks are scaled within that benchmark while labels keep the original scores.

8 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.

BenchmarkGemini 3.6 FlashCurrentGemini 3.5 FlashGemini 3.0 Flash
DeepSWE
编程与软件工程
49.00Thinking Enabled | Tools
37.00Thinking Level · Medium | Tools
--
SWE-Bench Pro - Public
编程与软件工程
58.70Thinking Enabled | Tools
55.10Thinking Level · High | Tools
49.60Thinking Level · High | Tools
MLE-Bench
AI Agent - 工具使用
63.90Thinking Enabled | Tools
49.70Thinking Enabled | Tools
--
OSWorld-Verified
AI Agent - 工具使用
83.00Thinking Enabled | Tools
78.40Thinking Level · High | Tools
--
TerminalBench 2.1
AI Agent - 工具使用
78.00Thinking Enabled | Tools
76.20Thinking Level · High | Tools
58.00Thinking Level · High | Tools
GDPval-AA v2
生产力知识
1421.00Thinking Enabled
1349.00Thinking Enabled
--
91.80Thinking Enabled
77.30Thinking Enabled
--
54.00Thinking Enabled
26.60Thinking Enabled
--

Single-Benchmark Version Trend

Viewing: DeepSWE · 编程与软件工程

Benchmark
NormalNormal + ToolsThinkingThinking + ToolsDeepDeep + Tools

X-axis shows model and release date, Y-axis shows score; solid lines connect the same mode across versions, while dotted guides align modes within the same generation.

Standard API Pricing Across the Gemini 3.6 Flash Series

Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.

Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens

ModelSupplierStandard inputStandard outputBase price applies to
Gemini 3.6 Flash
DeepMind$1.5 / 1M tokens$7.5 / 1M tokens
Gemini 3.5 Flash
DeepMind$1.5 / 1M tokens$9 / 1M tokens
Gemini 3.0 Flash
Google Deep Mind$0.5 / 1M tokens$3 / 1M tokens