See key specs and per-benchmark scores for each model/mode. Scroll horizontally for all columns. 当前对比 2 个模型的评测数据与核心参数。

Gemini 3.6 Flash
Google Deep Mind
Best overall
Grok 4.5 · 67.00
Best single
Grok 4.5 · TerminalBench 2.1 83.30
Modality coverage
Gemini 3.6 Flash · 4 modalities
Head to head
3
Benchmarks
0
Wins
3
Losses
-5.10
Average diff
Compare benchmark results across thinking modes and tool usage.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
Complete scores for each model/mode across selected benchmarks.
3 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | Gemini 3.6 Flash | Grok 4.5 |
|---|---|---|
DeepSWE 编程与软件工程 | 49.00Thinking Enabled | Tools | 53.00Thinking Level · High | Tools |
SWE-Bench Pro - Public 编程与软件工程 | 58.70Thinking Enabled | Tools | 64.70Thinking Level · High | Tools |
TerminalBench 2.1 AI Agent - 工具使用 | 78.00Thinking Enabled | Tools | 83.30Thinking Level · High | Tools |
Side-by-side input/output token pricing
Licensing, MoE architecture, and multi-modality support.
| Features & specs | Gemini 3.6 FlashGoogle Deep Mind | Grok 4.5xAI |
|---|---|---|
Core specsRelease | 2026-07-21 | 2026-07-08 |
Context length | 1M | 500K |
Max output | 65536 | Not provided |
MoE | No | No |
LicenseCode Open Source | Not provided | Not provided |
Weights Open Source | Not provided | Not provided |
Commercial use | 不开源 | 不开源 |
Modality supportText Input/Output | / | / |
Image Input/Output | / | / |
Audio Input/Output | / | Not provided |
Video Input/Output | / | Not provided |
ResourcesPaper / report | Gemini 3.6 Flash | Introducing Grok 4.5 |