See key specs and per-benchmark scores for each model/mode. Scroll horizontally for all columns. 当前对比 2 个模型的评测数据与核心参数。

Claude Sonnet 5
Anthropic

Gemini 3.6 Flash
Google Deep Mind
Best overall
Claude Sonnet 5 · 71.87
Best single
Gemini 3.6 Flash · OSWorld-Verified 83.00
Modality coverage
Gemini 3.6 Flash · 4 modalities
Head to head
3
Benchmarks
2
Wins
1
Losses
+1.87
Average diff
Compare benchmark results across thinking modes and tool usage.
Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology
Complete scores for each model/mode across selected benchmarks.
3 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | Claude Sonnet 5 | Gemini 3.6 Flash |
|---|---|---|
DeepSWE 编程与软件工程 | 54.00Deep Thinking Mode | Tools | 49.00Thinking Enabled | Tools |
OSWorld-Verified AI Agent - 工具使用 | 81.20Thinking Level · Extra High | Tools | 83.00Thinking Enabled | Tools |
TerminalBench 2.1 AI Agent - 工具使用 | 80.40Thinking Level · Extra High | Tools | 78.00Thinking Enabled | Tools |
Side-by-side input/output token pricing
Licensing, MoE architecture, and multi-modality support.
| Features & specs | Claude Sonnet 5Anthropic | Gemini 3.6 FlashGoogle Deep Mind |
|---|---|---|
Core specsRelease | 2026-06-30 | 2026-07-21 |
Context length | 1M | 1M |
Max output | 128000 | 65536 |
MoE | No | No |
Supported modes | 常规模式(Non-Thinking Mode)思考模式(Thinking Mode)深度思考(Deeper Thinking Mode) | No mode data |
LicenseCode Open Source | Not provided | Not provided |
Weights Open Source | Not provided | Not provided |
Commercial use | 不开源 | 不开源 |
Modality supportText Input/Output | / | / |
Image Input/Output | / | / |
Audio Input/Output | Not provided | / |
Video Input/Output | Not provided | / |
ResourcesPaper / report | Introducing Claude Sonnet 5 | Gemini 3.6 Flash |