Gemini 3.7 Flash Benchmark Details
Gemini 3.7 Flash currently shows benchmark results led by GPQA Diamond (1 / 253, score 94.82), GDP.pdf (1 / 122, score 34), MMMU-Pro (5 / 229, score 85.50). This page also compares it with 3 competitor models and 3 predecessor or same-series models, including performance and pricing views when available.
Benchmark Results
Benchmark Results
Abstract Generalization
6 evaluationsScientific Reasoning
4 evaluationsMemory & Persistence
3 evaluationsAgentic Development
4 evaluationsRepository Engineering
3 evaluationsCapability Indices
3 evaluationsVisual Understanding
3 evaluationsDocuments & Charts
4 evaluationsScientific Computing
3 evaluationsService Workflows
4 evaluationsLegal
3 evaluationsFinance
3 evaluationsMathematics
3 evaluationsCode Generation & Editing
2 evaluationsBiology & Genomics
3 evaluationsClinical Workflows
2 evaluationsCompetitor Comparison
Benchmark scores for Gemini 3.7 Flash compared against top models in its class
12 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | Gemini 3.7 FlashCurrent | Claude Sonnet 5 | Qwen3.8-27B | DeepSeek-V4-Flash |
|---|---|---|---|---|
14.30Thinking Level · High | 16.90Thinking Level · High | 5.40Thinking Level · Extra High | 7.10Thinking Level · High | |
94.82Thinking Level · High | 90.53Thinking Level · Extra High | 89.20Thinking Enabled | 71.20Standard Mode | |
1722.00Standard Mode | 1790.50Standard Mode | 1668.40Standard Mode | 1555.70Standard Mode | |
95.95Thinking Level · High | 79.53Thinking Level · High | 93.98Thinking Level · Medium | 69.42Thinking Enabled | |
85.80Thinking Enabled | Tools | 80.40Thinking Level · Extra High | Tools | 73.00Thinking Enabled | Tools | -- | |
14.90Thinking Level · High | Tools | -- | -- | 7.60Thinking Level · High | Tools | |
13.60Thinking Level · High | Tools | 12.42Thinking Level · High | Tools | 5.60Thinking Level · Extra High | Tools | 3.00Thinking Level · High | Tools | |
65.49Thinking Level · Medium | Tools | 54.00Deep Thinking Mode | Tools | 42.20Thinking Enabled | Tools | 53.32Thinking Level · High | Tools | |
157.72Thinking Level · High | 156.34Thinking Level · High | 149.38Thinking Level · High | 146.10Thinking Level · High | |
59.31Thinking Level · High | Tools | 59.61Thinking Level · High | Tools | 48.48Thinking Level · Extra High | Tools | -- | |
26.30Thinking Level · Medium | Tools | -- | 20.40Thinking Enabled | Tools | 25.20Thinking Level · High | Tools | |
30.40Thinking Enabled | Tools | -- | -- | 37.70Thinking Level · High | Tools |
Standard API Pricing: Gemini 3.7 Flash vs. Peer Models
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
Gemini 3.7 Flash | Google DeepMind | $0.75 / 1M tokens | $3.75 / 1M tokens | — |
Claude Sonnet 5 | Anthropic | $2 / 1M tokens | $10 / 1M tokens | — |
DeepSeek-V4-Flash | DeepSeek-AI | $0.14 / 1M tokens | $0.28 / 1M tokens | — |
Version History
How each version of the Gemini 3.7 Flash series stacks up on benchmark tests
12 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.
| Benchmark | Gemini 3.7 FlashCurrent | Gemini 3.6 Flash | Gemini 3.5 Flash | Gemini 3.0 Flash |
|---|---|---|---|---|
95.50Thinking Level · High | 91.17Thinking Level · High | Tools | 92.50Thinking Level · High | Tools | -- | |
84.58Thinking Level · High | 60.42Thinking Level · High | Tools | 72.08Thinking Level · High | Tools | 33.60Thinking Enabled | |
14.30Thinking Level · High | 10.60Thinking Level · High | 13.10Thinking Level · High | 8.60Thinking Enabled | |
94.82Thinking Level · High | 94.13Thinking Level · High | 92.80Thinking Level · High | 90.40Thinking Enabled | |
1722.00Standard Mode | 1599.80Standard Mode | -- | -- | |
95.95Thinking Level · High | 88.78Thinking Level · High | 77.19Thinking Level · High | -- | |
85.80Thinking Enabled | Tools | 78.00Thinking Enabled | Tools | -- | 58.00Thinking Level · High | Tools | |
13.60Thinking Level · High | Tools | 7.10Thinking Level · High | Tools | 6.60Thinking Level · High | Tools | -- | |
65.49Thinking Level · Medium | Tools | 49.00Thinking Enabled | Tools | 37.00Thinking Level · Medium | Tools | -- | |
157.72Thinking Level · High | 154.36Thinking Level · High | 154.55Thinking Level · High | 151.83Thinking Level · High | |
59.31Thinking Level · High | Tools | 55.35Thinking Level · High | Tools | 53.08Thinking Level · High | Tools | -- | |
85.50Thinking Level · High | 83.20Thinking Level · High | 84.30Thinking Level · High | 79.90Thinking Enabled |
Single-Benchmark Version Trend
Viewing: ARC-AGI-1 · Abstract Generalization
Standard API Pricing Across the Gemini 3.7 Flash Series
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
Gemini 3.7 Flash | Google DeepMind | $0.75 / 1M tokens | $3.75 / 1M tokens | — |
Gemini 3.6 Flash | Google DeepMind | $1.5 / 1M tokens | $7.5 / 1M tokens | — |
Gemini 3.5 Flash | Google DeepMind | $1.5 / 1M tokens | $9 / 1M tokens | — |
Gemini 3.0 Flash | Google DeepMind | $0.5 / 1M tokens | $3 / 1M tokens | — |