Gemini 3.1 Flash-Lite Benchmark Details
Gemini 3.1 Flash-Lite currently shows benchmark results led by SAGE (13 / 64, score 49.54), PinchBench v2 (10 / 45, score 80.50), MedCode (23 / 64, score 47.60). This page also compares it with 2 competitor models and 1 predecessor or same-series models, including performance and pricing views when available.
Benchmark Results
Benchmark Results
Scientific Reasoning
3 evaluationsAgentic Development
2 evaluationsTool Orchestration
2 evaluationsCapability Indices
2 evaluationsService Workflows
2 evaluationsLegal
3 evaluationsFinance
2 evaluationsClinical Workflows
2 evaluationsCompetitor Comparison
Benchmark scores for Gemini 3.1 Flash-Lite compared against top models in its class
12 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | Gemini 3.1 Flash-LiteCurrent | DeepSeek-V4-Flash | Qwen3.6-35B-A3B |
|---|---|---|---|
8.64Thinking Level · High | 51.50Thinking Level · High | Tools | 21.40Thinking Enabled | |
1.10Thinking Enabled | 7.10Thinking Level · High | 0.30Thinking Enabled | |
81.82Thinking Level · High | 71.20Standard Mode | 84.85Standard Mode | |
61.68Thinking Level · High | 65.48Standard Mode | -- | |
24.20Thinking Enabled | Tools | 38.60Thinking Level · High | Tools | 34.80Thinking Enabled | Tools | |
0.50Thinking Enabled | Tools | 3.00Thinking Level · High | Tools | -- | |
80.50Thinking Level · High | 81.74Thinking Level · High | -- | |
144.45Thinking Level · High | 146.10Thinking Level · High | 143.92Thinking Level · High | |
75.50Thinking Enabled | -- | 75.00Thinking Enabled | |
43.40Thinking Enabled | 45.30Thinking Level · High | 36.60Thinking Enabled | |
9.70Thinking Enabled | Tools | 30.90Thinking Level · High | Tools | 9.30Thinking Enabled | Tools | |
31.15Thinking Enabled | Tools | 81.33Thinking Level · High | Tools | -- |
Standard API Pricing: Gemini 3.1 Flash-Lite vs. Peer Models
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
Gemini 3.1 Flash-Lite | Google DeepMind | $0.25 / 1M tokens | $1.5 / 1M tokens | — |
DeepSeek-V4-Flash | DeepSeek-AI | $0.14 / 1M tokens | $0.28 / 1M tokens | — |
Version History
How each version of the Gemini 3.1 Flash-Lite series stacks up on benchmark tests
5 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.· Click a row to view its trend chart.
| Benchmark | Gemini 3.1 Flash-LiteCurrent | Gemini 2.5 Flash-Lite |
|---|---|---|
8.64Thinking Level · High | 6.90Standard Mode | |
81.82Thinking Level · High | 66.70Standard Mode | |
61.68Thinking Level · High | 42.56Thinking Level · High | |
24.20Thinking Enabled | Tools | 4.50Thinking Enabled | Tools | |
75.50Thinking Enabled | 58.20Thinking Enabled |
Single-Benchmark Version Trend
Viewing: HLE · Knowledge Exams
Standard API Pricing Across the Gemini 3.1 Flash-Lite Series
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier. · USD / 1M tokens
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
Gemini 3.1 Flash-Lite | Google DeepMind | $0.25 / 1M tokens | $1.5 / 1M tokens | — |
Gemini 2.5 Flash-Lite | Google DeepMind | $0.1 / 1M tokens | $0.4 / 1M tokens | — |