DataLearner logo

Gemini 3.8 FlashvsGemini 3.7 Flash

Across 11 shared benchmarks, Gemini 3.8 Flash leads overall: Gemini 3.8 Flash wins 10, Gemini 3.7 Flash wins 1, with 0 ties and an average score difference of +5.83.

Google Deep Mind
Gemini 3.8 Flash

Google Deep Mind · 2026-09-02 · Multimodal model

Google Deep Mind
Gemini 3.7 Flash

Google Deep Mind · 2026-08-13 · Multimodal model

Gemini 3.8 Flash10 wins(91%)(9%)1 winGemini 3.7 Flash

Benchmark scores

Grouped by capability, sorted by largest gap within each. 11 shared benchmarks.

AI Agent - Tool Usage

Gemini 3.8 Flash 5/5
BenchmarkGemini 3.8 FlashGemini 3.7 FlashDiff
BioMysteryBench (Human-difficult)56.501 / 2Thinking Medium (With Tools + Internet)43.502 / 2Thinking Medium (With Tools + Internet)+13
OSWorld 2.0594 / 7Thinking (With Tools)47.907 / 7Thinking Medium (With Tools)+11.10
LABBench286.201 / 2Thinking Medium (With Tools + Internet)82.102 / 2Thinking Medium (With Tools + Internet)+4.10
Terminal-Bench 2.189.401 / 48Thinking (With Tools)85.8010 / 48Thinking (With Tools)+3.60
BioMysteryBench (Human-solvable)88.801 / 2Thinking Medium (With Tools + Internet)87.102 / 2Thinking Medium (With Tools + Internet)+1.70

Multimodal Understanding

Gemini 3.8 Flash 2/3
BenchmarkGemini 3.8 FlashGemini 3.7 FlashDiff
CharXiv RQ86.208 / 19Thinking Medium (No Tools)88.706 / 19Thinking Medium (With Tools)-2.50
LVBench87.801 / 4Thinking Medium (With Tools)85.403 / 4Thinking Medium (No Tools)+2.40
GDP.pdf351 / 2Thinking Medium (No Tools)342 / 2Thinking Medium (No Tools)+1

Coding and Software Engineer

Gemini 3.8 Flash 1/1
BenchmarkGemini 3.8 FlashGemini 3.7 FlashDiff
DeepSWE73.701 / 33Thinking (With Tools)65.3012 / 33Thinking High (With Tools)+8.40

General Knowledge

Gemini 3.8 Flash 1/1
BenchmarkGemini 3.8 FlashGemini 3.7 FlashDiff
HLE-Verified54.901 / 2Thinking Medium (No Tools)53.602 / 2Thinking Medium (No Tools)+1.30

Productivity Knowledge

Gemini 3.8 Flash 1/1
BenchmarkGemini 3.8 FlashGemini 3.7 FlashDiff
GDPval-AA v21,54513 / 24Thinking (No Tools)1,52515 / 24Thinking (No Tools)+20

Specs

FieldGemini 3.8 FlashGemini 3.7 Flash
PublisherGoogle Deep MindGoogle Deep Mind
Release date2026-09-022026-08-13
Model typeMultimodal modelMultimodal model
ArchitectureDenseDense
ParametersNot availableNot available
Context length1M1M
Max output64K64K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemGemini 3.8 FlashGemini 3.7 Flash
Text input$0.75 / 1M tokens$0.75 / 1M tokens
Text output$3.75 / 1M tokens$3.75 / 1M tokens
Cache read$0.075 / 1M tokens$0.075 / 1M tokens

Summary

  • Gemini 3.8 Flashleads in:AI Agent - Tool Usage (5/5), Multimodal Understanding (2/3), Coding and Software Engineer (1/1), General Knowledge (1/1), Productivity Knowledge (1/1)

On average across the 11 shared benchmarks, Gemini 3.8 Flash scores 5.83 higher.

Largest single-benchmark gap: GDPval-AA v2 — Gemini 3.8 Flash 1,545 vs Gemini 3.7 Flash 1,525 (+20).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.