Mistral Medium 3.5vsGemma 4 31B
Across 3 shared benchmarks, Gemma 4 31B leads overall: Mistral Medium 3.5 wins 1, Gemma 4 31B wins 2, with 0 ties and an average score difference of +4.25.
Mistral Medium 3.5
MistralAI · 2026-05-01 · Chat model
Gemma 4 31B
DeepMind · 2026-04-02 · Chat model
Mistral Medium 3.51 win(33%)(67%)2 winsGemma 4 31B
Benchmark scores
Grouped by capability, sorted by largest gap within each. 3 shared benchmarks.
Coding and Software Engineer
Gemma 4 31B 1/1| Benchmark | Mistral Medium 3.5 | Gemma 4 31B | Diff |
|---|---|---|---|
| SciCode | 39.58104 / 130Thinking (No Tools) | 45.5087 / 130Thinking (No Tools) | -5.92 |
Multimodal Understanding
Gemma 4 31B 1/1| Benchmark | Mistral Medium 3.5 | Gemma 4 31B | Diff |
|---|---|---|---|
| GDP.pdf | 2.80102 / 118Thinking (No Tools) | 691 / 118Thinking (No Tools) | -3.20 |
Productivity Knowledge
Mistral Medium 3.5 1/1| Benchmark | Mistral Medium 3.5 | Gemma 4 31B | Diff |
|---|---|---|---|
| Harvey Lab-AA | 69.1034 / 43Thinking (With Tools) | 47.2340 / 43Thinking (With Tools) | +21.87 |
Specs
| Field | Mistral Medium 3.5 | Gemma 4 31B |
|---|---|---|
| Publisher | MistralAI | DeepMind |
| Release date | 2026-05-01 | 2026-04-02 |
| Model type | Chat model | Chat model |
| Architecture | Dense | Dense |
| Parameters | Not available | 30.7B |
| Context length | Not available | 256K |
| Max output | Not available | 32K |
Summary
- Mistral Medium 3.5leads in:Productivity Knowledge (1/1)
- Gemma 4 31Bleads in:Coding and Software Engineer (1/1), Multimodal Understanding (1/1)
On average across the 3 shared benchmarks, Mistral Medium 3.5 scores 4.25 higher.
Largest single-benchmark gap: Harvey Lab-AA — Mistral Medium 3.5 69.10 vs Gemma 4 31B 47.23 (+21.87).
Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.