Gemini 3.5 Flash-LitevsDeepSeek-V4-Flash
Gemini 3.5 Flash-Lite 与 DeepSeek-V4-Flash 在 1 个同模式 benchmark 中整体接近:Gemini 3.5 Flash-Lite 领先 0 项,DeepSeek-V4-Flash 领先 0 项,持平 1 项,平均分差 0。另有 11 项测试模式不同,仅供参考。
Gemini 3.5 Flash-Lite
Google DeepMind · 2026-07-21 · 多模态大模型
DeepSeek-V4-Flash
DeepSeek-AI · 2026-04-24 · 推理大模型
Gemini 3.5 Flash-Lite0 项(0%)持平1(0%)0 项DeepSeek-V4-Flash
评测分数
按能力类目分组,每组内按分差大小排列;共 1 项。
写作与创意表达
胶着 1/1| 评测项 | Gemini 3.5 Flash-Lite | DeepSeek-V4-Flash | 分差 |
|---|---|---|---|
| Creative Writing | 1,55948 / 110Normal (No Tools) | 1,55948 / 110Normal (No Tools) | 持平 |
测试模式不同的成绩
共 11 项。两款模型的公开成绩来自不同测试模式,分数不直接可比,不计入胜负和平均分差。
| 评测项 | Gemini 3.5 Flash-Lite | DeepSeek-V4-Flash |
|---|---|---|
| Terminal-Bench 2.1 | 54Thinking (With Tools) | 61.80Extra-High (With Tools) |
| Terminal-Bench 4.0 | 1Thinking (With Tools) | 3Thinking High (With Tools) |
| LiveBench | 63.94Thinking High (No Tools) | 65.48Normal (No Tools) |
| GDP.pdf | 13.60Thinking (No Tools) | 10.80Max (No Tools) |
| Context Arena | 72.57Thinking High (No Tools) | 26.47Normal (No Tools) |
| APEX-SWE | 31.20Thinking High (With Tools) | 46.90Max (With Tools) |
| SWE-Bench Pro - Public | 54.20Thinking (With Tools) | 49.10Normal (With Tools) |
| SciCode | 41.30Thinking (No Tools) | 45.30Max (No Tools) |
| GPQA Diamond | 83.33Thinking High (No Tools) | 71.20Normal (No Tools) |
| τ³-Banking | 17.50Thinking (With Tools) | 30.90Max (With Tools) |
| Toolathlon-Verified | 57.10Thinking High (With Tools) | 49.70Max (With Tools) |
规格对比
| 字段 | Gemini 3.5 Flash-Lite | DeepSeek-V4-Flash |
|---|---|---|
| 发布机构 | Google DeepMind | DeepSeek-AI |
| 发布时间 | 2026-07-21 | 2026-04-24 |
| 模型类型 | 多模态大模型 | 推理大模型 |
| 架构 | 稠密模型 | MoE 架构 |
| 参数规模 | 暂无数据 | 2840亿 |
| 上下文长度 | 1M | 1M |
| 最大输出 | 64K | 384K |
API 调用价格
价格优先使用 DataLearner 配置的 API 记录;缺失项不做推测。
| 价格项 | Gemini 3.5 Flash-Lite | DeepSeek-V4-Flash |
|---|---|---|
| 文本输入 | $0.3 / 1M tokens | $0.14 / 1M tokens |
| 文本输出 | $2.5 / 1M tokens | $0.28 / 1M tokens |
| 缓存读取 | $0.03 / 1M tokens | $0.0028 / 1M tokens |
小结
1 个同模式 benchmark 上,两款模型平均分差接近。
单项差距最大的 benchmark:Creative Writing — Gemini 3.5 Flash-Lite 1,559,DeepSeek-V4-Flash 1,559(分差 0)。
本页正文由结构化模型、价格与 benchmark 数据生成,不使用实时 LLM 撰写。