DataLearner 标志

Qwen3.8-MaxvsKimi K3

在 3 个共同 benchmark 中,Kimi K3 整体领先:Qwen3.8-Max 领先 1 项,Kimi K3 领先 2 项,持平 0 项,平均分差 -0.79。

阿里巴巴
Qwen3.8-Max

阿里巴巴 · 2026-08-03 · 推理大模型

Moonshot AI
Kimi K3

Moonshot AI · 2026-07-16 · 推理大模型

Qwen3.8-Max1 项(33%)(67%)2 项Kimi K3

评测分数

按能力类目分组,每组内按分差大小排列;共 3 项。

跨能力综合测试

胶着 2/2
评测项Qwen3.8-MaxKimi K3分差
SuperCLUE71.483 / 13Reported best (effort unspecified)70.685 / 13Reported best (effort unspecified)+0.80
LiveBench78.4620 / 157Reported best (effort unspecified)79.1916 / 157Reported best (effort unspecified)-0.73

跨能力聚合指数

Kimi K3 领先 1/1
评测项Qwen3.8-MaxKimi K3分差
Vals Index v2.148.2728 / 37Reported best (effort unspecified)50.7025 / 37Reported best (effort unspecified)-2.43

规格对比

字段Qwen3.8-MaxKimi K3
发布机构阿里巴巴Moonshot AI
发布时间2026-08-032026-07-16
模型类型推理大模型推理大模型
架构MoE 架构MoE 架构
参数规模2.4万亿2.8万亿
上下文长度1M1M
最大输出128K1M

API 调用价格

价格优先使用 DataLearner 配置的 API 记录;缺失项不做推测。

价格项Qwen3.8-MaxKimi K3
文本输入¥12 / 1M tokens¥20 / 1M tokens
文本输出¥36 / 1M tokens¥100 / 1M tokens
缓存读取¥1.5 / 1M tokens¥2 / 1M tokens

小结

  • Kimi K3在以下类目领先:跨能力聚合指数 (1/1)
  • 胶着类目:跨能力综合测试

3 个共同 benchmark 上,Kimi K3 平均高出 0.79 分。

单项差距最大的 benchmark:Vals Index v2.1 — Qwen3.8-Max 48.27,Kimi K3 50.70(分差 -2.43)。

本页正文由结构化模型、价格与 benchmark 数据生成,不使用实时 LLM 撰写。