DataLearner logo

Kimi K2.6vsQwen3.6-Max-Preview

Across 3 shared benchmarks, Kimi K2.6 leads overall: Kimi K2.6 wins 2, Qwen3.6-Max-Preview wins 1, with 0 ties and an average score difference of -4.81.

Moonshot AI
Kimi K2.6

Moonshot AI · 2026-04-20 · Reasoning model

阿里巴巴
Qwen3.6-Max-Preview

阿里巴巴 · 2026-04-18 · Chat model

Kimi K2.62 wins(67%)(33%)1 winQwen3.6-Max-Preview

Benchmark scores

Grouped by capability, sorted by largest gap within each. 3 shared benchmarks.

Coding and Software Engineer

Kimi K2.6 2/2
BenchmarkKimi K2.6Qwen3.6-Max-PreviewDiff
SWE-bench Multilingual76.708 / 29Thinking (With Tools)73.8013 / 29Thinking (With Tools)+2.90
SWE-bench Verified80.2014 / 116Thinking (With Tools)78.8021 / 116Thinking (With Tools)+1.40

Text Embedding

Qwen3.6-Max-Preview 1/1
BenchmarkKimi K2.6Qwen3.6-Max-PreviewDiff
Context Arena51.8886 / 126Normal (No Tools)70.6261 / 126Normal (No Tools)-18.74

Specs

FieldKimi K2.6Qwen3.6-Max-Preview
PublisherMoonshot AI阿里巴巴
Release date2026-04-202026-04-18
Model typeReasoning modelChat model
ArchitectureMoEDense
Parameters1TNot available
Context length256K262K
Max outputNot available64K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemKimi K2.6Qwen3.6-Max-Preview
Text input$0.95 / 1M tokens$1.3 / 1M tokens
Text output$4 / 1M tokens$7.8 / 1M tokens
Cache read$0.16 / 1M tokensNot public
Cache write$0.95 / 1M tokensNot public

Summary

  • Kimi K2.6leads in:Coding and Software Engineer (2/2)
  • Qwen3.6-Max-Previewleads in:Text Embedding (1/1)

On average across the 3 shared benchmarks, Qwen3.6-Max-Preview scores 4.81 higher.

Largest single-benchmark gap: Context Arena — Kimi K2.6 51.88 vs Qwen3.6-Max-Preview 70.62 (-18.74).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.