DataLearner logo

Qwen3.6-27BvsHaiku 4.5

Across 12 shared benchmarks, Qwen3.6-27B leads overall: Qwen3.6-27B wins 10, Haiku 4.5 wins 2, with 0 ties and an average score difference of +15.29.

阿里巴巴
Qwen3.6-27B

阿里巴巴 · 2026-04-22 · Reasoning model

Anthropic
Haiku 4.5

Anthropic · 2025-10-15 · Multimodal model

Qwen3.6-27B10 wins(83%)(17%)2 winsHaiku 4.5

Benchmark scores

Grouped by capability, sorted by largest gap within each. 12 shared benchmarks.

Agent Level Benchmark

Even 2/2
BenchmarkQwen3.6-27BHaiku 4.5Diff
τ²-Bench - Telecom93.6046 / 264Normal (With Tools)32.50204 / 264Normal (With Tools)+61.10
Terminal Bench Hard21.20145 / 244Normal (With Tools)27.30120 / 244Normal (With Tools)-6.10

Multimodal Understanding

Qwen3.6-27B 2/2
BenchmarkQwen3.6-27BHaiku 4.5Diff
MMMU-Pro71.70118 / 227Normal (No Tools)55.10188 / 227Normal (No Tools)+16.60
GDP.pdf1176 / 118Thinking (No Tools)3.8099 / 118Thinking (No Tools)+7.20

Claw-style Agent Evaluation

Haiku 4.5 1/1
BenchmarkQwen3.6-27BHaiku 4.5Diff
Claw Bench72.4027 / 29Thinking (With Tools)89.4011 / 29Thinking (With Tools)-17

Coding and Software Engineer

Qwen3.6-27B 1/1
BenchmarkQwen3.6-27BHaiku 4.5Diff
SciCode42.8096 / 130Thinking (No Tools)42.2098 / 130Thinking (No Tools)+0.60

General Evaluation

Qwen3.6-27B 1/1
BenchmarkQwen3.6-27BHaiku 4.5Diff
GPQA Diamond84.85159 / 462Normal (No Tools)60.50375 / 462Normal (No Tools)+24.35

General Knowledge

Qwen3.6-27B 1/1
BenchmarkQwen3.6-27BHaiku 4.5Diff
LiveBench64.0354 / 117Normal (No Tools)45.33105 / 117Normal (No Tools)+18.70

Instruction Following

Qwen3.6-27B 1/1
BenchmarkQwen3.6-27BHaiku 4.5Diff
IF Bench45.70177 / 282Normal (No Tools)42202 / 282Normal (No Tools)+3.70

Long Context

Qwen3.6-27B 1/1
BenchmarkQwen3.6-27BHaiku 4.5Diff
AA-LCR66.70118 / 170Normal (No Tools)49.70139 / 170Normal (No Tools)+17

Productivity Knowledge

Qwen3.6-27B 1/1
BenchmarkQwen3.6-27BHaiku 4.5Diff
Harvey Lab-AA82.3027 / 43Thinking (With Tools)61.0536 / 43Thinking (With Tools)+21.25

Text Embedding

Qwen3.6-27B 1/1
BenchmarkQwen3.6-27BHaiku 4.5Diff
Context Arena53.7882 / 126Normal (No Tools)17.68124 / 126Normal (No Tools)+36.10

Specs

FieldQwen3.6-27BHaiku 4.5
Publisher阿里巴巴Anthropic
Release date2026-04-222025-10-15
Model typeReasoning modelMultimodal model
ArchitectureDenseDense
Parameters27BNot available
Context length128K200K
Max output16K64K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemQwen3.6-27BHaiku 4.5
Text inputNot public$1 / 1M tokens
Text outputNot public$5 / 1M tokens
Cache readNot public$0.1 / 1M tokens
Cache writeNot public$1.25 / 1M tokens

One or both models have incomplete public pricing.

Summary

  • Qwen3.6-27Bleads in:Multimodal Understanding (2/2), Coding and Software Engineer (1/1), General Evaluation (1/1), General Knowledge (1/1), Instruction Following (1/1), Long Context (1/1), Productivity Knowledge (1/1), Text Embedding (1/1)
  • Haiku 4.5leads in:Claw-style Agent Evaluation (1/1)
  • Tied in:Agent Level Benchmark

On average across the 12 shared benchmarks, Qwen3.6-27B scores 15.29 higher.

Largest single-benchmark gap: τ²-Bench - Telecom — Qwen3.6-27B 93.60 vs Haiku 4.5 32.50 (+61.10).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.