Qwen3.5-27BvsQwen3-32B
Across 3 shared benchmarks, Qwen3.5-27B leads overall: Qwen3.5-27B wins 3, Qwen3-32B wins 0, with 0 ties and an average score difference of +18.27.
Qwen3.5-27B
阿里巴巴 · 2026-02-25 · Reasoning model
Qwen3-32B
阿里巴巴 · 2025-04-28 · Reasoning model
Qwen3.5-27B3 wins(100%)(0%)0 winsQwen3-32B
Benchmark scores
Grouped by capability, sorted by largest gap within each. 3 shared benchmarks.
General Evaluation
Qwen3.5-27B 1/1| Benchmark | Qwen3.5-27B | Qwen3-32B | Diff |
|---|---|---|---|
| GPQA Diamond | 84.20174 / 462Normal (No Tools) | 54.60401 / 462Normal (No Tools) | +29.60 |
General Knowledge
Qwen3.5-27B 1/1| Benchmark | Qwen3.5-27B | Qwen3-32B | Diff |
|---|---|---|---|
| HLE | 13.90348 / 563Normal (No Tools) · Text only | 4.10519 / 563Normal (No Tools) · Text only | +9.80 |
Instruction Following
Qwen3.5-27B 1/1| Benchmark | Qwen3.5-27B | Qwen3-32B | Diff |
|---|---|---|---|
| IF Bench | 46.90171 / 282Normal (No Tools) | 31.50262 / 282Normal (No Tools) | +15.40 |
Specs
| Field | Qwen3.5-27B | Qwen3-32B |
|---|---|---|
| Publisher | 阿里巴巴 | 阿里巴巴 |
| Release date | 2026-02-25 | 2025-04-28 |
| Model type | Reasoning model | Reasoning model |
| Architecture | Dense | Dense |
| Parameters | 27B | 32B |
| Context length | 1010K | 128K |
| Max output | 248320 | 16K |
API pricing
Prices use DataLearner records when available; missing fields are not inferred.
| Item | Qwen3.5-27B | Qwen3-32B |
|---|---|---|
| Text input | Not public | ¥0.0012 / 1K tokens |
| Text output | Not public | ¥0.0048 / 1K tokens |
One or both models have incomplete public pricing.
Summary
- Qwen3.5-27Bleads in:General Evaluation (1/1), General Knowledge (1/1), Instruction Following (1/1)
On average across the 3 shared benchmarks, Qwen3.5-27B scores 18.27 higher.
Largest single-benchmark gap: GPQA Diamond — Qwen3.5-27B 84.20 vs Qwen3-32B 54.60 (+29.60).
Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.