DataLearner logo

Qwen3.6-27BvsGemini 3.0 Flash

Qwen3.6-27B and Gemini 3.0 Flash are tied across 10 shared benchmarks: Qwen3.6-27B leads on 5, Gemini 3.0 Flash leads on 5, with 0 ties and an average score difference of +3.27.

阿里巴巴
Qwen3.6-27B

阿里巴巴 · 2026-04-22 · Reasoning model

Google Deep Mind
Gemini 3.0 Flash

Google Deep Mind · 2025-12-17 · Chat model

Qwen3.6-27B5 wins(50%)(50%)5 winsGemini 3.0 Flash

Benchmark scores

Grouped by capability, sorted by largest gap within each. 10 shared benchmarks.

General Knowledge

Qwen3.6-27B 2/3
BenchmarkQwen3.6-27BGemini 3.0 FlashDiff
LiveBench64.0354 / 117Normal (No Tools)56.3581 / 117Normal (No Tools)+7.68
CritPt0.90158 / 200Normal (No Tools)1.40135 / 200Normal (No Tools)-0.50
HLE15.10333 / 563Normal (No Tools) · Text only15334 / 563Normal (No Tools) · Text only+0.10

Agent Level Benchmark

Even 2/2
BenchmarkQwen3.6-27BGemini 3.0 FlashDiff
τ²-Bench - Telecom93.6046 / 264Normal (With Tools)43.30186 / 264Normal (With Tools)+50.30
Terminal Bench Hard21.20145 / 244Normal (With Tools)31.80100 / 244Normal (With Tools)-10.60

AI Agent - Tool Usage

Qwen3.6-27B 1/1
BenchmarkQwen3.6-27BGemini 3.0 FlashDiff
Terminal Bench 2.059.3020 / 48Thinking (With Tools)47.6039 / 48Thinking (With Tools)+11.70

Claw-style Agent Evaluation

Gemini 3.0 Flash 1/1
BenchmarkQwen3.6-27BGemini 3.0 FlashDiff
Claw Bench72.4027 / 29Thinking (With Tools)85.7015 / 29Thinking (With Tools)-13.30

General Evaluation

Qwen3.6-27B 1/1
BenchmarkQwen3.6-27BGemini 3.0 FlashDiff
GPQA Diamond84.85159 / 462Normal (No Tools)81.20215 / 462Normal (No Tools)+3.65

Instruction Following

Gemini 3.0 Flash 1/1
BenchmarkQwen3.6-27BGemini 3.0 FlashDiff
IF Bench45.70177 / 282Normal (No Tools)55.10133 / 282Normal (No Tools)-9.40

Multimodal Understanding

Gemini 3.0 Flash 1/1
BenchmarkQwen3.6-27BGemini 3.0 FlashDiff
MMMU-Pro71.70118 / 227Normal (No Tools)78.6055 / 227Normal (No Tools)-6.90

Specs

FieldQwen3.6-27BGemini 3.0 Flash
Publisher阿里巴巴Google Deep Mind
Release date2026-04-222025-12-17
Model typeReasoning modelChat model
ArchitectureDenseDense
Parameters27BNot available
Context length128K2000K
Max output16K64K

API pricing

Prices use DataLearner records when available; missing fields are not inferred.

ItemQwen3.6-27BGemini 3.0 Flash
Text inputNot public$0.5 / 1M tokens
Text outputNot public$3 / 1M tokens

One or both models have incomplete public pricing.

Summary

  • Qwen3.6-27Bleads in:General Knowledge (2/3), AI Agent - Tool Usage (1/1), General Evaluation (1/1)
  • Gemini 3.0 Flashleads in:Claw-style Agent Evaluation (1/1), Instruction Following (1/1), Multimodal Understanding (1/1)
  • Tied in:Agent Level Benchmark

On average across the 10 shared benchmarks, Qwen3.6-27B scores 3.27 higher.

Largest single-benchmark gap: τ²-Bench - Telecom — Qwen3.6-27B 93.60 vs Gemini 3.0 Flash 43.30 (+50.30).

Page generated from structured model, pricing and benchmark records. No real-time LLM is used to write the prose.