DataLearner logo
QW

Qwen3-235B-A22B

Reasoning modelQwen3

Qwen3-235B-A22B

Release date: 2025-04-28Updated: 2026-07-17Views: 2,797
Parameters
235B
Context length
128K
Multilingual
Supported
Reasoning ability
3/5

Qwen3-235B-A22B is a reasoning model from Alibaba, released on 2025-04-28. It accepts text input and returns text output. The cataloged parameter count is 235B, with 22B active parameters per inference. The recorded context window is 128K, and the recorded maximum output is 16K. Cataloged capabilities include Reasoning model and Multilingual. The checkpoint is listed under the Apache 2.0 license, with the license link included in the model references. Use the linked references to confirm current access, licensing, and provider-specific limits.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

Qwen3-235B-A22B

Model basics

Reasoning traces
Supported
Thinking modes
Standard Mode (Default)Thinking Mode
Context length
128K tokens
Max output length
16K tokens
Model type
Reasoning model
Modality (in / out)
Text → Text
Release date
2025-04-28
Model file size
470GB
MoE architecture
Yes
Total params / Active params
235B / 22B
Knowledge cutoff
No data
Qwen3-235B-A22B

Open source & experience

Code license
Weights license
Apache 2.0- Commercial use permitted
Live demo
Qwen3-235B-A22B

Official resources

Paper
DataLearnerAI blog
Qwen3-235B-A22B

API details

API speed
3/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-¥0.0020/ 1K tokens¥0.0080/ 1K tokens

“—” means the modality is not billed in that direction, or the vendor has not published a price for it.

Qwen3-235B-A22B

Benchmark Results

Qwen3-235B-A22B currently shows benchmark results led by GSM8K (2 / 70, score 96.40), MATH-500 (7 / 45, score 98), MMLU (38 / 124, score 85.80). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage

General Knowledge

7 evaluations
Benchmark / mode
Score
Rank/total
BBH
Standard Mode
88.87
8 / 21
MMLU
Standard Mode
85.80
38 / 124
MMLU Pro
Standard Mode
72.90
93 / 176
HLE
Standard Mode
7.60
427 / 563
HLE
Standard Mode
4.20
513 / 563
HLE
Thinking Mode
11
377 / 563
ARC-AGI-1
Standard Mode
4.30
145 / 147

Math and Reasoning

7 evaluations
Benchmark / mode
Score
Rank/total
MATH-500
Standard Mode
96.20
19 / 45
MATH-500
Thinking Mode
98
7 / 45
GSM8K
Standard Mode
96.40
2 / 70
AIME 2024
Standard Mode
85.70
20 / 62
AIME 2024
Thinking Mode
85.70
20 / 62
AIME2025
Standard Mode
24.70
193 / 215
AIME2025
Thinking Mode
81.50
92 / 215

Reading Comprehension

1 evaluations
Benchmark / mode
Score
Rank/total
DROP
Standard Mode
88.70
5 / 9

General Evaluation

2 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Standard Mode
71.10
314 / 462
GPQA Diamond
Thinking Mode
71.10
314 / 462

Common Sense

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleQA
Standard Mode
11
40 / 47

Coding and Software Engineer

3 evaluations
Benchmark / mode
Score
Rank/total
LiveCodeBench
Standard Mode
70.70
86 / 250
LiveCodeBench
Thinking Mode
70.70
86 / 250
SWE-bench Verified
Standard Mode
34.40
109 / 116

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleBench
Thinking Mode
31
74 / 93

Agent Level Benchmark

6 evaluations
Benchmark / mode
Score
Rank/total
Aider-Polyglot
Standard Mode
59.60
23 / 59
τ²-Bench
Thinking ModeTools
34.40
42 / 44
τ²-Bench - Telecom
Standard ModeTools
27.20
222 / 264
τ²-Bench - Telecom
Thinking ModeTools
24
233 / 264
Terminal Bench Hard
Standard ModeTools
6.10
204 / 244
Terminal Bench Hard
Thinking ModeTools
6.10
204 / 244

Instruction Following

2 evaluations
Benchmark / mode
Score
Rank/total
IF Bench
Standard Mode
36.60
237 / 282
IF Bench
Thinking Mode
38.70
227 / 282
Qwen3-235B-A22B

Publisher

Qwen3-235B-A22B

Model Overview

Qwen3-235B-A22B is a reasoning model from Alibaba, released on 2025-04-28.

It accepts text input and produces text output. Its cataloged capabilities include Reasoning model and Multilingual. The cataloged parameter count is 235B, with 22B active parameters per inference. The recorded context window is 128K, and the recorded maximum output is 16K.

The checkpoint is listed under the Apache 2.0 license, with the license link included in the model references. The page records 2 API pricing rules from the listed provider; current provider pricing and conditions should be checked before deployment. The evaluation section contains 24 cataloged benchmark results with their recorded modes and scores. The page links 5 release, model-card, repository, or provider references for checking the underlying claims. Specifications, availability, and prices can change; undisclosed values are intentionally left unstated.

Qwen3-235B-A22B

FAQ

What is Qwen3-235B-A22B?

Qwen3-235B-A22B is a reasoning model from Alibaba, released on 2025-04-28. It accepts text input and returns text output. The cataloged parameter count is 235B, with 22B active parameters per inference. The recorded context window is 128K, and the recorded maximum output is 16K. Cataloged capabilities include Reasoning model and Multilingual. The checkpoint is listed under the Apache 2.0 license, with the license link included in the model references. Use the linked references to confirm current access, licensing, and provider-specific limits.

What input and output modalities does Qwen3-235B-A22B support?

The current model record lists text as input and text as output.

What are the main recorded specifications for Qwen3-235B-A22B?

The cataloged parameter count is 235B, with 22B active parameters per inference. The recorded context window is 128K, and the recorded maximum output is 16K. Fields without a source-backed value remain undisclosed.

Does Qwen3-235B-A22B have API pricing?

The page records 2 API pricing rules from the listed provider; current provider pricing and conditions should be checked before deployment.

Are benchmark results available for Qwen3-235B-A22B?

The evaluation section contains 24 cataloged benchmark results with their recorded modes and scores. Compare only results that use the same benchmark version and evaluation mode.

Is Qwen3-235B-A22B open source?

The checkpoint is listed under the Apache 2.0 license, with the license link included in the model references. Review the linked license text before commercial or derivative use.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code