DataLearner logo
QW

Qwen3-4B-Thinking-2507

Chat modelQwen3

Qwen3-4B-Thinking-2507

Release date: 2025-08-06Updated: 2025-08-07Views: 993
Parameters
4B
Context length
256K
Multilingual
Supported
Reasoning ability
2/5

Qwen3-4B-Thinking-2507 is an AI model published by Alibaba, released on 2025-08-06, for Chat model, with 4B parameters, and 256K context length, requiring about 8.05GB storage, with a 81.30 score on AIME2025.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

Qwen3-4B-Thinking-2507

Model basics

Reasoning traces
Supported
Thinking modes
Thinking Mode (Default)
Context length
256K tokens
Max output length
16K tokens
Model type
Chat model
Modality (in / out)
Text → Text
Release date
2025-08-06
Model file size
8.05GB
MoE architecture
No
Total params / Active params
4B / Not applicable
Knowledge cutoff
No data
Qwen3-4B-Thinking-2507

Open source & experience

Code license
Weights license
Apache 2.0- Commercial use permitted
Live demo
Qwen3-4B-Thinking-2507

Official resources

Paper
DataLearnerAI blog
N/A
Qwen3-4B-Thinking-2507

API details

API speed
4/5
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-$0.110/ 1M$1.26/ 1M

“—” means the modality is not billed in that direction, or the vendor has not published a price for it.

Qwen3-4B-Thinking-2507

Benchmark Results

Qwen3-4B-Thinking-2507 currently shows benchmark results led by AIME2025 (55 / 106, score 81.30), LiveCodeBench (87 / 127, score 55.20), GPQA Diamond (210 / 270, score 65.80). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking

General Evaluation

1 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Thinking Mode
65.80
210 / 270

Coding and Software Engineer

1 evaluations
Benchmark / mode
Score
Rank/total
LiveCodeBench
Thinking Mode
55.20
87 / 127

Math and Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
AIME2025
Thinking Mode
81.30
55 / 106
Qwen3-4B-Thinking-2507

Publisher

Qwen3-4B-Thinking-2507

Model Overview

Qwen3-4B-Thinking-2507 is an AI model published by Alibaba, released on 2025-08-06, for Chat model, with 4B parameters, and 256K context length, requiring about 8.05GB storage, with a 81.30 score on AIME2025.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code