DataLearner logo
QW

Qwen2.5-3B

Foundation modelQwen2.5

Qwen2.5-3B

Release date: 2024-09-18Updated: 2024-09-21 11:23:261,315
Parameters
3B
Context length
32K
Chinese support
Supported
Reasoning ability

Qwen2.5-3B is an AI model published by Alibaba, released on 2024-09-18, for Foundation model, with 3B parameters, and 32K context length, requiring about 6GB storage, with a 79.10 score on GSM8K.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

Qwen2.5-3B

Model basics

Reasoning traces
Not supported
Thinking modes
Thinking modes not supported
Context length
32K tokens
Max output length
No data
Model type
Foundation model
Modality (in / out)
No data
Release date
2024-09-18
Model file size
6GB
MoE architecture
No
Total params / Active params
3B / N/A
Knowledge cutoff
No data
Qwen2.5-3B

Open source & experience

Code license
Weights license
- Commercial use permitted
Live demo
No live demo
Qwen2.5-3B

Official resources

Paper
DataLearnerAI blog
No blog post yet
Qwen2.5-3B

API details

API speed
No data
💡Default unit: $/1M tokens. If vendors use other units, follow their published pricing.
Standard
TypeConditionInputOutput
Text-¥0.0003/ 1K¥0.0009/ 1K
Qwen2.5-3B

Benchmark Results

Qwen2.5-3B currently shows benchmark results led by GSM8K (24 / 70, score 79.10), MBPP (30 / 70, score 57.10), HumanEval (58 / 101, score 42.10). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking

General Knowledge

3 evaluations
Benchmark / mode
Score
Rank/total
MMLU
Standard Mode
65.60
84 / 124
BBH
Standard Mode
56.30
17 / 21
MMLU Pro
Standard Mode
34.60
130 / 133

Math and Reasoning

2 evaluations
Benchmark / mode
Score
Rank/total
GSM8K
Standard Mode
79.10
24 / 70
MATH
Standard Mode
42.60
37 / 42

Coding and Software Engineer

2 evaluations
Benchmark / mode
Score
Rank/total
MBPP
Standard Mode
57.10
30 / 70
HumanEval
Standard Mode
42.10
58 / 101

General Evaluation

1 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Standard Mode
24.30
221 / 224

Compare with other models

No curated comparisons for this model yet.

Want a custom combination? Open the compare tool

Qwen2.5-3B

Publisher

Qwen2.5-3B

Model Overview

Qwen2.5-3B is an AI model published by Alibaba, released on 2024-09-18, for Foundation model, with 3B parameters, and 32K context length, requiring about 6GB storage, with a 79.10 score on GSM8K.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code