DataLearner logo
PH

Phi-4-mini-instruct (3.8B)

Chat modelPhi-4

Phi-4-mini-instruct (3.8B)

Release date: 2025-02-27Updated: 2025-02-28Views: 1,144
Live demoGitHubHugging FaceCompare
Parameters
3.8B
Context length
128K
Multilingual
Supported
Reasoning ability
No data

Phi-4-mini-instruct (3.8B) is an AI model published by Microsoft Azure, released on 2025-02-27, for Chat model, with 3.8B parameters, and 128K context length, requiring about 7.67GB storage, with a 88.60 score on GSM8K.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

Phi-4-mini-instruct (3.8B)

Model basics

Reasoning traces
No data
Thinking modes
Thinking modes not supported
Context length
128K tokens
Max output length
No data
Model type
Chat model
Modality (in / out)
Text → Text
Release date
2025-02-27
Model file size
7.67GB
MoE architecture
No
Total params / Active params
3.8B / Not applicable
Knowledge cutoff
No data
Phi-4-mini-instruct (3.8B)

Open source & experience

Code license
Weights license
MIT License- Commercial use permitted
GitHub repo
N/A
Live demo
N/A
Phi-4-mini-instruct (3.8B)

Official resources

Paper
DataLearnerAI blog
Phi-4-mini-instruct (3.8B)

API details

API speed
No data
No public API pricing yet.
Phi-4-mini-instruct (3.8B)

Benchmark Results

Phi-4-mini-instruct (3.8B) currently shows benchmark results led by GSM8K (16 / 70, score 88.60), HumanEval (47 / 140, score 74.40), MBPP (45 / 96, score 65.30). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage

General Knowledge

3 evaluations
Benchmark / mode
Score
Rank/total
MMLU
Standard Mode
67.30
80 / 124
MMLU Pro
Standard Mode
52.80
134 / 176
HLE
Standard Mode
4.50
500 / 563

Math and Reasoning

5 evaluations
Benchmark / mode
Score
Rank/total
GSM8K
Standard Mode
88.60
16 / 70
MATH-500
Standard Mode
71.80
45 / 45
MATH
Standard Mode
64
27 / 42
AIME 2024
Standard Mode
10
60 / 62
AIME2025
Standard Mode
6.70
211 / 215

Coding and Software Engineer

3 evaluations
Benchmark / mode
Score
Rank/total
HumanEval
Standard Mode
74.40
47 / 140
MBPP
Standard Mode
65.30
45 / 96
LiveCodeBench
Standard Mode
12.60
244 / 250

General Evaluation

1 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Standard Mode
36
439 / 462

Agent Level Benchmark

1 evaluations
Benchmark / mode
Score
Rank/total
τ²-Bench - Telecom
Standard ModeTools
8.20
264 / 264

Instruction Following

1 evaluations
Benchmark / mode
Score
Rank/total
IF Bench
Standard Mode
21.10
278 / 282

AI Agent - Tool Usage

1 evaluations
Benchmark / mode
Score
Rank/total
Terminal-Bench 2.1
Standard ModeTools
0.40
191 / 192
Phi-4-mini-instruct (3.8B)

Publisher

Phi-4-mini-instruct (3.8B)

Model Overview

Phi-4-mini-instruct (3.8B) is an AI model published by Microsoft Azure, released on 2025-02-27, for Chat model, with 3.8B parameters, and 128K context length, requiring about 7.67GB storage, with a 88.60 score on GSM8K.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code