DataLearner logo
NE

Nemotron 3 Ultra

Reasoning modelLong context

NVIDIA Nemotron 3 Ultra 550B-A55B

Release date: 2026-06-04Updated: 2026-08-23Knowledge cutoff: 2026-05Views: 927
Parameters
550B
Context length
1M
Multilingual
Supported
Reasoning ability
5/5

Nemotron 3 Ultra is a reasoning model from NVIDIA, released on 2026-06-04. It accepts text input and returns text output. The cataloged parameter count is 550B, with 55B active parameters per inference. The recorded context window is 1M. Cataloged capabilities include Reasoning model and Multilingual. The checkpoint is listed under the Free Commercial license. Use the linked references to confirm current access, licensing, and provider-specific limits.

Data sourced primarily from official releases (GitHub, Hugging Face, papers), then benchmark leaderboards, then third-party evaluators. Learn about our data methodology

Nemotron 3 Ultra

Model basics

Reasoning traces
Supported
Thinking modes
Thinking Mode (Default)Standard Mode
Context length
1M tokens
Max output length
No data
Model type
Reasoning model
Modality (in / out)
Text → Text
Release date
2026-06-04
Model file size
No data
MoE architecture
Yes
Total params / Active params
550B / 55B
Knowledge cutoff
2026-05
Nemotron 3 Ultra

Open source & experience

Code license
Free commercial use
Weights license
Free commercial use- Commercial use permitted
Nemotron 3 Ultra

Official resources

Paper
DataLearnerAI blog
N/A
Nemotron 3 Ultra

API details

API speed
4/5
No public API pricing yet.
Nemotron 3 Ultra

Benchmark Results

Nemotron 3 Ultra currently shows benchmark results led by IF Bench (4 / 282, score 81.70), IMO-AnswerBench (1 / 24, score 92.30), Pinch Bench (2 / 38, score 90). This page also consolidates core specs, context limits, and API pricing so you can evaluate the model from benchmark results and deployment constraints together.

Thinking
Tool usage
Internet

General Knowledge

8 evaluations
Benchmark / mode
Score
Rank/total
GPQA
Thinking Mode
87
1 / 17
MMLU-Pro
Thinking Mode
86.80
17 / 176
LiveBench
Standard Mode
51.78
90 / 117
HLE
Thinking Mode
28.40
222 / 565
HLE
Thinking Mode
26.70
236 / 565
HLE
Thinking ModeTools
37.40
154 / 565
CritPt
Thinking Mode
3.10
111 / 201

General Evaluation

1 evaluations
Benchmark / mode
Score
Rank/total
GPQA Diamond
Thinking Mode
86.70
137 / 463

Coding and Software Engineer

4 evaluations
Benchmark / mode
Score
Rank/total
LiveCodeBench
Thinking Mode
89
15 / 251
SWE-bench Verified
Thinking ModeTools
70.70
58 / 116
SWE-bench Multilingual
Thinking ModeTools
67.70
26 / 30
SciCode
Thinking Mode
44.60
93 / 131

Writing and Creative Capabilities

1 evaluations
Benchmark / mode
Score
Rank/total
Creative Writing
Standard Mode
1689.30
27 / 106

Common Sense Reasoning

1 evaluations
Benchmark / mode
Score
Rank/total
SimpleBench
Thinking Mode
41.70
62 / 93

Agent Level Benchmark

3 evaluations
Benchmark / mode
Score
Rank/total
τ²-Bench - Telecom
Thinking ModeTools
83.30
102 / 264
Terminal Bench Hard
Thinking ModeTools
36.40
67 / 244
τ³-Banking
Thinking ModeTools
22.60
85 / 167

Instruction Following

1 evaluations
Benchmark / mode
Score
Rank/total
IF Bench
Thinking Mode
81.70
4 / 282

AI Agent - Information Search

1 evaluations
Benchmark / mode
Score
Rank/total
BrowseComp
Thinking ModeToolsInternet
44.40
49 / 58

Math and Reasoning

2 evaluations
Benchmark / mode
Score
Rank/total
IMO-AnswerBench
Thinking Mode
88.60
6 / 24
IMO-AnswerBench
Thinking ModeTools
92.30
1 / 24

Long Context

2 evaluations
Benchmark / mode
Score
Rank/total
AA-LCR
Thinking Mode
79.30
54 / 171
LongBench v2
Standard Mode
61.90
5 / 14

Claw-style Agent Evaluation

2 evaluations
Benchmark / mode
Score
Rank/total
Pinch Bench
Thinking ModeTools
90
2 / 38
89.92
3 / 45

AI Agent - Tool Usage

2 evaluations
Benchmark / mode
Score
Rank/total
Terminal-Bench 2.1
Thinking ModeTools
56.40
122 / 196
Terminal-Bench 4.0
Thinking ModeTools
0.50
86 / 91

Productivity Knowledge

4 evaluations
Benchmark / mode
Score
Rank/total
GDPval-AA v2
Thinking ModeTools
1091
84 / 107
AA-Briefcase
Thinking ModeTools
878
71 / 85
Harvey Lab-AA
Thinking ModeTools
81.72
30 / 44
AA-AnalystAgent
Thinking ModeToolsInternet
6.25
29 / 29

Multimodal Understanding

1 evaluations
Benchmark / mode
Score
Rank/total
GDP.pdf
Thinking Mode
5
97 / 119

Compare with other models

Nemotron 3 Ultra

Publisher

NVIDIA Nemotron 3 Ultra 550B-A55B

Model Overview

Nemotron 3 Ultra is a reasoning model from NVIDIA, released on 2026-06-04.

It accepts text input and produces text output. Its cataloged capabilities include Reasoning model and Multilingual. The cataloged parameter count is 550B, with 55B active parameters per inference. The recorded context window is 1M.

The checkpoint is listed under the Free Commercial license. The evaluation section contains 2 cataloged benchmark results with their recorded modes and scores. The page links 4 release, model-card, repository, or provider references for checking the underlying claims. Specifications, availability, and prices can change; undisclosed values are intentionally left unstated.

Nemotron 3 Ultra

FAQ

What is Nemotron 3 Ultra?

Nemotron 3 Ultra is a reasoning model from NVIDIA, released on 2026-06-04. It accepts text input and returns text output. The cataloged parameter count is 550B, with 55B active parameters per inference. The recorded context window is 1M. Cataloged capabilities include Reasoning model and Multilingual. The checkpoint is listed under the Free Commercial license. Use the linked references to confirm current access, licensing, and provider-specific limits.

What input and output modalities does Nemotron 3 Ultra support?

The current model record lists text as input and text as output.

What are the main recorded specifications for Nemotron 3 Ultra?

The cataloged parameter count is 550B, with 55B active parameters per inference. The recorded context window is 1M. Fields without a source-backed value remain undisclosed.

Why is no API price shown for Nemotron 3 Ultra?

Open-weight checkpoint; NVIDIA does not publish a model-specific first-party token price for self-hosting this checkpoint.

Are benchmark results available for Nemotron 3 Ultra?

The evaluation section contains 2 cataloged benchmark results with their recorded modes and scores. Compare only results that use the same benchmark version and evaluation mode.

Is Nemotron 3 Ultra open source?

The checkpoint is listed under the Free Commercial license. Review the linked license text before commercial or derivative use.

DataLearner on WeChat

Follow DataLearner on WeChat for AI model updates and research notes.

DataLearner WeChat QR code