Nemotron 3 Ultra Benchmark Details
Nemotron 3 Ultra currently shows benchmark results led by IF Bench (4 / 282, score 81.70), IMO-AnswerBench (1 / 24, score 92.30), Pinch Bench (2 / 38, score 90). This page also compares it with 3 competitor models, including performance and pricing views when available.
Benchmark Results
Benchmark Results
General Knowledge
8 evaluationsCoding and Software Engineer
4 evaluationsWriting and Creative Capabilities
1 evaluationsAgent Level Benchmark
3 evaluationsMath and Reasoning
2 evaluationsClaw-style Agent Evaluation
2 evaluationsAI Agent - Tool Usage
2 evaluationsProductivity Knowledge
4 evaluationsCompetitor Comparison
Benchmark scores for Nemotron 3 Ultra compared against top models in its class
12 benchmarks with comparable scores. Each model shows its best score; mode label is displayed below.
| Benchmark | Nemotron 3 UltraCurrent | GLM-5.2 | DeepSeek-V4-Pro | MiniMax M3 |
|---|---|---|---|---|
38.32Thinking Enabled | -- | -- | 45.40Thinking Enabled | |
3.10Thinking Enabled | 20.90Thinking Level · High | 12.90Thinking Level · High | 3.70Thinking Enabled | |
37.40Thinking Enabled | Tools | 54.70Thinking Enabled | Tools | 48.20Thinking Level · Extra High | Tools | 39.00Thinking Enabled | |
51.78Standard Mode | 73.18Standard Mode | 71.57Standard Mode | 67.26Deep Thinking Mode | |
86.80Thinking Enabled | -- | 87.50Thinking Level · High | -- | |
86.70Thinking Enabled | 91.86Thinking Level · High | 90.50Thinking Level · High | 92.90Thinking Enabled | |
89.00Thinking Enabled | -- | 93.50Thinking Level · High | -- | |
44.60Thinking Enabled | 51.20Thinking Level · High | 50.80Thinking Level · High | 45.37Thinking Enabled | |
67.70Thinking Enabled | Tools | -- | 76.20Thinking Level · Extra High | Tools | -- | |
70.70Thinking Enabled | Tools | -- | 80.60Thinking Level · Extra High | Tools | -- | |
1689.30Standard Mode | 1752.80Standard Mode | 1552.10Standard Mode | -- | |
41.70Thinking Enabled | 58.80Standard Mode | 50.90Standard Mode | 45.80Thinking Enabled |
Standard API Pricing: Nemotron 3 Ultra vs. Peer Models
Shows standard text input and output pricing side by side for each model. If extended-context pricing exists, the chart keeps the base rate and explains the threshold below.
Source: DataLearnerAI. Standard text prices shown here use the default supplier.
These models use different currencies or billing units, so the page falls back to raw price values instead of a shared bar chart.
| Model | Supplier | Standard input | Standard output | Base price applies to |
|---|---|---|---|---|
GLM-5.2 | 智谱AI | $1.4 / 1M tokens | $4.4 / 1M tokens | — |
DeepSeek-V4-Pro | DeepSeek-AI | $0.435 / 1M tokens | $0.87 / 1M tokens | — |
MiniMax M3 | MiniMaxAI | ¥2.1 / 1M tokens | ¥8.4 / 1M tokens | — |