NVIDIA logo

Nemotron 3 Ultra (550B A55B)

NVIDIA

Nemotron 3 Ultra is NVIDIA's frontier-scale open model with 550B total / 55B active parameters, built for agentic reasoning, long-context analysis, tool use, and high-stakes RAG.

Key Specifications

Parameters
550.0B
Context
-
Release Date
June 4, 2026
Average Score
64.8%

Timeline

Key dates in the model's history
Announcement
June 4, 2026
Last Update
August 29, 2026

Technical Specifications

Parameters
550.0B
Training Tokens
20.0T tokens
Knowledge Cutoff
September 30, 2025
Family
-
Capabilities
MultimodalZeroEval

Benchmark Results

Model performance metrics across various tests and benchmarks

Programming

Programming skills tests
SWE-Bench Verified
Self-reported
70.7%

Reasoning

Logical reasoning and analysis
GPQA
No toolsSelf-reported
87.0%

Other Tests

Specialized benchmarks
AA-LCR
Self-reported
65.4%
Apex
Shortlist (with tools)Self-reported
84.8%
BrowseComp
With searchSelf-reported
44.4%
CritPT
No toolsSelf-reported
3.1%
Finance Agent
Vals.ai Financial Agent 1.1 (with web search)Self-reported
53.7%
Finance Agent v2
Verified
37.5%
GDPval
Self-reported
46.7%
Humanity's Last Exam
With toolsSelf-reported
37.4%
IFBench
Prompt looseSelf-reported
81.7%
IMO-AnswerBench
With toolsSelf-reported
92.3%
LiveCodeBench v6
Self-reported
89.0%
LongBench v2
≤ 1MSelf-reported
61.9%
MMLU-Pro
Self-reported
86.8%
MMLU-ProX
Average over languagesSelf-reported
83.0%
Multi-Challenge
Self-reported
63.8%
OmniScience
Non-HallucinationSelf-reported
78.7%
PinchBench
Self-reported
90.0%
ProfBench
SearchSelf-reported
56.0%
RULER
1MSelf-reported
94.7%
SciCode
SubtaskSelf-reported
44.6%
SWE-bench Multilingual
Self-reported
67.7%
TAU3-Bench
BankingSelf-reported
22.6%
Terminal-Bench 2.1
Self-reported
56.4%
WMT24++
en→xxSelf-reported
83.7%

License & Metadata

License
openmdw_license_v1_1
Announcement Date
June 4, 2026
Last Updated
August 29, 2026

Compare Nemotron 3 Ultra (550B A55B)

All comparisons

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.