NVIDIA logo

Nemotron 3 Ultra (550B A55B)

NVIDIA

Nemotron 3 Ultra is NVIDIA's frontier-scale open model with 550B total / 55B active parameters, built for agentic reasoning, long-context analysis, tool use, and high-stakes RAG.

Key Specifications

Parameters
550.0B
Context
-
Release Date
June 4, 2026
Average Score
64.8%

Timeline

Key dates in the model's history
Announcement
June 4, 2026
Last Update
August 29, 2026
Today
September 10, 2026

Technical Specifications

Parameters
550.0B
Training Tokens
20.0T tokens
Knowledge Cutoff
September 30, 2025
Family
-
Capabilities
MultimodalZeroEval

Benchmark Results

Model performance metrics across various tests and benchmarks

Programming

Programming skills tests
SWE-Bench Verified
Self-reported
70.7%

Reasoning

Logical reasoning and analysis
GPQA
No toolsSelf-reported
87.0%

Other Tests

Specialized benchmarks
AA-LCR
Self-reported
65.4%
Apex
Shortlist (with tools)Self-reported
84.8%
BrowseComp
With searchSelf-reported
44.4%
CritPT
No toolsSelf-reported
3.1%
Finance Agent
Vals.ai Financial Agent 1.1 (with web search)Self-reported
53.7%
Finance Agent v2
Verified
37.5%
GDPval
Self-reported
46.7%
Humanity's Last Exam
With toolsSelf-reported
37.4%
IFBench
Prompt looseSelf-reported
81.7%
IMO-AnswerBench
With toolsSelf-reported
92.3%
LiveCodeBench v6
Self-reported
89.0%
LongBench v2
≤ 1MSelf-reported
61.9%
MMLU-Pro
Self-reported
86.8%
MMLU-ProX
Average over languagesSelf-reported
83.0%
Multi-Challenge
Self-reported
63.8%
OmniScience
Non-HallucinationSelf-reported
78.7%
PinchBench
Self-reported
90.0%
ProfBench
SearchSelf-reported
56.0%
RULER
1MSelf-reported
94.7%
SciCode
SubtaskSelf-reported
44.6%
SWE-bench Multilingual
Self-reported
67.7%
TAU3-Bench
BankingSelf-reported
22.6%
Terminal-Bench 2.1
Self-reported
56.4%
WMT24++
en→xxSelf-reported
83.7%

License & Metadata

License
openmdw_license_v1_1
Announcement Date
June 4, 2026
Last Updated
August 29, 2026

Similar Models

All Models

Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.