Key Specifications
Parameters
550.0B
Context
-
Release Date
June 4, 2026
Average Score
64.8%
Timeline
Key dates in the model's history
Announcement
June 4, 2026
Last Update
August 29, 2026
Technical Specifications
Parameters
550.0B
Training Tokens
20.0T tokens
Knowledge Cutoff
September 30, 2025
Family
-
Capabilities
MultimodalZeroEval
Benchmark Results
Model performance metrics across various tests and benchmarks
Programming
Programming skills tests
SWE-Bench Verified
• Self-reported
Reasoning
Logical reasoning and analysis
GPQA
No tools • Self-reported
Other Tests
Specialized benchmarks
AA-LCR
• Self-reported
Apex
Shortlist (with tools) • Self-reported
BrowseComp
With search • Self-reported
CritPT
No tools • Self-reported
Finance Agent
Vals.ai Financial Agent 1.1 (with web search) • Self-reported
Finance Agent v2
• Verified
GDPval
• Self-reported
Humanity's Last Exam
With tools • Self-reported
IFBench
Prompt loose • Self-reported
IMO-AnswerBench
With tools • Self-reported
LiveCodeBench v6
• Self-reported
LongBench v2
≤ 1M • Self-reported
MMLU-Pro
• Self-reported
MMLU-ProX
Average over languages • Self-reported
Multi-Challenge
• Self-reported
OmniScience
Non-Hallucination • Self-reported
PinchBench
• Self-reported
ProfBench
Search • Self-reported
RULER
1M • Self-reported
SciCode
Subtask • Self-reported
SWE-bench Multilingual
• Self-reported
TAU3-Bench
Banking • Self-reported
Terminal-Bench 2.1
• Self-reported
WMT24++
en→xx • Self-reported
License & Metadata
License
openmdw_license_v1_1
Announcement Date
June 4, 2026
Last Updated
August 29, 2026
Compare Nemotron 3 Ultra (550B A55B)
All comparisonsSimilar Models
All ModelsNemotron 3 Super (120B A12B)
NVIDIA
120.0B
Best score:0.8 (GPQA)
Released:Mar 2026
Llama 3.1 Nemotron Ultra 253B v1
NVIDIA
253.0B
Best score:0.8 (GPQA)
Released:Apr 2025
LongCat-Flash-Thinking-2601
Meituan
560.0B
Best score:1.0 (TAU)
Released:Jan 2026
Kimi K2 Instruct
Moonshot AI
1.0T
Best score:0.9 (HumanEval)
Released:Jan 2025
Price:$0.57/1M tokens
Mistral Large 2
Mistral AI
123.0B
Best score:0.9 (HumanEval)
Released:Jul 2024
Price:$2.00/1M tokens
LongCat-Flash-Chat
Meituan
560.0B
Best score:0.9 (MMLU)
Released:Aug 2025
MiMo-V2-Flash
Xiaomi
309.0B
Best score:0.8 (GPQA)
Released:Dec 2025
Llama 3.1 405B Instruct
Meta
405.0B
Best score:1.0 (ARC)
Released:Jul 2024
Price:$3.50/1M tokens
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.