Nemotron 3.5 Lightning (30B A3B)
Nemotron 3.5 Lightning (30B A3B) is NVIDIA's open-weight mixture-of-experts model with 30 billion total parameters (3B active), built for long-running agentic tasks. It delivers efficient agentic coding performance, ranking second on PinchBench, with a 262K token context window.
Key Specifications
Parameters
30.0B
Context
262.1K
Release Date
August 11, 2026
Average Score
78.5%
Timeline
Key dates in the model's history
Announcement
August 11, 2026
Last Update
August 27, 2026
Technical Specifications
Parameters
30.0B
Training Tokens
-
Knowledge Cutoff
-
Family
-
Capabilities
MultimodalZeroEval
Pricing & Availability
Input (per 1M tokens)
$0.05
Output (per 1M tokens)
$0.20
Max Input Tokens
262.1K
Max Output Tokens
262.1K
Supported Features
Function CallingStructured OutputCode ExecutionWeb SearchBatch InferenceFine-tuning
Benchmark Results
Model performance metrics across various tests and benchmarks
Reasoning
Logical reasoning and analysis
GPQA
NeMo Gym / NeMo Evaluator; BF16; GPQA Diamond (no tools) • Self-reported
Other Tests
Specialized benchmarks
PinchBench
NeMo Gym / NeMo Evaluator; BF16 • Self-reported
MMLU-Pro
NeMo Gym / NeMo Evaluator; BF16 • Self-reported
IFBench
NeMo Gym / NeMo Evaluator; BF16; loose • Self-reported
License & Metadata
License
openmdw_license_v1_1
Announcement Date
August 11, 2026
Last Updated
August 27, 2026
Compare Nemotron 3.5 Lightning (30B A3B)
All comparisonsSimilar Models
All ModelsNemotron 3 Nano (30B A3B)
NVIDIA
32.0B
Best score:0.8 (GPQA)
Released:Dec 2025
Price:$0.06/1M tokens
Llama 3.1 Nemotron 70B Instruct
NVIDIA
70.0B
Best score:0.8 (MMLU)
Released:Oct 2024
Llama-3.3 Nemotron Super 49B v1
NVIDIA
49.9B
Best score:0.7 (GPQA)
Released:Mar 2025
Llama 3.1 Nemotron Ultra 253B v1
NVIDIA
253.0B
Best score:0.8 (GPQA)
Released:Apr 2025
Nemotron 3 Super (120B A12B)
NVIDIA
120.0B
Best score:0.8 (GPQA)
Released:Mar 2026
Mistral NeMo Instruct
Mistral AI
12.0B
Best score:0.7 (MMLU)
Released:Jul 2024
Price:$0.15/1M tokens
Magistral Small 2506
Mistral AI
24.0B
Best score:0.7 (GPQA)
Released:Jun 2025
LongCat-Flash-Lite
Meituan
68.5B
Best score:0.9 (MMLU)
Released:Feb 2026
Recommendations are based on similarity of characteristics: developer organization, multimodality, parameter size, and benchmark performance. Choose a model to compare or go to the full catalog to browse all available AI models.