Inkling-Small vs Nemotron 3.5 Lightning (30B A3B): Specs & Benchmark Comparison

CharacteristicInkling-SmallNemotron 3.5 Lightning (30B A3B)
CompanyThinking Machines LabNVIDIA
Release DateJuly 30, 2026August 11, 2026
Parameters276B30B
MultimodalYesNo
Context (input)256K262K
Context (output)256K262K
Input Price / 1M$0.30$0.05
Output Price / 1M$1.20$0.20
Average Score0.60.8
Benchmarks
GPQA0.90.8
IFBench0.80.7

Visual Benchmark Comparison

Inkling-Small
Nemotron 3.5 Lightning (30B A3B)
GPQA0.9 vs 0.8
0.9
0.8
IFBench0.8 vs 0.7
0.8
0.7

Verdict

Nemotron 3.5 Lightning (30B A3B) leads in 2 out of 4 comparison categories.

Overall Performance

Both models show comparable average scores: Inkling-Small — 0.6, Nemotron 3.5 Lightning (30B A3B) — 0.8.

API Cost

Nemotron 3.5 Lightning (30B A3B) is 6.0x cheaper: input $0.05/1M vs $0.30/1M tokens.

Context Window

Nemotron 3.5 Lightning (30B A3B) supports a larger context: 262K vs 256K tokens.

Recency

Both models were released around the same time: 7/30/2026 and 8/11/2026.

More About These Models

Related Comparisons

Frequently Asked Questions

Which is better for coding — Inkling-Small or Nemotron 3.5 Lightning (30B A3B)?
Direct comparison on the SWE-Bench benchmark is not available. We recommend reviewing other metrics on the comparison page.
Which model is cheaper — Inkling-Small or Nemotron 3.5 Lightning (30B A3B)?
Nemotron 3.5 Lightning (30B A3B) is cheaper for input: $0.05 per 1M tokens vs $0.30.
Which has a larger context window — Inkling-Small or Nemotron 3.5 Lightning (30B A3B)?
Nemotron 3.5 Lightning (30B A3B) supports a larger context: 262,100 tokens vs 256,000.

The Inkling-Small and Nemotron 3.5 Lightning (30B A3B) comparison is updated for 2026. Data includes benchmark results, API pricing, context window size and other specifications. For more detailed information, visit the Inkling-Small or Nemotron 3.5 Lightning (30B A3B) page. See also the complete list of AI model comparisons.