Pyyan / Compare / Nemotron 3.5 Lightning vs DeepSeek V4-Pro

Nemotron 3.5 Lightning vs DeepSeek V4-Pro

2 of 5

Open-Weight Models · verified 13 Sept 2026

Nemotron 3.5 LightningNVIDIAcurrent
DeepSeek V4-ProDeepSeekcurrent
3 slots left
SpecificationNemotron 3.5 LightningDeepSeek V4-Pro
SummaryHybrid Mamba and transformer, 3B active, and the training data ships with it.Best all-round open model of 2026.
Context128K—
Input ($/Mtok)Weights only—
Output ($/Mtok)Weights only—
LicenceOpenMDW-1.1MIT
Max output32K—
Parameters30B total, 3B active—
Released11 Aug 2026—
Model IDnvidia/nemotron-3.5-lightning—
CategoryOpen-Weight ModelsOpen-Weight Models
Official—DeepSeek ↗

Highlighted rows are where these differ.

Nemotron 3.5 Lightning

  • 30B mixture of experts with roughly 3B active per token, hybrid Mamba and transformer
  • Weights, training data and recipes all released, under NVIDIA's OpenMDW-1.1 licence
  • Scores 24 on the Artificial Analysis Intelligence Index, up from 15 for Nemotron 3 Nano

Best for long-running agents.

Full spec sheet →

DeepSeek V4-Pro

  • Tops open leaderboards on agentic coding and reasoning
  • MIT licence — no conditions
  • Pioneered aggressive cache-hit pricing at ~$0.07/M
  • DeepSeek routes all V4-Pro requests to V4.1 Flash from 14 September 2026

Best for everything open.

Full spec sheet →