Pyyan / Compare / Nemotron 3.5 Lightning vs MiMo-V2.6-Pro

Nemotron 3.5 Lightning vs MiMo-V2.6-Pro

2 of 5

Open-Weight Models · verified 29 Sept 2026

Nemotron 3.5 LightningNVIDIAcurrent
XMiMo-V2.6-ProXiaomicurrent
3 slots left
SpecificationNemotron 3.5 LightningMiMo-V2.6-Pro
SummaryHybrid Mamba and transformer, 3B active, and the training data ships with it.A trillion parameters under the MIT licence.
Context128K1M
Input ($/Mtok)Weights onlyWeights only
Output ($/Mtok)Weights onlyWeights only
LicenceOpenMDW-1.1MIT
Max output32KNot published
Parameters30B total, 3B active1.02T total, 42B active
Released11 Aug 202621 Sep 2026
Model IDnvidia/nemotron-3.5-lightningXiaomiMiMo/MiMo-V2.6-Pro
CategoryOpen-Weight ModelsOpen-Weight Models
Official—Xiaomi ↗

Highlighted rows are where these differ.

Nemotron 3.5 Lightning

  • 30B mixture of experts with roughly 3B active per token, hybrid Mamba and transformer
  • Weights, training data and recipes all released, under NVIDIA's OpenMDW-1.1 licence
  • Scores 24 on the Artificial Analysis Intelligence Index, up from 15 for Nemotron 3 Nano

Best for long-running agents.

Full spec sheet →

MiMo-V2.6-Pro

  • 1.02T total parameters, 42B active per token, a 4.1% activation ratio
  • Text, image, video and audio in, with a 1M token context
  • Shipped with 7,000+ reinforcement learning environments and the framework that trained it

Best for multimodal work on your own cluster.

Full spec sheet →