Pyyan / Compare / Nemotron 3.5 Lightning vs MiMo-V2.6-Flash

Nemotron 3.5 Lightning vs MiMo-V2.6-Flash

2 of 5

Open-Weight Models · verified 29 Sept 2026

Nemotron 3.5 LightningNVIDIAcurrent
XMiMo-V2.6-FlashXiaomicurrent
3 slots left
SpecificationNemotron 3.5 LightningMiMo-V2.6-Flash
SummaryHybrid Mamba and transformer, 3B active, and the training data ships with it.The small one, and still multimodal.
Context128K1M
Input ($/Mtok)Weights onlyWeights only
Output ($/Mtok)Weights onlyWeights only
LicenceOpenMDW-1.1MIT
Max output32KNot published
Parameters30B total, 3B active309B total, 15B active
Released11 Aug 202621 Sep 2026
Model IDnvidia/nemotron-3.5-lightningXiaomiMiMo/MiMo-V2.6-Flash
CategoryOpen-Weight ModelsOpen-Weight Models
Official—Xiaomi ↗

Highlighted rows are where these differ.

Nemotron 3.5 Lightning

  • 30B mixture of experts with roughly 3B active per token, hybrid Mamba and transformer
  • Weights, training data and recipes all released, under NVIDIA's OpenMDW-1.1 licence
  • Scores 24 on the Artificial Analysis Intelligence Index, up from 15 for Nemotron 3 Nano

Best for long-running agents.

Full spec sheet →

MiMo-V2.6-Flash

  • 309B total parameters, 15B active per token
  • Same 1M context and five layer speculative decoder as Pro
  • MIT licence, weights on Hugging Face

Best for multimodal work on your own cluster.

Full spec sheet →