Pyyan / Compare / Nemotron 3.5 Lightning vs Qwen3.8-Flash-Next

Nemotron 3.5 Lightning vs Qwen3.8-Flash-Next

2 of 5

Open-Weight Models · verified 2 Sept 2026

Nemotron 3.5 LightningNVIDIAcurrent
Qwen3.8-Flash-NextAlibabacurrent
3 slots left
SpecificationNemotron 3.5 LightningQwen3.8-Flash-Next
SummaryHybrid Mamba and transformer, 3B active, and the training data ships with it.The most downloaded thing on Hugging Face right now, and a stated Qwen4 preview.
Context128K262K native, ~1M extended
Input ($/Mtok)Weights only$0.15
Output ($/Mtok)Weights only$0.47
LicenceOpenMDW-1.1Open weights
Max output32K32K
Parameters30B total, 3B active125B total, 6B active
Released11 Aug 202626 Aug 2026
Model IDnvidia/nemotron-3.5-lightningQwen/Qwen3.8-Flash-Next
CategoryOpen-Weight ModelsOpen-Weight Models
Official——

Highlighted rows are where these differ.

Nemotron 3.5 Lightning

  • 30B mixture of experts with roughly 3B active per token, hybrid Mamba and transformer
  • Weights, training data and recipes all released, under NVIDIA's OpenMDW-1.1 licence
  • Scores 24 on the Artificial Analysis Intelligence Index, up from 15 for Nemotron 3 Nano

Best for long-running agents.

Full spec sheet →

Qwen3.8-Flash-Next

  • 125B total parameters with only 6B active per token
  • Alibaba describes it as a preview of the Qwen4 architecture
  • Top of the Hugging Face trending list, with a GGUF build close behind it

Best for open weights at speed.

Full spec sheet →