Pyyan / Compare / MiMo-V2.6-Flash vs Qwen3.8-Flash-Next

MiMo-V2.6-Flash vs Qwen3.8-Flash-Next

2 of 5

Open-Weight Models · verified 29 Sept 2026

XMiMo-V2.6-FlashXiaomicurrent
Qwen3.8-Flash-NextAlibabacurrent
3 slots left
SpecificationMiMo-V2.6-FlashQwen3.8-Flash-Next
SummaryThe small one, and still multimodal.The most downloaded thing on Hugging Face right now, and a stated Qwen4 preview.
Context1M262K native, ~1M extended
Input ($/Mtok)Weights only$0.15
Output ($/Mtok)Weights only$0.47
LicenceMITOpen weights
Max outputNot published32K
Parameters309B total, 15B active125B total, 6B active
Released21 Sep 202626 Aug 2026
Model IDXiaomiMiMo/MiMo-V2.6-FlashQwen/Qwen3.8-Flash-Next
CategoryOpen-Weight ModelsOpen-Weight Models
OfficialXiaomi ↗—

Highlighted rows are where these differ.

MiMo-V2.6-Flash

  • 309B total parameters, 15B active per token
  • Same 1M context and five layer speculative decoder as Pro
  • MIT licence, weights on Hugging Face

Best for multimodal work on your own cluster.

Full spec sheet →

Qwen3.8-Flash-Next

  • 125B total parameters with only 6B active per token
  • Alibaba describes it as a preview of the Qwen4 architecture
  • Top of the Hugging Face trending list, with a GGUF build close behind it

Best for open weights at speed.

Full spec sheet →