Pyyan / Compare / MiMo-V2.6-Flash vs MiMo-V2.6-Pro

MiMo-V2.6-Flash vs MiMo-V2.6-Pro

2 of 5

Open-Weight Models · verified 29 Sept 2026

XMiMo-V2.6-FlashXiaomicurrent
XMiMo-V2.6-ProXiaomicurrent
3 slots left
SpecificationMiMo-V2.6-FlashMiMo-V2.6-Pro
SummaryThe small one, and still multimodal.A trillion parameters under the MIT licence.
Context1M1M
Input ($/Mtok)Weights onlyWeights only
Output ($/Mtok)Weights onlyWeights only
LicenceMITMIT
Max outputNot publishedNot published
Parameters309B total, 15B active1.02T total, 42B active
Released21 Sep 202621 Sep 2026
Model IDXiaomiMiMo/MiMo-V2.6-FlashXiaomiMiMo/MiMo-V2.6-Pro
CategoryOpen-Weight ModelsOpen-Weight Models
OfficialXiaomi ↗Xiaomi ↗

Highlighted rows are where these differ.

MiMo-V2.6-Flash

  • 309B total parameters, 15B active per token
  • Same 1M context and five layer speculative decoder as Pro
  • MIT licence, weights on Hugging Face

Best for multimodal work on your own cluster.

Full spec sheet →

MiMo-V2.6-Pro

  • 1.02T total parameters, 42B active per token, a 4.1% activation ratio
  • Text, image, video and audio in, with a 1M token context
  • Shipped with 7,000+ reinforcement learning environments and the framework that trained it

Best for multimodal work on your own cluster.

Full spec sheet →