Pyyan / Compare / MiMo-V2.6-Pro vs MiMo-V2.6-Flash

MiMo-V2.6-Pro vs MiMo-V2.6-Flash

2 of 5

Open-Weight Models · verified 29 Sept 2026

XMiMo-V2.6-ProXiaomicurrent
XMiMo-V2.6-FlashXiaomicurrent
3 slots left
SpecificationMiMo-V2.6-ProMiMo-V2.6-Flash
SummaryA trillion parameters under the MIT licence.The small one, and still multimodal.
Context1M1M
Input ($/Mtok)Weights onlyWeights only
Output ($/Mtok)Weights onlyWeights only
LicenceMITMIT
Max outputNot publishedNot published
Parameters1.02T total, 42B active309B total, 15B active
Released21 Sep 202621 Sep 2026
Model IDXiaomiMiMo/MiMo-V2.6-ProXiaomiMiMo/MiMo-V2.6-Flash
CategoryOpen-Weight ModelsOpen-Weight Models
OfficialXiaomi ↗Xiaomi ↗

Highlighted rows are where these differ.

MiMo-V2.6-Pro

  • 1.02T total parameters, 42B active per token, a 4.1% activation ratio
  • Text, image, video and audio in, with a 1M token context
  • Shipped with 7,000+ reinforcement learning environments and the framework that trained it

Best for multimodal work on your own cluster.

Full spec sheet →

MiMo-V2.6-Flash

  • 309B total parameters, 15B active per token
  • Same 1M context and five layer speculative decoder as Pro
  • MIT licence, weights on Hugging Face

Best for multimodal work on your own cluster.

Full spec sheet →