Pyyan / Compare / Hy4 preview vs MiMo-V2.6-Flash

Hy4 preview vs MiMo-V2.6-Flash

2 of 5

Open-Weight Models · verified 29 Sept 2026

Hy4 previewTencentcurrent
XMiMo-V2.6-FlashXiaomicurrent
3 slots left
SpecificationHy4 previewMiMo-V2.6-Flash
Summary770B parameters under Apache 2.0, and it helped optimise its own training.The small one, and still multimodal.
Context1M1M
Input ($/Mtok)$0.834Weights only
Output ($/Mtok)$2.501Weights only
LicenceApache 2.0MIT
Max output64KNot published
Parameters770B total, 49B active309B total, 15B active
Released28 Aug 202621 Sep 2026
Model IDtencent/Hy4-previewXiaomiMiMo/MiMo-V2.6-Flash
CategoryOpen-Weight ModelsOpen-Weight Models
Official—Xiaomi ↗

Highlighted rows are where these differ.

Hy4 preview

  • 770B total parameters with 49B active, 78 layers, 256 routed experts plus one shared
  • Apache 2.0, with an FP8 quantised build shipped alongside it
  • Tencent used it during its own development, and report a 31.8% inference throughput gain from it

Best for the largest open weights.

Full spec sheet →

MiMo-V2.6-Flash

  • 309B total parameters, 15B active per token
  • Same 1M context and five layer speculative decoder as Pro
  • MIT licence, weights on Hugging Face

Best for multimodal work on your own cluster.

Full spec sheet →