Pyyan / Compare / Qwen3.8-27B vs MiMo-V2.6-Pro

Qwen3.8-27B vs MiMo-V2.6-Pro

2 of 5

Open-Weight Models · verified 29 Sept 2026

Qwen3.8-27BAlibabacurrent
XMiMo-V2.6-ProXiaomicurrent
3 slots left
SpecificationQwen3.8-27BMiMo-V2.6-Pro
Summary27B dense under Apache 2.0, and nearly five million downloads.A trillion parameters under the MIT licence.
Context262K native, ~1M extended1M
Input ($/Mtok)Weights onlyWeights only
Output ($/Mtok)Weights onlyWeights only
LicenceApache 2.0MIT
Max output32KNot published
Parameters27B dense1.02T total, 42B active
Released5 Aug 202621 Sep 2026
Model IDQwen/Qwen3.8-27BXiaomiMiMo/MiMo-V2.6-Pro
CategoryOpen-Weight ModelsOpen-Weight Models
Official—Xiaomi ↗

Highlighted rows are where these differ.

Qwen3.8-27B

  • 27B dense rather than a mixture of experts, so it behaves predictably on one machine
  • Apache 2.0, weights only, no hosted price
  • One of the most downloaded models on Hugging Face, with quantised builds everywhere

Best for running locally.

Full spec sheet →

MiMo-V2.6-Pro

  • 1.02T total parameters, 42B active per token, a 4.1% activation ratio
  • Text, image, video and audio in, with a 1M token context
  • Shipped with 7,000+ reinforcement learning environments and the framework that trained it

Best for multimodal work on your own cluster.

Full spec sheet →