Pyyan / Compare / Qwen3.8-27B vs Qwen3.8-Flash-Next

Qwen3.8-27B vs Qwen3.8-Flash-Next

2 of 5

Open-Weight Models · verified 2 Sept 2026

Qwen3.8-27BAlibabacurrent
Qwen3.8-Flash-NextAlibabacurrent
3 slots left
SpecificationQwen3.8-27BQwen3.8-Flash-Next
Summary27B dense under Apache 2.0, and nearly five million downloads.The most downloaded thing on Hugging Face right now, and a stated Qwen4 preview.
Context262K native, ~1M extended262K native, ~1M extended
Input ($/Mtok)Weights only$0.15
Output ($/Mtok)Weights only$0.47
LicenceApache 2.0Open weights
Max output32K32K
Parameters27B dense125B total, 6B active
Released5 Aug 202626 Aug 2026
Model IDQwen/Qwen3.8-27BQwen/Qwen3.8-Flash-Next
CategoryOpen-Weight ModelsOpen-Weight Models
Official——

Highlighted rows are where these differ.

Qwen3.8-27B

  • 27B dense rather than a mixture of experts, so it behaves predictably on one machine
  • Apache 2.0, weights only, no hosted price
  • One of the most downloaded models on Hugging Face, with quantised builds everywhere

Best for running locally.

Full spec sheet →

Qwen3.8-Flash-Next

  • 125B total parameters with only 6B active per token
  • Alibaba describes it as a preview of the Qwen4 architecture
  • Top of the Hugging Face trending list, with a GGUF build close behind it

Best for open weights at speed.

Full spec sheet →