Pyyan / Compare / Hy4 preview vs Qwen3.8-Flash-Next

Hy4 preview vs Qwen3.8-Flash-Next

2 of 5

Open-Weight Models · verified 2 Sept 2026

Hy4 previewTencentcurrent
Qwen3.8-Flash-NextAlibabacurrent
3 slots left
SpecificationHy4 previewQwen3.8-Flash-Next
Summary770B parameters under Apache 2.0, and it helped optimise its own training.The most downloaded thing on Hugging Face right now, and a stated Qwen4 preview.
Context1M262K native, ~1M extended
Input ($/Mtok)$0.834$0.15
Output ($/Mtok)$2.501$0.47
LicenceApache 2.0Open weights
Max output64K32K
Parameters770B total, 49B active125B total, 6B active
Released28 Aug 202626 Aug 2026
Model IDtencent/Hy4-previewQwen/Qwen3.8-Flash-Next
CategoryOpen-Weight ModelsOpen-Weight Models
Official——

Highlighted rows are where these differ.

Hy4 preview

  • 770B total parameters with 49B active, 78 layers, 256 routed experts plus one shared
  • Apache 2.0, with an FP8 quantised build shipped alongside it
  • Tencent used it during its own development, and report a 31.8% inference throughput gain from it

Best for the largest open weights.

Full spec sheet →

Qwen3.8-Flash-Next

  • 125B total parameters with only 6B active per token
  • Alibaba describes it as a preview of the Qwen4 architecture
  • Top of the Hugging Face trending list, with a GGUF build close behind it

Best for open weights at speed.

Full spec sheet →