Pyyan / Compare / Qwen3.8-Flash-Next vs Inkling
Open-Weight Models · verified 6 Sept 2026
| Specification | Qwen3.8-Flash-Next | Inkling |
|---|---|---|
| Summary | The most downloaded thing on Hugging Face right now, and a stated Qwen4 preview. | 975B under Apache 2.0, shipped as a starting point rather than a finished product. |
| Context | 262K native, ~1M extended | 1M |
| Input ($/Mtok) | $0.15 | Weights only |
| Output ($/Mtok) | $0.47 | Weights only |
| Licence | Open weights | Apache 2.0 |
| Max output | 32K | Not published |
| Parameters | 125B total, 6B active | 975B total, 41B active |
| Released | 26 Aug 2026 | 15 Jul 2026 |
| Model ID | Qwen/Qwen3.8-Flash-Next | Not published |
| Category | Open-Weight Models | Open-Weight Models |
| Official | — | Thinking Machines Lab ↗ |
Highlighted rows are where these differ.
Best for open weights at speed.
Best for fine tuning on your own data.