Pyyan / Compare / DeepSeek V4.1 Flash vs DeepSeek V4-Pro

DeepSeek V4.1 Flash vs DeepSeek V4-Pro

2 of 5

Open-Weight Models · verified 13 Sept 2026

DeepSeek V4.1 FlashDeepSeekcurrent
DeepSeek V4-ProDeepSeekcurrent
3 slots left
SpecificationDeepSeek V4.1 FlashDeepSeek V4-Pro
SummaryTwice the parameters of the Flash it replaces, and cheaper to run.Best all-round open model of 2026.
Context1M
Input ($/Mtok)$0.3
Output ($/Mtok)$1.2
LicenceMITMIT
Max output256K
Parameters552B total, 8B / 16B active
Released10 Sep 2026
Model IDdeepseek-flash
CategoryOpen-Weight ModelsOpen-Weight Models
OfficialDeepSeekDeepSeek

Highlighted rows are where these differ.

DeepSeek V4.1 Flash

  • 552B total, 8B active while reading the prompt and 16B while generating, with a 1M token context
  • DeepSeek says it beats V4-Pro; V4-Pro requests route to it from 14 September 2026
  • MIT licence, native image input, and a KV cache about a quarter the size of V4-Flash's

Best for cheap agentic coding with images.

Full spec sheet →

DeepSeek V4-Pro

  • Tops open leaderboards on agentic coding and reasoning
  • MIT licence — no conditions
  • Pioneered aggressive cache-hit pricing at ~$0.07/M
  • DeepSeek routes all V4-Pro requests to V4.1 Flash from 14 September 2026

Best for everything open.

Full spec sheet →