DeepSeek-V4-Flash
DeepSeek · 2026-04-24 · 292B parameters
DeepSeek-V4-Flash is the smaller, efficiency-focused open-weight MoE model in DeepSeek's V4 preview release, marketed at 284B total and 13B activated parameters. Released April 24, 2026, it supports a 1M-token context and Non-Think, Think High, and Think Max reasoning modes; its 43-layer attention stack begins with two sliding-window layers, then interleaves Compressed Sparse Attention (CSA) and Heavily Compressed Attention (HCA). Its weights are released under the MIT license.