DEV Community

AI Tech Connect
AI Tech Connect

Posted on • Originally published at aitechconnect.in

DeepSeek V4 Flash Ships GA and Beats Its Own Pro Preview

Originally published on AI Tech Connect.

What changed on 31 July DeepSeek pushed DeepSeek-V4-Flash-0731 into public beta on 31 July 2026. The model is a sparse Mixture-of-Experts with 284B total parameters and 13B active parameters per token, a context window of 1,048,576 tokens and a maximum output of 65,536 tokens. It sits in the V4 family alongside the Pro line, sharing that family's hybrid attention design — Compressed Sparse Attention paired with Heavily Compressed Attention, mHC connections and the Muon optimiser. The most interesting engineering detail in the release note is what did not change. The 0731 build keeps the exact structure and size of V4-Flash-Preview. No new expert count, no re-shaped attention, no different tensor layout. DeepSeek redid the post-training and nothing else. For anyone who has already sized a…


Read the full article on AI Tech Connect →

Top comments (0)