DeepSeek V4-Flash-0731 Is Official: The Budget Model That Beat Its Own Pro DeepSeek released DeepSeek-V4-Flash-0731 on July 31, an updated version of its existing deepseek-v4-flash API model that now outperforms DeepSeek's V4-Pro-Preview on all nine agent and coding benchmarks published by the company. The update, delivered through the same endpoint, model name, and price, applies a new post-training pipeline to the same 284B-parameter Mixture-of-Experts architecture with 13B active parameters and a 1M-token context window, focusing on coding and AI agents. On July 31, DeepSeek quietly updated every existing deepseek-v4-flash API integration. Same endpoint. Same model name. Same price. What changed is that the model now beats DeepSeek’s own V4-Pro-Preview on all nine agent and coding benchmarks the company published. If you were already calling Flash, you already got the upgrade. What Changed in the 0731 Build This is not a new model. DeepSeek-V4-Flash-0731 is the same April Preview architecture — a 284B-parameter Mixture-of-Experts model with 13B active parameters per token and a 1M-token context window. What changed is the post-training pipeline. DeepSeek rebuilt post-training around four targets: coding, AI agents, … The post