DeepSeek launches V4.1-Flash model with 552B parameters and a million-token context window
DeepSeek launched V4.1-Flash on September 10, a 552-billion-parameter Mixture-of-Experts model with a 1-million-token context window that activates roughly 8 billion parameters for input and 16 billio…