DeepSeek Launches V4.1-Flash With Lower Memory and API Costs
DeepSeek launched V4.1-Flash on Thursday, a 552-billion-parameter mixture-of-experts model that activates roughly 8 billion parameters per input token and 16 billion per output token, with a one-milli…