Alibaba releases Qwen3.8-Flash-Next, targeting "ultimate cost efficiency" Alibaba's Qwen team released Qwen3.8-Flash-Next, a mixture-of-experts model that activates 6 of 125 billion parameters per token, claiming one-ninth the training cost of competitors while outperforming DeepSeek-V4-Flash and Claude Opus 4.6 on coding and office benchmarks. The model, previewing the Qwen4 architecture, intensifies pricing pressure on OpenAI and Anthropic. Alibaba's Qwen team is previewing the Qwen4 architecture with Qwen3.8-Flash-Next, a mixture-of-experts model that activates just 6 out of 125 billion parameters per token. At one-ninth the training cost, it beats much larger competitors like DeepSeek-V4-Flash and Claude Opus 4.6 on coding and office benchmarks, adding more pricing pressure on OpenAI and Anthropic. The article Alibaba releases Qwen3.8-Flash-Next, targeting "ultimate cost efficiency" https://the-decoder.com/alibaba-releases-qwen3-8-flash-next-targeting-ultimate-cost-efficiency/ appeared first on The Decoder https://the-decoder.com .