Experiment with Qwen3.8-Flash-Next on NVIDIA GB300 NVL72 for Agentic Coding
Alibaba released the model weights for Qwen3.8-Flash-Next, a multimodal mixture-of-experts model with 125B parameters plus 51B N-gram embeddings and 6B activated per token, as a preview of the Qwen4 a…