Alibaba’s Qwen Team Releases Qwen3.8-Flash-Next: A 125B Multimodal MoE With 6B Active Parameters Previewing the Qwen4 Architecture
Alibaba’s Qwen workforce has launched Qwen3.8-Flash-Next, an open-weight multimodal Mixture-of-Experts mannequin constructed for price per token. The checkpoint pairs a 125B spine with a 51B N-gram embedding desk and a 4B multi-token prediction module. Only 6B parameters activate per token. The workforce positions it as an early preview of the structure that may underpin Qwen4,…
