Qwen has released Qwen3.8-Flash-Next, a new multimodal Mixture-of-Experts (MoE) model that serves as an early preview of the architecture intended for Qwen4. This large model, with 125B tokens but only 6B active, achieves a significant performance boost.
Source: Simon Willison