
Qwen has announced Qwen3.8-Flash-Next, a multimodal mixture-of-experts (MoE) model that serves as an early preview of the architecture planned for Qwen4. It follows Qwen3-Next, whose hybrid Gated DeltaNet (GDN) and Gated Attention design was later used across the Qwen3.5, Qwen3.6, Qwen3.7 and Qwen3.8 series. Continue reading “Qwen3.8-Flash-Next announced with 125B parameters, up to 1M-token context and Qwen4 architecture preview”
