The Qwen3.8-Flash-Next multimodal MoE model uses 6B active parameters out of 125B total tokens. This release serves as an early architectural preview for Qwen4. Early tests on quantized versions show high reasoning capabilities. Practitioners can now evaluate this efficiency-focused design before the full next-generation model arrives.