Qwen3.8-Flash-Next Previews Qwen4 Architecture With 6B Active Parameters
Covered by 5 sources
Read full postAlibaba's Qwen team unveiled Qwen3.8-Flash-Next, an experimental 125B-parameter model activating only 6B per token, previewing the Qwen4 architecture focused on cost-efficient inference with hybrid attention and long context support.


