Qwen3.8-Flash-Next Previews Qwen4 Architecture With 6B Active Parameters
Covered by 5 sources
Read full postAlibaba's Qwen team unveiled Qwen3.8-Flash-Next, an experimental 125B-parameter model activating only 6B per token, previewing the Qwen4 architecture focused on cost-efficient inference with hybrid attention and long context support.


.png?disable=upscale&width=1200&height=630&fit=crop)
