Qwen3.8-Flash-Next Previews Qwen4 Architecture With 6B Active Parameters

Qwen3.8-Flash-Next Previews Qwen4 Architecture With Hybrid Attention and 6B Active Parameters Alibaba’s Qwen team released Qwen3.8-Flash-Next on August 26, 2026, an open-weight experimental model that previews the architecture intended to underpin Qwen4. The model carries 125B parameters but activates only 6B per token, a configuration the team frames as a step toward what it calls ultimate cost-efficiency. The release is the first public model built on this design. The Qwen3.8-Flash-Next model…

Leave a Reply

Your email address will not be published. Required fields are marked *

Back To Top