The Next Web
Qwen4’s architecture is here early, firing 6B parameters out of 125B

Alibaba’s Qwen team released Qwen3.8-Flash-Next as an open-weight preview of the architecture slated for Qwen4, which will contain 125 billion parameters while activating only 6 billion per token. The model’s license is uncertain to qualify for the EU AI Act’s open-source exemption.