Alibaba to Release Qwen 3.8-Flash-Next as a Preview of What Qwen 4 Will Offer
Compiled by KHAO Editorial — aggregated from 1 source. See llms.txt for citation guidance.
★ Tier-1 Source
Alibaba is set to release Qwen 3.8-Flash-Next on Wednesday, a 125-billion-parameter model that activates 6 billion per token.
Key facts
- Alibaba is set to release Qwen 3.8-Flash-Next on Wednesday, a 125-billion-parameter model that activates 6 billion per token
- Qwen 3.8 Flash Next is releasing Tomorrow. 125B paramters +51B N-gram and 6B active
- Hugging Face, where the weights also live, also describes it as "a preview of the Qwen 4 architecture
- Alibaba’s Qwen team does describe the model as multimodal and built on the upcoming Qwen 4 architecture, and says it shipped the early build so developers can prepare for the full family
Summary
Alibaba's Qwen team is set to release Qwen 3.8-Flash-Next on Wednesday, a Mixture-of-Experts model described as a preview of the Qwen4 architecture. The team's pre-release briefing cites 125 billion total parameters with only 6 billion active per token. Hard benchmark scores haven't been published yet, and the weights aren't live on ModelScope as of this writing. There is no official information on the model, but based on rumors, it will likely be a mixture-of-experts system, a design that splits the network into many specialized sub-models and lights up only the relevant ones for each task.