Alibaba to Release Qwen 3.8-Flash-Next as a Preview of What Qwen 4 Will Offer
Alibaba's Qwen team is set to release Qwen 3.8-Flash-Next, a 125-billion-parameter model that activates just 6 billion per token. The model is described as a preview of the next-generation Qwen 4 architecture.
Intelligence analysis by Llama

Alibaba's Qwen team is releasing Qwen 3.8-Flash-Next, a 125-billion-parameter model that activates just 6 billion per token. The model is a preview of the next-generation Qwen 4 architecture.
Imagine you have a super smart friend who can do lots of things, like math and language. Qwen 3.8-Flash-Next is like a big computer program that can do lots of things too, but it's not as smart as your friend. It's like a preview of a new, even smarter friend that Alibaba is working on.
Analysis
Qwen 3.8-Flash-Next: A Glimpse into the Future of AI Models
The Qwen team at Alibaba is set to release Qwen 3.8-Flash-Next, a 125-billion-parameter model that activates just 6 billion per token. This model is described as a preview of the next-generation Qwen 4 architecture, which is expected to be a mixture-of-experts system.
The release of Qwen 3.8-Flash-Next is significant as it provides a glimpse into the architecture of the next-generation Qwen 4 model. This could have implications for the development of AI models in the future. The Qwen team has framed it as a preview of the next-generation Qwen 4 architecture, not a finished flagship.
There is no official information on the model, but based on rumors, it will likely be a mixture-of-experts system. This type of system is designed to combine the strengths of multiple models to produce a more accurate and robust output. The Qwen team has not released any hard benchmark scores for the model, and the weights are not live on ModelScope as of this writing.
The release of Qwen 3.8-Flash-Next is a significant step forward in the development of AI models. It provides a glimpse into the architecture of the next-generation Qwen 4 model and could have implications for the development of AI models in the future.
Key points
- Alibaba's Qwen team is releasing Qwen 3.8-Flash-Next, a 125-billion-parameter model that activates just 6 billion per token.
- The model is described as a preview of the next-generation Qwen 4 architecture.
- The release of Qwen 3.8-Flash-Next is significant as it provides a glimpse into the architecture of the next-generation Qwen 4 model.
- The Qwen team has not released any hard benchmark scores for the model, and the weights are not live on ModelScope as of this writing.
If Qwen 3.8-Flash-Next is successful, it could lead to the development of even more advanced AI models in the future. This could have significant implications for various industries and could lead to breakthroughs in areas such as healthcare and finance.
However, the development of Qwen 3.8-Flash-Next is still in its early stages, and there are many challenges that need to be overcome before it can be considered a success. If the model is not able to live up to its promises, it could lead to disappointment and a setback in the development of AI models.



