All NewsEducationTV
Equities & FundsCrypto & Digital AssetsAI & TechnologyBusiness & CorporateUS Politics & PolicyGeopolitics & Global RiskMacro, Rates & FXCommodities & EnergyEuropean Politics & MarketsAsia-PacificReal Estate & Property
All NewsHome
← Back to AI & Technology

Alibaba Previews Qwen 4 Architecture with Qwen 3.8-Flash-Next Model

Created at 25 Aug · 9:21 PM1 source↑ Market-relevant
IN SHORT

Alibaba's Qwen team will release Qwen 3.8-Flash-Next, a 125-billion-parameter model activating 6 billion per token, as a preview of its upcoming Qwen 4 architecture. The multimodal model is expected to utilize a mixture-of-experts design, allowing for efficient resource utilization.

Key Numbers

125 billiontotal parameters in Qwen 3.8-Flash-Next
6 billionactive parameters per token

Who's Involved

Alibaba
company releasing Qwen 3.8-Flash-Next
Qwen team
developer of the Qwen models
Alibaba Previews Qwen 4 Architecture with Qwen 3.8-Flash-Next Model

↳ Why This Matters

This release offers a glimpse into Alibaba's next-generation AI architecture, potentially making advanced AI capabilities more accessible and cost-effective for developers through its efficient mixture-of-experts design.

Key facts

  • Alibaba's Qwen team is releasing Qwen 3.8-Flash-Next on Wednesday.
  • The model has 125 billion total parameters, activating 6 billion per token.
  • It is described as a preview of the Qwen 4 architecture.
  • The model is multimodal and expected to use a mixture-of-experts design.
  • The weights are not yet live on ModelScope or Hugging Face.

Alibaba's Qwen team is set to release Qwen 3.8-Flash-Next on Wednesday, a model with 125 billion total parameters that activates only 6 billion per token. The team has framed this release as a preview of the upcoming Qwen 4 architecture, rather than a finished flagship product. While official benchmark scores have not yet been published, the model is described as multimodal and built upon the Qwen 4 architecture. It is expected to employ a mixture-of-experts (MoE) design, a system where the network is divided into specialized sub-models, with only the relevant ones activated for specific tasks. This MoE approach allows a large model to operate with the computational cost of a much smaller one, making advanced capabilities accessible on more common hardware. Alibaba has released this early build to allow developers to prepare for the full Qwen 4 family. The weights for the model are not yet available on platforms like ModelScope or Hugging Face.

Frequently asked questions

Qwen 3.8-Flash-Next is a new model from Alibaba's Qwen team, serving as a preview of the upcoming Qwen 4 architecture. It features 125 billion total parameters but only activates 6 billion per token.

A mixture-of-experts model splits its network into many specialized sub-models, activating only the relevant ones for each task. This allows for greater capability with reduced computational cost.

The model is scheduled for release on Wednesday.

As of this writing, the weights are not yet live on platforms like ModelScope or Hugging Face.

What Happens Next

01Qwen 3.8-Flash-Next will be released on Wednesday.
02Benchmark scores for Qwen 3.8-Flash-Next are expected.
03The full Qwen 4 architecture rollout is anticipated.

How It Developed

Alibaba's Qwen team will release Qwen 3.8-Flash-Next on Wednesday.
The model is described as a preview of the Qwen 4 architecture.
It features 125 billion total parameters with 6 billion active per token.
The model is multimodal and built on the upcoming Qwen 4 architecture.
Alibaba shipped the early build for developers to prepare for the full Qwen 4 family.

Sources

T1
Alibaba to Release Qwen 3.8-Flash-Next as a Preview of What Qwen 4 Will OfferDecrypt

Related Stories

Researchers Shrink AI Model, Making It Smarter Than Original
25 Aug · 8:51 PM
Apple unveils Mac Studio, Mac mini with M5 Ultra and M6 chips for AI
25 Aug · 1:12 PM
OpenAI's Jalapeño chip benchmarks show performance gains over Nvidia Blackwell
25 Aug · 2:56 PM
DeepSeek leads surge in low-cost Chinese open-weight models on US platform
25 Aug · 10:36 AM
AI Founders Launch Physics Model After Rejecting Bezos's Project Prometheus
25 Aug · 10:05 AM