BTC $78,837.13 +0.07%
ETH $2,452.59 -1.06%
BNB $696.23 -0.82%
XRP $1.44 -2.26%
SOL $97.24 -0.59%
TRX $0.3373 -1.98%
DOGE $0.0866 -3.55%
ADA $0.2118 -3.92%
BCH $266.10 -1.95%
LINK $11.35 -2.23%
HYPE $80.02 +2.58%
AAVE $126.89 -3.25%
SUI $0.7639 -4.56%
XLM $0.1850 -4.33%
ZEC $776.91 -6.64%
BTC $78,837.13 +0.07%
ETH $2,452.59 -1.06%
BNB $696.23 -0.82%
XRP $1.44 -2.26%
SOL $97.24 -0.59%
TRX $0.3373 -1.98%
DOGE $0.0866 -3.55%
ADA $0.2118 -3.92%
BCH $266.10 -1.95%
LINK $11.35 -2.23%
HYPE $80.02 +2.58%
AAVE $126.89 -3.25%
SUI $0.7639 -4.56%
XLM $0.1850 -4.33%
ZEC $776.91 -6.64%
first_img

Alibaba will release Qwen 3.8-Flash-Next and preview the Qwen 4 architecture

2026-08-26 05:20:40

The Alibaba Qwen team will release Qwen 3.8-Flash-Next on Wednesday, which is a mixture of experts (MoE) model with a total of 125 billion parameters, but only 6 billion parameters are activated for each token. The team positions it as a preview version of the next-generation Qwen 4 architecture, rather than the final flagship model.

The model is described as multimodal and is built on the upcoming Qwen 4 architecture. The team stated that the early release of this version is to prepare developers for the subsequent complete Qwen series. The model weights will be hosted on the Hugging Face and ModelScope platforms.

As of now, Qwen has not disclosed the benchmark results for this model, nor has it released comparative data with its own Qwen 3 series or overseas competitors. The specific figures of 125 billion and 6 billion parameters have not been officially confirmed. As an open-weight model, developers can download, fine-tune, and run it for free without sending data to a closed API, which helps reduce the cost of hosting the model.

app_icon
ChainCatcher Building the Web3 world with innovations.