Alibaba Releases Qwen3.8-Flash-Next as Early Architectural Preview of Qwen4
Alibaba's Qwen team released Qwen3.8-Flash-Next on August 26, 2026, an open-weight multimodal mixture-of-experts model designed as an architectural preview of the upcoming Qwen4 family. The model has 125 billion total parameters but activates only 6 billion per token, plus a separate 51-billion-parameter N-gram embedding layer that can run in system RAM rather than GPU memory. Alibaba claims training costs were roughly one-ninth those of the 397-billion-parameter Qwen3.7-Plus model. Vendor-reported benchmarks show strong performance on agentic coding tasks, though the model trails Claude Opus 5 significantly on computer-use evaluations such as OSWorld 2.0. The release follows Alibaba's stated strategy of publishing architectural experiments with public weights before launching a full model series.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.


Discussion (0)
Log in to join the discussion and vote.
Log in