Alibaba to Release Qwen3 Flash-Next 125B MoE Model Imminently
Alibaba's Qwen team is set to release a new AI model called Qwen3.8-Flash-Next as early as tomorrow. The model is a Mixture-of-Experts (MoE) architecture with 125 billion total parameters but only 8 billion active parameters per forward pass. The release was announced via a ModelScope model page, suggesting the weights will be publicly hosted on that platform. The post gained early traction on Hacker News, drawing attention from the AI developer community ahead of the official launch.
This is an AI-generated summary. ShortSingh links to the original source for the complete article.
Discussion (0)
Log in to join the discussion and vote.
Log in