Qwen 3.8 Max
Alibaba opened the weights of its closed flagship tier, and the file it published turned out to be a narrower thing than the product it was named after.
Alibaba previewed Qwen 3.8 Max on July 19, 2026, at the World AI Conference in Shanghai, and made the hosted model generally available on August 3. It is a mixture-of-experts design with 2.4 trillion total parameters and 95 billion activated per token, drawing on 512 experts of which ten routed and one shared fire for any given token. The published architecture is unusual in the open: 92 layers arranged as repeating blocks of gated DeltaNet and gated attention, each followed by a mixture-of-experts layer.
The Max tier had always been the closed half of Alibaba's catalogue. Smaller Qwen models had shipped with open weights for years, and the Max models were the ones you could only rent. In August the company published the weights as Qwen3.8-2.4T-A95B on Hugging Face under a licence named for the model, making this the first member of the Max class anyone could download.
The downloadable file is not the same object as the hosted service, and the model card says so. The released weights are text only and require thinking mode for every interaction, while the product announced in July was multimodal across images, video and documents. The context window is 262,144 tokens natively, extensible to roughly 1,010,000.
Key Facts
- 012.4 trillion total parameters, 95 billion activated per token, across 512 experts with ten routed and one shared active at a time.
- 02Previewed July 19, 2026 at the World AI Conference in Shanghai; the hosted model became generally available August 3, 2026.
- 03Weights published in August 2026 as Qwen3.8-2.4T-A95B, the first model in the Qwen-Max class to be made downloadable.
- 04Context of 262,144 tokens natively, extensible to about 1,010,000.
- 05The released weights are text only and require thinking mode; the hosted Max product was announced as multimodal.
Three weeks earlier Moonshot had published Kimi K3 at 2.8 trillion parameters, and the obvious reading of Qwen 3.8 Max is that a second Chinese laboratory matched the gesture. The more interesting fact is the one about tiers. Kimi K3 was an open model from a company whose models are open. This was a company reaching into the part of its catalogue that had always been rented and putting it on a download page, which is a decision about a business rather than about a release.
It also complicates what an open-weight release means. For most of the period this museum covers, weights published meant the thing itself, and a reader could assume the file and the service were the same model. Here they are not: one is multimodal and hosted, the other is text only and requires a particular mode to run at all. Asking whether a model is open is no longer a question with one answer, and the honest version of the question is which artifact, with which capabilities, under which licence.
Qwen3.8-2.4T-A95B model card
Qwen Team, Alibaba Cloud · 2026
https://huggingface.co/Qwen/Qwen3.8-2.4T-A95B
Qwen3.8
Qwen Team, Alibaba Cloud · 2026
https://qwen.ai/blog?id=qwen3.8
Alibaba's open-weight Qwen3.8-Max takes on long-horizon AI tasks with 2.4 trillion parameters
The Decoder · 2026
https://the-decoder.com/alibabas-open-weight-qwen3-8-max-takes-on-long-horizon-ai-tasks-with-2-4-trillion-parameters/