Alibaba Releases Qwen3.8-Max: A 2.4 Trillion Parameter MoE Model

Alibaba's Qwen team has released Qwen3.8-Max, a 2.4-trillion-parameter Mixture-of-Experts model positioned as the most capable model in the Qwen family to date. The release targets the frontier of open-weight model performance, directly competing with leading Western models on reasoning and instruction-following benchmarks. For developers, this represents a significant new open-weight option at the very top of the capability ladder, with potential deployment on owned infrastructure rather than via proprietary APIs. The scale of the MoE architecture means only a subset of parameters activate per inference, keeping compute requirements more manageable than a dense model of equivalent size. Teams building with Qwen or evaluating alternatives to GPT or Claude should benchmark this release against their specific workloads immediately.
Read original source ↗Part of the 2026-08-04 digest→