
Qwen 3.5 397B Beats Trillion-Param Rival Cheaply
Alibaba launched Qwen3.5-397B-A17B, an open-weight MoE model with 397B total parameters but only 17B active per token, outperforming its trillion-parameter Qwen3-Max on benchmarks. It offers 19x faster decoding at 256K context, 60% lower running costs, and native multimodal capabilities from scratch training on text, images, and video. The hosted version supports up to 1M tokens.


