SourceStalecollected in 1m

Alibaba Qwen3.5-Max Tops China, Trails US

Alibaba Qwen3.5-Max Tops China, Trails US
PostLinkedIn
🇭🇰Read original on SCMP Technology
#china-ai#benchmarks#model-previewqwen3.5-max-previewalibabaqwen3.5-max-previewarenaanthropicopenai

💡China's top LLM beats peers on benchmarks, closing US gap

⚡ 30-Second TL;DR

What Changed

Qwen3.5-Max-Preview tops Chinese AI models on Arena rankings

Why It Matters

Strengthens China's domestic AI ecosystem, offering practitioners a high-performing alternative to US models. May accelerate competition and innovation in multimodal LLMs.

What To Do Next

Test Qwen3.5-Max-Preview on Arena to benchmark against Claude and GPT models.

Who should care:Researchers & Academics

Key Points

  • Qwen3.5-Max-Preview tops Chinese AI models on Arena rankings
  • Lags behind US leaders like Anthropic, Google, OpenAI
  • Flagship of Alibaba's Qwen 3.5 family now available for preview
  • Positions Alibaba as China's AI frontrunner

🧠 Deep Insight

Background and context from public sources — not the original article. 7 sources cited.

🔑 Enhanced Key Takeaways

  • The Qwen 3.5 family introduces a novel 'Hybrid Mixture-of-Experts' architecture utilizing Gated DeltaNet (linear attention), which enables a 1-million token context window while delivering up to 19x higher decoding throughput than the previous Qwen 3 generation.
  • Alibaba has expanded linguistic support to 201 languages and dialects, utilizing a massive 250,000-token vocabulary that improves encoding efficiency by up to 60% for non-English scripts compared to the 150,000-token limit in Qwen 3.
  • The release follows a significant leadership exodus in early 2026, including the departure of technical lead Lin Junyang (Justin Lin) and head of post-training Yu Bowen, sparking industry debate over Alibaba's long-term commitment to its open-source strategy.
  • Qwen 3.5-Max-Preview features a dual-mode 'Thinking' vs. 'Fast' inference capability, where the model can engage in internal chain-of-thought reasoning (via tags) to match US rivals in complex logic while maintaining a low-latency mode for routine tasks.
📊 Competitor Analysis▸ Show
FeatureQwen 3.5-Max-PreviewGemini 3.1 ProClaude 4.6 OpusGPT-5.4
Arena Elo~1451150515031485
Context Window1M (Hosted) / 262K (Native)2M+200K128K
Architecture397B MoE (17B Active)Proprietary MoEProprietaryProprietary
LicenseApache 2.0 (Open-Weight)ProprietaryProprietaryProprietary
Multilingual201 Languages150+ Languages95+ Languages100+ Languages
Pricing (per 1M)~$0.10 (Est. API)$1.25 (Input)$3.00 (Input)$2.50 (Input)

🛠️ Technical Deep Dive

Detailed technical specifications for the Qwen 3.5-397B-A17B model:

  • Parameter Count: 397 billion total parameters with a sparse Mixture-of-Experts (MoE) routing that activates only 17 billion parameters per token.
  • Attention Mechanism: A hybrid layout consisting of 60 layers where 15 groups of 3 'Gated DeltaNet' (linear attention) layers are interleaved with 1 'Gated Attention' layer to optimize memory usage for long-context sequences.
  • Multimodal Integration: Native 'early-fusion' vision-language architecture where text and visual tokens are processed within the same transformer backbone rather than using a separate adapter.
  • Training Scale: Pre-trained on an estimated 36+ trillion tokens with a heavy emphasis on synthetic 'agentic' data and reinforcement learning (RL) scaled across million-agent environments.
  • Inference Optimizations: Native support for Multi-Token Prediction (MTP) and SGLang/vLLM acceleration engines, achieving near-100% multimodal training efficiency.

🔮 Future ImplicationsAI analysis grounded in cited sources

Alibaba will pivot toward a 'Cloud-First' proprietary model tier.
The increasing performance gap between the open-weight 397B model and the closed-source 'Plus' and 'Max-Preview' versions suggests Alibaba is prioritizing its Model Studio ecosystem over pure open-source parity.
Qwen will become the dominant foundation for non-English AI agents.
With support for 201 languages and superior performance on regional benchmarks like C-Eval, it is positioned as the primary alternative to US models in Southeast Asia and the Middle East.

Timeline

2023-04
Tongyi Qianwen (Qwen) Beta Launch
2024-06
Qwen 2 Series Released with 72B Flagship
2024-09
Qwen 2.5 Launch with Enhanced Reasoning
2025-04
Qwen 3 Family Debut (Apache 2.0 License)
2026-02
Qwen 3.5 and 397B MoE Open-Weight Release
2026-03
Qwen 3.5-Max-Preview Deployed on LMSYS Arena

📎 Sources (7)

Factual claims are grounded in the sources below. Forward-looking analysis is AI-generated interpretation.

  1. Google Search Source
  2. Google Search Source
  3. Google Search Source
  4. Google Search Source
  5. Google Search Source
  6. Google Search Source
  7. Google Search Source
📰

Weekly AI Recap

Read this week's curated digest of top AI events →

👉Related Updates

AI-curated news aggregator. All content rights belong to original publishers.
Original source: SCMP Technology

This is a summary, not the original. Read the source, or get the weekly briefing.

The weekly digest

One email a week. Unsubscribe anytime.