🇭🇰SCMP Technology•較早收集於 3m
阿里巴巴發布全新 Qwen 模型與自研 AI 晶片

💡阿里巴巴正在打造全端「AI 工廠」,旨在挑戰全球在模型訓練與推論基礎設施領域的領先地位。
⚡ 30-Second TL;DR
有什麼變化
推出新一代 Qwen 模型,強化訓練與推論效能。
為什麼重要
阿里巴巴垂直整合軟硬體的舉措,象徵其意圖主導中國 AI 基礎設施市場。這可能大幅降低依賴 Alibaba Cloud 進行大規模模型部署的開發者成本。
下一步行動
查閱 Alibaba Cloud 關於最新 Qwen 模型 API 的文件,評估其效能與目前開源替代方案的差異。
誰應關注:Developers & AI Engineers
關鍵要點
- •推出新一代 Qwen 模型,強化訓練與推論效能。
- •開發自研 AI 晶片以支援大規模模型運算負載。
- •轉型為「AI 工廠」模式,專注於透過運算密集型服務創造營收。
- •將雲端基礎設施與自主代理(Autonomous Agents)能力進行整合。
🧠 深度解析
Web-grounded analysis with 15 cited sources.
🔑 增強重點摘要
- •Alibaba's AI-related product revenue has reached an annualized run rate of CNY 35.8 billion (approximately USD 5.2 billion) and currently accounts for 30% of its Cloud Intelligence Group's external revenue, with expectations to exceed 50% within a year.
- •The newly introduced Zhenwu M890 AI chip, developed by Alibaba's T-Head subsidiary, delivers three times the performance of its predecessor, the Zhenwu 810E, and is specifically engineered to handle the high memory and communication demands of autonomous AI agent workloads.
- •Alibaba has committed to an investment exceeding RMB 380 billion (approximately USD 53 billion) over three years in cloud and AI infrastructure, indicating a significant capital expenditure to support its 'AI factory' ambitions.
- •The Qwen model family includes both open-source (e.g., Qwen 3.5-7B, MIT-licensed) and proprietary variants, with some open-weight models demonstrating superior performance in benchmarks like MMLU compared to competitors such as GPT-4o-mini, at significantly lower inference costs.
- •Alibaba's 'AI factory' strategy is driven by the rapid expansion of AI agent workloads, encompassing training, inference, and orchestration, with current demand for compute capacity reportedly outstripping available supply.
📊 競品分析▸ Show
| Feature/Metric | Alibaba Qwen 3.5-7B (Alibaba Cloud API) | OpenAI GPT-4o-mini | Anthropic Claude 3.5 Haiku | Google Gemma 3-9B |
|---|---|---|---|---|
| MMLU Score | 74.2% | 72.9% | N/A | N/A |
| Input Price (per 1M tokens) | $0.008 | $0.15 | $0.08 | $0.03 |
| Output Price (per 1M tokens) | $0.01 (Together AI) | N/A | N/A | N/A |
| License | MIT-licensed (for open-weight models) | Proprietary | Proprietary | Proprietary |
| On-device Inference | Yes (0.5B model on iPhone 15 Pro at 40 tokens/sec) | No (API-only) | No (API-only) | N/A |
Note: Qwen2.5-Max, an MoE model, is also positioned to compete with DeepSeek V3, OpenAI's GPT-4o, Anthropic's Claude-3.5-Sonnet, Meta's Llama-3.1–405B, and Google's Gemini 2.0 Flash, with claims of outperforming them in certain aspects.
🛠️ 技術深入
- Qwen Model Architecture: The Qwen family includes both dense transformer models (like Qwen 3.5, optimized for edge hardware) and Mixture-of-Expert (MoE) architectures (like Qwen 2.5-Max).
- Qwen Training Data: Qwen 2.5-Max was pretrained on over 20 trillion tokens, encompassing multi-lingual textual data and domain-specific corpora.
- Qwen Multimodality: Models like Qwen-Omni are end-to-end multimodal, processing text, images, audio, and video, and delivering real-time streaming responses.
- Qwen Thinking Modes: Qwen3 models feature hybrid 'Thinking' and 'Non-Thinking' modes, allowing flexible control over reasoning performance, speed, and costs.
- Zhenwu M890 AI Chip: This custom chip features 144 gigabytes of GPU memory, an upgrade from its predecessor's 96 gigabytes, enabling it to process significantly larger data for complex AI agent workloads.
- Hanguang 800 AI Chip (Predecessor): Alibaba's first AI inference chip, launched in 2019, was built on a 12-nm process with 17 billion transistors. It achieved a peak performance of 78,563 images per second (IPS) on ResNet-50 inference tests.
🔮 前景展望AI analysis grounded in cited sources
Alibaba's AI-related product revenue will become the primary driver of its Cloud Intelligence Group's external revenue.
Management expects AI-related product revenue to exceed 50% of external cloud revenue in about a year, pivoting the business more toward AI compute and agent services.
Alibaba will continue to rapidly advance its custom AI chip technology with a clear roadmap for future generations.
The company has outlined a multi-year chip roadmap, planning to launch the V900 in Q3 2027 and the J900 in Q3 2028, each expected to deliver significant performance gains.
The Qwen model family will solidify its position as a foundational 'operating system' for AI development, particularly in the open-source community.
Alibaba remains committed to open-sourcing Qwen models and aims to shape it into the 'operating system of the AI era,' empowering developers globally.
⏳ 時間線
2009-09
Alibaba Cloud officially established.
2017-Late
DAMO Academy, Alibaba's research institute, launched.
2018-09
T-Head, Alibaba's semiconductor division, spun out of DAMO Academy.
2019-09
Alibaba unveils Hanguang 800, its first AI inference chip.
2023-04
Alibaba launches beta of Tongyi Qianwen (Qwen) large language model.
2026-05
Alibaba unveils Zhenwu M890 AI chip and Qwen 3.7-Max model.
📎 來源 (15)
Factual claims are grounded in the sources below. Forward-looking analysis is AI-generated interpretation.
📰
AI 週報
閱讀本週精選 AI 大事摘要 →
👉相關動態
AI 策展新聞聚合。所有內容版權歸原始發布者所有。
原始來源: SCMP Technology ↗
