Alibaba unveils Zhenwu M890 as NVIDIA alternative

๐กA key development in the AI chip war: Alibaba's new GPU aims to bypass US export controls.
โก 30-Second TL;DR
What Changed
Zhenwu M890 is a new GPU-class AI chip from Alibaba's T-Head.
Why It Matters
This launch highlights China's push for semiconductor self-sufficiency, potentially altering the competitive landscape for AI hardware in the region.
What To Do Next
Evaluate the performance benchmarks of the M890 against NVIDIA A100/H100 if your enterprise operates in the Chinese market.
Key Points
- โขZhenwu M890 is a new GPU-class AI chip from Alibaba's T-Head.
- โขDesigned as a domestic alternative to NVIDIA chips in China.
- โขThe chip is already in scaled mass production.
๐ง Deep Insight
Web-grounded analysis with 26 cited sources.
๐ Enhanced Key Takeaways
- โขThe Zhenwu M890 delivers three times the performance of its predecessor, the Zhenwu 810E, and is specifically engineered for emerging 'AI agents' and agentic AI workloads that demand extensive memory and high-speed communication for complex, multi-step tasks.
- โขThe chip features 144 GB of GPU memory and 800 GB per second interchip bandwidth, indicating a focus on larger-context, memory-intensive applications like multimodal models and long-context Large Language Models (LLMs).
- โขAlibaba's T-Head unit has already shipped over 560,000 Zhenwu units (including previous generations) to more than 400 customers across 20 industries, demonstrating significant market adoption and scaled mass production.
- โขAlibaba has outlined a multi-year chip roadmap, planning to release the Zhenwu V900 in Q3 2027 and the Zhenwu J900 in Q3 2028, with each successor expected to deliver roughly threefold performance gains over its predecessor.
- โขThe Zhenwu M890 supports multiple data formats from FP32 down to FP4, demonstrating adaptability to the extreme quantization trends prevalent in the large model community.
๐ ๏ธ Technical Deep Dive
- The Zhenwu M890 is a 'training-inference unified AI chip' designed for both AI training and inference tasks.
- It features 144 GB of GPU memory.
- The chip provides 800 GB per second interchip bandwidth.
- It natively supports various data precisions, ranging from FP32 down to FP4.
- The M890 is paired with Alibaba's proprietary ICN Switch 1.0 interconnect chip, which enables full-bandwidth 128-card connectivity.
- This interconnect system reduces communication latency to the hundred-nanosecond level, allowing 128 AI chips to operate collaboratively as a single computing unit.
- The chip is purpose-built to handle the heavy memory and communication demands of agent workloads, where models must retain long stretches of context and coordinate in real time.
- The predecessor, Zhenwu 810E, was equipped with 96 GB of High Bandwidth Memory 2e (HBM2e) and an inter-chip bandwidth of up to 700 GB per second.
๐ฎ Future ImplicationsAI analysis grounded in cited sources
โณ Timeline
๐ Sources (26)
Factual claims are grounded in the sources below. Forward-looking analysis is AI-generated interpretation.
- business-standard.com
- investing.com
- biggo.com
- wtaq.com
- moomoo.com
- intellectia.ai
- benzinga.com
- kaohooninternational.com
- morningstar.com
- marketscreener.com
- letsdatascience.com
- tmtpost.com
- futunn.com
- trendforce.com
- yicaiglobal.com
- scmp.com
- cryptorank.io
- rcrwireless.com
- substack.com
- laweconcenter.org
- tekedia.com
- youtube.com
- tomshardware.com
- brookings.edu
- laweconcenter.org
- oplexa.com
Weekly AI Recap
Read this week's curated digest of top AI events โ
๐Related Updates
AI-curated news aggregator. All content rights belong to original publishers.
Original source: The Next Web (TNW) โ



