T-Head Unveils Zhenwu M890 AI Chip with 3x Performance
New high-performance AI chip from T-Head with 3x performance boost for large-scale training clusters.
30-Second TL;DR
What Changed
Zhenwu M890 features 144GB memory and 800GB/s interconnect bandwidth.
Why It Matters
The M890 significantly boosts Alibaba's infrastructure capabilities for Agentic AI, providing a competitive alternative for high-performance computing clusters in the Chinese market.
What To Do Next
Evaluate the Zhenwu M890's specs for your large-scale model training workloads if you are operating within the Alibaba Cloud ecosystem.
Key Points
- •Zhenwu M890 features 144GB memory and 800GB/s interconnect bandwidth.
- •Delivers 3x performance improvement over the Zhenwu 810E.
- •Supports diverse data precisions from FP32 to FP4 for training and inference.
- •Pairs with ICN Switch 1.0 for 64-card full-bandwidth cluster interconnects.
Deep Insight
Background and context from public sources — not the original article. 18 sources cited.
Enhanced Key Takeaways
- •The Zhenwu M890 is a core component of Alibaba's 'AI Golden Triangle' strategy, which integrates T-Head's proprietary chips with Alibaba Cloud's infrastructure and the Qwen large language model family to create a comprehensive AI supercomputing system.
- •The chip is specifically optimized for 'agentic AI workloads,' which are characterized by their demand for extensive working memory to retain context and high-speed communication capabilities for complex reasoning.
- •Alibaba Cloud has introduced a 128-card supernode server, named Pangu AL128, which incorporates the Zhenwu M890 and the ICN Switch 1.0, achieving communication latency as low as the hundred-nanosecond level for integrated computing.
- •T-Head has demonstrated significant market penetration, having delivered 560,000 Zhenwu units to over 400 enterprise customers across more than 20 industries as of May 2026.
- •The Zhenwu M890's predecessor, the Zhenwu 810E, was reported to offer performance comparable to Nvidia's China-compliant H20 accelerator and superior to the Nvidia A800, indicating T-Head's competitive standing in the domestic market.
Competitor Analysis
- T-Head Zhenwu M890 (2026)
- 144GB
- T-Head Zhenwu 810E (2026)
- 96GB HBM2e
- Nvidia H20 (China-compliant)
- 96GB (implied)
- Huawei Ascend 910B (2022)
- Double on-chip memory of 910, HBM2e (specific capacity not detailed)
- T-Head Zhenwu M890 (2026)
- 800GB/s
- T-Head Zhenwu 810E (2026)
- 700GB/s
- Nvidia H20 (China-compliant)
- 700GB/s (implied)
- Huawei Ascend 910B (2022)
- Higher bandwidth than 910 (specific value not detailed)
- T-Head Zhenwu M890 (2026)
- 3x Zhenwu 810E
- T-Head Zhenwu 810E (2026)
- Comparable to Nvidia H20, surpasses A800
- Nvidia H20 (China-compliant)
- Comparable to Zhenwu 810E
- Huawei Ascend 910B (2022)
- Matches/surpasses Nvidia A100 in some tests, 80% A100 efficiency in LLM training, 20% better in others
- T-Head Zhenwu M890 (2026)
- FP32 to FP4
- T-Head Zhenwu 810E (2026)
- Not specified
- Nvidia H20 (China-compliant)
- Not specified
- Huawei Ascend 910B (2022)
- FP16, INT8 (Ascend 910)
- T-Head Zhenwu M890 (2026)
- Not specified
- T-Head Zhenwu 810E (2026)
- Not specified
- Nvidia H20 (China-compliant)
- Not specified
- Huawei Ascend 910B (2022)
- SMIC's 7nm fabrication process
- T-Head Zhenwu M890 (2026)
- Proprietary parallel computing architecture
- T-Head Zhenwu 810E (2026)
- Proprietary parallel computing architecture
- Nvidia H20 (China-compliant)
- Not specified
- Huawei Ascend 910B (2022)
- DaVinci cores (25 active cores in 910B vs 32 in 910)
- T-Head Zhenwu M890 (2026)
- Not available
- T-Head Zhenwu 810E (2026)
- Alibaba Cloud pricing increased 5-34% (March 2026)
- Nvidia H20 (China-compliant)
- Not available
- Huawei Ascend 910B (2022)
- Up to 40% lower cost than imported alternatives (for T-Head PPU generally)
| Feature / Chip | T-Head Zhenwu M890 (2026) | T-Head Zhenwu 810E (2026) | Nvidia H20 (China-compliant) | Huawei Ascend 910B (2022) |
|---|---|---|---|---|
| Memory | 144GB | 96GB HBM2e | 96GB (implied) | Double on-chip memory of 910, HBM2e (specific capacity not detailed) |
| Interconnect B/W | 800GB/s | 700GB/s | 700GB/s (implied) | Higher bandwidth than 910 (specific value not detailed) |
| Performance | 3x Zhenwu 810E | Comparable to Nvidia H20, surpasses A800 | Comparable to Zhenwu 810E | Matches/surpasses Nvidia A100 in some tests, 80% A100 efficiency in LLM training, 20% better in others |
| Data Precision | FP32 to FP4 | Not specified | Not specified | FP16, INT8 (Ascend 910) |
| Manufacturing Process | Not specified | Not specified | Not specified | SMIC's 7nm fabrication process |
| Architecture | Proprietary parallel computing architecture | Proprietary parallel computing architecture | Not specified | DaVinci cores (25 active cores in 910B vs 32 in 910) |
| Pricing | Not available | Alibaba Cloud pricing increased 5-34% (March 2026) | Not available | Up to 40% lower cost than imported alternatives (for T-Head PPU generally) |
Technical Deep Dive
- The Zhenwu M890 features 144GB of memory and 800GB/s interconnect bandwidth.
- It supports diverse data precisions ranging from FP32 down to FP4 for both training and inference tasks.
- The chip is integrated into Alibaba Cloud's new 128-card supernode server, known as Pangu AL128.
- The M890 pairs with the ICN Switch 1.0, which facilitates 64-card full-bandwidth cluster interconnects and reduces communication latency to the hundred-nanosecond level.
- Its predecessor, the Zhenwu 810E, was built on a proprietary parallel computing architecture and utilized in-house inter-chip interconnect technology, complemented by a fully self-developed software stack for end-to-end vertical integration.
Future ImplicationsAI analysis grounded in cited sources
Timeline
- 2018-09Alibaba establishes T-Head (Pingtouge) as its semiconductor subsidiary.
- 2019-09Alibaba announces Hanguang 800, its first self-developed AI inference chip.
- 2025-09T-Head develops a new series of AI accelerators (PPU), with performance comparable to Nvidia's H20 and A800 GPUs, deployed by China Unicom.
- 2026-01-29T-Head officially unveils the Zhenwu 810E AI chip on its website, confirming performance comparable to Nvidia H20.
- 2026-03-20T-Head's Zhenwu 810E GPU reaches large-scale production, with 470,000 units delivered as of February 2026.
- 2026-05-20Alibaba's T-Head launches the Zhenwu M890 AI chip, offering 3x performance over its predecessor.
Sources (18)
Factual claims are grounded in the sources below. Forward-looking analysis is AI-generated interpretation.
Weekly AI Recap
Read this week's curated digest of top AI events →
AI-curated news aggregator. All content rights belong to original publishers.
Original source: 36氪 ↗
This is a summary, not the original. Read the source, or get the weekly briefing.
The weekly digest
One email a week. Unsubscribe anytime.