Latest AI Chips News & Updates
The silicon arms race: NVIDIAβs challengers, custom ASICs, memory bottlenecks and export-control fallout.
314 articles
DeepX Valuation Quadruples to $2.2B
South Korean AI chip designer DeepX has reportedly reached a valuation of approximately $2.2 billion, nearly four times its previous valuation. The company signed initial agreements for a Series D round, receiving 42 billion Korean won from existing investors BNW Investment and DS Asset Management.
Cerebras outperforms Nvidia by 21x in specific benchmarks
Cerebras has demonstrated significant performance advantages by utilizing wafer-scale chip technology. While technically superior in specific domains, the article notes the challenges of competing in the broader market.
Google Develops Frozen Chip for Gemini AI Efficiency
Google is developing a new 'Frozen' chip architecture designed to significantly enhance the operational efficiency of its Gemini AI models. This hardware innovation aims to optimize compute resources for large-scale model inference.
VeriSilicon Secures 6.4B RMB in AI-Focused Orders
VeriSilicon announced 6.413 billion RMB in new orders between April 30 and July 16, 2026. Over 90% of these orders are focused on AI computing power and data processing.
Apple seeks acquisitions for AI server chips
Apple is actively scouting for semiconductor startups and engaging with bankers to acquire AI server chip technology. This move aims to address performance issues within their internal AI server infrastructure.
Anthropic partners with Samsung to develop custom AI chips
Anthropic is in talks with Samsung to develop custom AI chips to secure computing sovereignty and optimize inference performance. This reflects a broader trend of large model companies moving upstream into hardware.
Weekly AI Roundup: DeepSeek, OpenAI, and Market Shifts
This week's roundup covers DeepSeek's custom inference chips, OpenAI's IPO rumors, and significant price competition from Chinese AI models. It also highlights new model developments from MiniMax and xAI.
Synopsys pivots from fab software to AI chip design
Synopsys is shifting its engineering focus away from semiconductor manufacturing control software to prioritize the more profitable AI chip design market. This strategic pivot aims to capture higher margins in the AI hardware sector.
Intel plans 14A Gen2 process to challenge 1.4nm rivals
Intel is reportedly upgrading its roadmap to include a '14A Gen2' process to compete with TSMC and Samsung's upcoming 1.4nm chip manufacturing technologies. This move emphasizes Intel's strategy to maintain competitiveness in the high-performance AI chip foundry market.
Inference chips are reshaping the AI compute landscape
The dominance of general-purpose GPUs is being challenged by specialized inference chips (ASICs) designed to optimize cost, latency, and power efficiency. Major AI firms are increasingly developing custom silicon to break free from GPU dependency and improve token economics.
Anthropic aims to dominate AI chip development
Anthropic, a leading AI research company, is reportedly shifting focus toward developing its own AI chips after 5 years of operation. This move aims to secure hardware sovereignty for their large language models.
Samsung Wins Meta AI Chip Contract for 2nm Process
Samsung Electronics has reportedly secured a contract worth over 10 trillion KRW to manufacture Meta's next-generation AI chips. The partnership focuses on producing 'MTIA' accelerators using Samsung's cutting-edge 2nm process.
Global supply chain faces 2021-level crisis
Global supply chains are under extreme pressure due to geopolitical conflicts, trade restructuring, and new regulatory costs like CBAM. AI chip supply remains highly constrained, while China's domestic AI chip self-sufficiency is rising rapidly.
Meituan Validates Large-Scale Domestic AI Compute Clusters
Meituan successfully deployed a 50,000-card domestic AI chip cluster (likely Huawei Ascend) for its 1.6 trillion parameter model, LongCat-2.0. This marks a shift in the domestic supply chain from single-chip breakthroughs to full-system commercial delivery.
Qualcomm Nears $4 Billion Deal for Modular
Qualcomm is in advanced talks to acquire AI chip startup Modular in a deal valued at approximately $4 billion. This move signals Qualcomm's intent to strengthen its AI hardware and software stack capabilities.
Google and MediaTek partner on next-gen TPU v9
Google is deepening its partnership with MediaTek to develop the 'Triggerfish' TPU v9 chip. The project focuses on advancing AI agents, reinforcement learning, and maximizing compute efficiency.
Co-Designed Chip Cuts DeepSeek V4 Inference Costs by 75%
A SemiAnalysis report reveals that the co-design of DeepSeek V4 and the Huawei Ascend 950DT AI accelerator has achieved a 75% reduction in inference costs. This highlights the efficiency gains possible through hardware-software co-optimization.
Jensen Huang: Nvidia Does Not Want to 'Lose'
Following an $81.6 billion revenue report, Jensen Huang expressed a strong determination to maintain Nvidia's competitive edge. The statement highlights the company's aggressive stance in the global AI hardware market.
Alibaba unveils Zhenwu M890 as NVIDIA alternative
Alibaba's T-Head unit has unveiled the Zhenwu M890, a new GPU-class AI chip. The chip is designed to serve as a domestic alternative to NVIDIA hardware amidst tightening US export controls.
Groq's AI inference chips disrupt market with massive valuation
Groq has gained significant market attention for its specialized AI inference chips, which offer high-speed performance that challenges Nvidia's dominance. The company's recent valuation surge highlights the growing demand for dedicated inference hardware in the AI ecosystem.
Huang: Top AI Firms Not Ditching CUDA
Nvidia CEO Jensen Huang dismisses claims that leading AI companies are abandoning CUDA. He states the premise is fundamentally wrong. A detailed 10,000-word transcript is provided.
Samsung Secures $200B Broadcom AI Infrastructure Deal
Samsung has signed a five-year memorandum of understanding with Broadcom worth $200 billion to develop next-generation AI infrastructure. This strategic partnership aims to challenge TSMC's dominance by offering comprehensive 2nm-class foundry solutions.
Alibaba open-sources SAIL to challenge Nvidia's CUDA dominance
Alibaba's T-Head unit has open-sourced SAIL, a comprehensive software stack designed for its Zhenwu AI chip series. This initiative aims to reduce developer dependency on Nvidia's CUDA ecosystem by facilitating easier migration to alternative hardware.
Alibaba open-sources SAIL stack to challenge Nvidia's CUDA
Alibaba's T-Head unit has open-sourced its SAIL software stack, which powers its Zhenwu AI chip series. This move aims to reduce developer reliance on Nvidia's proprietary CUDA ecosystem.
Suiyuan Technology's 4.7 billion loss and future
Suiyuan Technology faces significant financial losses while struggling to break into the high-end training chip market, relying heavily on inference products and Tencent as a primary customer.
HBM4 Prices Projected to Double by Late 2026
Due to surging AI demand and manufacturing complexities, HBM4 prices are expected to rise significantly by late 2026. The production process is constrained by low initial yields and high wafer consumption compared to standard DDR5 DRAM.
Enflame Technology IPO Registration Effective on STAR Market
Shanghai Enflame Technology Co., Ltd. has officially received approval for its IPO registration on the Shanghai Stock Exchange STAR Market. This marks a significant milestone for the domestic AI chip developer.
Enflame Technology Approved for STAR Market IPO
Enflame Technology has received official approval from the CSRC to proceed with its IPO on the Shanghai Stock Exchange's STAR Market. This marks a significant milestone for the Chinese AI chip developer as it prepares to enter the public capital markets.
Nvidia Partners with d-Matrix for AI Inference Systems
Nvidia is shifting its strategy by collaborating with AI chip startup d-Matrix. The partnership aims to integrate their respective hardware to create a new computing system specifically for large model inference.
DeepSeek and Zhipu AI Pivot to Custom Silicon
Leading Chinese AI labs DeepSeek and Zhipu AI are joining global peers in developing proprietary inference chips. This strategic shift aims to mitigate reliance on high-cost GPUs and optimize performance for their specific model architectures.