🐯Stalecollected in 5m

DeepSeek Financing Reveals Divergent AI Strategies of Tech Giants

DeepSeek Financing Reveals Divergent AI Strategies of Tech Giants
PostLinkedIn
🐯Read original on 虎嗅

💡Understand how Alibaba, Tencent, and ByteDance are using capital to shape the future of China's AI ecosystem.

⚡ 30-Second TL;DR

What Changed

Alibaba aims to integrate AI models into its ecosystem for control, while Tencent prefers a 'light' financial investment approach.

Why It Matters

The divergent strategies of these giants will shape the competitive landscape for AI infrastructure and application deployment in China.

What To Do Next

Evaluate whether your AI startup's growth strategy aligns with 'ecosystem integration' or 'independent infrastructure' before accepting capital from big tech.

Who should care:Founders & Product Leaders

Key Points

  • Alibaba aims to integrate AI models into its ecosystem for control, while Tencent prefers a 'light' financial investment approach.
  • ByteDance is aggressively investing 200 billion in AI, focusing on the C-end 'Doubao' app to capture user time.
  • DeepSeek maintains independence to avoid being 'locked' into any single giant's infrastructure.

🧠 Deep Insight

Web-grounded analysis with 20 cited sources.

🔑 Enhanced Key Takeaways

  • DeepSeek's recent financing round is expected to raise its valuation to as much as US$50 billion, a fivefold increase from initial reports, with participation from state-linked investors and Tencent.
  • DeepSeek's latest models, the DeepSeek-V4 series, are notably optimized to run on Huawei chips for inference, signaling a strategic move towards reducing reliance on US technology and fostering a sovereign AI ecosystem within China.
  • ByteDance's consumer-facing AI app, Doubao, which has surpassed 100 million daily active users, is now testing tiered subscription models, marking a significant industry shift in China from aggressive user acquisition to commercial viability and revenue generation.
  • Alibaba Cloud has committed RMB 380 billion ($53.4 billion) in AI and cloud infrastructure investments over the next three years, aiming to establish itself as a full-stack AI service provider with its Qwen models functioning as an 'operating system of the AI era.'
  • Tencent is not only a financial participant in DeepSeek's funding but is also actively developing its proprietary Hunyuan LLM, integrating it across its extensive product ecosystem like WeChat and QQ, and focusing on building AI agents, infrastructure, and knowledge bases.
📊 Competitor Analysis▸ Show

LLM Comparison: DeepSeek, Alibaba (Qwen), Tencent (Hunyuan), and ByteDance (Doubao)

Feature/ModelDeepSeek-V2DeepSeek-V4-ProAlibaba Qwen3-MaxTencent Hunyuan T1ByteDance Doubao (underlying model)
ArchitectureMixture-of-Experts (MoE)Mixture-of-Experts (MoE)Dense (over 1 trillion parameters)Proprietary (Deep-thinking model)Proprietary MoE multimodal model
Total Parameters236 Billion1.6 Trillion>1 TrillionNot specified, but aims for larger parametersNot specified, but aims for larger parameters
Activated Params21 Billion per token49 Billion per tokenAll parametersNot specifiedNot specified
Context Length128K tokens1 Million tokensNot specifiedNot specifiedNot specified
Key InnovationsMulti-head Latent Attention (MLA), DeepSeekMoEOptimized for Huawei chips, 1M context windowMultimodal (VL, Omni), Agent development platformMultimodal (Vision, Voice, 3D), AI AgentsAggressive low-cost pricing (initially), multimodal
Hardware Opt.Nvidia (historically), now Huawei for inferenceHuawei chips for inferenceOptimized for Alibaba Cloud infrastructureAligns with domestic chips, Nvidia H20 ordersPotentially in-house AI chips
Benchmark Perf.Comparable to GPT-4 (with fewer activated params)Strong performance on agentic coding (SWE-bench)On par with top closed-source models (SWE-Bench)Outperformed GPT-4.5 and DeepSeek R1 on benchmarksOn par with OpenAI's GPT-4o (Doubao-1.5-Pro)
Pricing (C-end)N/A (focus on open-source/API)N/A (focus on open-source/API)Free (Tongyi, Yuanbao)Free (Yuanbao)Tiered subscription (68-500 yuan/month)
StrategyIndependent, open-source, cost-efficientEcosystem control, full-stack AI, open-source QwenFinancial participation, internal product integrationDirect C-end competition, user time capture
Ecosystem FocusBroad developer adoptionAlibaba Cloud, enterprise software ecosystemTencent products (WeChat, QQ, Meeting)Doubao app, Douyin e-commerce, Doubao smartphone

🛠️ Technical Deep Dive

  • DeepSeek-V2: This is a Mixture-of-Experts (MoE) language model with 236 billion total parameters, activating only 21 billion parameters per token for efficient inference. It incorporates innovative architectures such as Multi-head Latent Attention (MLA) to significantly reduce KV cache requirements and DeepSeekMoE for economical training and inference. It supports a native context length of 128K tokens.
  • DeepSeek-V4 Series: Released in April 2026, this series includes DeepSeek-V4-Pro (1.6-trillion parameters, 49 billion activated per token) and DeepSeek-V4-Flash (284-billion parameters). Both models feature a substantial 1-million token context window. A key technical advancement is their optimization to run efficiently on Huawei chips, particularly for inference, marking a significant step towards hardware-software co-optimization within China.
  • DeepSeek-V3: This model is a 671 billion parameter MoE architecture, with 37 billion active parameters. It was noted for its performance and cost-effectiveness.
  • Training and Optimization: DeepSeek models are pre-trained on high-quality, multi-source corpora (e.g., 8.1T tokens for DeepSeek-V2) and undergo Supervised Fine-Tuning (SFT) and Reinforcement Learning (RL) to enhance their capabilities. The MoE structure dynamically selects relevant expert sub-networks for each input, reducing computational costs while maintaining high performance.

🔮 Future ImplicationsAI analysis grounded in cited sources

China's AI industry will see accelerated vertical integration and hardware-software co-optimization, driven by national strategic goals.
DeepSeek's optimization for Huawei chips and Alibaba's full-stack AI strategy, coupled with state-backed investments, indicate a strong push for domestic technological independence and control over the entire AI value chain.
The commercialization and monetization of consumer-facing AI applications in China will intensify, shifting from user growth to profitability.
ByteDance's Doubao app introducing tiered subscription models, despite its massive user base, signals a broader industry trend where Chinese AI developers are pressured to generate revenue to justify heavy investments and computing costs.
Independent AI research labs like DeepSeek will face increasing challenges in maintaining autonomy amidst the escalating capital requirements and ecosystem control ambitions of major tech giants.
DeepSeek's substantial funding rounds and the strategic interests of its investors (including state funds and Tencent) highlight the immense financial and infrastructural demands of frontier AI development, potentially making complete independence difficult to sustain long-term against ecosystem-building giants.

Timeline

2023-04
High-Flyer announced the launch of an AGI research lab.
2023-07
The AGI lab was spun off into an independent company, DeepSeek.
2023-11
DeepSeek LLM (67B parameters) was introduced, trained on 2 trillion tokens.
2025-03
DeepSeek released DeepSeek-V3-0324 under the MIT License.
2026-04
DeepSeek released a preview of its V4 series (V4-Pro, V4-Flash) with a 1-million token context window, optimized for Huawei chips.
2026-05
DeepSeek is anticipated to complete its first external financing round, potentially valuing the company at up to US$50 billion.
📰

Weekly AI Recap

Read this week's curated digest of top AI events →

👉Related Updates

AI-curated news aggregator. All content rights belong to original publishers.
Original source: 虎嗅