來源虎嗅•較早收集於 29m
DeepSeek-V4 基準測試領先,估值達百億美元

💡V4基準勝GPT-5.3;百億融資支持國產晶片轉移(28字)
⚡ 30 秒速覽
有什麼變化
洩露基準:MMLU-Pro 91.2 勝過GPT-5.3的88.4
為什麼重要
提升中國AI自主性,應對晶片禁令,若遷移成功可實現全球更廉價推理。吸引資金擴張,面對高期望。
下一步行動
使用DeepSeek-V4的SWE-bench 59.6分數基準測試您的程式碼任務。
誰應關注:Researchers & Academics
關鍵要點
- •洩露基準:MMLU-Pro 91.2 勝過GPT-5.3的88.4
- •3億美元融資輪,目標100億美元估值
- •V4延遲發布,因全面遷移至華為昇騰而非NVIDIA CUDA
- •強調超低Token成本,打造「Token工廠」商業化
🧠 深度解析
本篇為 AI 生成分析,非原文內容。
🔑 增強重點摘要
- •DeepSeek's migration to Huawei Ascend chips is part of a broader 'Project Sovereign' initiative aimed at insulating Chinese AI development from potential future US export control tightening on high-end NVIDIA hardware.
- •The $10B valuation reflects investor confidence in DeepSeek's proprietary 'Deep-MoE' routing algorithm, which reportedly achieves 40% higher compute efficiency than standard Mixture-of-Experts implementations.
- •Industry analysts suggest the 'Token factory' strategy aims to commoditize LLM inference by pricing tokens at sub-fractional costs, specifically targeting the integration of AI into low-power edge devices and IoT ecosystems.
📊 競品分析▸ Show
| Feature | DeepSeek-V4 | GPT-5.3 | Claude 3.5 Opus (Ref) |
|---|---|---|---|
| Architecture | Ascend-native MoE | NVIDIA-based Dense/MoE | Proprietary |
| MMLU-Pro | 91.2 | 88.4 | 86.7 |
| SWE-bench | 59.6 | 62.1 | 58.4 |
| Primary Market | China/Global (Low-cost) | Global (Premium) | Global (Enterprise) |
🛠️ 技術深入
- •Architecture: Evolution of the Deep-MoE (Mixture-of-Experts) framework, optimized for non-CUDA kernels.
- •Hardware Abstraction: Implementation of a custom software stack to map tensor operations directly to Huawei's CANN (Compute Architecture for Neural Networks) library.
- •Inference Optimization: Utilization of FP8 quantization across the entire model weight set to maximize throughput on Ascend 910B/C clusters.
- •Training Efficiency: Reported use of a novel 'Dynamic Load Balancing' technique to mitigate communication bottlenecks inherent in non-NVIDIA interconnects.
🔮 前景展望基於引用來源的 AI 分析
DeepSeek will trigger a price war in the Chinese LLM market.
The 'Token factory' commercialization strategy prioritizes extreme cost-efficiency over margin, forcing competitors to lower inference prices to retain market share.
Huawei Ascend chips will become the standard for domestic Chinese AI training.
DeepSeek's successful migration demonstrates the viability of the Ascend ecosystem, encouraging other major Chinese labs to reduce reliance on NVIDIA.
⏳ 時間線
2023-04
DeepSeek releases first open-source model series.
2024-01
DeepSeek-V2 launches with innovative MoE architecture.
2024-12
DeepSeek-V3 achieves parity with top-tier global models.
2026-02
DeepSeek initiates full-scale migration of training pipelines to Huawei Ascend infrastructure.
📰
AI 週報
閱讀本週精選 AI 大事摘要 →
👉相關動態
AI 策展新聞聚合。所有內容版權歸原始發布者所有。
原始來源: 虎嗅 ↗
每週電子報
每週一封,可隨時退訂。



