Investment Rating - The report does not explicitly state an investment rating for the industry. Core Insights - Huawei is exploring a "soft and hard integration" strategy to enhance its AI competitiveness, transitioning from merely catching up with industry SOTA models to customizing model architectures for its self-developed Ascend hardware [12][30]. - The evolution of the Pangu model series reflects a shift from parameter competition to a focus on efficiency and scalability, culminating in the adoption of the Mixture of Experts (MoE) architecture [12][30]. - The report highlights the introduction of innovative architectures like Pangu Pro MoE and Pangu Ultra MoE, which aim to maximize the utilization of Ascend hardware through structural and system-level optimizations [36][62]. Summary by Sections 1. Evolution of Pangu Models - The Pangu model series began with PanGu-α, a 200 billion parameter model, which established a technical route based on Ascend hardware [12][30]. - PanGu-Σ, launched in 2023, marked an early attempt at sparsification, exploring trillion-parameter models with a focus on efficiency [15][18]. - Pangu 3.0 introduced a "5+N+X" architecture aimed at deep industry applications, showcasing its capabilities in various sectors [22][23]. 2. Pangu Pro MoE and Pangu Ultra MoE - Pangu Pro MoE addresses the challenge of expert load imbalance in distributed systems through a new architecture called Mixture of Grouped Experts (MoGE) [36][37]. - The MoGE architecture ensures load balancing by structuring the selection of experts, thus enhancing efficiency in distributed deployments [45][46]. - Pangu Ultra MoE emphasizes system-level optimization strategies to explore the synergy between software and hardware, reflecting a practical application of the soft and hard integration concept [62]. 3. CloudMatrix Infrastructure - CloudMatrix serves as the physical foundation for AI infrastructure, enabling high-performance communication and memory management across distributed systems [5][10]. - The infrastructure supports the Pangu models by providing a unified addressing distributed memory pool, which reduces performance discrepancies in cross-node communication [5][10]. 4. Full-Stack Collaboration - Huawei's AI strategy is centered around full-stack collaboration, integrating open-source strategies to build an ecosystem around Ascend hardware [10][12]. - The architecture, systems, and operators form the three pillars of this full-stack collaboration, aimed at enhancing the overall efficiency and effectiveness of AI solutions [10][12].
产业深度:【AI产业深度】华为盘古大模型与昇腾AI计算平台,共同构建软硬一体的AI技术体系
GUOTAI HAITONG SECURITIES·2025-08-06 09:19