GPU cluster connection method and device, switch and storage medium
By using Leaf switches to interconnect server clusters within a GPU cluster, the connection rules and communication routes within and between clusters are optimized, solving the problems of high communication latency and low training efficiency in existing GPU cluster technologies, and realizing an efficient large-scale GPU cluster networking architecture.
Patent Information
- Application Number
- CN202510962123.1
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2025-07-11
- Publication Date
- 2025-10-28
AI Technical Summary
Existing GPU cluster networking architectures suffer from high communication latency and low training efficiency in large-scale parametric model training. In particular, the two-layer fat tree architecture increases network costs and communication hops, leading to load imbalance and network congestion.
A novel GPU cluster connection method is adopted to interconnect Leaf switches with the same sequence number from different server clusters, thereby expanding the scale of the GPU cluster. By optimizing the connection rules and communication routes within and between clusters, communication latency is reduced and training efficiency is improved.
Without increasing network layers, the GPU cluster size was expanded, communication latency was reduced, model training efficiency was improved, and network costs were saved.
Smart Images

Figure CN120849336A_ABST
Abstract
Citation Information
Cited By
Computer system, server optimization method, electronic device and storage medium
CN121217598A
Intelligent computing cluster networking topology method and device, electronic equipment and storage medium
CN121357190A
Artificial intelligence cluster networking topology method and device, electronic equipment and storage medium
CN121357190B