Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

15 results about "Collective communication" patented technology

Communication is the substratum which enables collective action between individual things. The corporation and bee colony both seem to exhibit collective consciousness. Collective consciousness results in aligned action involving multiple individuals, typically according to a sense of identity or goals.

Ensemble communication method, ensemble communication system and related device

The embodiment of the invention discloses a collective communication method, the method is applied to a collective communication system, the collective communication system comprises a management node and N computing nodes, the N computing nodes are used for processing a first communication task in parallel, N is an integer greater than 1, and the method comprises the following steps: the management node receives first information, the first information indicates that at least one first link breaks down, and each first link is a communication link used for processing the first communication task among the N computing nodes; the management node receives N pieces of second information from the N computing nodes, wherein each piece of second information is used for indicating whether the original data of the first communication task locally stored by one computing node is changed or not; and if the original data of the N computing nodes are not changed, the management node sends third information to the N computing nodes to instruct the N computing nodes to reprocess the first communication task based on the original data. In this way, the service recovery time can be shortened after the communication fault occurs.
Owner:HUAWEI TECH CO LTD

Ensemble communication unloading method, system, equipment and medium

ActiveCN121979690AResource allocationInference methodsCollective communicationComputer network
The invention discloses a set communication unloading method, system and device and a medium, and is applied to the technical field of computers, and the method comprises the steps: in a model deployment stage, a distributed reasoning controller generates a communication primitive blueprint based on a computational graph description file of a tensor parallel reasoning model and issues the communication primitive blueprint to a DPU; the DPU establishes a hardware-level communication context semantic environment based on the communication primitive blueprint; in the model reasoning stage, the GPU / NPU sends a trigger signal to the DPU when calculating to a communication boundary; and the DPU executes a DMA data pulling assembly line and an RDMA data sending assembly line in parallel based on a hardware-level communication context semantic environment, performs aggregation calculation on all tensor fragment data to be synchronized, and writes an aggregation calculation result back to the GPU / NPU. A hardware-level communication context semantic environment is established in advance to a model deployment stage, and double assembly lines are executed in the DPU in parallel, so that end-to-end communication delay is remarkably reduced, and zero participation of a host CPU is realized.
Owner:YIHUA TECHNOLOGY (BEIJING) CO LTD +1

Collective communication method and computing device cluster

PCT designated stageWO2026091617A1TransmissionCollective communicationData pack
The present application relates to the field of computer technology, and discloses a collective communication method and a computing device cluster. The method comprises: during the process of N computing units in a computing device cluster executing a collective communication operation, a first computing unit receives first data from a second computing unit and / or sends second data to a third computing unit, wherein when data in each computing unit among the N computing units is greater than a first threshold, the data in each computing unit is segmented into N pieces of slice data, the number of pieces of slice data in the second computing unit comprised in the first data is inversely proportional to a physical distance between the first computing unit and the second computing unit, and the number of pieces of slice data in the first computing unit comprised in the second data is inversely proportional to a physical distance between the first computing unit and the third computing unit. Thus, cross‑switch traffic conflicts during the process of N computing units executing a collective communication operation can be reduced, and the efficiency of collective communication can be improved.
Owner:HUAWEI TECH CO LTD

Time synchronized collective communication

ActiveUS12562994B2TransmissionData packCollective communication
Systems, methods, and devices that perform computing operations are provided. In one example, a system includes a least one node, the at least one node having one or more processors, each having associated memory, a clock, a scheduler, the scheduler monitoring one or more of rates, rates of lanes, rates at which packets are sent, times, latencies of packets, topology, communication states, nodes, and packets in the system, an attribute monitor that measures counters for one or more of congestion state, line rate, and communication attributes. A packet scheduler determines a destination node based on information from the scheduler and the attribute monitor, and sends at least a portion of a packet to the destination node.
Owner:MELLANOX TECHNOLOGIES LTD(IL)

RDMA (Remote Direct Memory Access) network performance abnormity diagnosis method under set communication

PendingCN121864639ASupport collective communication level diagnosticsTransmissionData packCollective communication
The invention provides an RDMA network performance abnormity diagnosis method under set communication. In a collective communication process, firstly, a monitor located at a server side executes algorithm disassembly in advance, monitors related performance information of a collective communication flow in real time according to steps, and reports the related performance information to an analysis server; the monitors inform the servers of waiting relations by sending Notification packets among the monitors, adaptively execute an appropriate anomaly detection triggering process when performance anomaly occurs, send a polling query data packet to the switches through the servers, trigger telemetry information collection on all the switches related to the anomaly, and report the telemetry information collection to an analysis server in sequence; and finally, the analysis server constructs a waiting relation graph and a network traceability graph, and comprehensively analyzes the waiting relation graph and the network traceability graph to give a diagnosis result. The problems that in a set communication scene, existing analysis modes such as a network traceability graph cannot give the overall performance bottleneck of set communication and cannot give contribution evaluation of flow to overall set communication performance abnormity and the like are solved.
Owner:BEIHANG UNIV

Optimizing tree-based collective communication operations by load-balancing network endpoints

PendingUS20260113274A1TransmissionBalancing networkCollective communication
The present disclosure generally relates to optimizing load-balancing of network endpoints using tree collectives representing a logical network communication topology for the network endpoints. Systems and methods described herein eliminate the previously restrictive conditions imposed on tree-based communication collectives by generating collective trees with any arity and representing any number of physical network endpoints. The resulting collective trees ensure that each represented network endpoint has a number of outgoing flows and a number of incoming flows that are no more than the arity of the collective tree. In this way, the described systems and methods inject significant efficiencies into communication collectives within networked compute nodes by eliminating communication bandwidth latencies and bottlenecks.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

NCCL set communication optimization method oriented to large model dynamic training environment

PendingCN121530849ATransmissionCollective communicationSimulation
An NCCL set communication optimization method oriented to a large model dynamic training environment comprises the following steps of (1) integrating open-source NCCL codes into PyTorch so as to be applied to actual distributed training; (2) a multi-node consistency guarantee mechanism is used, and system crash caused by strategy asynchronization is avoided; (3) respectively realizing a self-adaptive Channel allocation strategy on three levels of task scheduling, CUDAkernel and Proxy of the NCCL communication library; and (4) collecting, storing and visualizing multi-dimensional indexes. The method is compatible with mainstream deep learning ecology, and the communication efficiency and stability in a large model dynamic training environment are remarkably improved on the premise that hardware is not modified.
Owner:BEIJING UNIV OF POSTS & TELECOMM

Optimizing tree-based collective communication operations by load-balancing network endpoints

PCT designated stageWO2026089789A1TransmissionBalancing networkCollective communication
The present disclosure generally relates to optimizing load-balancing of network endpoints using tree collectives representing a logical network communication topology for the network endpoints. Systems and methods described herein eliminate the previously restrictive conditions imposed on tree-based communication collectives by generating collective trees with any arity and representing any number of physical network endpoints. The resulting collective trees ensure that each represented network endpoint has a number of outgoing flows and a number of incoming flows that are no more than the arity of the collective tree. In this way, the described systems and methods inject significant efficiencies into communication collectives within networked compute nodes by eliminating communication bandwidth latencies and bottlenecks.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Adaptive quantization and compression for variable-length collective communication

PendingUS20260180596A1Code conversionComputer hardwareCollective communication
A processing system dynamically and selectively quantizes or compresses variable-length input data for a collective operation to fit within a predetermined limit, executes the collective operation on the compressed data, and converts the data back to variable length results by dequantizing or decompressing the results of the collective operation.
Owner:ADVANCED MICRO DEVICES INC +1

A GPU collective communication method, system, device, medium and product

This invention discloses a GPU ensemble communication method, system, device, medium, and product. The method first maps ensemble communication interfaces from different GPU device manufacturers to a unified standard API interface based on GPU device manufacturer information, generating an interface mapping relationship. When responding to a ensemble communication request initiated by a user calling the API interface, the method translates the ensemble communication request into a corresponding manufacturer's ensemble communication interface request based on the GPU type used by the user and the interface mapping relationship, thereby invoking the corresponding manufacturer's ensemble communication interface to perform ensemble communication operations. Using this invention can reduce the development and maintenance complexity of multi-GPU ensemble communication systems and improve the overall performance and efficiency of distributed training of AI models.
Owner:CHINA MOBILE (SUZHOU) SOFTWARE TECH CO LTD +1

Collective communication method, collective communication system and related apparatus

PCT designated stageWO2026103582A1Fault recovery arrangementsRedundant operation error correctionCollective communicationCommunications system
Disclosed in the embodiments of the present application is a collective communication method. The method is applied to a collective communication system, wherein the collective communication system comprises a management node and N computing nodes, the N computing nodes being used for concurrently processing a first communication task, and N being an integer greater than 1. The method comprises: a management node receiving first information, wherein the first information indicates that at least one first link has failed, and each first link is a communication link between N computing nodes that is used for processing the first communication task; the management node receiving N pieces of second information from the N computing nodes, wherein each piece of second information is used for indicating whether original data of the first communication task that is locally stored in a computing node has been changed; and if the original data in the N computing nodes has not been changed, the management node sending third information to the N computing nodes, so as to instruct the N computing nodes to re-process the first communication task on the basis of the original data. In this way, the time required for task recovery after a communication fault occurs can be reduced.
Owner:HUAWEI TECH CO LTD

Parallel integrated collective communication and matrix multiplication operations

PendingCN122341962AComputational scienceCollective communication
A technique for efficiently performing integrated matrix multiplication to compute an output matrix based on two input matrices is described. Data corresponding to a first input tile of the first input matrix is ​​retrieved and stored in a local buffer. For each output tile of the output matrix, a non-blocking remote call is initiated to retrieve data corresponding to the next tile of the first input matrix into the local buffer, and concurrently with the processing of this remote call, the output tile is iteratively computed using the data from the first input matrix stored in the local buffer and a one-dimensional sequence of input tiles from the second input matrix. The output matrix is ​​generated based on the iteratively computed output tiles.
Owner:ADVANCED MICRO DEVICES INC