Water chilling unit fault diagnosis method and system for unbalanced data
Through the method based on the perturbation reconstruction mechanism and the spatiotemporal automatic encoder, high-quality balanced fault data is generated and dynamic topological timing feature integration network is used for feature fusion, which solves the problem of data imbalance in chiller fault diagnosis and the inefficiency of traditional methods, and achieves high-precision and efficient fault diagnosis.
Patent Information
- Application Number
- CN202510495841.2
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2025-04-21
- Publication Date
- 2025-05-16
- Estimated Expiration
- 2045-04-21
AI Technical Summary
There is a data imbalance problem in the chiller fault diagnosis, which leads to poor identification accuracy of the model for small sample fault types, and the traditional data generation method is inefficient, and the timing relationship and spatial structure of the generated data are biased from the original data.
The method based on the perturbation reconstruction mechanism is used to generate balanced fault data, and the time series features and spatial structures are extracted and modeled using the spatiotemporal automatic encoder. The dynamic topological timing feature integration network is used to perform feature fusion to achieve parallel generation and efficient diagnosis of multiple types of fault samples.
It significantly improves data generation efficiency, improves the quality and authenticity of generated data, improves the accuracy and adaptability of fault diagnosis, and can achieve high-precision real-time diagnosis on edge devices.
Smart Images

Figure CN120011867A_ABST
Abstract
Description
Technical Field
[0001] The invention belongs to the technical field of chiller fault diagnosis, and in particular relates to a chiller fault diagnosis method and system oriented to unbalanced data. Background Art
[0002] The statements herein merely provide background information related to the present invention and do not necessarily constitute prior art.
[0003] Chillers can accurately meet the cooling needs of buildings by efficiently coordinating refrigeration cycles. Since chillers are usually composed of multiple key equipment such as compressors, condensers, evaporators, valves and sensors, and operate under multi-variable coupling and time-varying load conditions for a long time, equipment aging, operating parameter mismatch or environmental changes may cause abnormal conditions. Continuous operation under fault conditions will not only lead to a decrease in refrigeration efficiency and a surge in energy consumption, but may also cause insufficient cooling effect due to system control failure, thereby affecting indoor temperature control and equipment safety. Therefore, it is necessary to explore an efficient and accurate method for diagnosing chiller faults.
[0004] Chillers have a variety of failure modes, such as compressor failure, valve leakage, and heat exchanger efficiency degradation, which lead to significant differences in the number of fault data for different fault types. Therefore, in the fault diagnosis of chillers, the unequal number of fault types leads to data imbalance, resulting in poor recognition accuracy of the training model for small sample fault types. More specifically, data imbalance leads to the problem of the model learning prior information about the proportion of samples in the training set, which leads to a bias towards the majority class in actual predictions. Therefore, for the fault diagnosis method and system of chillers, it is necessary to consider the data imbalance problem and adopt a suitable data generation method to solve the data imbalance problem.
[0005] The inventors found that the data generated by the data generation method used in the existing chiller fault diagnosis method is not ideal in terms of data stability. The main problems are as follows: (1) Due to the large difference in the number of samples of different fault types, the data distribution is unbalanced, which may lead to deviations in the fault diagnosis results; and the traditional data generation method usually adopts a serial generation method, that is, generating data for a single fault type one by one, which has the problem of low generation efficiency. (2) In the fault data generation stage, since the traditional method fails to fully capture and integrate the temporal dynamics in the original fault data and the spatial topological characteristics between sensors, the generated fault data of a certain type has a large deviation from the original fault data in terms of temporal relationship and spatial structure. (3) The traditional fault diagnosis method only focuses on the temporal or static structure of the chiller operation data, and cannot collaboratively model the temporal characteristics and dynamic topological characteristics, thus lacking the ability to adapt to real-time changing working conditions, which often leads to poor performance of the fault diagnosis model in practical applications. Summary of the invention
[0006] The purpose of the present invention is to overcome the shortcomings of the above-mentioned prior art and to provide a chiller fault diagnosis method and system for unbalanced data, which realizes the parallel generation of multiple types of fault samples, ensures that each type of generated fault data has similar spatiotemporal characteristics to the original fault data, can adapt to the real-time changing working conditions of the chiller, and realize the collaborative modeling of timing characteristics and dynamic topological characteristics in the complex multi-sensor system of the chiller, so as to achieve high-precision real-time diagnosis even on edge devices.
[0007] In order to achieve the above object, the present invention is implemented through the following technical solutions: On the one hand, the technical solution of the present invention provides a chiller fault diagnosis method for unbalanced data, comprising: Based on the disturbance reconstruction mechanism, the original data is used to restore the balanced fault data with the characteristics of the original data; The spatiotemporal autoencoder is used to extract and model the temporal characteristics and spatial structure of the balance fault data, reconstruct the original data, and obtain low-dimensional features that maintain the temporal dynamic characteristics and spatial topological structure of the original data; The dynamic topological temporal feature integration network is used to extract temporal features and topological features respectively, and the dynamic gated fusion mechanism is used for adaptive fusion. The fused features are classified to obtain the fault diagnosis results.
[0008] In at least one embodiment, the disturbance reconstruction mechanism is specifically: Gradually add noise to the original data to make it conform to Gaussian distribution; Through the denoising function, the noise is gradually removed and the balanced fault data with the characteristics of the original data is reconstructed.
[0009] In at least one embodiment, a label guidance layer is incorporated into the denoising process to generate data corresponding to specific labels, thereby achieving parallel generation of multiple types of data.
[0010] In at least one embodiment, the spatiotemporal autoencoder is composed of a first convolutional layer, a recurrent neural network layer, a graph attention network layer, and a second convolutional layer; The first convolutional layer is used to preprocess the original input data and extract local features; the recurrent neural network layer is used to capture the temporal dependencies in the time series and extract dynamically changing information; the graph attention network layer is used to use the self-attention mechanism to model the complex topological relationship between sensor nodes and effectively aggregate multi-hop information; the second convolutional layer is used to integrate and compress multi-scale features to generate high-quality hidden states.
[0011] In at least one embodiment, the dynamic topology timing feature integration network includes a dynamic topology extraction branch and a timing feature extraction branch; the dynamic topology extraction branch is used to capture the correlation between different sensors in the chiller that change with operating conditions; the timing feature extraction branch is used to model the long-term time dependence of sensor signals and predict fault evolution trends.
[0012] In at least one embodiment, the training process of the temporal feature extraction branch is: Input the timing fault data enhanced by the diffusion model; then use 1D convolution with a step size of 2 to reduce the sequence length and extract local features; then, input the local features extracted by convolution into the dynamic sparse attention block, use sparse attention and parameter sharing to model the global timing dependency, and output the timing features.
[0013] In at least one embodiment, the training process of the dynamic topology extraction branch is: The input is a multi-dimensional signal collected by the sensor at a fixed frequency and split with a sliding window; each sensor channel is independently standardized to construct a dynamic graph, which is input into a Chebyshev graph convolution with m layers for convolution processing, and a temporal convolution layer is used to capture the multi-scale temporal dependency of the sensor signal; finally, the importance weight of a single sensor node is calculated through the node attention pooling layer, and the weighted aggregation is aggregated into a global feature, and the topological feature is output.
[0014] In at least one embodiment, the adaptive fusion process is: Copy the topological features in time steps until they are aligned with the timing features. , and then generate dynamic gating weights and output gated fusion features.
[0015] In at least one embodiment, the fused features are input into a hierarchical classifier for classification to obtain a final fault diagnosis result.
[0016] On the other hand, the technical solution of the present invention also provides a chiller fault diagnosis system for unbalanced data, comprising: The balanced fault data parallel generation module is configured to: based on a disturbance reconstruction mechanism, use the original data to restore the balanced fault data with the original data characteristics; The data enhancement module is configured to: extract and model the temporal features and spatial structure of the balance fault data using a spatiotemporal autoencoder, reconstruct the original data, and obtain low-dimensional features that maintain the temporal dynamic features and spatial topological structure of the original data; The feature extraction and fusion module is configured to: use a dynamic topological temporal feature integration network to extract temporal features and topological features respectively and use a dynamic gated fusion mechanism to perform adaptive fusion; The fault diagnosis module is configured to classify the fused features to obtain a fault diagnosis result.
[0017] The beneficial effects of the technical solution of the present invention are as follows: 1) The present invention realizes the parallel generation of multiple types of fault samples, significantly improves the efficiency of data generation and shortens the generation time; at the same time, by balancing the data distribution, the quality of generated data is improved, the problem of imbalanced chiller fault data is effectively alleviated, and a more accurate and comprehensive training data foundation is provided for subsequent fault diagnosis.
[0018] 2) The fault data generated by the present invention is highly consistent with the original data in terms of temporal relationship and spatial structure, which significantly reduces the generation deviation, improves the authenticity and reliability of the data, provides higher quality data support for the fault diagnosis system, and further improves the accuracy of fault detection and diagnosis.
[0019] 3) The present invention realizes efficient analysis and accurate diagnosis of multi-source heterogeneous sensor data through the collaboration of dynamic topology extraction branches and timing feature extraction branches. It has high precision, low latency and strong generalization ability, can adapt to the real-time changing working conditions of the chiller, and realizes the collaborative modeling of timing features and dynamic topology features in the complex multi-sensor system of the chiller. It can also achieve high-precision real-time diagnosis on edge devices. BRIEF DESCRIPTION OF THE DRAWINGS
[0020] The accompanying drawings in the specification, which constitute a part of the present invention, are used to provide a further understanding of the present invention. The exemplary embodiments of the present invention and their descriptions are used to explain the present invention and do not constitute improper limitations on the present invention.
[0021] Figure 1 It is a chiller fault diagnosis method for unbalanced data and an overall schematic diagram of the system of the present invention; Figure 2 It is a schematic diagram of a module for parallel generation of balanced fault data of the present invention; Figure 3 is a schematic diagram of a data enhancement module of the present invention; Figure 4 It is a flow chart of the feature extraction and fusion module and the fault diagnosis module of the present invention. DETAILED DESCRIPTION
[0022] It should be noted that the following detailed descriptions are illustrative and intended to provide further explanation of the present invention. Unless otherwise specified, all technical and scientific terms used in the present invention have the same meanings as those commonly understood by those skilled in the art to which the present invention belongs.
[0023] As introduced in the background technology, the purpose of the present invention is to overcome the shortcomings existing in the above-mentioned prior art and to provide a chiller fault diagnosis method and system for unbalanced data, which realizes the parallel generation of multiple types of fault samples, ensures that each type of generated fault data has similar temporal and spatial characteristics to the original fault data, can adapt to the real-time changing working conditions of the chiller, and realize the collaborative modeling of timing characteristics and dynamic topological characteristics in the complex multi-sensor system of the chiller, so as to achieve high-precision real-time diagnosis even on edge devices.
[0024] Example 1 In a typical embodiment of the present invention, Figure 1 As shown, this embodiment discloses a chiller fault diagnosis method for unbalanced data, including: S100. Based on the disturbance reconstruction mechanism, the original data is used to restore the balanced fault data having the characteristics of the original data; S200. Extract and model the time series characteristics and spatial structure of the balance fault data using a spatiotemporal autoencoder, reconstruct the original data, and obtain low-dimensional characteristics that maintain the temporal dynamic characteristics and spatial topological structure of the original data; S300. Using a dynamic topological timing feature integration network to extract timing features and topological features respectively and using a dynamic gated fusion mechanism for adaptive fusion; S400. Classify the fused features to obtain fault diagnosis results.
[0025] A chiller fault diagnosis method oriented to unbalanced data is described in detail below in conjunction with specific embodiments.
[0026] S100. Based on the disturbance reconstruction mechanism, the original data is used to restore the balanced fault data with the characteristics of the original data.
[0027] In the fault diagnosis of chillers, data imbalance may cause the diagnostic model to be biased towards the majority class, reducing the ability to identify minority class faults. To address this challenge, a perturbation reconstruction mechanism is introduced in the latent space of the autoencoder to generate balanced fault data. The perturbation reconstruction mechanism is based on the Markov chain principle and a step-by-step denoising function is designed to gradually transform the noisy data into useful data, thereby recovering and generating fault data with original characteristics.
[0028] Specifically, the perturbation reconstruction mechanism includes two main steps: first, gradually add noise to the original data to make it conform to the Gaussian distribution; then, through the designed denoising function, gradually remove the noise and reconstruct the balanced fault data. At the same time, in order to improve the efficiency and quality of data generation, a label guidance layer is incorporated into the denoising process so that the model can generate data corresponding to specific labels, thereby realizing the parallel generation of multiple types of data. Figure 2 shown.
[0029] S101. Gradually add noise to the original data to make it into noise that conforms to Gaussian distribution.
[0030] During the noise injection process, for the original data Given an initial distribution , Gaussian noise is gradually added to this distribution in several steps, the number of steps is denoted by T , this process will eventually cause the data to be completely submerged by noise. The positive noise injection process is defined as follows: (1); in and Respectively represent the time step and Noise samples at time ; is a hyperparameter for scaling Gaussian noise, which varies with t Increase and increase; I represents the identity matrix, N Represents a Gaussian distribution.
[0031] S102. The noise is gradually removed through the designed denoising function to reconstruct balanced fault data.
[0032] Noise Removal Process By optimizing To complete the denoising task, is the set of model parameters that parameterize all key functions. The reverse process is defined as: (2); in, and denote the mean and variance of the noise removal process respectively.
[0033] According to the properties of Markov chain, we can conclude that: (3); Based on the probability density of the standard Gaussian distribution and Bayes' theorem, the posterior distribution is obtained The mean and variance of : (4); in , ,and Indicates the prediction of U-net at step t Added noise. Noise is sampled from a standard Gaussian distribution. By introducing the random variable , which can increase the diversity of data generated by DDPM.
[0034] By using the reparameterization technique, the final simplified objective function can be derived: (5); By using the objective function Training U-net can achieve accurate noise prediction.
[0035] S103. Incorporating a label guidance layer into the denoising process enables the model to generate data corresponding to specific labels, thereby achieving parallel generation of multiple types of data.
[0036] When training the reverse process of DDPM, U-net receives not only the generated data but also the steps t The embedded code is taken as input. t The embedding code is defined as follows: (6); in d represents the length of the embedded code, i Indicates the length index. t The embedding code of is input into U-net together with the noisy data, which can capture the temporal dependency of the noise.
[0037] In order to balance the problem of decreased generation quality due to noise guidance, this embodiment adopts a hybrid method of conditional guidance and unconditional guidance when embedding fault labels in the denoising process to guide data generation, and imposes weight restrictions on conditional guidance. The formula for conditional guidance is as follows: (7); in W is the similarity hyperparameter, t represents the time step, c represents label encoding, It is in step t The original vibration signal with noise added.
[0038] By gradually introducing noise to make the original data close to Gaussian distribution, and then using a denoising strategy to gradually restore the data, and incorporating label guidance, the parallel generation of multiple types of fault samples is achieved, which significantly improves the efficiency of data generation and shortens the generation time. At the same time, by balancing the data distribution, the quality of generated data is improved, and the problem of imbalance in chiller fault data is effectively alleviated. This solves the problems of imbalance in chiller fault data and low efficiency of traditional serial generation methods in the prior art, and provides a more accurate and comprehensive training data basis for subsequent fault diagnosis.
[0039] S200. Extract and model the time series characteristics and spatial structure of the balance fault data using a spatiotemporal autoencoder, reconstruct the original data, and obtain low-dimensional characteristics that maintain the temporal dynamic characteristics and spatial topological structure of the original data; In view of the significant time dependence and topological correlation of chiller data, this embodiment designs a spatiotemporal autoencoder framework, introduces a recurrent neural network layer and a graph attention network layer into the autoencoder, and extracts the temporal characteristics and spatial characteristics of the fault data respectively, so as to ensure the spatiotemporal characteristics of the hidden spatial data, so that the generated fault data is similar to the original data in spatiotemporal characteristics, so as to ensure that the extracted hidden state contains sufficient pattern features to support the reconstruction of chiller fault data. The spatiotemporal autoencoder can autonomously learn and extract representative intrinsic pattern features, and use low-dimensional hidden states to achieve reconstruction of chiller fault data.
[0040] In the encoder part, the spatiotemporal autoencoder consists of the first convolutional layer, the recurrent neural network layer, the graph attention network layer, and the second convolutional layer. Among them, the first convolutional layer is used to preprocess the original input data and extract local features; the middle recurrent neural network layer captures the temporal dependencies in the time series, thereby extracting dynamically changing information; the graph attention network layer that follows uses the self-attention mechanism to model the complex topological relationships between sensor nodes and effectively aggregate multi-hop information; the last second convolutional layer is used to further integrate and compress multi-scale features to generate high-quality hidden states. This design fully considers the characteristics of the chiller data, which has significant temporal dependence and topological correlation. The convolutional layer is used to capture local features, the recurrent neural network layer is used to model the temporal dependencies of the time series, and the graph attention network layer is used to capture the complex topological relationships between nodes in the system. Figure 3 As shown, the decoder in this embodiment adopts a network structure symmetrical to that of the encoder to ensure effective reconstruction of the original data.
[0041] S201. Extract and model the time series features of the balance fault data.
[0042] In view of the temporal dependency of the data, in the encoder, the information of each time step is first embedded through a 1×1 convolution to obtain For the chiller fault data, this embodiment uses a recurrent neural network to gradually capture the time dependency and compress the time step. By flexibly setting the hidden state dimension and number of layers of each layer, the model can integrate the information of all historical time steps with as little memory usage as possible. For the time series x, the recursive calculation using the multi-layer recurrent neural network unit R is formally defined as: (8); in Indicates tThe hidden state of time steps, is the ReLU activation function, and are the weight matrices from input to hidden layer and from hidden layer to hidden layer, is the bias term, and the output is .
[0043] S202. Extract and model the spatial structure of balanced fault data.
[0044] In view of the topological relevance of the data, this embodiment adds a relevance learning module - graph attention network in both the encoder and the decoder to aggregate multi-hop information. Adaptive adjacency matrix A and each layer hidden state As input, dynamic weighted aggregation of information is achieved through the self-attention mechanism.
[0045] Specifically, for each node i For example, the next hidden state is defined as: (9); in, Representation Node i The neighbor set of For the l The weight matrix of the layer, is the ReLU activation function; attention coefficient To measure the node j For Node i The influence of is calculated as: (10); in, is a learnable attention vector, || represents the concatenation operation of the vector, and the output is .
[0046] The encoder compresses the original chiller fault data of length T into a hidden state sequence through multi-layer RNN and GAT. , usually, the last hidden state is used As a representation of the input sequence, it is then passed through another 1×1 convolution to output the hidden state of the intrinsic pattern h In this embodiment, the decoder adopts a network structure symmetrical to that of the encoder to ensure efficient reconstruction of the original data.
[0047] In order to ensure that the generated data is accurate and the potential space distribution is close to the Gaussian distribution, the mean square error (MSE) and KL divergence are used as loss functions in this embodiment: (11); (12); In the formula, represents the reconstruction loss value of the spatiotemporal autoencoder, Represents the original input data of the encoder, represents the reconstructed data generated by the decoder; represents the KL divergence loss value, represents the potential variable in the latent space i The mean of the dimension, Represents the latent variable i The standard deviation of the dimension.
[0048] By introducing a recurrent neural network layer and a graph attention network layer into the autoencoder, the temporal features and spatial structure of the data are deeply extracted and dynamically modeled, respectively, to ensure that the data in the hidden space can faithfully reflect the temporal and spatial characteristics of the original data. Through this design, the problem of deviation between the temporal relationship and spatial structure of the fault data generated in the prior art and the original data is solved. The generated fault data is highly consistent with the original data in terms of temporal relationship and spatial structure, which significantly reduces the generation deviation, improves the authenticity and reliability of the data, provides higher quality data support for the fault diagnosis system, and further improves the accuracy of fault detection and diagnosis.
[0049] S300. Using a dynamic topological timing feature integration network to extract timing features and topological features respectively and using a dynamic gated fusion mechanism for adaptive fusion; In order to adapt to the real-time changing working conditions of the chiller and realize the collaborative modeling of the timing characteristics and dynamic topological characteristics of the chiller, this embodiment proposes a dynamic topological timing feature integration network to perform real-time fault diagnosis of the chiller. The dynamic topological timing feature integration network is mainly divided into two branches: a dynamic topological extraction branch and a timing feature extraction branch. The dynamic topological extraction branch can handle the dynamic changes of the graph structure and is used to capture the correlation between different sensors in the chiller as the working conditions change. The timing feature extraction branch models the long-term time dependence of the sensor signal and predicts the fault evolution trend. The fusion of the two branches is essentially the synergistic enhancement of spatial topological modeling and timing pattern capture, forming a complementary advantage of "1+1>2" in industrial fault diagnosis. The specific model structure is as follows: Figure 4 shown.
[0050] The joint training of the dynamic topology extraction branch and the temporal feature extraction branch is carried out in stages. First, the two branch networks are trained independently until convergence, and then the backbone network is frozen, and only the fusion gating weights and classifiers are trained. Next, the training process of the temporal feature extraction branch and the dynamic topology extraction branch is introduced respectively.
[0051] S301. Training of temporal feature extraction branch.
[0052] like Figure 4 As shown on the left, the time series feature extraction branch first inputs the time series fault data enhanced by the spatiotemporal autoencoder ,in L is the time step, D is the feature dimension. Then a 1D convolution with a stride of 2 is used to reduce the sequence length and extract local features: (13); in, is the weight parameter of the convolution kernel, is a learnable positional encoding. Output features , the sequence length is given by L Compress to L / 2, the feature dimension is D Expand to d model The convolution kernel size of 1D convolution is 3, and the sequence length is halved by step size 2. t The output is: (14) ; Subsequently, the local features extracted by the sliding window operation of the 1D convolution kernel in the time dimension are input into the dynamic sparse attention block. Each dynamic sparse attention block consists of a grouped query attention and a deep separable feedforward network (DS-FFN).
[0053] Grouped query attention is an improved self-attention mechanism, which is mainly used to optimize the computational efficiency of the model when processing large-scale input. Its core idea is to reduce the computational complexity by grouping queries. First, according to the input features of each time step, , projection generates query (Q), key (K), value (V) matrices. A total of 4 attention heads are used, and the 4 query heads are divided into 2 groups, each group sharing the key matrix and value matrix: (15); Among them, the weight matrix , ; , , each group contains two query heads.
[0054] Then, through the sparse attention mask, only the attention of each query and its five preceding and following keys is allowed to be calculated: (16); The goal of sparse attention masking is to limit the scope of interaction between queries and keys by setting a mask. The mask is usually a binary matrix, where the value indicates whether a query is allowed to interact with a key. Sparse attention masking can reduce unnecessary calculations, thereby improving the efficiency of the model when processing large-scale data. Thus, the single-head attention output is obtained: (17); like , using the key value of group 1: .like , using the key value of group 2: .
[0055] Then, the outputs of all heads are concatenated to obtain the output of grouped query attention: (18); Among them, the parameter matrix This design significantly reduces the computational resource requirements while maintaining the expressive power of multi-head attention through group-shared key-value projection and local sparse attention.
[0056] Using a residual connection, the output of the previous layer is added to the residual of the input: (19); This connection allows the gradients to flow directly through the network during backpropagation, thus alleviating the vanishing gradient problem.
[0057] The added output is normalized and then input into the deep separable feedforward network layer, which mainly consists of two parts: deep convolution layer and point-by-point convolution layer. The deep convolution layer is responsible for performing convolution operations on each channel of the input feature in the spatial dimension to extract local features; the point-by-point convolution layer uses 1D convolution to perform linear transformation on the features in the channel dimension to integrate the information between channels. The whole process can be expressed as: (20); Using the residual connection again, the output of the previous layer is added to the residual of the input: (twenty one); In order to reduce the number of model parameters, a cross-layer parameter sharing mechanism is adopted, and odd layers and even layers share the key / value projection matrix and convolution kernel parameters respectively.
[0058] Assumption n After processing with the stacked dynamic sparse attention blocks, the final output features of the model are , which is used for subsequent fusion with the output of the node attention pooling layer of the dynamic topology extraction branch.
[0059] S302. Training of dynamic topology extraction branch.
[0060] like Figure 4 As shown on the right, in the training process of the dynamic topology extraction branch, the multi-dimensional signal of the sensor is first input , the signal comes from the sensor and is collected at a fixed frequency.
[0061] Then split it with sliding window: (twenty two); w is the window length, which indicates the number of consecutive time steps contained in each window; t is the current time step, and the end position of the window is ; D is the sensor feature dimension; From the time step t−w arrive t −1 data matrix.
[0062] The sliding window ensures that when new data arrives, the oldest data is removed and the latest data is added. Each sensor channel is then normalized independently: (twenty three); Then, a dynamic graph is constructed. Specifically, a graph structure is constructed, each sensor corresponds to a graph node, and the Pearson correlation coefficient is used to calculate the dynamic adjacency matrix: (twenty four); For in-window sensors i The mean of .
[0063] The graph is sparsely processed, and only the most relevant node is retained for each node K Edge (such as K =3). The dynamic graph is The adjacency matrix is updated once every step. Using a double buffering mechanism, the background thread calculates the new adjacency matrix, and the foreground inference uses the old matrix.
[0064] Normalize the adjacency matrix symmetrically to improve numerical stability: (25); in, is the degree matrix. Then, the dynamic graph is input to the common m In the Chebyshev graph convolution layer, the ReLU activation function is used, and the definition of graph convolution is as follows: (26); is the scaled normalized Laplacian matrix, is the original Laplacian matrix, is the Chebyshev polynomial, is the trainable weight matrix of the k-th order polynomial in the l-th layer.
[0065] After the dynamic graph convolution, a temporal convolution layer (1D hole convolution) is used to further capture the multi-scale temporal dependencies of the sensor signal. Finally, the importance weight of a single sensor node is calculated through the node attention pooling layer, and the weighted aggregation is a global feature that can reflect the global influence of the sensor node in the structure: (27); (28); is the importance weight of a single sensor node, is the node feature of a single sensor node, N is the number of sensor nodes.
[0066] S400. Classify the fused features to obtain fault diagnosis results.
[0067] After the dynamic topology extraction branch network and the temporal feature extraction branch network are independently trained to converge, the dynamic gating fusion mechanism is used to adaptively fuse the temporal features extracted by the two branches. With topological features .
[0068] First, the topological features are copied in time steps to align with the temporal features output by the temporal feature extraction branch ,in, Indicates the length of the sequence after processing; N represents the number of sensor nodes; Represents the dimension of the feature vector.
[0069] Then perform dynamic gating weight generation: (29); is the 1D convolution kernel weight, is the bias term, the activation function is the Sigmoid function, and the output gate value The dynamic gated fusion mechanism ensures that the fusion weights are learned independently at each time step and each feature dimension.
[0070] Gating fuses the features output by two models: (30); when When it approaches 1, the feature of this time step mainly depends on the temporal feature extraction branch; when When it approaches 0, the structural features of the branch are extracted by relying on dynamic topology.
[0071] The fused features are finally input into the hierarchical classifier for classification to obtain the final fault diagnosis results: (31); In general, this time series dynamic integration method achieves efficient analysis and accurate diagnosis of multi-source heterogeneous sensor data through the collaboration of dynamic topology extraction branches and time series feature extraction branches. It has high precision, low latency and strong generalization capabilities, and provides a reliable edge intelligence solution for predictive maintenance of industrial equipment.
[0072] By introducing a fault diagnosis method based on a dynamic topological timing feature integration network, a timing feature extraction branch network is used to model global timing dependencies using sparse attention and parameter sharing. At the same time, a dynamic topology extraction branch network is used to build a graph structure according to the real-time sensor correlation and perform Chebyshev graph convolution to capture the dynamic topological changes of the chiller sensors. Finally, the features of the two models are fused through a dynamic gated fusion mechanism to output the fault diagnosis results. High-precision real-time diagnosis can also be achieved on edge devices, which can adapt to the real-time changing working conditions of the chiller and realize the collaborative modeling of timing features and dynamic topological features in the complex multi-sensor system of the chiller, solving the problem that traditional fault diagnosis methods lack the ability to adapt to real-time changing working conditions.
[0073] Example 2 In a typical implementation of the present invention, this embodiment discloses a chiller fault diagnosis system for unbalanced data, including: The balanced fault data parallel generation module is configured to: based on a disturbance reconstruction mechanism, use the original data to restore the balanced fault data with the original data characteristics; The data enhancement module is configured to: extract and model the temporal features and spatial structure of the balance fault data using a spatiotemporal autoencoder, reconstruct the original data, and obtain low-dimensional features that maintain the temporal dynamic features and spatial topological structure of the original data; The feature extraction and fusion module is configured to: use a dynamic topological temporal feature integration network to extract temporal features and topological features respectively and use a dynamic gated fusion mechanism to perform adaptive fusion; The fault diagnosis module is configured to classify the fused features to obtain a fault diagnosis result.
[0074] The above description is only a preferred embodiment of the present invention and is not intended to limit the present invention. For those skilled in the art, the present invention may have various modifications and variations. Any modification, equivalent replacement, improvement, etc. made within the spirit and principle of the present invention shall be included in the protection scope of the present invention.
Claims
1. A chiller fault diagnosis method for unbalanced data, characterized in that: include: Based on the disturbance reconstruction mechanism, the original data is used to restore the balanced fault data with the characteristics of the original data; The spatiotemporal autoencoder is used to extract and model the temporal characteristics and spatial structure of the balance fault data, reconstruct the original data, and obtain low-dimensional features that maintain the temporal dynamic characteristics and spatial topological structure of the original data; The dynamic topological temporal feature integration network is used to extract temporal features and topological features respectively, and the dynamic gated fusion mechanism is used for adaptive fusion. The fused features are classified to obtain the fault diagnosis results.
2. A chiller fault diagnosis method for unbalanced data according to claim 1, characterized in that: The disturbance reconstruction mechanism is specifically: Gradually add noise to the original data to make it conform to Gaussian distribution; Through the denoising function, the noise is gradually removed and the balanced fault data with the characteristics of the original data is reconstructed.
3. A chiller fault diagnosis method for unbalanced data according to claim 2, characterized in that: A label guidance layer is incorporated into the denoising process to generate data corresponding to specific labels and realize the parallel generation of multiple types of data.
4. A chiller fault diagnosis method for unbalanced data according to claim 1, characterized in that: The spatiotemporal autoencoder consists of a first convolutional layer, a recurrent neural network layer, a graph attention network layer, and a second convolutional layer; The first convolutional layer is used to preprocess the original input data and extract local features; the recurrent neural network layer is used to capture the temporal dependencies in the time series and extract dynamically changing information; the graph attention network layer is used to use the self-attention mechanism to model the complex topological relationship between sensor nodes and effectively aggregate multi-hop information; the second convolutional layer is used to integrate and compress multi-scale features to generate high-quality hidden states.
5. The method for diagnosing chiller faults based on unbalanced data according to claim 1, characterized in that: The dynamic topology temporal feature integration network includes a dynamic topology extraction branch and a temporal feature extraction branch; The dynamic topology extraction branch is used to capture the correlation between different sensors in the chiller that change with the operating conditions; the time series feature extraction branch is used to model the long-term time dependence of sensor signals and predict the fault evolution trend.
6. A chiller fault diagnosis method for unbalanced data according to claim 5, characterized in that: The training process of the temporal feature extraction branch is as follows: Input the timing fault data enhanced by the diffusion model; then use 1D convolution with a step size of 2 to reduce the sequence length and extract local features; then, input the local features extracted by convolution into the dynamic sparse attention block, use sparse attention and parameter sharing to model the global timing dependency, and output the timing features.
7. A chiller fault diagnosis method for unbalanced data according to claim 5, characterized in that: The training process of the dynamic topology extraction branch is: The input is a multi-dimensional signal collected by the sensor at a fixed frequency and split with a sliding window; each sensor channel is independently standardized to construct a dynamic graph, which is input into a Chebyshev graph convolution with m layers for convolution processing, and a temporal convolution layer is used to capture the multi-scale temporal dependency of the sensor signal; finally, the importance weight of a single sensor node is calculated through the node attention pooling layer, and the weighted aggregation is aggregated into a global feature, and the topological feature is output.
8. The method for diagnosing chiller faults based on unbalanced data according to claim 1, characterized in that: The adaptive fusion process is as follows: Copy the topological features in time steps until they are aligned with the timing features. , and then generate dynamic gating weights and output gated fusion features.
9. The method for diagnosing chiller faults based on unbalanced data according to claim 1, characterized in that: The fused features are input into the hierarchical classifier for classification to obtain the final fault diagnosis results.
10. A chiller fault diagnosis system for unbalanced data, characterized in that: include: The balanced fault data parallel generation module is configured to: based on a disturbance reconstruction mechanism, use the original data to restore the balanced fault data with the original data characteristics; The data enhancement module is configured to: extract and model the temporal features and spatial structure of the balance fault data using a spatiotemporal autoencoder, reconstruct the original data, and obtain low-dimensional features that maintain the temporal dynamic features and spatial topological structure of the original data; The feature extraction and fusion module is configured to: use a dynamic topological temporal feature integration network to extract temporal features and topological features respectively and use a dynamic gated fusion mechanism to perform adaptive fusion; The fault diagnosis module is configured to classify the fused features to obtain a fault diagnosis result.
Citation Information
Patent Citations
Fault quantitative diagnosis method based on lightweight attention deep reinforcement learning
CN116541793A
Fault diagnosis method based on enhanced multi-mode switching dynamic system model
CN119127539A
Improved GAN network and multi-modal feature fusion-based boat hanging frame sample imbalance fault diagnosis method and system
CN119152329A
Fault diagnosis method and device for heat exchange module
CN119378336A
Bearing fault noise immunity diagnosis method based on IPDSCS and Swin Transformer
CN119782890A
Cited By
Intelligent fault self-diagnosis method for evaporative condensation integrated screw water chilling unit
CN120317157A
Fault diagnosis method for carbon dioxide transcritical skid refrigerating unit
CN122153675A