Method and device for evaluating performance of heat treatment process of safety hook based on neural network

The neural network-based evaluation of safety hook hot processing addresses the low quality of human assessment by automating the process with data compression and self-encoding, ensuring accurate and efficient evaluation of safety hook performance.

CN120067603BActive Publication Date: 2025-07-15SHANDONG SHENLI RIGGING
View PDF 1 Cites 0 Cited by

Patent Information

Application Number
CN202510550447.4
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2025-04-29
Publication Date
2025-07-15
Estimated Expiration
2045-04-29

AI Technical Summary

Technical Problem

The heat treatment process evaluation method of safety hooks in the prior art relies on manual evaluation, resulting in low evaluation quality and it is difficult to accurately evaluate the complex nonlinear relationship between multiple parameters.

Method used

The performance evaluation method of safety hook heat treatment process based on neural network is adopted. By collecting a variety of process data to be evaluated, the pre-trained data compression model and the heat treatment process performance evaluation model are automatically evaluated, including data compression and feature extraction, data compression is used by the autoencoder, and optimization parameter updates are combined with inertial calibration and gradient conflict detection.

Benefits of technology

Automatic evaluation of the safety hook heat treatment process is realized, the evaluation accuracy is improved, the lack of manual evaluation is avoided, and the evaluation quality is ensured.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120067603B_ABST
    Figure CN120067603B_ABST
Patent Text Reader

Abstract

This application relates to the technical field of process monitoring, and in particular to a method and device for evaluating the heat treatment process performance of a safety hook based on a neural network. The method includes: during the process of treating the safety hook with the heat treatment process to be evaluated, collecting a variety of process data to be evaluated for the heat treatment process to be evaluated; inputting the variety of process data to be evaluated into a pre-trained data compression model to obtain the compressed data to be evaluated corresponding to the variety of process data to be evaluated; inputting the compressed data to be evaluated into a pre-trained heat treatment process performance evaluation model to obtain the evaluation result of the heat treatment process to be evaluated. This application can solve the problem of low evaluation quality in the existing manual evaluation method.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present application relates to the field of artificial intelligence technology, and particularly to a method and device for evaluating the performance of a heat treatment process of a safety hook based on a neural network. Background Art

[0002] In production and life, a safety hook is a force-bearing tool or force-bearing component widely used in infrastructure industries such as construction, mainly used to achieve the movement, tying, and stabilizing of spatial structures of objects. Safety hooks are widely used in industries and fields such as aviation, aerospace, ports, docks, logistics, mine traction and hoisting, and offshore platforms. In the manufacturing process of safety hooks, the heat treatment process is a very important key link, and its process quality directly affects the mechanical properties, service life, and safety of safety hooks. Therefore, it is necessary to evaluate the heat treatment process of safety hooks to determine whether the obtained safety hooks meet the usage requirements.

[0003] In the prior art, the heat treatment process of safety hooks is usually evaluated by manual evaluation. However, since the heat treatment process of safety hooks involves multiple parameters, such as temperature, pressure, atmosphere composition, applied force, and material hardness, and in addition, due to the complex non-linear relationship between the above parameters, the manual evaluation method has the problem of low evaluation quality. Summary of the Invention

[0004] In view of this, the purpose of the present application is to provide a method and device for evaluating the performance of a heat treatment process of a safety hook based on a neural network to solve the problem of low evaluation quality existing in the manual evaluation method in the prior art.

[0005] In the first aspect, the present application provides a method for evaluating the performance of a heat treatment process of a safety hook based on a neural network, and the method includes:

[0006] During the process of treating a safety hook with a heat treatment process to be evaluated, collect a variety of process data to be evaluated of the heat treatment process to be evaluated;

[0007] Input the variety of process data to be evaluated into a pre-trained data compression model to obtain compressed data to be evaluated corresponding to the variety of process data to be evaluated;

[0008] Input the compressed data to be evaluated into a pre-trained heat treatment process performance evaluation model to obtain an evaluation result of the heat treatment process to be evaluated.

[0009] In the second aspect, the present application provides a device for evaluating the performance of a heat treatment process of a safety hook based on a neural network, and the device includes: a collection module, a compression module, and an evaluation module;

[0010] The acquisition module is used to acquire a variety of to-be-evaluated process data of the to-be-evaluated heat treatment process during the process of treating the safety hook with the to-be-evaluated heat treatment process;

[0011] The compression module is used to input a variety of the to-be-evaluated process data into a pre-trained data compression model to obtain to-be-evaluated compressed data corresponding to the variety of the to-be-evaluated process data;

[0012] The evaluation module is used to input the to-be-evaluated compressed data into a pre-trained heat treatment process performance evaluation model to obtain an evaluation result of the to-be-evaluated heat treatment process.

[0013] Beneficial effects of the neural network:

[0014] The present application provides a method for evaluating the performance of a safety hook heat treatment process based on a neural network. The method includes: during the process of treating the safety hook with the to-be-evaluated heat treatment process, acquiring a variety of to-be-evaluated process data of the to-be-evaluated heat treatment process; inputting the variety of to-be-evaluated process data into a pre-trained data compression model to obtain to-be-evaluated compressed data corresponding to the variety of to-be-evaluated process data; inputting the to-be-evaluated compressed data into a pre-trained heat treatment process performance evaluation model to obtain an evaluation result of the to-be-evaluated heat treatment process. In summary, it can be seen that since the method for evaluating the performance of a safety hook heat treatment process based on a neural network provided by the present application can realize the automatic evaluation of the heat treatment process of the safety hook and avoid manual evaluation, and in addition, the to-be-evaluated process data is compressed during the automatic evaluation process, ensuring the accuracy of the evaluation. Therefore, it can solve the problem that the evaluation quality of the existing manual evaluation method is relatively low. BRIEF DESCRIPTION OF THE DRAWINGS

[0015] In order to more clearly illustrate the technical solutions of the embodiments of the present application, the following will briefly introduce the drawings required to be used in the embodiments of the present application. The following drawings only show some embodiments of the present application, so they should not be regarded as limiting the scope. For those of ordinary skill in the art, other related drawings can be obtained according to these drawings without creative efforts.

[0016] Figure 1 It is a schematic flow chart of the method for evaluating the performance of a safety hook heat treatment process based on a neural network provided by the embodiment of the present application;

[0017] Figure 2 It is an example diagram of the comparison of training losses of different training methods provided by the embodiment of the present application;

[0018] Figure 3 It is an example diagram of the influence of the autoencoder on the reconstruction accuracy provided by the embodiment of the present application;

[0019] Figure 4Example diagram of the impact of gradient conflict detection on parameter update provided by the embodiments of the present application;

[0020] Figure 5 Example comparison diagram of the energy flow provided by the embodiments of the present application for the feature retention ability in a noisy environment;

[0021] Figure 6 Schematic structural diagram of a device for evaluating the performance of a safety hook heat treatment process based on a neural network provided by the embodiments of the present application. Detailed implementation manners

[0022] To make the objectives, technical solutions, and advantages of the embodiments of the present application clearer, the technical solutions of the present application will be clearly and completely described below with reference to the accompanying drawings. Apparently, the described embodiments are some but not all of the embodiments of the present application. All other embodiments obtained by those of ordinary skill in the art based on the embodiments in the present application without creative efforts shall fall within the scope of protection of the present application.

[0023] First, the present application provides a method for evaluating the performance of a safety hook heat treatment process based on a neural network, as Figure 1 shown, Figure 1 Schematic flowchart of the method for evaluating the performance of a safety hook heat treatment process based on a neural network provided by the embodiments of the present application. The method includes: S110~S130, details are as follows:

[0024] S110: During the process of treating the safety hook with the heat treatment process to be evaluated, collect various process data to be evaluated of the heat treatment process to be evaluated.

[0025] Specifically, compared with the evaluation method by manual evaluation in the prior art, the technical solution of the present application can automatically evaluate the heat treatment process to be evaluated through the safety hook heat treatment process performance evaluation device, avoiding the problem of low evaluation quality caused by relying on manual evaluation.

[0026] In actual operation, the sources of the process data to be evaluated are mainly of two types; among them, the first is to directly collect at devices such as the temperature control system and mechanical testing equipment of the heat treatment furnace required for the heat treatment process. The collected data may include control system logs, mechanical operation records, and relevant experimental data, etc.; the second is obtained by collecting through multiple sensors arranged at the site of the heat treatment process to be evaluated. The sensors at least include, for example, temperature sensors, pressure sensors, and stress-strain sensors, etc.

[0027] In actual operation, the types of process data to be evaluated include: furnace temperature (unit: °C, representing the temperature inside the furnace during heat treatment), atmosphere temperature (unit: °C, representing the atmosphere temperature inside the furnace), applied force (unit: N, representing the force applied during heat treatment), heat treatment time (unit: seconds, representing the duration of a certain heat treatment cycle), material type identifier (representing the type of material being processed, such as steel, aluminum, etc.), material hardness (unit: HRC, representing the hardness of the material after heat treatment), strain rate of the material (unit: %, representing the deformation of the material during heat treatment), concentration of the furnace atmosphere (unit: ppm, representing the concentration of chemical components in the furnace atmosphere, such as oxygen, nitrogen, etc.), temperature gradient (unit: °C / min, representing the rate of temperature change inside the furnace), etc.; among them, it should be emphasized that the above data to be evaluated should all carry timestamps for integrating the data corresponding to the same timestamp into an input data.

[0028] During the process of collecting the process data to be evaluated, the real-time data of each sensor can be aggregated to the background server through serial or parallel transmission protocols (such as Modbus, CAN-bus) for enabling the background server to execute S120~S130; among them, the reading accuracy, data volume, and sampling frequency of the sensors are all dynamically adjusted by the background server to meet the requirements under different operating conditions.

[0029] S120: Input various process data to be evaluated into a pre-trained data compression model to obtain the compressed data to be evaluated corresponding to various process data to be evaluated.

[0030] Specifically, in actual application, the process data to be evaluated includes hundreds of data that can be collected by sensors, such as furnace temperature, applied force, and material hardness. If all the above hundreds of data are used as input data and input into the heat treatment process performance evaluation model, it will result in an extremely high feature dimension of the input data input into the heat treatment process performance evaluation model. High-dimensional data is prone to cause the curse of dimensionality, increase the calculation difficulty of the model, and further reduce the calculation efficiency of the model. In addition, the sparse data distribution may weaken the generalization ability of the model; thus, it can be seen that if all the above hundreds of data are used as input data and input into the heat treatment process performance evaluation model for evaluating the heat treatment process to be evaluated, it may lead to the heat treatment process performance evaluation model being unable to make accurate evaluation results.

[0031] To solve this problem, this application sets up a data compression model for compressing the process data to be evaluated collected by multiple sensors, and inputs the compressed data to be evaluated corresponding to various process data to be evaluated output by the data compression model into the heat treatment process performance evaluation model to obtain accurate evaluation results.

[0032] In one implementation, the network structure of the data compression model is an autoencoder, and the autoencoder includes an encoder and a decoder; during the iterative training of the autoencoder, the current training process before reaching the training stop condition includes steps (1) to (3), details are as follows:

[0033] Step (1): According to multiple training samples, determine the current sample input data corresponding to each training sample.

[0034] Among them, the training samples include various sample process data, and the sample process data is collected during the sample heat treatment process of the safety hook.

[0035] Specifically, in the embodiments of the present application, the task of the encoder in the trained autoencoder is to compress the process data to be evaluated from a high-dimensional space to a low-dimensional latent space, and the decoder attempts to reconstruct the process data to be evaluated from the low-dimensional representation.

[0036] Before using the autoencoder to determine the compressed data to be evaluated corresponding to multiple process data to be evaluated, iterative training needs to be performed on the autoencoder; in actual operation, in order to adapt to the complexity of the process data to be evaluated, for the training of the autoencoder, on the basis of using the gradient descent method for parameters, an inertial calibration method is used to optimize the gradient update step size of each round of training process. The inertial calibration method ensures smoother and more stable parameter updates by incorporating the influence of historical gradients into the parameter update process, avoiding the oscillation and convergence problems in traditional gradient descent.

[0037] It should be emphasized that during the iterative training of the autoencoder, the sample input data used in each training process should be the same. However, in order to distinguish different training rounds, the embodiments of the present application distinguish the sample input data corresponding to different training rounds. For example, "current sample input data" is used to represent the sample input data input to the autoencoder in the current training process, and "historical sample input data" is used to represent the sample input data input to the autoencoder in the historical training process that is temporally before the current training process.

[0038] Step (2): Input the current sample input data into the encoder to obtain the current low-dimensional latent space representation feature corresponding to the current sample input data.

[0039] Specifically, in the embodiments of the present application, the parameters of both the encoder and the decoder include weight matrices; among them, the mapping method of the encoder is as follows:

[0040] ;

[0041] In the formula, represents the The current low-dimensional latent space representation features output by the encoder during the current training process, including -dimensional features; denotes the encoder function, representing the process of mapping the current sample input data to the low-dimensional latent space; denotes the th current sample input data input to the encoder during the training process; denotes the th current weight matrix used by the encoder during the training process, which is a training parameter; denotes the th bias used by the encoder during the training process;

[0042] Step (3): Input the current low-dimensional latent space representation features into the decoder to obtain the current sample compressed data corresponding to the low-dimensional latent space representation features.

[0043] Specifically, the mapping method of the decoder is as follows:

[0044] ;

[0045] In the formula, denotes the th current sample compressed data output by the decoder during the training process; denotes the decoder function, representing the process of recovering from the current low-dimensional latent space representation features to the current sample compressed data; denotes the th weight matrix used by the decoder during the training process, which is the transpose of; denotes the th bias used by the decoder during the training process, which is the transpose of.

[0046] In actual operation, during the current training process before reaching the training stop condition of the autoencoder, after obtaining the current sample compressed data, it is also necessary to determine metrics such as the prediction accuracy of the autoencoder during the current training process based on the current sample input data and / or the current sample compressed data. In addition, it is also necessary to determine the parameters required by the autoencoder during the th training process based on the current sample input data and the current sample compressed data.

[0047] It should be emphasized that the training stop condition of the autoencoder can be determined according to actual needs, and the present application does not make specific limitations in this regard; in actual operation, the training stop condition can be to reach a preset maximum number of iterations. Preferably, the preset maximum number of iterations is set to 500 times, that is, when the training of the autoencoder reaches 500 times, the sample compression model can be obtained by stopping the training.

[0048] In one implementation manner, before step (1), S110 further includes: steps (4) to (8), the details are as follows:

[0049] Step (4): Determine the historical loss function value of the autoencoder during the historical training process according to the historical latent space representation features output by the encoder during the historical training process.

[0050] In one implementation manner, step (4) includes: steps (4.1) to (4.2), the details are as follows:

[0051] Step (4.1): Determine the feature weight weighting coefficient, the energy flow loss function value, and the adaptive feature interaction optimization loss function value according to the historical latent space representation features;

[0052] Among them, the feature weight weighting coefficient indicates the influence of the historical latent space representation features on the target increment; the energy flow loss function value indicates the influence of the energy flow of the historical latent space representation features on the target increment, and the energy flow indicates the amount of information of the historical latent space representation features; the adaptive feature interaction optimization loss function value indicates the influence of the features of each dimension in the historical latent space representation features on the target increment.

[0053] In one implementation manner, step (4.1) includes: steps (4.1.1) to (4.1.2), the details are as follows:

[0054] Step (4.1.1): Determine the feature contribution degree and information gain of the historical latent space representation features according to the historical latent space representation features;

[0055] Among them, the feature contribution degree indicates the contribution of the historical latent space representation features to the historical sample compression data output by the decoder during the historical training process; the information gain indicates the information amount difference of the historical latent space representation features.

[0056] Specifically, the feature contribution degree of the latent space representation features It is calculated by considering the matching degree between the dimensions of the process data to be evaluated and the overall distribution of the process data to be evaluated, in combination with the variance of the feature and the entropy value of the feature. For example, the variance of the data of a certain stress-strain sensor is high but the entropy value is low (the distribution is concentrated), indicating that it has a significant and stable process impact and a high contribution degree, while the experimental parameters with high entropy (such as discrete debugging records) may be suppressed. Therefore, the feature contribution degree of the latent space representation features can achieve a balance between the compression of redundant features (such as repeated log entries) and the expansion of key features (such as mutated temperature data), improving the dimensionality reduction efficiency.

[0057] In one implementation, step (4.1.1) includes: steps (4.1.1.1) to (4.1.1.3), details are as follows:

[0058] Step (4.1.1.1): Determine the variance and entropy of the low-dimensional latent space representation features according to the historical latent space representation features; where the entropy indicates the degree of discreteness of the low-dimensional latent space representation features.

[0059] Specifically, the entropy of the space representation ensures that features with a large degree of variation but high uncertainty can be appropriately compressed in the dimension expansion; where the entropy The calculation formula is as follows:

[0060] ;

[0061] In the formula, Represents the probability distribution of the th candidate value of the low-dimensional latent space representation feature determined according to the th training process; Represents the logarithmic function with base 10.

[0062] Step (4.1.1.2): Determine the feature contribution degree according to the preset second adjustment coefficient, historical latent space representation features, variance and entropy;

[0063] Among them, the second adjustment coefficient is used to adjust the relative relationship between the variance and the entropy.

[0064] Specifically, the feature contribution degree The calculation formula is as follows:

[0065] ;

[0066] In the formula, Represents the preset second adjustment coefficient; Represents the entropy of the low-dimensional latent space representation feature determined according to the th training process;

[0067] Step (4.1.1.3): Determine the information gain based on the variance.

[0068] Specifically, the information gain of the latent space representation features Automatically selects and enhances the representation of important features according to the feature distribution information of the low-dimensional latent space, while compressing redundant features. The information gain Is used to select features with the maximum amount of information.

[0069] For example, in the process data to be evaluated, for a multi-source sensor fusion scenario (such as simultaneously monitoring temperature, pressure, and strain), the data of multiple temperature sensors may exhibit different distribution characteristics. Some sensors are located at key process nodes (such as the material heating area), and their temperature fluctuation variances are relatively large, reflecting the core changes of the process. While other redundant sensors (such as ambient temperature monitoring) may have relatively small variances and repetitive information. Through the information gain Calculation, the autoencoder can quantify the information amount difference of each sensor feature, automatically expand the latent space dimension of high-gain features (such as key temperature nodes), and at the same time compress the redundant features of low gain (such as ambient temperature), avoiding information dilution caused by feature dimension expansion, and ensuring that the high-value sensor data can still dominate the latent space representation after dimensionality reduction.

[0070] Among them, the information gain The calculation formula is as follows:

[0071] ;

[0072] In the formula, Represents the feature of the th dimension in the low-dimensional latent space representation features output by the encoder during the th training process of the variance.

[0073] Step (4.1.2): Determine the feature weight weighting coefficient according to the preset first adjustment coefficient, historical latent space representation features, feature contribution degree, and information gain;

[0074] Among them, the first adjustment coefficient is used to adjust the compression and reconstruction of the historical sample input data by the encoder and decoder.

[0075] Specifically, different from the traditional autoencoder with a fixed latent dimension that cannot be dynamically adjusted according to the data distribution, the feature weight weighting coefficient Characterizes the dynamic dimension expansion and contraction constraint of the feature. By calculating the variance of the latent space representation features to judge the contribution degree of the feature, and adjust the latent space dimension accordingly.

[0076] For example, in the process data to be evaluated, if the characteristic variance of a certain temperature sensor is significant (reflecting the key temperature changes in the process), the dimension of its corresponding low-dimensional latent space will be expanded to enhance the representation ability, while redundant mechanical operation records (such as repetitive log entries) will be compressed.

[0077] Among them, the characteristic weight weighting coefficient is calculated as follows:

[0078] ;

[0079] In the formula, represents the exponential function; represents a preset first adjustment coefficient;

[0080] represents the variance of the low-dimensional latent space representation features output by the encoder during the th training process; represents the feature contribution degree of the low-dimensional latent space representation features determined according to the th training process; represents the information gain of the low-dimensional latent space representation features determined according to the th training process.

[0081] In one implementation, step (4.1) further includes: steps (4.1.3) to (4.1.4), details are as follows:

[0082] Step (4.1.3): Determine the energy flow of the low-dimensional latent space representation features and the deviation degree of the energy flow according to the historical latent space representation features.

[0083] Step (4.1.4): Determine the energy flow loss function value according to a preset third adjustment coefficient, the energy flow and the deviation degree; among them, the third adjustment coefficient is used to adjust the relative relationship between the energy flow and the entropy.

[0084] Specifically, in actual operation, the mechanical operation records and multi-sensor data in the process data to be evaluated may have redundancy (such as duplicate log entries) or noise (such as abnormal signals of equipment startup and shutdown). Redundant features waste computing resources, and noise interferes with the model's capture of key process parameters. That is, when the dimension of the process data to be evaluated is relatively high, conventional dimension expansion and compression methods are difficult to effectively balance information retention and feature selectivity. In addition, features with lower entropy have more concentrated information and stronger energy flow; while features with higher entropy have weaker energy flow due to their scattered information. For example, the energy flow weight of high-entropy pressure sensor data (distributed dispersedly) is reduced, while higher energy is given to low-entropy stable experimental parameters (such as constant stress values), so as to achieve selective retention of features.

[0085] To solve the above problems, the embodiments of the present application adopt a feature balance optimization method based on energy flow, which dynamically adjusts the importance of features during the feature compression process, so as to achieve efficient control of information volume and redundancy. The feature balance optimization method based on energy flow is inspired by the principle of energy conservation. The energy of each feature is quantified and dynamically allocated through a feature flow function, so that the feature information flow during the training process flows in a more reasonable way throughout the network. By optimizing the information flow between features, it is ensured that the energy flow intensity of important features is relatively large, while the energy flow of redundant features is effectively suppressed.

[0086] The energy flow of a feature not only considers the amplitude of the feature (measured by ), but also considers the entropy value of the feature to adjust the sparsity of the information volume. The energy flow loss function constructed according to the principle of dynamic energy balance in the embodiments of the present application dynamically adjusts the selection and compression process of features according to the energy flow of features. During the compression process, the selection of features not only depends on variance and entropy, but is also dynamically affected by the energy flow of features.

[0087] For example, high energy flow is given to the pressure sensor data with high-frequency changes to retain details, while the energy weight of the mechanical operation records with low information volume (such as periodic logs) is reduced. When processing multi-sensor fusion data (such as joint monitoring of temperature, pressure, and strain), it is possible to avoid a single feature dominating the latent space and ensure an equilibrium representation of multi-dimensional information.

[0088] Among them, the value of the energy flow loss function is calculated as follows:

[0089] ;

[0090] ;

[0091] ;

[0092] In the formula, Indicates the low-dimensional latent space representation features determined according to the th training process; the energy flow; Indicates a preset third adjustment coefficient. In actual operation, can be set to 0.2; Indicates the deviation degree of the energy flow determined according to the th training process; Indicates a preset reference energy flow. In actual operation, can be set to 10.

[0093] In one implementation, step (4.1) further includes: steps (4.1.5) to (4.1.7), details are as follows:

[0094] Step (4.1.5): Determine a feature interaction matrix according to the historical sample input data input to the encoder during the historical training process;

[0095] Among them, the feature interaction matrix indicates the interaction result between multiple features included in the historical sample input data.

[0096] Specifically, in actual operation, there are complex interactions among various feature data included in the process data to be evaluated (such as the non-linear coupling of temperature and stress affects material deformation), while traditional linear dimensionality reduction (such as principal component analysis) cannot model high-order interactions, resulting in the loss of key process relationships.

[0097] To solve this problem, the embodiments of the present application solve this problem through an adaptive feature interaction optimization method; the feature balance optimization method based on energy flow can adjust the intensity of features, but when the process data to be evaluated contains complex non-linear relationships, there is still a possibility that high-order interactions between some features are ignored. The embodiments of the present application adopt an adaptive feature interaction optimization mechanism to automatically identify and optimize high-order interaction relationships between features to enhance the model's ability to model non-linear and complex feature relationships.

[0098] Traditional compression methods such as principal component analysis and autoencoders usually ignore high-order interactions between features during compression; for example, the quadratic interaction of stress data and temperature data may reflect the material thermal expansion effect, and the cubic term can describe non-linear deformation characteristics. To better capture these interactions, the adaptive feature interaction optimization mechanism constructs a feature interaction matrix to describe second-order or higher-order interactions between features, enhances the ability to model complex process relationships (such as mechanical property changes under the cooperation of multiple sensors), makes up for the deficiencies of traditional linear dimensionality reduction methods, and assumes that there are features in the process data set to be evaluated input to the encoder.

[0099] Among them, the feature interaction matrix has the following calculation formula:

[0100] ;

[0101] In the formula, represents the sample input data input to the encoder during the th training process, and the th power of . The high-order power can be used to capture its high-order interaction effect.

[0102] Step (4.1.6): Perform normalization processing on the feature interaction matrix to obtain the normalized feature interaction matrix.

[0103] Specifically, in the embodiments of the present application, the feature interaction matrix is normalized to ensure that features of different orders can have balanced weights during the training process. The normalization method constructs a matrix based on the high-order power of the input data and performs L2 norm normalization to balance the weights of features of different orders and enhance the modeling ability for complex non-linear relationships.

[0104] Among them, the calculation formula of the normalized feature interaction matrix is as follows:

[0105] ;

[0106] In the formula, represents the L2 norm of the feature interaction matrix.

[0107] Step (4.1.7): Determine the value of the adaptive feature interaction optimization loss function according to the preset weight coefficient, the regularization coefficient of the preset adaptive feature interaction optimization loss function value , the low-dimensional latent space representation feature, and the historical weight matrix;

[0108] Among them, the weight coefficient adjusts the influence of the low-dimensional latent space representation feature on the autoencoder; the regularization coefficient is used to prevent overfitting.

[0109] Specifically, in the embodiments of the present application, in order to enable the model to adaptively learn the high-order interaction relationship between features, the adaptive feature interaction optimization loss function is adopted to automatically adjust the weight of its interaction relationship based on the contribution degree of each feature to the final prediction result; the role of the adaptive feature interaction optimization loss function is to constrain the autoencoder to not only minimize the prediction error, but also optimize the interaction relationship between features by adjusting the feature weights in the high-order interaction matrix, so as to effectively capture complex non-linear relationships.

[0110] Among them, the value of the adaptive feature interaction optimization loss function is calculated as follows:

[0111] ;

[0112] In the formula, represents the weight coefficient of the preset low-dimensional latent space representation feature used to control the influence of the latent space representation feature on the autoencoder. In actual operation, can be set to 0.2; represents the normalized feature interaction matrix determined according to the th training process; represents the label of the training sample; represents the preset value of the adaptive feature interaction optimization loss function regularization coefficient, used to prevent overfitting; represents the feature interaction matrix determined according to the th training process; represents the Softmax function.

[0113] Step (4.2): Determine the historical loss function value according to the historical sample input data input to the encoder during the historical training process, the preset sparsity weighting coefficient, the feature weight weighting coefficient, the energy flow loss function value, and the adaptive feature interaction optimization loss function value.

[0114] Specifically, in the feature reconstruction stage, through the dynamic dimension expansion and contraction method, the dimension of the latent space is gradually optimized. Different from the traditional autoencoder that simply reconstructs in the low-dimensional latent space, the dynamic dimension expansion and contraction method adopted in the embodiments of the present application can adaptively adjust the dimension of the low-dimensional latent space according to the characteristics of the process data to be evaluated. When the feature dimension is high or unevenly distributed, the decoder will dynamically expand the dimension of the low-dimensional latent space according to the contribution degree of each feature, while redundant or irrelevant features will be compressed.

[0115] In the embodiments of the present application, different from the traditional loss function that only considers the reconstruction error, the loss function of the autoencoder is calculated based on the norm of the reconstructed data, the energy flow loss, and the feature interaction loss, suppressing redundancy and capturing high-order interaction relationships while realizing the reconstruction of the data.

[0116] Among them, the value of the loss function of the autoencoder is calculated as follows:

[0117] ;

[0118] In the formula, represents the dimension of the features included in the sample input data; represents the feature weight weighting coefficient determined according to the th training process; represents a preset sparsity weighting coefficient. In actual operation, can be set to 0.3.

[0119] represents the L2 norm, which is the same as the calculation method of the Euclidean distance; represents the L1 norm.

[0120] represents the th data of the th dimension in the sample input data input to the encoder during the th training process; represents the th data of the th dimension in the sample compressed data output by the decoder during the th training process; represents the value of the energy flow loss function determined according to the th training process;

[0121] Step (5): Determine the historical gradient of the historical loss function value with respect to the historical weight matrix according to the historical loss function value and the historical weight matrix used by the encoder during the historical training process.

[0122] Step (6): Determine the conflict detection term of the target increment of the historical weight matrix according to the preset correction coefficient and the historical gradient;

[0123] wherein, the conflict detection term is used to adjust the update direction of the weight matrix.

[0124] Specifically, the embodiment of the present application adopts a gradient conflict detection method. The conflict detection term detects the gradient change in the low-dimensional latent space during each iteration, ensures that each gradient update is in the correct direction, especially corrects the gradient direction when encountering a gradient conflict, optimizes the learning rate selection, and avoids error accumulation leading to the model deviating from the optimal solution.

[0125] For example, when abnormal events in the control system log of the process data to be evaluated and normal stress data are input simultaneously, the gradient direction may conflict. The conflict detection term can adjust the update direction to avoid error accumulation and ensure the stable optimization of the model in complex scenarios (such as mechanical records of multi-device collaborative operations); wherein, the conflict detection term has the following formula:

[0126] ;

[0127] In the formula, represents the conflict detection term of the increment of the weight matrix determined according to the -th training process; represents a preset correction coefficient; represents the gradient of the weight matrix of the encoder determined according to the -th training process; represents the gradient of the weight matrix of the encoder determined according to the -th training process.

[0128] Step (7): Determine the target increment according to the preset inertia parameter, preset learning rate, historical gradient, and conflict detection term;

[0129] Among them, the inertia parameter is used to limit the influence of the update of the weight matrix in the historical training process on the target increment.

[0130] Step (8): Determine the target weight matrix used by the encoder and decoder in the current training process according to the historical weight matrix and the target increment.

[0131] Specifically, in actual operation, the traditional gradient descent method is prone to parameter update oscillation when there is noise data in the process data to be evaluated, and gradient conflict (such as when abnormal events coexist with normal data) causes the model to deviate from the optimal solution.

[0132] To solve this problem, the embodiments of the present application solve this problem by providing an inertia calibration mechanism and a gradient conflict detection method; in the embodiments of the present application, different from traditional optimization methods (such as Adam, RMSProp) that only rely on gradient magnitude or exponential averaging, the parameter update of the autoencoder (including the parameters of the encoder and decoder) adopts an inertia calibration mechanism, and the historical gradient information is used to correct each update process, so that the autoencoder can update the parameters more stably, avoiding the oscillation phenomenon in traditional gradient descent. For scenarios with a lot of noise such as mechanical operation records in the process data to be evaluated, for example, when there are irregular device start-stop signals in the control system log, the inertia calibration can smooth the parameter update path and avoid the model deviating from the optimal solution due to gradient mutation.

[0133] Among them, the target increment and the target weight matrix are calculated as follows:

[0134] ;

[0135] ;

[0136] In the formula, represents the increment of the weight matrix determined according to the th training process; represents a preset inertia coefficient; represents the learning rates of the preset encoder and decoder; represents the th loss function value of the autoencoder during the th training process for the gradient of the weight matrix; represents the weight matrix used by the encoder during the th training process.

[0137] It should be noted that in the embodiments of the present application, the update method of the bias of the autoencoder can refer to the update method of the weight matrix, which will not be elaborated here.

[0138] S130: Input the to-be-evaluated compressed data into the pre-trained heat treatment process performance evaluation model to obtain the evaluation result of the to-be-evaluated heat treatment process.

[0139] Specifically, in actual operation, the label should also be input into the heat treatment process performance evaluation model together with the training samples; in the embodiments of the present application, the content of the label may include different heat treatment qualification marks, indicating whether the heat treatment process meets the expected process standards; in actual operation, the process standards can be divided into 5 levels, and the higher the level, the higher the process standard reached.

[0140] In the embodiments of the present application, the network structure of the heat treatment process performance evaluation model can be a deep neural network, a support vector machine, a random forest, a gradient boosting tree, or a decision tree, etc., and the present application does not make specific limitations on this.

[0141] In actual operation, when the network structure of the heat treatment process performance evaluation model is a deep neural network, the deep neural network may include: an input layer, a hidden layer, and an output layer.

[0142] Among them, the input layer is used to receive the to-be-evaluated process data.

[0143] The hidden layer includes: a fully connected layer, a batch normalization layer (BatchNorm), and a Dropout layer; the number of neurons in the fully connected layer is 2 times the dimension of the low-dimensional latent space, the activation function of the fully connected layer is Leaky ReLU, and its negative slope coefficient is 0.01, which is used to alleviate the vanishing gradient; the batch normalization layer is used to accelerate convergence and suppress the deviation of the feature distribution; the dropout rate of the Dropout layer is 0.3, which is used to prevent overfitting and improve the generalization ability to noisy data.

[0144] The output layer also includes a fully connected layer, which includes 5 neurons corresponding to the 5 process standards mentioned above, and its activation function is Softmax, which is used to output the probability distribution.

[0145] In actual operation, when the network structure of the heat treatment process performance evaluation model is a deep neural network, the loss function of the deep neural network can use the cross-entropy loss function; during its training process, the compressed feature set is divided into a training set, a validation set, and a test set according to 7:2:1; among them, the validation set is used for early stopping and hyperparameter tuning; during its training process, the deep neural network adopts the AdamW optimizer, with an initial learning rate of 1e -3 , and a weight decay of 1e -4 , and dynamically adjusts the parameter update step size.

[0146] In actual operation, when the network structure of the heat treatment process performance evaluation model is a deep neural network, the deep neural network is used to output a 5D probability vector corresponding to each process data to be evaluated, representing the probability of belonging to each process standard. Further, the level corresponding to the maximum probability is taken as the final evaluation result; for example, if the 5D vector output by the deep neural network is [0.1, 0.05, 0.7, 0.1, 0.05], its final evaluation result should be the evaluation result corresponding to the probability value of 0.7.

[0147] In actual operation, in order to verify the superiority of the training method for the data compression model provided in the embodiments of the present application, that is, for the autoencoder, the present application now conducts the following comparative experiment on other existing training methods and the training method provided in the embodiments of the present application:

[0148] As Figure 2 shown, Figure 2 is an example diagram of the training loss comparison of different training methods provided in the embodiments of the present application. Figure 2 The abscissa of Figure 2 is the number of training iterations, and

[0149] The ordinate of

[0150] is the reconstruction loss. This comparative experiment aims to verify the convergence stability of the inertial calibration gradient update strategy compared with traditional optimization methods. The experiment compares the changes in the training losses of traditional gradient descent, momentum method, and the training method provided in the embodiments of the present application during 500 iterations. Figure 3 shown, Figure 3Example diagram of the impact of the autoencoder provided by the embodiments of the present application on the reconstruction accuracy Figure 3 The abscissa of Figure 3 is the dimension of the low-dimensional latent space, and the ordinate of

[0151] is the reconstruction error, which is represented by the mean squared error (MSE). This experiment aims to explore the impact of the latent space dimension on the reconstruction accuracy and verify the adaptive advantage of the dynamic dimension expansion and contraction method. The experimental results show that the reconstruction errors of the conventional autoencoder with a fixed dimension and the dynamic dimension method of the embodiments of the present application are compared under different latent dimensions. The conventional method has a high error due to excessive information compression at low dimensions (<50), and overfitting is caused by redundant features at high dimensions (>80). However, the embodiments of the present application maintain a stable low reconstruction error in the range of 10-100 dimensions by dynamically adjusting the dimension. It can be seen that the autoencoder provided by the embodiments of the present application can adaptively expand or compress the dimension according to the feature contribution degree, overcoming the defect that the dimension selection of the traditional method depends on prior knowledge and significantly improving the adaptability of the model to complex data.

[0152] As Figure 4 shown, Figure 4 is an example diagram of the impact of gradient conflict detection on parameter update provided by the embodiments of the present application. The abscissa of Figure 4 is the number of iterations, and the ordinate of Figure 4 is the gradient direction consistency, which is represented by the cosine similarity. This experiment verifies the optimization effect of the gradient conflict detection mechanism on parameter update by analyzing the gradient direction consistency. The experimental results show that the change in the cosine similarity of gradient update with and without conflict detection is compared. It is found that the gradient direction without conflict detection frequently deviates from the target (similarity <0.9), resulting in about 35% of invalid parameter updates. However, the embodiments of the present application dynamically correct the gradient direction through the conflict detection item, making the cosine similarity stable above 0.9 and increasing the number of effective parameter updates. It can be seen that the conflict detection mechanism reduces the random oscillation in the parameter update process by suppressing the gradient direction conflict, making the model more reliably approach the optimal solution.

[0153] As

[0154] shown, Figure 5 is an example diagram of the comparison of the energy flow on the feature retention ability in a noisy environment provided by the embodiments of the present application. The abscissa of Figure 5 is the noise level, and the ordinate of Figure 5 is the feature retention rate. This experiment evaluates the feature retention ability of the energy flow optimization method in a noisy environment and compares the feature retention rate and signal-to-noise ratio of the traditional principal component analysis method and the embodiments of the present application under different noise levels. Figure 5

[0155] The experimental results show that when the noise level exceeds 0.3, the principal component analysis method fails to distinguish noise from effective features, and the retention rate drops sharply to below 70%. In contrast, the embodiment of the present application suppresses redundancy by quantifying the feature energy through energy flow dynamics, and still maintains a feature retention rate of 82% under high noise (0.5), with the signal-to-noise ratio increased by 6 - 8 dB. It can be seen that the energy flow optimization mechanism provided by the embodiment of the present application can effectively identify and enhance the representation of important features, significantly improving the robustness of the model under noise interference.

[0156] Second, the present application provides a safety hook heat treatment process performance evaluation device based on a neural network, as Figure 6 shown Figure 6 is a structural schematic diagram of the safety hook heat treatment process performance evaluation device based on a neural network provided by an embodiment of the present application. The device includes: a collection module 210, a compression module 220, and an evaluation module 230;

[0157] The collection module 210 is configured to collect a variety of to-be-evaluated process data of the to-be-evaluated heat treatment process during the process of treating the safety hook with the to-be-evaluated heat treatment process;

[0158] The compression module 220 is configured to input the variety of to-be-evaluated process data into a pre-trained data compression model to obtain to-be-evaluated compressed data corresponding to the variety of to-be-evaluated process data;

[0159] The evaluation module 230 is configured to input the to-be-evaluated compressed data into a pre-trained heat treatment process performance evaluation model to obtain an evaluation result of the to-be-evaluated heat treatment process.

[0160] In one implementation, the network structure of the data compression model is an autoencoder, and the autoencoder includes an encoder and a decoder; the device further includes: a training module;

[0161] The training module is configured to perform iterative training on the autoencoder;

[0162] Among them, during the process of performing iterative training on the autoencoder, the current training process before reaching the training stop condition includes:

[0163] Determine the current sample input data corresponding to each training sample according to multiple training samples;

[0164] Among them, the training samples include a variety of sample process data, and the sample process data is collected during the sample heat treatment process for the safety hook;

[0165] Input the current sample input data into the encoder to obtain the current low-dimensional latent space representation feature corresponding to the current sample input data;

[0166] Input the current low-dimensional latent space representation feature into the decoder to obtain the current sample compression data corresponding to the low-dimensional latent space representation feature.

[0167] In one implementation, the parameters of both the encoder and the decoder include weight matrices; before inputting the sample input data into the encoder to obtain the low-dimensional latent space representation feature corresponding to the sample input data, the training module is further configured to determine the historical loss function value of the autoencoder during the historical training process according to the historical latent space representation features output by the encoder during the historical training process;

[0168] The training module is further configured to determine the historical gradient of the historical loss function value with respect to the historical weight matrix according to the historical loss function value and the historical weight matrix used by the encoder during the historical training process;

[0169] The training module is further configured to determine the conflict detection term of the target increment of the historical weight matrix according to a preset correction coefficient and the historical gradient;

[0170] Wherein, the conflict detection term is used to adjust the update direction of the weight matrix;

[0171] The training module is further configured to determine the target increment according to a preset inertia parameter, a preset learning rate, the historical gradient, and the conflict detection term;

[0172] Wherein, the inertia parameter is used to limit the influence of the update of the weight matrix during the historical training process on the target increment;

[0173] The training module is further configured to determine the target weight matrix used by the encoder and the decoder during the current training process according to the historical weight matrix and the target increment.

[0174] In one implementation, the training module is further configured to determine the feature weight weighting coefficient, the energy flow loss function value, and the adaptive feature interaction optimization loss function value according to the historical latent space representation features;

[0175] Wherein, the feature weight weighting coefficient indicates the influence of the historical latent space representation feature on the target increment; the energy flow loss function value indicates the influence of the energy flow of the historical latent space representation feature on the target increment, and the energy flow indicates the information amount of the historical latent space representation feature; the adaptive feature interaction optimization loss function value indicates the influence of each dimension feature in the historical latent space representation feature on the target increment;

[0176] The training module is further configured to determine the historical loss function value according to the historical sample input data input into the encoder during the historical training process, a preset sparsity weighting coefficient, the feature weight weighting coefficient, the energy flow loss function value, and the adaptive feature interaction optimization loss function value.

[0177] In one implementation, the training module is further configured to determine the feature contribution degree and information gain of the historical latent space representation features based on the historical latent space representation features;

[0178] Among them, the feature contribution degree indicates the contribution of the historical latent space representation features to the historical sample compressed data output by the decoder during the historical training process; the information gain indicates the information quantity difference of the historical latent space representation features;

[0179] The training module is further configured to determine the feature weight weighting coefficient according to a preset first adjustment coefficient, the historical latent space representation features, the feature contribution degree, and the information gain;

[0180] Among them, the first adjustment coefficient is used to adjust the compression and reconstruction of the historical sample input data by the encoder and the decoder.

[0181] In one implementation, the training module is further configured to determine the variance and entropy of the low-dimensional latent space representation features based on the historical latent space representation features; among them, the entropy indicates the degree of dispersion of the low-dimensional latent space representation features;

[0182] The training module is further configured to determine the feature contribution degree according to a preset second adjustment coefficient, the historical latent space representation features, the variance, and the entropy;

[0183] Among them, the second adjustment coefficient is used to adjust the relative relationship between the variance and the entropy;

[0184] The training module is further configured to determine the information gain according to the variance.

[0185] In one implementation, the training module is further configured to determine the energy flow and the deviation degree of the energy flow of the low-dimensional latent space representation features based on the historical latent space representation features;

[0186] The training module is further configured to determine the energy flow loss function value according to a preset third adjustment coefficient, the energy flow, and the deviation degree; among them, the third adjustment coefficient is used to adjust the relative relationship between the energy flow and the entropy.

[0187] In one implementation, the training module is further configured to determine the feature interaction matrix according to the historical sample input data input to the encoder during the historical training process;

[0188] Among them, the feature interaction matrix indicates the interaction result between multiple features included in the historical sample input data;

[0189] The training module is further configured to perform normalization processing on the feature interaction matrix to obtain the normalized feature interaction matrix;

[0190] The training module is further configured to according to the preset weight coefficient, the regularization coefficient for adaptively optimizing the loss function value of the feature interaction Determine the value of the adaptive feature interaction optimization loss function based on the low-dimensional latent space representation features and the historical weight matrix;

[0191] wherein, the weight coefficient adjusts the influence of the low-dimensional latent space representation features on the autoencoder; the regularization coefficient is used to prevent overfitting.

[0192] Thirdly, the present application also provides an electronic device, including a memory, a processor, and a computer program stored on the memory and executable on the processor. When the processor executes the computer program, the steps of S110 - S130 provided in the above embodiments are implemented.

[0193] Fourthly, the present application also provides a computer-readable storage medium, on which a computer program is stored. When the computer program is run by a processor, the steps of S110 - S130 in the above embodiments are executed.

[0194] Fifthly, the computer program product provided by the present application includes a computer-readable storage medium storing program code. The instructions included in the program code can be used to execute the methods in the foregoing method embodiments. For the specific implementation, reference can be made to the steps of S110 - S130 in the method embodiments, which will not be elaborated herein.

[0195] In the embodiments provided by the present application, it should be understood that the disclosed apparatus and method can be implemented in other ways. The apparatus embodiments described above are merely illustrative. For example, the division of the units is only a logical function division, and there may be other division methods in actual implementation. For another example, multiple units or components can be combined or integrated into another system, or some features can be ignored or not executed. Another point is that the displayed or discussed mutual coupling or direct coupling or communication connection can be through some communication interfaces. The indirect coupling or communication connection of the apparatus or unit can be in electrical, mechanical or other forms.

[0196] In addition, the units described as separate components may or may not be physically separated, and the components displayed as units may or may not be physical units, that is, they may be located in one place, or may be distributed to multiple network units. Some or all of the units can be selected according to actual needs to achieve the purpose of the solution of this embodiment.

[0197] Furthermore, in each embodiment of the present application, the functional modules can be integrated together to form an independent part, or each module can exist alone, or two or more modules can be integrated to form an independent part.

[0198] It should be noted that if a function is implemented in the form of a software functional module and sold or used as an independent product, it can be stored in a computer-readable storage medium. Based on this understanding, the technical solution of the present application, in essence, or the part that contributes to the prior art or a part of this technical solution can be embodied in the form of a software product. This computer software product is stored in a storage medium and includes several instructions for causing a computer device (which may be a personal computer, a server, or a network device, etc.) to execute all or part of the steps of the methods described in various embodiments of the present application. The aforementioned storage medium includes: various media such as USB flash drives, mobile hard disks, read-only memory (ROM), random access memory (RAM), magnetic disks, or optical discs that can store program codes.

[0199] In this text, relational terms such as "first" and "second" are only used to distinguish one entity or operation from another entity or operation, and do not necessarily require or imply any actual relationship or order between these entities or operations.

[0200] The above are only embodiments of the present application and are not used to limit the protection scope of the present application. For those skilled in the art, the present application can have various changes and modifications. Any modifications, equivalent replacements, improvements, etc. made within the spirit and principle of the present application shall be included in the protection scope of the present application.

Claims

1. A method for evaluating the heat treatment process performance of a safety hook based on a neural network, characterized in that The method includes: During the process of treating the safety hook with the heat treatment process to be evaluated, collecting various process data to be evaluated of the heat treatment process to be evaluated; Inputting the various process data to be evaluated into a pre-trained data compression model to obtain compressed data to be evaluated corresponding to the various process data to be evaluated; Wherein, the network structure of the data compression model is an autoencoder, and the autoencoder includes an encoder and a decoder; the parameters of both the encoder and the decoder include weight matrices; Inputting the compressed data to be evaluated into a pre-trained heat treatment process performance evaluation model to obtain an evaluation result of the heat treatment process to be evaluated; The method further includes: Determining a historical loss function value of the autoencoder during the historical training process according to the historical latent space representation features output by the encoder during the historical training process; Determining a historical gradient of the historical loss function value with respect to the historical weight matrix according to the historical loss function value and the historical weight matrix used by the encoder during the historical training process; Determining a conflict detection term of a target increment of the historical weight matrix according to a preset correction coefficient and the historical gradient; Wherein, the conflict detection term is used to adjust the update direction of the weight matrix; Determining the target increment according to a preset inertia parameter, a preset learning rate, the historical gradient, and the conflict detection term; Wherein, the inertia parameter is used to limit the influence of the update of the weight matrix during the historical training process on the target increment; Determining the target weight matrices used by the encoder and the decoder during the current training process according to the historical weight matrix and the target increment; Wherein, the current training process is the training process before reaching the training stop condition during the iterative training of the autoencoder.

2. The method according to claim 1, wherein During the iterative training of the autoencoder, the current training process before reaching the training stop condition includes: Determining current sample input data corresponding to each training sample according to a plurality of training samples; Wherein, the training samples include various sample process data, and the sample process data is collected during the sample heat treatment process of the safety hook; Inputting the current sample input data into the encoder to obtain a current low-dimensional latent space representation feature corresponding to the current sample input data; Inputting the current low-dimensional latent space representation feature into the decoder to obtain current sample compressed data corresponding to the low-dimensional latent space representation feature.

3. The method according to claim 2, wherein The determining the historical loss function value of the autoencoder during the historical training process according to the historical latent space representation features output by the encoder during the historical training process includes: Determining a feature weight weighting coefficient, an energy flow loss function value, and an adaptive feature interaction optimization loss function value according to the historical latent space representation features; Among them, the feature weight weighting coefficient indicates the influence of the historical latent space representation features on the target increment; the energy flow loss function value indicates the influence of the energy flow of the historical latent space representation features on the target increment, and the energy flow indicates the information amount of the historical latent space representation features; the adaptive feature interaction optimization loss function value indicates the influence of the features of each dimension in the historical latent space representation features on the target increment. Determine the historical loss function value according to the historical sample input data input to the encoder during the historical training process, a preset sparsity weighting coefficient, the feature weight weighting coefficient, the energy flow loss function value, and the adaptive feature interaction optimization loss function value.

4. The method according to claim 3, characterized in that, The determining the feature weight weighting coefficient according to the historical latent space representation features includes: Determine the feature contribution degree and information gain of the historical latent space representation features according to the historical latent space representation features. Among them, the feature contribution degree indicates the contribution of the historical latent space representation features to the historical sample compression data output by the decoder during the historical training process; the information gain indicates the information amount difference of the historical latent space representation features. Determine the feature weight weighting coefficient according to a preset first adjustment coefficient, the historical latent space representation features, the feature contribution degree, and the information gain. Among them, the first adjustment coefficient is used to adjust the compression and reconstruction of the historical sample input data by the encoder and the decoder.

5. The method according to claim 4, wherein The determining the feature contribution degree and information gain of the historical latent space representation features according to the historical latent space representation features includes: Determine the variance and entropy of the low-dimensional latent space representation features according to the historical latent space representation features; among them, the entropy indicates the discrete degree of the low-dimensional latent space representation features. Determine the feature contribution degree according to a preset second adjustment coefficient, the historical latent space representation features, the variance, and the entropy. Among them, the second adjustment coefficient is used to adjust the relative relationship between the variance and the entropy. Determine the information gain according to the variance.

6. The method according to claim 5, wherein The determining the energy flow loss function value according to the historical latent space representation features includes: Determine the energy flow of the low-dimensional latent space representation features and the deviation degree of the energy flow according to the historical latent space representation features. Determine the energy flow loss function value according to a preset third adjustment coefficient, the energy flow, and the deviation degree; among them, the third adjustment coefficient is used to adjust the relative relationship between the energy flow and the entropy.

7. The method according to claim 3, wherein The determining the adaptive feature interaction optimization loss function value according to the historical latent space representation features includes: Determine a feature interaction matrix according to the historical sample input data input to the encoder during the historical training process. Among them, the feature interaction matrix indicates the interaction result between multiple features included in the historical sample input data. Perform normalization processing on the feature interaction matrix to obtain a normalized feature interaction matrix. According to a preset weight coefficient and a regularization coefficient for adaptively optimizing a loss function value of feature interaction , and based on the low-dimensional latent space representation feature and the historical weight matrix, determine the value of the adaptive feature interaction optimization loss function; Among them, the weight coefficient adjusts the influence of the low-dimensional latent space representation features on the autoencoder; the regularization coefficient is used to prevent overfitting.

8. An evaluation device for the heat treatment process performance of a safety hook based on a neural network, characterized in that, For implementing the method according to claim 1, the apparatus comprises: an acquisition module, a compression module, and an evaluation module; The acquisition module is configured to acquire a plurality of process data to be evaluated of the heat treatment process to be evaluated during the process of treating a safety hook with the heat treatment process to be evaluated; The compression module is configured to input the plurality of process data to be evaluated into a pre-trained data compression model to obtain compressed data to be evaluated corresponding to the plurality of process data to be evaluated; The evaluation module is configured to input the compressed data to be evaluated into a pre-trained heat treatment process performance evaluation model to obtain an evaluation result of the heat treatment process to be evaluated.

9. The device according to claim 8, characterized in that, The network structure of the data compression model is an autoencoder, and the autoencoder comprises an encoder and a decoder; the apparatus further comprises: a training module; The training module is configured to perform iterative training on the autoencoder; Determine current sample input data corresponding to each training sample according to a plurality of training samples; Wherein, the training samples include a plurality of sample process data, and the sample process data is acquired during the process of performing a sample heat treatment process on a safety hook; Input the current sample input data into the encoder to obtain a current low-dimensional latent space representation feature corresponding to the current sample input data; Input the current low-dimensional latent space representation feature into the decoder to obtain current sample compressed data corresponding to the low-dimensional latent space representation feature.

Citation Information

Patent Citations

  • Thermal fatigue assessment method and device for safety hook rigging based on artificial intelligence

    CN119783009A