A data processing method for seabed-based monitoring equipment

Through technologies such as singular value decomposition, sliding time domain window segmentation, CUDA parallel processing flow and deep-sea environmental perception attention network model, the problem of indistinguishable environmental changes and equipment drift in seabed-based monitoring equipment is solved, and the accuracy and reliability of monitoring data is improved. It is suitable for marine scientific research and submarine resource development.

CN120217209BActive Publication Date: 2025-08-05青岛道万科技有限公司
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
CN202510668123.0
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2025-05-23
Publication Date
2025-08-05
Estimated Expiration
2045-05-23

AI Technical Summary

Technical Problem

Traditional seabed-based monitoring equipment is difficult to effectively distinguish data drift from environmental changes in long-term work, resulting in a decrease in monitoring accuracy and reliability.

Method used

Technical means such as singular value decomposition, sliding time domain window segmentation, CUDA parallel processing flow, tensor decomposition and deep-sea environment perception attention network model are used to construct a seabed environmental state tensor model, and distinguish environmental changes from equipment drift through multi-stage data processing flow.

Benefits of technology

It significantly improves the accuracy and reliability of long-term monitoring data, realizes stable monitoring of the seabed environment, and is suitable for marine scientific research and seabed resource development.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120217209B_ABST
    Figure CN120217209B_ABST
Patent Text Reader

Abstract

The present invention provides a data processing method for seabed-based monitoring equipment, belonging to the technical field of electrical digital data processing. In this invention, the collected data is normalized and outliers are removed, and the singular value decomposition method is used to extract key features, and then time series data blocks are formed by adaptive sliding time-domain window segmentation. The method innovatively uses the CUDA parallel processing stream to construct a disorder matrix and a drift matrix, and through tensor decomposition technology, the disorder pattern feature vectors, drift principal components and abnormal feature matrices are fused to construct a seabed environmental state tensor model. At the same time, the deep-sea environmental perception attention network model is used to learn the drift law in the long time series, and the real-time monitoring data is corrected and compensated. Finally, a hierarchical clustering algorithm is used to establish a seabed state evaluation standard, effectively solving the key technical problem of difficult distinction between environmental changes and equipment drift in seabed monitoring data.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present invention belongs to the technical field of electronic digital data processing. Specifically, it relates to a data processing method for seabed-based monitoring equipment. Background Art

[0002] Seabed-based monitoring equipment is widely used in fields such as marine scientific research, seabed resource exploration, and marine environmental protection. Traditional seabed-based monitoring technologies mainly rely on multi-sensor arrays to collect seabed environmental parameters, obtain raw data through a data acquisition unit, and then perform basic processing such as noise reduction and filtering through a signal processing module. These devices usually use simple outlier detection and statistical filtering methods to process data and can provide reliable monitoring results in a relatively stable environment.

[0003] However, traditional technologies face serious challenges in long-term seabed monitoring. Due to the complex and variable seabed environment, sensors will age and drift during long-term operation, resulting in a gradual shift of the measurement baseline. At the same time, it is difficult to distinguish between natural changes such as turbulence and tides in the seabed environment and data anomalies caused by sensor drift. Existing data processing methods often misjudge natural environmental changes as equipment failures or misinterpret sensor drift as environmental anomalies, seriously affecting monitoring accuracy and reliability.

[0004] Even more intractable is that traditional technologies lack effective means to simultaneously process the spatio-temporal correlation and long-term drift of multi-source data, especially in the seabed environment with limited computing resources, where real-time and efficient processing of a large amount of monitoring data cannot be achieved. The environmental changes and equipment drift in seabed data are intertwined, forming a complex mixed pattern. Traditional single processing methods are difficult to solve this core technical problem, seriously restricting the long-term stable application of seabed monitoring technologies. That is to say, there is a technical problem in the prior art that it is difficult to effectively distinguish the data drift and environmental change characteristics during the long-term operation of seabed-based monitoring equipment. Summary of the Invention

[0005] In view of this, the present invention provides a data processing method for seabed-based monitoring equipment, which can solve the technical problem in the prior art that it is difficult to effectively distinguish the data drift and environmental change characteristics during the long-term operation of seabed-based monitoring equipment.

[0006] The present invention is implemented as follows: The present invention provides a data processing method for seabed-based monitoring equipment, including: performing normalization processing on the data collected by the seabed-based monitoring equipment and eliminating abnormal data to determine an effective data set; using the singular value decomposition method to perform dimensionality reduction processing on the effective data set to extract key features and generate a feature matrix; implementing sliding time-domain window segmentation on the feature matrix to form multiple groups of time series data blocks; starting a CUDA parallel processing stream to construct a disorder matrix and a drift matrix; extracting disorder pattern feature vectors and drift principal components in the CUDA parallel processing stream; extracting abnormal data features in the CPU processing stream to generate an abnormal feature matrix; using tensor decomposition technology to fuse the disorder pattern feature vectors, drift principal components, and abnormal feature matrix to construct a seabed environmental state tensor model; constructing a data drift correction model based on a residual neural network to correct and compensate real-time monitoring data; and using a hierarchical clustering algorithm to classify the processed data to establish a seabed state evaluation criterion.

[0007] Among them, the singular value decomposition method specifically organizes the seabed monitoring data into a matrix form, and by calculating eigenvalues and eigenvectors, decomposes the data matrix into the product of three sub-matrices, retains the eigenvectors corresponding to the larger singular values, and discards noise and redundant information.

[0008] Among them, the disorder matrix specifically refers to a feature matrix constructed by calculating the volatility, uncertainty, and disorder degree of seabed monitoring data in the time and space dimensions, and is used to quantitatively describe the turbulence, perturbation, and nonlinear dynamic processes in the seabed environment.

[0009] Among them, the drift matrix specifically refers to a matrix constructed by comparing the deviation degree of the real-time readings of sensors with the historical reference values. Each element in the matrix represents the drift amount of the corresponding sensor at the corresponding time point, and is used to track and analyze the performance changes of sensors and the long-term evolution trend of the environment.

[0010] Among them, an adaptive spectrum optimization function is used in the process of constructing the disorder matrix, which is dynamically adjusted according to the frequency characteristic differences of the monitoring data under different seabed environmental conditions. By dynamically adjusting the analysis window length and frequency resolution, the accurate capture of data characteristics under different seabed environments is realized.

[0011] Among them, a pre-trained deep-sea environment perception attention network model is used in the data drift correction process. This model integrates the Transformer architecture and the recurrent neural network structure, and introduces a multi-scale drift perception attention mechanism designed for the characteristics of seabed data, which can identify and compensate complex drift patterns in seabed monitoring data at different time scales.

[0012] Among them, the specific structure of the deep - sea environmental perception attention network model is a hybrid architecture that combines a multi - layer bidirectional recurrent neural network and a multi - head self - attention mechanism. The bottom layer uses a convolutional neural network to extract features from multi - source sensor data. The middle layer uses six - layer Transformer encoders to extract temporal features and cross - sensor correlation features. The top layer uses a fully - connected network with skip connections for data drift estimation and compensation.

[0013] Among them, the deep - sea environmental perception attention network model introduces a sparse attention mechanism for seabed feature perception in the Transformer encoder. The sparse attention mechanism dynamically adjusts the attention range according to the physical characteristics of the seabed environment, and can simultaneously focus on short - term fluctuations and long - term drift patterns. The entire model adopts a hierarchical design, and processing modules are designed respectively for data characteristics under different depths and environmental conditions.

[0014] Among them, the pre - training process of the deep - sea environmental perception attention network model uses long - term monitoring data collected from the global seabed monitoring network, screens data segments containing complete environmental change cycles and sensor drift phenomena, and uses physical oceanography models to generate synthetic data to enhance the coverage of the dataset, forming a comprehensive training dataset covering various typical seabed environments such as shallow - sea areas, deep - sea plains, trench areas, and hydrothermal areas.

[0015] Among them, the seabed - based monitoring device mainly consists of a sensor array system, a data acquisition unit, a signal processing module, an energy supply system, a communication transmission module, and a protective housing. The sensor array system includes a pressure sensor, a temperature sensor, a flow velocity sensor, a seismic sensor, and a chemical substance detector.

[0016] Compared with the prior art, the present invention provides a data processing method for a seabed - based monitoring device. The present invention proposes a data processing method for a seabed - based monitoring device, which effectively solves the problem of difficult to distinguish environmental changes and device drift during long - term monitoring through a multi - stage data processing flow. This method first normalizes and removes outliers from the original data, and uses singular value decomposition to extract key features; then uses an adaptive sliding window to segment the time series, constructs a chaos matrix and a drift matrix; finally, through tensor decomposition and the deep - sea environmental perception attention network model, accurately distinguishes environmental changes and device drift.

[0017] The present invention effectively overcomes the limitations of traditional technologies. Through the collaborative work of CUDA parallel processing streams and CPU processing streams, it realizes the efficient processing of large - scale seabed monitoring data; the multi - dimensional tensor model successfully captures the high - order correlations between different sensor data; the pre - trained deep - sea environmental perception attention network can identify and compensate complex drift patterns on multiple time scales, effectively distinguishing natural environmental changes and device drift.

[0018] Through the organic combination of the above technical means, the present invention has successfully solved the core technical problem of the difficulty in distinguishing environmental changes from equipment drift in seabed-based monitoring data, significantly improved the accuracy and reliability of long-term monitoring data, provided a strong guarantee for the long-term stable monitoring of the seabed environment, and has important practical value for marine scientific research and seabed resource development. BRIEF DESCRIPTION OF THE DRAWINGS

[0019] Figure 1 It is a flowchart of the method of the present invention.

[0020] Figure 2 It is a schematic diagram of the composition of the seabed-based monitoring equipment in Embodiment 1. DETAILED DESCRIPTION OF THE EMBODIMENTS

[0021] To make the purpose, technical solutions and advantages of the embodiments of the present invention clearer, the technical solutions in the embodiments of the present invention will be clearly and completely described below in conjunction with the drawings in the embodiments of the present invention.

[0022] As Figure 1 shown, it is a flowchart of a data processing method for a seabed-based monitoring equipment provided by the present invention. This method includes the following steps:

[0023] S01. Perform normalization processing on the data collected by the seabed-based monitoring equipment, and use an outlier detection algorithm to identify and remove abnormal data outside the predetermined threshold range to determine an effective data set;

[0024] S02. Based on the effective data set, use the singular value decomposition method to perform dimensionality reduction processing on the original data matrix, eliminate data redundancy, extract key features, and generate a feature matrix;

[0025] S03. Perform sliding time-domain window segmentation on the feature matrix, and adaptively adjust the window width according to the characteristics of seabed environmental changes to form multiple groups of time series data blocks;

[0026] S04. Start a CUDA parallel processing stream to receive the time series data blocks, calculate the disorder degree index within each data block, and construct a disorder matrix; at the same time, count the degree of deviation of the sensor readings from the reference value to construct a drift matrix;

[0027] S05. Perform singular value decomposition on the disorder matrix in the CUDA parallel processing stream to extract disorder mode feature vectors; perform principal component analysis on the drift matrix to extract drift principal components;

[0028] S06. Extract abnormal data features in the CPU processing stream to generate an abnormal feature matrix, and align the abnormal feature matrix with the disorder mode feature vectors and the drift principal components;

[0029] S07. Use tensor decomposition technology to fuse the disordered pattern eigenvectors, the drift principal components, and the abnormal feature matrix, construct a seabed environmental state tensor model, and extract the coupling relationship between environmental factors and monitoring data;

[0030] S08. Based on a residual neural network, construct a data drift correction model. Input the drift principal components, train the network to learn the drift patterns in long time series, and perform correction and compensation on real-time monitoring data;

[0031] S09. Use a hierarchical clustering algorithm to classify the processed data, establish seabed state evaluation criteria, and generate multi-level seabed state evaluation results;

[0032] Among them, the singular value decomposition method specifically organizes the seabed monitoring data into a matrix form. By calculating eigenvalues and eigenvectors, the data matrix is decomposed into the product of three sub-matrices. Retain the eigenvectors corresponding to the larger singular values and discard noise and redundant information.

[0033] Among them, the sliding time-domain window segmentation specifically divides the continuous time series data into a series of overlapping data segments according to a preset time length. There is a certain proportion of overlap between adjacent windows, which is used to capture the time correlation and gradual change characteristics in the data. <~

[0034] Among them, the tensor decomposition technology specifically represents multi-dimensional data as a high-order tensor. Through multi-linear algebraic operations, the tensor is decomposed into the outer product of multiple low-dimensional factors, which can retain the high-order correlation between data and discover potential data structures.

[0035] Among them, the CUDA parallel processing stream specifically refers to a parallel computing process implemented using the Compute Unified Device Architecture technology on a graphics processing unit. By distributing large-scale matrix operations to thousands of computing cores for simultaneous execution, efficient and intensive computing is achieved.

[0036] Among them, data drift specifically refers to the phenomenon of slow deviation of the measurement baseline of seabed-based monitoring equipment during long-term operation due to sensor aging, environmental changes, or the characteristics of the equipment itself, which affects the accuracy of the data.

[0037] Among them, the disorder matrix specifically refers to a feature matrix constructed by calculating the volatility, uncertainty, and disorder degree of seabed monitoring data in the time and space dimensions, which is used to quantitatively describe the turbulence, perturbation, and nonlinear dynamic processes in the seabed environment.

[0038] Among them, the drift matrix specifically refers to a matrix constructed by comparing the deviation degree of the real-time readings of sensors with the historical reference values. Each element in the matrix represents the drift amount of the corresponding sensor at the corresponding time point, which is used to track and analyze the performance changes of sensors and the long-term evolution trend of the environment.

[0039] The seabed-based monitoring device mainly consists of a sensor array system, a data acquisition unit, a signal processing module, an energy supply system, a communication transmission module, and a protective housing. The sensor array system includes a pressure sensor, a temperature sensor, a flow velocity sensor, a seismic sensor, and a chemical substance detector, which conduct all-round monitoring of the seabed environment through the collaborative work of multi-source sensors. The data acquisition unit is responsible for sampling and preliminary processing of the signals from each sensor. The signal processing module contains a digital signal processor and an embedded computing unit for executing data processing algorithms. The energy supply system uses a combination of lithium batteries and a seawater energy harvesting device to provide long-term energy for the device. The communication transmission module is responsible for transmitting the processed data to the sea surface or shore-based stations through acoustic communication or an optical cable network. The entire system is protected by a waterproof and pressure-resistant titanium alloy housing to ensure stable operation in the high-pressure deep-sea environment.

[0040] The adaptive spectrum optimization function is used to optimize the construction process of the disorder matrix in S04, and dynamically adjusts according to the frequency characteristic differences of the monitoring data under different seabed environmental conditions. The inputs include a time series data block, an environmental type identification parameter, historical spectrum baseline data, a signal-to-noise ratio threshold, and a frequency range limit parameter. The output is an optimized disorder matrix, which includes the main frequency components, the distribution of disorder degree indicators, and the frequency drift quantification indicators. This function realizes the accurate capture of data characteristics under different seabed environments by dynamically adjusting the analysis window length and frequency resolution, and is especially suitable for processing seabed monitoring data with long-term slow drift.

[0041] The pre-trained deep-sea environment perception attention network model is used to optimize the data drift correction process in S08. This model integrates the Transformer architecture and the recurrent neural network structure, and introduces a multi-scale drift perception attention mechanism designed for the characteristics of seabed data, which can identify and compensate for complex drift patterns in seabed monitoring data at different time scales. The attention weight allocation parameters in the model need to be determined according to three key parameters: the sensor drift rate of the seabed monitoring device, the environmental change rate, and the data acquisition frequency.

[0042] The specific structure of the deep - sea environment perception attention network model is a hybrid architecture that combines a multi - layer bidirectional recurrent neural network and a multi - head self - attention mechanism. The bottom layer uses a convolutional neural network to extract features from multi - source sensor data. The middle layer uses a six - layer Transformer encoder to extract temporal features and cross - sensor correlation features. The top layer uses a fully - connected network with skip connections for data drift estimation and compensation. A sparse attention mechanism for seabed characteristics perception is introduced into the Transformer encoder. The sparse attention mechanism dynamically adjusts the attention range according to the physical characteristics of the seabed environment, and can simultaneously focus on short - term fluctuations and long - term drift patterns. The entire model adopts a hierarchical design, and processing modules are designed respectively for data characteristics under different depths and environmental conditions.

[0043] The steps for establishing the training data set during the pre - training process of the deep - sea environment perception attention network model specifically include collecting long - term monitoring data from the global seabed monitoring network, screening data segments containing complete environmental change cycles and sensor drift phenomena, expert - annotating each data segment to mark the normal environmental change and sensor drift parts, generating synthetic data using physical oceanography models to enhance the data set coverage, mixing real data and synthetic data in a ratio of 4:1 to construct the training set and validation set, constructing sub - data sets for transfer learning for different seabed environment types respectively, and finally forming a comprehensive training data set covering various typical seabed environments such as shallow - sea areas, deep - sea plains, trench areas, and hydrothermal areas.

[0044] The steps for pre - training the deep - sea environment perception attention network model specifically include first performing initialization training on synthetic data to learn basic temporal patterns, then performing supervised learning on the global data set to master the general seabed environmental change rules, then performing fine - tuning training on various types of seabed environment data to enhance the model's adaptability in the environment, using contrastive learning methods to train the model to distinguish environmental changes and sensor drift, introducing a curriculum learning strategy to gradually transition from simple patterns to complex drift patterns, using multi - task learning to simultaneously optimize the two tasks of drift detection and compensation, and finally using knowledge distillation technology to compress the large - scale model into a lightweight model suitable for deployment on the limited computing resources of seabed monitoring devices. The entire pre - training process is executed in parallel on multiple graphics processing unit servers using a distributed computing framework to ensure that the model can fully learn the complex patterns in long - temporal data.

[0045] The specific implementation manners of the above steps are described in detail below. The specific implementation manner of step S01 is to first perform normalization processing on the original data collected by the seabed-based monitoring equipment, unify the data with different dimensions into the range of 0 to 1. Specifically, the min-max normalization method is adopted. That is, for each sensor data sequence, the normalized value is calculated as the original value minus the minimum value divided by the difference between the maximum value and the minimum value, where the minimum value and the maximum value are the minimum value and the maximum value of the sensor data sequence respectively. Then, an improved local outlier factor algorithm is used to detect outliers. This algorithm identifies outliers by calculating the local density ratio of a sample point to its k nearest neighbor sample points. The value of k is adaptively selected according to the scale of the dataset, generally taking 5% to 10% of the total number of samples. The outlier determination threshold is set to 2.5. That is, when the local outlier factor of a certain data point is greater than 2.5, it is determined as an outlier and removed. The purpose of this step is to eliminate noise and outliers in the data, ensure the quality of the data for subsequent analysis and processing, and improve the reliability and accuracy of data processing.

[0046] The specific implementation manner of step S02 is to organize the effective dataset obtained in step S01 into a matrix form, where the rows represent time points and the columns represent the measurement values of different sensors. Then, singular value decomposition is applied to the matrix, and the original matrix is decomposed into the product of three matrices, including two orthogonal matrices and a diagonal matrix. The elements on the diagonal matrix are singular values. According to the magnitudes of the singular values, the first r singular values and the corresponding eigenvectors with a cumulative contribution rate reaching 95% are selected to form the feature matrix after dimensionality reduction. In practice, the value of r is usually selected such that the retained singular value energy accounts for more than 95% of the total energy. For typical seabed monitoring data, the value of r is generally between 10 and 20. The purpose of this step is to reduce the data dimension, eliminate redundant information and noise in the data, extract key information reflecting the main characteristics of the seabed environment, and provide a more compact and effective data representation for subsequent processing.

[0047] The specific implementation of step S03 is to perform time-domain segmentation on the feature matrix generated in step S02 using an adaptive sliding window method. First, the basic window width is set to 1 hour, and an adaptive adjustment coefficient is set according to the characteristics of seabed environment changes. This coefficient depends on the environmental change rate and data acquisition frequency, and generally ranges from 0.5 to 2.0. The variance change rate of real-time monitoring data is monitored. When the variance change rate exceeds a preset threshold (usually 20%), the window width is adjusted according to the formula of the basic window width multiplied by one plus the adjustment coefficient multiplied by the change rate, ensuring that the window width is reduced when the environment changes violently and enlarged when the environment is relatively stable. The overlap rate between adjacent windows is set to 50% to ensure data continuity and temporal correlation. Through this adaptive window segmentation method, the feature matrix is divided into multiple groups of time-series data blocks, and each data block contains all sensor feature data within a certain period of time. The purpose of this step is to adapt to the dynamic change characteristics of the seabed environment, ensure that subtle change information can be captured when the environment changes violently, and reduce data redundancy and improve processing efficiency when the environment is relatively stable.

[0048] The specific implementation of step S04 is to start a CUDA parallel processing stream and allocate computing resources to process multiple groups of time-series data blocks generated in step S03. For each data block, calculate the disorder degree indicators, including entropy value, Lyapunov exponent, and complexity, etc. Specifically, the sample entropy algorithm is used to calculate the complexity and irregularity of the time series, the embedding dimension is set to 2, and the similarity tolerance is set to 0.2 times the standard deviation. At the same time, calculate the Lyapunov exponent to evaluate the chaos degree of the sequence, the reconstruction delay is selected as 1 / 4 period, and the reconstruction dimension is determined by the false nearest neighbor method. These disorder degree indicators are organized into a disorder matrix, and each element in the matrix represents the disorder degree indicator of a certain sensor within a certain time window. In parallel, calculate the deviation degree of each sensor reading relative to the historical reference value, and construct a drift matrix, where each element represents the drift amount of a certain sensor within a certain time window. The drift amount is calculated using the exponentially weighted moving average method, the smoothing factor is set to 0.1, and the threshold is set to 3 times the standard deviation. This step uses the CUDA parallel computing architecture to distribute large-scale matrix operations to thousands of computing cores for simultaneous execution, significantly improving processing efficiency. The purpose is to quantitatively describe the turbulence, perturbation, and sensor drift in the seabed environment and provide a basis for subsequent analysis.

[0049] The specific implementation of step S05 is to perform singular value decomposition on the disorder matrix constructed in step S04 in the CUDA parallel processing stream, and decompose the matrix into the product of three matrices. Analyze the singular value distribution, and select the right singular vectors corresponding to the first k largest singular values as the disorder pattern feature vectors. The value of k is selected such that the cumulative contribution rate reaches 90%. At the same time, perform principal component analysis on the drift matrix. Calculate the covariance matrix as the transpose of the drift matrix multiplied by the drift matrix and then divided by the number of samples minus 1, where the number of samples is the number of time windows. Solve the characteristic equation to obtain the eigenvalues and eigenvectors, and select the principal components with an explained variance of 85% as the drift principal components. This step is executed in parallel on the GPU, making full use of the high-performance computing capabilities of the CUDA architecture. The purpose is to extract the main patterns and features from the disorder matrix and the drift matrix, reduce the data dimension, and extract the key information reflecting the disorder state of the seabed environment and the sensor drift trend.

[0050] The specific implementation of step S06 is to extract features from the abnormal data identified in step S01 in the CPU processing stream. First, group the abnormal data according to time windows, and calculate the statistical features of the abnormal data within each window, including frequency, amplitude, spatial distribution, and time distribution, etc. Use the kernel density estimation method to analyze the distribution characteristics of the abnormal data, and select the bandwidth parameter h as 0.1. Then extract the abnormal pattern features, and apply the random forest algorithm to identify the main types and features of the abnormal data. The number of trees is set to 100, and the minimum number of samples in the leaf nodes is set to 5. Organize the extracted abnormal features into an abnormal feature matrix A, and align A with the disorder pattern feature vectors and drift principal components extracted in step S05 in the time dimension through timestamp alignment to generate a unified time index, ensuring the temporal consistency of the three types of feature data. The purpose of this step is to extract the environmental information contained in the abnormal data and align it with the disorder pattern and drift principal components in the time dimension, providing a basis for subsequent data fusion.

[0051] The specific implementation of step S07 is to fuse the disorder pattern feature vectors, drift principal components, and anomaly feature matrices obtained in steps S05 and S06 using tensor decomposition technology. First, a third-order tensor X is constructed, with its three dimensions corresponding to time, sensor type, and feature type (disorder, drift, anomaly). Then, the Tucker decomposition is applied to decompose the tensor X into the product of a core tensor G and three factor matrices A, B, and C, i.e., X≈G×1A×2B×3C, where ×n represents the tensor-matrix product along the nth dimension. The dimension of the core tensor is selected based on empirical rules, generally taking 30% - 50% of the original size of each dimension. By analyzing the element sizes of the core tensor and their corresponding factor vectors, the coupling relationships between different features are identified, and a seabed environmental state tensor model is constructed. This model can express the complex associations among environmental factors, sensor characteristics, and data anomalies, providing a theoretical basis for multi-dimensional information fusion in seabed state assessment. The purpose of this step is to discover the hidden multi-dimensional association patterns in the data through high-order tensor decomposition and construct a tensor model that comprehensively reflects the seabed environmental state.

[0052] The specific implementation of step S08 is to construct a data drift correction model based on a residual neural network. The network structure includes 1 input layer, 4 residual blocks, and 1 output layer. Each residual block consists of 2 convolutional layers and 1 shortcut connection. The input layer receives the drift principal components extracted in step S05, and the output layer generates drift correction coefficients. The network training uses the Adam optimizer, with the initial learning rate set to 0.001 and adjusted using the cosine annealing strategy. The loss function uses a combination of mean squared error and L1 regularization, with the regularization coefficient set to 0.0001. The training data selects data segments with known drift patterns from historical monitoring data and divides them into a training set and a validation set in an 8:2 ratio. During the training process, when the validation loss does not decrease for 5 consecutive epochs, the early stopping mechanism is triggered. The trained model is applied to real-time monitoring data, and correction coefficients are generated according to the identified drift patterns to correct and compensate the data. The purpose of this step is to learn the long-term drift rules in seabed monitoring data, establish an effective drift correction model, and improve the accuracy and reliability of long-term monitoring data.

[0053] The specific implementation of step S09 is to classify the data processed in the previous steps using the hierarchical clustering algorithm. First, define the distance metric between data points, calculate the similarity between samples using the Mahalanobis distance, and consider the covariance structure of the data distribution. Then, use the Ward minimum variance method as the clustering criterion. This method selects the two clusters that result in the smallest increase in within-cluster variance for each merge step. The clustering process starts from individual samples, gradually merges the most similar clusters, and finally forms a hierarchical structure. Evaluate the quality of different clustering results by calculating the silhouette coefficient and the Davies-Bouldin index, and determine the optimal number of clusters, which is generally between 4 and 8. According to the knowledge of physical oceanography and combined with expert experience, assign physical meanings to each cluster, establish a seabed state assessment standard, and divide it into four levels: normal, slightly abnormal, moderately abnormal, and severely abnormal. Finally, generate a seabed state assessment report based on the clustering results of real-time monitoring data. The purpose of this step is to scientifically classify the processed data, establish an objective seabed state assessment standard, and provide decision-making support for seabed environmental monitoring and early warning.

[0054] The structure of the deep-sea environment perception attention network model adopts a hybrid architecture that combines a multi-layer bidirectional recurrent neural network and a multi-head self-attention mechanism. The bottom layer consists of 3 convolutional blocks, each convolutional block contains 2 one-dimensional convolutional layers and 1 max-pooling layer, the convolutional kernel sizes are 3 and 5 respectively, and the number of channels doubles layer by layer starting from 32. The middle layer consists of 6 layers of Transformer encoders, each layer contains a multi-head self-attention sublayer and a feed-forward neural network sublayer, the number of attention heads is set to 8, and the hidden layer dimension is 512. The top layer consists of 4 layers of fully connected networks, the number of neurons is 256, 128, 64, 32 in sequence, the activation function uses ReLU, and skip connections are added between the first 3 layers of the top layer to alleviate the problem of gradient disappearance. A sparse attention mechanism for seabed feature perception is introduced in the Transformer encoder, which dynamically adjusts the attention distribution according to the physical characteristics of the seabed environment, and the attention sparsity parameter is set to 0.7. The entire model adopts a hierarchical design, and processing modules are designed respectively for the data characteristics under different depths and environmental conditions. Shallow sea areas, deep-sea plains, trench areas, and hydrothermal areas correspond to different attention modules, and the module selection is controlled by the environmental type parameter.

[0055] The construction process of the training dataset for the deep - sea environment perception attention network model first collects long - term monitoring data from the public data in the global seabed monitoring network, including the monitoring site data in typical areas such as the Mid - Atlantic Ridge, the East Pacific Rise, and the Mariana Trench, with a time span of at least 3 years. Clean and pre - process the collected raw data, removing data segments with obvious errors and a large number of missing values. Then, screen data segments that contain a complete environmental change cycle and sensor drift phenomena, with the data segment length being at least 1 month. Invite more than 5 oceanography experts to annotate each data segment, distinguishing between normal environmental changes and sensor drift parts, and the annotation consistency requirement is that the Cohen's Kappa coefficient is not less than 0.8. Based on the physical oceanography model and the existing annotated data, generate synthetic data that is 10 times the amount of the original data to enhance the coverage of the dataset. Mix the real data and the synthetic data in a ratio of 4:1, and divide them into a training set, a validation set, and a test set in a ratio of 7:2:1. Construct sub - datasets for typical seabed environments such as shallow - sea areas (water depth < 200 meters), deep - sea plains (water depth 200 - 4000 meters), trench areas (water depth > 4000 meters), and hydrothermal areas. Each sub - dataset contains at least 100 valid data segments. Finally, form a comprehensive training dataset covering various typical seabed environments, with a total data volume of not less than 10TB, ensuring that the model can learn the data characteristics and drift rules under different seabed environments.

[0056] The following details the mathematical models or calculation processes involved in the present invention.

[0057] In step S01, the normalization formula for the data collected by the seabed - based monitoring device is expressed as follows:

[0058] ;

[0059] In the formula, is the normalized data value; is the original data value; is the minimum value of the sensor data sequence; is the maximum value of the sensor data sequence.

[0060] Outlier detection uses the Local Outlier Factor (LOF) algorithm, and its calculation formula is:

[0061] ;

[0062] In the formula, is the local outlier factor of point ; is the nearest neighbor point set of point ; is point Local reachability density; is the number of nearest neighbors.

[0063] The formula for local reachability density is:

[0064] [[ID=ll]];

[0065] In the formula, is the reachability distance from point to point and is defined as:

[0066] ;

[0067] In the formula, is the distance from point to its th nearest neighbor; is the Euclidean distance between point and point

[0068] In step S02, the matrix representation of singular value decomposition (SVD) is:

[0069] ;

[0070] In the formula, is the original data matrix, is the number of time points, is the number of sensors; and are orthogonal matrices respectively; is a diagonal matrix, and the elements on the diagonal are singular values.

[0071] The formula for the feature matrix after dimensionality reduction is:

[0072] ;

[0073] In the formula, contains the first columns of ; contains the diagonal matrix composed of the first singular values of ; contains the first columns of ; The selection of

[0074] ;

[0075] In the formula, is the​ Singular value

[0076] In step S03, the calculation formula for the adaptive sliding window width is:

[0077] ;

[0078] In the formula, is the actual window width; is the basic window width, set to 1 hour; is the adjustment coefficient, with a value range of 0.5 to 2.0; is the data variance change rate, and the calculation formula is:

[0079] ;

[0080] In the formula, is the data variance at the current moment; is the data variance at the previous moment.

[0081] In step S04, the calculation of the disorder index includes the calculation formula for sample entropy:

[0082] ;

[0083] In the formula, is the sample entropy value; is the embedding dimension, with a value of 2; is the similarity tolerance, with a value of 0.2 times the data standard deviation; is the data length; is at dimension the number of matching pattern pairs; is at dimension the number of matching pattern pairs.

[0084] The calculation process of the Lyapunov exponent first requires phase space reconstruction:

[0085] ;

[0086] In the formula, is the reconstructed state vector; is the th point of the time series; is the reconstruction delay, with a value of 1 / 4 of the period; is the reconstruction dimension, determined by the false nearest neighbor method.

[0087] The calculation formula for the Lyapunov exponent is:

[0088] ;

[0089] In the formula, is the Lyapunov exponent; is the th time point; is the distance between adjacent orbits in the phase space at time

[0090] The drift amount is calculated using the exponentially weighted moving average method, and the formula is:

[0091] ;

[0092] In the formula, is the smoothed value at time ; is the observed value at time ; is the smoothing factor, with a value of 0.1; is the smoothed value at time ;

[0093] The drift amount is defined as:

[0094] ;

[0095] In the formula, is the drift amount at time ; is the observed value at time ; is the smoothed value at time ;

[0096] In step S05, perform singular value decomposition on the disorder matrix :

[0097] ;

[0098] In the formula, is the disorder matrix, is the number of time windows, is the number of sensors; and are orthogonal matrices respectively; is a diagonal matrix, and the elements on the diagonal are singular values.

[0099] The selection of the disorder mode eigenvector satisfies:

[0100] ;

[0101] In the formula, is the th singular value; is the number of selected eigenvectors.

[0102] For the drift matrix Perform principal component analysis, and the covariance matrix calculation formula is:

[0103] ;

[0104] In the formula, is the covariance matrix; is the drift matrix, is the number of samples, is the number of sensors.

[0105] The selection of principal components satisfies:

[0106] ;

[0107] In the formula, is the covariance matrix of the th eigenvalue; is the number of selected principal components.

[0108] In step S06, kernel density estimation is used to analyze the distribution characteristics of abnormal data, and its calculation formula is:

[0109] ;

[0110] In the formula, is the kernel density estimate value; is the number of samples; is the bandwidth parameter, and the value is 0.1; is the kernel function, and usually the Gaussian kernel is selected, which is expressed as:

[0111] ;

[0112] In the formula, is the standardized distance.

[0113] In addition, in step S06, the random forest algorithm is used to identify the main types and characteristics of abnormal data, and its Gini impurity calculation formula is:

[0114] ;

[0115] In the formula, is the Gini impurity of the dataset ; is the number of classes; is the proportion of samples in the th class in the dataset.

[0116] The calculation formula of information gain is:

[0117] ;

[0118] In the formula, is the information gain of feature ; is the number of values of feature ; is the sample subset where the value of feature is ; and are the number of samples in the subset and the total dataset, respectively.

[0119] In step S07, the tensor decomposition technique applies Tucker decomposition, expressed as:

[0120] ;

[0121] In the formula, is the original third-order tensor, is the size of the time dimension, is the size of the sensor type dimension, is the size of the feature type dimension; is the core tensor, , , ; , , are the factor matrices; represents the tensor-matrix product along the th dimension.

[0122] The formula for dimension selection of the core tensor is:

[0123] ;

[0124] ;

[0125] ;

[0126] In the formula, represents rounding up; , , are adjustment parameters, and the value range is 0 to 0.2 times the original dimension size.

[0127] In step S08, the loss function of the residual neural network is:

[0128] ;

[0129] In the formula, is the total loss; is the mean square error, and the calculation formula is:

[0130] ;

[0131] Wherein, is the number of samples; is the true value; is the predicted value.

[0132] is the L1 regularization term, and its calculation formula is:

[0133] ;

[0134] Wherein, is the number of model parameters; is the th model parameter; is the regularization coefficient, and its value is 0.0001.

[0135] The learning rate is adjusted using the cosine annealing strategy, and the formula is:

[0136] ;

[0137] Wherein, is the learning rate of the th round; is the minimum learning rate, and its value is 0.00001; is the maximum learning rate, and its value is 0.001; is the total number of rounds; is the current round.

[0138] In step S09, the Mahalanobis distance calculation formula is:

[0139] ;

[0140] Wherein, is the Mahalanobis distance between samples and ; is the covariance matrix of the data.

[0141] The merging criterion of Ward's minimum variance method is:

[0142] ;

[0143] Wherein, is the variance increment after the merger of clusters and cluster ; and are the number of samples in clusters and cluster respectively; and are the clusters and clusters mean vector; represents the Euclidean norm.

[0144] The formula for the silhouette coefficient is:

[0145] ;

[0146] In the formula, is the silhouette coefficient of sample ; is the average distance between sample and other samples in the same cluster; is the average distance between sample and the nearest non - same - cluster.

[0147] The formula for the Davies - Bouldin index is:

[0148] ;

[0149] In the formula, is the Davies - Bouldin index; is the number of clusters; is the average distance from the samples of cluster to the cluster center; is the distance between the center of cluster and the center of cluster ;

[0150] When selecting the optimal number of clusters, the silhouette coefficient and the Davies - Bouldin index are considered comprehensively. The number of clusters with the largest silhouette coefficient and the smallest Davies - Bouldin index is taken, and at the same time, the knowledge of physical oceanography is combined for verification to ensure that the clustering results have practical physical significance.

[0151] The Local Outlier Factor algorithm detects outliers by comparing the density differences between sample points and their local neighborhoods. Compared with traditional distance - based or statistical methods, it is more suitable for processing data with complex distributions. This algorithm considers the local distribution characteristics of the data and can detect sample points that are abnormal in the local environment but not obvious globally, especially suitable for various complex fluctuations and abnormal phenomena existing in the seabed environment.

[0152] Singular value decomposition is a powerful matrix decomposition technique that realizes dimensionality reduction and denoising by extracting the main patterns of the data. In the processing of seabed monitoring data, SVD can effectively separate signals and noises and extract the key information reflecting the main characteristics of the seabed environment. The selection of singular values is based on the energy cumulative contribution rate, ensuring that the main information in the data is retained while redundant and noisy information is removed.

[0153] ​​​​​​The adaptive sliding window method dynamically adjusts the window width according to the characteristics of data changes. When the environment changes violently, a smaller window is used to capture subtle changes, and when the environment is relatively stable, a larger window is used to reduce redundancy. The innovation of this method lies in associating the window width with the rate of change of data variance, achieving an adaptive analysis of the dynamic changes in the seabed environment.

[0154] Sample entropy and Lyapunov exponent are important indicators for quantifying the complexity and chaos degree of time series. The larger the sample entropy, the higher the complexity and irregularity of the time series; a positive Lyapunov exponent indicates the existence of chaotic characteristics in the system, and the larger the value, the stronger the sensitivity of the system to the initial conditions. The combination of these two indicators can comprehensively describe the turbulent and disordered states in the seabed environment.

[0155] The exponentially weighted moving average method can smooth short-term fluctuations while reflecting long-term trends when calculating the drift amount. The selection of the smoothing factor balances the response speed to new data and the smoothing effect. This method is particularly suitable for dealing with the slow drift phenomenon existing in seabed monitoring data.

[0156] Principal component analysis realizes data dimensionality reduction and feature extraction by finding the main variation directions of data. When dealing with the drift matrix, PCA can identify the main drift patterns, reducing the data dimension while retaining key information. The selection of principal components is based on the proportion of explained variance, ensuring the capture of the main variations in the data.

[0157] Tucker decomposition is a high-order tensor decomposition technique that can handle multi-dimensional data and retain high-order correlations. In the seabed environmental state modeling, Tucker decomposition can consider the correlations in three dimensions of time, sensor type, and feature type simultaneously, achieving the fusion analysis of multi-source data. The dimension selection of the core tensor is based on the proportion of the original dimensions, balancing the model complexity and the expression ability.

[0158] The residual neural network adopts a loss function combining MSE and L1 regularization in the drift correction model. The MSE term ensures the prediction accuracy of the model, and the L1 regularization term promotes the sparsity and generalization ability of the model. The cosine annealing learning rate adjustment strategy can use a larger learning rate in the initial stage of training to quickly approach the optimal solution, and use a smaller learning rate in the later stage of training for fine-tuning, improving the training efficiency and performance of the model.

[0159] Mahalanobis distance takes into account the covariance structure of data when calculating sample similarity, can handle the correlation between features, and is more suitable for dealing with multivariate data compared to Euclidean distance. Ward's minimum variance method pursues a merging strategy with the minimum within-class variance in hierarchical clustering, which is conducive to forming a compact clustering structure. Silhouette coefficient and Davies-Bouldin index, as clustering evaluation metrics, evaluate the clustering quality from different perspectives, comprehensively considering within-class compactness and between-class separability, and providing an objective basis for determining the optimal number of clusters.

[0160] Specifically, the principle of the present invention is as follows: The technical principle of the present invention is based on a multi-level data processing and model fusion strategy. By combining feature extraction, multi-dimensional analysis, and deep learning of seabed monitoring data, it realizes the effective differentiation of environmental changes and equipment drift. First, the present invention uses the singular value decomposition method to perform dimensionality reduction on the original data. By retaining the eigenvectors with larger singular values, it effectively extracts the main features of the data, while eliminating noise and redundant information, laying a foundation for subsequent processing.

[0161] Secondly, the present invention innovatively proposes the concepts of disorder matrix and drift matrix, and realizes efficient calculation through CUDA parallel processing technology. The disorder matrix quantitatively describes the nonlinear dynamic processes such as turbulence and perturbation in the seabed environment, reflecting the natural change characteristics of the environment; while the drift matrix tracks and analyzes the change trend of sensor performance by comparing the deviation degree between the real-time readings of the sensor and the historical reference values. These two matrices describe the data characteristics from different dimensions, providing a theoretical basis for differentiating environmental changes and equipment drift.

[0162] The core of the present invention is to adopt tensor decomposition technology to fuse the disorder pattern feature vectors, drift principal components, and abnormal feature matrices, and construct a seabed environmental state tensor model. Compared with traditional matrix analysis methods, tensor decomposition can retain the high-order correlation between data and capture the structural features in multi-dimensional data more comprehensively. At the same time, the deep-sea environmental perception attention network model designed by the present invention integrates the Transformer architecture and the recurrent neural network structure. Through the multi-scale drift perception attention mechanism, it can simultaneously focus on short-term fluctuations and long-term drift patterns, and effectively learn and compensate for complex sensor drift.

[0163] By dynamically adjusting the analysis window length and frequency resolution through an adaptive spectrum optimization function, the present invention can accurately capture the data characteristics under different seabed environments; the pre-trained deep-sea environmental perception attention network differentiates environmental changes and sensor drift through contrastive learning, and realizes efficient processing under limited computing resources. This multi-level and multi-dimensional data processing method theoretically ensures the effective differentiation of environmental changes and equipment drift, and solves the problem of the accuracy of long-term monitoring data.

[0164] A specific embodiment 1 of the present invention is provided below. The specific implementation manners of each step in this embodiment 1 are described in detail as follows.

[0165] The specific implementation manner of step S01 is to normalize the original data collected by the seabed-based monitoring device, and uniformly convert the data with different dimensions into the interval [0, 1]. Specifically, the min-max normalization method is adopted. For each sensor data sequence x, the normalized value is calculated as follows:

[0166] ;

[0167] In the formula, is the normalized data value; is the original data value; is the minimum value of this sensor data sequence; is the maximum value of this sensor data sequence.

[0168] Then, the improved local outlier factor algorithm is used to detect outliers. This algorithm identifies outliers by calculating the ratio of the local density of a sample point to that of its k nearest neighbor sample points. The calculation formula is:

[0169] ;

[0170] In the formula, is the local outlier factor of point ; is the nearest neighbor point set of point ; is the local reachability density of point ; is the number of nearest neighbor points.

[0171] The calculation formula of the local reachability density is:

[0172] ;

[0173] In the formula, is the reachability distance from point to point , which is defined as:

[0174] ;

[0175] In the formula, is the distance from point to its th nearest neighbor point; is the distance from point and point the Euclidean distance between.

[0176] The k value is adaptively selected according to the scale of the dataset, generally taking 5% - 10% of the total number of samples. The outlier determination threshold is set to 2.5, that is, when the local outlier factor of a certain data point is greater than 2.5, it is determined as an outlier and removed. The purpose of this step is to eliminate noise and outliers in the data, ensure the quality of the data for subsequent analysis and processing, and improve the reliability and accuracy of data processing.

[0177] The specific implementation of step S02 is to organize the effective dataset obtained in step S01 into a matrix form X, where the rows represent time points and the columns represent the measurement values of different sensors. Then, apply singular value decomposition to matrix X:

[0178] ;

[0179] In the formula, is the original data matrix, is the number of time points, is the number of sensors; and are orthogonal matrices respectively; is a diagonal matrix, and the elements on the diagonal are singular values.

[0180] Sort according to the magnitudes of the singular values, and select the first r singular values and the corresponding eigenvectors whose cumulative contribution rate reaches 95% to form the reduced - dimensional feature matrix:

[0181] ;

[0182] In the formula, contains the first columns of ; contains the diagonal matrix composed of the first singular values of ; contains the first columns of . The selection of

[0183] ;

[0184] In the formula, is the th singular value. In practice, the r value is usually selected such that the energy of the retained singular values accounts for more than 95% of the total energy. For typical seabed monitoring data, the r value is generally between 10 and 20. The purpose of this step is to reduce the data dimension, eliminate redundant information and noise in the data, extract the key information reflecting the main characteristics of the seabed environment, and provide a more compact and effective data representation for subsequent processing.

[0185] The specific implementation of step S03 is for the feature matrix generated in step S02 , and an adaptive sliding window method is used for time-domain segmentation. First, the basic window width is set to 1 hour, and an adaptive adjustment coefficient α is set according to the characteristics of seabed environment changes. This coefficient depends on the environmental change rate and data acquisition frequency, and generally ranges from 0.5 to 2.0. The variance change rate of real-time monitoring data is monitored, and the adjustment formula for the window width is:

[0186] ;

[0187] In the formula, is the actual window width; is the basic window width, set to 1 hour; is the adjustment coefficient, with a value range of 0.5 to 2.0; is the data variance change rate, and the calculation formula is:

[0188] ;

[0189] In the formula, is the data variance at the current moment; is the data variance at the previous moment. When the variance change rate exceeds the preset threshold (usually 20%), the window width is adjusted according to the above formula to ensure that the window width is reduced when the environment changes violently and expanded when the environment is relatively stable. The overlap rate between adjacent windows is set to 50% to ensure data continuity and time correlation. Through this adaptive window segmentation method, the feature matrix is divided into multiple groups of time series data blocks, and each data block contains all sensor feature data within a certain period of time. The purpose of this step is to adapt to the dynamic change characteristics of the seabed environment, ensure that subtle change information can be captured when the environment changes violently, and reduce data redundancy and improve processing efficiency when the environment is relatively stable.

[0190] The specific implementation of step S04 is to start the CUDA parallel processing stream and allocate computing resources to process the multiple groups of time series data blocks generated in step S03. For each data block, disorder indexes are calculated, including entropy value, Lyapunov exponent, complexity, etc. The sample entropy algorithm is specifically used to calculate the complexity and irregularity of the time series:

[0191] ;

[0192] In the formula, is the sample entropy value; is the embedding dimension, with a value of 2; is the similarity tolerance, with a value of 0.2 times the data standard deviation; is the data length; is at dimension The number of matching pattern pairs below; For the dimension The number of matching pattern pairs below.

[0193] At the same time, calculate the Lyapunov exponent to evaluate the chaos degree of the sequence. First, perform phase space reconstruction:

[0194] ;

[0195] In the formula, Is the reconstructed state vector; Is the th point of the time series; Is the reconstruction delay, and the value is 1 / 4 of the period; Is the reconstruction dimension, determined by the false nearest neighbor method.

[0196] Then calculate the Lyapunov exponent:

[0197] ;

[0198] In the formula, Is the Lyapunov exponent; Is the th time point; Is The distance between adjacent orbits in the phase space at the moment.

[0199] Organize these disorder indexes into a disorder matrix M. Each element In the matrix represents the disorder index of the jth sensor within the ith time window. In parallel, calculate the deviation degree of each sensor reading relative to the historical reference value, using the exponentially weighted moving average method:

[0200] ;

[0201] In the formula, Is The smoothed value at the moment; Is The observed value at the moment; Is the smoothing factor, and the value is 0.1; Is The smoothed value at the moment.

[0202] The drift amount is defined as:

[0203] ;

[0204] In the formula, Is The drift amount at the moment; Is The observed value at the moment; Is Smoothing value of the moment.

[0205] Construct a drift matrix D, where each element in the matrix represents the drift amount of the j-th sensor in the i-th time window. The threshold is set to 3 times the standard deviation. This step uses the CUDA parallel computing architecture to distribute large-scale matrix operations to thousands of computing cores for simultaneous execution, significantly improving the processing efficiency. The purpose is to quantitatively describe the turbulence, perturbations, and sensor drifts in the seabed environment and provide a basis for subsequent analysis.

[0206] The specific implementation of step S05 is to perform singular value decomposition on the disorder matrix M constructed in step S04 in the CUDA parallel processing stream:

[0207] ;

[0208] In the formula, is the disorder matrix, is the number of time windows, is the number of sensors; and are orthogonal matrices respectively; is a diagonal matrix, and the elements on the diagonal are singular values.

[0209] Analyze the singular value distribution, and select the right singular vectors corresponding to the first k largest singular values as the disorder pattern feature vectors. The value of k is selected to satisfy:

[0210] ;

[0211] In the formula, is the -th singular value; is the number of selected feature vectors.

[0212] At the same time, perform principal component analysis on the drift matrix D, and calculate the covariance matrix:

[0213] ;

[0214] In the formula, is the covariance matrix; is the drift matrix, is the number of samples, is the number of sensors.

[0215] Solve the characteristic equation to obtain the eigenvalues and the eigenvectors , and select the first p principal components with an explained variance of 85% as the drift principal components. The selection satisfies:

[0216] ;

[0217] In the formula, is the covariance matrix of the th eigenvalue; is the number of selected principal components.

[0218] This step is executed in parallel on the GPU, making full use of the high-performance computing capabilities of the CUDA architecture. The purpose is to extract the main patterns and features from the disorder matrix and the drift matrix, reduce the data dimension, and extract the key information reflecting the disorder state of the seabed environment and the sensor drift trend.

[0219] The specific implementation of step S06 is to extract the features of the abnormal data identified in step S01 in the CPU processing stream. First, group the abnormal data according to the time window, and calculate the statistical features of the abnormal data in each window, including frequency, amplitude, spatial distribution, and time distribution, etc. Use the kernel density estimation method to analyze the distribution characteristics of the abnormal data:

[0220] ;

[0221] In the formula, is the kernel density estimate value; is the number of samples; is the bandwidth parameter, with a value of 0.1; is the kernel function, usually the Gaussian kernel, expressed as:

[0222] ;

[0223] In the formula, is the standardized distance.

[0224] Then extract the abnormal pattern features, and apply the random forest algorithm to identify the main types and features of the abnormal data. The Gini impurity calculation formula is:

[0225] ;

[0226] In the formula, is the Gini impurity of the dataset ; is the number of classes; is the th class sample proportion in the dataset.

[0227] The information gain calculation formula is:

[0228] ;

[0229] In the formula, is the information gain of the feature ; is a feature The number of values; is a feature The value is The sample subset; and are the number of samples in the subset and the total dataset respectively.

[0230] The number of trees in the random forest is set to 100, and the minimum number of samples in a leaf node is set to 5. The extracted abnormal features are organized into an abnormal feature matrix A, and A is aligned with the disorder pattern feature vector and the drift principal component extracted in step S05 in the time dimension through timestamp alignment to generate a unified time index, ensuring the temporal consistency of the three types of feature data. The purpose of this step is to extract the environmental information contained in the abnormal data and align it with the disorder pattern and the drift principal component in the time dimension, providing a basis for subsequent data fusion.

[0231] The specific implementation of step S07 is to fuse the disorder pattern feature vector, the drift principal component, and the abnormal feature matrix obtained in steps S05 and S06 using tensor decomposition technology. First, a third-order tensor X is constructed, and its three dimensions respectively correspond to time, sensor type, and feature type (disorder, drift, abnormal). Then, the Tucker decomposition is applied to decompose the tensor X into the product of a core tensor G and three factor matrices A, B, and C:

[0232] ;

[0233] In the formula, is the original third-order tensor, is the size of the time dimension, is the size of the sensor type dimension, is the size of the feature type dimension; is the core tensor, , , ; , , are the factor matrices; denotes the tensor-matrix product along the dimension.

[0234] The formula for selecting the dimensions of the core tensor is:

[0235] ;

[0236] ;

[0237] ;

[0238] In the formula, Denotes rounding up; , , is an adjustment parameter, and its value range is 0 to 0.2 times the original dimension size. By analyzing the element sizes of the core tensor and their corresponding factor vectors, the coupling relationships between different features are identified, and a seabed environmental state tensor model is constructed. This model can express the complex associations among environmental factors, sensor characteristics, and data anomalies, providing a theoretical basis for multi-dimensional information fusion in seabed state assessment. The purpose of this step is to discover the hidden multi-dimensional association patterns in the data through high-order tensor decomposition and construct a tensor model that comprehensively reflects the seabed environmental state.

[0239] The specific implementation of step S08 is to construct a data drift correction model based on a residual neural network. The network structure includes 1 input layer, 4 residual blocks, and 1 output layer. Each residual block consists of 2 convolutional layers and 1 shortcut connection. The input layer receives the drift principal components extracted in step S05, and the output layer generates drift correction coefficients. The network training uses the Adam optimizer, and the initial learning rate is set to 0.001. The cosine annealing strategy is adopted for adjustment:

[0240] ;

[0241] In the formula, is the learning rate of the th round; is the minimum learning rate, with a value of 0.00001; is the maximum learning rate, with a value of 0.001; is the total number of rounds; is the current round.

[0242] The loss function adopts a combination of mean squared error and L1 regularization:

[0243] ;

[0244] In the formula, is the total loss; is the mean squared error, and the calculation formula is:

[0245] ;

[0246] In the formula, is the number of samples; is the true value; is the predicted value.

[0247] is the L1 regularization term, and the calculation formula is:

[0248] ;

[0249] Wherein, is the number of model parameters; is the th model parameter; is the regularization coefficient, with a value of 0.0001.

[0250] The training data selects the data segments with known drift patterns in the historical monitoring data, and divides the training set and the validation set according to the ratio of 8:2. During the training process, when the validation loss does not decrease for 5 consecutive rounds, the early stopping mechanism is triggered. The trained model is applied to the real-time monitoring data, and the correction coefficient is generated according to the identified drift pattern to correct and compensate the data. The purpose of this step is to learn the long-term drift law in the seabed monitoring data, establish an effective drift correction model, and improve the accuracy and reliability of the long-term monitoring data.

[0251] The specific implementation of step S09 is to classify the data processed in the previous step by using the hierarchical clustering algorithm. First, define the distance metric between data points, and calculate the similarity between samples using the Mahalanobis distance:

[0252] ;

[0253] Wherein, is the Mahalanobis distance between samples and ; is the covariance matrix of the data.

[0254] Then, the Ward minimum variance method is used as the clustering criterion. This method selects the two clusters with the smallest increase in within-class variance for merging in each step:

[0255] ;

[0256] Wherein, is the variance increment after the merger of cluster and cluster ; and are the number of samples in cluster and cluster respectively; and are the mean vectors of cluster and cluster respectively; represents the Euclidean norm.

[0257] The clustering process starts from individual samples, gradually merges the most similar clusters, and finally forms a hierarchical structure. The quality of different clustering results is evaluated by calculating the silhouette coefficient and the Davies-Bouldin index to determine the optimal number of clusters. The formula for the silhouette coefficient is:

[0258] ;

[0259] In the formula, is the silhouette coefficient of the sample ; is the average distance between the sample and other samples in the same cluster; is the average distance between the sample and the nearest non - same - cluster.

[0260] The calculation formula of the Davies - Bouldin index is:

[0261] ;

[0262] In the formula, is the Davies - Bouldin index; is the number of clusters; is the average distance from the samples of the cluster to the cluster center; is the distance between the center of the cluster and the center of the cluster ;

[0263] Generally, the optimal number of clusters is between 4 and 8. According to the knowledge of physical oceanography and combined with expert experience, physical meanings are assigned to each cluster, an evaluation standard for seabed conditions is established, and it is divided into four levels: normal, slightly abnormal, moderately abnormal, and severely abnormal. Finally, according to the clustering results of real - time monitoring data, a seabed condition evaluation report is generated. The purpose of this step is to scientifically classify the processed data, establish an objective seabed condition evaluation standard, and provide decision - making support for seabed environmental monitoring and early warning.

[0264] Optionally, the seabed - based monitoring equipment adopts a combination of a multi - source sensor array system, a high - performance data - processing module, and an efficient energy - supply system to achieve all - around monitoring and data processing of the seabed environment. The data - processing and analysis methods in the above steps fully consider the characteristics of the seabed environment and the characteristics of monitoring data, and construct a complete set of seabed environmental monitoring data - processing methods through links such as normalization processing, anomaly detection, feature extraction, multi - dimensional data fusion, drift correction, and state evaluation. This method can effectively process multi - source heterogeneous data collected by seabed - based monitoring equipment, identify environmental changes and equipment drift, provide accurate and reliable seabed condition evaluation results, and provide data support and decision - making basis for marine scientific research and seabed resource development.

[0265] Furthermore, as Figure 2As shown, the seabed-based monitoring equipment in Example 1 employs a modular design. Its main structure consists of a titanium alloy shell, a flat cylindrical shape with a diameter of 75 cm, a height of 30 cm, and a total weight of approximately 120 kg. The shell utilizes a multi-layer composite structure, with an inner layer of electromagnetic shielding material, a middle layer of high-strength thermal insulation material, and an outer layer of corrosion-resistant titanium alloy. The equipment can withstand high-pressure environments at depths of 6,000 meters. The bottom is designed with an adjustable foot system, which is hydraulically controlled to ensure the equipment remains horizontal and stable on uneven seabeds. The sensor array system is evenly distributed around the device and primarily includes eight high-precision pressure sensors, 12 distributed temperature sensors, four three-dimensional acoustic Doppler current sensors, six triaxial seismic sensors, and a variety of chemical detectors, covering an area with a radius of 20 meters around the device. The sensors utilize a modular interface design, allowing for flexible configuration and replacement based on the characteristics of different marine environments.

[0266] The data acquisition unit, located in the center of the device, utilizes a 32-bit high-performance microcontroller as its core and is equipped with a high-precision analog-to-digital converter. The sampling rate can be dynamically adjusted according to monitoring needs, supporting up to 2,000 samples per second. It also has 256GB of solid-state storage capacity, capable of continuously storing 3-6 months of complete monitoring data. The signal processing module integrates a low-power FPGA and two DSP chips, forming a three-level processing architecture. The front-end FPGA is responsible for signal filtering and data preprocessing, the mid-end DSP performs complex algorithm calculations, and the back-end embedded computing unit is responsible for data management and decision analysis. This architectural design enables the device to process large amounts of sensor data in real time in a seabed environment, executing the complex algorithmic processes from S01 to S09. In particular, it can complete large-scale matrix calculation tasks with CUDA acceleration, enabling accurate assessment and prediction of the seabed environmental status.

[0267] The energy supply system utilizes an innovative hybrid energy design, primarily a high-energy-density lithium-ion battery pack with a total capacity of 600Wh. It is also equipped with a seawater temperature difference energy harvester and a micro-water flow generator. Under normal marine conditions, this system can provide approximately 15-20W of continuous energy, extending the device's operating life to 1-2 years. To cope with extreme conditions, the device is also equipped with an intelligent power management system that automatically adjusts the sampling frequency and data processing depth based on environmental conditions and power levels. In times of energy shortage, it switches to a low-power mode, retaining only core monitoring functions.

[0268] The communication transmission module includes two independent systems: one is a high-speed acoustic communication system with a working frequency range of 18 - 22 kHz, a data transmission rate of up to 10 kbps, and an effective communication distance of about 3 kilometers; the other is an optical and electrical composite cable system used to establish a high-speed data link with a nearby relay station or scientific research platform, with a transmission rate of up to 100 Mbps. The two systems work together to ensure the reliability of data transmission under various sea conditions. The transmitted content mainly includes the processed environmental status assessment results, key feature indicators, and system self-diagnosis information. For the raw data that needs in-depth analysis, it can be uploaded in batches after receiving the instruction.

[0269] The software system adopts a hierarchical architecture. The bottom layer is a real-time operating system that provides precise time synchronization and task scheduling; the middle layer is a data processing engine that implements all the algorithm processes from S01 to S09; the top layer is an intelligent decision-making system responsible for monitoring data anomaly pattern recognition, environmental status assessment, and warning generation.

[0270] To better understand and implement the present invention, an embodiment 2 of a specific application scenario of the present invention is provided below: When conducting deep-sea environmental research in the northern slope area of the South China Sea, researchers deployed a set of seabed-based monitoring equipment, which is located at a water depth of about 1200 meters and equipped with various sensors such as pressure sensors, temperature sensors, flow velocity sensors, seismic sensors, and chemical substance detectors. The monitoring equipment continuously works for 18 months and collects data at a frequency of once every 30 minutes, aiming to observe seabed environmental changes, cold seep activities, and potential geological disasters. Due to the obvious influence of the monsoon in the South China Sea and the existence of complex water mass exchange and internal wave activities, the monitoring data shows obvious seasonal changes and random perturbation characteristics. At the same time, under long-term working conditions, the sensors have shown varying degrees of drift, affecting the accuracy of the data. The researchers used the data processing method of the present invention to systematically process and analyze the collected seabed monitoring data.

[0271] First, the raw data collected by the seabed-based monitoring equipment is normalized and anomaly detected. The continuous monitoring for 18 months has generated approximately 26,280 data points at time points (18 months × 30 days × 24 hours × 2 times / hour), and each time point contains the readings of 5 sensors, totaling 131,400 data points. The minimum-maximum normalization method is applied to unify all sensor data into the [0, 1] interval, and then the local outlier factor algorithm is used for outlier detection. The k value is set to 6% of the total number of samples (about 7884 neighboring points), and the anomaly determination threshold is 2.5. As shown in Table 1, the statistical results of anomaly detection for different sensors are as follows:

[0272] Table 1 Statistical table of anomaly data detection results for each sensor in the South China Sea

[0273]

[0274] Anomaly detection found that the anomaly ratio of the flow velocity sensor was the highest, reaching 5.70%. This was mainly due to the frequent activities of internal waves in the South China Sea, which generated a large number of short-term flow velocity peaks when the internal waves passed by. The anomaly rate of the chemical substance detector ranked second, reaching 3.50%, which was related to the intermittent activities of cold seeps releasing chemical substances such as methane during the observation period.

[0275] Next, the researchers applied singular value decomposition to the effective data set for dimensionality reduction. The effective data was organized into a matrix form of 24782×5 (taking the effective time point data common to all sensors), and the singular value results shown in Table 2 were obtained through SVD decomposition:

[0276] Table 2 Results of Singular Value Decomposition of Data in the South China Sea

[0277]

[0278] According to the principle that the cumulative energy proportion reaches 95%, the first 3 singular values and the corresponding eigenvectors were selected to construct the dimensionality-reduced feature matrix, reducing the data dimension from 5 dimensions to 3 dimensions, reducing 40% of the data volume while retaining the main information.

[0279] In the third step, an adaptive sliding time-domain window segmentation was implemented for the dimensionality-reduced feature matrix. The initial basic window width was set to 2 hours, and the adaptive adjustment coefficient α was set to 1.2. According to the characteristics of the seabed environment in the South China Sea, special attention was paid to the semi-diurnal cycle changes caused by internal tides and internal waves, and at the same time, the seasonal changes caused by monsoons were taken into account. By monitoring the change rate of data variance, when the change rate exceeded the threshold of 25%, the window width was dynamically adjusted. As shown in Table 3, the adjustment of the window width in different environmental change stages:

[0280] Table 3 Adjustment of Adaptive Window Width in the South China Sea

[0281]

[0282] The window segmentation results showed that during the prevailing period of the summer monsoon in the South China Sea, due to the frequent typhoon activities and the enhancement of internal waves, the environmental changes were drastic, and the data variance change rate was as high as 58.3%. The window width was automatically adjusted to the minimum value of 0.60 hours, generating the largest number of data blocks; while during the transitional periods of spring and autumn, the environment was relatively stable, and the window width was expanded to about 2.5 hours, effectively reducing data redundancy.

[0283] Step 4: Start the CUDA parallel processing stream to receive time series data blocks, and calculate the disorder index and drift amount within each data block. Adopting the CUDA 11.4 parallel computing architecture on the NVIDIA RTX 3090 GPU, 18,830 data blocks are allocated to 10,496 CUDA cores for simultaneous processing. The sample entropy (embedding dimension m = 2, similarity tolerance r = 0.2×σ) and Lyapunov exponent (reconstruction delay τ = 8 hours, reconstruction dimension d = 4) are calculated to construct a disorder matrix. Meanwhile, the exponential weighted moving average method (smoothing factor α = 0.1) is used to calculate the drift amount of each sensor, and a drift matrix is constructed. Table 4 shows the average disorder index in different seasons:

[0284] Table 4 Statistical Table of Disorder Index in Different Seasons in the South China Sea

[0285]

[0286] Step 5: Perform SVD decomposition on the disorder matrix in the CUDA parallel processing stream to extract the disorder pattern eigenvectors; perform principal component analysis on the drift matrix to extract the drift principal components. The analysis results are shown in Table 5:

[0287] Table 5 Analysis Results Table of Disorder Patterns and Drift Principal Components in the South China Sea

[0288]

[0289] The results show that the disorder pattern can be described by 3 eigenvectors, with the cumulative explained variance reaching 91.28%; while the drift characteristics can be represented by 2 principal components, with the cumulative explained variance reaching 88.46%. Parallel computing greatly shortens the processing time, and all calculations are completed in only 8.9 seconds.

[0290] Step 6: Extract the abnormal data features in the CPU processing stream. Analyze the 4,336 abnormal data points identified in Step 1, adopt the kernel density estimation method (bandwidth parameter h = 0.1) to analyze the distribution characteristics of the abnormal data, and use the random forest algorithm (number of trees = 100, minimum number of samples in leaf nodes = 5) to identify the main types and characteristics of the abnormal data. The extracted abnormal feature types and their distributions are shown in Table 6:

[0291] Table 6 Analysis Results Table of Abnormal Data Features in the South China Sea

[0292]

[0293] Step 7: Use tensor decomposition technology to fuse the chaotic pattern feature vectors, drift principal components, and anomaly feature matrices. Construct a third-order tensor with dimensions [18830×5×3], and apply Tucker decomposition to decompose the tensor into the product of a core tensor and three factor matrices. According to the formula, the dimension of the core tensor is calculated to be [5649×2×1], 𝛥p is set to 0.1×I, and both 𝛥k and 𝛥ᵣ are set to 0. By analyzing the element sizes of the core tensor and their corresponding factor vectors, the coupling relationships between different features in the South China Sea area are identified, especially the correlation between internal wave activities and chemical substance releases, and a complete seabed environmental state tensor model is constructed.

[0294] Step 8: Build a data drift correction model based on a residual neural network. The network structure includes 1 input layer, 4 residual blocks, and 1 output layer. The network is trained using the Adam optimizer with an initial learning rate of 0.001, and the cosine annealing strategy is adopted for adjustment. The loss function uses a combination of mean squared error and L1 regularization, and the regularization coefficient λ = 0.0001. Select the known drift pattern data in the historical data for training, and divide the training set and validation set in a ratio of 8:2. After 350 rounds of training, the model converges, and the drift correction effects of each sensor are shown in Table 7:

[0295] Table 7 Drift Correction Effects of Each Sensor in the South China Sea Area

[0296]

[0297] Step 9: Use the hierarchical clustering algorithm to classify the processed data. The Mahalanobis distance is used to calculate the similarity between samples, and the Ward minimum variance method is adopted as the clustering criterion. The quality of different clustering results is evaluated by calculating the silhouette coefficient and Davies-Bouldin index, and the optimal number of clusters is determined to be 5. According to the oceanographic characteristics of the South China Sea, physical meanings are assigned to each cluster, and a seabed state evaluation standard is established. The clustering results are shown in Table 8:

[0298] Table 8 Clustering Results of Seabed State Evaluation in the South China Sea Area

[0299]

[0300] Traditional methods for processing South China Sea seabed monitoring data mainly rely on fixed threshold filtering and simple statistical analysis, making it difficult to cope with the complex and changeable environmental characteristics of the South China Sea area. Especially when dealing with long-term time series data, it is unable to effectively identify and correct sensor drift, resulting in a gradual decline in data quality; at the same time, traditional methods lack the ability to fuse multi-source heterogeneous data and are difficult to discover the complex relationships between different environmental factors; in addition, the computational efficiency is low, making it difficult to support real-time analysis of large-scale data, and it usually takes several days to complete all data processing.

[0301] The data processing method of the present invention shows significant advantages in the application in the South China Sea compared with traditional methods: First, through the adaptive window segmentation technology, it can dynamically adjust the analysis window according to the characteristics such as the alternation of the South China Sea monsoon and internal wave activities, improving the ability to capture short-term environmental changes; Second, the CUDA parallel processing technology reduces the data processing time from 72 hours of traditional methods to 3.5 hours, with the efficiency increased by about 20 times; Third, the tensor decomposition technology realizes the extraction of the correlation mode of internal wave-cold seep activities unique to the South China Sea, revealing the coupling relationship of environmental factors that was difficult to discover in the past; Fourth, the drift correction model constructed by the residual neural network reduces the average drift rate of the sensor from 3.84% to 0.67%, and effectively corrects the large drift (6.83%) of the chemical substance detector in particular; Fifth, the state evaluation method based on hierarchical clustering identifies five typical seabed states in the South China Sea, providing a scientific basis for the exploration of South China Sea resources and environmental monitoring. Compared with traditional methods, the present invention realizes high-precision, high-efficiency and high-integration of the processing of South China Sea seabed monitoring data, providing strong data support for the deep-sea research in the South China Sea.

[0302] It should be noted that the detailed explanations of the variables involved in the present invention are shown in Tables 9, 10, and 11 below.

[0303] Table 9 Variable Explanation Table (Part 1)

[0304]

[0305] Table 10 Variable Explanation Table (Part 2)

[0306]

[0307] Table 11 Variable Explanation Table (Part 3)

[0308]

[0309] As described above, it is only the specific implementation manner of the present invention, but the protection scope of the present invention is not limited thereto. Any person skilled in the art within the technical scope disclosed by the present invention can easily think of changes or substitutions, which should all be covered within the protection scope of the present invention.

Claims

1. A data processing method for seabed-based monitoring equipment, characterized in that: include: S01. Perform normalization processing on the data collected by the seabed-based monitoring equipment, and use an outlier detection algorithm to identify and eliminate abnormal data that exceeds a predetermined threshold range to determine a valid data set; S02. Based on the valid data set, use the singular value decomposition method to reduce the dimensionality of the original data matrix, eliminate data redundancy, extract key features and generate a feature matrix; S03, performing sliding time domain window segmentation on the feature matrix, adaptively adjusting the window width according to the changing characteristics of the seabed environment, and forming multiple groups of time series data blocks; S04, starting the CUDA parallel processing flow to receive the time series data blocks, calculating the disorder index in each data block, and constructing a disorder matrix; at the same time, calculating the degree to which the sensor readings deviate from the baseline value and constructing a drift matrix; S05. Performing singular value decomposition on the disordered matrix in a CUDA parallel processing flow to extract disordered mode eigenvectors; performing principal component analysis on the drift matrix to extract drift principal components; S06. Extract abnormal data features in the CPU processing flow, generate an abnormal feature matrix, and align the abnormal feature matrix with the disorder pattern feature vector and the drift principal component; S07. Using tensor decomposition technology to fuse the turbulence mode feature vector, the drift principal component, and the abnormal feature matrix, construct a seabed environmental state tensor model, and extract the coupling relationship between environmental factors and monitoring data; S08. Constructing a data drift correction model based on a residual neural network, inputting the drift principal component, training the network to learn the drift law in a long time series, and correcting and compensating the real-time monitoring data; S09. Using a hierarchical clustering algorithm to classify the processed data, establishing a seabed state assessment standard, and generating a multi-level seabed state assessment result; the turbulence index includes an entropy value and a Lyapunov index.

2. The data processing method of seabed-based monitoring equipment according to claim 1, characterized in that: The singular value decomposition method specifically organizes the seabed monitoring data into a matrix form, and decomposes the data matrix into the product of three sub-matrices by calculating the eigenvalues and eigenvectors.

3. The data processing method of seabed-based monitoring equipment according to claim 2, characterized in that: The chaos matrix specifically refers to a characteristic matrix constructed by calculating the volatility, uncertainty and disorder of seabed monitoring data in the time and space dimensions. It is used to quantitatively describe turbulence, disturbances and nonlinear dynamic processes in the seabed environment.

4. The data processing method of seabed-based monitoring equipment according to claim 3, characterized in that: The drift matrix is constructed by comparing the deviation between the real-time sensor readings and the historical baseline values. Each element in the matrix represents the drift of the corresponding sensor at a corresponding point in time. It is used to track and analyze changes in sensor performance and long-term environmental evolution trends.

5. The data processing method of seabed-based monitoring equipment according to claim 4, characterized in that: In the process of constructing the disorder matrix, an adaptive spectrum optimization function is used to dynamically adjust the frequency characteristics of the monitoring data under different seabed environmental conditions. By dynamically adjusting the analysis window length and frequency resolution, accurate capture of data characteristics in different seabed environments can be achieved.

6. The data processing method of seabed-based monitoring equipment according to claim 5, characterized in that: A pre-trained deep-sea environment perception attention network model is used in the data drift correction process. This model combines the Transformer architecture with the recurrent neural network structure and introduces a multi-scale drift-aware attention mechanism designed for the characteristics of seabed data. It is used to identify and compensate for complex drift patterns in seabed monitoring data at different time scales.

7. The data processing method of seabed-based monitoring equipment according to claim 6, characterized in that: The specific structure of the deep-sea environment perception attention network model is a hybrid architecture that combines a multi-layer bidirectional recurrent neural network with a multi-head self-attention mechanism. The bottom layer uses a convolutional neural network to extract features from multi-source sensor data. The middle layer uses a six-layer Transformer encoder to extract temporal features and cross-sensor correlation features. The top layer uses a fully connected network with skip connections to estimate and compensate for data drift.

8. The data processing method of seabed-based monitoring equipment according to claim 7, characterized in that: The deep-sea environment perception attention network model introduces a sparse attention mechanism for seabed characteristic perception in the Transformer encoder. The sparse attention mechanism dynamically adjusts the attention range according to the physical characteristics of the seabed environment, and is used to simultaneously focus on short-term fluctuations and long-term drift patterns. The entire model adopts a hierarchical design, with processing modules designed separately for data characteristics under different depths and environmental conditions.

9. The data processing method of seabed-based monitoring equipment according to claim 8, characterized in that: The pre-training process of the deep-sea environment perception attention network model uses long-term monitoring data collected from the global seabed monitoring network to screen data segments containing complete environmental change cycles and sensor drift phenomena, and uses physical oceanographic models to generate synthetic data to enhance the coverage of the data set, forming a comprehensive training data set covering a variety of typical seabed environments such as shallow sea areas, deep sea plains, trench areas, and hydrothermal areas.

10. The data processing method of seabed-based monitoring equipment according to claim 9, characterized in that: The seabed-based monitoring equipment is composed of a sensor array system, a data acquisition unit, a signal processing module, an energy supply system, a communication transmission module and a protective shell. The sensor array system includes a pressure sensor, a temperature sensor, a flow rate sensor, a seismic sensor and a chemical substance detector.

Citation Information

Patent Citations

  • Determination method of marine ecological environment damage causal-relationships

    CN108492007A

  • High-resolution temperature field reconstruction system based on optimized sonic sensor array

    CN117571152A