Server optimization method, system and device based on AI big data and readable storage medium

By acquiring hardware, software, and network layer data and performing time-aligned processing, a server stress index and topology affinity matrix are generated, solving the problem of inaccurate resource assessment caused by single-dimensional monitoring and achieving accurate resource assessment and efficient scheduling.

CN120929276AActive Publication Date: 2025-11-11CHANGSHA SHAOGUANG SEMICONDUCTOR CO LTD
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
CN202511453207.9
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-10-13
Publication Date
2025-11-11
Estimated Expiration
2045-10-13

AI Technical Summary

Technical Problem

Existing technologies rely on single-dimensional monitoring, which leads to inaccurate resource assessment and makes it impossible to allocate resources effectively under complex loads.

Method used

By acquiring data from the hardware, software, and network layers, performing time alignment processing to generate a heterogeneous dataset, calculating the server stress index and topology affinity matrix, performing dual-channel processing, and executing resource scheduling instructions.

Benefits of technology

It achieves multi-dimensional feature fusion analysis, accurately assesses resources, reduces energy consumption, shortens response time, and improves resource utilization.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120929276A_ABST
    Figure CN120929276A_ABST
Patent Text Reader

Abstract

The invention discloses a server optimization method, system and device based on AI big data and a readable storage medium, and the method comprises the steps: obtaining hardware layer collection data, software layer collection data and network layer collection data, and carrying out the time alignment processing of the hardware layer collection data, the software layer collection data and the network layer collection data, so as to obtain a heterogeneous data set; the method comprises the steps of calculating a server pressure index based on a heterogeneous data set, generating a topology affinity matrix, carrying out dual-channel processing based on the server pressure index and the topology affinity matrix, carrying out decision fusion on a processed result, and finally executing a resource scheduling instruction. According to the method, the resources can be accurately evaluated through multi-dimensional feature fusion analysis, energy consumption can be reduced and response time can be shortened through a dual-channel decision mechanism and adaptive resource scheduling, and the resource utilization rate is effectively improved.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] This invention relates to the field of cloud computing server management, and in particular to a server optimization method, system, device, and readable storage medium based on AI big data. Background Technology

[0002] Server optimization refers to improving server performance and efficiency by adjusting various server configurations and parameters. Server optimization includes multiple aspects such as hardware optimization, operating system optimization, and security optimization.

[0003] Currently, server optimization mainly involves monitoring CPU / memory to assess resources. However, monitoring only one dimension can lead to inaccurate resource assessment, making it impossible to allocate resources under complex loads.

[0004] The above content is only used to help understand the technical solution of the present invention and does not represent an admission that the above content is prior art. Summary of the Invention

[0005] The main objective of this invention is to provide a server optimization method, system, device, and readable storage medium based on AI big data, in order to solve the problem that current single-dimensional monitoring leads to inaccurate resource assessment, thus making it impossible to allocate resources under complex loads.

[0006] To achieve the above objectives, the present invention provides a server optimization method based on AI big data, the server optimization method based on AI big data comprising: Acquire hardware layer data, software layer data, and network layer data. The hardware layer data includes CPU temperature data and PDU real-time power consumption data. The software layer data includes the container's storage occupancy rate and API gateway request response time. The network layer data includes flow table statistics and server communication latency matrix. The hardware layer data, the software layer data, and the network layer data are time-aligned to obtain a heterogeneous dataset, wherein the heterogeneous dataset is the dataset obtained after time alignment of the hardware layer data, the software layer data, and the network layer data. The server stress index is calculated based on the heterogeneous dataset, and a topological affinity matrix is ​​generated. Dual-channel processing is performed based on the server stress index and the topology affinity matrix, wherein short-term prediction is performed based on the server stress index and topology analysis is performed based on the topology affinity matrix. The results of short-term prediction channel processing and topology analysis channel processing are fused together for decision-making, and resource scheduling instructions are executed.

[0007] Further, the step of performing time alignment processing on the hardware layer data, the software layer data, and the network layer data to obtain a heterogeneous dataset includes: The reference time source is extracted from the CPU temperature data, wherein the original timestamp of the hardware layer sensor clock is selected as the reference time source. Determine whether the PDU is a high-precision PDU. If the PDU is a high-precision PDU, directly obtain the power consumption data timestamp. If the PDU is a normal PDU, apply dynamic delay compensation to obtain the power consumption data timestamp. Here, the high-precision PDU refers to a PDU device equipped with a dedicated clock chip, and the normal PDU refers to a PDU device that relies on a software clock. Match the reference time source with the power consumption data timestamp to unify the hardware layer time axis; The target time point values ​​are calculated using cubic spline interpolation on the data collected by the software layer. Network layer data is aggregated using an exponentially weighted moving average.

[0008] Furthermore, if the PDU is a regular PDU, the step of applying dynamic delay compensation to obtain the power consumption data timestamp includes: If the PDU is a regular PDU, then the corrected timestamp is calculated using a first preset formula, and the corrected PDU timestamp is used as the power consumption data timestamp, wherein the first preset formula is: , The corrected PDU timestamp As a smoothing factor, For the precise local time of the data acquisition server, The network time obtained by the PDU device via the NTP protocol. This is the original timestamp of the PDU.

[0009] Furthermore, the step of matching the reference time source with the power consumption data timestamp to unify the hardware layer time axis includes: Based on the reference time source, the time residual is calculated on the power consumption data timestamp according to the second preset formula, wherein the second preset formula is: , The time residual is K, where K is the number of reference time points, and K is greater than or equal to 2. This refers to the time point of the power consumption data of the j-th PDU after the timestamp of the j-th PDU has undergone dynamic delay compensation. This is the base timestamp of the i-th CPU; Based on the time residual, a unified timestamp is calculated using a preset time mapping function to obtain a unified timestamp hardware layer dataset. This hardware layer dataset includes the CPU temperature data and the real-time power consumption data of the PDU after the unified timestamp. The time mapping function is: , The unified timestamp after alignment For time residuals, For dynamic compensation coefficients, The timestamp of the PDU to be mapped The timestamp of the current PDU The most recent CPU benchmark time point, This represents the local time deviation between the current PDU time point and the nearest CPU time point.

[0010] Furthermore, the step of calculating the target time point value using cubic spline interpolation on the data collected by the software layer includes: Obtain the set of data points and the target time point from the data collected by the software layer; The target time point value is calculated using a cubic spline function, wherein the cubic spline function is: , The target time point value, the For constant terms, The coefficient of the linear term, The coefficient of the quadratic term, The coefficient of the cubic term, As the independent variable, This represents the k-th data acquisition time.

[0011] Furthermore, the step of aggregating network layer data using an exponentially weighted moving average includes: Define the time window centered on the hardware reference time; Calculate the exponential weight for each network data point within the time window; The weighted aggregate value is calculated based on the exponential weights to obtain the time-aligned network layer acquisition data.

[0012] Furthermore, the step of calculating the server stress index based on the heterogeneous dataset and generating the topological affinity matrix includes: The server stress index , To calculate the weighting coefficients, , For memory weighting coefficients, , These are the network weight coefficients. , The CPU utilization is represented by (memory / 100), and the bus bandwidth usage percentage is represented by (memory / 100). For the number of packets retransmitted over the network, + + =1; The topological affinity matrix ,|delay i -Delay j | represents the communication delay from server i to j.

[0013] Furthermore, to achieve the above objectives, the present invention also provides a server optimization system based on AI big data, comprising: The acquisition module is used to acquire hardware layer data, software layer data, and network layer data. The hardware layer data includes CPU temperature data and PDU real-time power consumption data. The software layer data includes the container's storage occupancy rate and API gateway request response time. The network layer data includes flow table statistics and server communication latency matrix. The time alignment processing module is used to perform time alignment processing on the data collected by the hardware layer, the data collected by the software layer, and the data collected by the network layer to obtain a heterogeneous dataset, wherein the heterogeneous dataset is the dataset obtained after time alignment processing of the data collected by the hardware layer, the data collected by the software layer, and the data collected by the network layer. The calculation module is used to calculate the server stress index based on heterogeneous datasets and generate a topological affinity matrix; A dual-channel processing module is used to perform dual-channel processing based on the server stress index and the topology affinity matrix, wherein the short-term prediction channel is performed based on the server stress index, and the topology analysis channel is performed based on the topology affinity matrix. The decision fusion module combines the results of the short-term prediction channel with the results of the topology analysis channel to make decisions and execute resource scheduling instructions.

[0014] In addition, to achieve the above objectives, the present invention also provides a server optimization device based on AI big data. The server optimization device based on AI big data includes: a memory, a processor, and an AI big data-based server optimization program stored in the memory and executable on the processor. When the AI ​​big data-based server optimization program is executed by the processor, it implements the steps of the server optimization method based on AI big data as described above.

[0015] In addition, to achieve the above objectives, the present invention also provides a readable storage medium storing a server optimization program based on AI big data, wherein when the server optimization program based on AI big data is executed by a processor, it implements the steps of the server optimization method based on AI big data as described above.

[0016] This application acquires data from the hardware layer, software layer, and network layer, then performs time alignment processing on these data to obtain a heterogeneous dataset. Based on this heterogeneous dataset, a server stress index is calculated, and a topology affinity matrix is ​​generated. Dual-channel processing is then performed on the server stress index and the topology affinity matrix. The results of the short-term prediction channel processing and the topology analysis channel processing are fused for decision-making, and finally, resource scheduling instructions are executed. This application can accurately assess resources through multi-dimensional feature fusion analysis, and its dual-channel decision-making mechanism and adaptive resource scheduling can reduce energy consumption, shorten response time, and effectively improve resource utilization. Attached Figure Description

[0017] Figure 1 This is a flowchart illustrating an embodiment of server optimization based on AI big data in this application; The realization of the objective, functional features and advantages of the present invention will be further explained in conjunction with the embodiments and with reference to the accompanying drawings. Detailed Implementation

[0018] To make the objectives, technical solutions, and advantages of this application clearer, the following detailed description is provided in conjunction with the accompanying drawings and embodiments. It should be understood that the specific embodiments described herein are merely illustrative and not intended to limit the scope of this application.

[0019] This invention further provides a server optimization method based on AI big data. (Refer to...) Figure 1 , Figure 1 This is a flowchart illustrating an embodiment of the server optimization method based on AI big data according to the present invention.

[0020] In this embodiment, the execution entity of the AI-based big data server optimization method is an AI-based big data server optimization system. This system includes an AI-based big data server optimization device or equipment, which can be a PC, PDA, or other terminal device. This invention acquires hardware-layer, software-layer, and network-layer data, then performs time-alignment processing on these data to obtain a heterogeneous dataset. Based on this heterogeneous dataset, a server stress index is calculated, and a topology affinity matrix is ​​generated. Dual-channel processing is then performed on the server stress index and the topology affinity matrix. The results of the short-term prediction channel processing and the topology analysis channel processing are fused for decision-making, and finally, resource scheduling instructions are executed. This application can accurately assess resources through multi-dimensional feature fusion analysis, and through a dual-channel decision-making mechanism and adaptive resource scheduling, it can reduce energy consumption, shorten response time, and effectively improve resource utilization.

[0021] The steps of this AI-based big data-driven server optimization method include: Step S10: Acquire hardware layer data, software layer data, and network layer data; In this embodiment, the hardware layer collects data including CPU (central processing unit) temperature data and PDU (Power Distribution Unit) real-time power consumption data; the software layer collects data including container memory occupancy and API gateway request response time; and the network layer collects data including flow table statistics and server communication latency matrix. CPU temperature data can be obtained via the IPMI (Intelligent Platform Management Interface) protocol, and PDU real-time power consumption data can be directly read. The Prometheus exporter captures the container memory occupancy. This container refers to an isolated environment created based on operating system-level virtualization technology, achieving hardware resource isolation through cgroups and process, network, and file system isolation through namespaces. Each container shares the host kernel but has its own independent user space. The AI ​​big data server optimization system also records the API gateway request response time. By parsing the flow table, flow table statistics, such as OpenFlow flow tables, can be obtained, and a server communication latency matrix (N×N array) can be constructed.

[0022] In the above data acquisition sampling frequencies, the hardware layer sampling frequency can be 1Hz, the software layer sampling frequency can be 0.2Hz, and the network layer sampling frequency can be 5Hz.

[0023] Step S20: Perform time alignment processing on the data collected from the hardware layer, the data collected from the software layer, and the data collected from the network layer to obtain a heterogeneous dataset; In this embodiment, the heterogeneous dataset is a dataset obtained by time-aligning data collected from the hardware layer, software layer, and network layer. The hardware layer data can be given a unified time axis, the software layer data can be interpolated using cubic spline interpolation, and the network layer data can be processed using a moving average.

[0024] Step S20 includes: Step S21: Extract the reference time source from the CPU temperature data, wherein the original timestamp of the hardware layer sensor clock is selected as the reference time source. Step S22: Determine whether the PDU is a high-precision PDU. If the PDU is a high-precision PDU, directly obtain the power consumption data timestamp. If the PDU is a normal PDU, apply dynamic delay compensation to obtain the power consumption data timestamp. In one embodiment, the method of unifying the time axis for data acquisition at the hardware layer is as follows: the hardware layer sensor clock signal (raw timestamp) is selected as the reference time source from the CPU temperature data, and its clock error is less than ±1ms; it is determined whether the PDU is a high-precision PDU. If the PDU is a high-precision PDU, the power consumption data timestamp is directly obtained. If the PDU is a normal PDU, dynamic delay compensation is applied to obtain the power consumption data timestamp. Here, a high-precision PDU refers to a PDU device equipped with a dedicated clock chip, and a normal PDU refers to a PDU device that relies on a software clock.

[0025] The aforementioned dynamic delay compensation is performed by calculating the corrected timestamp using a first preset formula, and then using the corrected PDU timestamp as the power consumption data timestamp. The first preset formula is: , The corrected PDU timestamp As a smoothing factor, For the precise local time of the data acquisition server, The network time obtained by the PDU device via the NTP protocol. This is the original timestamp of the PDU.

[0026] Step S23: Match the reference time source with the power consumption data timestamp to unify the hardware layer time axis; The reference time source extracted from the CPU temperature data is matched with the power consumption timestamp of the PDU (corrected timestamp) to ultimately unify the hardware layer timeline.

[0027] Specifically, firstly, based on the reference time source, the time residual is calculated on the power consumption data timestamp according to the second preset formula, whereby: , The time residual is K, where K is the number of reference time points, and K is greater than or equal to 2. This refers to the time point of the power consumption data of the j-th PDU after the timestamp of the j-th PDU has undergone dynamic delay compensation. This is the base inter-stamp of the i-th CPU; Then, based on the time residual, a unified timestamp is calculated using a preset time mapping function to obtain a unified timestamp hardware layer dataset. The hardware layer dataset includes the CPU temperature data after the unified timestamp and the real-time power consumption data of the PDU after the unified timestamp.

[0028] The time mapping function is: , For the aligned unified timestamp, For time residuals, For dynamic compensation coefficients, The timestamp of the PDU to be mapped The timestamp of the current PDU The most recent CPU benchmark time point, This represents the local time deviation between the current PDU time point and the nearest CPU time point.

[0029] Step S24: Calculate the target time point value using cubic spline interpolation on the data collected by the software layer; In one embodiment, the method for time alignment processing of data acquired by the software layer includes cubic spline interpolation, which is used to calculate the target time point value.

[0030] Specifically, first, the data point set and target time point from the software layer's collected data are obtained. Then, the value of the target time point is calculated using a cubic spline function. The data point set from the software layer's collected data is used to construct the cubic spline function as follows: , The value at the target time point. This is a constant term (starting index value). It is the coefficient of the linear term (the instantaneous rate of change at the starting point). The coefficient of the quadratic term (initial curvature intensity). The coefficient of the cubic term (rate of curvature change). The independent variable (the moment when interpolation needs to be calculated). This represents the k-th data acquisition time.

[0031] The data point set in the software layer data collection is , Let i be the time of data collection for the i-th data point. For a moment Monitoring metrics (such as container memory usage), and the target time point for data collection at the software layer. This is the hardware layer reference time. Construct the above cubic spline function over the interval.

[0032] To eliminate endpoint oscillations, natural spline boundary conditions are used: in, The time of the first node in the data sequence. The time of the last node in the data sequence. Let be the second derivative at the left endpoint, and be the curvature at the starting point, indicating that the function at... The unevenness of the surface; Let be the second derivative at the right endpoint, and be the curvature at the endpoint, indicating that the function at... The unevenness of the surface.

[0033] The specific steps for constructing and solving tridiagonal equations are as follows: First, calculate the step size: , This represents the time interval between data points.

[0034] Then construct a system of equations: , ;in, , for nodes The second derivative at point; , where is the weighting coefficient for the left interval; , where is the weighting coefficient of the right interval. , is the right-hand term.

[0035] Boundary condition injection: transforming natural boundary conditions into... Then, the chasing method is used to solve the problem: forward elimination: Back-substitution solution: Then, the interpolation coefficients are calculated: [The results are incomplete in the original text.] Then calculate the coefficients for each interval. in, , For monitoring values ​​at the interval endpoints, The interval length is... , is the second derivative of the endpoints of the interval.

[0036] Step S25: Aggregate the network layer data using an exponentially weighted moving average.

[0037] In one embodiment, the method for time-aligning network layer acquisition data is exponentially weighted moving average aggregation. First, a time window is defined with the hardware reference time as the center. Then, an exponential weight is calculated for each network data point within the time window. Based on the exponential weight, a weighted aggregation value is calculated to obtain the time-aligned network layer acquisition data.

[0038] Specifically, firstly, based on hardware reference time Construct a symmetrical time window centered on the time window. ; Then, dynamic adjustments are made, with the following rules: base value. =0.5ms, if there is no data in the window: expand to =1.0ms, if network jitter rate >30%, shrink to =0.3ms, where, The hardware reference time is obtained through the CPU temperature sensor clock. The radius of the time window.

[0039] Next, time distance calculation is performed, calculating the absolute time distance for each data point within the window. , For absolute time distance, The data is timestamped on the network; conditions are imposed as follows: If the constraints are not met, then that point is excluded.

[0040] Then, the exponential weights are calculated using the exponential decay function: , The attenuation coefficient is [0.5, 1.0]; a weighted truncation mechanism is used to remove far-end noise points with a contribution of <1%, increasing the effective signal ratio by 23%. ,in The weights are the processed weights.

[0041] Finally, weighted aggregation is performed to output the aligned network layer acquisition data. This includes normalizing and weighting the valid data points within the window. , This is the aggregated output, representing the aligned network metrics. For valid points, satisfying The number of points, These are network metric values, which are the original network data.

[0042] Step S30: Calculate the server stress index based on the heterogeneous dataset and generate the topological affinity matrix; In this embodiment, the server stress index , To calculate the weighting coefficients, , For memory weighting coefficients, , These are the network weight coefficients. (Memory / 100) represents CPU utilization. For bus bandwidth usage ratio, For the number of packets retransmitted over the network, + + =1; Topological affinity matrix ,|delay i -Delay j | represents the communication delay from server i to j.

[0043] Step S40: Perform dual-channel processing based on the server stress index and topology affinity matrix; In this embodiment, short-term prediction channel processing is performed based on the server stress index, and topology analysis channel processing is performed based on the topology affinity matrix.

[0044] The short-term forecast channel processing based on the server stress index specifically includes: using the server stress index as data input, and the server stress index sequence within the time window is as follows: The window size is 6, the span is 5 minutes, and LSTM is used for modeling. Temperature sensing gating is introduced in the above process (the weight of the forget gate is increased when the CPU is >80℃).

[0045] Employing an attention mechanism: Predicted output: For example, the load probability distribution for the next 30 seconds (e.g., [CPU:0.85, Mem:0.72, Net:0.15]).

[0046] Topology analysis channel processing based on the topology affinity matrix specifically includes: constructing the affinity matrix. ; Then, the weights are calculated: ,in, This is the network distance weight (default 0.5). Weight for service call frequency (default 0.3). The weight for resource complementarity (default 0.2). .

[0047] Step S50: The results of the short-term prediction channel processing and the results of the topology analysis channel processing are fused together for decision-making, and resource scheduling instructions are executed.

[0048] In this embodiment, the decision-making process is determined by fusion decision triggering conditions, which are as follows: Enable topology decision-making, where, The threshold is dynamic (default 0.15).

[0049] Resource scheduling is performed using resource scheduling rules, which include expansion decisions, migration decisions, and energy-saving decisions.

[0050] In the expansion decision, priority is given to "backup clusters" in the same community, and the next priority is neighboring communities with an affinity of >0.6.

[0051] In the migration decision, high-pressure nodes are migrated to low-pressure communities. The migration delay must be less than the interruption time × 0.2. For example, the migration delay must be < 5ms × 0.2 = 1ms.

[0052] In energy-saving decision-making, the shutdown order is as follows: servers within the same PDU power supply unit, then in descending order of affinity (retaining highly connected nodes), ensuring that at least two nodes are retained in each community, including nodes that can be shut down. ∩{non-boundary nodes}, where For server node i, Let be the pressure index of node i.

[0053] Furthermore, this invention also proposes a server optimization system based on AI big data, which includes: The acquisition module is used to acquire hardware layer data, software layer data, and network layer data. The hardware layer data includes CPU temperature data and PDU real-time power consumption data. The software layer data includes the container's storage occupancy rate and API gateway request response time. The network layer data includes flow table statistics and server communication latency matrix. The time alignment processing module is used to perform time alignment processing on the data collected by the hardware layer, the data collected by the software layer, and the data collected by the network layer to obtain a heterogeneous dataset, wherein the heterogeneous dataset is the dataset obtained after time alignment processing of the data collected by the hardware layer, the data collected by the software layer, and the data collected by the network layer. The calculation module is used to calculate the server stress index based on heterogeneous datasets and generate a topological affinity matrix; A dual-channel processing module is used to perform dual-channel processing based on the server stress index and the topology affinity matrix, wherein the short-term prediction channel is performed based on the server stress index, and the topology analysis channel is performed based on the topology affinity matrix. The decision fusion module combines the results of the short-term prediction channel with the results of the topology analysis channel to make decisions and execute resource scheduling instructions.

[0054] Furthermore, this embodiment of the invention also provides a server optimization device based on AI big data. The server optimization device based on AI big data includes: a memory, a processor, and an AI big data-based server optimization program stored in the memory and executable on the processor. When the AI ​​big data-based server optimization program is executed by the processor, it implements the steps of the above-mentioned AI big data-based server optimization method.

[0055] In addition, this embodiment also provides a readable storage medium storing a server optimization program based on AI big data. When the server optimization program based on AI big data is executed by the processor, it implements the above-mentioned steps of the server optimization method based on AI big data.

[0056] The processor and memory can be connected via a bus or other means.

[0057] Memory, as a non-transitory computer-readable storage medium, can be used to store non-transitory software programs and non-transitory computer-executable programs. Furthermore, memory may include high-speed random access memory, and may also include non-transitory memory, such as at least one disk storage device, flash memory device, or other non-transitory solid-state storage device. In some embodiments, memory may optionally include memory remotely located relative to the processor, and these remote memories can be connected to the processor via a network. Examples of such networks include, but are not limited to, the Internet, intranets, local area networks, mobile communication networks, and combinations thereof.

[0058] It should be noted that, in this document, the terms "comprising," "including," or any other variations thereof are intended to cover non-exclusive inclusion, such that a process, method, article, or system that comprises a list of elements includes not only those elements but also other elements not expressly listed, or elements inherent to such a process, method, article, or system. Unless otherwise specified, an element defined by the phrase "comprising one..." does not exclude the presence of other identical elements in the process, method, article, or system that includes that element.

[0059] The sequence numbers of the above embodiments of the present invention are for descriptive purposes only and do not represent the superiority or inferiority of the embodiments.

[0060] Through the above description of the embodiments, those skilled in the art can clearly understand that the methods of the above embodiments can be implemented by means of software plus necessary general-purpose hardware platforms. Of course, they can also be implemented by hardware, but in many cases the former is a better implementation method. Based on this understanding, the technical solution of the present invention, or the part that contributes to the prior art, can be embodied in the form of a software product. This computer software product is stored in a storage medium (such as ROM / RAM, magnetic disk, optical disk) as described above, and includes several instructions to cause a terminal device (which may be a mobile phone, computer, server, air conditioner, or network device, etc.) to execute the methods described in the various embodiments of the present invention.

[0061] The above are merely preferred embodiments of the present invention and do not limit the scope of the patent. Any equivalent structural or procedural transformations made based on the description and drawings of the present invention, or direct or indirect applications in other related technical fields, are similarly included within the scope of patent protection of the present invention.

Claims

1. A server optimization method based on AI and big data, characterized in that, The server optimization method based on AI big data includes: Acquire hardware layer data, software layer data, and network layer data. The hardware layer data includes CPU temperature data and PDU real-time power consumption data. The software layer data includes the container's storage occupancy rate and API gateway request response time. The network layer data includes flow table statistics and server communication latency matrix. The hardware layer data, the software layer data, and the network layer data are time-aligned to obtain a heterogeneous dataset, wherein the heterogeneous dataset is the dataset obtained after time alignment of the hardware layer data, the software layer data, and the network layer data. The server stress index is calculated based on the heterogeneous dataset, and a topological affinity matrix is ​​generated. Dual-channel processing is performed based on the server stress index and the topology affinity matrix, wherein short-term prediction is performed based on the server stress index and topology analysis is performed based on the topology affinity matrix. The results of short-term prediction channel processing and topology analysis channel processing are fused together for decision-making, and resource scheduling instructions are executed.

2. The server optimization method based on AI big data as described in claim 1, characterized in that, The step of performing time alignment processing on the data collected from the hardware layer, the data collected from the software layer, and the data collected from the network layer to obtain a heterogeneous dataset includes: The reference time source is extracted from the CPU temperature data, wherein the original timestamp of the hardware layer sensor clock is selected as the reference time source. Determine whether the PDU is a high-precision PDU. If the PDU is a high-precision PDU, directly obtain the power consumption data timestamp. If the PDU is a normal PDU, apply dynamic delay compensation to obtain the power consumption data timestamp. Here, the high-precision PDU refers to a PDU device equipped with a dedicated clock chip, and the normal PDU refers to a PDU device that relies on a software clock. Match the reference time source with the power consumption data timestamp to unify the hardware layer time axis; The target time point values ​​are calculated using cubic spline interpolation on the data collected by the software layer. Network layer data is aggregated using an exponentially weighted moving average.

3. The server optimization method based on AI big data as described in claim 2, characterized in that, If the PDU is a regular PDU, the step of applying dynamic delay compensation to obtain the power consumption data timestamp includes: If the PDU is a regular PDU, then the corrected timestamp is calculated using a first preset formula, and the corrected PDU timestamp is used as the power consumption data timestamp, wherein the first preset formula is: , The corrected PDU timestamp As a smoothing factor, For the precise local time of the data acquisition server, The network time obtained by the PDU device via the NTP protocol. This is the original timestamp of the PDU.

4. The server optimization method based on AI big data as described in claim 3, characterized in that, The steps of matching the reference time source with the power consumption data timestamp to unify the hardware layer time axis include: Based on the reference time source, the time residual is calculated on the power consumption data timestamp according to the second preset formula, wherein the second preset formula is: , The time residual is K, where K is the number of reference time points, and K is greater than or equal to 2. This refers to the time point of the power consumption data of the j-th PDU after the timestamp of the j-th PDU has undergone dynamic delay compensation. This is the base timestamp of the i-th CPU; Based on the time residual, a unified timestamp is calculated using a preset time mapping function to obtain a unified timestamp hardware layer dataset. This hardware layer dataset includes the CPU temperature data and the real-time power consumption data of the PDU after the unified timestamp. The time mapping function is: , The unified timestamp after alignment For time residuals, For dynamic compensation coefficients, The timestamp of the PDU to be mapped The timestamp of the current PDU The most recent CPU benchmark time point, This represents the local time deviation between the current PDU time point and the nearest CPU time point.

5. The server optimization method based on AI big data as described in claim 2, characterized in that, The steps for calculating the target time point value using cubic spline interpolation on the data collected by the software layer include: Obtain the set of data points and the target time point from the data collected by the software layer; The target time point value is calculated using a cubic spline function, wherein the cubic spline function is: , The target time point value, the For constant terms, The coefficient of the linear term, The coefficient of the quadratic term, The coefficient of the cubic term, As the independent variable, This represents the k-th data acquisition time.

6. The server optimization method based on AI big data as described in claim 2, characterized in that, The step of aggregating network layer data using an exponentially weighted moving average includes: Define the time window centered on the hardware reference time; Calculate the exponential weight for each network data point within the time window; The weighted aggregate value is calculated based on the exponential weights to obtain the time-aligned network layer acquisition data.

7. The server optimization method based on AI big data as described in claim 2, characterized in that, The steps of calculating the server stress index based on heterogeneous datasets and generating the topological affinity matrix include: The server stress index , To calculate the weighting coefficients, , For memory weighting coefficients, , These are the network weight coefficients. , This represents CPU utilization, and (memory / 100) represents the bus bandwidth usage percentage. For the number of packets retransmitted over the network, + + =1; The topological affinity matrix ,|delay i -Delay j | represents the communication delay from server i to j.

8. A server optimization system based on AI big data, characterized in that, The AI-based big data server optimization system includes: The acquisition module is used to acquire hardware layer data, software layer data, and network layer data. The hardware layer data includes CPU temperature data and PDU real-time power consumption data. The software layer data includes the container's storage occupancy rate and API gateway request response time. The network layer data includes flow table statistics and server communication latency matrix. The time alignment processing module is used to perform time alignment processing on the data collected by the hardware layer, the data collected by the software layer, and the data collected by the network layer to obtain a heterogeneous dataset, wherein the heterogeneous dataset is the dataset obtained after time alignment processing of the data collected by the hardware layer, the data collected by the software layer, and the data collected by the network layer. The calculation module is used to calculate the server stress index based on heterogeneous datasets and generate a topological affinity matrix; A dual-channel processing module is used to perform dual-channel processing based on the server stress index and the topology affinity matrix, wherein the short-term prediction channel is performed based on the server stress index, and the topology analysis channel is performed based on the topology affinity matrix. The decision fusion module combines the results of the short-term prediction channel with the results of the topology analysis channel to make decisions and execute resource scheduling instructions.

9. A server optimization device based on AI big data, characterized in that, The AI-based big data server optimization device includes: a memory, a processor, and an AI-based big data server optimization program stored in the memory and executable on the processor. When the AI-based big data server optimization program is executed by the processor, it implements the steps of the method as described in any one of claims 1 to 7.

10. A readable storage medium, characterized in that, The readable storage medium stores a server optimization program based on AI big data, which, when executed by a processor, implements the steps of the method as described in any one of claims 1 to 7.

Citation Information

Patent Citations

  • Multi-source computing power data integration and intelligent scheduling system and method

    CN118916147A

  • Decision-making large model-oriented multi-level heterogeneous memory collaborative scheduling method

    CN119576555A