Intelligent scheduling method and system for optical communication network resources
By acquiring equipment operation and service performance data in the optical communication network for time synchronization, identifying and marking weak fluctuations, and performing causal correlation analysis, the problem of existing systems being unable to detect weak fluctuations is solved, and stable transmission and quality improvement of high-value services are achieved.
Patent Information
- Application Number
- CN202511926713.5
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2025-12-19
- Publication Date
- 2026-01-23
- Estimated Expiration
- 2045-12-19
AI Technical Summary
Existing intelligent scheduling systems for optical communication networks cannot effectively detect, quantify, and avoid weak and intermittent physical layer performance fluctuations, leading to a decline in the quality of high-value services and making them difficult to diagnose.
By acquiring equipment operation data, scheduling decision data, and business performance data for time synchronization, we can identify and mark minor fluctuations below the preset alarm threshold, analyze the causal relationship between minor fluctuations and performance degradation, and adjust business path selection preferences.
It improves the accuracy and reliability of optical communication network resource scheduling, ensures the stable transmission of high-value services, and enhances the quality of service.
Smart Images

Figure CN121397397A_ABST
Abstract
Description
TECHNICAL FIELD
[0001] The present application relates to the technical field of intelligent scheduling of optical communication network resources, and particularly relates to an intelligent scheduling method and system for optical communication network resources. BACKGROUND
[0002] Modern optical communication networks are the backbone of various digital services, and the core lies in an intelligent scheduling system that efficiently and automatically allocates network resources. The system continuously collects real-time running state information of various resources such as fiber links, wavelengths, optical power, and devices (such as optical transponders, optical amplifiers, and reconfigurable optical add-drop multiplexers (ROADM)) in the network, and constructs a dynamically updated network view. Based on this view, the system uses complex optimization methods to calculate a transmission path that meets service requirements (such as low latency and high bandwidth) and takes into account the overall network efficiency for new service requests, and automatically completes wavelength allocation and device configuration.
[0003] However, in actual operation, network devices may have characteristics that are difficult to detect, causing the network to appear normal on the surface while hiding risks that affect service quality. For example, after a routine software version upgrade, a specific batch of reconfigurable optical add-drop multiplexers (ROADM) devices deployed in the network begin to exhibit abnormal behavior that is difficult to detect due to a small timing compatibility problem between the new version firmware and some early production hardware modules. This abnormality is manifested as extremely weak and intermittent fluctuations in optical power loss when the internal optical switch matrix performs wavelength routing switching under certain load conditions, such as when the ROADM needs to handle a large number of wavelength channels for routing switching, or when the environmental temperature reaches a certain critical value. The fluctuation amplitude is usually between 0.1 dB and 0.3 dB, which is much lower than the system's preset alarm threshold (usually 0.5 dB or higher), so it does not trigger any alarm information, and these ROADM devices still display as "normal operation" on the network management system's monitoring interface.
[0004] Further, due to such intermittent fluctuations, the data collected by the internal optical power monitoring module of the affected ROADM device also deviates slightly when reporting its own state data to the network management system. For example, after a certain wavelength channel is allocated and activated, the actual output optical power value of the channel reported by the ROADM may be slightly lower than its nominal value for a short time, or the internal optical signal-to-noise ratio (OSNR) monitoring value may also have a slight instantaneous drop at some time. These deviations are also very subtle and are not enough to be identified as error data by the data checking mechanism of the network management system, because they are still within the "normal" measurement error range and are considered as normal measurement noise or system jitter. The data preprocessing module of the scheduling system has certain abnormal data filtering capability, but its design goal is to eliminate obvious outliers or format errors, and it is powerless for such "legal" but "inaccurate" weak fluctuations.
[0005] The prior art needs to be improved in view of the above problems. SUMMARY
[0006] The present application provides an optical communication network resource intelligent scheduling method and system, which aims to solve the problem that the existing optical communication network intelligent scheduling system cannot effectively perceive, quantify and avoid potential risk paths when facing weak and intermittent physical layer performance fluctuations, resulting in a decline in the quality of high-value business services and difficulty in diagnosis.
[0007] In a first aspect, to solve the above technical problems, the present application provides an optical communication network resource intelligent scheduling method and system, comprising: Obtain device operation data, scheduling decision data and service performance data and perform time synchronization; According to the device operation data after time synchronization, identify weak fluctuations below the preset alarm threshold and perform risk labeling; When the performance of the service is monitored to decline, obtain the service identifier and the time point of the performance decline, backtrack the optical path of the service, and according to the optical path, the risk label, the scheduling decision data and the service performance data, correlate and analyze the causal relationship between the weak fluctuation record and the performance decline; According to the causal relationship, adjust the path selection preference of the service.
[0008] Preferably, the correlation and analysis of the causal relationship between the weak fluctuation record and the performance decline according to the optical path, the risk label, the scheduling decision data and the service performance data comprises: Obtain the optical power and optical signal-to-noise ratio data and radio frequency interference intensity data of the device on the optical path within the time period of the performance decline; Evaluate whether the device performs zero-point calibration; If yes, an instruction is issued to the device so that the device switches to an idle state or a standby bypass within a preset time window; Within the preset time window, optical power and optical signal-to-noise ratio data of the device are collected, and the radio frequency interference intensity data are collected synchronously; According to the device data collected in the zero-point calibration and the radio frequency interference intensity data, a contribution ratio of the weak fluctuation is quantified; According to the optical path, the risk label, the scheduling decision data, the service performance data, and the contribution ratio, the cause-effect correlation is associated.
[0009] Preferably, when the performance of the service is monitored to decrease, an identification of the service with performance decrease and a time point of occurrence are obtained, and an optical path passed through by the service is traced back, comprising: When the performance of the service is monitored to decrease, according to the identification of the service and the time point of occurrence, a path tracing back procedure is triggered; An allocation record of the service is obtained, the allocation record comprising an allocation path, a wavelength channel, and an accurate time stamp; Path distribution information of actual transmission traffic of the service within the preset time window is obtained, the path distribution information comprising a traffic ratio and a time stamp of each path; The allocation record and the path distribution information of the actual transmission traffic are compared, and a switching time point and a switched path of traffic between multiple optical paths within the preset time window are identified; According to the switching time point and the switched path, a time series distribution diagram of the traffic on each optical path within the preset time window is constructed; According to the time series distribution diagram, in combination with a duration of performance decrease of the service, an optical path carrying main traffic of the service is taken as the optical path passed through by the service.
[0010] Preferably, according to the time series distribution diagram, in combination with a duration of performance decrease of the service, an optical path carrying main traffic of the service is taken as the optical path passed through by the service, comprising: During the performance decrease of the service, a traffic ratio of each optical path in the time series distribution diagram is divided into a fine-grained time window, and the traffic ratio in each time window is counted; All optical paths overlapping in the duration of performance decrease of the service are identified, and a corresponding weak fluctuation record is obtained for each overlapping optical path; According to the counting result and the weak fluctuation record, a potential contribution value of each overlapping optical path to the performance decrease of the service is calculated; identify the light path path with the highest potential contribution value as the light path path.
[0011] Preferably, the calculating the potential contribution value of each light path path to the performance degradation of the service according to the statistical result and the weak fluctuation record comprises: obtaining a sensitivity weight of the service according to the type of the service; dynamically adjusting the weights of the traffic proportion and the weak fluctuation record in the potential contribution value based on the sensitivity weight; applying the adjusted weights to the traffic proportion and the weak fluctuation record to calculate the potential contribution value of each light path path to the performance degradation of the service.
[0012] Preferably, the obtaining a sensitivity weight of the service according to the type of the service comprises: obtaining a service characteristic parameter of an emerging service or a customized service; simulating the transmission of the emerging service or the customized service under different combinations of bit error rate, delay and packet loss rate in a controlled network environment, and recording the performance and user experience feedback of the service under each combination; constructing a sensitivity curve of the emerging service or the customized service according to the service characteristic parameter, the performance and the user experience feedback; extracting the sensitivity weight of the emerging service or the customized service according to the sensitivity curve.
[0013] Preferably, the extracting the sensitivity weight of the emerging service or the customized service according to the sensitivity curve comprises: identifying a performance degradation point corresponding to a specific performance degradation threshold in the sensitivity curve; expanding from the performance degradation point to both sides along the sensitivity curve to identify a local interval containing the performance degradation point and having relatively stable curve shape; identifying an inflection point or a starting point of a platform period of the sensitivity curve according to the slope and curvature change trend of the sensitivity curve in the local interval; dividing the local interval into a plurality of sub-intervals according to the inflection point or the starting point of the platform period; calculating the average sensitivity or the dominant sensitivity corresponding to each sub-interval; weighting the sensitivity weight according to the average sensitivity or the dominant sensitivity of the sub-interval and the position of the specific performance degradation threshold in the local interval.
[0014] Preferably, the identifying the inflection point or the platform starting point according to the slope, curvature change trend of the sensitivity curve in the local interval comprises: performing a moving average processing on the sensitivity curve in the local interval; calculating a first-order difference and a second-order difference of the sensitivity curve after the moving average processing; identifying a point with maximum slope change rate or continuous change of second-order difference sign as the inflection point according to the continuous change trend of the first-order difference and the second-order difference; identifying a region with continuous stability of the first-order difference in a preset small range and a duration exceeding a preset threshold as the platform starting point.
[0015] Preferably, after the identifying the region with continuous stability of the first-order difference in a preset small range and a duration exceeding a preset threshold as the platform starting point, the method further comprises: performing a consistency check on the identified inflection point or the platform starting point, and selecting a point with maximum change amplitude as the final inflection point or platform starting point when multiple similar points meet the identification condition.
[0016] In a second aspect, the present application provides an optical communication network resource intelligent scheduling system, comprising: an input end configured to acquire device running data, scheduling decision data and service performance data and perform time synchronization, and identify weak fluctuations below a preset alarm threshold and perform risk marking according to the device running data after time synchronization; an analysis end configured to acquire a service identifier and a time point of performance degradation when monitoring performance degradation of the service, backtrack an optical path of the service, and analyze a causal relationship between the weak fluctuations and the performance degradation according to the optical path, the risk marking, the scheduling decision data and the service performance data; an adjustment end configured to adjust path selection preference of the service according to the causal relationship.
[0017] Compared with the prior art, the present application has the following beneficial effects: By acquiring device running data, scheduling decision data and service performance data and performing time synchronization, the network running state can be comprehensively mastered. On this basis, the method can identify weak fluctuations below the preset alarm threshold and mark the risks, effectively solving the problem that weak fluctuations are difficult to be found in the prior art. When the service performance is monitored to decline, the method can acquire the service identifier and the time point of performance decline, and trace back the optical path path passed by the service, and then according to the optical path path, the risk mark, the scheduling decision data and the service performance data, the causal relationship between the weak fluctuation record and the performance decline is associated and analyzed. This innovative causal correlation analysis mechanism breaks through the limitation of the existing scheduling system which only relies on surface "normal" data for decision-making, and can reveal the potential risks hidden behind the weak fluctuations. Finally, according to the causal correlation, the method adjusts the subsequent path selection preference of the service, so as to actively avoid the optical path path with potential risks and ensure the stable transmission of high-value services. Through the above technical scheme, the present application effectively solves the problem that the intelligent scheduling system in the prior art cannot perceive and quantify the cumulative influence of weak physical layer performance fluctuations on optical path transmission quality, leading to suboptimal resource allocation and hidden risks, and significantly improves the accuracy, reliability and service quality of optical communication network resource scheduling. BRIEF DESCRIPTION OF DRAWINGS
[0018] Figure 1 is a flow chart of an optical communication network resource intelligent scheduling method provided by an embodiment of the present application; Figure 2 is a flow chart of a method for associated analysis of weak fluctuation records and causal relationship provided by an embodiment of the present application; Figure 3 is a structural schematic diagram of an optical communication network resource intelligent scheduling system provided by an embodiment of the present application. DETAILED DESCRIPTION
[0019] The technical solutions in the embodiments of the present application will be clearly and completely described below with reference to the drawings in the embodiments of the present application. Obviously, the described embodiments are only part of the embodiments of the present application, not all. Based on the embodiments in the present application, all other embodiments obtained by those skilled in the art without creative labor are within the scope of protection of the present application.
[0020] Referring to Figure 1 , Figure 1 is a flow chart of an optical communication network resource intelligent scheduling method provided by an embodiment of the present application, comprising the following steps: S1, acquiring device running data, scheduling decision data and service performance data and performing time synchronization; S2, according to the device running data after time synchronization, identifying weak fluctuations below the preset alarm threshold and marking the risks; S3, when the performance of the service is monitored to decrease, the service identifier and the time point of performance decrease are obtained, the optical path path passed by the service is traced back, and the causal relationship between the weak fluctuation record and the performance decrease is associated and analyzed according to the optical path path, the risk label, the scheduling decision data and the service performance data; S4, according to the causal relationship, the path selection preference of the service is adjusted.
[0021] The existing intelligent scheduling system of the traditional optical communication network mainly relies on the real-time running state data reported by the equipment when allocating resources. However, these data may have weak and intermittent physical layer performance fluctuations, such as slight changes in optical power loss or transient decreases in signal-to-noise ratio. These fluctuations are often below the preset alarm threshold, causing the scheduling system to fail to perceive and quantify their potential impact on service transmission quality, which may allocate high-value services to paths with potential risks, ultimately leading to service performance degradation and difficulty in diagnosis.
[0022] To this end, the present application proposes an intelligent scheduling method for optical communication network resources, which can identify weak fluctuations below the preset alarm threshold and mark risks by obtaining equipment running data, scheduling decision data and service performance data and performing time synchronization. When the performance of the service is monitored to decrease, the method can trace back the optical path path passed by the service, and comprehensively analyze the causal relationship between the weak fluctuation record and the performance decrease according to the optical path path, the risk label, the scheduling decision data and the service performance data. Thus, the path selection preference of the service can be adjusted according to the causal relationship, thereby effectively avoiding potential risks and improving the stability and reliability of service transmission.
[0023] In order to better understand the intelligent scheduling method for optical communication network resources proposed by the present application, some key terms involved therein are first explained.
[0024] “Equipment running data” refers to real-time or historical data generated by various types of equipment (such as optical transponders, optical amplifiers, reconfigurable optical add-drop multiplexers ROADM, etc.) in the optical communication network during operation, including but not limited to physical layer parameters such as optical power, optical signal-to-noise ratio (OSNR), bit error rate (BER), temperature, voltage, current, as well as working mode, port state, alarm information, etc. of the equipment. These data are the basis for evaluating the health status and performance of network equipment.
[0025] “Scheduling decision data” refers to the strategies, rules and historical decision records that the intelligent scheduling system relies on when allocating network resources. This includes but is not limited to parameters of path selection algorithms, wavelength allocation strategies, service priority settings, network topology information, resource capacity limits and historical scheduling results, etc. These data provide guidance and reference for the system to make future scheduling.
[0026] “Service performance data” refers to the performance indicators of various services carried in the network during transmission, such as delay, jitter, packet loss rate, throughput, bit error rate, and user experience feedback, etc. These data directly reflect the quality of service (QoS) and user experience (QoE) of the service, and are an important basis for evaluating the service level of the network.
[0027] “Weak fluctuation” refers to abnormal changes in the device operation data that have small amplitude and may have intermittent nature, which are usually below the conventional alarm threshold and thus not easily detected by traditional monitoring systems. For example, a transient drop in optical power in the range of 0.1 dB to 0.3 dB, or a slight fluctuation in OSNR.
[0028] “Risk marking” refers to marking the identified weak fluctuations to indicate their potential impact on service performance in terms of risk level or type.
[0029] “Optical path” refers to a series of fiber links and optical network devices through which service data passes from the source point to the destination point in an optical communication network.
[0030] “Causal correlation” refers to determining whether a weak fluctuation is the cause or a significant factor affecting the decline in service performance by analyzing the temporal, spatial, and logical relationship between the weak fluctuation record and the decline in service performance.
[0031] The implementation environment of the present application is generally an intelligent scheduling system in an optical communication network, which has data collection, processing, analysis, decision-making, and control capabilities, and can interact with various optical devices in the network and issue instructions.
[0032] The optical communication network resource intelligent scheduling method proposed in the present application is based on the perception, analysis, and utilization of weak fluctuations in the network, thereby optimizing the path selection of services.
[0033] Firstly, device operation data, scheduling decision data, and service performance data need to be acquired and time-synchronized. Device operation data can be periodically collected from optical network devices through a network management system (NMS) or a dedicated performance monitoring agent. For example, a ROADM device can be configured to report its port optical power and OSNR data every 100 milliseconds. Scheduling decision data can be directly read from the database of an intelligent scheduling system, including the current effective scheduling strategy and historical scheduling records. Service performance data can be monitored in real-time by deploying performance probes on the network edge or service servers, such as sampling the delay, packet loss rate, and bit error rate of specific service flows. In order to ensure the accuracy of subsequent analysis, all acquired data needs to be time-synchronized. This can be achieved by using high-precision time protocols (such as NTP or PTP) to calibrate the timestamps of different data sources, ensuring that all data points can accurately correspond to the same time axis.
[0034] Secondly, according to the time-synchronized device operation data, weak fluctuations below the preset alarm threshold are identified and risk-labeled. Traditional alarm systems usually only focus on significant abnormalities that exceed a certain threshold. However, this application focuses on those "legal" but "inaccurate" weak fluctuations. For example, sliding window averaging can be performed on optical power data, and the deviation between each data point and the sliding average is calculated. If the deviation of a certain data point continuously exceeds a small preset threshold (e.g. 0.2 dB), but is below the regular alarm threshold (e.g. 0.5 dB), it can be identified as a weak fluctuation. For the identified weak fluctuations, different risk labels can be assigned according to factors such as their duration, amplitude, frequency of occurrence, and the type of equipment involved. For example, weak fluctuations with longer duration and larger amplitude can be labeled as "high risk", while short-lived and small-amplitude fluctuations can be labeled as "low risk". These risk labels will serve as important inputs for subsequent causal correlation analysis.
[0035] Furthermore, when the performance of a service is monitored to decline, the service identifier and the time point of the performance decline are acquired, the optical path of the service is traced back, and the causal relationship between the weak fluctuation records and the performance decline is analyzed according to the optical path, risk labels, scheduling decision data, and service performance data. The monitoring of service performance decline can be achieved by continuously analyzing service performance data. For example, when the bit error rate of a certain service significantly increases within a short period of time, or the delay exceeds the preset service level agreement (SLA) threshold, it is determined that the performance has declined. Once the performance decline is monitored, the system will immediately record the unique identifier of the service and the exact time point of the performance decline.
[0036] Subsequently, the system needs to trace back the optical path that the service has passed during the performance degradation. This can be achieved by querying the historical path allocation records of the scheduling system. For example, if a service is allocated to path A in a certain time period and path B in another time period, it is necessary to identify which path the service mainly carries on when the performance degradation occurs. After determining the optical path, the system will conduct comprehensive correlation analysis by combining the weak fluctuation records (with risk labels) of all devices on the path, historical scheduling decision data, and performance data of the service itself. For example, if there is a weak fluctuation labeled as “high risk” on the path passed by the performance-degraded service at a certain time point, and the fluctuation highly coincides with the time point of the service performance degradation, it can be preliminarily judged that there is a causal relationship between the two.
[0037] Finally, according to the causal correlation, the subsequent path selection preference of the service is adjusted. Once the causal relationship between the weak fluctuation and the service performance degradation is determined, the system will use this information to optimize future scheduling decisions. For example, if it is found that a weak fluctuation on a certain ROADM device leads to the performance degradation of a specific service, the system will reduce the priority of the path containing the ROADM device or completely avoid using the device when selecting the path for the service or similar sensitive services in the future. This can be achieved by introducing a “risk weight” or “blacklist” mechanism in the scheduling algorithm. For example, the cost factor of the device or link with a weak fluctuation can be increased, so that the scheduling algorithm tends to select other more stable paths when optimizing the objective function. This adjustment can be dynamic and continuously updated as the network state changes and new weak fluctuation records appear.
[0038] The optical communication network resource intelligent scheduling method proposed in the present application effectively solves the problem that the traditional scheduling system cannot perceive and avoid hidden network risks by introducing the identification and risk labeling of weak fluctuations below the preset alarm threshold.
[0039] The core innovation of the present application is that it breaks through the limitation of traditional scheduling systems that only focus on explicit alarms and incorporates the perception and quantification of “weak fluctuations” into the scheduling decision-making process. By performing fine-grained analysis on device operation data, the present application can identify and label weak fluctuations below the preset alarm threshold, revealing potential and hidden risk points in the network. Further, when the service performance degrades, the present application can trace back the optical path that the service has passed and conduct causal correlation analysis by combining the risk labels, scheduling decision data, and service performance data, thereby accurately locating the weak fluctuation that causes the performance degradation.
[0040] In some embodiments of the application described above, the causal relationship between the weak fluctuation record and the performance degradation is analyzed according to the optical path, the risk label, the scheduling decision data and the service performance data. However, in actual application, direct correlation analysis based on these data may not completely rule out the influence of device measurement errors or environmental factors (such as radio frequency interference) on the quality of the optical signal, resulting in inaccurate judgment of the causal relationship between the weak fluctuation and the performance degradation, or even misjudgment. If the above problems are not solved, it may lead to incorrect scheduling decisions, affecting the stability of the service and the user experience. To this end, the application further provides a more accurate causal correlation analysis method, which introduces a zero-point calibration mechanism to quantify the real contribution ratio of the weak fluctuation, thereby improving the accuracy of causal analysis.
[0041] Reference Figure 2 , Figure 2 is a method flowchart for analyzing the weak fluctuation record and the causal relationship provided by an embodiment of the application, specifically comprising: S31, obtaining the optical power and optical signal-to-noise ratio data and radio frequency interference intensity data of the device on the optical path within the time period of the performance degradation; S32, evaluating whether the device can perform zero-point calibration; S33, if yes, issuing an instruction to the device to switch to an idle state or a standby bypass within a preset time window; S34, within the preset time window, collecting the optical power and optical signal-to-noise ratio data of the device, and synchronously collecting the environmental radio frequency interference intensity data; S35, quantifying the contribution ratio of the weak fluctuation according to the device data collected during the zero-point calibration and the environmental radio frequency interference intensity data; S36, correlating the causal relationship according to the optical path, the risk label, the scheduling decision data, the service performance data and the contribution ratio.
[0042] Specifically, when the service performance degradation is monitored, the optical power data, the optical signal-to-noise ratio data and the radio frequency interference intensity data of the related device on the optical path within the performance degradation time period are first obtained. These data are used for preliminary analysis of the device running state and the external interference. Among them, the optical power data reflects the intensity of the optical signal, the optical signal-to-noise ratio data reflects the quality of the optical signal, and the radio frequency interference intensity data is used to evaluate the potential influence of the external electromagnetic environment on the optical communication link.
[0043] Further, it is necessary to evaluate whether the device has the ability to perform zero-point calibration. Zero-point calibration refers to measuring the output of the device under no signal input or known stable input to determine the inherent bias or background noise of the device. If the device supports zero-point calibration, an instruction is issued to the device to switch to an idle state or a standby bypass within a preset time window. The idle state means that the device does not carry any traffic, and the standby bypass means that the traffic is temporarily switched to other paths so that the target device can be tested independently. The preset time window is a pre-set time period to ensure the integrity of the calibration process and not to affect normal business operation.
[0044] Within the preset time window, the optical power and optical signal-to-noise ratio data of the device are collected, and the environmental radio frequency interference intensity data are collected synchronously. The optical power and optical signal-to-noise ratio data collected at this time reflect the baseline performance and internal noise level of the device under no traffic or no external optical signal input. The environmental radio frequency interference intensity data are collected synchronously to exclude or quantify the influence of environmental radio frequency interference on the measurement results of the device during calibration.
[0045] Therefore, according to the device data (i.e. baseline data) collected during the zero-point calibration and the environmental radio frequency interference intensity data, the contribution ratio of the weak fluctuation can be quantified. The contribution ratio represents the degree of signal quality decline actually caused by the weak fluctuation on the optical path after excluding the inherent bias of the device itself and the environmental radio frequency interference. For example, by comparing the measurement values during calibration as a reference with the measurement values during business operation, the real changes caused by the weak fluctuation can be stripped out.
[0046] Finally, the causal relationship is associated according to the optical path, the risk label, the scheduling decision data, the business performance data, and the contribution ratio. By introducing the quantified contribution ratio, the causal relationship between the weak fluctuation record and the business performance decline can be more accurately judged, and the inherent error of the device or the environmental interference is avoided to be misjudged as the performance problem caused by the weak fluctuation.
[0047] The scheme of the present application effectively solves the problem of insufficient precision in the correlation analysis of weak fluctuations and performance degradation by introducing a zero-point calibration mechanism. Specifically, after monitoring the performance degradation of a service, the optical power, optical signal-to-noise ratio, and radio frequency interference intensity data of the equipment on the optical path during the performance degradation period are first obtained, which provides a basis for preliminary analysis. Subsequently, by evaluating whether the equipment supports zero-point calibration, and in the case of support, the equipment is switched to an idle state or a standby bypass, so that baseline optical power and optical signal-to-noise ratio data of the equipment are collected within a preset time window, and environmental radio frequency interference intensity data are collected synchronously. It is precisely because these baseline data are collected in the absence of service or in a known stable state that the inherent bias of the equipment itself, internal noise, and the influence of environmental radio frequency interference can be separated from the measurement data during service operation. In this way, the real contribution ratio of weak fluctuations on the optical path to service performance degradation can be accurately quantified, thereby avoiding misjudgment of non-fluctuation factors (such as equipment aging and environmental interference) as the cause of weak fluctuations. This quantified contribution ratio makes the causal correlation analysis more accurate and reliable.
[0048] Through the above technical solution, the present application can significantly improve the accuracy of causal correlation analysis between weak fluctuations and service performance degradation in an optical communication network. Specifically, through zero-point calibration and synchronous collection of environmental radio frequency interference data, the measurement error of the equipment itself and external environmental interference can be effectively eliminated, making the quantification of the real contribution of weak fluctuations more accurate. This avoids misjudgment of non-fluctuation factors as the cause of performance degradation, thereby reducing false positives and unnecessary resource scheduling. As a result, based on more accurate causal correlation, subsequent service path selection preference adjustment will be more targeted and effective, and can more accurately avoid optical paths with real weak fluctuations, thereby improving the overall stability of the optical communication network and the quality of service, and reducing operating costs.
[0049] In some preferred embodiments, the following is described by a specific example. Assume that an optical path in an optical communication network carries high-priority services, and recently it is monitored that the performance of the service has intermittently decreased. Preliminary analysis shows that there are weak fluctuation records on the optical path that are lower than a preset alarm threshold. In order to accurately determine whether these weak fluctuations are the real cause of the performance degradation of the service, the scheme of the present application is enabled.
[0050] First, the system obtains the optical power, optical signal-to-noise ratio, and radio frequency interference intensity data of the relevant optical transmission equipment on the optical path during the performance degradation period. Subsequently, the system evaluates and finds that the optical transmission equipment supports zero-point calibration function. Therefore, the system issues an instruction to the equipment to switch the service traffic to a standby bypass within a preset time window (e.g., 2:00 to 2:15 in the early morning) during the night service low peak period, so that the equipment enters an idle state.
[0051] Within the preset time window, the device automatically collects its optical power and optical signal-to-noise ratio baseline data when there is no traffic flow, and at the same time, the environmental sensor deployed near the device synchronously collects the environmental radio frequency interference intensity data within the time period. After calibration, the system uses these baseline data and environmental radio frequency interference data to correct and analyze the data collected during the performance degradation of the service. For example, if the calibration data shows that the device itself has a 0.5 dBm optical power measurement deviation, and the environmental radio frequency interference causes a 0.2 dB signal-to-noise ratio drop in a specific frequency band, these factors will be deducted when quantifying the contribution proportion of weak fluctuations.
[0052] In this way, the system can accurately calculate the proportion of performance degradation actually caused by weak fluctuations on the optical path after excluding the inherent deviation of the device and environmental interference. For example, if the corrected weak fluctuations contribute up to 80% to the performance degradation, it can be concluded that weak fluctuations are the main cause; if the contribution proportion is low, further investigation of other factors may be needed. Ultimately, based on this more accurate contribution proportion, the system can more accurately associate the causal relationship between weak fluctuations and service performance degradation, and accordingly adjust the subsequent path selection preference of the service, such as permanently or temporarily switching the service to another more stable optical path, thereby effectively avoiding resource waste or service interruption caused by misjudgment.
[0053] In some embodiments of the above application, it is proposed that when the performance degradation of the service is monitored, the optical path path traversed by the service needs to be traced back for subsequent causal correlation analysis. However, in actual optical communication networks, traffic flow may not always be transmitted along a single preset path, especially when the network is dynamically adjusted, load balanced, or fault recovered, traffic flow may be switched between multiple optical path paths. If the main optical path path actually carrying the traffic flow during performance degradation cannot be accurately identified, the accuracy of causal correlation analysis may be reduced, thereby affecting the effectiveness of subsequent scheduling decisions.
[0054] In this regard, the present application further proposes that the step of tracing back the optical path path traversed by the service when the performance degradation of the service is monitored includes: When the performance degradation of the service is monitored, triggering a path tracing back program according to the service identifier and the occurrence time point; Obtaining the allocation record of each path of the service, the allocation record including the allocated path, wavelength channel, and accurate time stamp; Obtaining the path distribution information of the actual transmission flow of the service within the preset time window, the path distribution information including the flow proportion and time stamp of each path; comparing the allocation record with the path distribution information of the actual transmission traffic, identifying a switching time point and a switched path of the traffic between multiple optical path paths within the preset time window; constructing a time series distribution diagram of the traffic on each optical path path within the preset time window according to the switching time point and the switched path; combining the duration of performance degradation of the service, and taking the optical path path carrying the main traffic of the service as the optical path path passed through by the service according to the time series distribution diagram.
[0055] Specifically, when the performance of a service in an optical communication network is monitored to be degraded, the system immediately obtains the unique identifier of the service with performance degradation and the accurate time point of performance degradation. Based on this information, a special path backtracking program will be triggered, which aims to accurately determine the actual optical path path passed through by the service during performance degradation.
[0056] The path backtracking program first obtains the historical path allocation record of the service from the network management system. The allocation record details the optical path path to which the service is allocated at different time points, the wavelength channel used, and the corresponding accurate time stamp. These records reflect the "planned" path of the service in the network.
[0057] At the same time, the system also obtains the path distribution information of the actual transmission traffic of the service within the preset time window. The path distribution information is obtained through real-time monitoring or historical data analysis, which contains the distribution proportion of the service traffic on different optical path paths and the corresponding collection time stamp, reflecting the "actual" transmission situation of the service traffic.
[0058] Subsequently, by comparing the allocation record with the path distribution information of the actual transmission traffic, the system can identify the specific time point of switching of the traffic between multiple optical path paths and the optical path path used after switching within the preset time window. This comparison helps to find the difference between the actual traffic and the allocation record, especially the path deviation that may occur during network dynamic adjustment or fault recovery.
[0059] Based on the identified switching time point and switched path, the system will construct a time series distribution diagram of the traffic on each optical path path within the preset time window. The distribution diagram intuitively shows the dynamic distribution of the service traffic on different optical path paths, for example, the service traffic may be mainly concentrated in path A in a certain time period, and switched to path B in another time period.
[0060] Finally, according to the time series distribution diagram and in combination with the duration of the performance degradation of the service, the system will identify the optical path path carrying the main traffic of the service during the performance degradation period and take it as the actual optical path path passed by the service. This ensures that the optical path path based on the subsequent causal correlation analysis is the most consistent with the actual service traffic condition.
[0061] The scheme of the present application effectively solves the problem that the traditional method cannot accurately identify the actual transmission path of the service in a dynamic network environment by introducing a refined path backtracking mechanism. Specifically, when the performance of the service degrades, the system not only considers the preset path allocation record, but more importantly, it obtains the path distribution information of the actual transmission traffic and compares it with the allocation record, thereby accurately identifying the switching behavior of the service traffic between multiple optical path paths. Thus, by constructing a time series distribution diagram of the traffic on each optical path path and in combination with the duration of the performance degradation, the scheme can identify the optical path path carrying the main traffic of the service during the critical period. This method avoids inaccurate path identification caused by path switching or traffic dispersion and provides a more accurate and reliable basis for the subsequent causal correlation analysis of weak fluctuations and performance degradation.
[0062] Through the above technical scheme, the present application can overcome the limitation of the traditional method that the path backtracking is inaccurate due to the switching or dispersion of service traffic paths in a complex dynamic network environment. The present scheme accurately identifies the traffic switching point by comprehensively analyzing the allocation record and actual traffic distribution of the service and constructs a time series distribution diagram, thereby accurately determining the optical path path actually carrying the main traffic during the performance degradation of the service. This significantly improves the accuracy and reliability of optical path path identification, provides a solid foundation for the subsequent causal correlation analysis of weak fluctuations and performance degradation of the service, and further enables the intelligent scheduling system to make more accurate and effective path selection preference adjustments, ultimately improving the overall operation efficiency and service quality of the optical communication network.
[0063] In some preferred embodiments, the following is described by a specific example. Assume that in a certain optical communication network, a high-bandwidth video conference service is monitored to have serious video stuttering and packet loss from 10:30 to 10:45 in the morning, i.e., the performance of the service degrades.
[0064] First, the system will obtain the identifier of the video conference service (e.g., service ID: VC_20231027_001) and the occurrence time point of the performance degradation (10:30). Subsequently, the path backtracking program is triggered.
[0065] The procedure queries the history record and finds that the service was assigned to lightpath path A at 10:00 am using wavelength channel 1550 nm. However, through real-time traffic monitoring data, the system finds that the actual traffic distribution of the service within the preset time window of 10:00-10:45 is as follows: - 10:00-10:15: 100% traffic is transmitted on lightpath path A.
[0066] - 10:15-10:30: Due to network load balancing policy, 70% traffic is transmitted on lightpath path A and 30% traffic is shunted to lightpath path B.
[0067] - 10:30-10:45: Due to slight jitter of a device on lightpath path A, the system automatically switches all traffic of the service to lightpath path B.
[0068] By comparing the assignment record and the actual traffic distribution, the system identifies that traffic switching occurs at 10:15 and 10:30. Based on these switching points, the system constructs a time series distribution diagram of the traffic of the service on lightpath path A and lightpath path B during 10:00-10:45.
[0069] In combination with the duration of performance degradation of the service (10:30-10:45), the system analyzes the time series distribution diagram. It is found that during the entire duration of performance degradation, lightpath path B carries all or most of the traffic of the service. Therefore, the system finally identifies lightpath path B as the main lightpath path through which the video conference service passes during performance degradation. In this way, the subsequent causal correlation analysis will focus on the weak fluctuation record on lightpath path B, so as to more accurately locate the root cause of the problem.
[0070] In some embodiments of the above application, by constructing a time series distribution diagram of the traffic of the service on each lightpath path and identifying the lightpath path carrying the main traffic according to the distribution diagram, the lightpath path through which the service passes is traced back. However, in actual application, the cause of service performance degradation may not simply be caused by the lightpath path carrying the main traffic, especially when multiple lightpath paths carry the traffic of the service and a certain path has weak fluctuation, the judgment of the main traffic may not accurately identify the root cause of performance degradation. This may lead to misjudgment of the root cause of the fault, thereby affecting the accuracy of subsequent scheduling decisions.
[0071] In view of this, the application further provides the above method, comprising: During the performance degradation of the service, the traffic proportion of each light path in the time series distribution diagram is divided into fine-grained time windows, and the traffic proportion in each time window is counted; All light paths overlapping in the performance degradation duration of the service are identified, and for each overlapping light path, a corresponding weak fluctuation record is obtained; According to the statistical results and the weak fluctuation record, the potential contribution value of each overlapping light path to the performance degradation of the service is calculated; The overlapping light path with the highest potential contribution value is identified as the light path through which the service passes.
[0072] Specifically, during the performance degradation of the service, the system divides the traffic proportion of each light path in the time series distribution diagram into fine-grained time windows. The fine-grained time window can be understood as further dividing the entire performance degradation duration into smaller time periods, such as every second, every minute, or shorter time intervals, in order to more accurately capture the dynamic changes in traffic on different paths. In each fine-grained time window, the traffic proportion carried by each light path is counted, obtaining a detailed dynamic view of traffic distribution.
[0073] Among them, identifying all light paths overlapping in the performance degradation duration of the service means that all light paths that have ever or are currently carrying the traffic of the service are identified within the entire duration of the performance degradation of the service. For each identified overlapping light path, the system obtains its corresponding weak fluctuation record. These weak fluctuation records are weak fluctuations below the preset alarm threshold identified from time-synchronized device operation data and have been risk-labeled.
[0074] In practical applications, according to the statistical results and the weak fluctuation record, the potential contribution value of each overlapping light path to the performance degradation of the service is calculated. This potential contribution value aims to quantify the possible influence of each path in the performance degradation event of the service, which considers the traffic proportion of the path during the performance degradation (statistical results) and the existence of weak fluctuations on the path. For example, the higher the traffic proportion, the more frequent or more serious the weak fluctuations, the higher the potential contribution value.
[0075] Finally, the system identifies the overlapping light path with the highest potential contribution value as the light path through which the service passes. This path is considered to be the most important or most direct light path that causes the performance degradation of the service.
[0076] The scheme of the present application solves the problem of possible misjudgment caused by only judging the main traffic by introducing fine-grained time window division, identifying overlapping light path paths, and combining with weak fluctuation records. Specifically, by performing fine-grained time window division on the traffic proportion during the service performance decline period, the dynamic distribution of traffic on different paths can be captured more finely, avoiding the information loss that may be caused by coarse-grained statistics. At the same time, all overlapping light path paths in the performance decline duration are identified to ensure that even the paths carrying non-main traffic but having potential problems can be included in the consideration range. More importantly, by comprehensively analyzing the traffic proportion statistical results of each overlapping light path path and its corresponding weak fluctuation record, and calculating the potential contribution value, the system can quantify the actual impact of each path on the performance decline. This quantification method can more accurately evaluate which paths have a stronger causal relationship between the weak fluctuation and the service performance decline, thereby avoiding the limitation of only focusing on traffic size and ignoring potential risks. Thus, the root light path path causing the service performance decline can be more accurately located.
[0077] Through the above technical scheme, the present application can significantly improve the accuracy of service performance decline fault positioning in optical communication networks. By carefully analyzing the traffic distribution and weak fluctuation records, even in multi-path transmission or complex fault scenarios, the light path path with the highest potential contribution to the service performance decline can be accurately identified. This not only avoids the possible misjudgment of traditional methods, but also provides a more reliable and fine basis for subsequent causal relationship analysis and path selection preference adjustment, thereby effectively improving the efficiency and accuracy of network resource intelligent scheduling, and ensuring the stable operation of services and user experience.
[0078] In some preferred embodiments, the following is described by a specific example. Assume that a video conference service has a performance decline phenomenon of video freezing and audio interruption during 10:00-10:05 am. According to the above method, the system has constructed a time series distribution diagram of the traffic of the service on light path path A, light path path B, and light path path C during 10:00-10:05. Among them, light path path A carries about 70% of the traffic, light path path B carries 20%, and light path path C carries 10%.
[0079] In some embodiments of the application described above, when calculating the potential contribution of each overlapping path to the performance degradation of the service, the traffic proportion statistics and the weak fluctuation record are mainly used. However, in practical applications, different types of services have significant differences in sensitivity to network performance degradation. For example, real-time audio and video services are extremely sensitive to delay and packet loss rate, while offline data transmission services are relatively insensitive. If the type of service is not distinguished and the potential contribution value is simply calculated based on a unified weight, it may lead to inaccurate analysis of the performance degradation of some key or sensitive services, thereby affecting the adjustment effect of the subsequent path selection preference, and even may not be able to effectively guarantee the quality of service of high priority services in time.
[0080] To this end, the application further proposes that the above-mentioned calculation of the potential contribution of each overlapping path to the performance degradation of the service according to the statistics and the weak fluctuation record comprises: According to the type of the service, the sensitivity weight of the service is obtained; Based on the sensitivity weight, the weights of the traffic proportion and the weak fluctuation record in the potential contribution value are dynamically adjusted; The adjusted weights are applied to the traffic proportion and the weak fluctuation record to calculate the potential contribution of each overlapping path to the performance degradation of the service.
[0081] Specifically, the above-mentioned obtaining of the sensitivity weight of the service according to the type of the service refers to that the system queries or calculates the sensitivity of the type of service currently experiencing performance degradation, such as video conference, online game, VoIP call, data backup or web browsing, to network performance degradation. The sensitivity weight can be a quantitative value reflecting the degree of influence of the service on the quality of service or user experience when facing a specific performance degradation (such as delay increase, packet loss rate increase, bandwidth fluctuation, etc.). For example, the sensitivity weight of real-time interactive services is usually higher than that of non-real-time data transmission services.
[0082] Among them, the above-mentioned dynamic adjustment of the weights of the traffic proportion and the weak fluctuation record in the potential contribution value based on the sensitivity weight refers to that when calculating the potential contribution value, instead of using a fixed weight factor, the influence of the traffic proportion and the weak fluctuation record in the final contribution value calculation is dynamically adjusted according to the obtained sensitivity weight of the service. For example, for high-sensitivity services, the weight of the weak fluctuation record may be increased to highlight its impact on performance degradation; and for low-sensitivity services, the weight of the traffic proportion may remain unchanged or be slightly adjusted. This dynamic adjustment ensures that the contribution value calculation can better reflect the characteristics of the service.
[0083] In actual application, the application of the adjusted weight to the traffic proportion and the weak fluctuation record to calculate the potential contribution value of each overlapping optical path to the performance degradation of the service refers to multiplying the dynamically adjusted weight coefficient to the traffic proportion statistical value and the weak fluctuation record value of the corresponding optical path respectively, and then comprehensively calculating the weighted values to obtain the final potential contribution value of each overlapping optical path to the current performance degradation of the service. The calculation result can more accurately reflect the actual influence degree of the weak fluctuation on the different optical paths on the performance degradation of the specific service.
[0084] The scheme of the present application effectively solves the limitation of the traditional method that cannot fully consider the difference of the service types when evaluating the cause of the service performance degradation by introducing the service sensitivity weight and dynamically adjusting the weight of the traffic proportion and the weak fluctuation record in the calculation of the potential contribution value. Specifically, when the performance degradation of the service is monitored, the sensitivity weight of the service type is obtained first, and the weight can quantify the bearing capacity of different services to the network fluctuation. Then, the influence of the traffic proportion and the weak fluctuation record in the calculation of the potential contribution value is dynamically adjusted by using the sensitivity weight. For example, for the service (such as real-time video) that is highly sensitive to delay and packet loss rate, even if the absolute value of the weak fluctuation record is not large, the weight of the weak fluctuation record in the potential contribution value is also increased correspondingly to highlight the potential influence of the weak fluctuation on the service performance. Conversely, for the service (such as background data synchronization) that has low performance requirement, the weight of the weak fluctuation record is relatively low. Thus, through the differentiated weight adjustment mechanism, the finally calculated potential contribution value can more accurately reflect the actual influence degree of the weak fluctuation on the specific optical path on the performance degradation of the current service, thereby avoiding the evaluation deviation of all services.
[0085] Through the above technical scheme, the present application can realize more refined and personalized analysis of the cause of the service performance degradation in the optical communication network. By introducing the service sensitivity weight and dynamically adjusting the calculation parameters, the system can fully consider the differentiated requirements of different services to the network quality when identifying the causal relationship between the weak fluctuation and the service performance degradation. This not only improves the accuracy of the calculation of the potential contribution value, but also makes the root cause positioning more accurate, and helps the scheduling system to more intelligently identify and prioritize the problems that have the greatest impact on the high sensitivity service, thereby optimizing the resource scheduling strategy, effectively guaranteeing the service quality and user experience of the key service, and avoiding the waste of resources or service degradation caused by the general evaluation model.
[0086] In some preferred embodiments, the following is described by a specific example. It is assumed that there are two services: service A is real-time video conference, and service B is large file download. When the performance degradation of service A and service B is monitored, the system will obtain the sensitivity weights of service A and service B respectively.
[0087] For service A (real-time video conference), its sensitivity weight is set to a high value, e.g. 0.8, because it is very sensitive to delay and packet loss. At this time, when calculating the potential contribution value, the weight of the weak fluctuation record will be dynamically adjusted upwards based on 0.8, while the weight of the traffic proportion may be relatively adjusted downwards, to ensure that even a small weak fluctuation, as long as it occurs on the optical path carrying service A, its potential contribution value to the performance degradation of service A will be significantly amplified.
[0088] For service B (large file download), its sensitivity weight is set to a low value, e.g. 0.3, because it has a higher tolerance to short delay and small packet loss. At this time, when calculating the potential contribution value, the weight of the weak fluctuation record will be dynamically adjusted downwards based on 0.3, while the weight of the traffic proportion may be relatively adjusted upwards, to reflect its dependence on continuous large traffic transmission.
[0089] For example, if there is the same weak fluctuation record and traffic proportion on a certain optical path, in calculating the potential contribution value to service A, due to its high sensitivity weight, the contribution value of this optical path will be calculated to be higher; while in calculating the potential contribution value to service B, due to its low sensitivity weight, the contribution value of this optical path will be relatively lower. Thus, the system can more accurately identify the optical path problem that has the greatest impact on real-time video conference, and prioritize processing or adjusting its path selection preference, thereby effectively improving the user experience of high sensitivity services.
[0090] In some embodiments of the application described above, although the sensitivity weight of the service is obtained according to the type of the service, and the weights of the traffic proportion and the weak fluctuation record in the potential contribution value are dynamically adjusted based on the sensitivity weight, to calculate the potential contribution value of each overlapping optical path to the performance degradation of the service, for emerging services or customized services, their sensitivity characteristics may lack historical data or standard model support, resulting in inaccurate sensitivity weight acquisition, and further affecting the accuracy of causal correlation analysis.
[0091] To this end, the application further proposes that the sensitivity weight of the service is obtained according to the type of the service, specifically comprising: obtaining service characteristic parameters of an emerging service or a customized service; simulating transmission of the emerging service or the customized service under different combinations of bit error rate, delay and packet loss rate in a controlled network environment, and recording performance and user experience feedback of the service under each combination; constructing a sensitivity curve of the emerging service or the customized service according to the service characteristic parameters, the performance and the user experience feedback; extracting a sensitivity weight of the emerging service or the customized service according to the sensitivity curve.
[0092] Specifically, the service characteristic parameters of the emerging service or the customized service refer to collecting key attributes defined in the design, deployment or initial running of these services. For example, for the emerging real-time interactive VR service, its characteristic parameters can include strict requirements on bandwidth, delay, jitter and packet loss rate, and its sensitivity to user experience of visual fluency and interactive response speed. These parameters provide basic input for subsequent simulation and sensitivity curve construction.
[0093] In the controlled network environment, the transmission of the emerging service or the customized service under different combinations of error rate, delay and packet loss rate is simulated, and the performance and user experience feedback of the service under each combination are recorded, aiming to quantify the response of the service to network quality changes through systematic experiments. The controlled network environment can be a simulation platform or an isolated test network, in which network parameters such as error rate, delay and packet loss rate can be accurately controlled. Performance can include objective indicators such as throughput, success rate and response time, while user experience feedback can be obtained through questionnaire surveys, user behavior analysis or expert evaluation, such as user tolerance to video lag and voice interruption.
[0094] In practical applications, according to the service characteristic parameters, the performance and the user experience feedback, a sensitivity curve of the emerging service or the customized service is constructed, aiming to visualize and quantify the response of the service to network quality changes. The sensitivity curve can be a multi-dimensional function or chart, which depicts the trend of changes in service performance and user experience feedback under different network quality combinations. For example, a surface or curve cluster can be constructed with error rate, delay and packet loss rate as independent variables and service performance degradation or user satisfaction as dependent variables.
[0095] Further, according to the sensitivity curve, the sensitivity weight of the emerging service or the customized service is extracted, aiming to simplify the complex sensitivity characteristics into a numerical value that can be used for weight adjustment. The sensitivity weight can be a single numerical value or a set of weight vectors, which reflects the sensitivity of the service to changes in specific network parameters. For example, the weight can be calculated according to the slope, inflection point or performance degradation amplitude at a specific threshold of the curve, to ensure that in the subsequent calculation of potential contribution value, the real sensitivity of the service to weak fluctuations can be accurately reflected.
[0096] The scheme of the present application solves the problem that the traditional method is difficult to accurately obtain the sensitivity weight when lacking historical data by establishing a set of sensitivity evaluation mechanism for emerging or customized services. Specifically, first, the characteristic parameters of the service are obtained to provide basic information for subsequent simulation; second, the transmission of the service under various network quality conditions is simulated in a controlled environment, and its performance and user experience feedback are recorded comprehensively, thereby obtaining the real response data of the service under different network conditions; then, a curve reflecting the sensitivity characteristics of the service is constructed based on these data, and the complex service sensitivity is quantified; finally, the accurate sensitivity weight is extracted from the sensitivity curve. Due to this data-driven and systematic sensitivity weight acquisition method, even for emerging or customized services lacking historical data, the sensitivity to network fluctuations can be accurately evaluated, thereby providing more reliable input for subsequent causal correlation analysis.
[0097] Through the above technical scheme, the accuracy and robustness of the optical communication network resource intelligent scheduling method in processing emerging or customized services can be significantly improved. Specifically, by finely obtaining the sensitivity weight of these special services, the potential contribution of slight fluctuations to the performance decline of the service can be more accurately evaluated, avoiding misjudgment or omission due to inaccurate sensitivity weight. Thus, in the causal correlation analysis, the root cause of the performance decline of the service can be more accurately identified, thereby making the subsequent path selection preference adjustment more targeted and effective, ultimately improving the resource scheduling efficiency and service quality of the entire optical communication network, especially in the face of evolving and diversified business demands.
[0098] In some preferred embodiments, the following is described by a specific example. Suppose there is an emerging "remote high-precision surgery" service, which is extremely sensitive to network delay and packet loss rate.
[0099] First, the service characteristic parameters of the remote high-precision surgery service are obtained, such as the requirement that the end-to-end delay is less than 5 milliseconds, the packet loss rate is less than 0.001%, and the synchronization of video stream and control signal is extremely high.
[0100] Next, in a controlled network environment, the transmission of the service under different combinations of bit error rate, delay, and packet loss rate is simulated. For example, multiple experiments are set up to simulate the change of delay from 1 millisecond to 50 milliseconds and packet loss rate from 0% to 1%, and the real-time feedback delay, video frame freezing frequency, control instruction loss rate, and other performance data of the surgery operation are recorded. At the same time, professional surgeons are invited to perform simulated operations, and their user experience feedback on operation fluency, feedback timeliness, and visual clarity is collected.
[0101] Then, according to these service characteristic parameters, performance data and user experience feedback, a sensitivity curve of the remote high-precision surgery service is constructed. The curve can intuitively show that when the delay exceeds 10 milliseconds or the packet loss rate exceeds 0.01%, the service performance and user experience will decrease sharply.
[0102] Finally, according to the constructed sensitivity curve, the sensitivity weight of the remote high-precision surgery service is extracted. For example, it can be identified that when the delay and packet loss rate reach a certain threshold, the slope of the service performance decrease is very steep, indicating that its sensitivity to these parameters is extremely high, so it is given a higher sensitivity weight. This weight is then used to adjust the traffic ratio and the weight of the weak fluctuation record in the potential contribution value, ensuring that the special sensitivity of the service to network weak fluctuations can be fully considered when analyzing the causal relationship of the service performance decrease.
[0103] In some embodiments of the above application, a method for obtaining service sensitivity weight according to service type is proposed, in which the sensitivity weight is extracted from the constructed sensitivity curve. However, in actual application, simply extracting the weight from the sensitivity curve may not fully capture the sensitivity change characteristics of the service under different performance decreases, especially when the sensitivity curve presents nonlinearity or has multiple key turning points, which may lead to the extracted weight failing to accurately reflect the real sensitivity of the service to performance decrease, thereby affecting the adjustment accuracy of the subsequent path selection preference.
[0104] To this end, the application further proposes the above method for extracting the sensitivity weight of emerging services or customized services according to the sensitivity curve, which includes the following steps: In the sensitivity curve, identify the performance decrease point corresponding to a certain performance decrease threshold; Starting from the performance decrease point, expand to both sides along the sensitivity curve, and identify a local interval containing the performance decrease point and having relatively stable curve shape; In the local interval, identify the inflection point or the starting point of the platform period of the curve according to the slope, curvature change trend of the sensitivity curve; According to the inflection point or the starting point of the platform period, divide the local interval into multiple subintervals; For each subinterval, calculate the corresponding average sensitivity or dominant sensitivity; According to the average sensitivity or the dominant sensitivity of each subinterval and the position of the certain performance decrease threshold in the local interval, the sensitivity weight is weighted calculated.
[0105] Specifically, identifying the performance degradation point corresponding to a specific performance degradation threshold refers to determining a pre-set performance degradation degree on the constructed sensitivity curve (which usually reflects the relationship between service performance and user experience feedback), such as a certain critical value of bit error rate, a delay exceeding a certain upper limit, etc., and finding the point on the curve corresponding to the threshold. This point is usually considered as the starting point of the significant deterioration of service performance or the obvious damage to user experience.
[0106] Further, from the performance degradation point, expand to both sides along the sensitivity curve, and identify a local interval containing the performance degradation point and having relatively stable curve morphology. This local interval aims to focus on the area near the performance degradation point where the service sensitivity changes most critically. By expanding to both sides, the trend of service sensitivity change before and after the performance degradation point can be captured, and "relatively stable curve morphology" means that the sensitivity change in this interval has certain regularity, which is convenient for subsequent analysis.
[0107] Among the local interval, according to the slope, curvature change trend of the sensitivity curve, the inflection point or the starting point of the platform period of the curve is identified. The inflection point represents the point where the sensitivity change rate changes significantly, such as from slow decline to rapid decline, or vice versa. The starting point of the platform period represents that the sensitivity change tends to be flat and enters a relatively stable stage. These points are crucial for understanding the internal mechanism of service sensitivity, as they mark the transition of service response mode to performance degradation. For example, these points can be identified by calculating the first derivative (slope) and second derivative (curvature) of the sensitivity curve.
[0108] Thus, according to the inflection point or the starting point of the platform period, the local interval is divided into multiple sub-intervals. Each sub-interval represents different sensitivity characteristics of the service within a specific performance degradation range. This division helps to analyze the sensitivity of the service more finely and avoids considering the entire local interval as a single sensitivity mode.
[0109] For each sub-interval, the corresponding average sensitivity or dominant sensitivity is calculated. The average sensitivity can be obtained by averaging the sensitivity values of all points in the sub-interval, while the dominant sensitivity can be determined by identifying the point with the highest sensitivity value or the most drastic change in the sub-interval. Both of these methods aim to provide a representative sensitivity quantitative value for each sub-interval.
[0110] Finally, the sensitivity weight is calculated by weighting the average sensitivity of each sub-interval or the dominant sensitivity, and the position of the specific performance degradation threshold in the local interval. This means that the sensitivity contribution of different sub-intervals will be weighted according to their importance, such as the distance from the performance degradation threshold, the length of the sub-interval, or the size of its sensitivity value, so as to obtain a comprehensive and accurate weight value that reflects the overall sensitivity of the service.
[0111] The scheme of the present application overcomes the limitations that may exist in simple extraction of weight by fine analysis of the sensitivity curve. Specifically, first, the performance degradation point corresponding to the specific performance degradation threshold is identified, so that the analysis can focus on the key area where the service performance begins to be significantly affected. Then, by extending the identification of the local interval, the overall consideration of the sensitivity change before and after the performance degradation point is ensured. In the local interval, by analyzing the slope and curvature change trend of the sensitivity curve to identify the inflection point or the starting point of the plateau period, the transition of the service sensitivity mode can be accurately captured, such as from linear response to nonlinear response, or from rapid deterioration to stable deterioration. These inflection points and starting points of the plateau period serve as key division points to divide the local interval into multiple sub-intervals, so that each sub-interval can represent the unique sensitivity characteristics of the service in a specific performance degradation range. By calculating the average sensitivity or the dominant sensitivity of each sub-interval and combining the position of the specific performance degradation threshold in the local interval for weighted calculation, the scheme can generate a more fine, accurate and representative sensitivity weight. This weighted calculation ensures that the area with the greatest impact on service performance is given a higher weight, so that the final sensitivity weight can more truly reflect the sensitivity of the service to performance degradation.
[0112] Through the above technical scheme, the present application can provide a more accurate and detailed sensitivity weight extraction method. Compared with simply extracting weight from the sensitivity curve, the present application identifies the key performance degradation point, the local interval, the inflection point and the starting point of the plateau period, and performs sub-interval division and weighted calculation, which significantly improves the accuracy and representativeness of the sensitivity weight. This fine analysis makes the obtained sensitivity weight more truly reflect the sensitivity characteristics of the emerging service or the customized service under different performance degradation degrees, especially in the key areas where the performance deteriorates rapidly or tends to be stable. Therefore, in the subsequent path selection preference adjustment, decisions can be made based on more accurate sensitivity weight, so as to realize more intelligent and efficient optical communication network resource scheduling, effectively avoid improper resource allocation or further deterioration of service performance caused by inaccurate sensitivity weight, and ultimately improve user experience and network service quality.
[0113] Specifically, the above-mentioned identification of the inflection point or the starting point of the plateau period of the curve in the local interval according to the slope and curvature change trend of the sensitivity curve comprises: performing a moving average processing on the sensitivity curve within the local interval; calculating a first-order difference and a second-order difference of the moving average processed sensitivity curve; identifying a point with maximum slope change rate or continuous sign change of the second-order difference as the inflection point according to the continuous change trend of the first-order difference and the second-order difference; identifying a region with continuous stability of the first-order difference within a preset small range and a duration exceeding a preset threshold as the plateau starting point.
[0114] The moving average processing on the sensitivity curve within the local interval aims to smooth the curve data and reduce the impact of random noise on subsequent difference calculation, thereby improving the accuracy and stability of the identification of the inflection point and the plateau starting point. The moving average processing can be achieved by selecting a suitable moving window size, for example, a moving average filter can be used to weight and average the data points on the sensitivity curve.
[0115] Further, calculating the first-order difference and the second-order difference of the moving average processed sensitivity curve is a common method for mathematically analyzing the change trend of the curve. The first-order difference reflects the slope of the curve, i.e., the change rate; the second-order difference reflects the change rate of the slope, i.e., the curvature of the curve. These difference data provide quantitative basis for accurately identifying the curve features. For example, the first-order difference can be approximately calculated by the difference between adjacent data points, and the second-order difference can be approximately calculated by the difference of the first-order difference.
[0116] Specifically, according to the continuous change trend of the first-order difference and the second-order difference, a point with maximum slope change rate or continuous sign change of the second-order difference is identified as the inflection point. The inflection point is usually the place where the slope of the curve changes most sharply, which is manifested as a change in sign or a local extreme value of the first-order difference on the second-order difference. By monitoring these mathematical features, the inflection point can be accurately located.
[0117] In addition, a region with continuous stability of the first-order difference within a preset small range and a duration exceeding a preset threshold is identified as the plateau starting point. The plateau represents a stage where the business performance is no longer sensitive to the change of a certain parameter or the sensitivity change trend tends to be flat. On the first-order difference, this is manifested as a slope close to zero and stable. By setting a preset small range (e.g., a threshold close to zero) and a preset duration threshold, the starting point of the plateau can be effectively identified.
[0118] The scheme of the present application effectively suppresses the noise in the sensitivity curve by introducing a moving average process, providing a more smooth and reliable data basis for subsequent mathematical analysis. It is due to the data smoothing process that the calculation of the first-order difference and the second-order difference can more accurately reflect the inherent trend of the curve, rather than being disturbed by random fluctuations. Through accurate calculation of the first-order difference (slope) and the second-order difference (curvature) and analysis of the continuous change trend, the inflection point of the curve, i.e. the point with the maximum slope change rate or the point where the sign of the curvature changes, can be mathematically defined and identified. These points usually correspond to the key transition of the performance response mode in business sensitivity analysis. At the same time, by monitoring the region where the first-order difference is stable within a predetermined small range, the present application can accurately locate the starting point of the plateau, which is crucial to understanding when the business performance enters the saturation or insensitive state. Thus, the scheme overcomes the limitations of traditional methods in noisy environments, such as inaccurate identification or reliance on subjective judgment, and realizes the automation and high-precision identification of key feature points of the sensitivity curve.
[0119] Through the above technical scheme, the present application can significantly improve the identification accuracy and robustness of the inflection point and the starting point of the plateau in the sensitivity curve. Compared with the method of relying only on visual observation or simple trend judgment, the present application uses mathematical tools for quantitative analysis of the curve, effectively filtering out the influence of data noise, and ensuring the objectivity and consistency of the identification results. This precise identification capability makes the subsequent division of local intervals more reasonable, so that the sensitivity weight of the business can be calculated more accurately, providing a more reliable decision basis for intelligent scheduling of optical communication network resources. Ultimately, this helps to optimize the configuration of network resources, improve business performance, and improve user experience.
[0120] In some preferred embodiments, assuming that when constructing the sensitivity curve of a new video service, the curve shows a trend of rapid decline followed by flattening in a certain local interval, but there is certain measurement noise in the data. In order to accurately identify the inflection point of performance decline and the starting point of the plateau entering the plateau, the method proposed by the present application can be used.
[0121] Firstly, the sensitivity curve data of the local interval is subjected to a moving average process, for example, a 5-point moving average filter is used to eliminate high-frequency noise.
[0122] Next, based on the smoothed data, the first-order difference and the second-order difference are calculated. By analyzing the change rate of the first-order difference, it can be found that it reaches a maximum value at a certain point, and the sign of the second-order difference changes near this point, so that the point is accurately identified as the inflection point of performance decline.
[0123] Subsequently, the first-order difference is continuously observed, and when the value thereof remains stable within a preset small range, for example, -0.01 to 0.01, and the duration of the stable state is longer than, for example, 10 data points, the starting point of the region is identified as the starting point of the plateau period. In this way, the key turning points and stable periods of the video service in the performance decline process can be automatically and accurately identified, and fine sensitivity weights can be provided for subsequent adjustment of the resource scheduling strategy.
[0124] In some embodiments of the application described above, a method for identifying the inflection point or the starting point of the plateau period according to the slope and curvature change trend of the sensitivity curve in a local interval is proposed. However, in actual applications, due to the noise or local small fluctuations in network data, multiple similar data points may simultaneously meet the identification conditions of the inflection point or the starting point of the plateau period, thereby introducing the uncertainty of the identification result and affecting the accuracy of the subsequent sensitivity weight calculation. In this regard, the application further proposes an optimization scheme to improve the accuracy and robustness of the identification of the inflection point or the starting point of the plateau period, and to ensure that the identified key points can more accurately reflect the essential changes of the service sensitivity curve.
[0125] After identifying the region where the first-order difference remains stable within a preset small range and the duration is longer than a preset threshold as the starting point of the plateau period, the method further includes: performing consistency verification on the identified inflection point or the starting point of the plateau period, and when multiple similar points meet the identification conditions, selecting the point with the largest change amplitude as the final inflection point or the starting point of the plateau period.
[0126] Specifically, the consistency verification refers to further evaluation and screening of all points preliminarily identified as meeting the inflection point or the starting point of the plateau period conditions. The purpose is to exclude the misidentification caused by data noise or local small fluctuations and to ensure the reliability of the selected points. Among them, the multiple similar points can be understood as in a certain local interval of the sensitivity curve, there are multiple points within a preset distance threshold or time window, which all meet the identification standards of the inflection point or the starting point of the plateau period.
[0127] For example, in a certain sharp change region of the curve, there can be multiple adjacent data points whose first or second order difference reaches or approaches the identification threshold. In practical applications, the point with the largest change amplitude refers to the point among these similar candidate points whose slope change rate (for inflection points) or sensitivity value change (for plateau starting points) of the corresponding sensitivity curve is the most significant. For example, for an inflection point, the absolute value of its second order difference can be calculated, and the point with the largest absolute value is selected; for a plateau starting point, the change amount of its sensitivity values before and after can be investigated, and the point with the largest change amount is selected. The purpose is to identify the key turning point that has the most significant impact on business performance. Through the above verification and selection process, a most representative and influential point is finally determined as the inflection point or plateau starting point of the local interval, which is used for subsequent sensitivity weight calculation.
[0128] The scheme of the present application effectively solves the problem that multiple similar points can simultaneously meet the identification conditions of inflection points or plateau starting points in a complex data environment by introducing a consistency verification mechanism. Specifically, when multiple potential inflection points or plateau starting points are initially identified, the verification mechanism further evaluates the "importance" or "significance" of these points. By selecting the point with the largest change amplitude as the final inflection point or plateau starting point, it is ensured that the identified key point is the turning point that has the most significant impact on the change of business sensitivity in the local interval. Thus, false positives caused by noise or minor fluctuations are avoided, enabling subsequent division of the sensitivity curve and calculation of sensitivity weights to be based on more accurate and representative key points.
[0129] Through the above technical scheme, the present application can significantly improve the accuracy and robustness of inflection point or plateau starting point identification. Specifically, by performing consistency verification on the identified candidate points and selecting the point with the largest change amplitude, the interference of data noise and local minor fluctuations on the identification result is effectively eliminated, ensuring that the identified key point can more accurately reflect the essential characteristics of the business sensitivity curve. This makes the subsequent sensitivity curve division more reasonable, enabling more accurate sensitivity weights to be extracted, providing more reliable decision basis for intelligent scheduling of optical communication network resources, and further optimizing the accuracy of resource allocation and business path selection.
[0130] In some preferred embodiments, it is assumed that when analyzing the sensitivity curve of a new business, three adjacent data points A, B and C in a local interval are preliminarily identified to meet the identification condition of inflection point by calculating the first-order difference and the second-order difference. The absolute value of the second-order difference of point A is 0.05, the absolute value of the second-order difference of point B is 0.12, and the absolute value of the second-order difference of point C is 0.06. According to the scheme of the present application, consistency verification will be performed on the three points. Since the absolute value of the second-order difference of point B is the largest (0.12), which indicates that the corresponding curve slope changes most sharply, point B will be selected as the final inflection point of the local interval. In this way, even if there are multiple potential inflection points, the most representative key turning point can be accurately identified, avoiding the calculation deviation of sensitivity weight caused by selecting secondary points.
[0131] Reference Figure 3 , Figure 3 is a schematic diagram of an optical communication network resource intelligent scheduling system provided by an embodiment of the present application, comprising: an input end for acquiring device running data, scheduling decision data and service performance data and performing time synchronization; identifying weak fluctuations below a preset alarm threshold and performing risk marking according to the device running data after time synchronization; an analysis end for acquiring the service identifier and the time point of performance decline when monitoring the performance decline of the service, backtracking the optical path passed by the service, and associating and analyzing the causal relationship between the weak fluctuation record and the performance decline according to the optical path, the risk marking, the scheduling decision data and the service performance data; an adjustment end for adjusting the subsequent path selection preference of the service according to the causal association.
[0132] The conventional existing optical communication network intelligent scheduling system mainly relies on the real-time running state data reported by the device when performing resource allocation. However, these data may have weak and intermittent physical layer performance fluctuations. These fluctuations are often below the preset alarm threshold due to their small amplitude and short duration, which leads to the fact that the scheduling system cannot perceive and quantify their potential impact on service transmission quality, so that high-value services may be allocated to paths with potential risks, ultimately leading to service performance decline and difficulty in diagnosis.
[0133] To this end, the present application proposes an optical communication network resource intelligent scheduling system. Through its input end, device operation data, scheduling decision data and service performance data are obtained and time synchronization is performed. According to the time-synchronized device operation data, weak fluctuations below the preset alarm threshold are identified and risk labels are added. When the service performance is monitored to decline, the analysis end can obtain the service identifier and the time point of performance decline, backtrack the optical path path passed by the service, and comprehensively analyze the optical path path, risk label, scheduling decision data and service performance data to associate and analyze the causal relationship between the weak fluctuation record and the performance decline. Thus, the adjustment end can adjust the subsequent path selection preference of the service according to the causal association, thereby effectively avoiding potential risks and improving the stability and reliability of service transmission. Through the modular design, the data acquisition and preliminary processing, the core analysis and causal association, and the final scheduling strategy adjustment function are clearly divided, ensuring the professionalism and synergy of each functional module, and realizing the intelligent and risk-avoiding scheduling of optical communication network resources.
[0134] In order to better understand the optical communication network resource intelligent scheduling system proposed in the present application, some key components involved therein are described in detail below.
[0135] The input end is configured to obtain device operation data, scheduling decision data and service performance data and perform time synchronization. Specifically, the input end can be a data acquisition module that periodically acquires various device operation data such as optical power, optical signal-to-noise ratio (OSNR), bit error rate (BER), etc. through interface communication with the network management system (NMS), performance monitoring agent or directly with the optical network equipment. At the same time, the input end can also read scheduling decision data from the database of the scheduling system and obtain service performance data from the service performance probe or service server. In order to ensure the accuracy of data analysis, the input end is also responsible for time synchronization of all the obtained data, such as calibrating the time stamps of different data sources through high-precision time protocol (such as NTP or PTP).
[0136] Further, the input end is also configured to identify weak fluctuations below the preset alarm threshold and add risk labels according to the time-synchronized device operation data. For example, the input end can include a data preprocessing unit that performs sliding window average processing on the collected optical power data and calculates the deviation between each data point and the sliding average. When the deviation continuously exceeds a small preset threshold (e.g. 0.2 dB) but is below the regular alarm threshold (e.g. 0.5 dB), the unit identifies it as a weak fluctuation. For the identified weak fluctuations, the input end can assign different risk labels such as "high risk" or "low risk" according to factors such as their duration, amplitude, occurrence frequency and device type involved, and store these weak fluctuation records with risk labels for subsequent analysis.
[0137] Analysis end: configured to obtain the performance degradation of the service identification and the time point when the performance degradation is monitored, trace back the optical path path passed by the service, and analyze the cause and effect relationship between the weak fluctuation record and the performance degradation according to the optical path path, the risk mark, the scheduling decision data and the service performance data. Specifically, the analysis end can include a performance monitoring module, which continuously analyzes the service performance data. For example, when the bit error rate of a certain service increases significantly within a short time, or the delay exceeds the preset service level agreement (SLA) threshold, the module can determine that the performance is degraded, and record the unique identification of the service and the accurate time point of the performance degradation.
[0138] Further, the analysis end also includes a path tracing module, which identifies the optical path path passed by the service during the performance degradation by querying the historical path allocation record of the scheduling system. For example, the module can query the specific optical path path to which the service is allocated in different time periods according to the service identification and the time point. After determining the optical path path, the core function of the analysis end is its cause and effect analysis module. The module will combine the weak fluctuation record (with risk mark) of all devices on the path, the historical scheduling decision data and the performance data of the service itself for comprehensive correlation analysis. For example, if there is a weak fluctuation marked as "high risk" on the path passed by the performance degraded service at a certain time point, and the fluctuation is highly coincident with the time point of the service performance degradation, the module can preliminarily determine that there is a cause and effect relationship between the two. The module can use statistical analysis, machine learning model or expert system rule to identify the cause and effect relationship.
[0139] Adjustment end: configured to adjust the path selection preference of the service in the future according to the cause and effect relationship. Once the analysis end determines the cause and effect relationship between the weak fluctuation and the performance degradation of the service, the adjustment end will use this information to optimize the future scheduling decision. Specifically, the adjustment end can include a strategy updating module, which dynamically updates the path selection preference in the scheduling algorithm according to the result of the cause and effect analysis. For example, if it is found that the weak fluctuation on a certain ROADM device leads to the performance degradation of a specific service, then when selecting the path for the service or similar sensitive services in the future, the strategy updating module will reduce the priority of the path containing the ROADM device, or temporarily list it in the "blacklist". This can be realized by introducing a "risk weight" or "cost factor" in the scheduling algorithm, so that the scheduling algorithm tends to select other more stable paths when optimizing the objective function. This adjustment can be dynamic, continuously updated as the network state changes and new weak fluctuation records appear, ensuring that the scheduling system can continuously learn and adapt to changes in the network environment.
[0140] The optical communication network resource intelligent scheduling system provided in the present application effectively solves the problem that the traditional scheduling system cannot perceive and avoid hidden network risks by introducing the identification of weak fluctuations below the preset alarm threshold, risk marking, and scheduling adjustment mechanism based on causal association.
[0141] Compared with the prior art, the system of the present application has significant progress. The traditional optical communication network scheduling system usually relies on explicit alarms and conventional performance indicators reported by devices for resource allocation, and its system architecture often lacks special modules to process and analyze those weak fluctuations below the alarm threshold. When these hidden risks occur in the network, the traditional system cannot identify them, resulting in scheduling decisions based on incomplete or biased data, which may allocate high-value services to paths with potential performance risks. This not only leads to a decline in service performance, but also makes fault location and resolution extremely difficult, seriously affecting the quality of service and operation and maintenance efficiency of the network.
[0142] The system of the present application realizes fine management of weak fluctuations and causal association analysis through its clearly divided input end, analysis end and adjustment end. The input end is responsible for comprehensive and detailed data collection and identification of weak fluctuations, providing accurate risk information for subsequent analysis. The analysis end focuses on complex causal association analysis and can accurately locate the root cause of the decline in service performance. The adjustment end dynamically and intelligently adjusts the scheduling strategy based on the analysis results to actively avoid risky paths. The innovation of this system architecture lies in its ability to perceive and quantify the "physical reality" and embed it into the scheduling decision-making process, enabling the system to fundamentally avoid allocating services to paths with potential risks. In this way, the system of the present application can significantly improve the utilization efficiency of network resources and the stability of service transmission, reduce the complexity of manual intervention and fault troubleshooting, and thus provide users with higher quality and more reliable optical communication services.
[0143] The above only describes the embodiments of the present application and does not limit the protection scope of the present application. For those skilled in the art, the present application can have various modifications and changes. Any modification, equivalent replacement, improvement, etc. made within the spirit and principles of the present application shall be included in the protection scope of the present application.
Claims
1. A method for intelligent scheduling of resources in an optical communication network, characterized in that, include: Acquire equipment operation data, scheduling decision data, and business performance data, and synchronize them with time. Based on the time-synchronized device operation data, identify and mark the slight fluctuations below the preset alarm threshold. When a performance degradation of a service is detected, the service identifier and the time of occurrence of the performance degradation are obtained, the optical path of the service is traced back, and the causal relationship between the weak fluctuation records and the performance degradation is analyzed based on the optical path, the risk marker, the scheduling decision data, and the service performance data. Based on the causal relationship, the path selection preference of the service is adjusted.
2. The intelligent scheduling method for optical communication network resources according to claim 1, characterized in that, The step of analyzing the causal relationship between the weak fluctuation records and the performance degradation based on the optical path, the risk marker, the scheduling decision data, and the service performance data includes: Acquire optical power, optical signal-to-noise ratio, and radio frequency interference intensity data of the devices along the optical path during the time period of performance degradation; Assess whether the device performs zero-point calibration; If so, an instruction is sent to the device to switch to idle state or standby bypass within a preset time window; Within the preset time window, the optical power and optical signal-to-noise ratio data of the device are collected, and the radio frequency interference intensity data are collected simultaneously. The contribution ratio to the weak fluctuations is quantified based on the device data collected during the zero-point calibration and the radio frequency interference intensity data. The causal relationship is established based on the optical path, the risk marker, the scheduling decision data, the service performance data, and the contribution ratio.
3. The intelligent scheduling method for optical communication network resources according to claim 2, characterized in that, When a performance degradation of a service is detected, the process of obtaining the service identifier and the time of occurrence of the performance degradation, and tracing back the optical path traversed by the service, includes: When a performance degradation of the service is detected, a path backtracking procedure is triggered based on the service identifier and the time of occurrence. Obtain the allocation record of the service, the allocation record including the allocation path, wavelength channel and precise timestamp; Obtain the path distribution information of the actual transmission traffic of the service within the preset time window, the path distribution information including the traffic ratio and timestamp of each path; By comparing the allocation record with the path distribution information of the actual transmission traffic, the switching time point and the path after the switching of the traffic between multiple optical paths within the preset time window are identified. Based on the switching time point and the switched path, construct a time series distribution map of the traffic on each optical path within the preset time window; Based on the time series distribution map and the duration of the performance degradation of the service, the optical path carrying the main traffic of the service is taken as the optical path traversed by the service.
4. The intelligent scheduling method for optical communication network resources according to claim 3, characterized in that, The step of determining the optical path carrying the main traffic of the service as the optical path traversed by the service, based on the time series distribution map and the duration of the service performance degradation, includes: During the period of performance degradation of the service, the traffic ratio of each optical path in the time series distribution diagram is divided into fine-grained time windows, and the traffic ratio within each time window is statistically analyzed. Identify all overlapping optical paths during the duration of performance degradation of the service, and obtain the corresponding weak fluctuation record for each overlapping optical path; Based on the statistical results and the recorded slight fluctuations, calculate the potential contribution of each overlapping optical path to the performance degradation of the service. The overlapping optical path with the highest potential contribution value is identified as the optical path.
5. The intelligent scheduling method for optical communication network resources according to claim 4, characterized in that, The step of calculating the potential contribution of each overlapping optical path to the performance degradation of the service based on the statistical results and the weak fluctuation records includes: Based on the type of the service, obtain the sensitivity weight of the service; Based on the sensitivity weight, the weight of the traffic ratio and the slight fluctuation recorded in the potential contribution value are dynamically adjusted; The adjusted weights are applied to the traffic ratio and the weak fluctuation record to calculate the potential contribution of each overlapping optical path to the performance degradation of the service.
6. The intelligent scheduling method for optical communication network resources according to claim 5, characterized in that, The step of obtaining the sensitivity weight of the service based on its type includes: Obtain business characteristic parameters for emerging or customized businesses; In a controlled network environment, the transmission of the emerging service or the customized service is simulated under different combinations of bit error rate, latency, and packet loss rate, and the performance and user experience feedback of the service under each combination are recorded. Based on the business characteristic parameters, performance, and user experience feedback, a sensitivity curve for the emerging business or the customized business is constructed. Based on the sensitivity curve, extract the sensitivity weights of the emerging business or the customized business.
7. The intelligent scheduling method for optical communication network resources according to claim 6, characterized in that, The step of extracting the sensitivity weights of the emerging business or the customized business based on the sensitivity curve includes: In the sensitivity curve, identify the performance degradation point corresponding to a specific performance degradation threshold; Starting from the performance degradation point, expand along the sensitivity curve to both sides to identify a local interval that contains the performance degradation point and has a relatively stable curve shape. Within the local interval, the inflection point or the starting point of the plateau period of the curve is identified based on the slope and curvature change trend of the sensitivity curve. Based on the inflection point or the starting point of the plateau period, the local interval is divided into multiple sub-intervals; For each sub-interval, calculate the corresponding average sensitivity or dominant sensitivity; The sensitivity weight is calculated by weighting the average sensitivity or the dominant sensitivity of the sub-interval and the position of the specific performance degradation threshold within the local interval.
8. The intelligent scheduling method for optical communication network resources according to claim 7, characterized in that, Within the local interval, identifying the inflection point or plateau start point of the sensitivity curve based on the slope and curvature change trend includes: Within the local interval, the sensitivity curve is subjected to a moving average process; Calculate the first and second differences of the sensitivity curve after the moving average processing; Based on the continuous changing trend of the first-order difference and the second-order difference, the point where the slope change rate is the largest or the sign of the second-order difference changes continuously is identified as the inflection point. The region where the first-order difference remains stable within a preset small range and lasts for a duration exceeding a preset threshold is identified as the starting point of the plateau period.
9. The intelligent scheduling method for optical communication network resources according to claim 8, characterized in that, After identifying the region where the first-order difference remains stable within a preset small range for a duration exceeding a preset threshold as the starting point of the plateau period, the method further includes: Consistency verification is performed on the identified inflection point or the starting point of the plateau period. When multiple similar points meet the identification conditions, the point with the largest change is selected as the final inflection point or the starting point of the plateau period.
10. An intelligent scheduling system for optical communication network resources, characterized in that, include: The input terminal is used to acquire equipment operation data, scheduling decision data, and business performance data, and to synchronize them with time. Based on the time-synchronized device operation data, identify and mark the slight fluctuations below the preset alarm threshold. The analysis unit is used to obtain the service identifier and the time point of occurrence when a service performance degradation is detected, trace back the optical path of the service, and analyze the causal relationship between the slight fluctuation and the performance degradation based on the optical path, the risk marker, the scheduling decision data, and the service performance data. The adjustment terminal is used to adjust the path selection preference of the service based on the causal relationship.
Citation Information
Patent Citations
Data change identification method and device
CN110288003A
Network equipment light attenuation intelligent monitoring system and method based on AI algorithm
CN119814140A
Data transmission method and electronic equipment
CN119946728A
Winch safety performance evaluation method
CN119976685A
Data remote transmission method and system for intelligent force transducer based on Internet of Things
CN120880613A