Telecom Network Root Cause Analysis for Large-Scale Events
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Operational support systems (OSSs) and network operation centers (NOCs) face challenges in efficiently identifying the root cause of large-scale events (LSEs) in telecommunication networks, which are complex and span tens of thousands of cell sites, leading to time-consuming and labor-intensive processes for NOC technicians.
Innovation Solution
An incident management tool that analyzes data from disparate systems to rapidly identify the root cause of LSEs by correlating alarms with cell site attributes, presenting information in a segmented user interface, and automatically assigning incident reports to responsible parties, potentially completing the analysis in under five minutes.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If manual analysis methods are used by NOC technicians to identify root causes of large-scale events, then comprehensive analysis can be performed, but the process becomes time-consuming and labor-intensive
Solution Approach 1:
The system performs automated root cause analysis without requiring manual intervention from NOC technicians. The incident management application automatically retrieves alarm data, analyzes correlations, identifies root causes, and generates incident reports, enabling the system to serve itself in the analysis process and eliminating manual labor while maintaining comprehensive analysis capabilities
Solution Approach 2:
The patent replaces the mechanical manual analysis process with an automated electronic system. The incident management application uses computer-based algorithms to retrieve, correlate, and analyze alarm data from multiple sources, substituting human technicians' manual work with automated computational processes that are both faster and equally comprehensive
2Measurement precision
If comprehensive alarm data from multiple sources is analyzed to identify root causes, then accurate identification is achieved, but system complexity increases
Solution Approach 1:
The incident management application serves multiple functions within a single system: it retrieves alarm data from multiple sources, correlates alarms, analyzes data to identify root causes, generates incident reports, and assigns reports to appropriate groups. This multi-functionality consolidates what would otherwise require multiple separate systems into one universal platform, managing complexity while maintaining comprehensive analysis
Solution Approach 2:
The incident management application acts as an intermediary between various alarm sources and the incident reporting system. It retrieves data from multiple operational support systems, processes and correlates the information, then interfaces with the incident reporting system to create and assign incident reports, simplifying the interaction between complex subsystems
3Productivity
If automated incident management application is implemented to rapidly identify root causes, then analysis time is reduced, but implementation complexity increases
Solution Approach 1:
The system performs preliminary actions by pre-configuring the incident management application with the necessary interfaces and data retrieval capabilities before incidents occur. The application is预先 set up to automatically retrieve alarm data from multiple operational support systems, correlate alarms, and identify root causes, so that when incidents occur, the automated analysis can immediately begin without additional configuration complexity
Data Source
AI summary
A telecommunication network management system. The system comprises an incident reporting application that creates incident reports pursuant to alarms on network elements of a telecommunication network and wherein one of the incident reports is associated with a large-scale event (LSE), wherein the LSE incident report identifies alarms at a plurality of different network elements as associated with the LSE; and an incident management application that analyzes attributes of cell sites identified in the LSE incident report as impacted by the LSE, determines that at least 75% of the cell sites receive backhaul service from a same alternative access vendor (AAV) and that at least one backhaul circuit of the at least 75% of the cell sites is in an alarmed state, causes the incident reporting application to record a root cause of the LSE incident report as an AAV fault.


