Cluster operation and maintenance state diagnosis method, operation and maintenance monitoring system, terminal and storage medium

A state diagnosis and cluster technology, applied in the Internet field, can solve the problems of not being able to know the health status of the cluster in time, not being able to know the optimal configuration of the cluster, and the performance cluster not being able to exert the best performance, so as to improve the efficiency of cluster operation and maintenance, improve Operation and maintenance efficiency and success rate, and the effect of reducing operation and maintenance risks

CN111880993AActive Publication Date: 2020-11-03PING AN TECH (SHENZHEN) CO LTD
6 Cites 1 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Publication Date
2020-11-03

Smart Images

  • Figure 1
    Figure 1
  • Figure 2
    Figure 2
  • Figure 3
    Figure 3
Patent Text Reader

Abstract

The invention discloses a cluster operation and maintenance state diagnosis method which comprises the steps: acquiring the current actual value of at least one piece of cluster operation and maintenance data, wherein the cluster operation and maintenance data comprises at least one piece of daily operation and maintenance data, cluster monitoring data, use specification data and cluster log data;obtaining the operation and maintenance data volume of the cluster according to the at least one piece of cluster data, and adjusting at least one preset threshold value and a preset repair suggestion according to the operation and maintenance data volume; obtaining a standard threshold range corresponding to each piece of cluster operation and maintenance data, and judging whether the current actual value of each piece of cluster operation and maintenance data exceeds the corresponding standard threshold range or not; if yes, obtaining a corresponding final repair suggestion according to thecurrent actual value; and displaying a diagnosis result, wherein the diagnosis result comprises the name and meaning of each piece of cluster operation and maintenance data, whether the name and meaning exceed a corresponding standard threshold range and / or repair suggestions. The invention further discloses an operation and maintenance monitoring system, a terminal and a storage medium. The cluster operation and maintenance efficiency can be improved, and the operation and maintenance risk is reduced.
Need to check novelty before this filing date? Find Prior Art

Description

technical field

[0001] The invention relates to the field of the Internet, in particular to a cluster operation and maintenance state diagnosis method, an operation and maintenance monitoring system, a terminal, and a storage medium. Background technique

[0002] With the increasing use of commercial big data search scenarios, Elasticsearch has become the best choice among open source search engines with its excellent performance, rich functions, and complete ecosystem features. Therefore, the monitoring of the ES (Elasticsearch) cluster has become an important IT (Internet Technology, Internet technology) operation and maintenance work content to ensure the stable operation of the business.

[0003] In conventional ES operation and maintenance, usually only the infrastructure layer monitoring indicators of the cluster are collected, such as CPU (central processing unit, central processing unit) usage, memory usage, disk IO (Input / Output, input / output) interface usage, etc. ...

Examples

Embodiment Construction

[0023] The following will clearly and completely describe the technical solutions in the embodiments of the present invention with reference to the accompanying drawings in the embodiments of the present invention. Obviously, the described embodiments are only some, not all, embodiments of the present invention. Based on the embodiments of the present invention, all other embodiments obtained by persons of ordinary skill in the art without creative efforts fall within the protection scope of the present invention.

[0024] In conventional ES operation and maintenance, usually only the infrastructure layer monitoring indicators of the cluster are collected, such as CPU usage, memory usage, and disk IO interface usage. Due to the complexity of the ES technology stack, the lack of monitoring indicators in a dedicated dimension will make it impossible to know the health status of the cluster in a timely manner, and it is impossible to know the optimal configuration for the current ...