Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

692 results about "High availability" patented technology

High availability (HA) is a characteristic of a system, which aims to ensure an agreed level of operational performance, usually uptime, for a higher than normal period. Modernization has resulted in an increased reliance on these systems. For example, hospitals and data centers require high availability of their systems to perform routine daily activities. Availability refers to the ability of the user community to obtain a service or good, access the system, whether to submit new work, update or alter existing work, or collect the results of previous work. If a user cannot access the system, it is – from the users point of view – unavailable. Generally, the term downtime is used to refer to periods when a system is unavailable.

Processing environment switching and recovering method and device, equipment and medium

PendingCN121092357AFault responseRecovery methodMulti source data
The invention relates to the technical field of artificial intelligence, can be applied to business scenes such as financial science and technology and medical health, and discloses a processing environment switching and recovery method, device, equipment and medium. The method comprises the steps that multi-source heterogeneous data in a main processing environment and a standby processing environment are acquired, and the system fault probability is obtained through multi-model collaborative prediction; a dynamic threshold value is generated in combination with a historical service period mode and a real-time service load, when the fault probability exceeds the threshold value, a switching strategy is generated based on the fault scene knowledge base and the service priority, and flow scheduling between the main processing environment and the standby processing environment is executed; and monitoring the business index of the standby processing environment during the scheduling period, and triggering the fusing rollback when the business index is lower than the health standard. According to the method, the fault identification precision is improved through multi-source data fusion and multi-model prediction, adaptive scheduling is realized in combination with a dynamic threshold and a switching strategy, and fusing rollback is triggered to guarantee high availability and data consistency, so that the continuity and stability of key services are enhanced.
Owner:CHINA PING AN PROPERTY INSURANCE CO LTD

Server cluster monitoring system based on multi-node collaboration and implementation method thereof

The invention relates to a server cluster monitoring system based on multi-node collaboration and an implementation method thereof, a dynamic topology network module is configured to reconstruct a connection topology among monitoring nodes in real time according to node performance and link quality, support mixed configuration of a star type, a ring type and a net structure, and realize multi-node collaboration. Multi-dimensional data capture from a physical layer to an application layer is realized through a cross-level index acquisition module based on an integrated hardware sensor interface and a virtualization layer probe, and each node is enabled to perform collaborative reasoning through parameter encryption sharing through a decision model based on federated learning. A monitoring task fragmentation strategy is dynamically adjusted through an adaptive elastic fragmentation unit according to network delay and load fluctuation, and an abnormal event association rule base is updated in real time through an incremental knowledge graph construction unit. High availability and elastic expansion are realized through a multi-node collaborative architecture, the monitoring efficiency is improved in combination with dynamic load balancing and hybrid detection, and an intelligent multi-level response mechanism is constructed to guarantee the service continuity.
Owner:四川华鲲振宇智能科技有限责任公司

Reliable communication system based on publishing and subscribing mode in satellite-ground weak network environment

The invention discloses a reliable communication system based on a publishing and subscribing mode in a satellite-ground weak network environment, and aims to solve the problems of high round-trip delay, limited link bandwidth, high instantaneous packet loss rate, easiness in blocking and switching influence of links and the like in a data transmission process between a low-orbit satellite and a ground station. A publishing and subscribing model is introduced into an application layer to realize service logic decoupling, deep fusion with an underlying enhanced QUIC transmission protocol is realized, and mechanisms such as message persistence, breakpoint resume, forward error correction and connection migration are combined, so that stable, orderly, complete and safe transmission of service data can still be ensured under the conditions of link instability, quality fluctuation and communication interruption, and the service data transmission efficiency is improved. And high availability and continuity of a satellite-ground communication task under a weak network condition are realized.
Owner:ZHEJIANG LAB

Cloud computing extension cluster high availability method based on dynamic fault domain and intelligent scheduling

The invention relates to a cloud computing extension cluster high availability method based on a dynamic fault domain and intelligent scheduling, and relates to the field of cloud computing extension cluster high availability. According to the method, hardware health data and network performance indexes of physical nodes are collected in real time, and the fault correlation degree between the nodes is calculated based on the hardware health similarity and a network topology attenuation factor; dynamically generating and updating a logic fault domain topological structure according to the fault correlation degree and a preset threshold value; based on the logic fault domain topological structure and a preset SLA strategy library, executing a virtual machine scheduling decision of multi-objective optimization so as to balance business service quality, fault domain risk and resource cost; and in response to the detected fault event, triggering a corresponding hierarchical migration process according to the service priority and the fault level. According to the method, closed-loop linkage of real-time reconstruction and intelligent scheduling of the dynamic fault domain in the cross-region cloud computing extension cluster can be realized, and the disaster tolerance capability and the resource utilization rate of the system are remarkably improved.
Owner:JINAN INSPUR DATA TECH CO LTD

Cloud platform-based computing power resource dynamic scheduling and monitoring method

The invention relates to a computing power resource dynamic scheduling and monitoring method based on a cloud platform. The method is suitable for an intelligent scheduling scene of a high-performance GPU cluster. The method comprises seven steps of task portrait modeling, GPU node state acquisition, resource trend prediction, SLA tracking, scheduling scoring and deployment, operation monitoring and task migration, and SLA feedback optimization. According to the system, task semantics are represented by constructing task vectors, node health states, topology affinity, SLA historical performance conditions and resource prediction risks are fused, a multi-factor adjustable scheduling scoring mechanism is constructed, and second-level perception and task thermal migration of high-temperature nodes are achieved. Compared with a traditional Kubernetes static scheduling scheme, the method has the advantages that the GPU utilization rate, the task SLA achievement rate and the system stability are remarkably improved, the learning ability, the self-adaptive ability and the high availability are achieved, and the method is an intelligent scheduling closed-loop system oriented to AI reasoning and training scenes.
Owner:北京娱广科技有限公司

Cross-data center fault isolation and switching method and system in multi-tenant environment

The invention relates to the technical field of data center high availability, in particular to a cross-data center fault isolation and switching method and system in a multi-tenant environment. The resource mapping table is constructed by taking the tenants as the minimum control units, the affected tenants are accurately identified, logic isolation is executed, the problem of waste of full-tenant service migration resources caused by node-level or cluster-level switching in the prior art is avoided, and the fault influence range is minimized. The optimal data center is dynamically selected by combining the multi-factor objective function with the tenant SLA level, the defects that an existing scheduling strategy is opaque and tenant priorities cannot be distinguished can be overcome, and it is ensured that key tenant services are preferentially recovered. And finally, through combination of a gray takeover mechanism and health feedback confirmation and multi-dimensional recovery verification after migration, the condition that an existing recovery mechanism is extensive can be changed, the fault processing precision and the resource scheduling efficiency of the data center in a multi-tenant environment are remarkably enhanced, and the service continuity is guaranteed.
Owner:SHANGHAI DATA SOLUTION

Financial machine room AI operation and maintenance method and device and readable storage medium

The invention provides a financial machine room AI operation and maintenance method and device and a readable storage medium. The method comprises the steps of obtaining multi-modal data; preprocessing the multi-modal data to obtain preprocessed data; transform depth feature fusion is carried out on the preprocessed data, and a fusion feature vector is obtained; according to the fusion feature vector, LSTM time sequence prediction and evaluation are carried out, a prediction result is obtained, and the prediction result comprises at least one of the equipment health degree score, the fault occurrence probability, the business influence degree and the residual service life of the equipment. According to the application, the physical state monitoring of the equipment and the real-time state of the financial service are deeply fused, so that the fundamental transformation from the traditional passive response to the active prevention is realized. According to the invention, a business-aware intelligent prediction model and a transaction-aware-free resource scheduling mechanism are constructed especially for the high availability requirement and the transaction continuity guarantee requirement of a financial machine room.
Owner:CHINA UNITED NETWORK COMM GRP CO LTD

Cross-domain switching engine equipment load balancing method based on dynamic consistency hash

The invention relates to a cross-domain exchange engine equipment load balancing method based on dynamic consistency hash, and belongs to the technical field of load balancing. According to the method, a dynamic concept is introduced to enable a consistency Hash algorithm based on virtual nodes to adapt to a software and hardware load environment of cross-domain switching engine equipment, and key parameters such as a CPU utilization rate, a memory utilization rate, a disk IO rate, a network bandwidth occupancy rate and network time delay of the cross-domain switching engine equipment are comprehensively considered to reflect a current service load condition of the cross-domain switching engine equipment; therefore, requests are dynamically allocated to cross-domain exchange engine equipment on a hash ring, load balancing is achieved, request delay of a shared exchange system is effectively reduced, the system has higher availability, and the service capacity and the resource utilization rate of the system are effectively improved.
Owner:BEIJING INST OF COMP TECH & APPL

Test method, device and equipment and computer readable storage medium

The invention discloses a test method, device and equipment and a computer readable storage medium, which are applied to the technical field of computers, and comprise the following steps: obtaining multi-source data of a distributed storage cluster, the multi-source data comprising hardware state data, service associated operation data and business dependence data; constructing a relation dependency graph among the hardware nodes, the service nodes and the service nodes based on the multi-source data; adjusting the chaos test script according to the real-time state of the distributed storage cluster and the dynamic adjustment strategy to obtain a target chaos test script; and executing the target chaos test script, determining a fault range and a recovery mechanism based on the relation dependency graph, and outputting a test report. According to the method, comprehensive and accurate presentation of the dependency relationship of each node is realized, and the chaos test script can adapt to the dynamic change of the cluster, so that the test scene is more practical, the invalid test is avoided, the effectiveness, reliability and efficiency of the cluster chaos test are remarkably improved, and the high availability and service continuity of the cluster are powerfully guaranteed.
Owner:JINAN INSPUR DATA TECH CO LTD

OpenRAN intelligent dynamic CU-UP scaling solution

A system is disclosed for providing Open RAN CU-UP high availability, the system comprising: at least one active CU-CP; at least one active CU-UP in communication with the at least one active CU-CP; and at least one standby CU-UP in communication with the at least one active CU-CP; wherein when a message may be received from a CU-CP that detects a failure of the at least one active CU-UP, the at least one standby CU-UP may be configured to take over and become an active CU-UP, thereby providing failover redundancy for the at least one active CU-UP.
Owner:PARALLEL WIRELESS INC

Route preference based on link performance

Techniques are disclosed for computing a priority of routes advertised by nodes implementing Layer-3 (L3) Multi-Node High Availability (MNHA) for a Software-Defined Wide Area Network (SD-WAN) interconnecting a first network device and a second network device according to link adherence to performance requirements. In one example, a node computes a priority for a route to the second network device based at least in part on a comparison between (1) one or more performance measurements of a link between the node and the second network device and (2) one or more performance requirements for the link. In some examples, the computed priority for the route is further based in part on a preference for the node. The node exports, to the first network device, the route to the second network device, wherein the route comprises data specifying the computed priority for the route.
Owner:JUNIPER NETWORKS INC

Key value storage system indexing method oriented to NVM-NVMe SSD hybrid architecture

The invention discloses a key value storage system indexing method oriented to an NVM-NVMe SSD (Non-Volatile Memory-Non-Volatile Memory Express Solid State Disk) hybrid architecture, which comprises the following steps of: constructing a heterogeneous storage architecture taking cold and hot data perception as a core driving mechanism, coordinating and managing two types of storage media, namely a non-volatile memory NVM and a solid state disk NVMe SSD, and realizing efficient identification, layered writing and dynamic migration of cold and hot data. According to the method, the characteristics of low delay, durability and byte addressing of the NVM are utilized, the frequency and I / O overhead of Flush and Compaction operations are reduced in a data write-in path, meanwhile, the problem of mixed storage of cold and hot data is avoided, and therefore the overall performance and storage efficiency of a system are remarkably improved. In addition, by introducing an asynchronous migration module, the method can dynamically adapt to the change of data popularity along with time evolution, and effectively support high performance and high availability of the key value system in long-term operation.
Owner:ANHUI UNIV

Data center interconnection link fault self-recovery and path switching method

The invention provides a data center interconnection link fault self-healing and path switching method, which comprises the following steps of: acquiring multi-dimensional link performance monitoring data, establishing a standardized time sequence database, realizing link health degree multi-index trend prediction by combining an LSTM-Attention model, and determining the link health degree according to the LSTM-Attention model. Risk-aware dynamic path selection and switching for multiple business scenes are realized by combining business service sensitivity modeling, path inherent stability quantification, nonlinear adaptive switching criteria and a debounce switching process, the response accuracy and the business adaptation degree of path switching are improved, false triggering and service interruption risks can be remarkably reduced, and the path switching efficiency is improved. And the high availability and robustness of routing in a complex network environment are enhanced.
Owner:AVIC CLOUD SOFTWARE (GUANGZHOU) CO LTD

AI model intelligent training and reasoning integrated method and system

The invention provides an AI model intelligent training and reasoning integration method and system, and the method comprises the following steps: receiving a model training instruction, and carrying out the preprocessing of original data, and obtaining a training data set; executing model training based on the training data set, monitoring task priorities and resource requirements through a dynamic resource scheduling algorithm, and dynamically adjusting training resource allocation according to a real-time monitoring result; when the model reaches a preset performance index, performing model pruning and quantification to generate an optimization model, and performing parameter fine tuning on the optimization model to obtain a final deployment model; generating reasoning service configuration according to calculation characteristics of the final deployment model, migrating the model and the configuration to a reasoning environment, and starting reasoning service; and dynamically adjusting the number of reasoning nodes according to the real-time network flow of the reasoning service. By implementing the technical scheme provided by the invention, through bidirectional dynamic resource scheduling and model deep optimization, the computing resource utilization rate and the model deployment efficiency are improved, and the high availability of the reasoning service is guaranteed.
Owner:BEIJING HIZHI TECH CO LTD

Service monitoring and management system based on distributed Agent

The invention relates to the cross technical field of a distributed artificial intelligence system and micro-service governance, in particular to a service monitoring and management system based on a distributed Agent, which adopts a three-level architecture of a user layer, a host Agent layer and a service Agent layer: the user layer is provided with a unified entrance forwarding request; the host Agent layer is integrated with a service registry and other modules to realize global management and control; the service Agent layer comprises a registration module and the like to complete function and state feedback. Through four innovations of a dynamic observability model, an incremental heartbeat protocol, a fine-grained state machine and an intelligent routing engine, dynamic binding of a service function and a state is realized. The scheme has the advantages that the service observability is improved, and the problem positioning efficiency is optimized; the monitoring data transmission quantity is reduced; the service interruption time is shortened, and the availability is improved; and the method is suitable for scenes such as intelligent customer service and the like.
Owner:SHANGHAI TEGAO INFORMATION TECH CO LTD

Intelligent database switching method and device, computer equipment and storage medium

The invention discloses an intelligent database switching method and device, computer equipment and a storage medium, belongs to the technical field of big data, and is applied to a financial database downtime processing scene. According to the method, the node state of the database is monitored in real time, multi-dimensional load evaluation is combined, the optimal main and standby database combination is automatically selected, second-level fault sensing and rapid switching are achieved, and a minute-level time window needed by traditional database recovery is remarkably shortened. The database identifier is embedded in the service main key, so that the request can accurately position the corresponding database node, complex cross-database query and routing overhead are avoided, and the system response efficiency is improved. In addition, through data seamless transmission and automatic abnormal switching between the main database and the standby database, the risk of service interruption caused by database faults is reduced. According to the scheme, the high availability and the fault recovery speed of the database are improved, repeated construction of a multi-product-line database cluster is reduced through resource integration, and a large amount of hardware and operation and maintenance cost are saved.
Owner:CHINA PING AN PROPERTY INSURANCE CO LTD

Simulation deduction method and system based on distributed parallel scheduling and storage medium

The invention provides a simulation deduction method and system based on distributed parallel scheduling and a storage medium, and is applied to the technical field of data processing.The method comprises the steps that proxy services are deployed at a plurality of preset computing nodes, resource state information of all the nodes is dynamically collected, and a computing resource pool is obtained; arranging a pre-registered plug-in based on the data dependency relationship to obtain a simulation task process; in response to simulation scene configuration submitted by a user, generating a distributed scheduling task containing a fragmentation strategy; determining a plurality of sub-tasks corresponding to the distributed scheduling task based on the simulation task process; and based on the real-time load of the computing resource pool and the task priorities of the plurality of sub-tasks, respectively distributing the plurality of sub-tasks to a plurality of target computing nodes for task execution, and obtaining task execution result information. According to the method and the device, the time delay performance of distributed scheduling and the high availability performance of scheduling simulation tasks can be improved.
Owner:齐鲁空天信息研究院

Data storage method and device based on partition identifier mapping, equipment and medium

The invention relates to the technical field of distributed storage, can be applied to business scenes of financial science and technology, medical health and the like, and discloses a data storage method, device, equipment and medium based on partition identifier mapping, and the method comprises the following steps: receiving a file and dividing the file into data blocks to generate block identifiers; generating a partition identifier based on the file identifier and the block identifier through hash mapping; creating a partition disk pack mapping table and writing an initial relationship; querying the mapping table to obtain a disk group, and writing the data block into a physical disk to generate a copy; monitoring the health of the disk, keeping the partition identifier and the block identifier unchanged when a fault occurs, and updating the mapping relation to a new disk group; and receiving a read-write request of the target partition identifier, querying the mapping table to obtain the target disk pack, and executing access. According to the method, the partition identification and the block identification are kept unchanged, fast fault tolerance is achieved only by updating the mapping relation, metadata updating expenditure is reduced, bottom layer change is shielded through partition identification routing, and access transparency and high availability are achieved.
Owner:PING AN TECH (SHENZHEN) CO LTD

Fault switching method and device, electronic equipment and storage medium

The invention discloses a failover method and device, electronic equipment and a storage medium, and relates to the technical field of distributed storage, and the failover method comprises the following steps: determining a first disk resource group corresponding to a first storage node and a second disk resource group corresponding to a second storage node; and establishing a double-control relationship between the first disk resource group and the second disk resource group. And after the establishment of the double-control relationship is completed, carrying out fault switching on different disk resource groups under the double-control relationship based on the first mechanism and the second mechanism. The technical problem that in a distributed storage system, efficient and safe disk resource management and data access control of a fault controller cannot be effectively achieved under a double-controller architecture is solved, and the technical effects of improving rapidness, smoothness and automatic back-switching of fault switching and remarkably improving the high availability of the distributed storage system are achieved.
Owner:JINAN INSPUR DATA TECH CO LTD

Document collaboration platform based on cloud computing

The invention relates to the technical field of computers, and discloses a document collaboration platform based on cloud computing. Real-time semantic complementation, automatic typesetting optimization and content abstract generation are provided through the intelligent editing assistant module, the manual editing burden is remarkably reduced, and the efficiency is improved; a rule engine of an automatic operation module is combined to realize automatic management of document versions, intelligent resolution of conflicts and standardized review, and the problems of communication cost and conflicts in multi-person cooperation are thoroughly solved; the problem of information overload is solved by relying on the deep semantic processing capability, and accurate positioning of core content is realized; and meanwhile, based on the elastic expansion capability of the distributed cloud architecture, the stability and high availability of multi-terminal real-time cooperation are ensured. Finally, a document processing closed loop integrating intelligent assistance, an automatic process and efficient cooperation is formed, and the bottlenecks of a traditional platform in the aspects of editing efficiency, cooperation fluency, information processing and the automation level are comprehensively broken through.
Owner:WIN THE BID HUIKANG TECH CO LTD

Multi-agent system-oriented self-healing graph scheduling system and method

The invention provides a self-healing graph scheduling system and method for a multi-agent system. According to the method, in the multi-agent task flow graph, when any node fails or needs to be upgraded, bypass or hot replacement can be automatically completed within the time lower than a preset failure threshold value, and it is ensured that task topology is continuously acyclic, data are not lost, and services are not interrupted. According to the method, in a directed acyclic graph of a multi-agent task process, a main / standby node is configured for each edge, and the health state of the nodes is monitored in real time through dual-channel Gossip heartbeat; when the main node meets the failure condition, the flow is automatically redirected to the backup node at the millisecond level, the node state is recovered by using the XOR-delta snapshot, and the topology acyclic property is verified at the same time, so that the task is ensured to be continuous and traceable. Experimental results show that the average recovery time is reduced to 18 ms, the annual downtime is reduced by more than ten times, and the method is suitable for intelligent finance, industrial internet, automatic driving and other real-time scenes needing parallel agent collaboration and high availability.
Owner:SHANGHAI GREAT WISDOM INFORMATION TECH CO LTD

Data processing method and system based on network security service

The invention provides a data processing method and system based on network security service, relates to the technical field of network data security processing, and realizes structured expression and behavior extraction of potential threats in encrypted communication by establishing a new data structure and behavior association model. By introducing a context causal chain modeling mechanism, associating the originally isolated security events into a complete attack path; by establishing a set of event grading mechanism based on comprehensive scoring of service criticality, response cost and attack propagation path, threat processing does not distribute resources blindly and averagely, but performs intelligent scheduling according to priority, so that the resource utilization efficiency is improved; a structured and executable defense instruction or blocking strategy can be generated according to an analysis result, man-machine cooperation or full-automatic response is supported, and response efficiency and strategy adaptability are remarkably improved. And a set of intelligent data processing solution which is oriented to a network security service practical application scene and has high availability and high expansibility is constructed.
Owner:HUBEI JINCHU NETWORK TECH

Automatic rebalancing of container-based services for high availability

Techniques are described for enabling a container service of a cloud provider network to detect imbalances of container placements across a selected set of availability zones (AZs) and to automatically rebalance placement of the containers, if needed. An actual distribution of containers may differ from an expected distribution according to a configured placement strategy, e.g., due to an outage or other operational issue affecting one or more of the AZs, scaling operations over time, and the like. In these and other scenarios, the container service can rebalance placement of the containers by terminating one or more containers in one or more of the AZs and launching additional containers in one or more other AZs to restore a more desirable balance. The periodic rebalancing of containers in this manner improves a container-based application's ability to load balance demand and further improves application availability by ensuring sufficient capacity is spread across multiple AZs.
Owner:AMAZON TECH INC

Distributed database system migration method and device supporting active transaction online migration

The invention provides a distributed database system migration method and device supporting online migration of active transactions. The method and device are suitable for a distributed database to complete dynamic migration of fragmented or tenant-level data and transactions under the condition that services are not interrupted. The method comprises the following steps: generating a consistency snapshot of a source node and transmitting the consistency snapshot to a target node; capturing an incremental transaction log after snapshot in real time; a log is efficiently synchronized to a target node through a remote direct memory access (RDMA) mechanism; the logs are played back in parallel at the target node to construct a complete data state; identifying active transactions in the source node and migrating their context; the target node continues to execute the transaction until submission is completed; and finally updating system routing information to complete migration takeover. According to the method, migration delay is effectively reduced, online transaction migration is realized, and the method has the advantages of high availability, low overhead, high adaptability and the like, and is suitable for database scenes such as elastic capacity expansion and contraction, load balancing and the like.
Owner:EAST CHINA NORMAL UNIV

Road berth charging and control system and method for realizing multi-level fault tolerance

The invention discloses a road berth charging and control system for realizing multi-level fault tolerance and a method thereof, and belongs to the technical field of intelligent traffic systems and Internet of Things. The system adopts an edge computing and distributed sensing collaborative architecture, and is composed of an edge controller (ECU) and a sensing and control unit (SCU) deployed in a berth. High availability of the system is ensured through fault-tolerant design of three core dimensions: firstly, perception layer fault tolerance dynamically adjusts data fusion weights of geomagnetism, millimeter wave radar and visual AI according to real-time environment data such as rainfall and electromagnetic interference by integrating an environment perception sub-module; secondly, network layer fault tolerance utilizes an immutable transaction log and a Saga distributed transaction compensation mechanism to realize local charging and digital RMB double offline payment when cloud connection is interrupted, and data consistency is ensured after network recovery; and finally, a hardware layer establishes a neighborhood cooperation protocol based on a signature agent command through fault tolerance, and a healthy node is allowed to act as an agent fault node through safety verification to execute an unlocking instruction. According to the invention, the problems of environmental interference, network interruption, single-point hardware failure and the like in unattended parking management are effectively solved, and the robustness and financial security of the system are remarkably improved.
Owner:JIANGSU RUOLIN LINK TECH CO LTD

Distributed high-availability code warehouse management method and system and medium

The invention relates to a distributed high-availability code warehouse management method and system and a medium, and relates to the technical field of version control and code management.The method comprises the steps that a pre-configured external load balancer, a plurality of GitLab Web nodes, an internal load balancer, a PostgreSQL database cluster, a Redis cluster, a Gitaly cluster and object storage or file storage are obtained; carrying out load balancing on the plurality of GitLab Web nodes through the external load balancer; performing load balancing on the internal connection of the GitLab application program through the internal load balancer; the data of the GitLab Web and the data of the GitLab Pracect are stored through the PostgreSQL database cluster, and the data of the GitLab Web and the data of the GitLab Pracect are And providing access to a Git warehouse by using the Gitaly cluster. According to the invention, high availability, high performance and expandability of code warehouse management are realized through a distributed architecture and redundancy configuration.
Owner:ZHIJI AUTOMOTIVE TECH CO LTD

Distributed transaction coordination system for high-concurrency scene

The invention discloses a distributed transaction coordination system for a high-concurrency scene, and particularly relates to the technical field of distributed computing and transaction management. The system comprises a transaction manager, a resource agent group and a log storage module, the transaction manager receives a transaction request, generates a global transaction identifier and maintains a state machine, and an internal fragment coordinator routes the transaction to different processing fragments according to a consistent Hash algorithm for parallel coordination; the resource agent group is connected with each resource node, executes local operation and reports a state through an asynchronous batch mechanism; and the log storage module adopts a multi-copy distributed structure to persist transaction logs. Through fragment parallel processing, asynchronous state reporting and multi-copy log cooperation, the system throughput is effectively improved, the transaction delay is reduced, and the final consistency and high availability of distributed transactions are guaranteed.
Owner:HANGZHOU CHENSHENG INTELLIGENT TECHNOLOGY CO LTD

High availability system based on container platform, container scheduling method and electronic equipment

The invention provides a high-availability system based on a container platform, a container calling method and electronic equipment, and the method comprises the steps: pre-creating and activating a container copy through a container scheduling module according to the resource state of each container node and a priority constraint condition in the early stage of fault occurrence, and then when a fault event occurs, calling the container copy in advance; and in response to the fault event reminding message, calling an interface of the container scheduling module to create a new service instance in the container copy, carrying out synchronous processing on the container state based on the new service instance, and migrating the flow on the fault container instance to the new container instance. Therefore, in a high-concurrency scene, by selecting the embodiment of the invention, the flow is immediately taken over through the pre-activated standby container copy when a fault occurs, so that zero-shutdown container switching is realized, the high availability of a container platform is remarkably improved, and smooth operation of services is ensured.
Owner:DUXIAOMAN TECH (BEIJING) CO LTD

Runtime component for hot-standby and high availability

A duplex configuration for application state synchronicity. A primary controller actively monitors and controls a plant / process and a secondary controller takes over in case of a failure of the primary controller. Input data is received at respective inputs of the primary controller and the secondary controller. Determining which of the received input data is associated with or should be treated as External Events permits achieving application state synchronicity between the primary controller and the secondary controller by synchronizing the execution of events associated with the same input data.
Owner:SCHNEIDER ELECTRIC IND SAS +1

K8s-based JUPter Notebook container arrangement method

The invention relates to the technical field of containers and micro-service architecture, and discloses a K8s-based Juser Notebook container arrangement method, which comprises the following steps of: S1, packaging Notebook into a Docker mirror image file; s2, mounting a Notebook configuration file on the K8s cluster through a ConfigMap (Configuration Map); s3, deploying and defining the number of copies of the Notebook application through Deployment; S4, configuring a service network and an ingress network; S3, deploying and defining the number of copies of the Notebook application through Deployment; s5, a cloud storage mode is adopted, and calculation and storage are separated; and S6, K8s API automatic management is carried out. The deployment process is simplified, the maintenance and management difficulty is reduced, and the resource utilization rate is improved; a cloud storage mode is adopted, separation of calculation and storage is achieved, and environment configuration is simplified; the Ingress provides network functions such as flow management and reverse proxy, so that the Notebook service is easier to access and integrate; the K8s cluster can ensure the high availability of the Notebook service; and when a fault occurs, quick recovery is realized, service interruption time is reduced, monitoring and management are easy, and management and operation and maintenance efficiency is improved.
Owner:JIANGXI FASHION TECH