Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

891 results about "Concurrency" patented technology

In computer science, concurrency is the ability of different parts or units of a program, algorithm, or problem to be executed out-of-order or in partial order, without affecting the final outcome. This allows for parallel execution of the concurrent units, which can significantly improve overall speed of the execution in multi-processor and multi-core systems. In more technical terms, concurrency refers to the decomposability property of a program, algorithm, or problem into order-independent or partially-ordered components or units.

Big data distributed storage and parallel processing cooperation method based on cloud computing

The invention discloses a big data distributed storage and parallel processing collaboration method based on cloud computing. The method comprises the following steps: sensing data stream characteristics in real time through a self-adaptive dynamic partitioning engine, dynamically adjusting a partitioning strategy and generating a metadata label; constructing a node selection model through a comprehensive evaluation algorithm, selecting storage nodes to form an optimal storage cluster, and dynamically adjusting a resource matching weight coefficient based on a load state through a load optimization module; decomposing a data processing task into parallel subtask units, and constructing a dual-objective optimization model; triggering a dynamic rebalance mechanism through a distributed monitoring agent in combination with a hierarchical early warning strategy; and constructing a multi-level cache system to optimize a data access path, outputting a final result, pushing the final result to the user terminal, and updating the knowledge base. Through collaborative optimization of dynamic partitioning, multi-dimensional resource scheduling, elastic scaling and intelligent caching technologies, the resource utilization rate, the load balancing capacity and the stability in a high-concurrency scene of the system are remarkably improved.
Owner:CHINA THREE GORGES UNIV

Multi-agent non-cooperative game driven high-concurrency task reasoning method and system

The invention discloses a multi-agent non-cooperative game driven high-concurrency task reasoning method and system, and the method comprises the steps: decomposing a heterogeneous task into standardized task units, and carrying out the quantitative modeling of the calculation resource demands of the task; constructing a task auction mechanism of the multi-agent non-cooperative game model, performing task allocation by adopting an anti-strategic auction rule and a fragmentation asynchronous protocol, and constraining resource declaration behaviors of the agents through a credit pledge mechanism; the real-time resource utilization rate of the system is monitored, the task priority is dynamically adjusted in combination with the task preemption historical state, a resource soft preemption strategy is triggered, and a compensation queue is established; a self-adaptive model splitting strategy is adopted to distribute reasoning tasks to end side equipment and edge computing nodes, cooperative execution is achieved through data compression and transmission, and task rescheduling is conducted according to feedback of an execution result; according to the method, the task execution efficiency, the resource utilization rate and the system stability can be remarkably improved in a high-concurrency scene.
Owner:PANDA ELECTRONICS

Industrial scene-oriented data acquisition system and acquisition method thereof

The invention discloses an industrial scene-oriented data acquisition system and an industrial scene-oriented data acquisition method. According to the invention, through the multi-level dynamic optimization design, the data acquisition efficiency and the system reliability in a complex environment are significantly improved. The system dynamically allocates thread resources based on the real-time communication state of the equipment, realizes stable throughput under a high-concurrency scene in combination with core binding and polling scheduling strategies, and ensures that millisecond-level response is still maintained when 10000-level equipment is accessed. The memory preloading and protocol template technology eliminates jitter during operation, the hierarchical resource isolation mechanism provides deterministic guarantee for key control instructions, and interruption of the production process due to data delay or resource competition is avoided. The data flow delay is further reduced through multi-protocol efficient analysis and intelligent cache management, seamless compatibility of heterogeneous equipment in a hybrid networking scene is supported, and the strict requirements of continuous production industries such as steel and chemical engineering for real-time performance and stability are met.
Owner:云鼎科技股份有限公司

Real-time interactive digital human system supporting high concurrency and implementation method thereof

The invention discloses a real-time interactive digital human system supporting high concurrency and an implementation method thereof, and relates to the technical field of digital human interaction.The system comprises a model instance pool module used for loading the weight of a deep learning model to a shared memory area through the memory mapping technology; the multi-thread scheduling module is used for scheduling user requests by adopting a lock-free queue and a dynamic priority algorithm; the asynchronous pipeline processing module is composed of decoupled micro-services, and the modules are connected in series through asynchronous message queues; the client SDK is used for dynamically switching a rendering mode according to terminal hardware performance and network conditions; the elastic capacity expansion and contraction module is used for monitoring resource loads in real time and automatically adjusting the number of service instances; and the audio and video synchronization calibration module is used for ensuring that the synchronization error of the audio and video frames is lower than a preset threshold value through a timestamp alignment algorithm. According to the scheme, core challenges in a high-concurrency digital human interaction scene can be systematically solved, and innovative support is provided for large-scale real-time application.
Owner:LIANGSHENG DIGITAL ARTIFICIAL INTELLIGENCE (SHENZHEN) CO LTD

Dynamic rendering resource scheduling method and related equipment

The invention discloses a resource scheduling method for dynamic rendering and related equipment, and relates to the technical field of dynamic rendering, and the method comprises the following steps: obtaining time sequence data of a user touch event; based on the time sequence data, predicting a user behavior state through a behavior prediction model, and determining a current resource demand level; dynamically adjusting a resource allocation strategy and a video memory management strategy of the graphic processing unit according to the resource demand level; and executing a rendering task based on the adjusted resource allocation strategy, and performing picture output by adopting a double-buffer architecture so as to reduce delay caused by time sequence mismatch. According to the method and the device, by establishing closed-loop control of a user behavior-resource demand-supply strategy, accurate scheduling and efficient utilization of resources are realized, the problems of time sequence mismatch, resource waste and jamming in the traditional technology are solved, the rendering efficiency and the user experience are improved, and technical support is provided for high-concurrency and low-delay scenes.
Owner:启朔(深圳)科技有限公司

System for knowledge instantiation and evidence synthesis through human-computer workflow orchestration and neuromorphic prompting

The present invention discloses a modular AI-based system for knowledge instantiation and evidence synthesis, unifying five modules: Knowledge Instantiation, ingesting anchor knowledge via retrieval- augmented generation and neuromorphic prompting; Case-Based Agent Generation, orchestrating domain constraints, performance criteria, and sub-task decomposition; Secondary Epistemogenesis, spawning sub-agents with inherited heuristics; Workflow Orchestration and Evidence Synthesis, merging outputs, tracking metadata, and storing final artifacts; and Reinforcement Learning from Human Refinement, capturing feedback and edits. The method for output generation initiates by defining knowledge requirements, indexing public and proprietary knowledge, validating output structure, drafting content, and refining outputs via sub-agent spawning and human oversight. The system ensures domain alignment, concurrency control, and persistent versioning. By combining specialized knowledge sources, multi-agent concurrency, and iterative human feedback loops, it dynamically adapts to evolving knowledge requirements. This invention emphasizes reliability, traceability, and AI-driven outputs in regulated industries, delivering trustworthy outcomes anchored in expert knowledge.
Owner:RAO UJJWAL

High-availability testing method and system for task scheduling of computing power host management platform

The invention discloses a high-availability testing method and system for task scheduling of a computing power host management platform, belongs to the technical field of high-availability testing of computers, and aims to solve the technical problems that in the prior art, high-concurrency scene testing coverage rate is insufficient, cross-domain scheduling, failover and self-healing scene verification are imperfect, and testing efficiency is high. According to the technical scheme, the method structurally comprises the steps of hybrid load testing, wherein the elastic capacity expansion and contraction capacity and the performance stability of a system are verified by simulating a long-time stable load and burst flow scene; cross-regional intelligent scheduling: tasks are dynamically scheduled based on cost, time delay and regional strategies, and optimal resource allocation and service continuity are ensured; resource fault migration recovery: simulating node-level or region-level faults, and verifying the automatic detection, migration and recovery capabilities of the system; monitoring and data collection: collecting affairs, resources and stability indexes in real time, and providing multi-dimensional data support for verification; and analysis and optimization: identifying a bottleneck based on test data to generate optimization suggestions, and continuously improving the high availability of the system.
Owner:INSPUR COMM TECH CO LTD

CPU (Central Processing Unit), GPU (Graphic Processing Unit) and NPU (Network Processing Unit) resource allocation method and system for training and calculating integrated machine

The invention relates to the field of heterogeneous computing resource management, in particular to a CPU (Central Processing Unit), GPU (Graphic Processing Unit) and NPU (Network Processing Unit) resource allocation method and system of a training and reckoning all-in-one machine. The invention discloses a CPU, GPU and NPU resource allocation system of a training and reckoning all-in-one machine. The system comprises a schedulable resource analysis module, a load analysis module and a resource allocation module. According to the method, the reference parameters and the real-time dynamic indexes of heterogeneous hardware are fused, so that deep perception and efficient quantification of computing power resources are realized; on the basis of an NPU temperature attenuation experiment, calibrating a computing power loss coefficient, and dynamically correcting available weight values of a CPU, a GPU and an NPU in combination with suitability analysis of a GPU stream processor utilization rate and a task batch processing demand and collaborative efficiency evaluation of a CPU dominant frequency and an IPC value; and the accuracy of resource availability prediction is improved, so that the system can still guarantee the stability of computing power output in complex environments such as high temperature and high concurrency.
Owner:DONGGUAN HUAMING TENG TECH CO LTD

Data processing method and system based on cloud computing

The invention provides a data processing method and system based on cloud computing, belongs to the technical field of data processing, and realizes high concurrent processing and automatic load balancing by dynamically scheduling computing power through elastic resource allocation and a distributed computing framework and combining a storage and computing separation framework. The system supports cross-node disaster recovery backup and multi-layer encryption, and data security is guaranteed; and an on-demand payment mode is adopted, so that the hardware input cost is reduced, the resource utilization rate is improved, and the large-scale data processing efficiency and the system expandability are remarkably improved.
Owner:ZHONGKE NUOXIN BEIJING HI TECH

Server resource dynamic scheduling method for dealing with video stream high concurrent access

The invention discloses a server resource dynamic scheduling method for dealing with video stream high concurrent access, and particularly relates to the technical field of computer network and intelligent scheduling. A user access behavior data set is constructed; a deep learning model is adopted to train an access hot spot prediction model based on the time sequence features to predict a future access hot spot area and a peak trend; acquiring response delay, CPU / GPU occupancy rate and bandwidth load information of the heterogeneous server group, and generating a resource state multi-dimensional parameter set; carrying out joint modeling on the model and the parameter set, constructing a resource scheduling priority model by adopting a graph neural network, and generating an optimal scheduling path graph based on an A * improved algorithm; task transfer, instance elastic expansion, cache preheating and other scheduling operations are executed according to the model; according to the method, the resource utilization efficiency and the service quality of the video system in a high-concurrency scene can be improved, and the method has the advantages of being high in real-time performance, intelligent in scheduling and high in self-learning capability.
Owner:SBAIDA INTERNET OF THINGS TECH (BEIJING) CO LTD +1

Large language model reasoning method based on unloading assembly line

The invention discloses a large language model reasoning method based on an unloading pipeline, and the method comprises the steps: obtaining an optimal unloading reasoning strategy through calculation according to the obtained model structure information and model configuration information of a large language model, and the hardware specification information and system operation load information of reasoning equipment; and then tasks in the optimal unloading reasoning strategy are scheduled through a fine-grained pipeline to output an unloading reasoning result of the large language model. According to the method, reasoning tasks are automatically configured according to input large language model information and a system hardware environment, and hardware resource use and reasoning performance are optimized. For deployment of a large language model on local equipment, transmission optimization is carried out for unloading using a solid state disk, so that the data transmission speed is increased. Meanwhile, fine-grained assembly line task scheduling is carried out for unloading reasoning, and the reasoning concurrency and the model throughput can be remarkably improved through unloading reasoning of an assembly line model.
Owner:NANJING UNIV

Compute-in-memory chip, instruction scheduling method, and related apparatus

The present application discloses a compute-in-memory chip, an instruction scheduling method, and a related apparatus. The compute-in-memory chip comprises an instruction memory, an instruction scheduler, and at least one compute-in-memory memory; each compute-in-memory memory comprises at least one storage array; the instruction memory is used for acquiring a first tensor instruction to be executed; and the instruction scheduler is used for scheduling, on the basis of the association relationship between the first tensor instruction and a second tensor instruction and the state of a target storage array needing to be operated for executing the first tensor instruction, the first tensor instruction to the compute-in-memory memory to which the target storage array belongs so that the compute-in-memory memory executes the first tensor instruction. According to embodiments of the present application, diversified compute-in-memory computing can be supported, efficient out-of-order execution scheduling of a tensor instruction set is achieved on the basis of the compute-in-memory chip, and the requirements of compute-in-memory technology for high concurrency and high throughput rate are met.
Owner:HUAWEI TECH CO LTD

Dynamic tool integration method and system for intent-driven AI dialogue system

The invention provides a dynamic tool integration method and system for an intention-driven AI dialogue system, and the method comprises the steps: S5, extracting resource occupation information from a transaction log according to a state persistence record, carrying out the resource release and temporary context cleaning of each heterogeneous component, and obtaining a released resource list; s6, searching a breakpoint state snapshot and a transaction log from the distributed database according to the released resource list, and reconstructing a tool chain execution flow historical state to obtain a fault context analysis result; and S8, according to the context recovery state and the semantic feature vector, updating a tool chain execution stream, and reallocating computing resources through a heterogeneous component coordination mechanism to obtain a task continuity confirmation result. According to the method, the recovery efficiency and the resource utilization rate of the tool chain in an abnormal scene are remarkably improved, and the task continuity and the system robustness in a high-concurrency environment are guaranteed.
Owner:GUANGDONG SANDING INTELLIGENT INFORMATION TECH CO LTD

Method and system for calling service capability in application software

The invention provides a service capability calling method and system in application software. The method comprises the following steps of: monitoring real-time temperature data and load state data of each available service node in a server cluster in real time, and combining a preset load balancing strategy and a virtual machine dynamic migration technology to ensure the stable operation of the server cluster in a high-concurrency scene. And when it is detected that the real-time temperature of the available server node exceeds the preset temperature threshold value or the capacity of the server node reaches the capacity upper limit, the service capability of the available server node is migrated to the standby node, and service interruption is avoided. Meanwhile, the real-time temperature data, the load state data and the execution frequency of the migration operation are subjected to correlation analysis, and a load balancing strategy and a migration trigger threshold value are dynamically optimized, so that service continuity is maintained, and the temperature fluctuation range of hardware resources is controlled. According to the technical scheme provided by the invention, continuous response of service requests and temperature stability of hardware resources in a high-concurrency scene can be realized.
Owner:HANDING TECHNOLOGY (BEIJING) CO LTD

Large language model reasoning performance evaluation and optimization method, electronic equipment and storage medium

The invention discloses a big language model reasoning performance evaluation and optimization method, electronic equipment and a storage medium. The method comprises the following steps: initializing a test environment; dynamically generating concurrent user requests; collecting performance data based on the concurrent user requests; performing real-time monitoring and reporting based on the performance data; and generating a performance evaluation report based on the performance data. Through the above steps, the method can comprehensively evaluate the reasoning performance of the LLM in a high-concurrency scene, and ensures the optimal performance of the system in different hardware environments.
Owner:XINZHIHUIXIANG TECHNOLOGY (TIANJIN) CO LTD

Intelligent control method for heat supply digital twinning

The invention provides a heat supply digital twinborn intelligent regulation and control method, and relates to the technical field of intelligent regulation and control, and the method comprises the steps: executing millisecond-level high-concurrency numerical simulation according to a dynamic digital twinborn body, so as to obtain a simulation result; based on the simulation result, coupling the electricity price fluctuation data and the weather prediction data, performing future heat supply system multi-working-condition deduction, and dynamically optimizing a heat source unit start-stop combination and pipe network valve opening strategy to obtain a deduction result and an optimization strategy; generating an equipment regulation and control instruction set according to the deduction result and the optimization strategy; and an equipment regulation and control instruction set is issued to an edge computing node, a variable-frequency circulating pump and an electric control valve executing mechanism are controlled, and dynamic closed-loop regulation of the flow and the pressure of the heat supply pipe network is achieved. According to the invention, the pipe network flow and pressure regulation response time can be shortened to a second level.
Owner:YUANHUA YITONG HEAT SUPPLY SCI TECH DEV BEIJING

Multi-mode audio and video synchronous processing method and system based on distributed architecture

The invention relates to the technical field of computers, and discloses a multi-mode audio and video synchronous processing method and system based on a distributed architecture. The method comprises the following steps: generating a unique logic sequence identifier consisting of a device code, a media type and a frame number for each frame of audio and video data; distributing the frame carrying the identifier to a distributed node for processing and adding a local timestamp; and the sink node reconstructs an original frame sequence according to the identifier, calculates a synchronous offset by combining high-precision clock calibration, and realizes dynamic reordering and output through a double-buffer structure and a self-adaptive drift compensation algorithm. The system comprises a source end acquisition module, an identifier generation module, a task scheduling module, a distributed processing cluster module, a convergence synchronization module and a synchronous output module. According to the scheme, the system effectively guarantees the audio and video synchronization precision in a high-concurrency heterogeneous environment, the synchronization error is controlled within 20 milliseconds, and the playing quality and the user experience of scenes such as live broadcast and cloud games are remarkably improved.
Owner:SHENZHEN HAIWEI HENGTAI INTELLIGENT TECH CO LTD

Intelligent resource scheduling method and system based on task characteristics and system state perception

The invention discloses an intelligent resource scheduling method and system based on task characteristics and system state perception, and relates to an operating system kernel optimization technology. According to the method, resource demand types, importance labels and time sensitivity of tasks are captured in real time through an eBPF technology, system bottlenecks are recognized by combining / proc and sysfs file system data, a dynamic scheduling strategy is generated by adopting an improved Q-Learning model, and resource allocation is executed through a schemed * function and a cgroup v2 interface. The problems that a traditional scheduler is poor in self-adaptive capacity and resources are mismatched are solved, the actually-measured resource utilization rate is increased to 85%-92%, and the key task delay standard reaching rate gt is achieved; the method is suitable for edge computing, cloud servers and other high-concurrency scenes.
Owner:SHANGHAI AVCON INFORMATION TECH

Heterogeneous reasoning computing power scheduling method and device and storage medium

The invention provides a heterogeneous reasoning computing power scheduling method and device and a storage medium, and the method comprises the steps: receiving a reasoning request, splitting the reasoning request, and obtaining at least one reasoning task; according to the coding length of each reasoning task, scheduling the reasoning tasks to a long task queue or a short task queue; when it is monitored that reasoning tasks with to-be-processed task states exist in the long task queue and the short task queue, the working state of the reasoning service of the first processor is judged based on the number of tasks in processing of the first processor in the long task queue and the short task queue and the maximum concurrency of the first processor; the maximum concurrency of the first processor is obtained based on the current conditions of the long task queue and the short task queue; and on the basis of the working state of the first processor reasoning service, the reasoning task to be processed is scheduled to the first processor reasoning service or the second processor reasoning service so as to maximize and reasonably utilize the mixed computing power resource of the heterogeneous processor.
Owner:INSPUR SUZHOU INTELLIGENT TECH CO LTD

Fragmented storage and query optimization method and system for high-concurrency database

The invention belongs to the field of query optimization, and particularly relates to a fragmentation storage and query optimization method and system for a high-concurrency database, and the method comprises the steps: analyzing business features through a preset decision tree model, and selecting an optimal fragmentation key, constructing a fragmentation rule engine and initializing a database in combination with hash, range and list fragmentation rules and a fragmentation splitting-merging strategy; generating a query abstract syntax tree by using an ANTLR4 analyzer, and detecting whether a preset Cube is hit or not to directly return a result; if not, dynamically routing to a target fragment list based on a fragment key field, load balancing and a failover strategy, accelerating data retrieval in combination with a high-frequency index, and combining fragment-level results through a T-TopK algorithm; according to the method, the self-adaptive matching of the fragmentation strategy and the service requirement, the calculation push-down of the query process and the result optimization are realized, and the low delay and the high availability in a high-concurrency scene are ensured.
Owner:JIANGSU LINGHAO NETWORK TECH CO LTD

Dynamic task scheduling resource optimization method and device for distributed system

The invention discloses a method and a device for optimizing dynamic task scheduling resources of a distributed system. The method comprises the following steps of: monitoring a task dependency relationship and hardware resource state data in real time; a dependency release amount R representing the downstream task activation capability is generated based on the dependency relationship; generating an index A representing the resource utilization efficiency based on the resource state; a scheduling strategy is dynamically selected according to the load state: when the load is low, the task with the maximum R is preferentially allocated to improve the CPU parallelism degree, when the load is stable, allocation is carried out according to R and A weighted values to balance resource utilization, and when the load is high, the task with the maximum A is preferentially allocated to reduce the memory pressure; and finally driving task execution and updating a resource state. By dynamically sensing the task topological relation and the resource state, adaptive scheduling under load fluctuation is realized, the collaborative utilization efficiency of CPU and memory resources is effectively improved, and the problem of performance bottleneck of a traditional scheme under complex dependence and high-concurrency scenes is solved.
Owner:CHENGDU UNIV OF INFORMATION TECH

Printing task scheduling method and system based on main control chip and terminal

The invention relates to the technical field of printing task scheduling, in particular to a printing task scheduling method and system based on a main control chip and a terminal, and the method comprises the steps of initial task classification, dynamic priority adjustment, task queue recombination, resource allocation execution, termination mechanism triggering and the like. The high-priority task response speed is optimized through system load monitoring and a waiting time coefficient, and a real-time task pool and an elastic buffer pool are divided to improve the scheduling efficiency. According to the method, the real-time task pool and the elastic buffer pool are divided through an algorithm of'dynamic memory pool partitioning + task life cycle binding ', and the memory blocks are pre-allocated according to the task types, so that the real-time performance and the scheduling efficiency of the high-concurrency printing tasks are improved, rapid response of high-priority tasks can be realized, delay of key tasks is reduced, and the efficiency of the high-concurrency printing tasks is improved. The resource utilization rate and the scheduling fairness are improved, and the method is suitable for high-concurrency printing scenes.
Owner:ZHEJIANG CANGTIAN INTELLIGENT INFORMATION TECH CO LTD

Lightweight asynchronous message processing method and system for high-concurrency scene

The invention relates to a lightweight asynchronous message processing method and system for a high-concurrency scene, and the method comprises the steps: receiving a message request of a client, carrying out the preprocessing of the message request, and generating a to-be-processed message; distributing the to-be-processed message to a multi-level priority queue based on a priority scheduling mechanism; dynamically adjusting the number of core threads and the maximum number of threads of the thread pool according to the resource state data, and forming a thread pool resource configuration scheme of the current scheduling period; dynamically calculating the message processing quantity of the current scheduling batch based on the message accumulation quantity of each priority queue, the thread pool resource configuration scheme, the CPU utilization rate in the resource state data and the JVM heap memory utilization rate; and based on the out-of-heap memory buffer area reference, assigning the to-be-processed message object to a corresponding processing thread to execute a service processing logic until the processing operation of the messages in all the scheduling units is completed. The method has the effect of improving the resource adaptability of the asynchronous message processing architecture.
Owner:TONGCHENG NETWORK TECH CO LTD

Storage and calculation integrated data scheduling system and method for high-concurrency scene

The invention relates to the technical field of high-concurrency data processing, and discloses a storage and calculation integrated data scheduling system and method for a high-concurrency scene, and the system comprises a storage and calculation fusion architecture module, a distributed cache module, a dynamic fragmentation module, a parallel calculation module, a preloading module and a resource scheduling center. The method corresponds to the system. According to the method, by deeply fusing data storage and calculation and utilizing a distributed caching technology and a data preloading mechanism, I / O bottleneck is effectively reduced, dynamic data fragmentation and parallel calculation are supported, and the method is particularly suitable for large-scale data analysis scenes.
Owner:GLORYVIEW TECH INC

Resource scheduling method, device, system and storage medium

A resource scheduling method, device, and system, and a storage medium. The method includes: acquiring a target resource allocation object to be scheduled; allocating, from a plurality of candidate scheduler instances, a target scheduler instance for the target resource allocation object, and pre-allocating, by the target scheduler instance and from a resource node cluster, a resource node for the target resource allocation object, to obtain a pre-allocated resource node corresponding to the target resource allocation object; performing conflict detection on the pre-allocated resource node based on an optimistic concurrency strategy, and preempting the pre-allocated resource node after passing the conflict detection; and scheduling the target resource allocation object to run on the pre-allocated resource node.
Owner:BEIJING VOLCANO ENGINE TECH CO LTD

Efficient database reading method based on multi-level cache optimization and dynamic index fragmentation

The invention relates to the field of database reading, and particularly discloses an efficient database reading method based on multi-level cache optimization and dynamic index fragment.The efficient database reading method comprises the steps that high-frequency query data is accurately recognized through a sliding window algorithm, a dynamic cache loading and preheating strategy is implemented, query delay is effectively shortened, and the database burden is relieved; intelligent distribution of index fragments is realized by applying a consistent Hash algorithm, and the query efficiency is greatly improved in cooperation with distributed query routing and transaction processing optimization; the query load balancing ensures stable and efficient operation of the system in a high-concurrency scene by means of weighted polling and a minimum connection strategy in combination with real-time monitoring and an elastic capacity expansion and contraction mechanism. According to the method, the database access efficiency and the dynamic adaptive capacity are remarkably improved, and the method is particularly suitable for a large-scale distributed data processing environment.
Owner:YANTAI JIERUI NETWORK TRADING

Large language model high concurrency reasoning method and system

The invention relates to the technical field of artificial intelligence, and discloses a large language model high concurrency reasoning method and system, and the method comprises the steps: calculating the size of a video memory block through an actuator, and distributing a video memory space; the scheduler is used for converting the request sequence and putting the request sequence into a waiting queue of the scheduler; the scheduler allocates a corresponding video memory block to each request sequence until each request sequence can perform next reasoning; the scheduler calculates the video memory requirement of the request sequence in the waiting queue according to the priority sequence, and transfers the request sequence in the waiting queue to the running queue; according to the number of pre-filling types and the number of decoding types of the request sequence, distributing the number of video memory blocks used for executing pre-filling reasoning or decoding reasoning; therefore, according to the method, continuous batch processing, a dynamic space allocation mechanism and a task scheduling framework are adopted, the parallel reasoning capability of continuous batch processing is fully utilized, the concurrency and throughput of large model reasoning are improved, and the limitation that traditional continuous batch processing needs space pre-allocation is solved.
Owner:TROY INFORMATION TECHNOLOGY CO LTD

CPU-GPU (Central Processing Unit-Graphic Processing Unit) cooperative execution method and system for mobile terminal deep learning reasoning

The invention relates to the technical field of mobile terminal data processing, in particular to a CPU-GPU dual-core collaborative execution method and system for a mobile terminal deep learning reasoning task. The method comprises the following steps: in an offline stage, performing structural analysis on a target deep learning model, and predicting execution time and data transmission overhead of each operator in combination with hardware characteristics; in the online stage, the two key modules of the task scheduler and the heterogeneous execution engine ensure efficient CPU and GPU cooperative execution through the innovative dynamic scheduling and resource management technology. The reasoning tasks are dynamically distributed to the CPU and the GPU, flexible cooperative scheduling is achieved, the calculation advantages of all processors are fully played, the overall utilization rate is remarkably improved, and the idle or bottleneck of a single calculation resource is avoided. For resource-limited scenes such as a mobile terminal, the system optimizes the memory use efficiency, improves the task execution concurrency and stability and realizes efficient reasoning performance through a double-buffer structure and a data multiplexing and concurrent execution strategy.
Owner:HUNAN UNIV

Distributed high-concurrency real-time industrial data pushing method and system

The invention discloses a distributed high-concurrency real-time industrial data pushing method and system, relates to the technical field of high-concurrency data processing, and solves the problems that dynamic adjustment of traditional buried point logic is difficult, and decoupling with service data cannot be carried out. The industrial data pushing system comprises a gateway management and control layer, a message queue layer, a cache optimization layer, a data processing layer, a pushing service layer and a monitoring operation and maintenance layer, asynchronous decoupling is carried out on the client and the server through the message queue layer, and priority transmission and flow peak clipping are carried out on original industrial data; performing anomaly prediction on the original industrial data through a data processing layer according to an anomaly detection rule to obtain anomaly alarm data and an out-of-order anomaly event; according to the invention, the push service layer carries out connection management and abnormal alarm data and out-of-order abnormal event push on the client according to a scheduling push mechanism, and carries out data burying point judgment on a client request, thereby improving the dynamic flexibility of data burying points, and solving the problems of industrial data analysis and resource configuration optimization.
Owner:ZHENGZHOU JIACHEN ELECTRIC CO LTD

Concurrency control method and device, electronic equipment, storage medium and product

The invention discloses a concurrency control method and device, electronic equipment, a storage medium and a product. The method is applied to a concurrency control system comprising a processor and a bus controller, and comprises the steps that the processor responds to an atomic operation request of a target thread and sequentially sends a storage area switching instruction and a register operation instruction to the bus controller; the bus controller responds to the storage area switching instruction and the register operation instruction, executes storage area switching operation, generates a bus occupation maintaining signal and continuously executes register read-write operation in a target storage area; the processor sends a scheduling request to an operating system, so that the operating system improves the priority of the target thread to be a first priority; and the bus controller generates a bus release signal after the atomic operation of the target thread is finished. According to the scheme, through cooperation of the processor and the bus controller, atomized execution of thread Bank switching and register reading and writing is achieved, and the concurrency performance is improved while the reliability and consistency of data are ensured.
Owner:HUIZHOU DESAY SV AUTOMOTIVE