Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

205 results about "Instruction scheduling" patented technology

In computer science, instruction scheduling is a compiler optimization used to improve instruction-level parallelism, which improves performance on machines with instruction pipelines. The pipeline stalls can be caused by structural hazards (processor resource limit), data hazards (output of one instruction needed by another instruction) and control hazards (branching).

Instruction processing system, method and device, electronic equipment and storage medium

According to the instruction processing system, method and device provided by the invention, the instructions can be efficiently and smoothly distributed and run through the at least one processing module of the processor, and different instruction distribution logics are adopted based on different types of business processing instructions through two-stage instruction distribution, so that the service processing efficiency is improved. The business processing instructions are stored in the target storage area and then scheduled to the instruction execution module, so that quick distribution and hardware support of different types of business processing instructions are realized, the instruction distribution efficiency is greatly improved, complex calculation tasks can be flexibly and efficiently processed, instruction scheduling can be performed according to the priorities of the business processing instructions, and the service processing efficiency is improved. The key business operation is ensured to be responded and processed in time, and the overall performance and the operation efficiency of the database processor are effectively improved. The technical effect of flexibly and efficiently processing the database acceleration task is achieved.
Owner:INSPUR SUZHOU INTELLIGENT TECH CO LTD

Robot control instruction analysis method and system fusing continuous instructions

The invention provides a robot control instruction analysis method and system fusing continuous instructions, and relates to the technical field of electric digital data processing, and the method comprises the steps: carrying out the quantitative evaluation of the instruction execution quality and resource consumption according to prediction data, and generating a quality score and a consumption estimation; constructing a multi-objective optimization framework based on the quality score and the consumption estimation, and determining a candidate path set; switching cost and execution time are calculated according to the candidate path set, and optimized path schemes meeting constraint conditions are screened; the instruction execution priority and the resource allocation proportion are adjusted in combination with the real-time environment data, and a scheduling strategy is generated. According to the method, environment changes and task quality are monitored in real time, instruction priorities and resource allocation are dynamically adjusted, when execution efficiency is lower than a threshold value, online updating of a prediction model is triggered, model parameters are finely adjusted to adapt to new conditions, and through continuous optimization and monitoring, the execution efficiency and task quality of an instruction sequence can be improved, and the execution efficiency is improved. And intelligent instruction scheduling and resource allocation are realized.
Owner:SHENZHEN MINRRAY IND CORP LTD

Compute-in-memory chip, instruction scheduling method, and related apparatus

The present application discloses a compute-in-memory chip, an instruction scheduling method, and a related apparatus. The compute-in-memory chip comprises an instruction memory, an instruction scheduler, and at least one compute-in-memory memory; each compute-in-memory memory comprises at least one storage array; the instruction memory is used for acquiring a first tensor instruction to be executed; and the instruction scheduler is used for scheduling, on the basis of the association relationship between the first tensor instruction and a second tensor instruction and the state of a target storage array needing to be operated for executing the first tensor instruction, the first tensor instruction to the compute-in-memory memory to which the target storage array belongs so that the compute-in-memory memory executes the first tensor instruction. According to embodiments of the present application, diversified compute-in-memory computing can be supported, efficient out-of-order execution scheduling of a tensor instruction set is achieved on the basis of the compute-in-memory chip, and the requirements of compute-in-memory technology for high concurrency and high throughput rate are met.
Owner:HUAWEI TECH CO LTD

Vector kernel module of artificial intelligence chip and operation method thereof

The invention provides a vector core module of an artificial intelligence chip and an operation method of the vector core module, which are used for improving the efficiency of the vector core module in an application situation that a vector core is a host and a tensor core module is a slave. The vector core module comprises a vector core instruction execution pipeline, a tensor core instruction processing pipeline and an instruction scheduling unit. An instruction scheduling unit performs instruction classification to distinguish vector core instructions and tensor core instructions from a thread bundle. In response to the thread bundle including a vector core instruction, the instruction scheduling unit sends the vector core instruction to a vector core instruction execution pipeline. The vector core instruction execution pipeline executes the vector core instruction and stores an execution result in the memory module. In response to the thread bundle including a tensor core instruction, the instruction scheduling unit sends the tensor core instruction to a tensor core instruction processing pipeline. The tensor core instruction processing pipeline processes the tensor core instruction and sends a processing result to the tensor core module.
Owner:SHANGHAI BIREN TECH CO LTD

Data processor, method, electronic device and storage medium

The invention provides a data processor and method, electronic equipment and a storage medium, the data processor comprises an instruction scheduler and a copy engine, the instruction scheduler comprises a first tensor access register, and the first tensor access register is used for storing sub-tensor description information transmitted externally; the replication engine comprises a first queue structure, a second tensor access register and an instruction decoder, the second tensor access register is used for storing sub-tensor description information received from the first queue structure, and the instruction decoder is used for analyzing an instruction received from the first queue structure and generating a control signal to drive data handling; wherein the instruction scheduler is configured to dynamically detect a change parameter item of the sub-tensor description information, and write the change parameter item into the first queue structure, so that the copy engine incrementally updates the sub-tensor description information in the second tensor access register based on the change parameter item, therefore, the data transmission redundancy can be reduced, and the data transmission efficiency is improved. And the instruction processing efficiency is improved.
Owner:SUZHOU YIZHU INTELLIGENT TECH CO LTD

Multi-nozzle collaborative printing method

The invention discloses a multi-nozzle collaborative printing method, and particularly relates to the technical field of printing path planning. According to the method, a three-dimensional space mesh model is constructed by collecting structural feature information of a target printing area; calculating collaborative coverage parameters of the sub-regions according to the space coverage capability and the region curvature characteristics of the nozzle; constructing a priority map based on nozzle performance and task suitability, and performing task matching by adopting a minimum cross entropy objective function to generate an initial path allocation scheme; constructing a jet printing path conflict graph, extracting conflict feature points, performing local path reconstruction by using a graph convolutional neural network, and generating an optimized jet printing path; and finally, the multiple nozzles are driven to cooperatively execute a jet printing task through time synchronization and instruction scheduling. According to the invention, the conflict rate and waiting time of the nozzles can be obviously reduced, the path planning efficiency and the jet printing precision are improved, and the method has good engineering applicability.
Owner:YIXING HUALI TINPLATE PRINTING & CAN-MAKING CO LTD

Compiler optimization method and device and storage medium

The invention discloses a compiler optimization method and device and a storage medium, and belongs to the technical field of computers. The method comprises the steps of obtaining a vector length corresponding to an access operation in a compiling language; the vector length is related to the number of elements included in the memory address to be accessed; under the condition that the vector length is not the power of the preset numerical value and the vector length is greater than a preset threshold value, splitting the access operation according to the vector length to obtain a plurality of sub-access operations; and according to the plurality of sub-access operations, optimizing repeated operations in the compilation language to obtain a compilation optimization result. By standardizing the access operation in the intermediate language of the compiler and splitting the access operation into a plurality of sub-access operations, the repeated sub-access operations in the intermediate language can be eliminated; moreover, in the instruction scheduling process after register allocation, based on the instructions corresponding to the multiple sub-access operations, access conflicts specific to hardware can be eliminated, and the accuracy of compiler optimization results is improved.
Owner:SHENZHEN INTELLIFUSION TECHNOLOGIES CO LTD

Task scheduling method, task scheduling device, electronic equipment and readable storage medium

The invention discloses a task scheduling method, a task scheduling device, electronic equipment and a readable storage medium, and relates to the technical field of task scheduling, the method comprises the steps that a first copy instruction is scheduled through an instruction scheduling unit, and the first copy instruction is used for instructing a single copy engine to carry a target tensor; receiving the first copy instruction through the task splitting unit, splitting the target tensor indicated by the first copy instruction into a plurality of sub-tensors, generating a plurality of second copy instructions, and correspondingly scheduling the plurality of second copy instructions to the plurality of copy engines, all the copying engines complete tensor carrying of the target tensor together; wherein the second copying instruction is used for indicating the single copying engine to carry one sub-tensor obtained by splitting the target tensor. The method and the device have the effects of effectively improving the load balance between the replication engines and improving the parallel efficiency and the throughput at the same time.
Owner:SUZHOU YIZHU INTELLIGENT TECH CO LTD

Device protocol adaptive analysis and instruction scheduling method of intelligent central control system

The invention relates to the technical field of computers, and discloses an equipment protocol adaptive analysis and instruction scheduling method for an intelligent central control system, and the method comprises the steps: obtaining an original communication data flow of heterogeneous equipment, and extracting a protocol feature fingerprint to recognize a protocol type; calling a corresponding pluggable protocol parser to generate a standardized instruction object; analyzing instruction semantics and dynamically calculating priorities in combination with a system load; and performing non-blocking scheduling and execution monitoring on the instruction based on a resource token bucket mechanism. The system comprises a data acquisition module, a protocol identification module, a self-adaptive analysis module, a semantic analysis module, a priority calculation module, an instruction sorting module, a token management module, a scheduling execution module, an execution monitoring module and a response generation module. According to the method, through protocol self-learning, dynamic priority quantification and multi-dimensional resource isolation scheduling, the response certainty, expansibility and operation stability of the system in a high-concurrency scene are remarkably improved.
Owner:SHENZHEN HAIWEI HENGTAI INTELLIGENT TECH CO LTD

Digital memory computing accelerator, instruction set architecture and data processing method thereof

The invention discloses a digital in-memory computing accelerator, an instruction set architecture and a data processing method thereof, and belongs to the technical field of in-memory computing. The digital memory computing accelerator comprises a chip-level module, a core-level module and a unit-level module, the chip-level module is used for acquiring a target calculation task, splitting the target calculation task into a plurality of core-level calculation tasks according to a load distribution strategy and distributing the core-level calculation tasks to the core-level module; the core-level module comprises a control unit, a calculation unit and a cache unit, and the control unit comprises an instruction fetching module, a decoding module and an instruction scheduler; an instruction fetching module reads a to-be-executed instruction corresponding to the core-level calculation task, a decoding module decodes the to-be-executed instruction to generate a control signal and sends the control signal to a calculation unit, and the to-be-executed instruction is an instruction of a calculation instruction set architecture in a digital memory; the calculation unit executes corresponding calculation operation according to the control signal; the cache unit caches intermediate data generated by executing the corresponding calculation operation.
Owner:BEIHANG UNIV

Loading storage circuit and graphics processor

The invention provides a loading storage circuit and a graphics processor, and relates to the technical field of graphics processing. The loading storage circuit comprises an instruction scheduling module, an address generation module and a data service module, and is provided with a buffer area module which comprises an operand buffer area and an effective data buffer area and is used for caching information required by address calculation and data access; the instruction scheduling module is used for collecting a data access instruction and outputting instruction information; the address generation module calculates a target access address according to the instruction information and the operand; and the data service module executes corresponding data loading or storage operation on the target cache unit. According to the scheme, the operands and the valid data are pre-cached, so that overflow and stagnation caused by inconsistent processing rhythms among modules can be avoided, and the parallel processing capability of an assembly line is improved; through the independent address calculation and data access process, the stability of the access time sequence and the data processing efficiency can be improved.
Owner:MOORE THREADS TECH CO LTD

Store instruction scheduling method and apparatus, device, and storage medium

The present disclosure provides a store instruction scheduling method and apparatus, a device, and a storage medium. The solution comprises: acquiring a plurality of target cache entries; setting a status bit register and a countdown register for each target cache entry; within any clock cycle, decrementing the countdown of the register by one, and changing the status bit set for a target cache entry into which a store instruction is written; determining whether the number of occupied target cache entries is greater than or equal to a first threshold, determining whether the countdown of the countdown register is zero, and determining whether the status bit of the status bit register set for each target cache entry is equal to a second threshold, so as to obtain a second determination result, and on the basis of the second determination result, changing the status bit and resetting the countdown; and when the set status bit is equal to the second threshold, writing the store instruction cached in each target cache entry into a cache memory. Thus, there is no need to continuously occupy a cache port and bandwidth.
Owner:BEIJING VCORE TECH CO LTD

Instruction scheduling method and system for cold chain warehouse to control multiple devices

The invention discloses an instruction scheduling method and system for a cold chain warehouse to control multiple devices, and belongs to the field of cold chain warehouse control, and the method comprises the steps: collecting environment data and device state data in the cold chain warehouse, carrying out the preprocessing of the data, and sending the data to a cloud; the cloud constructs an equipment database in a time sequence according to the timestamps in the message format, and processes data in the equipment database through a cold chain storage cooperative control model; sorting all the generated task requests according to a predefined priority, generating a task queue of a scheduling core, and generating an atomic operation instruction for specific equipment through a rolling time domain optimization algorithm; and sending the operation instruction message to the edge gateway, and sending the native command to the master control of the corresponding equipment through a communication module of the edge gateway. According to the invention, intelligent monitoring and scheduling of the equipment state are realized, and regulation and control can be carried out in time according to environmental changes.
Owner:四川参盘供应链科技有限公司

Multipath power supply parallel automatic test system and test method

The invention relates to a multi-path power supply parallel automatic test system and a test method. The multi-path power supply parallel automatic test system comprises a main control module, an instruction scheduling module, a test instrument group and a plurality of test fixtures, each test fixture is used as an independent test channel with self-defined configuration; a request signal input end of the main control module is connected with a request signal output end of the independent test channel, a test signal input end of the main control module is connected with a test signal output end of the test instrument group, and a control signal output end of the main control module is connected with a control signal input end of the instruction scheduling module; the signal output end of the instruction scheduling module is connected with the signal input end of the test instrument group, and the test instrument group is used for executing corresponding test operation on the to-be-tested power supply according to the control signal; the testing efficiency is remarkably improved, manual operation is reduced, and the basic capability of automatic parallel testing of multiple power supplies is achieved.
Owner:SHENZHEN POWER INSTR ELECTRONICS CO LTD

Ultra-wide RISC-V long vector processor

The invention relates to the technical field of processor hardware design, and discloses an ultra-wide RISC-V long vector processor, which comprises a 64-bit scalar RISC-V core and a plurality of vector clusters, wherein each vector cluster comprises an instruction dispatcher and an instruction scheduler, the instruction dispatcher receives vector instructions sent from the scalar core and distributes the instructions to idle processing channels, and the instruction scheduler controls the execution sequence of the instructions in time; each channel is connected with the mask unit, the sliding unit and the vector read-write unit through the full-interconnection crossbar switch, and the full-interconnection crossbar switch and the mask unit carry out conditional execution on elements in a vector instruction based on the mask register. According to the method, the long vector can be quickly and efficiently calculated, the expandability problem of a full interconnection structure is solved by adopting a special layered pipeline interconnection structure, and then long vector support can be carried out on an extended vector processor architecture.
Owner:SHANDONG INSPUR SCI RES INST CO LTD

Virtual power plant distributed power generation instruction distribution method and system for power distribution network

The invention discloses a virtual power plant distributed power generation instruction distribution method and system for a power distribution network, and the method comprises the steps: firstly carrying out the linearization of a power distribution network AC power flow model, carrying out the modeling of the line active power and node voltage in the power distribution network, and representing the node voltage and the line active power as the mapping of the node power variation; setting constraint conditions including a power balance constraint, a node voltage constraint, a line capacity constraint, an adjustable unit power adjustment upper and lower limit constraint and a climbing rate constraint of each adjustable unit, and constructing a power generation instruction scheduling optimization model taking the highest frequency modulation performance and the lowest carbon emission as double optimization targets; and finally, solving the power generation instruction scheduling optimization model in each control period by adopting a Nesterov momentum acceleration-based distributed algorithm to realize rapid and accurate solving of the instruction, and setting a communication condition to reduce the communication pressure during distributed information exchange and improve the convergence speed of the model. The method is high in convergence speed and high in precision.
Owner:ZHEJIANG UNIV

Instruction-level parallel scheduling method and device in deep learning compiler

The invention provides an instruction-level parallel scheduling method and device in a deep learning compiler, and the method comprises the steps: decomposing a calculation graph of a deep learning task, decomposing a task represented by each calculation graph node on the calculation graph into a hardware calculation instruction according to a current task decomposition scheme, and obtaining an instruction set of the calculation graph; according to the data flow dependency relationship between the nodes on the calculation graph, obtaining a dependency relationship graph between the instructions in the instruction set; dividing the hardware calculation instructions in the instruction set into each sub-core in the processor; according to the dependency graph, determining a directed acyclic graph containing the dependency among the hardware calculation instructions on the sub-cores in the processor, and according to an emission model of the processor and the reciprocal throughput rate of the hardware calculation instructions, determining the execution sequence of the hardware calculation instructions divided to the sub-cores on the sub-cores, and the processor executes the instruction scheduling to obtain an execution result of the deep learning task.
Owner:INST OF COMPUTING TECH CHINESE ACAD OF SCI

Instruction scheduling system and method and electronic equipment

The invention discloses an instruction scheduling system and method and electronic equipment, and relates to the technical field of computers, a dispatch module writes a to-be-scheduled instruction and a renamed register index combination into a dispatch queue, and a dependency check module constructs a dependency linked list according to the dispatch queue so as to determine an instruction execution sequence and send the instruction according to the instruction execution sequence; according to the method, the sending sequence and the execution sequence of the instruction with RAW dependency are ensured to be matched, so that the technical problems of low RAW dependency processing efficiency, long instruction waiting time and waste of pipeline instruction storage resources in the instruction scheduling of the superscale out-of-order processor can be solved, and the execution delay of the instruction is reduced.
Owner:SHANDONG YUNHAI GUOCHUANG CLOUD COMPUTING EQUIP IND INNOVATION CENT CO LTD

Intelligent task planning and control method and system for multifunctional bionic robot

The invention discloses a multifunctional bionic robot intelligent task planning and control method and system, and the method comprises the following steps: S1, collecting and preprocessing multi-source heterogeneous sensing data, and constructing a multi-modal event flow; s2, inputting the multi-modal event stream into an event perception Transform model, and generating a semantic event vector sequence; s3, performing tasks based on the semantic event vector sequence, and constructing a structured task request graph; s4, constructing an optimization controller graph based on the idle execution module and the structured task request graph; s5, inputting the structured task request graph and the optimization controller graph into a distributed evolution graph control module to generate an optimal scheduling mapping relation; s6, generating a control instruction sequence based on the optimal scheduling mapping relation, distributing the control instruction sequence to an idle execution module, and executing action instruction scheduling and control parameter configuration operation; and S7, updating the multi-mode event flow, and repeatedly executing the steps S2 to S6. According to the invention, autonomous task understanding, efficient scheduling and closed-loop control of the bionic robot in a complex task scene are realized.
Owner:ZHIMOU (ZHEJIANG) TECHNOLOGY DEVELOPMENT CO LTD

Universal computing unit and instruction scheduling method

The invention provides a general purpose computing unit and an instruction scheduling method for the general purpose computing unit. The general-purpose computing unit comprises a plurality of execution units, wherein each execution unit is used for executing an operation instruction by taking a thread bundle as a unit; and an instruction scheduler for determining a priority of the plurality of operation instructions based on whether the multiplexing flag exists, and scheduling each operation instruction to one of the plurality of execution units based on the priority; wherein each execution unit comprises one or more reuse registers, each reuse register is used for registering an operand of a previous operation instruction and a label, and the label is used for indicating an address and a thread bundle of the operand; and the instruction decoding unit is used for decoding the received operation instruction to determine the reading position of the operand of the operation instruction. The instruction is preferentially scheduled by adding a reuse mark to the whole operation instruction, so that the hit rate of the operand is improved, and excessive instruction bit fields do not need to be occupied.
Owner:SHANGHAI BIREN TECH CO LTD

Serial port instruction intelligent processing method and device based on multi-frame cache and storage medium

The invention relates to a serial port instruction intelligent processing method and device based on multi-frame cache and a storage medium. The method comprises the steps that a system is initialized; starting serial port interruption to receive a serial port AT instruction; the received instruction is analyzed; after analysis is completed, the AT instructions are further classified; after instructions are classified and stored in corresponding queues, the system adopts a priority scheduling algorithm to determine which instruction is processed preferentially; and the system executes corresponding operation according to the specific content of the instruction. According to the method, a multi-level cache pool technology is adopted, a virtual cache queue is constructed on a software layer, the equivalent cache capacity is expanded, and the situation that a hardware buffer overflows or instructions are lost due to insufficient process processing capacity is avoided; a priority instruction scheduling algorithm is introduced, instruction grades are dynamically divided, key instructions are ensured to be processed preferentially, and the problem that a traditional single-frame mode cannot respond in time when the MCU is in a high load state is solved; and meanwhile, enhanced frame identification of the protocol is realized, and the problem of insufficient error correction capability of a traditional mode is solved.
Owner:TIANDI CHANGZHOU AUTOMATION +1

Power resource dynamic scheduling and safe tracing system and method based on edge computing and block chain fusion

The invention provides a power resource dynamic scheduling and security tracing system and method based on edge computing and block chain fusion, and the system employs a hierarchical collaborative architecture, and comprises an edge computing layer, a block chain layer and a cross-layer security module. The edge layer collects photovoltaic output and load demand data, and predicts short-term output and load through LSTM. The ACEGA algorithm is combined with the micro-service priority and the edge node resource state to generate a scheduling instruction; and the scheduling instruction hash value is subjected to uplink evidence storage, the energy storage device is triggered to charge and discharge after verification of the smart contract, and meanwhile, a green power consumption voucher is recorded. According to the invention, power resource dynamic scheduling optimization, data credibility evidence storage and security protection are realized, the problems of distributed new energy consumption, energy storage resource configuration and power transaction tracing are solved, and the defects of low edge computing scheduling efficiency, high data security risk and insufficient cross-technology collaboration in the prior art are overcome.
Owner:CHINA YANGTZE POWER

Python byte code obfuscation method and system based on dynamic screening and random replacement

The invention provides a Python bytecode obfuscation method and system based on dynamic screening and random replacement, and the method comprises the steps: calculating a plurality of safety indexes and a comprehensive safety index SDI of bytecode files, and screening out the bytecode files which have high safety requirements and need to be obfuscated; a linear congruence generator is constructed, a pseudo-random number sequence is provided for obfuscation operation of byte codes, and operation code dynamic replacement obfuscation operation is performed in byte code instructions by generating random seeds, constructing random dynamic operation code mapping tables and applying the independent operation code mapping tables to different byte code files needing to be obfuscated; and performing obfuscation state detection on the obfuscated bytecode file, if the detection result is the obfuscated bytecode file, modifying bytecode analysis and instruction scheduling logic, increasing a runtime de-obfuscation process, embedding an Opcode reflection mechanism, re-compiling to generate a customized Python interpreter supporting the execution of the obfuscated bytecode, and executing the obfuscated bytecode file. According to the invention, byte codes can be confused.
Owner:UNIV OF SCI & TECH BEIJING

Hardware resource scheduling method and hardware scheduler used in graphics processing unit

The invention provides a hardware resource scheduling method and a hardware scheduler used in a graphics processing unit. The hardware resource scheduling method used in the graphics processing unit comprises the following steps: for any hardware computing unit in the graphics processing unit, acquiring task or instruction queue state information indicating whether a task or instruction queue of the hardware computing unit is in an idle state or a full state from the hardware computing unit; receiving, from a driver system of the graphics processing unit, priority ranking information indicating a priority ranking between respective hardware computing units inside the graphics processing unit; and based on the task or instruction queue state information and the priority ranking information of the hardware computing unit, judging whether the task or instruction to be scheduled to the hardware computing unit is scheduled to the hardware computing unit, and if so, scheduling the task or instruction to be scheduled to the hardware computing unit to the hardware computing unit.
Owner:MOFFETT AI TECHNOLOGY SHENZHEN CO LTD

Bluetooth sound box voice control system based on Internet of Things

The invention discloses a Bluetooth sound box voice control system based on the Internet of Things, and particularly relates to the technical field of the Internet of Things, and the system comprises an audio collection module which collects a user voice signal, a voice recognition processing module which recognizes and extracts an instruction keyword, and an instruction analysis module which generates control instruction information. The instruction scheduling module calls a local program according to the control instruction information or sends the control instruction information to the cloud for processing through the Internet of Things communication module, the execution control module completes corresponding function operation according to a scheduling result, and data interaction and control between the Bluetooth sound box and the cloud server as well as between the Bluetooth sound box and the smart home equipment are realized; the high-frequency instruction words are stored in the local instruction scheduling unit, local identification and rapid processing are achieved, dependence on network transmission is remarkably reduced, the problem that response time is prolonged due to network fluctuation or delay is solved, the high-frequency instruction words and the low-frequency instruction words are subjected to hierarchical storage and dynamic scheduling, and the high-frequency instruction words and the low-frequency instruction words are dynamically scheduled. The frequent calling of cloud services is effectively reduced, and the overall data transmission demand and energy consumption of the system are reduced.
Owner:SHENZHEN ZUNTE DIGITAL CO LTD

Apparatus and method for configuring a warp of cooperative threads in a vector processing system

The present invention relates to an apparatus and method for configuring a warp in a vector operation system. The apparatus includes: general-purpose registers; an arithmetic logic unit; a warp instruction scheduler; and a plurality of warp resource registers. The warp instruction scheduler allows each of a plurality of warps to include a part of relatively independent instructions in the program core according to the warp allocation instruction in the program core, and allows each of the plurality of warps to access all or specified local data in the general-purpose registers through the arithmetic logic unit according to the configuration during software execution, and completes the operations of each of the above warps through the arithmetic logic unit. The present invention can more widely adapt to different applications, such as big data, artificial intelligence operations, etc., through the above-described components that can allow software to dynamically adjust and configure general-purpose registers for different warps.
Owner:SHANGHAI BIREN TECH CO LTD

Microgrid two-stage robust optimization scheduling method, device, equipment and medium

The embodiment of the invention discloses a micro-grid two-stage robust optimization scheduling method and device, equipment and a medium, and the method comprises the steps: building a mixed integer linear programming model for a target micro-grid according to the equipment information of the target micro-grid by taking the minimization of the total operation cost of the target micro-grid as a first target function; solving the mixed integer linear programming model to obtain a first scheduling scheme; minimizing the adjustment cost in the target scene as a second target function, and constructing a robust optimization scheduling model for the target micro-grid, the adjustment cost including the adjustment abandoning cost and the load reduction cost in the target scene; based on the first scheduling scheme, solving the robust optimization scheduling model by using a preset target algorithm to determine a target scheduling scheme; and sending a control instruction to the micro-grid according to the target scheduling scheme, and scheduling the target micro-grid to operate. According to the scheme, operation of the micro-grid under significant uncertainty can be managed, and multi-device collaborative robust optimization scheduling of the micro-grid is realized.
Owner:NANJING POWER PROPERTY MANAGEMENT CO LTD +1

Ai-based techniques for guiding an instruction scheduler

Using artificial intelligence (AI)-based techniques to guide instruction scheduling in a compiler can improve the efficiency and code generation quality of the compiler. AI-guided scheduling of a basic block of a computer program can include obtaining first and second representations of the basic block; selecting K instruction scheduling procedures from a set of N instruction scheduling procedures based on analysis of the first representation of the basic block by a model, where 1≤K<N and N≥2; generating K candidate schedules of the basic block, including applying the K instruction scheduling procedures to the second representation of the basic block, and ordering the instructions of the second representation of the basic block in accordance with a candidate schedule included in the K candidate schedules.
Owner:ADVANCED MICRO DEVICES INC