Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

92 results about "Network processing unit" patented technology

Network processors are typically software programmable devices and would have generic characteristics similar to general purpose central processing units that are commonly used in many different types of equipment and products.

High-speed I / O virtualization server based on FPGA and remote forwarding method

The invention relates to the technical field of computer hardware and network communication, in particular to a high-speed I / O (input / output) virtualization server based on an FPGA (field programmable gate array) and a remote forwarding method.The high-speed I / O virtualization server comprises an FPGA chip, a network processing unit, an I / O protocol processing unit, a protocol conversion and scheduling unit, a time sequence and transaction management unit and a data buffering and flow control unit are integrated in the FPGA chip; the optical fiber Ethernet interface module comprises a physical layer interface and an MAC layer processing and network configuration interface; the local I / O interface module comprises a USB interface circuit, a serial port interface circuit, a PCIe interface circuit and an interface expansion connector; the power supply and monitoring module comprises a multi-path power supply management system, a state monitoring and diagnosis system and an LED indication system; according to the self-adaptive Baud rate detection circuit, the Baud rate can be accurately locked under the condition that data with less than two characters is received, the number of data bits, the type of a check bit and the length of a stop bit are automatically identified, and compared with a traditional microcontroller scheme, the self-adaptive Baud rate detection circuit has higher precision and higher detection speed.
Owner:梁琦

Edge gateway computing power sharing method based on MQTT protocol

The invention provides an edge gateway computing power sharing method based on an MQTT protocol. An edge terminal sends a computing power request to an edge gateway through the MQTT protocol, wherein the computing power request comprises a computing task type identifier and a priority identifier; after receiving the computing power request, the edge gateway dynamically schedules heterogeneous computing resources of a built-in neural network processing unit (NPU) and a built-in central processing unit (CPU) according to a computing task type identifier and a priority identifier; and the edge gateway executes a corresponding calculation task by using the scheduled calculation resource, feeds back a processing result to the edge terminal initiating the request through an MQTT protocol, and synchronizes own calculation power state information.
Owner:FUJIAN ZHONGRUI NETWORK CO LTD

Flexible and scalable thermal test vehicle design for electronics cooling solutions

PCT designated stageWO2026084737A1Analog circuit testingDigital circuit testingTransistor arrayNetwork processing unit
The density and power consumption of modern integrated circuits, such as Graphic Processing Units (GPUs), Central Processing Units (CPUs), and Network Processing Units (NPUs) is growing rapidly, which necessitates designing advanced cooling systems. Existing solutions for characterizing and validating these cooling system are inadequate. A flexible, scalable Thermal Test Vehicle (TTV) is disclosed which is based on an array of power transistors, measurement / control circuitry, and onboard computer. The TTV is configured for characterizing the performance of electronic cooling solutions under a variety of operating conditions.
Owner:RGT UNIV OF CALIFORNIA +1

Neural network processing unit scheduling system and method

The invention provides a neural network processing unit scheduling system and method, relates to the field of artificial intelligence accelerators, and can set a task progress register in a neural network processing unit for recording the task processing progress of the neural network processing unit for processing and calculating sub-tasks and exposing the task processing progress to a software layer. A scheduler and a global shared task queue can be arranged in the runtime module. The scheduler is used for scheduling the calculation subtasks to the neural network processing units, screening the slow processing units according to the task processing progress of the neural network processing units, determining the remaining subtasks which are not completed by the slow processing units, taking the remaining subtasks as the calculation subtasks again, and sending the calculation subtasks to the neural network processing units; and re-issuing to the idle neural network processing unit through the global shared task queue. Therefore, the remaining sub-tasks which are not completed by the low-speed processing unit can be re-issued to the idle neural network processing unit, so that idle computing resources can be avoided, and the processing efficiency is improved.
Owner:CCORE TECH CO LTD

Large model reasoning system and method based on combination of flash memory controller and NPU

The invention relates to the technical field of cross of storage controllers and artificial intelligence acceleration, and discloses a large model reasoning system and method based on combination of a flash memory controller and an NPU (Network Processing Unit), and the large model reasoning system comprises the flash memory controller, the NPU and a flash memory array, the flash memory controller integrates a host interface module, a flash memory interface module, an independent AI acceleration interface module and an AI management engine, and the AI management engine autonomously completes NPU initialization, model weight direct loading, KV Cache hierarchical management, RAG knowledge base retrieval and model switching; the flash memory array is divided into a firmware partition, an AI special partition and a user storage partition, and different data storage requirements are met. According to the method, large model reasoning with low delay and low CPU dependence can be realized, and the model loading delay is reduced from 5-30 seconds to lt; after 500 milliseconds, the CPU occupancy rate of the host is reduced from 15-25% to lt; 2%, and concurrent operation of 4-8 models is supported. According to the invention, integration of storage and calculation is realized, and edge end, data center and mobile equipment scenes are adapted.
Owner:YEESTOR MICROELECTRONICS CO LTD

Hybrid precision interface in neural network processing units

An apparatus and method are disclosed for a fabric interface of a processing entity (PE) within a PE array optimized for neural network computations. The fabric interface includes a channel extension circuit and a channel compression circuit. The channel extension circuit is configured to receive first data from another PE, extend the channel dimension of the first data to produce extended data with an increased dimension that aligns with the minimum data granularity of an associated memory, and store the extended data in the memory. Conversely, the channel compression circuit is configured to fetch second data from the memory, compress the channel dimension of the second data to obtain compressed data with a reduced dimension, and output the compressed data from the PE.
Owner:MOFFETT TECH CO LTD

Epilepsy prediction system based on multivariate weighted joint recursion and graph attention network

The application provides an epilepsy prediction system based on multi-element weighted joint recursion and graph attention network, which can be applied to the fields of biomedical signal processing and artificial intelligence technology. The system comprises a brain function imaging module configured to acquire electroencephalogram signal data of a target object under authorization of the target object; a processor comprising a multi-element weighted joint recursion processing unit configured to obtain a phase space trajectory vector of each channel of the electroencephalogram signal data according to the electroencephalogram signal data, construct a recursion graph according to the phase space trajectory vector of each channel, obtain a channel correlation coefficient between any two channels based on the recursion graph of each channel, and construct a weighted adjacency matrix representing the brain function of the target object according to the channel correlation coefficient between any two channels; and a graph attention network processing unit configured to perform spatiotemporal feature processing according to the weighted adjacency matrix and the electroencephalogram signal data and output an epilepsy seizure prediction result.
Owner:TIANJIN POLYTECHNIC UNIV

Network device for performing map packet forwarding with the aid of network processing unit and / or software map information table

A network device includes a storage device and a network processing unit (NPU). The storage device stores a mapping of address and port (MAP) information table, wherein the MAP information table includes a plurality of table entries that store a plurality of second internet protocol (IP) headers each complying with a second IP version. The NPU receives a first packet with a first IP header that complies with a first IP version different from the second IP version, and generates and forwards a second packet that includes a second IP header retrieved from the storage device and at least a portion of the first packet.
Owner:AIROHA TECH (SUZHOU) LTD

Airborne satellite communication processing terminal

The invention relates to the technical field of avionics communication. The invention provides an airborne satellite communication processing terminal. The airborne satellite communication processing terminal comprises a data control unit, a satellite data processing unit and a network processing unit, the satellite data processing unit sends data sent by a satellite to the data control unit; the data control unit performs communication protocol analysis and data format conversion on the received data; the data control unit sends the data subjected to communication protocol analysis and data format conversion to the network processing unit; the network processing unit encapsulates the received data into an Ethernet frame and then sends the Ethernet frame to the built-in network; the built-in network sends the received data to the data control unit through the Ethernet communication interface; and the data control unit encrypts the received data, converts the encrypted data into satellite signals and sends the satellite signals to a satellite. Satellite communication is converted into a universal Ethernet communication interface, and the compatibility problem of multi-platform interfaces is solved.
Owner:SHENYANG HANGSHENG TECH CO LTD

Network configuration method and apparatus, related device, storage medium and computer program product

The application discloses a network configuration method and device, related equipment, a storage medium and a computer program product, and applies to the technical field of cloud computing. The method comprises the following steps: in response to an application request of a user applying for a bare metal server, network configuration information is respectively allocated to a plurality of neural network processing units (NPUs) in a target bare metal server according to link layer discovery protocol information of the plurality of NPUs, and parameter plane network information of the target bare metal server is obtained; based on the parameter plane network information, a configuration driver file of the target bare metal server is updated; and the target bare metal server is controlled to execute the configuration driver file, and the configuration driver file triggers the target bare metal server to perform network configuration on the plurality of NPUs when being executed.
Owner:CHINA MOBILE (SUZHOU) SOFTWARE TECH CO LTD +1

Intelligent depth reconstruction architecture and method for strip-shaped scintillator

The invention relates to an intelligent depth reconstruction architecture and method for a strip-shaped scintillator, and the architecture specifically comprises a position sensitive strip-shaped scintillator which is used for converting incident radiation particles into photon signals, and converting the photon signals into pulse signals through silicon photomultipliers coupled to the two ends of the scintillator; the analog front-end circuit is used for amplifying and shaping the pulse signal, then completing high-speed sampling through a high-speed analog-to-digital conversion circuit module and converting an analog signal into a digital signal; the field programmable gate array data processing module is used for digital imaging preprocessing, amplitude extraction and time marking operation, and meanwhile, key information is added to each event data; an auto-encoder model based on an artificial neural network is deployed in a data processing module of the neural network processing unit, and the neural network processing unit is used for carrying out feature extraction and mode matching analysis on the input signals and outputting reconstruction parameters corresponding to interaction positions in the strip-shaped scintillators. And localized intelligent processing of the nuclear pulse signal is realized.
Owner:CHENGDU UNIVERSITY OF TECHNOLOGY

Ultrasonic guided wave online defect detection method for shell-and-tube heat exchanger and heat exchanger

The invention provides an ultrasonic guided wave online defect detection method for a shell-and-tube heat exchanger and the heat exchanger. The method comprises the steps that a sensor array is permanently installed in a specific annular area of the outer surface of a tube plate; exciting and receiving the non-dispersion torsional guided wave to obtain an echo signal; performing noise reduction, frequency dispersion compensation and feature extraction on the signal; intelligently identifying defect types by using a pre-trained attention mechanism neural network model, quantifying the size and evaluating the confidence coefficient; using a fatigue or corrosion model to predict the remaining life based on the defect size and generating a report; a corresponding shell-and-tube heat exchanger is integrated with the detection system, the detection system comprises a sensor array and a built-in ultrasonic circuit, and through a network processing unit NPU special chip, an NB-IoT communication module and an embedded monitoring unit managed by an intelligent power supply, non-stop and online automatic defect detection and safety early warning of the heat exchanger are achieved.
Owner:四川凌耘建科技有限公司

Quantum error correction hardware decoder and chip

A quantum error correction hardware decoder and chip, relating to the fields of artificial intelligence and quantum technology, includes: an instruction storage device configured to store computer instructions; a control unit configured to read the computer instructions from the instruction storage device and control at least one neural network processing unit based on the computer instructions; the at least one neural network processing unit configured, in response to control by the control unit, to decode error syndrome information of a quantum circuit based on a neural network model to obtain an output result of the neural network model, where the error syndrome information indicates an error syndrome obtained by performing error measurements on the quantum circuit; and an error batch search unit configured to determine error information based on the output result, where the error information indicates a quantum bit in the quantum circuit where an error has occurred and an error type corresponding to the error. The programmable hardware architecture provided in this application can ensure decoding performance while sufficiently reducing decoding delay and improving scalability and flexibility.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

An automatic sorting robot system based on embedded vision module

This invention discloses an automated sorting robot system based on an embedded vision module, belonging to the field of computer vision technology. Its features include: an image acquisition module for acquiring image information of workpieces to be sorted, comprising at least two cameras and two lasers emitting parallel laser lines to form a binocular structured light ranging field; an embedded main control module electrically connected to the image acquisition module, comprising a neural network processing unit (NPU), a model inference unit, a depth calculation unit, and a coordinate transformation unit; a sorting execution module communicatively connected to the embedded main control module for performing gripping and placement operations based on the workpiece's three-dimensional coordinates and dimensions output by the embedded main control module; and a central server module connected to several embedded main control modules via an industrial Ethernet or message queue communication protocol.
Owner:河北工业职业技术大学

Multi-network processing unit cooperative MPLS (Multi-Protocol Label Switching) message processing method and system

The invention discloses a multi-network processing unit cooperative MPLS message processing method and system, the system at least comprises a first network processing unit located at an entrance and a second network processing unit located at an exit port, and the network processing units are sequentially connected through internal loopback ports; wherein the first network processing unit is used for identifying an MPLS (Multi-Protocol Label Switching) message and adding a custom mark and a traceability VLAN to the MPLS message; redirecting the MPLS message added with the custom mark and the traceability VLAN to an internal loopback port, and enabling the MPLS message to finally reach a second network processing unit after loopback; the second network processing unit is used for further processing according to the custom mark in the MPLS message, including stripping and complementing the label and / or modifying the traceability VLAN; and redirecting the processed MPLS message to a final output port. According to the method, the function limitation and the performance bottleneck of a single processing unit are effectively overcome, and the efficient and flexible forwarding of the complex MPLS flow is ensured.
Owner:YUNHE ZHIWANG (SHANGHAI) TECHNOLOGY CO LTD

Dental triage and clinical risk assessment system

A dental triage and clinical risk assessment system (100), the system (100) comprising: a first user device (102) configured to receive patient-related dental information for remote dental assessment; a first user interface (104) within the first user device (102) that is configured to guide the patient's interaction by presenting structured symptom questionnaires, image acquisition instructions, and status notifications; a communication network (110) that is connected to the first user device (102) and configured to securely transmit patient-related dental information and control signals; a processing unit (112) connected to the first user device (102) via the communication network (110) and configured to perform automated dental triage and clinical risk assessment through coordinated data analysis decision logic and routing operations, wherein the processing unit (112) comprises: a data acquisition module (114) configured to receive patient-related dental information from the first user device (102); a data preprocessing module (116) configured to validate, normalize and structure the received dental information for analytical processing; a triage intelligence module (124) configured to analyze the structured dental information to generate a quantified urgency score and risk classification level using combined analytical logic; a decision support module (130) configured to assign the urgency score and risk classification level to a recommended dental care pathway; a nursing navigation module (134) configured to generate routing instructions corresponding to emergency referrals, scheduled consultations, teleconsultations or automated guidance; a learning and analysis module (144) configured to update internal analytical parameters based on confirmed clinical results; a storage unit (154) connected to the processing unit (112) and configured to store patient data, analytical results, priority points, routing decisions and result feedback for continuous system operation; a second user device (158), connected to the processing unit (112) via the communication network (110) and configured to receive triage outputs and recommendations for care pathways; and a second user interface (160) in the second user device (158), configured to display risk indicators, urgency points, patient data summaries and override controls for interaction with the clinician.
Owner:ALOTHMANI OSAMA S DR +7

AI processor and method based on storage and calculation integration, three-dimensional integration and memory sharing

The invention relates to an AI processor based on storage and calculation integration, three-dimensional integration and memory sharing and a control method of the AI processor. The calculation unit is coupled with the at least one storage unit through a three-dimensional integration technology, and the calculation unit comprises a neural network processing unit which is used for executing neural network correlation calculation; the interface device is arranged on the computing unit and is used for performing communication coupling on the AI processor and external main control equipment; the control module is arranged in the storage unit and used for switching the storage unit between a first working mode and a second working mode, and in the first working mode, the storage unit establishes communication coupling with external main control equipment through the interface device and executes storage operation in response to the read-write request; in the second working mode, the storage unit establishes communication coupling with the neural network processing unit, and responds to a calculation request of the neural network processing unit to assist in executing calculation operation; and the functions of high-performance AI acceleration and general storage are considered through working mode switching and time division multiplexing.
Owner:WUXI MICRONANO CORE ELECTRONIC TECH CO LTD +1

Photovoltaic power generation power edge prediction method

The invention discloses a photovoltaic power generation power edge prediction method, and the method comprises the steps: firstly collecting photovoltaic output power and multi-dimensional meteorological data, and constructing a training set and a test set after normalization; then, key features are extracted by adopting a random forest method based on the training set, and a prediction model is built and trained by utilizing a BiTCN-BiGRU hybrid neural network architecture; performing pruning optimization on the trained model, converting the model into an RKNN format, performing quantification processing and performance evaluation on the model, and performing further optimization according to an evaluation result to obtain a lightweight model; and finally, the lightweight model is deployed in an edge energy management unit, hardware acceleration is performed by means of a neural network processing unit, and millisecond-level online reasoning and real-time prediction of the photovoltaic power generation power are realized. Through fusion of feature selection and a bidirectional time sequence convolution-recurrent neural network, the limitation of a traditional method in processing high-dimensional features is effectively overcome, the prediction precision and generalization ability are improved, and efficient and low-delay power prediction is realized through edge side deployment.
Owner:SHENZHEN KAIFA TECH (CHENGDU) CO LTD

Writing analysis and integrity evaluation system and method thereof

The present disclosure relates to a writing analysis and integrity evaluation system and method thereof (100) comprising a user device (102), a user interface (104) integrated within the user device (102), a communication network (106) configured to transmit the writing sample from the user device (102) and return analyzed data from a plurality of system components to the user device (102), a processing unit (108) operatively connected to the user device (102) through the communication network (106), the processing unit (108) configured to manage input reception, perform analysis, and generate feedback, the processing unit (108) comprising; a natural language processing module (110), a coherence evaluation module (112), a grammar and syntax module (114), a fluency scoring module (116), a structural evaluation module (118), a similarity comparison module (120), a feedback generation module (122), an integrity scoring module (124), a storage unit (126) connected to the processing unit (108).
Owner:JOHNSON TIMOTHY T

Network device using network processing unit and hardware acceleration circuit to meet speed test requirements of high-speed network and associated network speed test method

A network device includes a storage device, a central processing unit (CPU), a hardware acceleration circuit, and a network processing unit (NPU). The storage device stores program codes. The CPU loads and executes the program codes to deal with a control function of a network speed test. The hardware acceleration circuit provides hardware-accelerated packet forwarding. The NPU interacts with the control function performed by the CPU, and deals with processing of data packets used for the network speed test. Transmission of the data packets between the network device and another network device is performed through the NPU and the hardware acceleration circuit, without intervention of the CPU.
Owner:AIROHA TECH (SUZHOU) LTD

Eyeball tracking and positioning method based on artificial intelligence large model and automatic high-risk area (blood vessel and like) avoidance

An AI-driven ophthalmic treatment system (20) includes a laser radiation source (48), a controller (44) integrated with an artificial intelligence processing unit, a neural network processing unit (NPU), an electronic control unit, and a storage module. The controller is configured to perform the following artificial intelligence-based operations: automatically identify and calibrate a plurality of local target regions (84) within an eyeball region (25) of a patient (22) by a deep learning model, and assign an appropriate predetermined laser energy to each target region, respectively, using a machine learning algorithm; controlling the radiation source to perform laser irradiation on at least the first target area; after the irradiation of the first target area is completed, recognizing the change of an eyeball structure through a real-time image analysis and computer vision method; and based on the identified structural change, dynamically inhibiting or adjusting predetermined energy output corresponding to the non-irradiated second target area through an AI decision model. Other implementation modes adopting artificial intelligence and machine learning technologies are also covered.
Owner:SMART EYES (BEIJING) TECHNOLOGY CO LTD

Scheduling controller and method for neural network processing unit

A scheduling controller and method for neural network processing unit are disclosed. The neural network processing unit is used for running a neural network model and comprises multiple cores and each core comprises multiple processing elements (PEs). The scheduling controller comprises a top level module including a top RISC-V CPU with Vector extension (RVV) and a chip level scheduler (ChLS), a core level module including a core RVV and a core level scheduler (CoLS), and a PE level module including a PE RVV and a PE level scheduler (PELS).
Owner:MOFFETT TECH CO LTD

Method for processing multiple data, real display device, equipment, medium and product

The application provides a multi-path data processing method, a real display device, equipment, a medium and a product. Pre-set row data output by at least two image signal processors is acquired, and the pre-set row data output by the at least two image signal processors is processed in parallel through network reasoning. The demand for image segmentation processing of the whole image is cancelled to improve the accuracy of image processing, and the requirement for the overall computing power and data cache space of the system is reduced. In addition, the pre-set row data output by at least part of the image signal is divided into multiple pieces, and network reasoning is performed on each piece of data according to a pre-set strategy. The pre-set strategy is that the mobile terminal starts from a first end, moves to a second opposite end, returns to the first end after reaching the second end, and continues until the piece data reasoning is completed. In this way, the memory overhead of the neural network processing unit is reduced, the ability of a single neural network processing unit to process multiple paths of image data in parallel is improved, high frame rate and high resolution are supported, and network reasoning delay is reduced.
Owner:GRAVITYXR ELECTRONICS & TECH CO LTD

Network packet processing device with explicit congestion notification marking aided by network processing unit and related network packet processing method

A network packet processing device includes a hardware-accelerated forwarding circuit and a network processing unit (NPU). The hardware-accelerated forwarding circuit is used to receive a plurality of L4S packets from a network port and send the plurality of L4S packets to another network port through hardware-accelerated forwarding without intervention of a central processing unit (CPU). The NPU is used to perform ECN marking on at least a portion of the plurality of L4S packets.
Owner:AIROHA TECH (SUZHOU) LTD

Cross-domain data transmission method and system and vehicle

The invention relates to the technical field of data transmission, and discloses a cross-domain data transmission method and system and a vehicle, each domain controller and a central controller of the cross-domain data transmission method and system are each provided with a network processing unit which undertakes protocol conversion, data encapsulation and cross-domain routing functions, end-to-end protocol conversion is achieved, and the data transmission efficiency is improved. A first network processing unit in the first domain controller encapsulates communication data sent by the intra-domain interface into a target protocol message, different protocols are converted into standard protocols for Ethernet transmission, bus redundancy of different protocols is reduced, a traditional bus is replaced by an Ethernet wire harness, data transmission efficiency is improved, and data transmission cost is reduced. And the first network processing unit sends the target protocol message to the Ethernet and transmits the target protocol message to a second network processing unit in the second domain controller or a third network processing unit in the central controller, so that the second network processing unit or the third network processing unit analyzes the received target protocol message and sends the analyzed target protocol message to the Ethernet. And obtaining communication data.
Owner:CHONGQING CHANGAN AUTOMOBILE CO LTD

An edge computing-based security camera interference target filtering system

The application discloses an edge-computing-based security camera interference target filtering system, and belongs to the technical field of intelligent security protection. The system comprises an edge computing module, a multi-modal feature extraction module, an adaptive feature fusion module, an interference target classification module, a real-time filtering execution module and an edge-cloud collaborative optimization module. The edge computing module is arranged at the security camera end and integrates a lightweight neural network processing unit. The multi-modal feature extraction module simultaneously extracts visual features, motion features and environment perception features. The adaptive feature fusion module dynamically allocates the weights of different modal features by adopting an attention mechanism. The interference target classification module constructs a multi-level classification system. The real-time filtering execution module adopts a three-level filtering strategy. The edge-cloud collaborative optimization module realizes global optimization and parameter updating of the model. The application effectively solves the problems of low single-feature recognition accuracy, large network delay and fixed threshold value that cannot adapt to scene changes in the prior art.
Owner:ZHEJIANG CHANGCHUN TECH CO LTD

High-efficiency GPNPU (General Purpose Network Processing Unit) chip architecture integrating memory and calculation

The invention discloses a high-efficiency GPNPU (General Purpose Network Processing Unit) chip architecture with integrated memory and calculation, which belongs to the technical field of NPU chip architectures and comprises a plurality of array PE (Provider Edge) units, each PE unit comprises a local memory and a calculation unit and can independently execute vector and tensor calculation tasks, and the plurality of PE units are connected through data channels to realize direct exchange and sharing of data. Therefore, dependence on an external memory is reduced, and data processing efficiency is improved. According to the method, various existing AI network models can be supported, particularly, hardware, large-scale and instruction processing is carried out on operators in a recently popular LLM (Large Language Model), the AI calculation requirement of a large data volume is met, parallel calculation is achieved through specific instruction combination, the AI calculation efficiency and speed are improved, and complex AI tasks can be rapidly processed.
Owner:LONGQINGWEI (SHANGHAI) INTELLIGENT TECHNOLOGY CO LTD +1

Storage controller, storage system, computer device and model reasoning method

The invention is suitable for the technical field of artificial intelligence, and relates to a storage controller, a storage system, computer equipment and a model reasoning method. The storage controller comprises a controller chip, and the controller chip comprises a task input and output unit which is used for receiving a task execution instruction from a calculation unit; the neural network processing unit is used for executing a calculation task of at least one expert network in the hybrid expert model according to the task execution instruction to obtain a first calculation result; and the task input and output unit is also used for writing the first calculation result into a destination address indicated by the task execution instruction, and is used for the calculation unit to obtain a reasoning result in combination with a second calculation result of other expert networks in the self-execution hybrid expert model. The method can meet the performance requirements of a large-scale model, especially a hybrid expert model, for high-concurrency and low-delay reasoning.
Owner:YEESTOR MICROELECTRONICS CO LTD