Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

89 results about "Instruction memory" patented technology

Memory reference instruction An instruction that has one or more of its operand addresses referring to a location in memory, as opposed to one of the CPU registers or some other way of specifying an operand. Pick a style below, and copy the text for your bibliography.

Compute-in-memory chip, instruction scheduling method, and related apparatus

The present application discloses a compute-in-memory chip, an instruction scheduling method, and a related apparatus. The compute-in-memory chip comprises an instruction memory, an instruction scheduler, and at least one compute-in-memory memory; each compute-in-memory memory comprises at least one storage array; the instruction memory is used for acquiring a first tensor instruction to be executed; and the instruction scheduler is used for scheduling, on the basis of the association relationship between the first tensor instruction and a second tensor instruction and the state of a target storage array needing to be operated for executing the first tensor instruction, the first tensor instruction to the compute-in-memory memory to which the target storage array belongs so that the compute-in-memory memory executes the first tensor instruction. According to embodiments of the present application, diversified compute-in-memory computing can be supported, efficient out-of-order execution scheduling of a tensor instruction set is achieved on the basis of the compute-in-memory chip, and the requirements of compute-in-memory technology for high concurrency and high throughput rate are met.
Owner:HUAWEI TECH CO LTD

Enhanced Harvard Architecture Reduced Instruction Set Computer (RISC) with Debug Mode Access of Instruction Memory within a Unified Memory Space

A Harvard-architecture computer simultaneously reads instructions using an instruction bus and accesses data over a separate data bus. A debug module has an external interface to an external debugger that initiates a debugging session by setting a debug bit in a debug mode register. When the debug bit is set, a first mux to disconnect an instruction pointer and an instruction buffer from the instruction bus and instead connects the data bus to the instruction bus. A second mux disconnects a load / store unit in an execution core from the instruction bus and instead connects the debug module to the instruction bus. The external debugger sees a unified memory space and writes addresses to the debug module that are sent through the second mux to access the data memory, and instruction addresses are sent through the second mux and the first mux to the instruction memory to read or write instructions.
Owner:HIGH TECH TECH LTD

Controlling a quantum processor via quantum programming field payloads

A system comprises pulse instruction memory and pulse generation circuitry, wherein the pulse generation circuitry is operable to retrieve a pulse instruction from the pulse instruction memory, and concurrently generate one or more analog pulses based on a first one or more fields present in the pulse instruction, and one or more digital pulses based on a second one or more fields present in the pulse instruction.
Owner:Q M TECH LTD

Controller area network extra-long (CAN-XL) low latency hardware and software partitioned architecture for message handler

Apparatuses and computer-implemented methods for implementing a message-based protocol interface with a communication bus are provided. An example apparatus for implementing a message-based protocol interface with a communication bus may include message handler core circuitry having a transmit message buffer, wherein the transmit message buffer is configured to store a portion of a transmit message. The apparatus may further include receive handler circuitry configured to store a portion of a received message. The apparatus further includes a message handler processor comprising a processor and an instruction memory including program code, the instruction memory and program code configured to, with the processor, cause the message handler processor to transmit at least the portion of the transmit message from a transmit data memory to the message handler core circuitry and receive the received message from the receive handler circuitry into a receive data memory.
Owner:STMICROELECTRONICS INT NV

System for hardware-based compliance with legal regulations in blockchain smart contracts

A system for autonomous compliance with legal regulations in smart contracts; the system includes: a regulatory data collection unit comprising a network interface controller physically connected to an external communications port, a hardware-based public key infrastructure circuit configured to validate digital certificates of regulatory servers, and a direct memory access controller configured to transfer authenticated regulatory update packets from the network interface controller to a volatile buffer memory without processor intervention; a Compliance Code Conversion Unit with a hardware lexicon scanner implemented as a finite state machine and embedded in reconfigurable FPGA logic blocks, a microcontroller executing a hardware-based lexical analysis pipeline stored in firmware registers, and a translation cache memory configured to temporarily store tokenized compliance rules before writing them to a non-volatile instruction memory; a policy evaluation and enforcement unit comprising a transaction verification processor connected to a secure enclave memory, a hardware comparator circuit configured to compare transaction parameters with compliance thresholds stored in the secure enclave memory, and a logic gate circuit configured to generate execution gate signals that selectively allow, block, or modify smart contract execution signals transmitted to a blockchain execution processor; and an audit logging unit with a cryptographic hashing circuit configured to generate block hashes of enforcement records, a timestamp oscillator configured to generate temporal signatures of compliance enforcement actions, and a Merkle tree generation circuit configured to create tamper-proof hierarchical hash structures, with the enforcement records being passed to a decentralized storage interface for anchoring in the chain or distributed ledger.
Owner:KEMPAIAH MADHURA GAYATHRI BENGALURU +3

Training System and Method for High-Speed Parallel Port IP

The present application relates to the field of integrated circuit technologies and provides a training system and method for a high-speed parallel port IP. The system includes: a high-speed transceiver for transmitting and receiving data and command signals; a delay and reference voltage regulator for performing delay adjustment and reference voltage adjustment; a multi-branch generator for separately generating command combinations, data sequences, and adjustment signals; a microprocessor for controlling the generation of the multi-branch generator based on a first instruction combination; and an instruction memory for storing the first instruction combination. The microprocessor is further configured to generate, based on the first instruction combination, a first scanning strategy including a first scanning object, which is determined based on a first high-speed parallel port IP. The microprocessor controls the generation of the multi-branch generator to utilize the delay and reference voltage regulator to implement the first scanning strategy for the first scanning object in the data and command signals so as to complete eye diagram adaptation. In this way, flexibility and compatibility are achieved.
Owner:XIN YAOHUI TECH CO LTD

Intelligent heterogeneous radar hardware accelerator applied to rail transit

The invention discloses an intelligent heterogeneous radar hardware accelerator applied to rail transit, and belongs to the technical field of radar signal processing and integrated circuits. The accelerator is based on an SoC FPGA platform, a microprocessor soft core based on an RISC-V instruction set is constructed at an FPGA end, and an FFT calculation module and a CFAR calculation module are integrated to serve as hardware acceleration units. And the microprocessor soft core dynamically schedules each hardware acceleration unit by executing a program stored in the instruction memory, constructs a reconfigurable data path, and completes a radar signal processing flow. The method overcomes the defects of power consumption, communication bandwidth and flexibility of traditional DSP and FPGA schemes, and is particularly suitable for efficient and real-time detection and identification of roadbed structure damage in rail transit.
Owner:TIANJIN EMBEDTEC

Softmax instruction set extension method and system based on RISC-V

The invention belongs to the field of neural network hardware acceleration, and provides a Softmax instruction set extension method and system based on RISC-V. The Softmax instruction set extension method based on the RISC-V. The Softmax instruction set extension method based on the RISC-V. The Softmax instruction set extension method based on the RISC-V. The Softmax instruction set extension method based on the RISC-V. The Softmax instruction set extension method comprises an instruction fetching stage, in the decoding stage, corresponding instruction functions are analyzed for Opcode, Funct7 and Funct3 of the Softmax instruction, and corresponding control signals are generated. In the execution stage, data to be subjected to Softmax operation is taken out from the data memory and transmitted into the Softmax calculation unit according to a control signal; executing Softmax calculation according to the following formula, and writing a Softmax calculation result back to the target register; the calculation process is divided into four sub-modules of maximum solution, index calculation, summation and normalization, and a hardware acceleration strategy is designed, so that the operation delay and the resource overhead are greatly reduced, the degree of parallelism, the precision and the energy efficiency ratio of calculation are effectively improved, and the method is suitable for a high-performance neural network reasoning acceleration scene.
Owner:SHANDONG LINGNENG ELECTRONIC TECH CO LTD

Angle-based solution system based on CORDIC instructions

This invention belongs to the field of computer technology and relates to an angle calculation system based on CORDIC instructions. The system includes an IFU (Instruction Fetch Unit) and an EXU (Execution Unit). During the instruction fetch phase, the IFU reads CORDIC instructions from the instruction memory according to the address of the PC. The EXU includes a decoding module, an arithmetic logic unit, a load-memory unit, and a write-back unit. This invention extends the CORDIC angle calculation instruction set based on the RISC-V architecture and designs a complete pipelined hardware structure, supporting trigonometric function operations. It has a complete CORDIC operation path hardware circuit structure, reusing the same operation path for different operation processes, improving hardware resource utilization. It overcomes the limitation of convergence of the calculation result range when performing angle calculation based on the CORDIC algorithm by pre-storing special angle results to reduce calculation time and optimize the circuit operating frequency.
Owner:NORTH CHINA ELECTRIC POWER UNIV

Integrated circuit system-on-a-chip for a neural interface

In one aspect, a system-on-a-chip (SoC) includes a plurality of programmable channels. The SoC includes an analog front end in communication with the plurality of programmable channels. The SoC includes a processing element array configured to receive output signals from the analog front end. The SoC includes an instruction memory supporting an instruction set architecture for supervising computing task of the processing element array, wherein the instruction memory comprises infinite impulse response instructions, discrete Fourier transform instructions, convolutional layer instructions, and fully connected layer instructions, wherein the processing element array is configured to execute instructions in the instruction memory, which, when executed, cause the processing element array to perform functions of a confusion matrix based teach-student convolutional neural network for low power and low latency, and perform functions of a sparsity controller for lower power.
Owner:NORTHWESTERN UNIV

Performance counting device and chip

The invention relates to a performance counting device and a chip, and the performance counting device comprises a performance counting module which is used for receiving at least one instruction memory address from a program counter associated with the performance counting module, and according to the at least one instruction memory address and a pre-configured instruction memory address interval, counting the at least one instruction memory address; obtaining a statistical value of a performance signal of at least one instruction memory address in the instruction memory address interval; the storage module is used for storing the statistical value of the performance signal of the at least one instruction memory address of the performance statistical module; the write-in control module is used for writing the statistical value, obtained by the performance statistical module, of the performance signal of the at least one instruction memory address into the storage module; and the configuration module is used for configuring an instruction memory address interval of the performance statistics module. According to the method and the device, the configuration flexibility of performance statistics can be improved, performance statistics of a plurality of specified program segments can be realized at the same time on the basis, and parallel processing of the performance statistics and accurate positioning of the program segments are further realized.
Owner:SHANGHAI BIREN TECH CO LTD

Printed circuit board (PCB) wiring automatic design method and system

A printed circuit board (PCB) wiring automatic design method in a PCB wiring automatic design system, the system including at least one processor and at least one memory including an instruction, and the PCB wiring automatic design method performed in cooperation with the instruction, the memory, and the processor, includes: receiving PCB data including a net list and information on a plurality of terminals; updating a cost of a wiring exploration region so that a cost of at least some of the wiring exploration region is increased based on a preset constraint condition; exploring a shortest path for wiring the plurality of terminals according to the net list in a cost-updated wiring exploration region; and performing wiring of the plurality of terminals based on wiring probability according to the shortest path.
Owner:LG MANAGEMENT DEV INST CO LTD

Flash-based data access method

The application discloses a FLASH-based data access method, comprising the following steps: obtaining a next data storage starting address, executing a data storage instruction by a FLASH memory, and detecting whether there is a data reading instruction; detecting the data reading instruction, stopping the data storage instruction execution by the FLASH memory and executing the data reading instruction; and if the data reading instruction is not detected, continuing the data storage instruction execution by the FLASH memory. The FLASH-based data access method, the address cyclic storage mode in the address space, compared with erasing the FLASH each time the address is updated, reduces the number of times of FLASH erasing in the address space, increases the service life of the FLASH, organizes the storage and reading data, and maximizes the use of the FLASH storage space to store more data.
Owner:SHAANXI LINGYUN TECH

Histogram operation

A digital data processor (100) comprising: an instruction memory (121) storing instructions, each of the instructions specifying a data processing operation and at least one data operand field; an instruction decoder (113) coupled to the instruction memory for sequentially retrieving instructions from the instruction memory and determining the data processing operation and the at least one data operand; and at least one arithmetic unit (110) coupled to a data register file (123) and to the instruction decoder for performing a data processing operation on at least one operand corresponding to an instruction decoded by the instruction decoder and storing the result of the data processing operation. The arithmetic unit is configured to increment a histogram value in response to a histogram instruction by incrementing a bin entry at a specified location in at least one histogram of a specified number.
Owner:TEXAS INSTRUMENTS INC

Histogram operation

The invention relates to histogram operation. A digital data processor (100) includes: an instruction memory (121) storing instructions each specifying a data processing operation and at least one data operation digit segment; an instruction decoder (113) coupled to the instruction memory for sequentially invoking instructions from the instruction memory and determining the data processing operation and the at least one data operand; and at least one arithmetic unit (110) coupled to the data register file (123) and to an instruction decoder to perform a data processing operation on at least one operand corresponding to an instruction decoded by the instruction decoder and to store a result of the data processing operation. The arithmetic unit is configured to incrementing a histogram value in response to a histogram instruction by incrementing a bin entry at a specified location in at least one histogram of a specified number.
Owner:TEXAS INSTRUMENTS INC

Multi-agent instruction execution engine for neural inference processing

Multi-agent instruction execution engines for neural inference processing are provided. In various embodiments, a neural core is provided. The neural core includes an instruction memory. The instruction memory comprises a plurality of instruction streams, each instruction stream associated with one of a plurality of agents. The instruction memory further comprises a plurality of shared functional units. The neural core is adapted to concurrently execute the plurality of instruction streams on the plurality of associated agents. The execution includes maintaining a separate program counter for each of the plurality of agents, determining a plurality of operations from the instructions of each instruction stream, and directing the operations to the shared functional units. The instructions of each instruction stream are statically scheduled prior to runtime to ensure their execution is conflict free.
Owner:INTERNATIONAL BUSINESS MACHINE CORPORATION

Information processing device

To provide an information processing device that can reduce development man-hours and lower vulnerabilities. [Solution] The memory 10 of the information processing device 1 stores a secure function module 111a that contains instructions to be executed by the processor 20 in the performance of a function and is restricted from external interference. The memory 10 also stores a processing function module 131 that contains instructions common to the secure function module 111a, has less restrictive external interference than the secure function module 111a, and is processed into a disabled state in which the processor 20 cannot execute instructions. The memory 10 also stores a release module 112 that has more restricted external interference than the processing function module 131 and can release the disabled state in the processing function module 131. The memory 10 also stores a security request determination flag 115 that defines the security level required in the performance of a function.
Owner:DENSO CORP

Memory control circuit of micro computer

To provide a memory control circuit of a micro computer that can enhance a processing speed of the micro computer in comparison with conventional arts.SOLUTION: A micro computer memory control circuit comprises: a memory read control circuit that stores an address where a processor reads a command from a memory, and a next address in a memory address register, stores a command of the address and a command of a next address in a command memory, and stores the next address and a command of the next address in a first command cache; and a cache controller that, when the processor reads a next command from the memory, if the address stored in the memory address register and the address stored in the first command cache are mismatched, stores the address stored in the memory address register, the command of the address and a frequency of the mismatch in a second command cache.SELECTED DRAWING: Figure 1
Owner:ROHM CO LTD

Quantum-computing control system and method with non-sequential mid-circuit control flow

A quantum-computing control system includes a script processor, an instruction memory, an instruction fetcher, and a waveform player. The instruction memory stores a plurality of local instructions forming a local instruction set. The script processor executes a local script and transmits to the instruction fetcher a sequence of fetch commands that is based on the local script. The instruction fetcher retrieves, based on each fetch command of the sequence of fetch commands, a fetched instruction from the instruction memory. The fetched instruction belongs to the local instruction set. The instruction fetcher also transmits, to the waveform player, the fetched instruction as one of a sequence of fetched instructions. The waveform player executes the sequence of fetched instructions to generate a digital waveform. The control system may be used to implement mid-circuit (i.e., during runtime) non-sequential control flow of a quantum computer or other type of quantum-engineered system.
Owner:ATOM COMPUTING INC

An FPGA-based target detection neural network accelerator and a target detection system

The application discloses an FPGA-based target detection neural network accelerator and a target detection system, the accelerator is arranged on an FPGA and comprises a control module, an input buffer module, a parameter buffer module, a systolic array module, a post-processing module and a pooling module; the control module comprises an instruction memory and a controller; the instruction memory is used for storing an instruction set of a neural network; the controller is used for generating control instructions according to the instruction set and controlling the operation of each module; the input buffer module is used for storing input feature data blocks of the target detection neural network; the parameter buffer module is used for storing bias data and weight data of the target detection neural network; the systolic array module is used for performing convolution operation; the post-processing module is used for processing the convolution calculation result; the pooling module is used for performing maximum pooling operation on data; and the output buffer module is used for reading and writing cache of data. The application improves the inference efficiency of the accelerator by configuring parameters of the target detection neural network, and realizes the high-performance and high-flexibility target detection neural network accelerator.
Owner:GUANGDONG UNIV OF TECH

Debugging of accelerator circuit for mathematical operations using packet limit breakpoint

Embodiments of the present disclosure relate to debugging of an accelerator circuit using a packet limit breakpoint. A vector circuit reads a subset of instruction packets from an instruction memory and receives a portion of input data from a data memory corresponding to the subset of instruction packets. The vector circuit executes a set of vector operations in accordance with multiple instruction packets from the subset using data from the received portion of input data identified in the multiple instruction packets to generate output data. A program counter control circuit coupled to the instruction memory triggers a breakpoint in a program stored in the instruction memory causing the accelerator circuit to stop executing remaining instruction packets in the program following the multiple instruction packets responsive to a number of instruction packets executed in the program from a time instant of an event reaching a predetermined number.
Owner:APPLE INC

Circuits and methods for linear memory access control table switching for fine-grained partitioning

Circuits and methods for implementing one or more handover subprocess instructions are described. In some examples, a hardware processor (e.g., a core) includes (e.g., a coupling to) a memory management circuit to control memory access based on a memory tag stored in a memory tag data structure and a memory tag based on a pointer to the memory; decoder circuitry to decode an instruction into a decoded instruction, the instruction comprising an operand to identify a memory tag data structure for a sub-process of a plurality of memory tag data structures for a corresponding sub-process of a process and an opcode, the operand to identify a memory tag data structure for the sub-process of the process. The opcode is used for indicating the execution circuitry to switch from another memory tag data structure for another sub-process of the process to a memory tag data structure for the sub-process; and execution circuitry to execute the decoded instruction according to the opcode. The memory tag data structure may be resumed for providing access control permissions for sub-processes per memory particle.
Owner:INTEL CORP

A system that uses multiple lightweight processors to implement asymmetric algorithm multi-core parallel architecture using a single instruction memory

The present invention discloses a system for implementing an asymmetric algorithm multi-core parallel architecture using a single instruction memory for multiple lightweight processors. The system comprises an AXI bus, an interface terminal, an ultra-high-speed interface, a high-speed interface, a master processor, an instruction register, an asymmetric key memory, and multiple lightweight processors. Each lightweight processor is connected to an asymmetric algorithm core and an asymmetric interface memory, and each lightweight processor is connected to the instruction register and the asymmetric key memory. The asymmetric key memory, asymmetric interface memory, interface terminal, ultra-high-speed interface, high-speed interface, and master processor are all connected to the AXI bus. Advantages of the system include: using one instruction memory to provide instruction reading for multiple lightweight processors and one asymmetric key memory to provide asymmetric keys for multiple lightweight processors to implement an asymmetric algorithm chip with a reduced area, thereby reducing costs and preventing excessive heat generation during chip operation.
Owner:GUANGZHOU WANXIETONG INFORMATION TECH CO LTD

Variable register block structure for supporting different scene applications of MCU (Microprogrammed Control Unit) chip

The invention discloses a variable register block structure supporting different scene applications of an MCU chip, and belongs to the technical field of chip design, the variable register block structure is applied to the MCU chip, the MCU chip comprises a core, an instruction memory, a data memory and a variable register block, the core is used as a main device, and the instruction memory is used as an auxiliary device. The register block module, the instruction memory, the data memory and the variable register block are used as slave devices; wherein the core is respectively connected with the instruction memory, the data memory and the variable register group through a bus; and the variable register group is used as slave equipment and accesses registers in the variable register group through the core. The variable register block disclosed by the invention can be used as different functional modules in different working scenes; by using the variable register block structure, more flexible and changeable applications can be provided, and meanwhile, the relatively fixed area and power consumption can be reduced; and a certain protection effect is also achieved for the exposure risk of starting the ROM.
Owner:CANXIN SEMICON (SUZHOU) CO LTD

MEMS hot switch testing system

Embodiments disclose a system for testing a micro-electromechanical-system (MEMS) switch under hot switching conditions. The system includes a processor, and an instruction memory with computer code instructions stored thereon, configured to cause the system cyclically open and close while a voltage is applied to at least one contact of the MEMS switch. The system measures and stores in a characteristic memory one or more characteristic values associated with the MEMS switch during the cyclical opening and closing. The system records an operational status of the MEMS switch during one or more cycles of the MEMS switch being opened and closed. The operational status of the MEMS switch is either an operational state or a failure state. The system then calculates a life expectancy of the MEMS switch utilizing the measured characteristic values associated with the MEMS switch and the operational status of the MEMS switch.
Owner:MENLO MICROSYSTEMS INC

Memory device and method including a loop instruction memory queue

A memory device includes a memory bank including one or more bank arrays, a PIM circuit configured to perform an operation logic processing operation, and an instruction memory including a first instruction queue segment to an mth instruction queue segment configured in a circular instruction queue to store instructions provided by a host, wherein the instructions stored in the first instruction queue segment to the mth instruction queue segment are executed in response to an operation request from the host, and each new instruction provided by the host is updated on the fully executed instructions in the circular instruction queue.
Owner:SAMSUNG ELECTRONICS CO LTD

Apparatus, a method and a non-transitory machine-readable storage medium

It is provided an apparatus comprising interface circuitry, machine-readable instructions, and processing circuitry to execute the machine-readable instructions. The machine-readable instructions include instructions to record an entry into a virtual integrity register. The entry is based on a measurement of a module. The module is a part of a confidential computing environment. The machine-readable instructions further include instructions to load the module into a memory. The memory is accessible by the confidential computing environment. The machine-readable instructions further include instructions to record an entry of a log. The log comprises a load history. The entry comprises information about the loading of the module. The machine-readable instructions further include instructions to record an entry into an integrity measurement register. The entry being based on the recorded virtual integrity register entry.
Owner:XING BIN +2

Controller area network extra-long (can-XL) low latency hardware and software partitioned architecture for message handler

Apparatuses and computer-implemented methods for implementing a message-based protocol interface with a communication bus are provided. An example apparatus for implementing a message-based protocol interface with a communication bus may include message handler core circuitry having a transmit message buffer, wherein the transmit message buffer is configured to store a portion of a transmit message. The apparatus may further include receive handler circuitry configured to store a portion of a received message. The apparatus further includes a message handler processor comprising a processor and an instruction memory including program code, the instruction memory and program code configured to, with the processor, cause the message handler processor to transmit at least the portion of the transmit message from a transmit data memory to the message handler core circuitry and receive the received message from the receive handler circuitry into a receive data memory.
Owner:STMICROELECTRONICS (GRAND OUEST) SAS +1

Cooperative instruction prefetch on multicore system

Aspects of the disclosure are directed to methods, systems, and apparatuses using an instruction prefetch pipeline architecture that provides good performance without the complexity of a full cache coherent solution deployed in conventional CPUs. The architecture can include components which can be used to construct an instruction prefetch pipeline, including instruction memory (TiMem), instruction buffer (iBuf), a prefetch unit, and an instruction router.
Owner:GOOGLE LLC

Clock aware simulation vector processor

A processing system for validating a circuit design, the processing system includes a flow processor, and an evaluation system coupled with the flow processor. The flow processor generates instructions from the circuit design. The evaluation system includes instruction memory circuitry receives the instructions from the flow processor and generate control signals, and interconnect circuitry receives the control signals routes a plurality of values based on the control signals. Each of the plurality of values having one of four states. The evaluation further includes operation circuitry that receives the plurality of values and the control signals, performs one or more operations of the circuit design with the plurality of values based on the control signals, and outputs operation values based on performing the one or more operations, the operation values indicative of an error within the circuit
Owner:SYNOPSYS INC