Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

42 results about "Scratchpad memory" patented technology

Scratchpad memory (SPM), also known as scratchpad, scratchpad RAM or local store in computer terminology, is a high-speed internal memory used for temporary storage of calculations, data, and other work in progress. In reference to a microprocessor ("CPU"), scratchpad refers to a special high-speed memory circuit used to hold small items of data for rapid retrieval. It is similar to the usage and size of a scratchpad in life: a pad of paper for preliminary notes or sketches or writings, etc.

Dataflow architecture processor statically reconfigurable to perform N-dimensional affine transformation in parallel manner by replicating copies of input image across multiple scratchpad memories

A statically reconfigurable dataflow architecture processor (SRDAP) performs an N-dimensional affine transform specified by a matrix on an input image to produce an output image includes L address pattern memory units (PMUs) comprising a memory arranged as a vector of L banks, and L corresponding data PMUs. Each data PMU receives a copy of the input image. In parallel: each address PMU writes an L-vector of addresses of input pixels to the vector of L banks and reads a single address of the written L-vector of addresses from a predetermined bank corresponding to a PMU number of the address PMU among the L address PMUs, and each data PMU receives the single address from the corresponding address PMU and uses it to read a single input pixel from the data PMU memory. A tree of pattern compute units coalesces the L single input pixels into an L-vector of input pixels.
Owner:SAMBANOVA SYSTEMS INC

Floating point arithmetic device and method of operating the same

A floating point arithmetic device with two floating point operands and its operation method are disclosed. The floating point arithmetic device includes an exponent subtraction circuit, an exponent calculation circuit, a mantissa calculation circuit, and a conversion circuit. The exponent subtraction circuit calculates the difference between the exponents of the two operands and generates a sign bit and an exponent difference. The exponent calculation circuit generates the post-operation exponent bits according to the larger one of the exponents of the two operands. The mantissa calculation circuit aligns the mantissa bits of the two operands and performs one of addition and subtraction on the aligned mantissa bits. To improve the calculation efficiency and reduce the power consumption, the floating point arithmetic device can complete the floating point addition or subtraction operation in one step (one clock cycle) without moving the intermediate floating point data between the registers and the functional circuit units as in the multi-step operation.
Owner:XINLIJIA INTEGRATED CIRCUIT (SHANGHAI) CO LTD

Prescient computing

Techniques for static instruction decoupling for data movement and computer are described. In some examples, hardware support at least includes a plurality of instruction queues to store instructions, wherein each instruction queue of the plurality of instruction queues is dedicated to a separate thread; a local memory to store instructions and / or data for a first thread; a scratchpad memory, coupled to the local memory, to store instructions and / or data for a second thread; and execution resources, coupled to the scratchpad memory, to execute one or more mathematic and / or logical instructions for a third thread.
Owner:INTEL CORP

Chip and its design method and failure analysis method

A chip, a design method thereof, and a fault analysis method thereof, wherein the chip, also known as an integrated circuit, includes a configuration register that is usually volatile, and / or at least one on-chip non-volatile memory m, which usually includes at least one reserved memory location that can be reserved to store the content of at least one of the configuration registers that is usually volatile; and / or a write unit that is configured to store a value indicating the content of at least one of the configuration registers that is usually volatile at least once in, for example, the at least one reserved memory location in the on-chip non-volatile memory.
Owner:NUVOTON

On-package accelerator complex (AC) for integrating accelerator and IOS for scalable ran and edge cloud solution

Methods and apparatus for on-package accelerator complex (AC) for integrating accelerator and IOs for scalable RAN and edge cloud solutions. The AC comprises one or more dies including an IO interface tile that is coupled to multiple intellectual property (IP) blocks that may be integrated on the same die as the IO interface tile or separate dies that are coupled to the IO interface tile via die-to-die or chiplet-to-chiplet interconnects. The IP blocks may include a network interface (e.g., Ethernet) and one or more accelerators. The package further includes a central processing unit (CPU) that is coupled to the AC via a die-to-die or chiplet-to-chiplet interconnect. The IO interface tile includes integrated shared scratchpad memory that is shared among the IP blocks and the CPU cores. The IO interface tile further includes an interface controller for scheduling IP blocks and configuring data transfers between the IP blocks, such as used by a RAN pipeline.
Owner:INTEL CORP

Test circuit and electronic device

ActiveCN113345508BStatic storageMemory circuitsScratchpad memory
The present application provides a test circuit and an electronic device for testing a memory circuit, which includes a controller, a pattern generating circuit, a comparison circuit and a register. The controller is used to generate a plurality of internal test signals and receive a test result. The pattern generating circuit writes a test data into a memory block of the memory circuit according to the internal test signals and reads the memory block to generate a read data. The comparison circuit compares the test data and the read data to generate the test result. The register is used to store the test result. The controller judges whether the memory circuit is normal according to the test result stored in the register.
Owner:VANGUARD INTERNATIONAL SEMICONDUCTOR CORPORATION

Non-integer frequency divider and flash memory controller

The present invention relates to a non-integer frequency divider and a flash memory controller. The non-integer frequency divider includes a plurality of registers, a counter, a control signal generator, and a clock gating circuit. With respect to the plurality of registers, at least a portion of the plurality of registers is set to have a value. The counter is configured to sequentially generate a plurality of count values, wherein the plurality of count values ​​respectively correspond to the at least a portion of the registers, and the plurality of count values ​​are repeatedly generated. The control signal generator is configured to generate a control signal based on the received count value and the corresponding register value. The clock gating circuit is configured to mask or unmask an input clock signal based on the control signal to generate an output clock signal.
Owner:SILICON MOTION INC

Non-integer frequency divider and flash memory controller

The invention relates to a non-integer frequency divider and a flash memory controller. The non-integer frequency eliminator comprises a plurality of temporary memories, a counter, a control signal generator and a clock pulse gating circuit. At least a portion of the plurality of registers is set to have a value with respect to the plurality of registers. The counter is used for sequentially generating a plurality of count values, the plurality of count values respectively correspond to the at least one part of the temporary register, and the plurality of count values are repeatedly generated. The control signal generator is used for generating a control signal according to the received count value and the value of the corresponding register. The clock gating circuit is used for referring to the control signal to shield or not shield an input clock signal so as to generate an output clock signal.
Owner:SILICON MOTION INC

Floating point arithmetic device and operating method thereof

The invention discloses a floating-point arithmetic device with two floating-point number operands and an operation method of the floating-point arithmetic device. The floating-point arithmetic device comprises an index subtraction circuit, an index calculation circuit, a mantissa calculation circuit and a conversion circuit. The exponent subtraction circuit is used for calculating the difference between the exponents of the two operands to generate a sign bit and an exponent difference. The exponent calculation circuit generates a post-operation exponent bit according to a larger exponent in the exponents of the two operands. The mantissa calculation circuit aligns mantissa bits of the two operands, and performs one of addition and subtraction on the aligned mantissa bits. In order to improve the calculation efficiency and reduce the power consumption, the floating-point arithmetic device can complete floating-point addition or subtraction operation in one step (one clock cycle) without moving intermediate floating-point number data between a temporary memory and each functional circuit unit like multi-step operation.
Owner:XINLIJIA INTEGRATED CIRCUIT (SHANGHAI) CO LTD

3D in-pipeline private scratchpad memory for SIMT compute cores

A processor core (1) of the SIMT type and a related method of operation are disclosed. The processor core comprises frontend circuitry (10), a plurality of lanes (20a-f), a multi-banked and arbitration-free scratchpad memory (40), and an interconnect (50) to couple active ones of the plurality of lanes to different ones of the scratchpad memory banks. There are at least as many banks as lanes in the processor core. Each lane comprises a thread-private register file (21a-f) and address generation logic (22a-f) to independently calculate an effective address of an operand within the scratchpad memory. An execution pipeline of the processor core, associated with the thread-parallel execution of vector memory instructions, comprises at least the address generation logic of the different lanes, the interconnect, and the different banks as pipeline components. The processor core is implemented as a stack of dies (2a, 2b) including the frontend circuitry and the plurality of lanes on a first die (2a) and the scratchpad memory on a second die (2b).
Owner:INTERUNIVERSITAIR MICRO ELECTRONICS CENT (IMEC VZW)

Shared scratchpad memory with parallel load-store

Methods, systems, and apparatus, including computer-readable media, are described for a hardware circuit configured to implement a neural network. The circuit includes a first memory, respective first and second processor cores, and a shared memory. The first memory provides data for performing computations to generate an output for a neural network layer. Each of the first and second cores include a vector memory for storing vector values derived from the data provided by the first memory. The shared memory is disposed generally intermediate the first memory and at least one core and includes: i) a direct memory access (DMA) data path configured to route data between the shared memory and the respective vector memories of the first and second cores and ii) a load-store data path configured to route data between the shared memory and respective vector registers of the first and second cores.
Owner:GOOGLE LLC

Field programmable memory array, writing method and reading method

The invention provides a field programmable memory array, a writing method and a reading method. The field programmable memory array comprises a single readable and writable memory array, an addressing circuit and an access circuit. The readable and writable memory array includes m memory cells in 2i rows and 2j columns to perform a digital circuit function, and has a plurality of digital inputs and a plurality of digital outputs. The addressing circuit is configured to receive an n-bit input vector and is coupled between the readable and writable memory array and the access circuit. The n-bit input vectors are selected from a register conversion layer logic table, and the logic table comprises a plurality of n-bit input vectors and a plurality of m-bit output vectors. The FPMA utilizes the character line / bit line multiplexer characteristics of the memory array to realize the operation of a plurality of binary inputs and any number of binary outputs. The FPMA has the advantages that the connection complexity is reduced, the number of repeated flip-flops is reduced, and the grain area in an integrated circuit chip is saved.
Owner:XINLIJIA INTEGRATED CIRCUIT (SHANGHAI) CO LTD

Memory management method, apparatus, device, and medium

The application discloses a memory management method, device, equipment and medium. The method comprises the following steps: obtaining pre-defined temporary register initialization parameters during initialization; applying memory space to an operating system once according to the temporary register initialization parameters, and obtaining initialized temporary registers; obtaining a first value corresponding to memory application of a target application program during running of the target application program; traversing all the initialized temporary registers, and calculating a difference value between a memory space value corresponding to at least one memory block in each temporary register and the first value; searching for an optimal temporary register in all the temporary registers according to the difference value; when the optimal temporary register is found, allocating at least one memory block in an unused state from the optimal temporary register to the target application program, and marking a state of each allocated memory block as a used state. The method can uniformly allocate and manage memory blocks by creating temporary registers in advance, reduce the number of times of directly calling system memory allocation, and improve memory use efficiency.
Owner:JOINT WARFARE COLLEGE NAT DEFENSE UNIV OF THE CHINESE PEOPLES LIBERATION ARMY

Temporary main memory module special for luminous server

The utility model relates to a temporary memory main memory module special for a light-emitting server. The temporary memory main memory module comprises a printed circuit board, a plurality of dynamic random access memories, a temporary clock pulse driver and at least one light-emitting diode, the dynamic random access memories are arranged on the printed circuit board. The temporary clock pulse driver is arranged on the printed circuit board. The at least one light emitting diode is arranged on the printed circuit board. Therefore, the main memory module of the temporary memory special for the light-emitting server can emit light through the light-emitting diode in the state of stable current, highlights the aesthetic feeling, and meets the requirements of artificial intelligence creators, game developers and display type workstations or servers.
Owner:V COLOR TECH INC

Inter-integrated circuit digital audio standard audio device and method for skew recovery thereof

PendingCN122511316AShift registerHemt circuits
An inter-integrated digital audio standard (I2S) audio device and its misalignment recovery method are provided, including a first-in-first-out (FIFO) circuit and a shift register. The FIFO circuit has a memory, a FIFO synchronization logic circuit, and a FIFO misalignment flag. The FIFO synchronization logic circuit, in response to changes in the logic bits of the left and right clock signals, shifts data out of the memory or into the memory according to the data's position. In response to a misalignment in the FIFO circuit, the logic bit of the corresponding FIFO misalignment flag is set to the first logic bit. The shift register outputs data from the memory or shifts data into the memory according to changes in the logic bits of the bit clock signals. In response to the FIFO misalignment flag being set to the first logic bit, the I2S audio device is reset to recover the misalignment.
Owner:NUVOTON

Method, System, Electronic Device and Medium for Improving CPU Execution Efficiency

The present invention relates to a method, system, electronic device, and medium for improving CPU execution efficiency, comprising: obtaining an access request, the access request including a first access address of target data, the first access address representing the storage location of the target data in a temporary register; obtaining the target data from the temporary register according to the first access address, and if the target data is not obtained from the temporary register, determining a second access address corresponding to the target data according to the first access address, the second access address representing the storage location of the target data in a main memory; obtaining the target data from the main memory according to the second access address; and storing the target data in the temporary register according to the target storage address, the target storage address representing the storage address of the target data in the temporary register. The method solves the problem of low CPU execution efficiency when the current CPU fails to obtain data in the temporary register.
Owner:WUHAN MENGXIN TECH CO LTD

Graphics processor instruction processing method and device based on risc-v instruction set

ActiveCN116342368BGraphicsInstruction set design
The application discloses a kind of based on RISC-V instruction set's graphic processor instruction processing method and device, comprising: the current instruction is decoded, and initial decoding result is obtained;According to the initial decoding result, it is judged whether the current instruction is register expansion instruction;If the current instruction is not the register expansion instruction, then according to the effective bit information of target temporary register, the instruction processing category is determined;Wherein, the instruction processing category includes RISC-V standard decoding category or splicing decoding category;For the splicing decoding category, target expansion data is read from the target temporary register;Wherein, the target expansion data is the expansion data corresponding to last register expansion instruction;The target expansion data and the initial decoding result are spliced and handled, and target decoding result is generated.The application can improve the instruction processing performance of graphic processor designed by using RISC-V instruction set.
Owner:INTERNATIONAL INNOVATION CENTER OF TSINGHUA UNIVERSITY SHANGHAI +1

Asynchronous first-in first-out memory verification method, device, equipment and medium

The invention discloses an asynchronous first-in first-out memory verification method, device and equipment and a medium, and relates to the technical field of integrated circuit design and verification, and the method comprises the steps: analyzing a temporary memory transfer language code, detecting whether a statement is complete or not and whether bit width and depth parameters of a memory unit are consistent with design specifications or not, and obtaining a detection result; verifying the clock domain affiliation of the read-write pointer generation logic to obtain a first verification result; verifying the synchronization chain level, the Gray code coding and the clock domain isolation to obtain a second verification result; verifying the data integrity and the error processing mechanism through the test case to obtain a third verification result; determining a coverage rate heat map according to the coverage rate data, and verifying the completeness of the test case to obtain a fourth verification result; and generating a verification report according to the detection result, the first verification result, the second verification result, the third verification result, the fourth verification result and the coverage rate heat map to complete verification. And automatic verification of the asynchronous first-in first-out memory is realized.
Owner:JINAN MAIWEI INTELLIGENT TECHNOLOGY CO LTD

General padding support for convolution on systolic arrays

Methods and systems, including computer programs encoded on a computer storage medium. In one aspect, a method includes the actions of receiving a request to perform convolutional computations for a neural network on a hardware circuit having a matrix computation unit, the request specifying the convolutional computation to be performed on a feature tensor and a filter and padding applied to the feature tensor prior to performing the convolutional computation; and generating instructions that when executed by the hardware circuit cause the hardware circuit to perform operations comprising: transferring feature tensor data from a main memory of the hardware circuit to a scratchpad memory of the hardware circuit; and repeatedly performing the following operations: identifying a current subset of the feature tensor; and determining whether a memory view into the scratchpad memory for the current subset is consistent with a memory view of the current subset in the main memory.
Owner:GOOGLE LLC

Method and system for binary search

The application provides a binary search method and system. The method comprises the following steps: providing a storage device with M entries, each of which stores a value; providing index register registers including N registers, wherein the N registers divide the storage device into N-1, N or N+1 search areas; wherein M and N are integers and N
Owner:OPTICORE TECH INC

Static timing analysis method and static timing analysis system

This invention discloses a static timing analysis method and a static timing analysis system. The static timing analysis method includes: obtaining a standard component library file describing multiple standard components; performing circuit structure analysis on the standard component library file to identify a target sequential component from the standard components, the target sequential component including logic gates, selection circuits, and register circuits; executing a logic test program to identify pin combinations with non-mutually controllable relationships, and treating the timing constraints related to these pin combinations in the standard component library file as redundant timing constraints to remove them from the standard component library file, thereby generating an optimized standard component library file; and performing static timing analysis on the target circuit design based on the optimized standard component library file.
Owner:REALTEK SEMICON CORP

Temporary main memory module special for luminous server

The invention relates to a temporary memory main memory module special for a light-emitting server. The temporary memory main memory module comprises a printed circuit board, a plurality of dynamic random access memories, a temporary clock pulse driver and at least one light-emitting diode, the dynamic random access memories are arranged on the printed circuit board. The temporary clock pulse driver is arranged on the printed circuit board. The at least one light emitting diode is arranged on the printed circuit board. Therefore, the temporary memory main memory module special for the light-emitting server can emit light through the light-emitting diode under the state of stable current, the aesthetic feeling is highlighted, and the requirements of an artificial intelligence creator, a game developer and a display type work station or a server shell are met.
Owner:V COLOR TECH INC

Reliable programming of one-time programmable (OTP) memory during device testing

A method includes providing one or more signals to an electronic device for performing a test procedure that involves programming a One-Time Programmable (OTP) memory in the electronic device. A verification is made as to whether connection of the one or more signals to the electronic device is stable, by performing a sequence of one or more iterations, each iteration including (i) determining, from among a set of scratchpad addresses in the OTP memory, an address that is available for programming, (ii) writing a test value to the address, and then (iii) reading the test value from the address. If the read test value differs from the written test value, re-tuning of the connection of the one or more signals is initiated. Only when the connection is verified as stable by the sequence of iterations, the OTP memory is programmed in accordance with the test procedure.
Owner:NUVOTON

Data processing method and device combined with access memory cell array

This invention discloses an apparatus including a storage cell array, multiple page registers, and a processing element. The storage cell array is divided into multiple blocks. The multiple page registers are coupled to the edge blocks of the storage cell array, and the multiple page registers access page data based on a page copying method. The processing element is connected to the multiple page registers to process conditionally and / or directly read data from the multiple page registers to perform multiplication and / or accumulation in machine learning of an artificial intelligence system.
Owner:PIECEMAKERS TECH

Interference signal correction and calculation method based on high-scribed-line density grating

The invention relates to an interference signal correction and calculation method based on a high-scribed-line-density grating. The method comprises the steps that 1, four paths of grating interference signal data are collected through a photoelectric detector; 2, subtracting a self mean value from four paths of grating interference original signals received by each photoelectric detector, and performing direct current elimination processing; 3, carrying out filtering processing on the grating interference signal after the direct current elimination processing by adopting a mean value smooth filtering method; 4, setting a temporary memory for the filtered four-path grating interference signals for stack type storage, and setting a data acquisition and preprocessing module and an interference signal correction module to be in dual-thread synchronous operation; 5, correcting the displacement nonlinearity of the grating interference signal read from the temporary memory by using a correction algorithm based on RANSAC-Heydemann; and 6, calculating the displacement. According to the method, a double-closed-loop thread processing mode is designed and is innovatively combined with an RANSAC-Heydemann correction algorithm, so that the measurement stability and accuracy are greatly improved.
Owner:SHANGHAI METROLOGY & TESTING TECHNOLOGY RESEARCH INSTITUTE CO LTD

Large language model hardware description language conversion method oriented to logic gate netlist generation

The invention discloses a logic gate netlist generation oriented large language model hardware description language conversion method, electronic equipment and a computer readable storage medium, and relates to the technical field of digital circuits and generative AI and large language models.The method comprises the steps that an LLM is trained to analyze and rewrite an input original HDL; converting the state machine description into a standardized state machine description with a uniform format; analyzing the description of the standardized state machine through logic Boolean, and generating a Boolean expression of the sum of products corresponding to each bit of a state register, an internal register and an output signal in the circuit as an intermediate expression; and generating the logic gate netlist based on an SOP Boolean expression and a predefined logic gate unit library. The problem that a traditional EDA tool is prone to falling into local optimum due to the fact that the traditional EDA tool depends on a curing heuristic algorithm is solved.
Owner:SHANGHAI JULU CHUANGXIN TECH CO LTD

Vector reductions using shared scratchpad memory

Methods, systems, and apparatus, including computer-readable media, are described for performing vector reductions using a shared scratchpad memory of a hardware circuit having processor cores that communicate with the shared memory. For each of the processor cores, a respective vector of values is generated based on computations performed at the processor core. The shared memory receives the respective vectors of values from respective resources of the processor cores using a direct memory access (DMA) data path of the shared memory. The shared memory performs an accumulation operation on the respective vectors of values using an operator unit coupled to the shared memory. The operator unit is configured to accumulate values based on arithmetic operations encoded at the operator unit. A result vector is generated based on performing the accumulation operation using the respective vectors of values.
Owner:GOOGLE LLC

Memory reuse method, device and equipment of temporary memory and storage medium

The invention discloses a memory reuse method, device and equipment for a temporary memory and a storage medium, and the method comprises the steps: obtaining an artificial intelligence operator program containing the storage demand of the temporary memory, carrying out the activity analysis of a plurality of input data blocks associated with the temporary memory in the operator program, and obtaining the life cycle of each data block; judging whether an input data block with a dynamic shape exists in the operator program or not; if yes, inserting a main control code containing memory multiplexing logic of each data block in an operator program compiling stage, and completing memory multiplexing of the temporary memory according to the main control code; and if not, according to the life cycle of each data block, generating a temporary memory allocation and reuse strategy in an operator program compiling stage, and completing the memory reuse of the temporary memory. According to the technical scheme provided by the embodiment of the invention, the time cost and the labor cost consumed by manually allocating the memory resources of the temporary storage can be avoided, the allocation difficulty of the memory resources of the temporary storage is reduced, and the memory utilization rate of the temporary storage is improved.
Owner:SHANGHAI ENFLAME INTELLIGENCE TECH CO LTD