Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

50 results about "Single cycle" patented technology

A single cycle processor is a processor that carries out one instruction in a single Clock cycle. MIPS architecture, MIPS-32 architecture. DLX, a very similar architecture designed by John L. Hennessy (creator of MIPS) for teaching purposes.

Single-period delay compensation method and device for magnetic suspension bearing switch power amplifier

The invention belongs to the technical field of active magnetic suspension bearings, and discloses a single-cycle delay compensation method and device for a magnetic suspension bearing switch power amplifier, and the method comprises the steps: building a unipolar single-cycle control mathematical model, and calculating the duty ratio of a single cycle; setting a state switching criterion of a charging and discharging period, and judging a charging and discharging state of a single period; and establishing a linear duty ratio prediction model, calculating a linear derivation duty ratio dt and a calculation duty ratio dj to obtain a total duty ratio of a corresponding switch tube in a period, and outputting the total duty ratio to the corresponding switch tube. According to the method, target tracking can be achieved under the static condition, steady-state errors do not exist, and the problem that time delay exists in a traditional digital single-cycle control algorithm can be effectively solved under the dynamic condition.
Owner:SHAANXI UNIV OF SCI & TECH

Instruction processing method and device, terminal equipment and program product

The invention is suitable for the technical field of computer processor design, and provides an instruction processing method and device, terminal equipment and a program product, and the method comprises the steps: obtaining instruction information of a to-be-processed instruction set in a processor; generating an index vector corresponding to the to-be-processed instruction set based on the original effective vector; each index value in the index vector is subjected to one-hot code conversion, a one-hot code matrix corresponding to the instruction set to be processed is generated, and each row of one-hot code vector in the one-hot code matrix corresponds to each path of input instruction; and according to the original effective vector and the one-hot code matrix, sorting operation codes of each path of input instruction to obtain processed operation code information, so that a processor executes corresponding operation based on the processed operation code information. The method can meet the requirement that the high-performance processor completes instruction processing in a single cycle, the processing delay is low, and therefore the instruction execution efficiency of the processor is improved.
Owner:GUANGDONG LEAPFIVE TECH CO LTD

Serialized data link with zero-cycle or short-cycle paths

Systems and methods are disclosed for reduced-power serial data links. A clock-forwarded serial link carries a clock lane and one or more data lanes. Every active serial data cycle is accompanied by its own serial clock edge: a clock delay allows the same clock edge to drive data at a transmitter and latch data at a receiver. Power is saved by idling the serial clock when data is not being transmitted. A valid signal can be omitted, providing a space saving. At the destination, similar clock-forwarding and delay enables a single parallel clock edge to drive data to the boundary of its clock domain, e.g. from a deserializer to a FIFO. The data link exhibits zero-cycle entry and exit. Variations with half- or single-cycle entry or exit are disclosed.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Single cycle request arbiter

A memory array includes a plurality of memory devices, each of which includes a memory configured to store packet data and a request arbiter configured to interface with other memory devices of the memory array. The request arbiter filters invalid requests from a plurality of requestors, and determines a bitvector representing a sequence of the plurality of requestors, the bitvector indicating whether each of the plurality of requestors has a valid request. The request arbiter outputs an indication of a first request to be serviced by the memory device, and shifts the bitvector to determine a second request to be serviced by the memory device.
Owner:MARVELL ASIA PTE LTD

Multi-period modulation driving assembly, method, system and product

The invention provides a multi-cycle modulation driving assembly, method, system and product. A control plane and hardware are decoupled by utilizing an SDN technology, a remote controller can realize storage unit data configuration and circuit logic device programming of different devices according to codes of controlled devices, and various periodic / non-periodic time sequence driving signals required by various device chips are provided; and the time sequence requirements of array driving and data reading of different equipment chips are met. The multi-cycle modulation driving assembly comprises a state division assembly, a synchronous multi-cycle configuration assembly and a single-cycle modulation assembly. The state division component is used for performing functional decoding according to the function mode of the equipment and generating a trigger signal matched with a time sequence according to the specific function mode; the synchronous multi-cycle configuration component is used for generating multiple types of cycle / non-cycle enable signals under different dominant frequencies according to the configuration information so as to realize multi-cycle time sequence driving; the single-cycle modulation assembly can be used as a universal modulation unit, and each unit can generate periodic in-compact time sequence driving signals according to configuration information and multi-cycle driving signals, so that the driving requirements of digital and analog reading units of different equipment chips are met. According to the invention, the overall time sequence structure of the driving chip can be greatly simplified, and the remote controller can realize the output of multi-period modulation driving signals in batches by using the programmable interconnection connection sub-module array generated by software; and the requirements of different devices on flexible configuration and real-time switching of periodic and non-periodic driving signals in multiple modes can be met without additional independent customized design. The design complexity is reduced, each module subunit can be configured with an independent time sequence outside a chip, the error-tolerant rate and flexibility of chip time sequence design are improved, and meanwhile the design period is shortened through reusability of multiple modules.
Owner:NANJING VPS SEMICONDUCTOR TECHNOLOGY CO LTD

Computer architecture 3D bus interrupt

A multiple CPU pseudo 3D structure is provided that allows single clock cycle interrupt latency, requires no context storage while taking only a single cycle away from normal programs. Multiple interrupts are given flexible vectored parallel computing responses without timing interactions.
Owner:ICAT LLC D B A TURING MICRO

A storage coherent hub chip based on core particle integration, a storage coherent arbitration device and an adaptive control method

The present application relates to the technical field of multi-core heterogeneous computing, and discloses a storage coherent hub chip based on core integration, a storage coherent arbitration device and an adaptive control method, aiming to solve the technical problems of high cross-core memory access latency, low coupling degree of cache coherence maintenance and memory scheduling, and slow response of operation strategy adjustment under the existing multi-core architecture. The present application takes an independently packaged storage coherent core as a globally consistent unique maintenance node, integrates a single-cycle static addressing architecture, adopts a cache coherence state machine and a memory scheduling controller with deep fusion of logic layers, realizes automatic switching between robust mode and aggressive mode through a pure hardware MHM monitoring unit, and the atomic withdrawal process is transparent to the upper layer. The present application can reduce the cross-core memory access latency by more than 40%, improve the storage access throughput by 25%, while guaranteeing 99.999% operation reliability, adapting to various heterogeneous interconnection protocols, and being applicable to various application scenarios such as servers, high-frequency financial transactions, AR / VR wearable devices, edge computing, etc.
Owner:胡青

Gene data matching coding optimization method and device based on in-memory calculation

The invention provides a gene data matching coding optimization method and device based on in-memory calculation, and the method comprises the steps: selecting data in a complete gene symbol sequence through a sliding window to obtain sliding window data, and storing the sliding window data in an SRAM array which comprises N * M rows and N * A columns; the one-time iteration process comprises the following steps: starting from a starting symbol in sliding window data, simultaneously matching a symbol sequence to be matched and corresponding symbols of all columns in an SRAM (Static Random Access Memory) array in a single period by using OneCycleByteSearch, stopping matching until effective result vectors are all 0, recording a result of a result vector FindPos in a previous period as a longest matching starting position, and carrying out one-time iteration on the result vector FindPos in the previous period as a longest matching starting position; subtracting 1 from the number of the currently used time periods as the longest matching length; if the last symbol in the sliding window data is already carried out, the initial position of the longest matching is the result of the result vector FindPos of the current period, and the longest matching length is N.
Owner:INST OF COMPUTING TECH CHINESE ACAD OF SCI

Processing circuit architecture supporting multi-calculation precision dynamic switching

The invention relates to the technical field of integrated circuit design, in particular to a processing circuit architecture supporting multi-calculation precision dynamic switching, which comprises a control port module, a partial product generation module, a symbol compression module, an addition compression tree module and a final adder module. And single-cycle dynamic switching of various precisions is realized. The partial product generation module adopts a Booth coding algorithm to split an operation vector, and cooperates with boundary symbol selection logic to solve the problem of symbol expansion and auxiliary bit overlapping; the symbol compression module compresses the redundant extension bits through a preset coding mode; the addition compression tree module is formed by cascading multiple stages of compressors and inserting carry blocking logic; the final adder module is composed of a plurality of carry lookahead adders, and the output bit width is dynamically controlled through blocking logic. The architecture optimizes the parallel operation efficiency and the resource utilization rate, adapts to the deep learning training and reasoning full scene, and has the advantages of real-time performance and low power consumption.
Owner:GUANGDONG INST OF INTELLIGENT SCI & TECH

Micro-controller chip containing multi-protocol communication interface peripheral and operation method thereof

Disclosed are a micro-controller chip containing a multi-protocol communication interface peripheral and an operation method thereof. The micro-controller chip comprises a multi-protocol communication interface peripheral, the multi-protocol communication interface peripheral is connected to a system bus, the multi-protocol communication interface peripheral is connected with an I / O port, the multi-protocol communication interface peripheral comprises an exclusively used RISC instruction set micro-kernel, a code memory and a code program stored on the code memory and executable by the RISC instruction set micro-kernel, the code program at least comprises two bit operation instructions of 1 setting and 0 clearing, the instructions are single-cycle instructions, and when the RISC instruction set micro-kernel executes the code program, the I / O port outputs 1 or 0.
Owner:NANJING QINHENG MICROELECTRONICS CO LTD

A multi-level branch predictor supporting branch target buffer compression and a prediction method

The application discloses a multi-stage branch predictor supporting branch target buffer compression and a prediction method. The predictor comprises a three-stage structure. The first-stage branch predictor is a single-cycle predictor, which is used for giving a prediction result in one cycle. The second-stage branch predictor is a double-cycle predictor, which is used for giving a prediction result in the second cycle. The third-stage branch predictor is a triple-cycle predictor, which is used for giving a prediction result in the third cycle. A per-cycle PC multiplexer simultaneously sends the selected PC into the three-stage branch predictor, and the three-stage branch predictor respectively gives corresponding prediction results. The low delay of the low-stage predictor guarantees the continuous flow of instructions in the pipeline. The high accuracy of the high-stage predictor reduces the pipeline refresh times caused by branch prediction failure. The combination of the short prediction cycle of the low-stage branch predictor and the high prediction accuracy of the high-stage branch predictor optimizes and compresses the branch target buffer structure of the branch predictor, and reduces the area of the branch target buffer.
Owner:NORTHWESTERN POLYTECHNICAL UNIV

Multistage branch predictor supporting branch target cache compression and prediction method

The invention discloses a multi-stage branch predictor supporting branch target cache compression and a prediction method, the predictor comprises a three-stage structure, the first-stage branch predictor is a single-period predictor and is used for giving a prediction result in a period, the second-stage branch predictor is a double-period predictor and is used for giving a prediction result in a second period, and the second-stage branch predictor is used for giving a prediction result in a third period. And the third-stage branch predictor is a three-period predictor and gives a prediction result in a third period. In each period, the PC multiplexer sends the selected PC to the three-stage branch predictor at the same time, the three-stage branch predictor gives corresponding prediction results, the low delay of the low-stage predictor is utilized to ensure that the instruction is not interrupted in the assembly line, and the high accuracy of the high-stage predictor is utilized to reduce the number of times of assembly line refreshing caused by branch prediction failure. The short prediction period of the low-level branch predictor and the high prediction precision of the high-level branch predictor are combined, the branch target cache structure of the branch predictor is optimized and compressed, and the area of the branch target cache is reduced.
Owner:NORTHWESTERN POLYTECHNICAL UNIV

Scheduling control method, scheduling control device, electronic equipment and storage medium

The invention provides a scheduling control method, a scheduling control device, electronic equipment and a storage medium, and relates to the technical field of intelligent warehouse management, and the method comprises the steps: generating a plurality of groups of operation basic sequences in a full arrangement form based on Monte Carlo random sampling; aiming at each group of operation basic sequences, generating an initial operation sequence of each crane corresponding to the group of operation basic sequences, and replacing tail end tasks of each initial operation sequence; performing discrete event simulation on each operation sequence after replacement by taking a rigid demand period of a power generation side as a step length, and performing screening by taking single-period consumption meeting the power generation side in the whole process as a constraint condition to obtain a candidate operation sequence meeting the constraint condition; and selecting a sequence meeting a preset optimization index from the candidate operation sequences as a final scheduling sequence, and controlling each crane to operate based on the final scheduling sequence. According to the method and the device, the final scheduling sequence is generated to coordinate the operation of the multiple cranes, so that the congestion is avoided, and the working efficiency is improved.
Owner:BEIJING SHIDAI CHONGSHU TECHNOLOGY CO LTD

Single-cycle control method and related device for PFC circuit

The present invention discloses a single-cycle control method for a PFC circuit and related devices. The PFC circuit includes a PFC main loop, a voltage loop controller, and a control loop. The method, executed by the control loop, includes: obtaining a first regulated voltage value output by the voltage loop controller, and output voltage sampled values, output current sampled values, input voltage value, and input current value of the PFC main loop; determining a feedforward control variable based on the output voltage sampled values, output current sampled values, and input voltage value; summing the first regulated voltage value and the feedforward control variable to obtain a second regulated voltage value; determining a duty cycle based on the input voltage value, input current value, output voltage sampled value, and second regulated voltage value; and generating a pulse width modulation signal based on the duty cycle, the pulse width modulation signal being used to regulate and control the PFC main loop. The present invention enables rapid regulation and control of the PFC circuit, improves control effectiveness, and facilitates miniaturization of the PFC circuit.
Owner:SHINRY TECH

32-bit processor based on RISC-V instruction set architecture

The invention discloses a 32-bit processor based on an RISC-V instruction set architecture, and belongs to the field of computer system structures and microprocessor design. According to the technical scheme adopted by the invention, the 32-bit processor based on the RISC-V instruction set architecture comprises an instruction set extension module, two new instructions of'aggresgate 'and'disaggresgate' are introduced, an R-type coding format is adopted, and the 32-bit processor is used for realizing grouping and splitting operation of data bits; the data path unit comprises a general register group, an arithmetic logic unit ALU and a special adder and is used for executing various data operations and processing; the control unit is used for generating a control signal according to an instruction decoding result and controlling operation of the data path unit and the memory; according to the method, by adding the special instruction and optimizing the data flow control mechanism, the complex bit operation is completed in a single period, the cache pressure is reduced, and the method has the advantages of improving the bit operation efficiency, reducing the instruction cache pressure and reducing the dynamic power consumption.
Owner:济南晶谷研究院 +1

A multi-cycle modulation driving component, method, system, and product

This invention proposes a multi-cycle modulation driving component, method, system, and product. The multi-cycle modulation driving component includes a state division component, a synchronous multi-cycle configuration component, and a single-cycle modulation component. The state division component performs functional decoding based on the device's functional mode and generates timing-matched trigger signals according to the specific functional mode. The synchronous multi-cycle configuration component generates various types of periodic / aperiodic enable signals at different clock frequencies based on configuration information, realizing multi-cycle timing drive. The single-cycle modulation component can serve as a general-purpose modulation unit; each unit can generate periodic composite on-chip timing drive signals based on configuration information and multi-cycle drive signals, thereby meeting the driving requirements of digital and analog readout units of different device chips. This invention can greatly simplify the overall timing structure of the driving chip.
Owner:NANJING VPS SEMICONDUCTOR TECHNOLOGY CO LTD

A signal acquisition method and a chip verification platform

The present application provides a signal acquisition method and a chip verification platform. The method includes: when an exception occurs during the process of running the bitstream file of the design under test of the chip, using the clock counting module in the bitstream file to record the time point when the exception occurs; during the exception debugging process, using the clock control module to control the design under test to run within a first running time period, and the right endpoint value of the first running time period is not greater than the time point; when the running duration of the design under test reaches the first running time period, controlling the design under test to run within a second running time period; during the process of the design under test running within the second running time period, using the clock control module to control the signal acquisition module to acquire the signals to be observed during the exception debugging process in a single-cycle manner.
Owner:NEW H3C SEMICON TECH CO LTD

Single-cycle multi-data packet verification and retransmission method, electronic device, and medium

The present invention relates to the field of communication technology, and in particular to a single-cycle multi-data packet verification and retransmission method, electronic device, and medium. The method can perform parallel verification of data packets to be transmitted, and the verification process of each data packet to be transmitted has no mutual constraints. The logical operation is simple, which greatly reduces the time overhead of the logical operation and reduces the delay of single-cycle multi-data packet verification and retransmission. In addition, compared with the traditional verification method, the data verification method of the present invention does not require additional storage space, and provides the possibility for high-frequency, high-bandwidth, low-power, and small-area data transmission operation circuits. Compared with the traditional operation method, the present invention has small structural changes and simple integration, can effectively improve the operation efficiency of the communication system, and reduce data transmission delay. The larger the data bandwidth and the more data packets are transmitted in a single clock cycle, the more obvious the performance improvement. The present invention reduces the power consumption and area of ​​the verification and retransmission circuit, and improves the performance of the communication system.
Owner:SHANGHAI UNIVISTA IND SOFTWARE GRP CO LTD +2

Debugging method and system for RISCV system memory

The application discloses a kind of RISCV system memory debugging method and system, its method includes: the interrupt request function of RISCV kernel is configured, enter debugging action and exit debugging action do not trigger the exception mechanism of RISCV processor;With host computer connection RISCV system, enter debugging state;The single-cycle valid signal triggered by host computer execution enter debugging action is converted into long-time valid signal;When long-time valid signal is valid, judge whether RISCV kernel has read-write, if yes, determine that the read-write signal of debugging module is invalid, set the state flag bit of interactive register, notify host computer system busy;If no, determine that the read-write signal of debugging module is valid, debugging module executes read-write to memory;Debugging module completes memory read-write and receives host computer exit debugging signal after exit debugging action, and make long-time valid signal invalid.The application utilizes the gap that RISCV kernel does not read nor write, realizes the read-write of debugging module to memory, realizes no interrupt memory debugging.
Owner:XIAMEN XINSIWANG INTEGRATED CIRCUIT TECH CO LTD

A single cycle delay compensation method and device for a magnetic suspension bearing switching power amplifier

This invention belongs to the field of active magnetic levitation bearing technology, and discloses a single-cycle delay compensation method and device for a magnetic levitation bearing switching power amplifier. The method includes: establishing a unipolar single-cycle control mathematical model and calculating the duty cycle of a single cycle; setting a state switching criterion for the charging and discharging cycle and determining the charging and discharging state of a single cycle; establishing a linear duty cycle prediction model and deriving the duty cycle d linearly through calculation. t And calculate the duty cycle value d j The total duty cycle of the corresponding switching transistor within the cycle is obtained and output to the corresponding switching transistor. This invention enables the tracking of the target under static conditions without steady-state error, and under dynamic conditions, it can effectively reduce the time delay problem of traditional digital single-cycle control algorithms.
Owner:SHAANXI UNIV OF SCI & TECH

Data separation method and separation system

The present invention provides a data separation method and a separation system, wherein the data separation method comprises the following steps: obtaining a full-cycle processing curve during workpiece processing, wherein the full-cycle processing curve comprises a plurality of single-cycle processing curves; selecting one of the single-cycle processing curves in the full-cycle processing curve as a template curve, and the other single-cycle processing curves as processing curves to be matched; determining a standard sample curve in the template curve; judging whether the standard sample curve matches each processing curve to be matched; judging whether to update the standard sample curve based on the matching results between n adjacent processing curves to be matched that match the standard sample curve and the standard sample curve; and separating the template curve before the update and all processing curves to be matched that match the standard sample curve in the template curve.
Owner:SIGER

Serialized data link with zero-cycle or short-cycle paths

Systems and methods are disclosed for reduced-power serial data links. A clock-forwarded serial link carries a clock lane and one or more data lanes. Every active serial data cycle is accompanied by its own serial clock edge: a clock delay allows the same clock edge to drive data at a transmitter and latch data at a receiver. Power is saved by idling the serial clock when data is not being transmitted. A valid signal can be omitted, providing a space saving. At the destination, similar clock-forwarding and delay enables a single parallel clock edge to drive data to the boundary of its clock domain, e.g. from a deserializer to a FIFO. The data link exhibits zero-cycle entry and exit. Variations with half- or single-cycle entry or exit are disclosed.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

A RISC-V-based hardware development system

The application provides a RISC-V-based hardware development system, comprising an instruction tracking module, a waveform slicing module, a deadlock detection module, an emulated storage tracking module, a function call tracking module, a peripheral access tracking module, an instruction emulator, a differential debugging comparison module, a comprehensive debugging module and a calling module; wherein the waveform period recorded by the waveform slicing module each time can be configured by itself, greatly reducing the space occupied by the generated waveform file; the differential debugging comparison module can adaptively process the execution period of each instruction of a processor, that is, the module can be used for a processor with single period or a processor using multi-stage pipeline technology. Through the use of the RISC-V-based hardware development system of the application for RISC-V processor design, the development and debugging efficiency of the processor can be greatly improved.
Owner:SHENZHEN UNIV

In-memory computing device, in-memory computing method, processing device, tile module and accelerator

The application discloses a memory-computing integrated device, a memory-computing method, a processing device, a tile module and an accelerator, and relates to the technical field of electronic circuits, and comprises: each memory-computing integrated array realizes parallel computation within the array, the memory-computing integrated array can simultaneously perform multiplication operation on single-bit input pulses at each moment in the current cycle buffer and corresponding weights, and synchronously generate membrane potential increment values at each moment; and the parallel accumulation of the increment values by a summation tree forms an efficient computing link of "parallel multiplication + parallel accumulation". Compared with the delay caused by the step-by-step waiting of the row-by-row serial processing of a computing task, the parallel architecture greatly shortens the processing time of the "multiplication-accumulation" whole process in a single cycle under the premise of ensuring high computing precision, further improves the computing efficiency under unit energy consumption due to the reduction of repeated data scheduling and state switching overhead in serial computation, and finally realizes the double optimization of computing delay reduction and energy efficiency ratio.
Owner:SHANDONG YUNHAI GUOCHUANG CLOUD COMPUTING EQUIP IND INNOVATION CENT CO LTD

System and method for single cycle dynamic hysteresis error control

A signal quantization module comprising: a combining module configured to combine an input signal and a feedback signal to generate a combined signal; an integrator module coupled to the combining module and configured to generate an integrated signal using the combined signal; and a quantizer configured to generate an output signal based at least in part on the integrated signal, wherein the output signal is fed back to provide the feedback signal to the combining module and the integrator module, and wherein the quantizer is further configured to, when the quantizer is triggered at a first point in time, compare an input to the quantizer to a threshold hysteresis, and in response to the input to the quantizer reaching the threshold hysteresis, adjust the threshold hysteresis by an amount that characterizes a difference between a target hysteresis and the input to the quantizer at the first point in time.
Owner:JAMES HAMMOND PTY LTD

An entropy decoding decoder suitable for HEVC and its optimization method

The present invention discloses an entropy decoding decoder optimization method suitable for HEVC. According to the characteristics of different syntax elements, multiple conventional arithmetic decoders are grouped and parallelized to form a five-way output device, thereby realizing the parallel calculation of multiple bins in a single-cycle clock, achieving parallel calculation of entropy decoding at the circuit level, and having good versatility. The steps of parallelizing multiple conventional arithmetic decoders are as follows: obtaining the bin currently undergoing conventional arithmetic decoding, and outputting the corresponding conventional arithmetic decoder drive signal when the current bitstream pointer points to the syntax element required by the corresponding module. For the renormalization operation in the conventional arithmetic unit, the conventional arithmetic unit input is clipped to the upper 5 bits of the current 8-bit bitstream, and the ivlCurrRange interval size comparison is improved to a shift judgment selector.
Owner:GUANGZHOU XINHUA TECHNICAL SERVICE CO LTD

Dynamic load balancing mapping method and system based on interrupt transaction combination

The invention discloses a dynamic load balancing mapping method and system based on interrupt transaction combination. The method comprises the following steps: an interrupt synchronization processing step; an interruption mark setting step; a load state acquisition step; a dynamic combination control step; a channel mapping step; a target source processing step; and a closed loop feedback step. The invention also comprises a system for implementing the method. According to the method, the transformation of an interrupt scheduling normal form is realized through'dynamic load perception + hardware closed-loop architecture + elastic scheduling strategy ', firstly, post statistics is replaced by real-time quantitative load perception, and the problem that a scheduling decision and a real-time demand are disjointed is solved; secondly, due to full-hardware implementation, software intervention overhead is eliminated, and scheduling delay is reduced to a single-cycle level; and finally, an elastic combination mechanism and a priority collaboration strategy ensure the deterministic response of the key interruption.
Owner:HUNAN GREAT WALL GALAXY TECH CO LTD

Bounded-Carry Fixed-Delay FP8 Accumulator for Pipelined Mixed-Precision Processing Elements

PendingUS20260203021A1Computer architectureCarry propagation
A bounded-carry partial-sum adder for FP8 accumulation in a pipelined mixed-precision processing element is disclosed. The adder limits carry propagation to a predetermined depth, ensuring deterministic fixed delay independent of operand magnitude. The accumulator forms the second stage of a two-stage fused multiply-add pipeline and supports a one-cycle initiation interval while accumulating products of FP4 weight operands and FP8 activation operands. The architecture enables uniform high-frequency timing across dense arrays of processing elements.
Owner:SILVEBROOK KIA

Encoder, Encoding Method, and Chip

PendingUS20260205142A1AlgorithmForward error correction
An encoder includes: a feedforward module configured to receive to-be-encoded data with a parallelism of n symbols per cycle, and perform calculation in a finite field based on a symbol of the to-be-encoded data in a current cycle, to obtain a single-cycle polynomial corresponding to the current cycle; and a feedback module configured to receive the single-cycle polynomial corresponding to the current cycle output by the feedforward module, and perform calculation in the finite field based on the single-cycle polynomial corresponding to the current cycle and a first polynomial indicating a symbol received in a historical cycle, to obtain a second polynomial, where the second polynomial is used to determine a target polynomial indicating the to-be-encoded data, and the target polynomial is to generate a check sequence of the to-be-encoded data, to obtain a codeword obtained by encoding the to-be-encoded data based on a forward error correction (FEC) encoding scheme.
Owner:HUAWEI TECH CO LTD