Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

18 results about "Block floating-point" patented technology

Block floating point (BFP) is a method used to provide an arithmetic approaching floating point while using a fixed-point processor. The algorithm will assign an entire block of data an exponent, rather than single units themselves being assigned an exponent, thus making them a block, rather than a simple floating point. Block floating-point algorithm operations are done through a block using a common exponent, and can be advantageous to limit the space use in the hardware to perform the same functions as floating-point algorithms.

Stacked apparatus using in-memory compute chiplet devices for inference-time compute acceleration

A stacked apparatus using in-memory compute (IMC) chiplet devices for inference-time compute acceleration. The apparatus is configured to accelerate the workload computations for neural network models, such as those for Large Language Models (LLMs) and reasoning models. The apparatus achieves high throughput and low latency using a chiplet design, digital IMC (DIMC) based engines, efficient die-to-die (D2D) interconnects, block floating point (BFP) numerics, and large high bandwidth on-chip memories. With modular chiplets in stacked configurations with memory devices and efficient interconnects, the accelerator apparatus can be easily scaled to accelerate workloads for models of different sizes. The DIMC configuration within the chiplet slices also improves computational performance and reduces power consumption by integrating computational functions and memory fabric. And by dynamically switching between precision levels based on real-time analysis of a target workload, computational efficiency can be optimized while maintaining accuracy.
Owner:D-MATRIX CORP

Dynamic range channelization receiver method and system based on block floating point and AGC

The invention relates to the technical field of channelized receiver management, in particular to a dynamic range channelized receiver method and system based on a block floating point and AGC (Automatic Gain Control), in the system, an automatic gain control module is used for detecting the power of an obtained intermediate frequency analog signal, generating an AGC voltage corresponding to the intermediate frequency analog signal and feeding back the AGC voltage to an analog front end; and dynamically updating the gain of the variable gain amplifier in the radio frequency signal preprocessing process. According to the invention, each channel has independent block floating point gain control, and the channels do not interfere with each other, so that parallel signal-to-noise ratio processing is realized; meanwhile, a complete gain control history is constructed by jointly recording the AGC gain and block floating point processing, data support is provided for recovering the original absolute power value of the signal in each channel, and the contradiction between gain control and information reservation is solved to a certain extent.
Owner:NANJING NAT ELECTRONIC TECH CO LTD

Block floating point calculation device and differential equation calculation system

The invention provides a block floating point arithmetic device and a differential equation calculation system, and relates to the field of calculation devices, and the differential equation calculation system manages the flow of data organized in a differential format among storage hierarchies through a buffer controller. And a processing unit array synchronously executes Gaussian-Seidel iterative calculation for differential format optimization based on a wavefront sequence, then compresses a result through a block floating point quantizer, and outputs a solution matrix through a differential reduction unit after convergence, a plurality of parallel processing units in the processing unit array combine precision configuration with a block floating point input format to form a collaborative optimization calculation mode, the complexity of mantissa processing is further controlled in the calculation process by sharing index compression data and dynamic precision parameter permission, and the problems of resource occupation requirements, memory access times and high energy consumption are solved.
Owner:NANJING UNIV

Data coding method and device based on block chain, equipment and storage medium

The invention discloses a data coding method and device based on a block chain, equipment and a storage medium. The data coding method comprises the following steps: reading floating point data in a block body of a first block; encoding the floating point data to obtain block floating point data; and storing a sign bit and a mantissa bit of the block floating point data into a block body of the first block, and storing a sharing index of the block floating point data into a block head of the first block. According to the application, the floating point data in the block body of the first block is encoded to obtain the block floating point data; the sign bit and the mantissa bit of the block floating point data are stored in the block body of the first block, and the sharing index of the block floating point data is stored in the block head of the first block, so that the storage space of the index bit is released, the on-chain storage overhead and the network transmission load of the block chain can be reduced, the data transmission speed is improved, and the user experience is improved. And the performance of the block chain is optimized.
Owner:SHENZHEN LINZHOU TECHNOLOGY CO LTD

Neural network activation compression with non-uniform mantissa

Apparatuses and methods for training a neural network accelerator using quantization precision data formats are disclosed, and in particular for storing activation values from a neural network in a compressed format having lossy or non-uniform mantissa for use during forward and backward propagation training of the neural network. In certain examples of the disclosed technology, a computing system includes a processor, a memory, and a compressor in communication with the memory. The computing system is configured to perform forward propagation for a layer of a neural network to produce first activation values in a first block floating point format. In some examples, the activation values generated by the forward propagation are converted by the compressor to a second block floating point format having a non-uniform and / or lossy mantissa. The compressed activation values are stored in the memory, where they can be retrieved for use during backward propagation.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Enhanced block floating point compression for open radio access network fronthaul

A distributed unit (DU) signals a maximum IQ data bit width for downlink communications associated with a zone identifier (ID) to a radio unit (RU). The DU signals a per physical resource block (PRB) bit width parameter for downlink communications to the RU. The DU transmits the downlink communications based on at least one of the maximum IQ data bit width or the bit width parameter for the PRB, and the RU receives the downlink communications based on at least one of the maximum IQ data bit width or the bit width parameter for the PRB. For uplink communications, the DU transmits a first indication of a maximum IQ data bit width in a control plane message to the RU. The RU transmits a second indication of a per PRB bit width parameter for uplink communications to the DU. The RU transmits the uplink communications based on at least one of the maximum IQ data bit width or the bit width parameter for the PRB, and the DU receives the uplink communications based on at least one of the maximum IQ data bit width or the bit width parameter for the PRB.
Owner:QUALCOMM INC

Cross-chain data sending method and receiving method, electronic equipment and storage medium

The invention discloses a cross-chain data sending method and receiving method, electronic equipment and a storage medium. The cross-chain data sending method comprises the following steps: performing block coding on first floating point data to be transmitted to obtain first block floating point data; performing structured packaging on the first block of floating point data to obtain a second block of floating point data; and performing cross-chain transmission on the second block of floating point data to send the second block of floating point data to the second block chain. In a cross-chain scene, data is compressed through block coding, the on-chain storage overhead and the network transmission load of the block chain are reduced, and the data transmission speed and the cross-chain communication efficiency are improved; through structured packaging, the problem of inconsistent analysis of traditional floating point data in a multi-chain environment is avoided, and the exchange efficiency and consistency of the floating point data between multi-chain systems are improved. The method is suitable for cross-chain floating point calculation or data synchronization scenes, in particular to traceability chains, cross-chain model calculation, credible floating point data verification and other scenes with low data precision requirements but large data volume.
Owner:SHENZHEN LINZHOU TECHNOLOGY CO LTD

Floating point tensor data auditing method and device, equipment and storage medium

The invention discloses a floating point tensor data auditing method and device, equipment and a storage medium. The floating point tensor data auditing method comprises the following steps: dividing floating point tensor data to be processed into a plurality of blocks with fixed sizes; performing block coding on the floating point tensor data in each block to obtain block floating point tensor data; and in a block chain node, auditing the block floating point tensor data. According to the method, the floating-point tensor data is subjected to block processing, the floating-point tensor data is compressed and calculated in a block floating-point representation mode, and the data scale is remarkably compressed, so that the on-chain storage and bandwidth overhead is reduced, and the on-chain auditing efficiency is improved. The method can be implemented on the intelligent contract and consensus mechanism of the existing block chain platform (such as Ethereum, Fabric and the like), and has relatively high landing capability.
Owner:SHENZHEN LINZHOU TECHNOLOGY CO LTD

A hardware implementation method and device of a GAN network, a storage medium and a terminal

The application discloses a kind of hardware implementation method, device, storage medium and terminal of GAN network, method includes: training GAN network obtains the floating-point number model after convergence, and exports the floating-point number model after convergence;According to the network parameter of GAN network, floating-point number model is converted into block floating-point model;The block floating-point convolution structure of block floating-point model is deployed on hardware, and convolution operation is carried out based on the block floating-point convolution structure after deployment.The hardware implementation method provided in the application has the advantages of simple implementation steps, less network parameter precision loss, low operation complexity and significantly reduced storage unit requirements, solves the problem that existing generative adversarial network has large resource overhead and slow inference speed when deployed on hardware.
Owner:INST OF MICROELECTRONICS CHINESE ACAD OF SCI LTD

Acceleration chip, data processing method and application

The invention discloses an acceleration chip which comprises an off-chip storage unit used for storing input activation data and weight matrix data of a block floating point data structure; the on-chip cache unit is used for caching block floating point data to be calculated; the block floating point number direct carrying unit is used for carrying block floating point data between the off-chip storage unit and the on-chip cache unit; the block floating-point number matrix multiply-add unit is used for executing mantissa fixed-point multiply-add operation based on the block floating-point data structure and generating a floating-point result; a data format conversion unit for converting between a floating point format and a block floating point data structure; and the scheduling control unit is used for coordinating the work of the block floating point number direct carrying unit, the block floating point matrix multiplication and addition unit and the data format conversion unit. The invention further discloses a data processing method which has a wide application prospect.
Owner:SHANGHAI QUSU CHAOWEI TECHNOLOGY CO LTD

Neural network device performing floating point operation and operating method thereof

ActiveCN113807493BDigital data processing detailsAnalogue-digital convertersComputer hardwareMultiply–accumulate operation
A neural network device performing floating point operations and an operating method thereof are provided. The neural network device performs a multiply-accumulate (MAC) operation for a product of a fraction of a weight and an input activation in a block floating point format by using an analog crossbar array, performs an addition operation for a shared exponent of the weight and the input activation in the block floating point format by using a digital computing circuit, and outputs a partial sum of a floating point output activation by combining a result of the MAC operation with a result of the addition operation.
Owner:SAMSUNG ELECTRONICS CO LTD

BLOCK FLOATTING POINT CALCULATIONS WITH COMMON EXPONS

ActiveDE602019086343T2Block floating-pointStructural engineering
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Neural network activation compression with non-uniform mantissas

Apparatuses and methods are disclosed for training neural network accelerators using quantized precision data formats, and in particular for storing activation values from neural networks in a compressed format with lossy or non-uniform mantissas for use during forward and backward propagation training of neural networks. In some examples of the disclosed technology, a computing system includes a processor, a memory, and a compressor in communication with the memory. The computing system is configured to perform forward propagation for layers of a neural network to generate a first activation value in a first block floating point format. In some examples, an activation value generated by forward propagation is converted by a compressor into a second block floating point format having a non-uniform and / or lossy mantissa. The compressed activation value is stored in a memory in which the compressed activation value may be retrieved for use during backward propagation.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Block floating point operation device and differential equation calculation system

The application provides a block floating point operation device and a differential equation calculation system, and relates to the field of calculation devices. The differential equation calculation system manages the flow of data organized in a difference format between storage levels through a buffer controller, synchronously performs Gauss-Seidel iteration calculation optimized for the difference format based on a wavefront sequence by a processing unit array, compresses the result through a block floating point quantizer, and outputs a solution matrix through a difference reduction unit after convergence. In the processing unit array, multiple parallel processing units combine precision configuration with a block floating point input format to form a cooperatively optimized calculation mode. By sharing index compression data, dynamic precision parameters allow further control of the complexity of the mantissa during the calculation process, solving the problems of high resource occupation demand, memory access frequency and energy consumption.
Owner:NANJING UNIV

Data compression method and device, data decompression method and device, communication device and medium

The invention provides a data compression method and device, a data decompression method and device, a communication device and a medium, which can be applied to the technical field of communication. The method is applied to a first communication device and comprises the following steps: performing block floating point compression on first data to be transmitted to obtain a plurality of index factors and mantissa data corresponding to the index factors respectively; performing decimal bit quantization on the mantissa data to obtain a plurality of quantized data blocks; the bit width of the quantized data included in the quantized data block is smaller than that of the mantissa data; second data is transmitted to a second communication device, the second data including the plurality of exponential factors and the plurality of quantized data blocks. The compression mode provided by the invention has good compression capability, reduces the transmission bandwidth demand, and improves the bandwidth utilization rate.
Owner:RUIJIE NETWORKS CO LTD

Compression and storage of neural network activations for backpropagation

Apparatus and methods for training a neural network accelerator using quantized precision data formats are disclosed, and in particular for storing activation values from a neural network in a compressed format for use during forward and backward propagation training of the neural network. In certain examples of the disclosed technology, a computing system includes processors, memory, and a compressor in communication with the memory. The computing system is configured to perform forward propagation for a layer of a neural network to produced first activation values in a first block floating-point format. In some examples, activation values generated by forward propagation are converted by the compressor to a second block floating-point format having a narrower numerical precision than the first block floating-point format. The compressed activation values are stored in the memory, where they can be retrieved for use during back propagation.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

Adjusting precision and topology parameters for neural network training based on a performance metric

Apparatus and methods for training neural networks based on a performance metric, including adjusting numerical precision and topology as training progresses are disclosed. In some examples, block floating-point formats having relatively lower accuracy are used during early stages of training. Accuracy of the floating-point format can be increased as training progresses based on a determined performance metric. In some examples, values for the neural network are transformed to normal precision floating-point formats. The performance metric can be determined based on entropy of values for the neural network, accuracy of the neural network, or by other suitable techniques. Accelerator hardware can be used to implement certain implementations, including hardware having direct support for block floating-point formats.
Owner:MICROSOFT TECHNOLOGY LICENSING LLC

System and method for tiny machine learning using block floating point

A system includes a first FP-to-BFP converter, a second FP-to-BFP converter, an 8-bit integer multiplier, an adder, an accumulator, and a BFP-to-FP converter. The first and second FP-to-BFP converters receive 32-bit floating-point pixel and filter data, respectively, reducing their mantissas to 8-bit BFP format. The 8-bit integer multiplier processes these BFP values via multiply-accumulate operations, generating a 16-bit product. The adder accumulates multiple 16-bit products into a 64-bit sum, which the accumulator further aggregates. The BFP-to-FP converter transforms the 64-bit accumulated sum into a 32-bit floating-point output.
Owner:HONG KONG APPLIED SCI & TECH RES INST