Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

49 results about "Block transform" patented technology

Video encoding and decoding

Some aspects of the disclosure provide a method of video decoding. In some examples, values of one or more target quantization coefficients in a target region of quantization coefficients of a current block are obtained. A sub-block transform (SBT) mode of the current block is derived based on the values of the one or more target quantization coefficients in the target region. The current block is reconstructed based on the SBT mode of the current block. Apparatus and non-transitory computer-readable storage medium counterpart embodiments are also contemplated.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Two-stage point cloud attribute encoding scheme with nested local and global transformations

PendingCN122349731APoint cloudAlgorithm
Some embodiments of a method can include obtaining a point cloud, which can include a first set of information describing geometry of the point cloud and a second set of information describing attributes of the point cloud; performing geometry encoding of the first set of information to generate a geometry bitstream; performing a two-stage attribute compression process to generate an attribute bitstream, wherein a first stage of the two-stage attribute compression process includes performing a block transform on each node of a set of nodes of the point cloud, and wherein a second stage of the two-stage attribute compression process includes performing hierarchical encoding on the set of nodes of the point cloud, and outputting an output bitstream including the geometry bitstream and the attribute bitstream.
Owner:INTERDIGITAL CE PATENT HOLDINGS SAS

Video coding and decoding method and apparatus

PCT designated stageWO2026087173A1Digital video signal modificationVideo encodingBlock transform
A method and apparatus are provided for video encoding and decoding. Regarding a method for encoding, the method includes encoding a coding unit using a sub-block transform including a plurality of transform units. Encoding the coding unit includes defining a residual block for a first transform unit of the plurality of transform units. The first transform unit extends across a medial portion of the coding unit to at least one of a first pair of opposed edges of the coding unit while being spaced by transform blocks of second and third transform units of the plurality of transform units from a second pair of opposed edges of the coding unit. The method also includes causing storage and / or transmission of motion information, a prediction block identified by the motion information to be associated with the coding unit, and information regarding the residual block for the first transform unit.
Owner:NOKIA TECHNOLOGIES OY

Subblock transform for intra prediction coding block

An example method of video decoding includes receiving a video bitstream that includes multiple blocks, including a current block. The method also includes identifying a partial region of the current block, where residual data for the current block outside of the partial region is zero, and reconstructing the current block by applying a subblock transform to the partial region of the current block. Instructions for the example method may be stored in a computer system or storage medium.
Owner:TENCENT AMERICA LLC

A deep fake face detection method based on dual domain

The present application belongs to the field of computer vision and image processing, aiming at the double bottleneck of "semantic overfitting" and "space-frequency feature mislocation" existing in the existing deep fake detection, the present application proposes a deep fake face detection method based on dual domain. The present application constructs a dual domain complementary architecture of dual domain deep fake detection network DDSF-Net based on semantic suppression and space-frequency alignment, the micro mark difference module in the network significantly enhances the residual subtle fake artifacts between pixels by eliminating redundant face attributes. And the space-frequency mapping module uses the block transformation strategy to adaptively mine the abnormal distribution of frequency domain, realizes the multi-dimensional enhancement of the fake features. At the same time, through the complementary fusion of dual domain features, the detection accuracy and generalization are effectively balanced. Extensive experiments on multiple benchmark datasets show that the performance of the method in the same dataset and cross dataset scenarios is significantly better than that of the existing advanced technology.
Owner:TAIYUAN UNIVERSITY OF SCIENCE AND TECHNOLOGY +1

INTERACTION BETWEEN SECONDARY TRANSFORMATION AND SUB-BLOCK TRANSFORMATION

PendingID202606448ABlock transformVideo decoder
The present invention relates to a device (e.g., a decoder for video decoding) that can determine that a sub-block transform (SBT) is enabled for a video block. The device can determine whether a secondary transform (e.g., a low-frequency non-separable transform (LFNST) or a non-separable primary transform (NSPT)) is enabled for the video block. The device can determine a transform function to decode the video block based on determining whether the secondary transform is enabled for the video block. For example, a first type of transform function (e.g., discrete cosine transform (DCT) II) can be used to decode the video block if the LFNST or NSPT is enabled for the video block. A second type of transform function can be used to decode the video block if the secondary transform is disabled for the video block. The device can decode the video block using the transform function.
Owner:INTERDIGITAL CE PATENT HOLDINGS SAS

Encoder, decoder, system, method for decoding, method for encoding, data stream and computer program using a code word for the position of a transform coefficient

A decoder for decoding a one-dimensional digital waveform signal from a data stream by use of block-wise transform decoding, the decoder configured to decode a sequence of transform coefficients of a predetermined block of the one-dimensional digital waveform signal from the data stream by deriving a position information from the data stream, locating, using the position information, a position in the sequence of transform coefficients; and decoding one or more first transform coefficients of the sequence of transform coefficients from the data stream, which precede the position in the sequence of transform coefficients, and attributing a predetermined value to one or more second transform coefficients of the sequence of transform coefficients, which follow the position in the sequence of transform coefficients. Further disclosed is an encoder, a system, methods, a data stream and a computer program.
Owner:FRAUNHOFER GESELLSCHAFT ZUR FORDERUNG DER ANGEWANDTEN FORSCHUNG EV

Image and video coding and decoding

A method of encoding images into a bitstream, the method comprises determining a residual subblock comprising an area of a residual block for a block to be encoded, and encoding transform coefficients of a transform block to be encoded based on the residual subblock area. A corresponding method comprises decoding images from a bitstream. The methods may comprise obtaining a prediction information portion of a block, comprising a plurality of samples for the block and determining the residual subblock based on the prediction samples. The methods may comprise inferring the position of the residual subblock. The process may be an Advanced Subblock Transform (ASBT) process. The method may comprise selectively encoding an orientation of the transform block in the bitstream, and wherein the orientation of the transform block is context-adaptive binary arithmetic coding (CABAC) coded; wherein at least one CABAC context is associated with the orientation of the transform block. [Fig. 14]
Owner:CANON KK

A millimeter wave radar echo signal processing method and system

The present application relates to the technical field of radar sensor, and relates to a millimeter wave radar echo signal processing method and system. When the surrounding environment is detected by a millimeter wave radar, a transmitting array generates and transmits FMCW radar wave detection pulses, a receiving array receives echo signals of the radar wave detection pulses reflected by targets in the environment, and the echo signals enter a receiving signal processor. The received signal pulses are converted into radar digital baseband signal sequences by a radio frequency front end module, a radar echo digital baseband intermediate signal sequence is obtained through first step digital baseband signal processing, the radar echo digital baseband intermediate signal sequence is transmitted to hardware for performing second step digital baseband processing through a high-speed data interface, the second step digital baseband processing is performed, CFAR and AoA detection are finally performed, and information of the detected target is obtained. Compared with a traditional scheme, the data amount of inter-chip transmission is greatly reduced, so that the baseband data scale that can be processed by the radar system is increased, and the radar detection capability is improved.
Owner:GUIBU MICROELECTRONICS (NANJING) CO LTD

Subblock transform for intra prediction coding block

An example method of video decoding includes receiving a video bitstream that includes multiple blocks, including a current block. The method also includes identifying a partial region of the current block, where residual data for the current block outside of the partial region is zero, and reconstructing the current block by applying a subblock transform to the partial region of the current block. Instructions for the example method may be stored in a computer system or storage medium.
Owner:TENCENT AMERICALLC

Quantification method and device of feature data in model, storage medium and electronic equipment

PendingCN122242595ABiological modelsInference methodsData packBlock transform
This disclosure provides a method, apparatus, storage medium, and electronic device for quantizing feature data in a model. The method includes: acquiring feature data to be quantized from a target model deployed on a target device, wherein the feature data includes at least two data blocks; performing an inter-block transformation on the feature data using a pre-constructed inter-block orthogonal transformation matrix to obtain inter-block transformed data; performing an intra-block transformation on each data block included in the inter-block transformed data using a pre-constructed intra-block orthogonal transformation matrix to obtain intra-block transformed data; and quantizing the intra-block transformed data to obtain quantized feature data. This disclosure can evenly distribute the energy of the feature data across all blocks, eliminating the influence of high-variance blocks on the quantization sharing index, and redistributing the intra-block data so that the intra-block data is evenly mapped to each codeword interval of the codebook, solving the codebook collapse problem and thus significantly improving quantization accuracy.
Owner:NANJING HOUMO TECH CO LTD

Method and device for processing video signal

Embodiments in the present specification provide a method and device for processing video signal including steps of: obtaining a sub-block transform (SBT) flag indicating whether a SBT is applied, wherein the SBT represents a transform applied to a transform unit corresponding to one of subblocks split from a coding unit; determining a variable value related to a transform target region based on the SBT flag and a size value of the transform unit, wherein if the SBT flag equals to 1 and the size value is greater than a threshold value, the variable value is determined as the threshold value; obtaining position information of a last significant coefficient according to a scan order within the transform unit, wherein the position information of the last significant coefficient is coded based on the variable value; and performing an inverse-transform based on the position information of the last significant coefficient.
Owner:LG ELECTRONICS INC

On Subblock-Transform For Intra Video And Image Coding and Subblock-Transform Information Inferring

A mechanism for processing video data is disclosed. The mechanism includes determining to apply a sub-block transform (SBT) to an intra-predicted coding unit (CU). A conversion is performed between a visual media data and a bitstream based on the intra-predicted CU. For example, a CU can be split into subblocks for intra prediction and then the SBT can be applied to the residual of each subblock.
Owner:DOUYIN VISION CO LTD +1

Video coding method, apparatus and device

The application discloses a video coding method, device and equipment, and belongs to the technical field of audio and video. The method comprises the following steps: decoding a target coding unit to obtain a quantization coefficient matrix corresponding to the target coding unit; determining first reference information according to quantization coefficients in the quantization coefficient matrix; obtaining the value of a transform flag corresponding to the first reference information, wherein the transform flag refers to a flag of a sub-block transform position, and the sub-block transform position refers to the position of a sub-block in the coding unit which needs to be transformed and quantized; and determining the sub-block transform position of the target coding unit according to the obtained value of the transform flag. The application implicitly indicates the flag of the sub-block transform position in the target coding unit through the quantization coefficients in the quantization coefficient matrix corresponding to the coding unit, avoids explicit coding of the sub-block transform position, reduces the number of bits occupied by a video code stream, and improves the video coding efficiency.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Method and apparatus for video coding, device and computer readable medium

ActiveCN117499653BDigital video signal modificationAlgorithmBlock transform
Embodiments of the present disclosure relate to a method and apparatus for video coding, device and computer readable medium. A method for video decoding includes determining a maximum transform size associated with a video sequence based on a first high level syntax element of the video sequence, the maximum transform size corresponding to a maximum transform unit area. When a sub-block transform (SBT) mode is enabled, a maximum block size of the SBT mode is allowed to be constrained by the maximum transform size. Wherein the maximum block size of the SBT mode is allowed to be constrained by the maximum transform size includes allowing a width and a height of a maximum coding unit (CU) of the SBT mode to be derived as a minimum value between the maximum transform size indicated by the first high level syntax element and a width and a height of a maximum CU of the SBT mode allowed to be indicated in a second high level syntax element.
Owner:TENCENT AMERICA LLC

A Radar Signal Modulation Recognition Method Based on Temporal Modeling and Transvariable Fusion

PendingCN122307491AFeature extractionBlock transform
This invention provides a radar signal modulation recognition method based on temporal modeling and cross-variable fusion, comprising: preprocessing the radar IQ signal to be identified into blocks to obtain block tensors; constructing a hybrid neural network model, which includes a block transform temporal encoder, a modern temporal convolutional network cross-variable fusion module, and a classification head. The block transform temporal encoder is used to capture long-range temporal dependencies, and the modern temporal convolutional network cross-variable fusion module is used to perform cross-variable fusion on the extracted temporal feature tensors; inputting the block tensors into the trained hybrid neural network model for processing using the block transform temporal encoder, the modern temporal convolutional network cross-variable fusion module, and the classification head, and outputting radar signal modulation recognition results. This improves the model's adaptability and generalization ability; compensates for the shortcomings of local segmentation methods in global modeling capabilities; and achieves joint feature extraction of multi-channel radar signals.
Owner:XIDIAN UNIV

Sub-block transform in transform skip mode

A method for video processing includes determining, based on a first indication, whether a sub-block residual coding scheme is applied to residual of a current video block in a transform skip mode, the sub-block residual coding scheme splitting the residual of the current video block into multiple sub-blocks and a subset of the multiple sub-blocks have non-zero coefficients; determining, based on a second indication, a specific split pattern to be applied to the residual of the current video block, in response to the sub-block residual coding scheme being applied to the residual of the current video block; deriving, based on a third indication, the subset of the multiple sub-blocks which have non-zero coefficients; and performing a conversion on the residue of the current video block based on the determined subset of sub-blocks having non-zero coefficients.
Owner:DOUYIN VISION CO LTD +1

Method and apparatus for sub-block transform partitioning in video coding

According to one aspect of the present disclosure, a method of decoding is provided. The method may include, in response to a sub-block transform (SBT) with corner partitioning or center partitioning being enabled for a current block, determining, by a processor, whether a non-separable transform is enabled for the current block. The method may include, in response to the non-separable transform being enabled, determining, by the processor, the non-separable transform from among a plurality of non-separable transforms based on a size of a SBT partition. The method may include decoding, by the processor, the current block based on the non-separable transform.
Owner:GUANGDONG OPPO MOBILE TELECOMMUNICATIONS CORP LTD

Method and apparatus for sign coding of transform coefficients in video coding system

A method and apparatus for joint sign prediction of transform coefficients of residual blocks in a video coding system are disclosed. At the encoder side, a transform coefficient region or an index value range is determined according to coding context associated with the current block. A set of signs associated with a set of selected transform coefficients are determined for joint sign prediction are determined according to the transform coefficient region or the index value range. Joint sign prediction for the set of signs is determined by selecting a hypothesis from a group of hypotheses for the set of signs that achieves a minimum cost. The sign prediction is then used for coding the set of signs. A corresponding method and apparatus for the decoder side is also disclosed.
Owner:MEDIATEK INC

Coding concepts for transformed representations of sample blocks

ActiveJP7797609B2Digital video signal modificationAlgorithmBlock transform
To provide an encoder and a decoder which perform a choice between multiple pre-defined transform types for a block of a picture.SOLUTION: If a first coded coefficient position 102 is located inside a predetermined subarea 106 of a transform coefficient block 104, and if an underlayer transform is within a first set 132 of available transforms, a decoder decodes coefficients along a first coefficient scan order 110. If the transform is within a second set 134 of available transforms, the decoder decodes coefficients located within the predetermined subarea along a second coefficient scan order 114, and infers that coefficients located outside the predetermined subarea are zero. The first coefficient scan order is such that coefficients outside the predetermined subarea are scanned between two transform coefficients located inside the predetermined subarea.SELECTED DRAWING: Figure 7
Owner:FRAUNHOFER GESELLSCHAFT ZUR FORDERUNG DER ANGEWANDTEN FORSCHUNG EV

Video encoding method and video decoding method

PendingCN122248166ADigital video signal modificationBlock transformVideo encoding
This application relates to a video encoding and decoding method, apparatus, chip, computer device, computer-readable storage medium, and computer program product. The video encoding method includes: determining the target sub-block partitioning method of the current coding unit and the target transform kernel corresponding to the target sub-block partitioning method; partitioning sub-blocks from the current coding unit according to the target sub-block partitioning method; performing sub-block transform on the inter-frame prediction residuals corresponding to the sub-blocks using the target transform kernel to obtain transform coefficients; and obtaining the bitstream of the current coding unit based on the transform coefficients. The target sub-block partitioning method is one of multiple sub-block partitioning methods, which at least include: the sub-block being located in the middle region of the current coding unit in the vertical or horizontal direction, and the pixel height or pixel width of the sub-block being the pixel height of the current coding unit, and the pixel width or pixel height of the sub-block being 1 / 2 or 1 / 4 of the pixel height of the current coding unit. Using this method can improve the encoding effect.
Owner:HISENSE VISUAL TECH CO LTD

On subblock-transform for intra video and image coding and subblock-transform information inferring

PCT designated stageWO2026007990A1Digital video signal modificationAlgorithmBlock transform
A mechanism for processing video data is disclosed. The mechanism includes determining to apply a sub-block transform (SBT) to an intra-predicted coding unit (CU). A conversion is performed between a visual media data and a bitstream based on the intra-predicted CU. For example, a CU can be split into subblocks for intra prediction and then the SBT can be applied to the residual of each subblock.
Owner:DOUYIN VISION CO LTD +1

Coding method, decoder, decoding chip

PendingCN122349023ADecoding methodsBlock transform
This application provides an encoding / decoding method, a decoder, and a decoding chip that can reduce signaling overhead. The decoding method includes: determining the estimated residual of the current block; based on the estimated residual, calculating the distortion magnitude corresponding to the SBT modes of multiple sub-blocks in a first category, wherein the SBT modes in the first category have the same division ratio and division shape; and determining the SBT mode with the minimum distortion as the target SBT mode used by the current block.
Owner:HISENSE VISUAL TECH CO LTD

Method and device for processing video signal

PendingUS20260101064A1Digital video signal modificationAlgorithmBlock transform
Embodiments in the present specification provide a method and device for processing video signal including steps of: obtaining a sub-block transform (SBT) flag indicating whether a SBT is applied, wherein the SBT represents a transform applied to a transform unit corresponding to one of subblocks split from a coding unit; determining a variable value related to a transform target region based on the SBT flag and a size value of the transform unit, wherein if the SBT flag equals to 1 and the size value is greater than a threshold value, the variable value is determined as the threshold value; obtaining position information of a last significant coefficient according to a scan order within the transform unit, wherein the position information of the last significant coefficient is coded based on the variable value; and performing an inverse-transform based on the position information of the last significant coefficient.
Owner:LG ELECTRONICS INC

Video signal processing method and apparatus using multi-assumption prediction

PendingEP4637146A3Digital video signal modificationAlgorithmBlock transform
A video signal processing method comprises: obtaining a first syntax element indicating whether combined prediction is applied to the current block, wherein the combined prediction is a prediction mode that combines inter-prediction and intra-prediction; and if the first syntax element indicates that the combined prediction is applied to the current block, reconstructing the current block based on a combined prediction block, characterised in that the method further comprises: if the first syntax element indicates that the combined prediction is not applied to the current block, obtaining a second syntax element indicating whether sub-block transform is applied to the current block, wherein the sub-block transform indicates a transform mode that applies transform to one of sub-blocks of the current block divided in a horizontal direction or in a vertical direction; and if the second syntax element indicates that the sub-block transform is applied to the current block, reconstructing the current block based on the sub-block transform; wherein the combined prediction block is obtained by performing a weighted-sum of an inter-prediction block and an intra-prediction block, wherein the inter-prediction block is obtained by using the inter-prediction for the current block, and the intra-prediction block is obtained by using a planar mode for the current block, wherein a weight for the weighted-sum is determined based on a prediction mode of each neighbouring locations, and wherein the neighbouring locations are (xCb-1, yCb-1+cbHeight) and (xCb-1+cbWidth, yCb-1), wherein (xCb, yCb) is a coordinate of the top-left of the current block, the cbHeight is a height of the current block and the cbWidth is a width of the current block.
Owner:WILUS INSTITUTE OF STANDARDS & TECHNOLOGY INC

Video decoding method, video coding method and device

The invention provides a video decoding method and device and a video coding method and device, and relates to the technical field of video coding and decoding. The video decoding method comprises the following steps: acquiring coded data of a current coding unit; according to the coded data, determining whether sub-block transformation is performed on the current coding unit based on an SBT mode in a preset SBT mode set, the preset SBT mode set comprising at least one SBT mode used for selecting an edge center block of the current coding unit as a non-zero residual TU, the midpoint of the target edge of the edge center block coincides with the midpoint of the target edge of the current coding unit, and the size of the target edge of the edge center block is smaller than the size of the target edge of the current coding unit; if yes, acquiring an SBT mode corresponding to the current coding unit according to the coding data; and reconstructing the current coding unit based on the SBT mode corresponding to the current coding unit. Some embodiments of the invention are used for improving the coding effect based on sub-block transformation.
Owner:HISENSE VISUAL TECH CO LTD