Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

112 results about "Coding tree unit" patented technology

Coding tree unit (CTU) is the basic processing unit of the High Efficiency Video Coding (HEVC) video standard and conceptually corresponds in structure to macroblock units that were used in several previous video standards. CTU is also referred to as largest coding unit (LCU).

Encoder, decoder, and medium

An encoder includes circuitry and memory coupled to the circuitry. In operation, the circuitry encodes subpicture information in which a horizontal position and a vertical position of a region of a subpicture are represented in a unit of a coding tree unit (CTU). The subpicture is a rectangular region in a picture. The horizontal position is represented by a first position of a first CTU in the subpicture, the first position being relative to a left end of the picture. The vertical position is represented by a second position of a second CTU in the subpicture, the second position being relative to a top end of the picture.
Owner:PANASONIC INTELLECTUAL PROPERTY CORP OF AMERICA

Target detection enhanced video code rate control method

The invention relates to a target detection enhanced video code rate control method, and belongs to the technical field of video compression. The method comprises the following steps: preprocessing each frame of image of an input video, dividing a plurality of coding tree units in each frame of image, and calculating gradient features of the image and the coding tree units; the video is coded through different quantization parameters, and the code rate of the coded video and the distortion degree of the video are calculated; fitting a cubic logarithmic model to establish a relationship between the distortion degree and the code rate of the video by combining the gradient characteristics of each frame and the distortion degree and the code rate of the video; marking the ROI region of each image frame of the input video through a target detection network, and dividing the compensation level of the coding tree unit in combination with the ROI region and the gradient feature of the coding tree unit; and in combination with the compensation level of each coding tree unit, through a code rate allocation mechanism, allocating a coding code rate to each coding tree unit in each image frame of the input video. According to the invention, the video compression process can be optimized, and the overall performance of the compressed video is improved.
Owner:CHONGQING UNIV OF POSTS & TELECOMM

Zero-delay panoramic video bit rate control method considering temporal distortion propagation

A zero-delay panoramic video bit rate control method that considers distortion temporal propagation aims to optimize target bit allocation and achieve global rate-distortion optimization in coding. It encompasses coding tree unit (CTU)-level bit rate control and temporal global rate-distortion optimization. CTU-level bit rate control involves optimizing target bit allocation and updating bit rate control parameters. Temporal global rate-distortion optimization utilizes the reconstruction error and motion compensation prediction error information of the previous coded frame to estimate the temporal dependence between CTUs in the current coding frame and those in the previous coded frame to adjust the coding parameters of the current CTU. The coding parameters are further fine-tuned according to the area stretching ratio encountered during the projection of the panoramic video from a 3D spherical surface to a 2D plane, taking into consideration the detrimental impact of interpolated redundant pixels on the coding process.
Owner:UNIV OF ELECTRONICS SCI & TECH OF CHINA

A method and system for fast decision of coding unit partitioning based on vvc

The present application relates to the technical field of video coding, and especially relates to a coding unit division fast decision method and system based on multi-functional video coding, comprising constructing a quadtree division prediction network, using the network to judge whether a coding tree unit needs to perform quadtree division, and obtaining a coding unit after performing the quadtree division; constructing a mixed tree division prediction network, using the network to judge whether the coding unit needs to be divided, stopping division if the coding unit does not need to be divided, otherwise judging whether the coding unit adopts binary tree division or ternary tree division, and taking a smaller coding unit division prediction value in parallel branches as a final prediction result. The present application improves coding efficiency, and can balance the relationship between visual quality and coding efficiency by selecting different threshold values according to specific scene needs.
Owner:CHONGQING UNIV OF POSTS & TELECOMM

Video coding method

The invention provides a video coding method, which comprises the following steps of: inputting a coding tree unit to be coded, and extracting texture features of a coding block, the texture features comprising local binary pattern similarity of all pixel points in the coding block and direction dispersion of each pixel point; inputting the texture features into the trained multi-type division prediction architecture, and sequentially judging whether to execute quadtree division, horizontal or vertical division, horizontal binary tree or ternary tree division and vertical binary tree or ternary tree division or not through a binary classification problem method; and calculating the rate distortion cost under each division mode to obtain an optimal division mode and a suboptimal division mode of the coding block, and outputting an optimal mode or a combination of the optimal mode and the suboptimal mode. According to the invention, the accuracy of the prediction mode in coding is improved.
Owner:SHANGHAI FULLHAN MICROELECTRONICS

Image processing device, image processing method, and bitstream transmission device

The invention provides an image processing apparatus, an image processing method, and a bitstream transmission apparatus. An image processing device for processing a picture including a sub-picture including one or more CTUs (Code Tree Units), the image processing device comprising: a memory; and a circuit connected to the memory, the circuit performing a process of determining whether (i) a third value, which is a product of a CTU size and a sum of a first value indicating a left end of the sub-picture and a second value indicating a width of the sub-picture, and (ii) the width of the picture is smaller, and if the third value is smaller than the width of the picture, determining whether the width of the picture is smaller or not, and if the third value is smaller than the width of the picture, the width of the sub-picture is smaller than the width of the sub-picture, the width of the sub-picture is smaller than the width of the sub-picture. When the width of the picture is smaller than the third value, the product of the second value and the CTU size is derived as the width of the sub-picture, and when the width of the picture is smaller than the third value, the difference between the width of the picture and the product of the first value and the CTU size is derived as the width of the sub-picture.
Owner:PANASONIC INTELLECTUAL PROPERTY CORP OF AMERICA

History-based motion vector prediction

PendingUS20260254970A1Video bitstreamCoding block
A method of video decoding is provided. In the method, a coded video bitstream of a current picture is received. The current picture is divided into a plurality of rows of coding tree units (CTUs). Coding blocks in a first CTU row of the plurality of rows of CTUs are decoded based on motion vector information stored in a history-based motion vector prediction (HMVP) buffer. The motion vector information stored in the HMVP buffer is updated as the coding blocks in the first CTU row are decoded. The HMVP buffer is updated to include the motion vector information of the coding blocks in a first CTU of the first CTU row after the coding block at an end of the first CTU row is decoded.
Owner:TENCENT AMERICA LLC

Method, apparatus and system for encoding and decoding a block of video samples

Decoding an image frame from a bitstream, the image frame being divided into a plurality of coding tree units. The method comprises decoding a maximum transform block size constraint and or a maximum coding tree unit (CTU) size constraint from the bitstream; and decoding a maximum enabled transform block size and / or a maximum enabled CTU size from the bitstream. The decoded maximum enabled transform block size is less than or equal to the decoded maximum transform block size constraint. The decoded maximum enabled CTU size is less than or equal to the decoded maximum CTU size constraint. Determining each of the one or more transform blocks for each of the plurality of coding tree units according to the decoded maximum enabled transform block size, maximum enabled CTU and split flags decoded from the bitstream; and decode each of the determined one or more transform blocks from the bitstream.
Owner:CANON KK

Method for decoding video from video bitstream, method for encoding video, video decoder, and video encoder

A method for decoding a video from a video bitstream is provided and includes: accessing a binary string representing a partition of the video, the partition comprising a plurality of coding tree units (CTUs) forming one or more CTU rows; for each CTU of the plurality of CTUs in the partition, determining whether the CTU is the first CTU in a slice or a tile; in response to determining that the CTU is not the first CTU in a slice or a tile, determining whether parallel decoding is enabled and the CTU is the first CTU in a CTU row of a tile; in response to determining that the parallel decoding is enabled and the CTU is the first CTU in a CTU row of a tile, determining an available flag for a top neighboring block of the CTU.
Owner:GUANGDONG OPPO MOBILE TELECOMMUNICATIONS CORP LTD

List structure improvement

This disclosure generally describes embodiments relating to video coding. [Solution] The neighboring blocks of the current block include a first block adjacent to one of the top edge, upper left corner, and upper right corner of the current block, and a second block adjacent to one of the left edge and lower left corner of the current block. It is determined whether one or more of the first blocks and the current block are in the same coding tree unit (CTU). Based on whether one or more of the first blocks and the current block are in the same CTU, each intra-mode associated with one or more of the first blocks is added to the most likely mode (MPM) list of the current block based on the sequence of conditions. Each intra-mode associated with each of the second blocks is added to the MPM list based on the sequence of conditions.
Owner:TENCENT AMERICA LLC

Adaptive prediction cost estimation for video encoding

Systems and methods for adaptive prediction cost estimation in video encoding are provided. The techniques improve early cost estimation and reduce the number of candidates for the later decision stages and final RDO stage. In particular, an adaptive sum of absolute transformed differences (SATD) is determined for each candidate, and, based on the adaptive SATD values, a subset of candidates is selected for mode decision search to determine block partitioning, motion vectors, and encoding modes, The adaptive SATD combines a weighted DC component of the SATD and the AC component of the SATD. The weighting factor is selected from a DC adjustment ratio table based on the spatial variation and the QP for a respective coding tree unit. The techniques improve cost estimation accuracy, reduce encoding complexity, and are hardware-friendly for integration into video codecs such as HEVC, AV1, VVC, and AV2.
Owner:INTEL CORP

Method, apparatus and system for encoding and decoding a block of video samples

Decoding an image frame from a bitstream, the image frame being divided into a plurality of coding tree units. The method comprises decoding a maximum transform block size constraint and or a maximum coding tree unit (CTU) size constraint from the bitstream; and decoding a maximum enabled transform block size and / or a maximum enabled CTU size from the bitstream. The decoded maximum enabled transform block size is less than or equal to the decoded maximum transform block size constraint. The decoded maximum enabled CTU size is less than or equal to the decoded maximum CTU size constraint. Determining each of the one or more transform blocks for each of the plurality of coding tree units according to the decoded maximum enabled transform block size, maximum enabled CTU and split flags decoded from the bitstream; and decode each of the determined one or more transform blocks from the bitstream.
Owner:CANON KK

Video encoding method and apparatus, and video decoding method and apparatus

Disclosed are a video decoding method and apparatus, and a video encoding method and apparatus. The video decoding method according to the present disclosure acquires division information relating to a division structure of a coding tree unit from a bitstream, divides the coding tree unit into at least one coding unit based on the division information, and performs decoding based on the at least one coding unit. The segmentation structure of the coding tree unit comprises a second coding unit obtained by segmenting a first coding unit in coding units segmented from the coding tree unit according to levels, the first coding unit is a coding unit obtained by performing quadtree segmentation on a coding unit of an upper layer depth or one of coding units obtained by performing multi-type segmentation on the coding unit of the upper layer depth, and if the first coding unit is one of the coding units obtained by performing multi-type segmentation on the coding unit of the upper layer depth; the second coding unit can comprise a coding unit obtained by performing quadtree segmentation on the first coding unit.
Owner:INTELLECTUAL DISCOVERY CO LTD

History-based rice parameter derivations for wavefront parallel processing in video coding

In some embodiments, a video decoder decodes a video from a bitstream of the video using a history-based Rice parameter derivation along with the wavefront parallel processing (WPP). The video decoder accesses a binary string representing a partition of the video and processes each coding tree unit (CTU) in the partition to generate decoded coefficient values in the CTU. The process includes prior to decoding the CTU, determining whether WPP is enabled and the CTU is the first CTU of a current CTU row in the partition, and if so, setting a history counter to an initial value. The process further includes decoding the CTU by calculating the Rice parameters for transform units (TUs) in the CTU based on the value of the history counter and decoding the binary string corresponding to the TUs in the CTU into coefficient values of the TUs based on the calculated Rice parameters.
Owner:GUANGDONG OPPO MOBILE TELECOMMUNICATIONS CORP LTD

Method of decoding an image frame from a bitstream, video decoding device, storage medium and computer program product

PendingCN122476196AAlgorithmCoding tree unit
The present disclosure relates to a method of decoding an image frame from a bitstream, a video decoding device, a storage medium, and a computer program product. The image frame is partitioned into a plurality of coding tree units. The method includes decoding a maximum transform block size constraint and / or a maximum coding tree unit (CTU) size constraint from the bitstream; and decoding a maximum enabled transform block size and / or a maximum enabled CTU size from the bitstream. The decoded maximum enabled transform block size is less than or equal to the decoded maximum transform block size constraint. The decoded maximum enabled CTU size is less than or equal to the decoded maximum CTU size constraint. The method operates to determine, from the decoded maximum enabled transform block size, the maximum enabled CTU, and a split flag decoded from the bitstream, each transform block of one or more transform blocks of each coding tree unit of the plurality of coding tree units; and decode each transform block of the determined one or more transform blocks from the bitstream to decode the image frame.
Owner:CANON KK

Video stream processing method, electronic device, storage medium and program product

The embodiment of the invention relates to the technical field of video coding and decoding, and provides a video stream processing method, an electronic device, a storage medium and a program product, and the method comprises the steps: determining an APS subset used by a coding tree unit in a video image for adaptive loop filtering, writing filtering parameter information in a coded video stream according to the APS subset, the filtering parameter information includes signaling for determining whether the coding tree unit uses a default subset of APS indicated by the high-level signaling. According to the embodiment of the invention, whether each CTU uses the default APS subset or not is indicated through the signaling, and the default APS subset indication information is carried through the high-level signaling, so that repeated writing of the default APS subset indication information in a plurality of CTUs can be reduced, redundant information in a video coding stream is reduced, and the coding or decoding efficiency is improved.
Owner:ZTE CORP

Video decoding method, encoding method, decoding apparatus, and encoding apparatus

In some embodiments, a video decoder uses history-based rice parameter derivation and wavefront parallel processing (WPP) to decode a video from a bitstream of the video. The video decoder accesses a bin string representing a partition of the video and processes each coding tree unit (CTU) in the partition to generate decoded coefficient values in the CTU. The process includes determining, before decoding the CTU, whether the CTU is the first CTU of a current CTU row in the partition, and if so, setting a history counter to an initial value. The process further includes decoding the CTU by calculating a rice parameter for a transform unit (TU) in the CTU based on a value of the history counter, and decoding the bin string corresponding to the TU in the CTU into coefficient values of the TU based on the calculated rice parameter.
Owner:GUANGDONG OPPO MOBILE TELECOMMUNICATIONS CORP LTD

History-based image coding method and apparatus

To provide a method and apparatus for improving image coding efficiency.SOLUTION: The method includes deriving a history-basedMotionVectorPrediction (HMVP) buffer for a current block, constructing a motion information candidate list based on HMVP candidates included in the HMVP buffer, deriving motion information of the current block based on the motion information candidate list, deriving a reference picture index of the current block based on the motion information, deriving a motion vector of the current block based on the motion information, generating prediction samples for the current block based on the reference picture index and the motion vector, and generating reconstructed samples based on the prediction samples. The current picture includes one or more tiles and includes a plurality of tile columns and tile rows, and a tile is a rectangular region of coding tree units (CTUs) within a specific tile column and a specific tile row in the current picture.SELECTED DRAWING: Figure 27
Owner:LG ELECTRONICS INC

Method for decoding video from video bitstream, method for encoding video, video decoder, and video encoder

A method for decoding a video from a video bitstream is provided and includes: accessing a binary string representing a partition of the video, the partition comprising a plurality of coding tree units (CTUs) forming one or more CTU rows; for each CTU of the plurality of CTUs in the partition, determining whether the CTU is the first CTU in a slice or a tile; in response to determining that the CTU is the first CTU in a slice or a tile, initializing context variables for context-adaptive binary arithmetic coding (CABAC) according to a first context variable initialization process; in response to determining that the CTU is not the first CTU in a slice or a tile, determining whether parallel decoding is enabled and the CTU is the first CTU in a CTU row of a tile.
Owner:GUANGDONG OPPO MOBILE TELECOMMUNICATIONS CORP LTD

Independent history-based Rice parameter derivations for video coding

In some embodiments, a video decoder decodes a video from a bitstream of the video using a history-based rice parameter derivation. The video decoder accesses a binary string representing a partition of the video and processes each coding tree unit (CTU) in the partition to generate decoded coefficient values in the CTU. The process includes updating a replacement variable for a transform unit (TU) in the CTU for calculating rice parameters independently of the previous TU or CTU. The process further includes calculating the rice parameters for TU in the CTU based on the value of the replacement variable and decoding the binary string corresponding to the TU into coefficient values based on the calculated rice parameters. Pixel values of the TU can be determined from the decoded coefficient values for output.
Owner:GUANGDONG OPPO MOBILE TELECOMMUNICATIONS CORP LTD

Initialization process for video encoding

ActiveCN121151567BPredictor variableAlgorithm
In some embodiments, a video decoder decodes a video from a bitstream. The video decoder accesses a bin string representing a partition of the video and processes each coding tree unit (CTU) in the partition to generate decoded values in the CTU. The process includes initializing context variables for context adaptive binary arithmetic coding (CABAC), Rice parameter variables, and palette predictor variables only when the CTU is the first CTU in a tile, or the CTU is the first CTU in a slice, or parallel coding is enabled and the CTU is the first CTU in a CTU row of a tile. No other initialization is performed for these variables. The video decoder decodes the CTU based on the initialized context variables, Rice parameter variables, and palette predictor variables.
Owner:GUANGDONG OPPO MOBILE TELECOMMUNICATIONS CORP LTD

Operation method of multi-codec circuit

A method of operating a multi-codec circuit configured to support a plurality of codecs, may include selecting a current coding tree unit (CTU) from a plurality of CTUs partitioned from an input image, selecting a current coding unit (CU) from a plurality of CUs partitioned from the current CTU, determining availability of neighbor CUs of the current coding unit based on a coding unit bitmap, and performing signal processing on the current CU using at least one of the neighbor CUs that is determined to be available. The coding unit bitmap may include a plurality of bits, each indicating whether signal processing has been performed for a corresponding coding unit, among the plurality of CUs. Each of the plurality of bits may correspond to a minimum CU supported by each of the plurality of codecs.
Owner:SAMSUNG ELECTRONICS CO LTD

Video encryption method, device, equipment, medium and product

The invention provides a video encryption method and device, equipment, a medium and a product. The method comprises the following steps: acquiring a to-be-encrypted first video; initializing a key stream generator using the first key and the first initial vector; and generating a first key stream by using a key stream generator, and encrypting the brightness discrete cosine transform coefficient and the chroma discrete cosine transform coefficient in each coding tree unit in each video frame in the first video through the first key stream to obtain an encrypted first video. In the invention, the brightness discrete cosine transform coefficient and the chroma discrete cosine transform coefficient in the video coding are selectively encrypted, so that the video content can be effectively confused, the coding format is not damaged, the encrypted video can be normally decoded by a standard decoder, the calculation burden of equipment is obviously reduced, and the security of the encrypted video is also improved; and the encrypted video is still compatible with the standard encoder, the original encoding format is not damaged, and normal decoding and playing of the video are ensured.
Owner:CHINA MOBILE (JIANGXI) VIRTUAL REALITY TECH CO LTD +3

A video processing method, apparatus and medium

Methods, systems, and devices are described that use intra block copy (IBC) with non-adjacent neighboring blocks. One example method of video processing includes, for a conversion between a video comprising a current video block and a bitstream of the video, determining, based on a coding tree unit (CTU) row comprising a non-adjacent neighboring block and the current video block, whether to use a block vector of the non-adjacent neighboring block to predict a block vector of the current video block, and performing the conversion based on the determination.
Owner:DOUYIN VISION CO LTD

An airport monitoring video transmission method and device, a storage medium and an electronic device

The application provides an airport monitoring video transmission method and device, a storage medium and electronic equipment, comprehensive complexity corresponding to each coding tree unit in each frame of image in an image group is acquired, and based on a preset condition, a current frame of image can be accurately divided into a key area and a non-key area; based on the comprehensive complexity of the kth merging area in the current frame of image and the total bit budget of the current frame of image, the number of bits allocated to the kth merging area in the current frame of image is determined; based on the number of bits allocated to the kth merging area, the pixel points in the kth merging area are encoded; after all the frames of image in the image group are encoded, the encoded data is transmitted. According to the dynamic bit allocation based on the comprehensive complexity of the key area and the non-key area, the coding accuracy of the key area is improved, and the definition of the key area and the overall monitoring efficiency are significantly improved.
Owner:HANGZHOU XIAOSHAN INT AIRPORT +2

Method, apparatus and system for encoding and decoding a block of video samples

Decoding an image frame from a bitstream, the image frame being divided into a plurality of coding tree units. The method comprises decoding a maximum transform block size constraint and or a maximum coding tree unit (CTU) size constraint from the bitstream; and decoding a maximum enabled transform block size and / or a maximum enabled CTU size from the bitstream. The decoded maximum enabled transform block size is less than or equal to the decoded maximum transform block size constraint. The decoded maximum enabled CTU size is less than or equal to the decoded maximum CTU size constraint. Determining each of the one or more transform blocks for each of the plurality of coding tree units according to the decoded maximum enabled transform block size, maximum enabled CTU and split flags decoded from the bitstream; and decode each of the determined one or more transform blocks from the bitstream.
Owner:CANON KK

Fast h.266 / VVC-based intra coding unit (CU) partitioning method for screen content based on multi-task learning and device

An H.266 / VVC-based intra coding unit partitioning method for screen content based on multi-task learning and a device, the method includes: partitioning a 128×128 coding tree unit into 64×64 coding units, a multi-task learning network model comprises a trunk network configured to extract CU features, a first sub-network, and a second sub-network, inputting the CU features into the first sub-network and the second sub-network to predict a CU partitioning type and a coding mode, determining the predicted result in combination with the coding mode, a corresponding predicted probability of the coding mode, and a partitioning type of an adjacent CU, inputting the 64×64 CUs into the model to obtain a first predicted result, partitioning each of the 64×64 CUs into four 32×32 CUs in response to determining that the first predicted result is partition, inputting the four 32×32 CUs into the model to obtain a second predicted result.
Owner:HUAQIAO UNIVERSITY

A vvc inter coding rate control method based on linear model

ActiveCN121397223BAlgorithmVideo encoding
This invention relates to a VVC inter-frame coding rate control method based on a linear model, belonging to the field of video coding technology. This invention combines convex optimization techniques and rate control to achieve a significantly improved and enhanced rate control scheme. It transforms the original bitrate-distortion relationship into a linear model, converting the problem into a convex optimization problem for optimal solution. Using the improved rate control scheme, quantization parameters are calculated at both the frame level and the coding tree unit level, achieving control over both bitrate and quality. This invention achieves bitrate and quality control by performing frame-level and coding tree unit-level bitrate control on each image group of the initial video. Extensive experiments demonstrate that the algorithm provides significantly improved and enhanced rate control coding performance, illustrating the effectiveness of this method.
Owner:BEIJING INST OF COMP TECH & APPL