Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

907 results about "Video bitstream" patented technology

Method, apparatus, and medium for video processing

Embodiments of the disclosure provide a solution for video processing. A method for video processing is proposed. The method includes: determining, for a conversion between a video unit of a video and a bitstream of the video, a neural-network post-filter (NNPF) is activated for a set of pictures; apply the NNPF to one or more pictures in the set of pictures according to an order; and performing the conversion based on the NNPF.
Owner:DOUYIN VISION CO LTD +1

Method, apparatus, and medium for video processing

Embodiments of the present disclosure provide a solution for video processing. A method for video processing is proposed. In the method, for a conversion between a current video block of a video and a bitstream of the video, at least one target cross component prediction (CCP) model for the current video block is determined based on a history table of CCP models or a list of CCP candidates. A prediction of the current video block is determined based on CCP information of the at least one target CCP model. The conversion is performed based on the prediction.
Owner:BYTEDANCE INC

Method, apparatus, and medium for video processing

Embodiments of the present disclosure provide a solution for video processing. A method for video processing is proposed. The method comprises: obtaining, for a conversion between a current video block of a video and a bitstream of the video, first information regarding whether to filter a prediction for a chroma component of the current video block, wherein the prediction for the chroma component is determined with a cross-component prediction (CCP) mode, and the first information is dependent on coding information associated with the current video block; and performing the conversion based on the first information.
Owner:DOUYIN VISION CO LTD +1

Techniques of intra prediction

A coded video bitstream is received. The coded video bitstream includes coded information of a current block in a current picture, the coded information indicates that the current block is coded by an intra prediction using an affine intra mode (AIM) model. The affine intra mode model for applying on the current block is determined. At least a first intra angular mode for a first sample in the current block and a second intra angular mode for a second sample in the current block are determined according to the affine intra mode model, the first intra angular mode is different from the second intra angular mode. The current block is reconstructed based on the intra prediction using the affine intra mode model, the first sample is predicted based on the first intra angular mode and the second sample is predicted based on the second intra angular mode.
Owner:TENCENT AMERICA LLC

Film grain neural network post filter

Syntax elements allow to enable applying film grain using a neural network post-processing filter. Syntax elements define parameters related to the film grain neural network post- processing filter. An encoded video bitstream carries such syntax elements from an encoding device to a decoding device, thus allowing the encoding device to specify how to apply film grain using a neural network post-processing filter and the decoding device to apply the film grain neural network post-processing filter when displaying the decoded video. Additional syntax elements specify parameters comprising a film grain style, a film grain intensity, a film grain purpose or a region of interest where to apply the film grain.
Owner:INTERDIGITAL CE PATENT HOLDINGS SAS

Techniques for transform kernel set selection / derivation

An aspect of the disclosure provides a method of video decoding. For example, a coded video bitstream is received. The coded video bitstream includes coded information of a plurality of pictures. Based on coded information of a current block in a current picture, it is determined that the current block is coded using an intra prediction that generates prediction samples of the current block based on reference samples in the current picture. At least a first transform kernel set is determined based on a histogram of occurrence (HoC) of intra prediction modes in neighboring samples of the current block. Based on the coded information of the current block, a residual block of the current block is calculated according to at least the first transform kernel set. The current block is reconstructed based on the residual block and the intra prediction of the current block.
Owner:TENCENT AMERICA LLC

Cross-component sample offset edge direction derivations

An example method of video coding includes receiving a video bitstream comprising a plurality of frames, including a current frame comprising a current block. The method also includes determining an edge direction by performing edge detection for the current block, and determining a difference between a current sample of the current block and a neighboring sample based on the edge direction. The method further includes applying a filter to the current block based on the determined difference.
Owner:TENCENT AMERICA LLC

Intra template matching prediction at subblock level

An apparatus of video decoding is provided. The apparatus includes processing circuitry. The processing circuitry is configured to receive a video bitstream including coded information of a current block and a template of the current block in a current picture. The template of the current block includes samples adjacent to the current block, and the current block includes a first subblock and a second subblock adjacent to the first subblock. The processing circuitry is configured to determine a template of the second subblock based on at least one of (i) the template of the current block and (ii) reconstructed samples of the first subblock. The processing circuitry is configured to reconstruct the second subblock based on the determined template according to intra template matching prediction (intraTMP).
Owner:TENCENT AMERICA LLC

System and method for adaptive motion vector prediction list construction

An example method of video coding includes receiving a video bitstream including a plurality of blocks. The method further includes determining a scan order of a motion vector list of a first block of the plurality of blocks based on one or more of: a number of neighboring blocks of the current block having a corresponding temporal motion vector, a number of neighboring blocks of the current block encoded in an inter prediction mode, a mode of the current block, and a reference frame index of the current block. The method further includes generating a motion vector list according to the scan order, and identifying a motion vector predictor of the current block from the motion vector list. The method also includes decoding the current block using the identified motion vector predictor.
Owner:TENCENT AMERICA LLC

WIFI coupled low vision camera

A low vision aid system including a camera module to be used with a mobile device, including a macro camera configured to focus on near-field objects that is integrated within the low vision aid camera module and an illumination source structured to illuminate a field of view associated with the macro camera, a battery power supply and a wireless communication module. The wireless communication module establishes an isolated Wi-Fi connection to the mobile device. The isolated Wi-Fi connection prevents the mobile device from communicating over other Wi-Fi networks while the isolated Wi-Fi connection is maintained. The wireless communication module receives, via the isolated Wi-Fi connection, instructions from the mobile device, that are associated with operation of the macro camera, and sends via the isolated Wi-Fi connection, a video bitstream to the mobile device based on the instructions.
Owner:ESCHENBACH OPTIK OF AMERICA

Method and apparatus for adaptive multi-hypothesis probability model for arithmetic coding

A method performed by at least one processor of a video decoder includes receiving a coded video bitstream including at least one picture and one or more syntax elements encoded in accordance with multi-hypothesis arithmetic coding. The method further includes decoding each syntax element from the one or more syntax elements based on the multi-hypothesis arithmetic coding. The method further includes selecting a probability update rate from a plurality of probability update rates based on a predetermined condition, the plurality of probability update rates including a first probability update rate that is higher than a second probability update rate. The method further includes updating at least one probability model utilized in the multi-hypothesis arithmetic coding based on the selected probability update rate. The method further includes decoding at least one block in the at least one picture based on the decoded one or more syntax elements.
Owner:TENCENT AMERICA LLC

Constrained position dependent intra prediction combination (PDPC)

A second level intra prediction mode can be combined with one or more of sixty-seven JVET intra prediction modes during encoding of a coding unit in a video bitstream. Embodiments include making a position dependent intra prediction combination (PDPC) mode available as the second level intra prediction mode. In embodiments, when a PDPC (position dependent intra prediction combination) mode is enabled, the second level intra prediction is combined with one of the 67 selected intra predictor modes. In embodiments, the PDPC mode is only enabled or available for a predetermined subset of intra prediction modes (out of 67 possible modes), in order to reduce encoder complexity and potentially improve coding efficiency. The PDPC mode may be identifies as enabled or available by a list of modes or signaling in the video bitstream.
Owner:ARRIS ENTERPRISES LLC

Method and apparatus for video coding

Aspects of the disclosure provide methods and apparatuses for video encoding / decoding. In some examples, an apparatus for video decoding includes processing circuitry. For example, the processing circuitry decodes prediction information of a current block from a coded video bitstream. The prediction information indicates that a prediction of the current block is at least partially based on an inter prediction. Then, the processing circuitry reconstructs at least a sample of the current block as a combination of a result from the inter prediction and neighboring samples of the block that are selected based on a position of the sample.
Owner:TENCENT AMERICA LLC

Method, apparatus, and medium for video processing

Embodiments of the disclosure provide a solution for video processing. A method for video processing is proposed. The method includes: generating, for a conversion between a video unit of a video and a bitstream of the video, a prediction value of the video unit based on a cross-component prediction candidate; modifying the prediction value of the video unit; obtaining a reconstructed sample value based on the modified prediction value; and performing the conversion based on the reconstructed sample value.
Owner:BYTEDANCE INC

Cross-component sample offset (CCSO)

Various implementations described herein include methods and systems for coding video. In one aspect, a video bitstream includes a current image frame and a first syntax element for a CCSO mode. When the CCSO mode is enabled, a plurality of candidate luma sets are identified in a filter range that includes a first luma sample and neighboring luma samples. Each candidate luma set includes respective luma samples having positions symmetric with respect to a position of the first luma sample, and each luma sample located in the filter range is used in at least one of the candidate luma sets. A set of target luma samples is selected from the candidate luma sets. A loop filter is applied to combine the set of target luma samples and the first luma sample to generate the first sample offset of a first color sample collocated with the first luma sample.
Owner:TENCENT AMERICA LLC

Motion vector derivation of subblock-based template-matching for subblock based motion vector predictor

A video bitstream is received. The video bitstream includes a current block comprising a plurality of subblocks and a template region of the current block comprising a plurality of template subblocks adjacent to at least one of a top side and a left side of the current block. A motion vector (MV) located in a center position of the current block is determined. The MV is determined based on at least one MV of the plurality of subblocks of the current block. A MV for each of the plurality of template subblocks is determined based on the MV located in the center position of the current block and a respective MV of a corresponding subblock of the plurality of subblocks that is adjacent to the respective template subblock. The current block is reconstructed based on the determined MVs for the plurality of template subblocks.
Owner:TENCENT AMERICA LLC

Method for signaling of reference picture resampling with resampling picture size indication in video bitstream

A method, device, and computer-readable medium for decoding an encoded video bitstream using at least one processor, including obtaining a flag indicating that a conformance window is not used for reference picture resampling; based on the flag indicating that the conformance window is not used for the reference picture resampling, determining whether a resampling picture size is signaled; based on determining that the resampling picture size is signaled, determining a resampling ratio based on the resampling picture size; based on determining that the resampling picture size is not signaled, determining the resampling ratio based on an output picture size; and performing the reference picture resampling on a current picture using the resampling ratio.
Owner:TENCENT AMERICA LLC

Method, apparatus, and medium for video processing

Embodiments of the disclosure provide a solution for video processing. A method for video processing is proposed. The method includes: determining, for a conversion between a video unit of a video and a bitstream of the video, whether to apply a combination of intra block copy (IBC) and intra prediction (CIBCIP) to the video unit based on at least one of: coding information, whether an indication of CIBCIP mode is indicated, or one or more syntax elements; deriving a prediction of the video unit by combining an IBC predicted signal and an intra predicted signal; and performing the conversion based on the prediction of the video unit.
Owner:DOUYIN VISION CO LTD +1

Method, apparatus, and medium for video processing

Embodiments of the present disclosure provide a solution for video processing. A method for video processing is proposed. In the method, for a conversion between a current video block of a video and a bitstream of the video, a cross component prediction (CCP) generated prediction is determined for the current video block. The CCP generated prediction is inherited from a CCP candidate. Filtering information regarding the CCP generated prediction is determined. The filtering information comprises at least one of: whether to apply a filtering process on the CCP generated prediction, or how to apply the filtering process on the CCP generated prediction. The conversion is generated based on the filtering information.
Owner:BYTEDANCE INC

Immersive video coding method and system based on 3DGS

The invention provides an immersive video coding method and system based on 3DGS, and the method comprises the steps: carrying out the sparse reconstruction of each frame of multi-view video data, and obtaining an initial point cloud; determining anchor points of three-dimensional Gaussian distribution in each frame; extracting spatial context information of the anchor points through multi-resolution hash coding, and splicing the spatial context information with the features of the anchor points to form fusion features; predicting parameters of three-dimensional Gaussian distribution corresponding to each anchor point through a neural network by using the fusion features and camera parameters, and obtaining a rendered image; calculating the color loss between the rendered image and the original video frame, and optimizing the parameters of the anchor points; performing quantization and entropy coding on the optimized parameters of the anchor points; in combination with the color loss and the coding rate, performing rate distortion optimization on the quantization parameter to generate compressed three-dimensional scene representation; and repeating the steps for each frame of the multi-view video, and finally outputting a compressed immersive video code stream. According to the invention, high-quality representation and efficient compression of the immersive video are realized, and the rate-distortion performance is improved.
Owner:SHANGHAI JIAOTONG UNIV

Handling different NAL types in video sub-bitstream extraction

Examples of video encoding methods and apparatus and video decoding methods and apparatus are described. An example method of video processing includes performing a conversion between a video and a bitstream of the video. The bitstream includes network abstraction layer (NAL) units for multiple video layers according to a rule. The rule defines a sub-bitstream extraction process by which NAL units are removed from the bitstream to generate an output bitstream, and specifies to remove all supplemental enhancement information (SEI) NAL units that contain a non-scalable-nested SEI message with a particular payload type, responsive to a list of NAL unit header layer identifier values in an output layer set (OLS) with a target OLS index not including all values of NAL unit header layer identifiers in all video coding layer (VCL) NAL units in the bitstream that is input to the sub-bitstream extraction process.
Owner:BYTEDANCE INC

Method, apparatus, and medium for video processing

Embodiments of the present disclosure provide a solution for video processing. A method for video processing comprises: determining, from a plurality of filter shapes during a conversion between a current video block of a video and a bitstream of the video, a first filter shape for coding a first sample of the current video block; and performing the conversion based on the first filter shape. Compared with the conventional solution, the proposed method can advantageously improve the performance of the filtering tool.
Owner:DOUYIN VISION CO LTD +1

Intra predictor and intra mode coding

A video bitstream including coded information of a current block in a current picture is received. The coded information indicates a plurality of candidate intra prediction modes for the current block. Two or more predictors are determined based on the plurality of candidate intra prediction modes for the current block according to a pre-defined condition. At least one of the two or more predictors is generated based on a matrix-multiplication mode of the plurality of candidate intra prediction modes such that the at least one of the two or more predictors is obtained by a matrix multiplication of a matrix of weight coefficients and neighboring reconstructed samples in a template of the current block. The current block is reconstructed based on a weighted combination of the two or more predictors.
Owner:TENCENT AMERICA LLC

Method, apparatus, and medium for video processing

Embodiments of the present disclosure provide a solution for video processing. A method for video processing is proposed. In the method, for a conversion between a current video block of a video and a bitstream of the video, a base candidate of the current video block is determined. Whether to inherit a flip type for the base candidate is based on a candidate type of the base candidate. A target candidate of the current video block is determined based on the base candidate. The target candidate includes at least one of: an intra block copy (IBC) merge mode with block vector differences (IBC-MBVD) candidate, an IBC template matching (IBC-TM) merge candidate, or an IBC-TM advanced motion vector prediction (AMVP) candidate. The conversion is performed based on the target candidate.
Owner:DOUYIN VISION CO LTD +1

Cross-component residual prediction by using residual template

The various implementations described herein include methods and systems for coding video. In one aspect, a video bitstream includes a current image frame having a current coding block and signals a first syntax element for a residual template cross-component residual model (RT-CCRM) mode. When the RT-CCRM mode is enabled, the computing system identifies, in the current coding block, a first chroma sample and one or more luma samples corresponding to the first chroma sample, determines one or more residuals of the one or more luma samples in the current coding block, and applies a residual filter corresponding to the RT-CCRM mode to generate a first residual of the first chroma sample based on the residuals of the one or more luma samples. The computing system reconstructs the current image frame by compensating a predicted chroma sample with at least the first residual to reconstruct the first chroma sample.
Owner:TENCENT AMERICA LLC

Intra template matching prediction signaling

An apparatus for video decoding is provided. The apparatus includes processing circuitry. The processing circuitry is configured to receive a video bitstream including coded information associated with a current block that includes a luma component and a chroma component. The coded information indicates whether intra template matching prediction (intraTMP) is applied to the current block. When the coded information indicates that the intraTMP is applied to the current block, the processing circuitry is configured to determine (i) a BV for the luma component of the current block based on the intraTMP and (ii) a BV for the chroma component of the current block based on the determined BV for the luma component of the current block. The processing circuitry is configured to reconstruct the current block based on the determined BV for the luma component and the determined BV for the chroma component.
Owner:TENCENT AMERICA LLC

Supporting mixed IRAP and non-IRAP pictures within an access unit in multi-layer video bitstreams

A decoding method implemented by a video decoder is provided, the method comprising: receiving a bitstream including Coded Video Sequence Start (CVSS) access units (AUs), the CVSS AUs including picture units (PUs) for each layer, and a coded picture in each PU being a Coded Layer Video Sequence Start (CLVSS) picture; identifying a coded picture from one of the layers based on a picture order count (POC) value; and decoding the coded picture to obtain a decoded picture.
Owner:HUAWEI TECH CO LTD

Video decoder, video encoder, method for decoding video content, method for encoding video content, computer program, and video bitstream

To provide a video decoder, a video encoder, a method for decoding video content, a method for encoding video content, a computer program, and a video bitstream for implementing arithmetic encoding and decoding with optimal coding efficiency.SOLUTION: A video decoder comprises an arithmetic decoder for providing a decoded binary sequence on the basis of an encoded representation of a binary sequence. The arithmetic decoder is configured to determine a first source statistic value using a first estimation parameter and to determine a second source statistic value using a second estimation parameter. The arithmetic decoder is configured to determine a combined source statistic value on the basis of the first source statistic value (at) and on the basis of the second source statistic value. The arithmetic decoder is configured to determine one or more range values for an interval subdivision, which is used for mapping the encoded representation of the binary sequence onto the decoded binary sequence, on the basis of the combined source statistic value.SELECTED DRAWING: Figure 1
Owner:FRAUNHOFER GESELLSCHAFT ZUR FORDERUNG DER ANGEWANDTEN FORSCHUNG EV

Spatial resampling in video coding and decoding systems

This disclosure relates generally to video coding / decoding and particularly for spatial downsampling and / or resampling in video coding and / or decoding systems. One method includes obtaining, by a device, a coded video bitstream; determining, by the device from the coded video bitstream, a spatial resampling flag for a picture frame; and when the spatial resampling flag indicates that spatial resampling is enabled for the picture frame: determining, by the device from the coded video bitstream, an index indicating a spatial resampling filter, and decoding, by the device, the coded video bitstream by generating spatial resampling data based on the spatial resampling filter.
Owner:TENCENT AMERICA LLC

Method, apparatus, and medium for video processing

Embodiments of the present disclosure provide a method for video processing. The method comprises: determining, during a conversion between a current video block of a video and a bitstream of the video, whether to insert a target motion candidate for the current video block into a prediction list for the current video block based on a comparison between the target motion candidate and at least one existing motion candidate in the prediction list; and performing the conversion based on the determination. The proposed method can advantageously fill the prediction list more efficiently.
Owner:BYTEDANCE INC +1