Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

542 results about "Video bitstream" patented technology

Cross-component sample offset edge direction derivations

An example method of video coding includes receiving a video bitstream comprising a plurality of frames, including a current frame comprising a current block. The method also includes determining an edge direction by performing edge detection for the current block, and determining a difference between a current sample of the current block and a neighboring sample based on the edge direction. The method further includes applying a filter to the current block based on the determined difference.
Owner:TENCENT AMERICA LLC

WIFI coupled low vision camera

A low vision aid system including a camera module to be used with a mobile device, including a macro camera configured to focus on near-field objects that is integrated within the low vision aid camera module and an illumination source structured to illuminate a field of view associated with the macro camera, a battery power supply and a wireless communication module. The wireless communication module establishes an isolated Wi-Fi connection to the mobile device. The isolated Wi-Fi connection prevents the mobile device from communicating over other Wi-Fi networks while the isolated Wi-Fi connection is maintained. The wireless communication module receives, via the isolated Wi-Fi connection, instructions from the mobile device, that are associated with operation of the macro camera, and sends via the isolated Wi-Fi connection, a video bitstream to the mobile device based on the instructions.
Owner:ESCHENBACH OPTIK OF AMERICA

Constrained position dependent intra prediction combination (PDPC)

A second level intra prediction mode can be combined with one or more of sixty-seven JVET intra prediction modes during encoding of a coding unit in a video bitstream. Embodiments include making a position dependent intra prediction combination (PDPC) mode available as the second level intra prediction mode. In embodiments, when a PDPC (position dependent intra prediction combination) mode is enabled, the second level intra prediction is combined with one of the 67 selected intra predictor modes. In embodiments, the PDPC mode is only enabled or available for a predetermined subset of intra prediction modes (out of 67 possible modes), in order to reduce encoder complexity and potentially improve coding efficiency. The PDPC mode may be identifies as enabled or available by a list of modes or signaling in the video bitstream.
Owner:ARRIS ENTERPRISES LLC

Cross-component sample offset (CCSO)

Various implementations described herein include methods and systems for coding video. In one aspect, a video bitstream includes a current image frame and a first syntax element for a CCSO mode. When the CCSO mode is enabled, a plurality of candidate luma sets are identified in a filter range that includes a first luma sample and neighboring luma samples. Each candidate luma set includes respective luma samples having positions symmetric with respect to a position of the first luma sample, and each luma sample located in the filter range is used in at least one of the candidate luma sets. A set of target luma samples is selected from the candidate luma sets. A loop filter is applied to combine the set of target luma samples and the first luma sample to generate the first sample offset of a first color sample collocated with the first luma sample.
Owner:TENCENT AMERICA LLC

Method, apparatus, and medium for video processing

Embodiments of the present disclosure provide a solution for video processing. A method for video processing is proposed. In the method, for a conversion between a current video block of a video and a bitstream of the video, a cross component prediction (CCP) generated prediction is determined for the current video block. The CCP generated prediction is inherited from a CCP candidate. Filtering information regarding the CCP generated prediction is determined. The filtering information comprises at least one of: whether to apply a filtering process on the CCP generated prediction, or how to apply the filtering process on the CCP generated prediction. The conversion is generated based on the filtering information.
Owner:BYTEDANCE INC

Immersive video coding method and system based on 3DGS

The invention provides an immersive video coding method and system based on 3DGS, and the method comprises the steps: carrying out the sparse reconstruction of each frame of multi-view video data, and obtaining an initial point cloud; determining anchor points of three-dimensional Gaussian distribution in each frame; extracting spatial context information of the anchor points through multi-resolution hash coding, and splicing the spatial context information with the features of the anchor points to form fusion features; predicting parameters of three-dimensional Gaussian distribution corresponding to each anchor point through a neural network by using the fusion features and camera parameters, and obtaining a rendered image; calculating the color loss between the rendered image and the original video frame, and optimizing the parameters of the anchor points; performing quantization and entropy coding on the optimized parameters of the anchor points; in combination with the color loss and the coding rate, performing rate distortion optimization on the quantization parameter to generate compressed three-dimensional scene representation; and repeating the steps for each frame of the multi-view video, and finally outputting a compressed immersive video code stream. According to the invention, high-quality representation and efficient compression of the immersive video are realized, and the rate-distortion performance is improved.
Owner:SHANGHAI JIAOTONG UNIV

Method, apparatus, and medium for video processing

Embodiments of the present disclosure provide a solution for video processing. A method for video processing comprises: determining, from a plurality of filter shapes during a conversion between a current video block of a video and a bitstream of the video, a first filter shape for coding a first sample of the current video block; and performing the conversion based on the first filter shape. Compared with the conventional solution, the proposed method can advantageously improve the performance of the filtering tool.
Owner:DOUYIN VISION CO LTD +1

Spatial resampling in video coding and decoding systems

This disclosure relates generally to video coding / decoding and particularly for spatial downsampling and / or resampling in video coding and / or decoding systems. One method includes obtaining, by a device, a coded video bitstream; determining, by the device from the coded video bitstream, a spatial resampling flag for a picture frame; and when the spatial resampling flag indicates that spatial resampling is enabled for the picture frame: determining, by the device from the coded video bitstream, an index indicating a spatial resampling filter, and decoding, by the device, the coded video bitstream by generating spatial resampling data based on the spatial resampling filter.
Owner:TENCENT AMERICA LLC

Method, apparatus, and medium for video processing

Embodiments of the present disclosure provide a method for video processing. The method comprises: determining, during a conversion between a current video block of a video and a bitstream of the video, whether to insert a target motion candidate for the current video block into a prediction list for the current video block based on a comparison between the target motion candidate and at least one existing motion candidate in the prediction list; and performing the conversion based on the determination. The proposed method can advantageously fill the prediction list more efficiently.
Owner:BYTEDANCE INC +1

Method, apparatus, and medium for video processing

Embodiments of the present disclosure provide a solution for video processing. A method for video processing is proposed. The method comprises: determining, during a conversion between a target video block of a video and a bitstream of the video, a cost metric for a target motion candidate for the target video block at least based on a matching cost of the target motion candidate; and performing the conversion based on a comparison of the cost metric and a further matching cost for the target video block. Compared with the conventional solution, the proposed method can advantageously improve the coding effectiveness and coding efficiency.
Owner:DOUYIN VISION CO LTD +1

Setting a value range of a number of sub-block based merge candidates

A method for video encoding includes determining a parameter corresponding to the coded video bitstream based on a calculated maximum number of candidates. The parameter is in a range from 0 to 5-sps_sbtmvp_enabled_flag, where the sps_sbtmvp_enabled_flag equal to 1 specifies that subblock based temporal motion vector predictors are used. The sps_sbtmvp_enabled_flag equal to 0 specifies that the subblock based temporal motion vector predictors are not used. In response to a current block being in a subblock based prediction mode, the method includes encoding samples of the current block based on a candidate selection from a constructed subblock based merge candidate list of the current block. The constructed subblock based merge candidate list of the current block is constrained by the maximum number of candidates in the subblock based merge candidate lists.
Owner:TENCENT AMERICA LLC

Temporal motion vector predictor

An example method of video coding includes receiving a video bitstream comprising a plurality of blocks, including a current block. The method also includes populating a motion vector list for the current block with one or more temporal motion vectors. At most N positions are scanned when fetching the one or more temporal motion vectors, and the N positions include at least one position outside of a block area of the current block and at least one position inside of the block area of the current block. The method further includes identifying, from the motion vector list, a motion vector predictor for the current block, and decoding the current block using the identified motion vector predictor.
Owner:TENCENT AMERICA LLC

Video decoding method, video encoding method, storage medium, electronic device and product

The present application discloses a video decoding method, a video encoding method, a storage medium, an electronic device and a product. The video decoding method comprises: obtaining a video bitstream, wherein the video bitstream comprises decoding indication information, and the decoding indication information comprises at least one filter flag bit used for indicating a filter parameter; on the basis of a value of the filter flag bit, determining tap coefficient information in the filter parameter; and on the basis of the tap coefficient information, determining the filter parameter.
Owner:TENCENT TECHNOLOGY (SHENZHEN) CO LTD

Improvements on signaling inter prediction mode

The various implementations described herein include methods and systems for coding video. In one aspect, a method includes receiving a video bitstream comprising a plurality of blocks including a current block. The method includes determining that the current block is encoded using motion information from a first reference block and a second reference block. The method includes (i) when the first and second reference blocks are in different reference frames, selecting a compound inter prediction mode for the current block from a first set of compound inter prediction modes, and (ii) when the first and second reference blocks are in a same reference frame, selecting the compound inter prediction mode for the current block from a second set of compound inter prediction modes. The method includes reconstructing the current block using the compound inter prediction mode and the motion information from the first and second reference blocks.
Owner:TENCENT AMERICA LLC

Video coding method, device and medium

The invention provides a video coding method and device and a medium. The apparatus comprises computer code for causing one or more processors to perform: obtaining a video bitstream; encoding the video bitstream at least in part through a neural network; determining topological information and parameters of the neural network; the determined topology information and parameters of the neural network are signaled in a plurality of syntax elements associated with the encoded video bitstream.
Owner:TENCENT AMERICA LLC

Method, apparatus, and medium for video processing

Embodiments of the disclosure provide a solution for video processing. A method for video processing is proposed. The method includes: deriving, for a conversion between a video unit of a video and a bitstream of the video, a cross-component prediction (CCP) model of a current block of the video unit, based on a previous CCP model coded block; and performing the conversion based on the CCP model.
Owner:DOUYIN VISION CO LTD +1

Disallowing unnecessary layers in multi-layer video bitstreams

A method of decoding is provided. The method includes receiving, by the video decoder, a video bitstream including a video parameter set (VPS) and a plurality of layers, where no layer is neither an output layer of at least one OLS nor a direct reference layer of any other layer; and decoding, by the video decoder, a picture from one of the plurality of layers. A method of encoding is also provided. The method includes generating, by the video encoder, a plurality of layers and a VPS specifying one or more output layer sets (OLSs), where no layer is neither an output layer of at least one OLS nor a direct reference layer of any other layer; encoding, by the video encoder, the plurality of layers and the VPS into a video bitstream; and storing, by the video encoder, the video bitstream for communication toward a video decoder.
Owner:HUAWEI TECH CO LTD

Bit depth shift control method and apparatus

A bit-depth shift control method and apparatus. The method includes receiving a video bitstream including an encoded video sequence and bit-depth signaling information, decoding the encoded video sequence to generate a decoded video sequence, and performing bit-depth shift processing on the decoded video sequence based on the bit-depth signaling information.
Owner:TENCENT AMERICA LLC

Interlayer prediction signaling in video bitstreams

To provide a method and an apparatus for signaling of inter layer prediction in a video bitstream.SOLUTION: A method according to the present invention causes one or more processors to perform: parsing at least one video parameter set (VPS) including at least one syntax element indicating whether at least one layer in a scalable bitstream is one of a dependent layer of the scalable bitstream and an independent layer of the scalable bitstream; determining the number of dependent layers including the dependent layer, of the scalable bitstream, on the basis of multiple flags included in the VPS; decoding a picture in the dependent layer by parsing and interpreting an inter-layer reference picture list; and decoding a picture in an independent layer without parsing and interpreting the inter-layer reference picture list.SELECTED DRAWING: Figure 3
Owner:TENCENT AMERICA LLC

Quantization parameter (QP) coding for video compression

There is provided a method (600) for decoding a current coded picture from a video bitstream. The method comprises deriving a list of delta quantization parameter, QP, values from parameter set syntax elements in the video bitstream. The method comprises deriving an index value, IV, from a slice header, a segment header or a picture header, associated with the current coded picture. The method comprises deriving a delta QP value for the current coded picture using the derived list of delta QP values and the IV. The method comprises using the derived delta QP value to derive an initial QP value, QPi, for the current coded picture. The method comprises using the initial QP value in a decoding process to decode the current coded picture or segment thereof.
Owner:TELEFONAKTIEBOLAGET LM ERICSSON (PUBL)

Method, apparatus and computer program for decoding a coded video stream

To provide an encoding and decoding method capable of changing the resampling and resolution of a reference picture.SOLUTION: The decoding method includes obtaining a sequence parameter set (SPS) network abstraction layer (NAL) unit, a picture parameter set (PPS) NAL unit, a picture header (PH) NAL unit, and a slice NAL unit from a coded video bitstream, and decoding a coded picture based thereon, wherein the SPSNAL unit is available to at least one processor before the PPSNAL unit, and wherein the PPSNAL unit is available to the at least one processor before the PHNAL unit and the at least one coded slice NAL unit.SELECTED DRAWING: Figure 21
Owner:TENCENT AMERICA LLC

Systems and methods for smooth mode predictions

The various implementations described herein include methods and systems for encoding and decoding video. In one aspect, a method of video decoding includes receiving video data that includes a first block from a video bitstream, where the first block is encoded in a smooth mode. The method further includes identifying a set of reference samples for the first block and deriving a first prediction value for the first block. The method also includes deriving a refined first prediction value for the first block using a weighted sum of a first reference sample of the set of reference samples and the first prediction value and decoding the first block based on the refined first prediction value.
Owner:TENCENT AMERICA LLC

Method, apparatus, and program for decoding a video bitstream

To provide a method for signaling a picture header in an encoded video stream.SOLUTION: A method of decoding an encoded video bitstream using at least one processor includes: obtaining a video coding layer (VCL) network abstraction layer (NAL) unit; determining whether the VCL NAL unit is a first VCL NAL unit of a picture unit (PU) containing the VCL NAL unit; based on determining that the VCL NAL unit is the first VCL NAL unit of the PU, determining whether the VCL NAL unit is a first VCL NAL unit of an access unit (AU) containing the PU; and based on determining that the VCL NAL unit is the first VCL NAL unit of the AU, decoding the AU based on the VCL NAL unit.SELECTED DRAWING: Figure 5
Owner:TENCENT AMERICA LLC

Transmission of volumetric images in multiplane imaging format

Methods and apparatus for transmission of volumetric images in the MPI format. According to an example embodiment, texture and alpha layers of multiplane images are packed, as tiles, into a sequence of video frames. The sequence of video frames is then compressed to generate a video bitstream, which is transmitted together with a metadata bitstream specifying at least the parameters of the packing arrangement for the tiles in the sequence of video frames. Example packing arrangements include various selectable spatial and temporal arrangements for texture layers, alpha layers, and camera views. In some examples, the metadata bitstream is implemented using a SEI message and includes parameters selected from the group consisting of a size of the reference view, the number of layers in the multiplane image, the number of simultaneous views, one or more characteristics of the packing arrangement, layer merging information, dynamic range adjustment information, and reference view information.
Owner:DOLBY LABORATORIES LICENSING CORP

Video decoding method, video encoding method, device, computer device and storage medium

Embodiments of the present application disclose a video decoding method, a video encoding method, a device, computer equipment and a storage medium. The video decoding method comprises: decoding encoded information of at least one block from an encoded video bitstream, the encoded information indicating whether a super-resolution encoding mode is applied to the at least one block, wherein the super-resolution encoding mode is applied in response to the at least one block being down-sampled from a high spatial resolution to a low spatial resolution by an encoder; and when the encoded information indicates that the super-resolution encoding mode is applied to the at least one block, generating a reconstructed block by using the super-resolution encoding mode to up-sample information of a first block in the at least one block, wherein the first block has the low spatial resolution, the reconstructed block has the high spatial resolution which is higher than the low spatial resolution, the at least one block comprises transform coefficients, and the reconstructed block comprises sample values in a spatial domain.
Owner:TENCENT AMERICA LLC

Encoder, a decoder and corresponding methods of chroma intra mode derivation

A method of coding implemented by a decoding device, comprising obtaining a video bitstream; decoding the video bitstream to obtain an initial intra prediction mode value for chroma component of a current coding block; determining whether a ratio between a width for luma component of the current coding block and a width for chroma component of the current coding block is equal to a threshold or not; obtaining a mapped intra prediction mode value for chroma component of the current coding block according to a predefined mapping relationship and the initial intra prediction mode value, when it's determined that the ratio is equal to the threshold; obtaining a prediction sample value for chroma component of the current coding block according to the mapped intra prediction mode value.
Owner:HUAWEI TECH CO LTD

Signaling of down-sampling information for video bitstreams

A video processing method includes determining information included in a bitstream that indicates whether down-sampling is performed on a video unit, and performing a conversion between the video unit and the bitstream based on the bitstream.
Owner:BYTEDANCE INC +1

Method and apparatus for signaling decoded data using high-level syntax elements

A method and apparatus for signaling decoded data using a high-level syntax element. A method (800, 900, 1600, 1700) and apparatus (2100) for signaling decoded data in a video bitstream in which a syntax element is used that indicates whether the signalling data is explicitly coded in the video bitstream or inferred from previous data of the video bitstream. A bitstream, a computer readable storage medium, and a computer program product are also described.
Owner:INTERDIGITAL VC HOLDINGS INC

Method and apparatus for improving performance of neural network filter based video coding

A video decoding method includes: receiving an encoded video bitstream, and decoding a first block. The encoded video bitstream includes data to be decoded as the first block of pixels in a picture, and the first block includes a luma block and at least one chroma block. Decoding the first block includes: determining whether to apply a neural network (NN) filter on the luma block and the at least one chroma block according to an NN filter mode of the luma block and at least one NN filter mode of the at least one chroma block.
Owner:MEDIATEK INC

Method, apparatus, and medium for video processing

Embodiments of the present disclosure provide a solution for video processing. A method for video processing is proposed. In the method, a conversion between a current video unit of a video and a bitstream of the video is performed. Performing the conversion comprises: applying a down-sampling filter to the current video unit to obtain an internal video unit; applying at least one neural network (NN) -based filter to the internal video unit to obtain a filtered video unit; and applying an up-sampling filter to the filtered video unit to obtain a reconstructed video unit of the current video unit.
Owner:DOUYIN VISION CO LTD +1