Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

798 results about "Filter (video)" patented technology

A video filter is a software component that performs some operation on a multimedia stream. Multiple filters can be used in a chain, known as a filter graph, in which each filter receives input from its upstream filter, processes the input and outputs the processed video to its downstream filter.

Method and apparatus for adaptive motion compensated filtering

Methods for video decoding and encoding, apparatuses and non-transitory computer-readable storage media thereof are provided. In one method for video decoding, a decoder may obtain a plurality of prediction blocks based on a current inter coding block; obtain a current template of the current inter coding block, wherein the current template comprises a plurality of reconstructed samples neighboring to the current inter coding block; obtain a plurality of template predictions of the current template respectively corresponding to the plurality of prediction blocks of the current inter coding block; obtain at least one filter based on the plurality of template predictions and the current template; and obtain a filtered block based on the at least one filter and the plurality of prediction blocks.
Owner:BEIJING DAJIA INTERNET INFORMATION TECH CO LTD

Method and apparatus for video coding using improved inloop filter for chroma component

A method and an apparatus are disclosed for video coding using an improved inloop filter for chroma components. In the disclosed embodiments, a video decoding device performs obtaining a reconstructed frame that includes luma samples and chroma samples, subsequently obtaining a scale value representing a resolution difference between the reconstructed frame and an original frame. Responsive to the reconstructed frame having a resolution that is based on the scale value and is less than a resolution of the original frame, the video decoding device generates an upsampled chroma frame by inputting the luma samples and the chroma samples into a cross component resampling inloop filter (CC-RIF).
Owner:HYUNDAI MOTOR CO LTD +2

Method and apparatus for cross-component prediction for video coding

The present disclosure provides a method for decoding video data, comprising: obtaining a video block from a bitstream, obtaining internal luma sample values of the video block, external luma sample values and external chroma sample values of an external region of the video block, determining, by using values based on the external luma sample values and the external chroma sample values, a set of weighting coefficients corresponding to a filter shape, wherein the filter shape and the set of weighting coefficients are configured for predicting a chroma sample value based on a plurality of corresponding luma sample values, and the values comprise non-downsampled values of the external luma sample values, or the non-downsampled values of the external luma sample values along with a downsampled value of at least one of the external luma sample values, predicting, with the filter shape and the set of weighting coefficients, internal chroma sample values of the video block based on the internal luma sample values, and obtaining decoded video block using the predicted internal chroma sample values.
Owner:BEIJING DAJIA INTERNET INFORMATION TECH CO LTD

Privacy preserving online video capturing and recording

Systems and methods are provided herein for only including portions of a user's environment that have been approved by a user in a video conference while excluding portions that have not been approved. This may be accomplished by a device receiving a policy identifying one or more approved objects of a scene of a video stream. The device may then generate a filtered video stream by only including portions of the scene that comprise the one or more objects that were approved by the policy in the filtered video stream. The filtered video stream may be combined with other video streams to generate a video conference that is transmitted and / or stored by one or more devices participating in the video conference.
Owner:ADEIA GUIDES INC

Video transmission high-definition image intelligent splicing method and system

The invention relates to the technical field of image stitching, and discloses a video transmission high-definition image intelligent stitching method and system. The system comprises an image preprocessing module, a feature point matching and screening module, an image optimal splicing seam generation module and a video output module. The method comprises the following steps: firstly, acquiring a video, carrying out dynamic range equalization, carrying out filtering processing by using an adaptive noise reduction method, and carrying out image correction; secondly, extracting feature points by using a self-adaptive corner detection algorithm, matching the feature points based on a multi-dimensional spatial data index to obtain feature point matching pairs, and screening the feature points; searching an overlapping region, and introducing a dynamic search algorithm to generate an optimal image splicing path to obtain an optimal image splicing seam; and finally, dividing the image according to the dynamic grid, and realizing video output by using priority ranking. According to the method, the video images are processed and spliced, the purpose of intelligent image splicing is achieved, and the method is accurate and objective.
Owner:GUANGZHOU WEITUXIN ELECTRONIC TECH CO LTD

A method, an apparatus and a computer program product for video encoding and video decoding

The embodiments relate to a method for encoding / decoding. The encoding method (900) comprises receiving a video sequence (1505) comprising a first frame and a second frame; encoding (1510) the first frame into a first coded frame using a first coding method (901); reconstructing (1515) a first decoded frame corresponding to the first coded frame; deriving (1520) one or more optimizing parameters to adjust a traditional filter (1110), wherein the optimizing parameters reduce distortion of the first decoded frame to produce a first filtered frame; filtering (1525) the first decoded frame with the traditional filter (1010, 1110); encoding (1530) the second frame into a second coded frame by a second set of algorithms of the second coding method (906, 1210) and by using the first filtered frame directly or indirectly for prediction; and signalling (1535) said one or more optimizing parameters. The embodiments also relate to apparatuses for encoding / decoding.
Owner:NOKIA TECHNOLOGIES OY

Film grain neural network post filter

Syntax elements allow to enable applying film grain using a neural network post-processing filter. Syntax elements define parameters related to the film grain neural network post- processing filter. An encoded video bitstream carries such syntax elements from an encoding device to a decoding device, thus allowing the encoding device to specify how to apply film grain using a neural network post-processing filter and the decoding device to apply the film grain neural network post-processing filter when displaying the decoded video. Additional syntax elements specify parameters comprising a film grain style, a film grain intensity, a film grain purpose or a region of interest where to apply the film grain.
Owner:INTERDIGITAL CE PATENT HOLDINGS SAS

Interpolation filter clipping for sub-picture motion vectors

A video coding mechanism is disclosed. The mechanism includes receiving a bitstream comprising a current picture including a sub-picture coded according to inter-prediction. A motion vector for a block of the sub-picture is determined. A clipping function is applied to sample locations in a reference block to support application of an interpolation filter when the motion vector points outside of the sub-picture and when a flag is set to indicate the sub-picture is treated as a picture. The interpolation filter is applied to results of the clipping function to obtain a predicted sample value. The block is decoded based on the predicted sample value. The block is forwarded for display as part of a decoded video sequence.
Owner:HUAWEI TECH CO LTD

Adaptive non-linear mapping for sample offset

A method for in-loop sample offset filtering in a video decoder is disclosed. The method includes obtaining at least one statistical property associated with reconstructed samples of at least a first color component in a current reconstructed data block of a video stream, selecting a target sample offset filter among a plurality of sample offset filters based on the at least one statistical property, the target sample offset filter comprising a nonlinear mapping between sample delta measures and sample offset values, and filtering a current sample in a second color component of the current reconstructed data block using the target sample offset filter and reference samples in a third color component of the current reconstructed data block to generate a filtered reconstructed sample of the current sample.
Owner:TENCENT AMERICA LLC

Cross-component sample offset edge direction derivations

An example method of video coding includes receiving a video bitstream comprising a plurality of frames, including a current frame comprising a current block. The method also includes determining an edge direction by performing edge detection for the current block, and determining a difference between a current sample of the current block and a neighboring sample based on the edge direction. The method further includes applying a filter to the current block based on the determined difference.
Owner:TENCENT AMERICA LLC

Production line multi-source heterogeneous stream data acquisition system and method

The invention discloses a production line multi-source heterogeneous stream data acquisition system, and the system comprises a video data preprocessing module which intercepts continuous frames from a video and extracts effective features from the video frames; the time synchronization module is used for uniformly managing timestamps of the multi-source heterogeneous data and identifying a time sequence relationship among the data; the protocol analysis module is used for analyzing the data frame of the communication protocol and carrying out effective data interaction with different devices or systems; the data fusion module is used for integrating the data into a unified data system model; the data storage module is used for collecting, processing and distributing data in real time; and the data display module provides data display, dynamically screens and filters data, and has an alarm function. The problem that data timestamps are inconsistent is effectively solved, all collected data can be compared and analyzed under the unified time reference, diversified data formats from different devices and systems are efficiently processed, and integration and unified management of multi-source heterogeneous data are achieved.
Owner:BEIJING JIAOTONG UNIV

Adaptive loop filter with samples before deblocking filter and samples before sample adaptive offsets

A device for decoding video data determines a pre-filtered reconstructed block of video data; applies one or more of a deblocking filter or a sample adaptive offset filter to the pre-filtered reconstructed block to determine a filtered reconstructed block; applies an adaptive loop filter (ALF) to the filtered reconstructed block to determine a final filtered reconstructed block, wherein to apply the ALF to the filtered reconstructed block, the device is further configured to determine a difference value based on a difference between a value of a current sample of the filtered reconstructed block and a value of a pre-filtered neighboring sample; apply a filter to the difference value to determine a sample modification value; and determine a final filtered sample value based on the sample modification value.
Owner:QUALCOMM INC

Adaptive filter for decoder-side intra mode derivation

Methods and apparatuses for video decoding and video encoding and a method of processing visual media data are disclosed. The apparatus for video decoding includes processing circuitry that receives coded information indicating that a current block in a current picture is coded with a decoder-side intra mode derivation (DIMD) mode. A template of the current block includes reconstructed samples in the current picture and is adjacent to the current block. The template includes one of a left template and a top template. The processing circuitry determines a filter type from a plurality of filter types associated with the one of the left template and the top template, applies the DIMD mode to the template based on the determined filter type to determine one or more intra prediction modes for the current block, and reconstructs the current block according to the one or more intra prediction modes.
Owner:TENCENT AMERICA LLC

Method, apparatus, and medium for video processing

Embodiments of the disclosure provide a solution for video processing. A method for video processing is proposed. The method comprises: performing, for a conversion between a video unit of a video and a bitstream of the video, a refinement on a prediction or reconstruction of the video unit by applying a filter for the video unit; and performing the conversion based on the refined prediction or refined reconstruction.
Owner:DOUYIN VISION CO LTD +1

Video encoding using reconstruction of spatially decimated frames

Approaches presented herein provide for the high quality, high resolution reconstruction of a sequence of decimated images, such as may be useful for remote desktop applications. The video frames can be sub-sampled or decimated such that each encoded frame only includes a fraction (e.g., ΒΌ) of the total pixel values for the full resolution frame. A server can analyze the current and previous video frames at full resolution to determine motion or actionable changes, and can apply a lowpass filter to those values based on the type of interpolation to be performed on the client. Static values can remain where no motion is detected. When a client receives the decimated and encoded video frames, the client can determine areas of motion and can perform bilinear interpolation for only those portions of the image where motion is detected, and can otherwise perform weaving of the static pixel values received over a limited sequence of decimated video frames.
Owner:NVIDIA CORP

Parameter signaling for CNN-based in-loop filters with multiple sets of neural network tools and contexts for video coding

A video encoder is configured to determine to filter video data using a neural network (NN)-based filter and a fixed block size inference, and encode a flag that indicates the fixed block size inference is used for the NN-based filter. Reciprocally, a video decoder is configured to decode a flag that indicates whether a fixed block size inference is used for an NN-based filter, and filter video data using the NN-based filter based on the flag. The flag may be signaled at a sequence parameter set (SPS) level.
Owner:QUALCOMM INC

Method and system for filtering a panoramic video signal using visual fixation

Methods and apparatus for filtering panoramic video signals using visual fixation are disclosed. A panoramic video signal organized into indexed frames is displayed and durations of visual fixations of a viewer of the display are measured by means of detecting the viewer's gaze positions and identifying clusters of adjacent gaze positions. For each visual fixation attaining a prescribed duration threshold, a segment of the panoramic video signal corresponding to a respective view region of prescribed shape and dimensions is extracted to form a filtered video signal. Contents of successive frames of the panoramic signal are cyclically supplied to a bank of content-filtering units operating concurrently to produce individual content-filtered frames which are concatenated for transmission to respective destinations.
Owner:3649954 CANADA INC

Encoding device, decoding device, and non-transitory machine-readable medium for coding video data

PCT designated stage expiredWO2025150539A1Digital video signal modificationAlgorithmTemplate based
An electronic device for decoding / encoding video data is provided. The electronic device includes at least one processor and at least one non-transitory computer-readable medium coupled to the at least one processor and storing one or more computer-executable instructions that, when executed by the at least one processor, cause the electronic device to: receive the video data; determine a block unit from a current frame included in the video data; determine at least one extrapolation area type and at least one extrapolation filter type; derive extrapolation filter models, each derived based on a corresponding extrapolation area type and a corresponding extrapolation filter type; determine template matching costs, each calculated using a corresponding extrapolation filter models; determine an arrangement of the extrapolation filter models based on the template matching costs; and reconstruct the block unit based on the arrangement. In addition, a non-transitory machine-readable medium for coding video data is also provided.
Owner:SHARP KK

Affine motion model restrictions for memory bandwidth reduction of enhanced interpolation filter

A method for coding a video implemented in an encoder or a decoder including the enhanced interpolation filter, EIF, for motion compensation, the method comprising: i) determining control point motion vectors, CPMVs, for a block according to affine inter-prediction, the block being an affine block or a sub-block of the affine block; ii) for a predefined sub-block size determining a reference area for a sub-block with the predefined sub-block size according to values of the CPMVs; iii) comparing the determined reference area with a predefined threshold; iv) applying EIF for motion compensation, comprising deriving the pixel-based motion vector field for the block; wherein if the determined reference area is larger than the threshold, deriving the pixel-based motion vector field for the block further comprises motion vector clipping, wherein motion vector clipping range is determined based on motion model of the block and the size of the block.
Owner:HUAWEI TECH CO LTD

Method, apparatus, and medium for video processing

Embodiments of the present disclosure provide a solution for video processing. A method for video processing is proposed. The method comprises: determining, for a conversion between a video unit of a video and a bitstream of the video unit, whether to apply at least one neural network (NN) filter model or determine a rate distortion cost during a rate distortion optimization (RDO) process of the video unit based on at least one of: a distortion without NN filter model, a distortion with n-th NN filter model, a combination of distortions of a plurality of NN filter models, or coding statistics of the video unit, and wherein n is an integer number; determining a coding mode of the video unit based on a rate distortion optimization (RDO) criterion in the RDO process; and performing the conversion based on the coding mode.
Owner:DOUYIN VISION CO LTD +1

Vehicle violation manned intelligent detection method and system based on CLIP and D-Fine cascade framework

The invention discloses a vehicle violation manned intelligent detection method and system based on a CLIP and D-Fine cascade framework, and the method comprises the steps: collecting a real-time video frame image of a preset traffic monitoring network, processing the real-time video frame image, inputting a monitoring image into a CLIP-ILP model, employing the CLIP-ILP model as a filter, and combining with a Top-K screening strategy, and screening out candidate images; and inputting the candidate image into a constructed context enhanced D-FINE detection model, carrying out positioning and multi-class detection on a preset target, outputting a multi-class detection result, and executing verification of a spatial co-occurrence rule, a license plate position heuristic rule, a size consistency rule and a context consistency rule. And marking the result which does not pass the verification as suspicious or rejecting the result, and outputting the result which passes the verification as an illegal manned detection result to the target terminal. According to the method, the problem of balance between the recall rate and the precision in manned detection can be solved.
Owner:YUNNAN MINZU UNIV +1

Multiple input sources based extended taps for adaptive loop filter in video coding

A mechanism for processing video data is disclosed. The mechanism includes determining to apply an adaptive looper filter (ALF) with an extended tap to a picture in a video. An intermediate filtering result of a second filter is used as input for the extended tap. A conversion is performed between a visual media data and a bitstream based on the ALF.
Owner:BYTEDANCE INC +1

Intra prediction based on extrapolation filter

Aspects of the present disclosure include methods and apparatus for video decoding and video encoding, and methods of processing visual media data. An apparatus for video decoding includes processing circuitry configured to: receive prediction information indicating that a current block in a current picture is predicted using an extrapolation filter-based intra prediction (EIP) mode; determining gradient information associated with a current sample in the current block; determining a prediction value of the current sample based on an initial prediction value predicted using the EIP mode and additional information including gradient information; and reconstructing the current sample according to the predicted value of the current sample.
Owner:TENCENT AMERICA LLC

Privacy preserving online video recording

Systems and methods are provided herein for only including portions of a user's environment that have been approved by a user in a video conference while excluding portions that have not been approved. This may be accomplished by a device receiving a policy identifying one or more approved objects of a scene of a video stream. The device may then generate a filtered video stream by only including portions of the scene that comprise the one or more objects that were approved by the policy in the filtered video stream. The filtered video stream may be combined with other video streams to generate a video conference that is transmitted and / or stored by one or more devices participating in the video conference.
Owner:ADEIA GUIDES INC

Neural network video coding in-loop filtering in transform domain

A method of coding video data, the method comprising: obtaining input data, wherein the input data includes one or more of predicted video data, reconstructed video data, quantization parameter data, boundary strength data, or prediction mode data; converting the input data from an input domain to a transform domain to generate converted video data; applying a neural network (NN)-based in-loop filter (ILF) to the converted video data to generate filtered video data; and converting the filtered video data from the transform domain to the input domain.
Owner:QUALCOMM INC

Neural network-based in-loop filter architectures for video coding

A video encoder and video decoder are configured to perform neural network (NN)-based filtering. The video encoder and video decoder may receive a picture of video data, reconstruct the picture of video data, and perform an NN-based filter process on one or more blocks of the reconstructed picture of video data using an NN-based filter, wherein the NN-based filter includes a pair of backbone blocks, each of the pair of backbone blocks comprising a three-component one-dimensional (1D) decomposition of a multi-dimensional convolution, wherein the 1D decomposition includes at least one layer with feature channel reduction.
Owner:QUALCOMM INC

Sample position and block characteristic dependent intra-prediction for video coding

An example device for decoding video data includes: a memory for storing video data; and a processing system implemented in circuitry and configured to: generating a prediction block for a current block of the video data using a sample position-dependent intra-prediction mode, including, for one or more samples of the prediction block: select a filter for the sample according to a shape of the current block and a position of the sample; and predict the sample using the selected filter; decode a residual block for the current block of the video data; and combine the prediction block with the residual block to decode the current block of the video data.
Owner:QUALCOMM INC