Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

71 results about "Rate distortion" patented technology

Rate distortion can be described in terms of Lagrangian multipliers. It can also be described by the Principle of Equal Slopes, which states that the coding parameters should be selected so that the rate of change of distortion with respect to bit rate is the same for all parts of the system.

Communication segmentation learning system and method for adaptive channel compression, and medium

The invention provides a communication segmentation learning system and method for adaptive channel compression, and a medium, and relates to the technical field of segmentation learning and communication compression. The system comprises a plurality of clients, a server side and a channel compression device arranged between the clients and the server side, the channel compression device comprises a channel sensitivity modeling module, a rate distortion adaptive compression module and a cross-client fair coordination module. In a channel sensitivity modeling module, channel importance is dynamically evaluated through intermediate layer activation value fusion; differentiated quantization and compression strategies are designed on the basis of sensitivity scores in a rate-distortion self-adaptive compression module, so that key information is reserved while communication overhead is reduced; and meanwhile, the cross-client fair coordination module realizes balanced distribution of communication resources among multiple clients through a fairness regularization and dual optimization mechanism, so that the influence of excessive compression on global convergence is avoided. According to the method, the communication efficiency and the training stability of segmentation learning in a complex heterogeneous environment are remarkably improved.
Owner:XIAMEN UNIV OF TECH

End-to-end image compression method and system based on window local attention and generalized checkerboard space channel context

The embodiment of the invention provides an end-to-end image compression method and system based on window local attention and generalized chessboard space channel context, and belongs to the technical field of image processing. The method comprises the following steps: constructing a transformation network based on an attention module and a stacked residual block; the transformation network based on the attention module and the stacked residual block is used for executing adaptive transformation of contents through dynamic representation and neighborhood information embedding to obtain potential features; establishing a generalized chessboard space channel context model; the generalized chessboard space channel context model is used for carrying out entropy coding on the potential features; and obtaining image compression data according to the transformation network based on the attention module and the stacked residual block and the generalized chessboard space channel context model. According to the method, redundancy can be eliminated to the maximum extent, excellent rate distortion performance is achieved, and meanwhile high-throughput parallel computing efficiency is ensured.
Owner:SUN YAT SEN UNIV

VVC code rate control algorithm based on deep reinforcement learning

The invention discloses a VVC code rate control algorithm based on deep reinforcement learning, and the algorithm comprises the following steps: importing a video sequence into an encoder, and enabling the video sequence to enter initial frame coding; after the encoder completes the default encoding of the first two frames, the subsequent frame prediction firstly extracts the encoding state information of the previous prediction frame, and a greedy strategy is adopted to perform action selection; an overall reward value is obtained through CTU-level code rate control and an actual coding process in sequence; observing the next state and the last reward, performing TD iteration on the Q value, and adding the Q value into the Q value network after TD iteration; setting an experience playback pool; after the capacity of the experience playback pool reaches a threshold value, randomly sampling from the experience playback pool in batches; resetting the Q value network; and carrying out coding test based on the Q value network obtained by training. The VVC code rate control algorithm based on deep reinforcement learning has a good capability of guiding code rate control coding, and compared with standard code rate control in VTM13.0, the VVC code rate control algorithm based on deep reinforcement learning can bring high rate distortion performance and improve code control precision.
Owner:HAINAN NORMAL UNIV

Three-dimensional point cloud prediction geometric coding method based on deep learning

The invention relates to a three-dimensional point cloud prediction geometric coding method based on deep learning, and the method comprises the steps: 1, forming a prediction tree through points collected by each laser transmitter, and converting an original laser radar point cloud LPC into a plurality of prediction trees for representation; 2, respectively designing different predictors and entropy encoders for each component of the coordinates of the points, and compressing each component of the coordinates of the points by applying different quantization step lengths; 3, selecting a quantization step size for quantifying each component by adopting a quantization step size selection strategy; 4, adopting an entropy model to model the probability distribution of the residual error of each component so as to encode the residual error entropy; and step 5, decoding is realized through a reverse process from the step 1 to the step 4. Compared with other methods, the method provided by the invention obtains the best rate distortion performance.
Owner:SHANDONG UNIV

Video coding method

The invention provides a video coding method, which comprises the following steps of: inputting a coding tree unit to be coded, and extracting texture features of a coding block, the texture features comprising local binary pattern similarity of all pixel points in the coding block and direction dispersion of each pixel point; inputting the texture features into the trained multi-type division prediction architecture, and sequentially judging whether to execute quadtree division, horizontal or vertical division, horizontal binary tree or ternary tree division and vertical binary tree or ternary tree division or not through a binary classification problem method; and calculating the rate distortion cost under each division mode to obtain an optimal division mode and a suboptimal division mode of the coding block, and outputting an optimal mode or a combination of the optimal mode and the suboptimal mode. According to the invention, the accuracy of the prediction mode in coding is improved.
Owner:SHANGHAI FULLHAN MICROELECTRONICS

Rate-distortion optimization with motion information

Systems and methods directed to a rate-distortion optimization process. One example is directed to a video encoding device including an electronic processor. The electronic processor is configured to obtain motion information associated with a video block and determine, as a function of the motion information, a subset of coding tools from a plurality of coding tools for rate-distortion performance evaluation of the video block.
Owner:INTERDIGITAL CE PATENT HOLDINGS SAS

Inter-plane prediction

A better rate distortion ratio is achieved by making interrelationships between coding parameters of different planes available for exploitation for the aim of redundancy reduction despite the additional overhead resulting from the need to signal the inter-plane prediction information to the decoder. In particular, the decision to use inter plane prediction or not may be performed for a plurality of planes individually. Additionally or alternatively, the decision may be done on a block basis considering one secondary plane.
Owner:DOLBY VIDEO COMPRESSION LLC

Decoder-complexity aware rate distortion optimization

In one implementation, an encoder obtains a first value indicating a number of bits used to encode a current block in a picture, under a coding option, and obtains a second value indicating a distortion between a reconstructed version and an original version of the current block associated with the coding option. The encoder obtains a cost function based on the first value and the second value for the current block. The encoder also obtains another cost function for the current block, associated with a current best coding option. By comparing the cost function and the another cost function, wherein the cost function or the another cost function is scaled by a complexity factor indicating computational complexity associated with the coding option, the encoder updates the another cost function and the current best coding option based on the comparison result.
Owner:INTERDIGITAL CE PATENT HOLDINGS SAS

Method and apparatus for loop filtering

The invention provides a method and apparatus for loop filtering. The loop filtering method comprises the following steps: acquiring an MCTF (Motion Compensation Time Domain Filtering) filtered image of a current image after MCTF; determining the similarity between the MCTF filtering image and a filtering reference image when the MCTF is carried out on the current image; and determining whether to perform rate distortion optimization of loop filtering using an original image of the current image or the MCTF filtered image based on the similarity.
Owner:BEIJING DAJIA INTERNET INFORMATION TECH CO LTD

Method and apparatus for encoding responsive to a lagrangian rate-distortion cost modified for adaptive loop filtering

In various implementations, method and devices are disclosed that involves video encoding based on a Lagrangian rate-distortion cost using a Lagrangian parameter. In a variant, a set of coding modes responsive to a Lagrangian rate-distortion cost comprises an adaptive in-loop filtering mode and wherein determining an adaptive in-loop filtering mode based on a Lagrangian rate-distortion cost using a Lagrangian parameter further comprises modifying a Lagrangian parameter for determining the adaptive in-loop filtering mode.
Owner:INTERDIGITAL CE PATENT HOLDINGS SAS

A video compression transmission method and related device

The application discloses a video compression transmission method and related equipment, the method comprises the following steps: obtaining a video sequence to be compressed, taking the first frame and the last frame as conditional frames; inputting the video sequence into the encoder of a pre-trained variational autoencoder to obtain a target latent space representation; processing the target latent space representation through the downsampling module of the compressor to generate an extreme compression representation; transmitting the conditional frames and the extreme compression representation to the receiving end to enable the receiving end to reconstruct the video through the upsampling module of the compressor, the generation model and the decoder of the variational autoencoder to obtain a reconstructed video sequence; the application provides time sequence boundary information through the conditional frames, and in combination with the reconstruction capability of the generation model, can effectively reduce the block effect, blur and high-frequency detail loss commonly seen in traditional methods; the conditional frames and the extreme compression representation significantly reduce the transmission code rate and bandwidth demand, can significantly improve the rate distortion performance, and can be widely applied to the technical field of video compression.
Owner:CHINA TELECOM ARTIFICIAL INTELLIGENCE TECHNOLOGY (BEIJING) CO LTD

Deep learning based video compression using optimized rate-distortion

This invention discloses deep learning-based video compression using optimized rate distortion. Systems and methods are provided for enhancing image patches of frames in a video to be compressed. These systems and methods enhance image patches to improve the compression of the video. In at least one embodiment, systems and methods are provided for enhancing video content using a learnable pre-filtering network trained with a joint loss function to improve compression.
Owner:NVIDIA CORP

A preprocessing algorithm for remote sensing image coding

The application discloses a kind of pre-processing algorithms for remote sensing image coding.The algorithm is mainly to extract image texture feature parameters, extract image texture feature parameters, according to the texture feature parameter bypasses the prediction mode and the division mode with smaller probability.The algorithm mainly includes image texture feature extraction;Fast rate distortion cost estimation;Intra-frame division prediction;Determine candidate prediction mode and division set.The present application, especially for hardware difficult to realize the H.265 encoding algorithm of remote sensing image.The algorithm saves hardware resources, reduces hardware implementation complexity and difficulty under the premise of trying to guarantee image compression quality, effectively utilizes hardware storage resources, and also reduces the time delay of compression and decoding.
Owner:SHENZHEN DIVIMATH SEMICON CO LTD

Inter-plane prediction

A better rate distortion ratio is achieved by making interrelationships between coding parameters of different planes available for exploitation for the aim of redundancy reduction despite the additional overhead resulting from the need to signal the inter-plane prediction information to the decoder. In particular, the decision to use inter plane prediction or not may be performed for a plurality of planes individually. Additionally or alternatively, the decision may be done on a block basis considering one secondary plane.
Owner:DOLBY VIDEO COMPRESSION LLC

Method and apparatus for encoding responsive to a lagrangian rate-distortion cost modified for adaptive loop filtering

In various implementations, method and devices are disclosed that involves video encoding based on a Lagrangian rate-distortion cost using a Lagrangian parameter. In a variant, a set of coding modes responsive to a Lagrangian rate-distortion cost comprises an adaptive in-loop filtering mode and wherein determining an adaptive in-loop filtering mode based on a Lagrangian rate-distortion cost using a Lagrangian parameter further comprises modifying a Lagrangian parameter for determining the adaptive in-loop filtering mode.
Owner:INTERDIGITAL CE PATENT HOLDINGS SAS

An intra prediction method, an encoder, a decoder and a storage medium

The embodiment of the application provides a kind of intra prediction method, encoder, decoder and storage medium, comprising: traversing intra prediction mode, determine the initial prediction value of the initial prediction block corresponding to current block.It is respectively carried out intra prediction filtering and intra prediction smoothing filtering processing to initial prediction block, obtain first type prediction value and second type prediction value;Intra prediction smoothing filtering is the process that a plurality of adjacent reference pixels in each adjacent reference pixel set in at least two adjacent reference pixel sets are filtered to current block.Using initial prediction value, first type prediction value and second type prediction value, rate distortion cost calculation is carried out with the original pixel value of current block, determine the current prediction mode corresponding to optimal rate distortion cost.Using current prediction mode, intra prediction is carried out to current block.The index information of current prediction mode and filter identification are written in code stream, and filter identification represents the identification corresponding to intra prediction filtering and / or intra prediction smoothing filtering.
Owner:GUANGDONG OPPO MOBILE TELECOMMUNICATIONS CORP LTD

Rate-distortion prediction based method and system for rate control of depth video encoder

The application provides a rate distortion prediction-based deep video encoder code rate control method and system, which comprises the following steps: step 1, training a prediction module; step 2, inputting a video frame into the prediction module to obtain a prediction point set; step 3, fitting a code rate and quality model according to the prediction point set; step 4, obtaining a frame-level code rate allocation ratio through a code rate control algorithm; and step 5, determining the corresponding encoding parameters of each frame and inputting the encoding parameters into an encoder for encoding. The application directly utilizes a neural network and an original video to predict the code rate model and the quality model of each frame for the first time, without pre-encoding; the video frame is down-sampled to a fixed small resolution before being inputted into the neural network, so that the efficiency is improved and the generalization is enhanced; and the application realizes code rate control at a mini-GOP level for the first time. Compared with the existing code rate control methods, the application can realize the same code rate control accuracy and finer code rate control granularity at a faster speed.
Owner:NANJING UNIV

Rate distortion optimization for time varying textured mesh compression

Apparatuses and methods are disclosed for encoding mesh data. Techniques disclosed include receiving a sequence of frames, each of which includes mesh data. For a frame in the sequence, techniques disclosed for encoding the mesh data of the frame according to a static path and according to a motion path of a multipath encoder, computing a static path cost of the encoding according to the static path and a motion path cost of the encoding according to the motion path, where the costs are computed by optimizing a rate-distortion cost function, and selecting, based on the computed motion path cost and static path cost, a bitstream generated by the encoding according to the motion path or a bitstream generated by the encoding according to the static path.
Owner:INTERDIGITAL CE PATENT HOLDINGS SAS

Quantization parameter selection based on rd costs approximation for adaptive quantization parameter

PCT designated stageWO2026175994A1AlgorithmRate distortion
Methods and apparatus are provided for adaptive quantization parameter selection based on a cost using a neural network. In one embodiment, the neural network can be trained from a training database comprising numerous coding units or blocks. In another embodiment, a rate distortion cost is determined for combinations of quantization parameter and split configuration. The quantization parameter with a minimal rate distortion cost for a particular split is used for encoding. In another embodiment, rate and distortion are separately approximated, resulting in lowest rate cost and lowest distortion cost for each possible type of split.
Owner:INTERDIGITAL CE PATENT HOLDINGS SAS

Adaptive bitrate optimization for content delivery

Techniques and systems for adaptive bitrate optimization for live content delivery are disclosed. In some embodiments, such techniques may include extracting, in real-time, features from content segments of a video stream; for each content segment of the video stream, identifying a rate-distortion (RD) cluster from a mapping of extracted features to RD clusters using one or more trained machine learning models; and transcoding the video stream by applying, to the content segments, a transcoding ladder having bitrate information corresponding to each identified RD cluster.
Owner:AMAZON TECH INC

Machine vision rate-distortion encoding method for license plate recognition task

A kind of machine vision rate distortion encoding method of license plate recognition task, by obtaining license plate image data set, extracting saliency feature, modulation adaptive quantization parameter, rate distortion optimization and fast division, license plate recognition step composition.The present application extracts background, vehicle, license plate three-level area feature using target detection network, calculates normalized perception intensity, uses hyperbolic tangent function to dynamically adjust the quantization parameter of coding unit, and optimizes the coding division process by combining subsection depth constraint strategy.Solves the technical problems of license plate character target blur and machine recognition rate decline under the conditions of application of video coding on unmanned aerial vehicle and limited code rate, reduces the encoding code rate and computational complexity, preserves the high-frequency microscopic details of license plate characters, improves the recognition accuracy and robustness of machine vision.The present application has the advantages of balancing code rate, low computational complexity, high license plate recognition accuracy, and can be used for unmanned aerial vehicle long-distance license plate recognition.
Owner:XIAN UNIV OF POSTS & TELECOMM

Method and apparatus for estimating motion vector of inter-frame coding

A method and apparatus for predicting motion vector for inter-frame encoding are provided. An implementation scheme of the method includes: acquiring first and second sets of motion vectors; dividing, in response to the number of valid adjacent PUs being greater than or equal to a preset number, the second set of motion vectors into at least one motion vector subset; calculating a correlation between the first set of motion vectors and each motion vector subset respectively, to obtain a priority of each adjacent PU; calculating, sequentially according to the priority in descending order, a rate distortion based on a motion vector of each adjacent PU, and stop calculating until a rate distortion smaller than a predetermined threshold is obtained; and determining a motion vector of an adjacent PU used when the rate distortion smaller than the predetermined threshold is obtained as the motion vector of the current PU.
Owner:BEIJING BAIDU NETCOM SCI & TECH CO LTD

Division decision method and device based on coefficient features, equipment and storage medium

The embodiment of the invention discloses a coefficient feature-based division decision method and device, equipment and a storage medium. According to the technical scheme provided by the embodiment of the invention, the skipping judgment threshold is determined according to the size information of the current coding block and the number of the non-zero coefficients by determining the number of the non-zero coefficients in the transformation coefficients of the current coding block and the rate distortion cost of the optimal mode; determining a first non-block division skipping strategy for the current coding block according to the rate distortion cost and a skipping judgment threshold, the first non-block division skipping strategy being used for indicating whether to skip non-block division prediction for the current coding block, and performing division decision on the current coding block based on the first non-block division skipping strategy, the skipping judgment threshold value is flexibly determined according to the number of non-zero coefficients in the transformation coefficients of the current coding block and the rate distortion cost of the optimal mode, so that a good video coding effect can be ensured while the video coding efficiency is improved.
Owner:BIGO TECH PTE LTD

Processing method and apparatus based on versatile video coding, computer device, and storage medium

The application provide a processing method based on versatile video coding, including: performing prediction at a first resolution for a current coding unit and determining a rate-distortion cost as a current best rate-distortion cost; and performing prediction at another resolution according to following steps: constructing a prediction motion vector list of a current traversal resolution based on a reference image; determining a start search point of the current traversal resolution based on the prediction motion vector list, and determining a first rate-distortion cost; when the first rate-distortion cost and the current best rate-distortion cost meet a preset condition, skipping motion estimation and motion compensation at the current traversal resolution, and using a prediction motion vector as an actual motion vector to calculate a second rate-distortion cost of the current traversal resolution; updating the current best rate-distortion cost based on the second rate-distortion cost and the current best rate-distortion cost.
Owner:SHANGHAI BILIBILI TECH CO LTD

Video coding mode determination method and device, encoder, medium and product

The embodiment of the invention discloses a video coding mode determination method and device, a coder, a medium and a product, and the method comprises the steps: determining a preset condition corresponding to a coding stage in a video coding process; wherein the nth preset condition corresponds to the nth coding stage, and is matched with the characteristics of the nth coding stage; n is smaller than or equal to N; n is the total number of the coding stages; under the condition that the coding parameter of the to-be-processed coding tree unit meets a target preset condition in the preset conditions, determining a coding mode of the to-be-processed coding tree unit as a coding mode corresponding to the target preset condition, and executing a skipping operation based on the coding mode corresponding to the target preset condition; the coding parameter comprises coding cost and / or a motion vector; the coding stage at least comprises any one of the following stages: a down-sampling motion estimation stage, an integer pixel motion estimation stage, a sub-pixel motion estimation stage and a rate distortion optimization stage.
Owner:MOORE THREADS TECH CO LTD

Three-dimensional video compact representation method and device, equipment and storage medium

The embodiment of the invention provides a three-dimensional video compact representation method and device, equipment and a storage medium, and relates to the technical field of video processing. The method comprises the following steps: determining a plurality of anchor point primitives according to a key frame sequence, carrying out quantization coding on the anchor point primitives to obtain a key frame coding code stream which comprises corresponding decoding anchor point primitives, selecting non-key frame sequences one by one as a processing frame sequence, obtaining a reference frame sequence according to a previous acquisition moment of the processing frame sequence, and carrying out decoding on the reference frame sequence according to a corresponding decoding anchor point primitive. Performing rate distortion optimization on the decoding anchor point primitive corresponding to the reference frame sequence to generate a coding motion parameter, determining a coding residual error anchor point primitive according to the decoding motion parameter corresponding to the coding motion parameter, and obtaining a decoding anchor point primitive corresponding to the processing frame sequence according to the decoding residual error anchor point primitive corresponding to the coding residual error anchor point primitive, and sending the coding code streams of all the frames to a decoding end for decoding operation. The video coding efficiency and expression compactness can be improved, and the overall coding complexity can be reduced.
Owner:PENG CHENG LAB

Encoding and decoding method, code stream, encoder, decoder and storage medium

The embodiment of the invention discloses a coding and decoding method, a code stream, a coder, a decoder and a storage medium, and the method comprises the steps: at a decoding end, the decoder determines a prediction mode corresponding to a current point; when the prediction mode is the first mode, determining a neighbor point corresponding to the current point in the current frame, and determining a reference point corresponding to the current point in a reference frame corresponding to the current frame; and determining a predicted value of the current point according to the first predicted value of the reference point and the second predicted value of the neighbor point. At an encoding end, an encoder determines a prediction mode corresponding to a current point according to a rate distortion optimization algorithm; or determining the prediction mode corresponding to the current point according to the correlation between the current point and the reference point corresponding to the current point in the reference frame; wherein the prediction mode comprises a first mode and a second mode.
Owner:GUANGDONG OPPO MOBILE TELECOMMUNICATIONS CORP LTD

Optimization of multi-layer rate distortion decision

Systems, methods, and instrumentalities are disclosed for selecting an inter-layer mode as the coding mode for encoding a portion of a video block. A plurality of candidate modes may be determined for encoding a portion of a video block. For each candidate mode of the plurality of candidate modes, a first rate-distortion cost may be calculated using at least a first rate-distortion multiplicator. It may be determined whether an inter-layer mode among the plurality of candidate modes has the lowest rate-distortion cost. The inter-layer mode may be selected as the coding mode, if the inter-layer mode has the lowest rate-distortion cost. An indication of the selected coding mode and encode the portion of the video block may be coded according to the selected coding mode.
Owner:INTERDIGITAL CE PATENT HOLDINGS SAS

Three-dimensional scene reconstruction method and apparatus, and device and storage medium

PCT designated stageWO2026144340A1Pattern recognitionVisual technology
The embodiments of the present application relate to the technical field of computer vision. Provided are a three-dimensional scene reconstruction method and apparatus, and a device and a storage medium. The method comprises: acquiring an anchor-based initial three-dimensional representation model corresponding to a target scene; for each anchor, performing compression encoding on anchor attributes to obtain an encoded bitstream corresponding to the target scene; performing entropy decoding on the encoded bitstream to obtain a decoded feature attribute, a decoded size attribute and a decoded offset attribute; on the basis of the decoded feature attributes and a feature attribute mean value, performing weighted prediction to obtain weighted feature attributes, and performing channel-wise grouping on the weighted feature attributes to obtain rendered feature attributes; and finally, on the basis of the rendered feature attributes, the decoded size attributes and the decoded offset attributes, generating a three-dimensional representation model of the target scene. By means of a weighted prediction process and a channel-wise grouping process, feature attributes of each anchor are simplified, so as to improve the rate distortion performance of an initial three-dimensional representation model while reducing a storage space.
Owner:PENG CHENG LAB