Patents
Literature
Patsnap Eureka AI that helps you search prior art, draft patents, and assess FTO risks, powered by patent and scientific literature data.

18 results about "Variable bit rate vbr" patented technology

Variable bit rate (VBR) encoding is an alternative to constant bit rate encoding (CBR) and is supported by some codecs. Where CBR encoding strives to maintain the bit rate of the encoded media, VBR strives to achieve the best possible quality of the encoded media.

Dynamic self-adaptive edge-cloud cooperative computing segmentation point and compression strategy adjustment method

The invention discloses a dynamic self-adaptive edge-cloud cooperative computing segmentation point and compression strategy adjustment method. The core of the method is to deploy a dynamic decision engine on edge equipment. The engine senses a network state (such as bandwidth and delay) and a task priority specified by an application in real time, and according to a preset decision logic (such as a look-up table or a lightweight prediction model), a deep learning model which supports multiple segmentation points and is equipped with a variable bit rate hierarchical quantization feature compression model for each segmentation point is used for performing a multi-segmentation-point hierarchical quantization feature compression model on the basis of the multi-segmentation-point hierarchical quantization feature compression model, so that the multi-segmentation-point hierarchical quantization feature compression model is obtained. And dynamically selecting an optimal model segmentation point and feature compression level combination. The combination is transmitted to the cloud, and the cloud loads the corresponding back-end model to complete calculation. Through the dynamic adjustment mechanism, the real-time optimal balance among the reasoning delay, the task precision and the network overhead is realized, and the performance stability, the resource utilization rate and the adaptability of the system in a variable environment are remarkably improved.
Owner:杭州智元研究院有限公司

Token stream guide variable code rate image compression method oriented to unification of perception and understanding

The invention discloses a token stream guide variable code rate image compression method oriented to unification of perception and understanding. The method comprises the following steps of: 1, acquiring a two-dimensional image, and processing the two-dimensional image to acquire a one-dimensional token sequence; 2, processing the one-dimensional token sequence through a variable token mask to generate a binary compressed bit stream; 3, decoding the compressed bit stream to obtain an unmasked token sequence, and carrying out dynamic prediction on the unmasked token sequence to recover a complete token sequence; and 4, realizing human perception through the receiving end I after the complete token sequence is recovered, realizing machine perception through the receiving end II after the complete token sequence is recovered, and finally realizing image compression. The method has the capability that a single model supports the continuous variable bit rate, supports a large language model to directly perform semantic understanding based on the compressed code stream, and has the characteristics of low calculation complexity and low delay.
Owner:XIDIAN UNIV

Lightweight scalable bit-rate multi-view image compression method and model

The application provides a lightweight variable bit rate multi-view image compression method and system, and relates to the technical field of image compression. The method comprises: downsampling, feature extraction and feature scaling of a single-view image to obtain a latent representation corresponding to a target bit rate; quantization and lossless entropy coding of the latent representation to obtain a final compressed bit stream; subsequent lossless entropy decoding and inverse scaling to restore the latent representation; feature fusion and upsampling of the restored latent representations of different views to generate a reconstructed compressed image. The model comprises a main encoder, a feature scaling module, a quantization module, an autoregressive entropy model, an arithmetic encoder, an arithmetic decoder, a feature inverse scaling module and a decoder. The application can efficiently compress image data, reduce computational complexity and storage space occupation while preserving image details and quality, thereby providing faster speed and lower bandwidth requirement for image transmission and storage.
Owner:WUHAN UNIV OF TECH

Autoencoder and method for adaptive learned image compression with configurable encoder and decoder for variable bitrate applications

An autoencoder and a method for adaptive learned image compression with configurable encoder and decoder for variable bitrate applications are described, wherein the encoder comprises: - a pre-trained transformer-based architecture operable to convert input image data into quantized latent representations (9); - one or more low-rank adapter modules (W a,W b) integrated within said encoder (g a), each adapter module (W a,W b) being configured to adjust the encoding process for different target bitrates while maintaining pre-trained model parameters; - wherein said one or more low-rank adapter modules (W a,W b) are incorporated into fully connected layers (120,121) of the encoder's multi-layer perception modules (MLP), enabling said encoder (g a) to achieve efficient rate-distortion performance across a range of bitrates by fine-tuning only parameters of said adapter modules (W a,W b); - a mechanism adapted to merge said adapter modules (W a,W b) with pre-trained model weights following adaptation, thereby restoring the encoder's complexity to its original level.
Owner:SISVEL TECH +2

Truncateable predictive coding

A method, system, and computer program to encode and decode a channel coherence parameter applied on a frequency band basis, where the coherence parameters of each frequency band form a coherence vector. The coherence vector is encoded and decoded using a predictive scheme followed by a variable bit rate entropy coding.
Owner:TELEFONAKTIEBOLAGET LM ERICSSON (PUBL)

A large model-based digital hybrid coding method

The application discloses a kind of digital hybrid coding methods based on large model, including the following steps: S1, pre-processing to be encoded data;S2, feature learning method driven by diffusion model is used to feature coding to pre-processing data;S3, neural network coding is carried out to feature coding data, and based on variable bit rate quantization method dynamically adjusts encoding bit number;S4, adaptive entropy coding is carried out to neural network coding data, based on normalized stream entropy coding method carries out probability density transformation;S5, discrete transform coding is carried out to entropy coding data, and hierarchical sub-block quantization method is used to execute hierarchical quantization to different frequency components;S6, based on A3C optimization transform coding data's code length distribution and entropy coding parameter, extract lightweight coding model from diffusion model using knowledge distillation method, generate final optimization coding data.The application improves data compression efficiency and coding quality significantly by hybrid coding and dynamic optimization strategy.
Owner:广州企通云网络科技有限公司

Variable bit rate quantization speech recognition method for intelligent speech controller

The invention relates to a variable bit rate quantization speech recognition method for an intelligent speech controller. The variable bit rate quantization speech recognition method is applied to a first end comprising a target sensing array. The first end dynamically collects noisy voice signals through the voice collection module, and when the signals are detected, time domain features and frequency domain features are collected through the time domain feature unit and the frequency domain feature unit respectively. The features are sent to a second end, the second end fuses the features to generate an enhanced speech signal, variable bit rate quantization coding is carried out, and then a text transcription result is output through a transformation decoding structure. The method realizes efficient and accurate speech recognition in a complex noise environment, has adaptive bit rate allocation capability, remarkably improves transmission efficiency and recognition robustness, and is suitable for scenes such as smart home and the like.
Owner:SHENZHEN JINGHUA DISPLAY

Method and apparatus for pre-buffer media storage

Image capture devices and methods may be used to pre-buffer media storage. The pre-buffering method includes recording an image capture segment of a variable bitrate input stream in a circular buffer. The circular buffer includes a number of recordable segments. The method includes recording a next image capture segment in a next adjacent recordable segment of the circular buffer if the next adjacent recordable segment of the predetermined number of recordable segments is available. The method includes overwriting an oldest recordable segment if the next adjacent recordable segment of the predetermined number of recordable segments is not available.
Owner:GOPRO INC

Systems, methods, and devices for optimizing streaming bitrate based on multiclient display profiles

Systems, methods, and devices are provided for optimizing streaming bitrate during multiclient streaming sessions based, at least in part, on display profiles associated with client media receivers to which different video streams are concurrently provided. The method may be carried-out by a streaming media server in communication with first and second client media receivers over a network. In various embodiments, the method may include establishing at the streaming media server first and second bandwidth allotment thresholds based, at least in part, on display profiles assigned to display devices associated with the client media receivers. During an ensuing multiclient streaming session, the streaming media server further encodes segments of video streams at variable bitrates regulated in accordance with the established bandwidth allotment thresholds. Additionally, the streaming media server transmits the encoded segments of the video streams over the network to the client media receivers for presentation on the display devices.
Owner:SLING MEDIA PVT LTD

Dynamic adaptive edge-cloud collaborative computing split point and compression strategy adjustment method

The application discloses a dynamic self-adaptive edge-cloud collaborative computing segmentation point and compression strategy adjustment method. The core of the method is to deploy a dynamic decision engine on the edge device. The engine real-time perceives network state (such as bandwidth, delay) and application specified task priority, and according to the preset decision logic (such as query table or lightweight prediction model), dynamically selects the optimal model segmentation point and feature compression level combination from a deep learning model supporting multiple segmentation points and equipped with variable bit rate layered quantization feature compression model for each segmentation point. The combination is transmitted to the cloud, and the corresponding backend model is loaded to complete the calculation. Through the dynamic adjustment mechanism, the application realizes the real-time optimal balance among inference delay, task accuracy and network overhead, and significantly improves the performance stability, resource utilization and adaptability of the system in a variable environment.
Owner:杭州智元研究院有限公司

Progressive generative face video compression with bandwidth intelligence

Methods and systems implement a progressive generative face video compression framework with bandwidth intelligence, hierarchically accommodating variable bitrate video communication and implementing high-fidelity face reconstruction towards overall bandwidth coverage. Heterogeneous-granularity facial description regularizes long-term dependencies between video frames and compensates for motion estimation errors caused by compact representations of motion information, achieving satisfactory human visual perception and bandwidth intelligence in a progressive fashion. High efficiency for heterogeneous-granularity signal compression is achieved by two different entropy-based signal compression methods: heterogeneous-granularities feature representation from the key-reference frame as hyperpriors to optimize the entropy model for compressing heterogeneous-granularity feature from subsequent inter frames, and a feature difference operation for heterogeneous-granularities feature representation between key-reference and subsequent inter frames, such that the entropy model only compresses heterogeneous-granularities feature residual for redundancy reduction.
Owner:ALIBABA (CHINA) CO LTD

Compression for split neural network computing to accommodate varying bitrate

Various systems and methods for providing variable bitrate compression for split deep neural network (DNN) computing are described herein. A system may be configured to manage a split DNN, the split DNN configured to operate on a compute system and a second system over a communication network. The system may access a performance metric; determine, based on the performance metric, a split point of the split DNN, the split point defining a head portion of the split DNN and a tail portion of the split DNN; determine, based on the performance metric, a bottleneck layer configuration for a bottleneck layer at the split point, the bottleneck layer including a bottleneck encoder and a bottleneck decoder; execute the head portion of the DNN and the bottleneck encoder on the compute system; and recurrently access an updated performance metric and determine a revised split point or a revised bottleneck layer configuration based on the updated performance metric.
Owner:INTEL CORP

Variable bit rate dynamic point cloud geometric lossy coding and decoding method and device based on feature scaling and resolution adaptive network

The invention discloses a variable bit rate dynamic point cloud geometric lossy coding and decoding method and device based on feature scaling and a resolution adaptive network, and the method comprises the steps: carrying out the voxelization of a to-be-coded point cloud frame and a reference point cloud frame, and obtaining a to-be-coded point cloud grid and a reference point cloud grid; carrying out lightweight down-sampling based on sparse convolution, and carrying out lossy mapping on the to-be-coded point cloud grid and the reference point cloud grid to a feature space to obtain voxelization features; obtaining a first hidden variable according to the voxelization feature; generating a binary mask code stream according to the voxelization features based on spatial prediction; generating a transmission code stream according to the first hidden variable and the binary mask code stream; and decoding the transmission code stream to obtain a reconstructed point cloud. The method can improve the coding and decoding efficiency, and can be widely applied to the technical field of point cloud compression.
Owner:SUN YAT SEN UNIV

Methods and systems for providing variable bitrate content

Provided are methods and systems for providing variable bitrate content (e.g., video content, audio content, multimedia content, etc.). The content may comprise a plurality of portions (e.g., frames, segments, fragments, etc.). Each portion of the content may be tagged and / or associated with a content element. A content element may be associated with and / or indicate one or more of an encoding parameter associated with a respective portion of the content, attributes of the content (e.g., a scene transition, scene change, etc.) associated with the respective portion of the content, or additional content related items (e.g., one or more advertisements, etc.) associated with the respective portion of the content. The content element may be used to determine a bitrate to associate with the respective portion of the content. The content may be received / retrieved according to the determined bitrates.
Owner:COMCAST CABLE COMM LLC

Single model based variable bit rate learned image and video compression using adaptive quantization offset

Systems, methods, and instrumentalities are disclosed herein for quantization processes used within image or video compression that perform end-to-end learning with additional parameters (e.g., quantization offsets). A video decoding and / or encoding device may include a processor. The device may be configured to obtain latent variables associated with a neural network (NN) model. The device may obtain a quantization offset associated with the latent variable. The device may reconstruct a quantized latent variable associated with the NN model based on the quantization offset. The quantization offset associated with the latent variable may be obtained based on at least one of a lookup table, using a second neural network, or an estimated standard deviation based on a distribution associated with the latent variable.
Owner:INTERDIGITAL VC HOLDINGS INC

A variable rate 4D gaussian compression method

ActiveCN120034657BDigital video signal modificationKey frameVariable bit rate vbr
The application provides a variable bit rate 4D Gaussian compression method, comprising the following steps: S1: using a complete 3D Gaussian to represent a first frame (key frame); S2: using a motion grid and two shared light-weight multi-layer perceptron (MLP) to perform motion estimation on an inter-frame Gaussian cell; S3: using a sparse compensation Gaussian to perform motion compensation on an inter-frame change region; S4: using a complete Gaussian expression of a previous frame in a buffer, a motion grid and a sparse compensation Gaussian of a current frame to reconstruct a complete Gaussian expression of the current frame, and storing the complete Gaussian expression in the buffer for reconstruction of a next frame; and S5: quantizing and entropy encoding the motion grid and the sparse compensation Gaussian which have been trained, compressing a model size, obtaining a variable bit rate code stream, and realizing streaming. The application realizes a wide variable bit rate through a single model, while keeping superior rate distortion performance, and can be used for streaming to adapt to different requirements.
Owner:SHANGHAI JIAOTONG UNIV

Controller state control method and device, controller and electric vehicle

The application discloses a controller state control method and device, a controller and an electric vehicle, and relates to the technical field of automobile control. The method comprises the following steps: receiving a first message, wherein the first message is a message on a controller area network (CAN) with flexible data rate (CANFD) bus supporting a variable bit rate; in the case that it is determined that the first message is not a network management message, controlling the controller to be in a sleep state; and in the case that it is determined that the first message is a network management message, controlling the controller to be in a working state. According to the scheme, in the case that the received message is not a network management message, the controller can keep in a sleep state without sending a message to a network or performing an unexpected output, interference to a whole vehicle network is avoided, and the energy consumption of the whole vehicle is reduced.
Owner:BEIJING ELECTRIC VEHICLE

Truncateable predictive coding

A method, system, and computer program to encode and decode a channel coherence parameter applied on a frequency band basis, where the coherence parameters of each frequency band form a coherence vector. The coherence vector is encoded and decoded using a predictive scheme followed by a variable bit rate entropy coding.
Owner:TELEFONAKTIEBOLAGET LM ERICSSON (PUBL)