SEI Multiview Position Signaling for Accurate View Rendering
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing video coding technologies face challenges in efficiently signaling multiview view positions, leading to suboptimal compression and rendering of video data, particularly in applications requiring multiple perspectives.
Innovation Solution
The implementation of a supplemental enhancement information (SEI) message in the bitstream to provide multidimensional coordinates for multiple views, allowing for the determination and rendering of vertical and horizontal view positions, and reordering of pictures based on these coordinates.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If traditional video coding techniques are used without multiview position signaling, then the bitstream structure remains simple and decoding is straightforward, but the compression efficiency is suboptimal and multiview rendering quality deteriorates
Solution Approach 1:
The patent segments the bitstream by introducing a separate SEI message structure dedicated to multiview position information. This separates the position signaling from the main video data, allowing efficient compression without complicating the core decoding process. The SEI message contains organized arrays of view position data that can be processed independently.
Solution Approach 2:
The SEI message acts as an intermediary carrier for multiview position information. It mediates between the need for detailed position data and the desire for simple bitstream structure by providing a standardized, self-contained message format that can be attached to the bitstream without altering the core video coding structure.
2Measurement precision
If multiview position information is not signaled, then the bitstream size remains small and transmission bandwidth is conserved, but rendering accuracy for multiple perspectives deteriorates
Solution Approach 1:
The patent uses parameter changes by encoding view positions using coordinate systems and differential encoding techniques. Position information is represented through structured parameters (view identifiers, position coordinates, dimensional information) that can be efficiently compressed while maintaining precision. The use of arrays and structured data formats allows compact representation of multiple view positions.
3Adaptability or versatility
If existing SEI message formats are used without multiview position extensions, then the message structure remains simple and parsing is easy, but multiview video coding and rendering capabilities are insufficient
Solution Approach 1:
The patent extends the SEI message structure by adding dimensional information for multiview positioning. It introduces arrays of view position data with horizontal and vertical coordinate dimensions, transforming the traditional scalar position representation into a multi-dimensional structured format that accommodates multiple perspectives while maintaining message organization.
Data Source
AI summary
Aspects of the disclosure provide methods and apparatuses for video encoding/decoding. In some examples, an apparatus for video decoding includes processing circuitry. The processing circuitry decodes pictures associated with views in a bitstream and determines, from a supplemental enhancement information (SEI) message, positions of multidimensional coordinates in a multidimensional space respectively for the views. In an example, a position of a view is defined in the SEI message by at least a vertical view position and a horizontal view position in the multidimensional space. Further, the processing circuitry determines a rendering picture from the pictures based on a rendering view that is defined by at least the vertical view position and the horizontal view position in the multidimensional space.


