Patch-Based Reshaping for 3D Video Point Cloud Compression
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current 3D video coding technologies face challenges in efficiently encoding and decoding 3D video content, particularly in consumer devices with limited codecs, leading to incorrect interpretation and visible artifacts in rendered images.
Innovation Solution
The implementation of patch-based reshaping and metadata techniques for video point cloud compression (V-PCC) and visual volumetric video-based coding (V3C), which involve generating and reshaping patches from 3D point clouds, encoding operational parameters, and transmitting them as metadata to enable accurate reconstruction of 3D video content across devices.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Productivity
If 3D video content is encoded in a format not supported by consumer devices, then coding efficiency and compression can be improved, but device compatibility and ease of operation deteriorate
Solution Approach 1:
The patent segments 3D video content into multiple 2D video views that can be independently encoded and decoded by consumer devices. Each view represents a different angular perspective of the 3D scene, allowing devices to decode whichever views they support while maintaining compatibility with existing 2D video codecs
Solution Approach 2:
The patent creates a universal 3D video format that can be decoded by both advanced devices with 3D capabilities and consumer devices with limited codecs. The encoded bitstream contains multiple views that can be interpreted differently by different device types, making the same encoded content universally compatible
2Manufacturing precision
If advanced 3D video coding formats are used, then manufacturing precision and representation accuracy can be improved, but device complexity increases
Solution Approach 1:
Instead of requiring devices to perform complex 3D decoding operations, the patent inverts the approach by pre-processing 3D content into multiple 2D views during encoding. This shifts the complexity from the decoding side to the encoding side, allowing simple consumer devices to decode the content using standard 2D video decoders
3Speed
If 3D video content is decoded without proper format support, then processing speed can be improved, but measurement precision and interpretation accuracy deteriorate
Solution Approach 1:
The patent creates multiple 2D view copies of the 3D scene from different angles, which can be decoded independently and quickly by consumer devices. Each view is a complete representation that can be decoded at full speed without requiring complex 3D reconstruction, while still maintaining accurate representation of the original 3D content
Data Source
AI summary
An input 3D point cloud including a spatial distribution of points is received. Patches including pre-reshaped patch data are generated from the input 3D point cloud. Encoder-side reshaping is performed on the pre-reshaped patch data to generate reshaped patch data for the patches. The reshaped patch data is encoded into a 3D video signal, which a recipient device of the 3D video signal can decode to generate a reconstructed 3D point cloud that approximates the input 3D point cloud.


