Free Viewpoint Video Encoding via Planar Splicing

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing video encoding technologies face delays and high computational pressures when handling multiple client viewpoints, limiting scalability and viewer experience.

Innovation Solution

The method involves generating a planar splicing image and splice information from multiple single-viewpoint videos at a server, encoding this information along with camera side information to create a planar splicing video bit stream, which is then decoded at the client side to synthesize virtual viewpoints, reducing delay and computational load.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Reliability

If viewpoint information is transmitted from client to server for real-time rendering, then viewing experience is improved, but transmission delay increases

Engineering Contradiction:
Improveviewing experienceVSAvoidtransmission delay
Core Design Contradiction:
ReliabilityVSLoss of time

Solution Approach 1:

Instead of transmitting viewpoint information from client to server and rendering at server, the patent inverts the approach by pre-rendering multiple viewpoints at the server and transmitting only the necessary view synthesis data to the client for local rendering. This reverses the traditional client-server rendering responsibility and eliminates transmission delay for viewpoint information.

Inventive Principle:
Principle #13The other way round (Inversion)

Solution Approach 2:

The patent performs preliminary rendering of multiple viewpoints (left-eye and right-eye views) at the server before transmission. By pre-processing the video content into multiple viewpoints and encoding them in advance, the system avoids real-time rendering delays at the server and enables faster client-side playback.

Inventive Principle:
Principle #10Preliminary action

2Adaptability or versatility

If multiple client viewpoints are processed at server, then viewing flexibility is improved, but computational pressure increases

Engineering Contradiction:
Improveviewing flexibilityVSAvoidcomputational pressure
Core Design Contradiction:
Adaptability or versatilityVSPower

Solution Approach 1:

The patent segments the rendering task by dividing the field of view into left-eye and right-eye view components. Each viewpoint is processed independently and encoded separately, allowing the server to handle multiple clients simultaneously with reduced computational overhead per client while maintaining viewing flexibility.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent creates copies of video content for different viewpoints (left and right views) at the server side. By pre-generating these viewpoint copies and transmitting them to clients, the server avoids real-time computational pressure when multiple clients request different viewpoints, as the rendering work is already completed during encoding.

Inventive Principle:
Principle #26Copying

3Adaptability or versatility

If all possible viewpoints are generated for storage, then viewing freedom is improved, but storage requirements increase

Engineering Contradiction:
Improveviewing freedomVSAvoidstorage requirements
Core Design Contradiction:
Adaptability or versatilityVSQuantity of substance

Solution Approach 1:

The patent applies partial action by generating only the essential viewpoints needed for free viewpoint video (typically left-eye and right-eye views) rather than all possible viewpoints. This selective generation approach provides sufficient viewing freedom for stereoscopic and free viewpoint applications while keeping storage requirements manageable.

Inventive Principle:
Principle #16Partial or excessive action

Data Source

PatentUS11330301B2Method and device of encoding and decoding based on free viewpoint
Publication Date: 2022.05.10 PEKING UNIV SHENZHEN GRADUATE SCHOOL
  • US11330301B2 patent drawing
  • US11330301B2 patent drawing
  • US11330301B2 patent drawing

AI summary

The present application provides a method and a device of encoding and decoding based on free viewpoint, and relates to the technical field of video encoding. The method includes: generating a planar splicing image and splice information based on multiple single-viewpoint videos at a server side; generating a planar splicing video based on the planar splicing image; generating camera side information of the planar splicing video based on camera side information existing in the multiple single-viewpoint videos; and encoding the planar splicing video, the splice information and the camera side information of the planar splicing video to generate a planar splicing video bit stream, and decoding a planar splicing video bit stream to acquire a virtual viewpoint according to viewpoint information of a viewer at client side.