Multiplane Video Layering for Immersive 3D Videoconferencing

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current videoconferencing systems lack the ability to create immersive, three-dimensional-like experiences on two-dimensional screens, limiting user interaction and engagement, as they do not allow participants to move their videos within a virtual environment or customize the background and foreground layers dynamically.

Innovation Solution

A videoconferencing system that employs a multiplane camera view by transmitting multiple video streams in different layers, enabling participants to adjust and interact with these layers, including full-body segmentation and gesture recognition, to create a dynamic and immersive experience.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If traditional single-plane video streaming is used, then system complexity is low, but the three-dimensional effect and user immersion are insufficient

Engineering Contradiction:
Improvethree-dimensional effectVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The video environment is segmented into multiple independent layers (foreground layer, background layer, participant video layer) that can be transmitted and processed separately. Each layer contains specific visual elements that can be independently controlled and customized by users, enabling three-dimensional spatial arrangement without requiring complete system redesign.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system transitions from traditional two-dimensional video display to multi-dimensional spatial arrangement by adding depth perception through layered composition. Users can adjust the Z-axis positioning of different video layers, creating a three-dimensional virtual environment that enhances immersion while maintaining compatibility with standard 2D display devices.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

2Adaptability or versatility

If multiple video layers are transmitted, then user customization and interaction are enhanced, but network bandwidth and processing requirements increase

Engineering Contradiction:
Improveuser customizationVSAvoidnetwork bandwidth
Core Design Contradiction:
Adaptability or versatilityVSUse of energy by moving object

Solution Approach 1:

Instead of transmitting complete high-resolution video for all layers simultaneously, the system transmits partial video data at reduced resolutions for background and foreground layers, while maintaining higher quality for participant videos. Users can dynamically adjust the detail level of each layer based on their needs, reducing overall bandwidth consumption while preserving essential customization capabilities.

Inventive Principle:
Principle #16Partial or excessive action

3Ease of operation

If full-body segmentation and gesture recognition are implemented, then natural interaction is improved, but computational requirements and processing time increase

Engineering Contradiction:
Improvenatural interactionVSAvoidprocessing requirements
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The system automatically performs full-body segmentation and gesture recognition without requiring manual user input or configuration. AI algorithms continuously analyze video feeds to identify body parts and gestures, automatically adjusting layer compositions and interactions based on detected user actions, thereby simplifying operation while managing processing through efficient automated pipelines.

Inventive Principle:
Principle #25Self-service

Data Source

PatentUS11843897B1Method of using online, real-time, interactive, multiplane camera view to enhance videoconferencing platforms
Publication Date: 2023.12.12 SLOTZNICK BENJAMIN
  • US11843897B1 patent drawing
  • US11843897B1 patent drawing
  • US11843897B1 patent drawing

AI summary

A user interface display is provided to a participant who is viewing an event via a communication system that provides videoconferencing. The communication system provides a composite video stream that includes a plurality of different video layers, each video layer providing a different portion of the composite video stream. The participant has a participant computer for allowing the participant to receive the composite video stream for display on the user interface display. A plurality of participants view the event via user interface displays of their respective participant computers. The layers include a participant layer that displays video streams of the participants, a foreground layer, an event layer that includes video of the event, an audience layer, a background layer, and an immersive layer.