Volumetric Video Aggregation for Personalized Free-Viewpoint Rendering

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional volumetric video systems limit viewers to a fixed perspective and do not allow for personalized selection or aggregation of content from multiple sources, restricting immersive experiences.

Innovation Solution

The implementation of a volumetric video module that allows for the aggregation of two or more source videos based on user-defined aggregation rules, enabling personalized rendering of objects and attributes, such as object selection, modification, and combination of multiple volumetric videos.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Area of stationary object

If conventional 360° video stitching is used to create immersive content, then the field of view is expanded to cover the entire sphere, but the viewer's movement is restricted to a fixed camera position

Engineering Contradiction:
Improvefield of viewVSAvoidviewer movement freedom
Core Design Contradiction:
Area of stationary objectVSAdaptability or versatility

Solution Approach 1:

The patent transitions from 2D spherical video stitching to 3D volumetric video representation by incorporating depth information through photogrammetry or light field arrays. This dimensional enhancement allows viewers to move freely in three-dimensional space while maintaining immersive field of view, resolving the contradiction between expanded view area and movement freedom.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

2Adaptability or versatility

If multiple camera angles are captured simultaneously to create volumetric video, then depth information and three-dimensional parameters are obtained, but the system complexity increases

Engineering Contradiction:
Improvedepth information captureVSAvoidcamera system complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent employs photogrammetry techniques that enable a single camera system to capture multiple viewpoints and depth information simultaneously through multi-functional operation. The same camera hardware serves multiple purposes: capturing images from different angles, extracting depth data, and generating three-dimensional parameters, thereby reducing overall system complexity while maintaining volumetric video capabilities.

Inventive Principle:
Principle #6Universality (Multi-functionality)

3Adaptability or versatility

If photogrammetry or light field arrays are used to capture depth information, then volumetric video with free viewpoint is achieved, but the manufacturing complexity and cost increase

Engineering Contradiction:
Improvefree viewpoint capabilityVSAvoidsystem implementation difficulty
Core Design Contradiction:
Adaptability or versatilityVSEase of manufacture

Solution Approach 1:

The patent uses photogrammetry to create virtual copies of the physical scene from multiple angles and depths. Instead of requiring complex multi-camera light field arrays, the system captures images from limited viewpoints and computationally generates three-dimensional representations that replicate the appearance and depth of objects from any viewpoint, significantly simplifying the manufacturing and implementation process.

Inventive Principle:
Principle #26Copying

4Adaptability or versatility

If conventional video aggregation is used to combine multiple sources, then content variety is increased, but personalization and user control over content selection are lost

Engineering Contradiction:
Improvecontent varietyVSAvoiduser control capability
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The patent implements dynamic, user-configurable aggregation rules that allow viewers to personalize their experience in real-time. Users can dynamically adjust parameters such as time of day preferences, object selection criteria, and content source weighting, enabling both content variety from multiple sources and full personalization control through an adaptable aggregation system.

Inventive Principle:
Principle #15Dynamics

Data Source

PatentUS12505670B2Personalized aggregation of volumetric videos
Publication Date: 2025.12.23 INTERNATIONAL BUSINESS MACHINE CORPORATION
  • US12505670B2 patent drawing
  • US12505670B2 patent drawing
  • US12505670B2 patent drawing

AI summary

An embodiment includes selecting, using a first attribute of a first object, the first object in a first volumetric video. The embodiment also includes selecting, using a second attribute of a second object, the second object in a second volumetric video, where the first attribute and the second attribute satisfy an aggregation rule. The embodiment also includes generating an aggregated volumetric video from the first volumetric video and the second volumetric video, where the generating of the aggregated video comprises rendering the first object and the second object simultaneously in the aggregated volumetric video based on the aggregation rule.