3D Video Generation Using Foreground Segmentation and Occlusion

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing methods for generating 3D visual effects in videos often result in image blurring, leading to loss of information and incomplete image transfer.

Innovation Solution

A method and apparatus for generating and playing videos with a 3D effect by segmenting raw images to isolate moving objects, determining occlusion methods for target occlusion images based on the moving object's track, and adding these occlusion images to the raw images to create a modified video with a 3D effect.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Illumination intensity

If blurring is applied to generate 3D visual effect, then depth perception is improved, but image information integrity deteriorates

Engineering Contradiction:
Improvedepth perceptionVSAvoidimage information integrity
Core Design Contradiction:
Illumination intensityVSLoss of information

Solution Approach 1:

The raw image is segmented into foreground image (containing moving object) and background image using image processing techniques. This segmentation allows different processing methods to be applied to different regions, preserving foreground information while creating depth effect through background blurring.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

Different quality levels are applied to different regions of the image. The background region is blurred to create depth perception, while the foreground region containing the moving object maintains original clarity. This local differentiation resolves the contradiction by applying blurring only where it serves the depth effect without compromising important information.

Inventive Principle:
Principle #3Local quality

2Ease of manufacture

If occlusion images are added to create 3D effect, then visual impressiveness is improved, but processing complexity increases

Engineering Contradiction:
Improvevisual impressivenessVSAvoidprocessing complexity
Core Design Contradiction:
Ease of manufactureVSDevice complexity

Solution Approach 1:

The method determines in advance which frames need occlusion images added by analyzing the moving track of objects in the foreground image sequence. Target frames are identified before processing, and occlusion images are pre-selected and prepared based on the moving object's trajectory, reducing real-time processing complexity.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

Occlusion images serve as intermediary elements that are strategically placed between the viewer and the moving object in target frames. These intermediary images create the 3D depth effect by simulating objects at different depths, enhancing visual impressiveness without requiring complex multi-layer compositing.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS20250157493A1Method and Apparatus for Generating Video with 3D Effect, Method and Apparatus for Playing Video with 3D Effect, and Device
Publication Date: 2025.05.15 TENCENT TECHNOLOGY (SHENZHEN) CO LTD
  • US20250157493A1 patent drawing
  • US20250157493A1 patent drawing
  • US20250157493A1 patent drawing

AI summary

A method and an apparatus for generating a video with a three-dimensional (3D) effect, a method and an apparatus for playing a video with a 3D effect, and a device are provided. The method includes: obtaining an original video; segmenting at least one frame of raw image of the original video to obtain a foreground image sequence including a moving object, the foreground image sequence including at least one frame of foreground image; determining, based on the foreground image sequence, a target raw image in which a target occlusion image is to be placed and an occlusion method of the target occlusion image in the target raw image; adding the target occlusion image to the target raw image based on the occlusion method to obtain a final image; and generating a target video with a 3D effect based on the final image and the original video.