3D Image Generation from Single Video and Photo
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing technologies require multiple images or videos to provide 3D images, which is inconvenient for users.
Innovation Solution
An image provision device and method that generates similar images corresponding to a viewer's gaze position using a first image from an original video and a second image from an original image, without requiring multiple images or videos.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If multiple images or videos are prepared to provide 3D images according to viewer's gaze position, then the 3D image provision capability is improved, but the ease of operation deteriorates due to considerable inconvenience of use
Solution Approach 1:
The patent uses image synthesis technology to generate synthetic images that replicate the appearance of multiple gaze-position images. Instead of requiring users to prepare multiple actual images or videos, the system creates virtual copies through deep learning models that synthesize appropriate images based on a single input image and gaze position data, thereby maintaining 3D image provision capability while dramatically improving ease of operation
Solution Approach 2:
The patent changes the fundamental parameter from requiring multiple source images/videos to using only a single source image. By utilizing gaze position information as a controlling parameter and applying image synthesis algorithms, the system transforms one input image into multiple gaze-specific output images, resolving the contradiction between adaptability and ease of operation
2Manufacturing precision
If multiple images or videos are prepared for different gaze positions, then the stereoscopic image quality is improved, but the device complexity increases
Solution Approach 1:
The patent replaces the mechanical approach of physically preparing and storing multiple images or videos with a computational image synthesis system. Instead of requiring a complex library of pre-prepared multi-gaze images, the system uses algorithms and deep learning models to generate images on-demand from a single source image and gaze position data, thereby reducing device complexity while maintaining stereoscopic image quality
Solution Approach 2:
The patent performs preliminary action by extracting facial features and preparing a single source image in advance, then uses this pre-processed data along with gaze position information to generate multiple stereoscopic images. This preliminary preparation simplifies the overall system complexity by avoiding the need to manage and process multiple source images simultaneously
Data Source
AI summary
An image provision device and an image provision method are disclosed. The image provision device includes a first pre-processor configured to generate a first image by detecting a first face region from each frame of an original video, a second pre-processor configured to generate a second image by detecting a second face region from an original image, and a similar image generator configured to generate similar images respectively corresponding to a viewer's gaze positions based on the first image and the second image.


