3D Image Generation from Single Video and Photo

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing technologies require multiple images or videos to provide 3D images, which is inconvenient for users.

Innovation Solution

An image provision device and method that generates similar images corresponding to a viewer's gaze position using a first image from an original video and a second image from an original image, without requiring multiple images or videos.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If multiple images or videos are prepared to provide 3D images according to viewer's gaze position, then the 3D image provision capability is improved, but the ease of operation deteriorates due to considerable inconvenience of use

Engineering Contradiction:
Improve3D image provision capabilityVSAvoidease of operation
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The patent uses image synthesis technology to generate synthetic images that replicate the appearance of multiple gaze-position images. Instead of requiring users to prepare multiple actual images or videos, the system creates virtual copies through deep learning models that synthesize appropriate images based on a single input image and gaze position data, thereby maintaining 3D image provision capability while dramatically improving ease of operation

Inventive Principle:
Principle #26Copying

Solution Approach 2:

The patent changes the fundamental parameter from requiring multiple source images/videos to using only a single source image. By utilizing gaze position information as a controlling parameter and applying image synthesis algorithms, the system transforms one input image into multiple gaze-specific output images, resolving the contradiction between adaptability and ease of operation

Inventive Principle:
Principle #35Parameter changes

2Manufacturing precision

If multiple images or videos are prepared for different gaze positions, then the stereoscopic image quality is improved, but the device complexity increases

Engineering Contradiction:
Improvestereoscopic image qualityVSAvoiddevice complexity
Core Design Contradiction:
Manufacturing precisionVSDevice complexity

Solution Approach 1:

The patent replaces the mechanical approach of physically preparing and storing multiple images or videos with a computational image synthesis system. Instead of requiring a complex library of pre-prepared multi-gaze images, the system uses algorithms and deep learning models to generate images on-demand from a single source image and gaze position data, thereby reducing device complexity while maintaining stereoscopic image quality

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The patent performs preliminary action by extracting facial features and preparing a single source image in advance, then uses this pre-processed data along with gaze position information to generate multiple stereoscopic images. This preliminary preparation simplifies the overall system complexity by avoiding the need to manage and process multiple source images simultaneously

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS20250159129A1Image provision device and image provision method
Publication Date: 2025.05.15 SK HYNIX INC
  • US20250159129A1 patent drawing
  • US20250159129A1 patent drawing
  • US20250159129A1 patent drawing

AI summary

An image provision device and an image provision method are disclosed. The image provision device includes a first pre-processor configured to generate a first image by detecting a first face region from each frame of an original video, a second pre-processor configured to generate a second image by detecting a second face region from an original image, and a similar image generator configured to generate similar images respectively corresponding to a viewer's gaze positions based on the first image and the second image.