Automatic Still Image Extraction from Video Frames

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional camera technologies require significant time and effort to prepare for capturing images, leading to missed opportunities due to the need for manual activation, focusing, and stabilization, especially in dynamic or unexpected situations.

Innovation Solution

An image capture system that automatically captures video and generates still images from a series of frames, allowing for rapid response and image selection based on user intent, movement characteristics, and sensor data, without the need for manual operation or stabilization.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Productivity

If manual activation and preparation operations are used, then device complexity and control precision are maintained, but time consumption increases and productivity decreases

Engineering Contradiction:
Improveimage capture speedVSAvoidpreparation time
Core Design Contradiction:
ProductivityVSLoss of time

Solution Approach 1:

The system performs preliminary actions by continuously capturing video frames and pre-processing them before actual image capture is needed. The processor analyzes motion patterns, detects objects of interest, and prepares still images from video frames in advance, so when capture is triggered, the system is already ready to quickly produce and transmit images without requiring manual preparation time.

Inventive Principle:
Principle #10Preliminary action

2Loss of time

If automatic capture is implemented, then time consumption decreases and productivity increases, but device complexity increases

Engineering Contradiction:
Improvecapture preparation timeVSAvoidsystem complexity
Core Design Contradiction:
Loss of timeVSDevice complexity

Solution Approach 1:

The system achieves multi-functionality by using a single camera to perform both video capture and still image capture functions. The same camera hardware processes both continuous video streaming and selective still image extraction, eliminating the need for separate capture devices and reducing overall system complexity despite the automatic capture capabilities.

Inventive Principle:
Principle #6Universality (Multi-functionality)

3Ease of operation

If manual focusing and stabilization are required, then image quality control is maintained, but operation ease decreases and time consumption increases

Engineering Contradiction:
Improvecapture operation simplicityVSAvoidfocus adjustment time
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The system performs self-service by automatically detecting motion patterns in video frames and selecting appropriate still images without requiring manual focus adjustment or stabilization operations. The processor autonomously analyzes motion vectors, identifies objects of interest, and selects frames that best represent the captured moment, eliminating the need for user intervention in focusing and stabilization processes.

Inventive Principle:
Principle #25Self-service

Data Source

PatentEP3323236B1Image production from video
Publication Date: 2020.09.16 GOOGLE LLC
  • EP3323236B1 patent drawingFigure 1
  • EP3323236B1 patent drawingFigure 2A~2C
  • EP3323236B1 patent drawingFigure 3a

AI summary

Implementations generally relate to producing a still image from a video or series of continuous frames. In some implementations, a method includes receiving the frames that a capture device shot while moving in at least two dimensions. The method further includes analyzing the frames to determine changes of positions of objects in at least two of the frames due to movement of the objects in the scene relative to changes of positions of objects due to the movement of the capture device during the shoot time. The method further includes determining, based at least in part on the variability of the objects, one or more target subjects which the capture device captures during the shoot time. One or more still images are generated from the plurality of frames having at least a portion of the target subject.