Dynamic Collage Generation from Mobile Video Key Frames

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Users face challenges in efficiently organizing, retrieving, and sharing user-generated video content on mobile devices due to the complex and unstructured nature of video sequences, making it difficult to access and manipulate video content for creating personalized products like collages and highlight videos.

Innovation Solution

A system and method that automatically extracts key frames from video sequences using advanced video processing algorithms, allowing users to interactively select and dynamically adjust collage outputs on mobile devices, enabling easy creation and sharing of personalized content by leveraging key frames as indexing points for video summarization and collage generation.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If users manually sort and select video frames to create collages, then they can have full control over content selection, but the process becomes time-consuming and complex

Engineering Contradiction:
Improveease of video content manipulationVSAvoidtime for manual video editing and sorting
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The system automatically extracts key frames from video sequences before the user needs to create collages. By pre-processing the video content and identifying important frames in advance, the system eliminates the need for users to manually review entire video sequences, significantly reducing the time and effort required for collage creation while maintaining user control over the final selection.

Inventive Principle:
Principle #10Preliminary action

2Productivity

If the system automatically extracts key frames from video sequences, then collage creation becomes faster and easier, but users lose manual control over frame selection

Engineering Contradiction:
Improvecollage creation speedVSAvoiduser control over content selection
Core Design Contradiction:
ProductivityVSEase of operation

Solution Approach 1:

The system provides interactive feedback mechanisms that allow users to review automatically extracted key frames, accept or reject selections, and adjust the extraction parameters. This feedback loop enables users to maintain control over content selection while benefiting from the speed of automated extraction, as they can quickly approve or modify the system's choices rather than manually sorting through all frames.

Inventive Principle:
Principle #23Feedback

3Area of stationary object

If users store all video files directly on mobile devices, then they have easy access to content, but device storage becomes full and management becomes difficult

Engineering Contradiction:
Improvedevice storage capacityVSAvoidvideo content access and retrieval
Core Design Contradiction:
Area of stationary objectVSEase of operation

Solution Approach 1:

The system extracts essential information from video files in the form of key frames and metadata, storing only these extracted elements on the mobile device while keeping the full video files in cloud storage or other external locations. This extraction approach significantly reduces the storage space required on mobile devices while maintaining easy access to video content through the stored key frames that serve as representatives of the full videos.

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS11880918B2Method for dynamic creation of collages from mobile video
Publication Date: 2024.01.23 KODAK ALARIS LLC
  • US11880918B2 patent drawing
  • US11880918B2 patent drawing
  • US11880918B2 patent drawing

AI summary

A method and system for automated creation of collages from a video sequence is provided. The system and method dynamically extracts key frames from a video sequence, which are used to create the collages shown on a display. By changing the position of cursors, a new set of key frames is extracted and the content of the collage will change correspondingly reflecting the changes in the key frames selection. The method also includes user interface elements to allow a user to change the layout design (e.g., rotate, overlap, white space, image scaling) of the video frames on the collage. In addition, the method includes different ways for browsing the video sequence by using automatically detected semantic concepts in various frames as indexing points. Further, the method includes defining swiping attributes based on motion characteristics (e.g., zoom/pan, motion of object), audio activity, or facial expression for browsing the video sequence.