Key Frame Extraction for Collaborative Video Navigation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Navigating recorded media content from electronic meetings and webinars, which often include static non-video content like slide presentations, is inefficient due to the presence of duplicate frames and unchanging content, requiring exhaustive manual searching using a video scrubber.

Innovation Solution

The technique involves extracting key frames from media content at a predetermined rate, modifying frames to remove non-video areas, deduplicating frames by removing unchanging frames, and differentiating frames to identify slide-type frames, which are then recorded and displayed as clickable thumbnail images for easy access.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If a video scrubber with progress bar is used to navigate recorded media content, then the media player application can play back recorded video, but users must manually and exhaustively search along the progress bar to access static non-video content such as slides

Engineering Contradiction:
Improveease of navigationVSAvoidtime to access desired content
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

The patent segments the continuous video recording into discrete key frames that represent significant moments or slides. Instead of navigating through the entire continuous video stream, the system extracts and indexes individual key frames at predetermined rates, allowing users to jump directly to specific content segments rather than manually searching through the entire recording.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The system performs preliminary extraction and organization of key frames before the user needs to navigate the content. By pre-processing the video recording to identify and extract key frames representing slides and significant moments, the system prepares an optimized navigation structure in advance, eliminating the need for users to manually search through the entire video during playback.

Inventive Principle:
Principle #10Preliminary action

2Loss of information

If all frames from the video recording are retained including duplicate and unchanging frames, then complete video content is preserved, but the quantity of frames increases making navigation more difficult

Engineering Contradiction:
Improvecompleteness of contentVSAvoidnumber of frames
Core Design Contradiction:
Loss of informationVSQuantity of substance

Solution Approach 1:

The patent extracts only the essential key frames from the complete video recording, removing duplicate and unchanging frames. By applying deduplication logic that identifies and eliminates consecutive duplicate frames and frames that haven't changed significantly, the system retains only the necessary frames that represent unique content moments, slides, or significant changes in the presentation.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The system changes the parameter of frame selection from retaining all frames to selecting only key frames based on specific criteria such as frame difference thresholds, predetermined extraction rates, and deduplication rules. This parameter change transforms the frame set from a complete but redundant collection to a optimized subset that preserves information while reducing quantity.

Inventive Principle:
Principle #35Parameter changes

3Productivity

If key frames are extracted and displayed as clickable thumbnails, then users can quickly access desired slide content, but additional processing steps are required compared to simple video playback

Engineering Contradiction:
Improvespeed of content accessVSAvoidcomplexity of media player
Core Design Contradiction:
ProductivityVSDevice complexity

Solution Approach 1:

The patent creates thumbnail copies of the extracted key frames and displays them as a navigable gallery or timeline. Instead of requiring users to watch the entire video or manually scrub through it, the system generates visual copies (thumbnails) of key frames that users can click to quickly navigate to specific content, thereby speeding up access while keeping the interface intuitive.

Inventive Principle:
Principle #26Copying

Data Source

PatentUS10990828B2Key frame extraction, recording, and navigation in collaborative video presentations
Publication Date: 2021.04.27 GOTO GRP INC
  • US10990828B2 patent drawing
  • US10990828B2 patent drawing
  • US10990828B2 patent drawing

AI summary

Techniques for performing key frame extraction, recording, and navigation in collaborative video presentations. The techniques include extracting a plurality of frames from media content at a predetermined rate, removing frame areas that do not correspond to a screen area for displaying electronic meeting/webinar content, de-duplicating the plurality of frames, identifying frames that correspond to the “slide type” or similar type of frames, and extracting key frames from the slide type of frames. The key frames can be recorded in a slide deck or other similar collection of key frames, as well as displayed as clickable thumbnails in a UI. By clicking or otherwise selecting a thumbnail representation of a selected key frame in the UI, or clicking-and-dragging a handle of a key frame locator bar to navigate the thumbnails to the selected key frame, users can quickly and more efficiently access desired slide presentation content from an electronic meeting/webinar.