Multi-Channel Video Compilation via Audio Fingerprint Alignment

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current video-sharing platforms like YouTube and Vimeo do not provide convenient mechanisms to present distinct but related video segments in a continuous video stream, due to storage and upload limitations, making it difficult for users to watch events such as concerts or sporting events without interruptions.

Innovation Solution

A system and method for compiling and playing multi-channel videos by aggregating distinct video segments from different sources into a continuous stream, using a video aggregation interface, alignment engine, and multi-channel video player to align and switch between video segments automatically or manually, providing a seamless viewing experience with additional perspectives.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Quantity of substance

If multiple video segments are uploaded to video-sharing websites, then the coverage of events is improved, but the convenience of continuous playback deteriorates due to upload limits and storage constraints

Engineering Contradiction:
Improvevideo segment coverageVSAvoidcontinuous playback convenience
Core Design Contradiction:
Quantity of substanceVSEase of operation

Solution Approach 1:

The system divides a complete event recording into multiple video segments captured by different cameras or recording devices. Each segment is independently uploaded and stored, then automatically reassembled through audio fingerprint correlation to form a continuous multi-channel video playback experience.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces an intermediary processing system that uses audio fingerprint correlation to match and align video segments from different sources. This intermediary layer automatically synchronizes segments based on their audio content, enabling seamless continuous playback without requiring manual intervention or breaking the viewing experience.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Quantity of substance

If video segments are divided due to storage limits, then the storage requirement per device is reduced, but the viewing experience is fragmented requiring manual selection of multiple segments

Engineering Contradiction:
Improvestorage capacity per deviceVSAvoidtime for manual segment selection
Core Design Contradiction:
Quantity of substanceVSLoss of time

Solution Approach 1:

The system performs preliminary actions by pre-processing video segments to extract audio fingerprints and store them in a correlated structure. This preparation work is done in advance so that when playback is requested, the system can automatically assemble the correct sequence of segments without requiring real-time manual selection or user intervention.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent implements a feedback mechanism where the system automatically correlates audio fingerprints from different video segments, identifies their temporal relationships, and uses this information to automatically sequence and playback segments in the correct order, eliminating the need for manual segment selection by the user.

Inventive Principle:
Principle #23Feedback

3Quantity of substance

If multiple video sources are aggregated, then the completeness of event coverage is improved, but the complexity of aligning and synchronizing segments increases

Engineering Contradiction:
Improveevent coverage completenessVSAvoidalignment and synchronization complexity
Core Design Contradiction:
Quantity of substanceVSDevice complexity

Solution Approach 1:

The patent replaces complex mechanical or manual alignment processes with an automated audio fingerprint correlation system. Instead of manually synchronizing video segments based on visual cues or timestamps, the system uses digital audio fingerprinting technology to automatically identify and align segments based on their audio content, significantly reducing the complexity of the alignment process.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

Solution Approach 2:

The system changes the parameter used for alignment from visual or temporal metadata to audio fingerprint correlation. By transforming the alignment problem into an audio pattern matching problem, the system can automatically handle multiple video sources with different timestamps and synchronization references, simplifying the overall alignment complexity.

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS9344606B2System and method for compiling and playing a multi-channel video
Publication Date: 2016.05.17 RADICAL URBAN LLC
  • US9344606B2 patent drawing
  • US9344606B2 patent drawing
  • US9344606B2 patent drawing

AI summary

A system and method including retrieving a multi-channel video file with a first video segment and a second video segment; assigning a rank to the first video segment and second video segment; and rendering at least the first and second video segment in a player interface, synchronized to a timeline of the multi-channel video file, and comprising: playing the first video segment in an active stream, progressing a timeline of the multi-channel video file when video is played in the active stream, when the timeline progresses to a synchronized time of the second video segment, playing the second video segment in the active stream and displaying the first video segment as a selectable video channel if the rank of the second video segment is greater than the rank of the first video segment, and otherwise displaying the second video segment as a selectable video channel.