Media Stream Scene Location via Closed Caption Hashing

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional systems fail to accurately determine the exact location of a scene in a media stream, especially when extraneous content is present or the stream is edited, making it difficult for viewers to post comments or updates specific to a scene, and limiting opportunities for targeted advertisements.

Innovation Solution

A system that extracts closed caption data from the media stream, calculates the order of duplication of caption data strings, and generates hash values to uniquely identify the current location, allowing precise determination of the scene position and enabling scene-specific social media updates and targeted advertisements.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If conventional systems are used to determine scene location in media streams, then the system is simple to operate, but the measurement precision of scene location is insufficient

Engineering Contradiction:
Improvescene location precisionVSAvoidsystem complexity
Core Design Contradiction:
Measurement precisionVSDevice complexity

Solution Approach 1:

The patent introduces closed caption data as an intermediary element to bridge the gap between media stream content and scene location identification. By extracting caption data strings and using them as reference markers, the system achieves precise scene location without requiring complex video analysis algorithms. The caption data serves as a mediator that simplifies the measurement process while improving precision.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The patent replaces complex video frame analysis mechanisms with a text-based caption data processing system. Instead of analyzing visual content to determine scene location, the system uses caption data strings as identifiers to locate scenes precisely. This substitution of mechanical/video-based analysis with data-processing approaches simplifies the system while improving measurement precision.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Adaptability or versatility

If the media stream includes extraneous content or editing, then the content is more flexible and adaptable, but the reliability of scene location determination deteriorates

Engineering Contradiction:
Improvecontent flexibilityVSAvoidscene location determination reliability
Core Design Contradiction:
Adaptability or versatilityVSReliability

Solution Approach 1:

The patent extracts closed caption data strings from the media stream as independent reference markers. By separating and extracting the caption data from the overall media stream structure, the system can reliably identify scene locations even when extraneous content or editing is present. The extracted caption strings serve as stable anchors that maintain their identifying function regardless of surrounding content modifications.

Inventive Principle:
Principle #2Taking out (Extraction)

Solution Approach 2:

The patent performs preliminary extraction and indexing of closed caption data strings before the actual scene location determination. By pre-processing the caption data to create a reference database, the system prepares the information needed for reliable scene identification in advance. This preliminary action ensures that when scene location determination is needed, the system can reliably match current stream captions against the pre-indexed references.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS9215496B1Determining the location of a point of interest in a media stream that includes caption data
Publication Date: 2015.12.15 GOOGLE TECHNOLOGY HOLDINGS LLC
  • US9215496B1 patent drawing
  • US9215496B1 patent drawing
  • US9215496B1 patent drawing

AI summary

A method and computing device for determining the location of a point of interest in a media stream. The method receives an order of duplication for a media stream and a sequence of caption data strings associated with the media stream. The method computes a hash value for a selected string in the sequence. The hash value representing the selected string, and a number of strings in the sequence that immediately precede the selected string, where the order of duplication determines the number of strings. The method receives a media stream time for the selected string based on the hash value, and determines a time at a point of interest in the media stream relative to the media stream time for the selected string.