Automated Content Recognition for Interactive Media Overlay

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Smart televisions struggle to leverage interactive functionality when presenting non-interactive content, as this type of content does not include instructions for executing specific functions of the smart television.

Innovation Solution

The system employs automated content recognition services to identify video segments being displayed, allowing for the transmission of notifications to the display device requesting audio input. This facilitates the detection of audio segments and the presentation of associated objects, enabling interactive functionality with non-interactive media.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If non-interactive media is presented, then media playback function is maintained, but interactive functionality cannot be leveraged

Engineering Contradiction:
Improveinteractive functionalityVSAvoidmedia playback compatibility
Core Design Contradiction:
Adaptability or versatilityVSReliability

Solution Approach 1:

The patent introduces an intermediary system consisting of a media device, automated content recognition service, and processing system that acts as a mediator between non-interactive media and interactive functionality. The media device receives non-interactive media, the processing system identifies media segments through automated content recognition, and conditional interactive functions are executed based on user audio input during playback, thus enabling adaptability without compromising media playback compatibility

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The system changes the state of non-interactive media from passive to interactive by detecting user audio input parameters during media playback. When audio input matching specific criteria is detected, the system transitions to executing conditional interactive functions associated with identified media segments, thereby transforming the media consumption experience without altering the original media content

Inventive Principle:
Principle #35Parameter changes

2Adaptability or versatility

If interactive functions are added to non-interactive media, then media adaptability improves, but system complexity increases

Engineering Contradiction:
Improvemedia interactivityVSAvoidsystem architecture
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent segments the media content into discrete media segments identified by the automated content recognition service. Each segment can be independently processed and associated with specific interactive functions. This segmentation allows the complex interactive functionality to be broken down into manageable units that can be conditionally executed based on user input, reducing the complexity burden on the overall system architecture

Inventive Principle:
Principle #1Segmentation

3Measurement precision

If automated content recognition is used, then media identification accuracy improves, but processing time increases

Engineering Contradiction:
Improvemedia segment identificationVSAvoidprocessing delay
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The system performs preliminary identification of media segments using automated content recognition before executing interactive functions. By pre-identifying media segments and preparing associated conditional functions in advance, the system minimizes processing delays during actual media playback, as the identification work is completed beforehand rather than in real-time during user interaction

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS20250140252A1Systems and methods for voice-based trigger for supplemental content
Publication Date: 2025.05.01 VIZIO INC
  • US20250140252A1 patent drawing
  • US20250140252A1 patent drawing
  • US20250140252A1 patent drawing

AI summary

A media device may perform contextual processing based on media segments that are presented. The media device may receive an identification of a video segment being displayed by a display device from an automated content recognition service. The media device may transmit a notification including information associated with the video segment and a request for audio input. Upon detecting one or more audio segments associated with the notification, the media device may facilitate a presentation of an object associated with the video segment.