Video Object Tagging System for Interactive Media Playback

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current technologies lack an efficient method to associate and interact with objects within video content, such as products or faces, to provide additional information or advertisements during playback, limiting user engagement and interactive experiences.

Innovation Solution

A system that allows users to tag objects in video content with references to media items, such as URLs or videos, enabling users to access additional information or advertisements by selecting the tagged objects during playback, using preprocessing to identify and associate these objects across frames.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If video content is made interactive with tagged objects, then user engagement is improved, but device complexity increases

Engineering Contradiction:
Improveuser engagementVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The system performs preprocessing of video content to identify and tag objects before playback. Tags are associated with objects in advance, storing references to additional media content. During playback, the system only needs to retrieve and display pre-tagged content, reducing real-time processing complexity while maintaining interactive capabilities.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent introduces tags as intermediary elements that connect video objects with additional media content. These tags act as mediators, storing reference information that enables interaction without requiring complex direct linking mechanisms between video frames and external content, thereby simplifying the overall system architecture.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Ease of operation

If object tagging is performed during video playback, then user interaction is enabled, but processing time increases

Engineering Contradiction:
Improveuser interactionVSAvoidprocessing time
Core Design Contradiction:
Ease of operationVSLoss of time

Solution Approach 1:

Object identification and tagging are performed during a preprocessing phase before video playback. The system analyzes video frames, identifies objects, and associates tags with these objects in advance. This eliminates the need for real-time object recognition during playback, reducing processing time while maintaining full interactive functionality.

Inventive Principle:
Principle #10Preliminary action

3Loss of information

If multiple tags are associated with video frames, then information availability is improved, but data storage requirements increase

Engineering Contradiction:
Improveinformation availabilityVSAvoiddata storage
Core Design Contradiction:
Loss of informationVSQuantity of substance

Solution Approach 1:

The patent extracts only the essential reference information from media content and stores it as compact tags associated with video objects. Instead of storing complete media files or extensive metadata, the system stores concise reference data that points to additional content, reducing storage requirements while maintaining information availability during interaction.

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS9596516B1Video object tag creation and processing
Publication Date: 2017.03.14 GOOGLE LLC
  • US9596516B1 patent drawing
  • US9596516B1 patent drawing
  • US9596516B1 patent drawing

AI summary

Methods, and systems, including computer programs encoded on computer-readable storage mediums, including a method for presenting a video content item in a first display area; concurrently presenting, with the video content item in the first display area, objects that are displayed during the presentation of the video content item in a second display area, wherein the objects persist in the second display area after the object is no longer displayed during the presentation of the video content item in the first display area; receiving an indication identifying one of the objects presented in the first display area or the second display area; and processing a tag associated with the object, the tag comprising a reference to a media item, wherein the processing comprises: accessing the media item referenced by the tag; and presenting the media item at least partially in the first display area or the second display area.