Real-Time Video Stream Object Modification

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Current telecommunications devices face challenges in modifying video streams in real-time during communication sessions, requiring physical manipulation of the device, which limits the ability to enhance video communications and modify images within the stream effectively.

Innovation Solution

The system identifies objects, such as faces, within a video stream and applies scaled graphical representations, like virtual glasses, by analyzing two-dimensional coordinates and tracking movement, allowing for real-time modification and presentation of the video stream within a user interface.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If physical manipulation of the device is used to perform operations, then the device can be operated, but the ability to enhance video communications and modify images within the stream effectively is limited

Engineering Contradiction:
Improvedevice operationVSAvoidvideo stream modification capability
Core Design Contradiction:
Ease of operationVSAdaptability or versatility

Solution Approach 1:

The patent replaces physical mechanical manipulation of the device with automated software-based object detection and modification systems. The system automatically identifies objects in video streams using image recognition algorithms and applies graphical representations without requiring users to physically manipulate the device, thereby enhancing video communication capabilities while maintaining ease of operation.

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Adaptability or versatility

If real-time modification of video streams is implemented, then user interaction is enhanced, but device complexity increases

Engineering Contradiction:
Improvevideo stream modification capabilityVSAvoidsystem complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent introduces an intermediary processing layer that sits between the video capture and display components. This intermediary system handles object detection, tracking, and graphical representation application, effectively managing the complexity of real-time video modification while preserving the enhanced user interaction capabilities. The intermediary layer abstracts the complex processing requirements from the core device operations.

Inventive Principle:
Principle #24Intermediary (Mediator)

3Adaptability or versatility

If multiple objects are modified simultaneously within a video stream, then user interaction is enhanced, but processing requirements increase

Engineering Contradiction:
Improvemulti-object modification capabilityVSAvoidprocessing power
Core Design Contradiction:
Adaptability or versatilityVSPower

Solution Approach 1:

The patent segments the video stream processing into distinct phases: object detection, object tracking, and graphical representation application. By dividing the processing of multiple objects into these separate segments that can be handled independently and concurrently, the system enhances multi-object modification capabilities while managing processing power requirements through efficient resource allocation across different processing stages.

Inventive Principle:
Principle #1Segmentation

Data Source

PatentUS11551425B2Modifying multiple objects within a video stream
Publication Date: 2023.01.10 SNAP INC
  • US11551425B2 patent drawing
  • US11551425B2 patent drawing
  • US11551425B2 patent drawing

AI summary

Systems, devices, media, and methods are presented for presentation of modified objects within a video stream. The systems and methods receive a set of images within a video stream and identify at least a portion of a face in a first subset of images. The systems and methods determine face characteristics by analyzing the portion of the face in the first subset of images. The systems and methods apply a graphical representation of glasses to the face based on the face characteristics and cause presentation of a modified video stream including the portion of the face with the graphical representation of the glasses in a second subset of images of the set of images while receiving the video stream.