Video Text Overlay Synchronized with Face Posture

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing video playback technologies fail to effectively display text information, such as bullet-chat screens and captions, in a manner that clearly associates the text with the target object in a video, leading to poor pertinence and flexibility in the display.

Innovation Solution

A method and apparatus for video playback that displays text in a region associated with the face of a target object within a video playback interface, adjusting the text's display posture to match the face posture changes, ensuring the text remains relevant and flexible in its presentation.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If text information is displayed in a specific region of the video picture (e.g., top or bottom), then the display layout is simple and stable, but the text cannot be effectively associated with the target object, reducing display pertinence

Engineering Contradiction:
Improvedisplay pertinenceVSAvoiddisplay layout complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent applies local quality by dynamically positioning text information in different regions of the video picture based on the location of target objects. Instead of using a fixed display region, the text is placed in a region associated with the specific target object it describes, making the display adapt to local content requirements and improving pertinence without excessive complexity

Inventive Principle:
Principle #3Local quality

Solution Approach 2:

The patent implements dynamics by making the text display position and posture change dynamically according to the target object's position and orientation in the video. The text follows the target object's movements and adjusts its angle to maintain proper association, transforming a static display into an adaptive one that responds to video content changes

Inventive Principle:
Principle #15Dynamics

2Adaptability or versatility

If text is displayed in a fixed position, then the display structure is simple, but the text cannot adapt to changes in target object position and posture, reducing display flexibility

Engineering Contradiction:
Improvedisplay flexibilityVSAvoiddisplay control complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent makes the text display position and posture dynamic by tracking the target object's position and orientation changes in real-time. The text automatically adjusts its location and angle to maintain proper association with the target object, providing display flexibility that adapts to varying video content without requiring complex manual control

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The system implements self-service by automatically detecting target objects, determining their positions and postures, and accordingly adjusting the text display parameters without external intervention. The display system serves itself by using video content analysis to control its own presentation, reducing the need for complex external control mechanisms

Inventive Principle:
Principle #25Self-service

3Productivity

If text display posture is fixed, then the rendering is simple and fast, but the text does not change with face posture, reducing user understanding and engagement

Engineering Contradiction:
Improveprocessing speedVSAvoidassociation information
Core Design Contradiction:
ProductivityVSLoss of information

Solution Approach 1:

The patent applies dynamics by making the text posture change dynamically in response to the target object's face posture changes. The text orientation and angle are continuously adjusted to match the face posture, maintaining proper visual association and preventing loss of contextual information while using efficient rendering techniques to maintain processing speed

Inventive Principle:
Principle #15Dynamics

Data Source

PatentUS12034996B2Video playing method, apparatus and device, storage medium, and program product
Publication Date: 2024.07.09 TENCENT TECHNOLOGY (SHENZHEN) CO LTD
  • US12034996B2 patent drawing
  • US12034996B2 patent drawing
  • US12034996B2 patent drawing

AI summary

This application provides a video playing method performed by a computer device. The method includes: playing a target video in a playing interface; when a video picture of the target video comprises a target object and there is a text associated with the target object, displaying the text in a display region associated with a face posture of the target object in a process of playing the target video; and adjusting a display posture of the text with the face posture when the face posture of the target object changes in a process of displaying the text.