Dynamic Subtitle Region Avoidance in Video Frames

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional subtitle displaying methods often result in subtitles blocking key content in video frames, degrading the viewer's experience due to fixed subtitle positions and lack of dynamic adjustment.

Innovation Solution

A method and apparatus that determine subtitle information and identify key regions in video frames using deep learning algorithms to dynamically adjust subtitle display regions, ensuring subtitles do not overlap with key content, thereby improving viewing experience.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of manufacture

If subtitles are displayed in a fixed position of a video frame, then the subtitle display implementation is simple, but the subtitle may block key content in the video frame, degrading viewing experience

Engineering Contradiction:
Improvesubtitle display implementation simplicityVSAvoidsubtitle blocking key content
Core Design Contradiction:
Ease of manufactureVSObject-affected harmful factors

Solution Approach 1:

The patent applies dynamics by transforming the fixed subtitle display position into a dynamic one. The system identifies key regions in video frames and adjusts subtitle positions in real-time to avoid blocking important content. This is achieved through continuous analysis of video content and adaptive repositioning of subtitles, making the display system flexible rather than static.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The patent implements feedback mechanisms by analyzing video content to identify key regions and using this information to adjust subtitle positions. The system continuously monitors the video frame, detects important content areas, and feeds this information back to the subtitle display system to optimize positioning, creating a closed-loop control system.

Inventive Principle:
Principle #23Feedback

2Object-affected harmful factors

If subtitle position is dynamically adjusted to avoid key content, then viewing experience is improved, but the system complexity increases due to need for key region identification

Engineering Contradiction:
Improvesubtitle blocking key contentVSAvoidsystem complexity
Core Design Contradiction:
Object-affected harmful factorsVSDevice complexity

Solution Approach 1:

The patent introduces an intermediary component - a key region identification module - that acts as a mediator between the video content and the subtitle display system. This intermediary analyzes the video content, identifies important regions, and provides this information to the subtitle positioning system, thereby decoupling the complexity of content analysis from the subtitle display logic.

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The patent applies segmentation by dividing the video frame into key regions and non-key regions. The system separately processes these regions, identifying key content areas and then positioning subtitles in non-key areas. This segmentation strategy simplifies the overall problem by breaking it down into manageable parts: content analysis, region classification, and subtitle positioning.

Inventive Principle:
Principle #1Segmentation

3Measurement precision

If deep learning algorithm is used to identify key regions, then key content identification accuracy is improved, but processing time and computational resources increase

Engineering Contradiction:
Improvekey content identification accuracyVSAvoidprocessing time
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The patent applies preliminary action by pre-training deep learning models on large datasets to recognize key content patterns. The models are prepared in advance with learned features and knowledge, enabling them to quickly identify key regions during actual video playback without requiring extensive processing time for each frame. The heavy computational work is done beforehand.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The patent utilizes parameter changes by optimizing deep learning model parameters such as model architecture, learning rates, and processing batch sizes to balance accuracy and speed. The system adjusts these parameters to achieve the desired level of key content identification accuracy while minimizing processing time and computational resource consumption.

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS10645332B2Subtitle displaying method and apparatus
Publication Date: 2020.05.05 ALIBABA GROUP HOLDING LTD
  • US10645332B2 patent drawing
  • US10645332B2 patent drawing
  • US10645332B2 patent drawing

AI summary

A subtitle displaying method including determining subtitle information of a video frame in a video when a play request for the video is received; identifying a key region in the video frame; determining a subtitle display region in a region in the video frame other than the key region; and displaying subtitle content in the subtitle information in the subtitle display region during playing the video. In the example embodiments of the present disclosure, the subtitle display region is determined to be placed in the region other than the key region, such that display content in the key region is prevented from being blocked by subtitles, thus improving viewing experience of viewers.