Dynamic Subtitle Region Avoidance in Video Frames
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional subtitle displaying methods often result in subtitles blocking key content in video frames, degrading the viewer's experience due to fixed subtitle positions and lack of dynamic adjustment.
Innovation Solution
A method and apparatus that determine subtitle information and identify key regions in video frames using deep learning algorithms to dynamically adjust subtitle display regions, ensuring subtitles do not overlap with key content, thereby improving viewing experience.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Ease of manufacture
If subtitles are displayed in a fixed position of a video frame, then the subtitle display implementation is simple, but the subtitle may block key content in the video frame, degrading viewing experience
Solution Approach 1:
The patent applies dynamics by transforming the fixed subtitle display position into a dynamic one. The system identifies key regions in video frames and adjusts subtitle positions in real-time to avoid blocking important content. This is achieved through continuous analysis of video content and adaptive repositioning of subtitles, making the display system flexible rather than static.
Solution Approach 2:
The patent implements feedback mechanisms by analyzing video content to identify key regions and using this information to adjust subtitle positions. The system continuously monitors the video frame, detects important content areas, and feeds this information back to the subtitle display system to optimize positioning, creating a closed-loop control system.
2Object-affected harmful factors
If subtitle position is dynamically adjusted to avoid key content, then viewing experience is improved, but the system complexity increases due to need for key region identification
Solution Approach 1:
The patent introduces an intermediary component - a key region identification module - that acts as a mediator between the video content and the subtitle display system. This intermediary analyzes the video content, identifies important regions, and provides this information to the subtitle positioning system, thereby decoupling the complexity of content analysis from the subtitle display logic.
Solution Approach 2:
The patent applies segmentation by dividing the video frame into key regions and non-key regions. The system separately processes these regions, identifying key content areas and then positioning subtitles in non-key areas. This segmentation strategy simplifies the overall problem by breaking it down into manageable parts: content analysis, region classification, and subtitle positioning.
3Measurement precision
If deep learning algorithm is used to identify key regions, then key content identification accuracy is improved, but processing time and computational resources increase
Solution Approach 1:
The patent applies preliminary action by pre-training deep learning models on large datasets to recognize key content patterns. The models are prepared in advance with learned features and knowledge, enabling them to quickly identify key regions during actual video playback without requiring extensive processing time for each frame. The heavy computational work is done beforehand.
Solution Approach 2:
The patent utilizes parameter changes by optimizing deep learning model parameters such as model architecture, learning rates, and processing batch sizes to balance accuracy and speed. The system adjusts these parameters to achieve the desired level of key content identification accuracy while minimizing processing time and computational resource consumption.
Data Source
AI summary
A subtitle displaying method including determining subtitle information of a video frame in a video when a play request for the video is received; identifying a key region in the video frame; determining a subtitle display region in a region in the video frame other than the key region; and displaying subtitle content in the subtitle information in the subtitle display region during playing the video. In the example embodiments of the present disclosure, the subtitle display region is determined to be placed in the region other than the key region, such that display content in the key region is prevented from being blocked by subtitles, thus improving viewing experience of viewers.


