Real-Time Video Text Translation in Dedicated Display Regions
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing video translation methods interrupt the normal watching experience by pausing the video to present translated text, leading to poor user experience.
Innovation Solution
A method and apparatus that convert text in a video frame to a target language in real-time and present it in a designated region, either a pre-defined text box or dynamically determined based on the frame's content, allowing seamless translation without interrupting the video playback.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Loss of information
If translated text is presented during video playback, then translation availability is improved, but video viewing continuity deteriorates
Solution Approach 1:
The patent segments the video frame into multiple regions, including a text box region for translation display and a picture content region for video viewing. This spatial segmentation allows translation information to be presented without interrupting the video playback, as the text is displayed in a dedicated area that does not block the video content.
Solution Approach 2:
The patent transitions from temporal dimension (pausing video to show translation) to spatial dimension (displaying translation in a separate region during continuous playback). By using a text box region positioned within the video frame, the system presents translation information in another spatial dimension, allowing both video and translation to coexist without interruption.
2Loss of information
If translated text is displayed in the video frame, then translation visibility is improved, but video content obstruction increases
Solution Approach 1:
The patent applies local quality by creating a specific text box region with distinct visual characteristics (such as background color, border, or transparency) that differentiates it from the video content region. This localized styling ensures that the translation text is visible and distinguishable while maintaining a clear separation from the video content, preventing obstruction.
Solution Approach 2:
The text box region acts as an intermediary element between the translation text and the video content. It provides a dedicated space that accommodates the translation display without directly overlapping with or blocking the video picture content, thus mediating between translation visibility and video content preservation.
Data Source
AI summary
Provided are a video processing method, an electronic device and a storage medium. The method includes steps described below. In a process of playing a target video, to-be-converted text in a to-be-processed video frame is determined in response to a triggering operation by a target user on the target video; and the to-be-converted text is converted into translated text of a target language type, and the translated text is presented in a target region in the to-be-processed video frame; where the target region includes a text box region to which the to-be-converted text belongs, or the target region is dynamically determined based on a picture content of the to-be-processed video frame.


