Video Caption Re-overlay for Small-Screen Readability
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Video captions become too small to be readable when high-resolution videos are downscaled for display on small-screen devices like mobile phones, as existing methods focus on visual enhancement rather than text size adjustment.
Innovation Solution
A system and method that detects caption text using computer vision algorithms, crops and separately resizes it, and overlays it back onto downscaled video frames with a smaller downscaling ratio, ensuring proportionally larger and more visible captions through re-layout and post-processing to merge seamlessly with the background.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Area of stationary object
If video is downscaled for small-screen devices, then the video fits the display, but the caption text becomes too small to be readable
Solution Approach 1:
The patent segments the video processing into distinct components: video frame downscaling and caption text processing. The caption detection block identifies caption regions, which are then separately processed through reformatting and re-overlay operations, allowing independent optimization of text readability while maintaining overall video fit.
Solution Approach 2:
The patent applies different processing qualities to different parts of the video. Caption text regions receive enhanced treatment with separate reformatting that preserves or enlarges text size, while the rest of the video frame is downscaled to fit the small screen. This local differentiation ensures captions remain readable while the video adapts to the display.
2Manufacturing precision
If caption text is kept at original size, then readability is maintained, but the caption does not scale proportionally with the video
Solution Approach 1:
The patent implements dynamic processing where the caption reformating block adjusts caption size and positioning based on the downscaling ratio applied to the video. The system calculates appropriate caption dimensions that maintain readability while adapting to different video resolutions and aspect ratios, enabling versatile adaptation across various small-screen devices.
Solution Approach 2:
The patent changes key parameters of the caption text including size, position, and formatting based on the video downscaling operation. The reformatting process modifies caption parameters to ensure they remain proportionally larger than the video content while adapting to the target display dimensions, resolving the conflict between maintainability and adaptability.
3Manufacturing precision
If caption text is processed separately from video, then text size can be optimized, but the processing complexity increases
Solution Approach 1:
The patent divides the processing system into specialized blocks: video reformatting block, caption detection block, caption reformatting block, and re-overlay block. This segmentation allows each component to be optimized for its specific function while working together in a coordinated pipeline, managing complexity through modular design.
Solution Approach 2:
The patent introduces intermediate processing steps where detected captions are extracted, reformatted, and prepared before being overlaid back onto the downscaled video. This intermediary processing stage enables optimized text handling while maintaining integration with the overall video adaptation workflow, balancing complexity with functionality.
Data Source
AI summary
In accordance with an embodiment, a method of processing an electronic image having caption text includes receiving the electronic source image, detecting the caption text in the electronic source image, reformatting the electronic source image, reformatting the caption text, and overlaying the reformatted caption text on the reformatted electronic image to form a resultant image.


