Video Caption Re-overlay for Small-Screen Readability

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Video captions become too small to be readable when high-resolution videos are downscaled for display on small-screen devices like mobile phones, as existing methods focus on visual enhancement rather than text size adjustment.

Innovation Solution

A system and method that detects caption text using computer vision algorithms, crops and separately resizes it, and overlays it back onto downscaled video frames with a smaller downscaling ratio, ensuring proportionally larger and more visible captions through re-layout and post-processing to merge seamlessly with the background.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Area of stationary object

If video is downscaled for small-screen devices, then the video fits the display, but the caption text becomes too small to be readable

Engineering Contradiction:
Improvevideo display areaVSAvoidcaption text readability
Core Design Contradiction:
Area of stationary objectVSManufacturing precision

Solution Approach 1:

The patent segments the video processing into distinct components: video frame downscaling and caption text processing. The caption detection block identifies caption regions, which are then separately processed through reformatting and re-overlay operations, allowing independent optimization of text readability while maintaining overall video fit.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent applies different processing qualities to different parts of the video. Caption text regions receive enhanced treatment with separate reformatting that preserves or enlarges text size, while the rest of the video frame is downscaled to fit the small screen. This local differentiation ensures captions remain readable while the video adapts to the display.

Inventive Principle:
Principle #3Local quality

2Manufacturing precision

If caption text is kept at original size, then readability is maintained, but the caption does not scale proportionally with the video

Engineering Contradiction:
Improvecaption text readabilityVSAvoidvideo adaptation
Core Design Contradiction:
Manufacturing precisionVSAdaptability or versatility

Solution Approach 1:

The patent implements dynamic processing where the caption reformating block adjusts caption size and positioning based on the downscaling ratio applied to the video. The system calculates appropriate caption dimensions that maintain readability while adapting to different video resolutions and aspect ratios, enabling versatile adaptation across various small-screen devices.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The patent changes key parameters of the caption text including size, position, and formatting based on the video downscaling operation. The reformatting process modifies caption parameters to ensure they remain proportionally larger than the video content while adapting to the target display dimensions, resolving the conflict between maintainability and adaptability.

Inventive Principle:
Principle #35Parameter changes

3Manufacturing precision

If caption text is processed separately from video, then text size can be optimized, but the processing complexity increases

Engineering Contradiction:
Improvecaption text readabilityVSAvoidprocessing system complexity
Core Design Contradiction:
Manufacturing precisionVSDevice complexity

Solution Approach 1:

The patent divides the processing system into specialized blocks: video reformatting block, caption detection block, caption reformatting block, and re-overlay block. This segmentation allows each component to be optimized for its specific function while working together in a coordinated pipeline, managing complexity through modular design.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces intermediate processing steps where detected captions are extracted, reformatted, and prepared before being overlaid back onto the downscaled video. This intermediary processing stage enables optimized text handling while maintaining integration with the overall video adaptation workflow, balancing complexity with functionality.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS8754984B2System and method for video caption re-overlaying for video adaptation and retargeting
Publication Date: 2014.06.17 FUTUREWEI TECHNOLOGIES INC
  • US8754984B2 patent drawing
  • US8754984B2 patent drawing
  • US8754984B2 patent drawing

AI summary

In accordance with an embodiment, a method of processing an electronic image having caption text includes receiving the electronic source image, detecting the caption text in the electronic source image, reformatting the electronic source image, reformatting the caption text, and overlaying the reformatted caption text on the reformatted electronic image to form a resultant image.