Pre-rendered Caption Overlay for Video Uniformity

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing video caption rendering technologies face challenges in maintaining uniformity across different user devices due to varying software renderers and the inability to render certain captions, such as vertical text or non-standard glyphs, and require changes to the video file when updating or adding new languages, which complicates debugging and editing.

Innovation Solution

A method where pre-rendered visual representations of captions, such as images, are sent out-of-band with video information, allowing user devices to directly overlay these images onto the video without rendering text, ensuring uniformity and compatibility across devices, and enabling changes to captions without modifying the video file.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If software renderers are used to render captions on different user devices, then caption rendering flexibility is improved, but display uniformity deteriorates due to varying renderer implementations

Engineering Contradiction:
Improvecaption rendering flexibilityVSAvoiddisplay uniformity
Core Design Contradiction:
Adaptability or versatilityVSManufacturing precision

Solution Approach 1:

The patent pre-renders captions into visual representations (images) before transmission to user devices. This preliminary rendering action ensures that captions are rendered consistently in advance, eliminating variations caused by different software renderer implementations on different devices. The pre-rendered visual representations are then transmitted to user devices for direct display, ensuring uniformity across all devices.

Inventive Principle:
Principle #10Preliminary action

2Stability of the object's composition

If captions are multiplexed with video in the same file (in-band), then integration is improved, but flexibility for updating captions deteriorates

Engineering Contradiction:
ImproveintegrationVSAvoidflexibility for updating captions
Core Design Contradiction:
Stability of the object's compositionVSEase of manufacture

Solution Approach 1:

The patent separates captions from video into distinct streams. Captions are transmitted as visual representations in a separate caption stream from the video stream. This segmentation allows independent modification of captions without affecting the video file, enabling easy updates and additions of languages while maintaining integration through the caption overlay process.

Inventive Principle:
Principle #1Segmentation

3Ease of manufacture

If pre-rendered visual representations of captions are sent out-of-band with video, then caption update flexibility is improved, but data transmission complexity increases

Engineering Contradiction:
Improvecaption update flexibilityVSAvoiddata transmission complexity
Core Design Contradiction:
Ease of manufactureVSDevice complexity

Solution Approach 1:

The patent uses visual representations (pre-rendered images) as intermediaries to carry caption information. Instead of transmitting raw text or multiplexed data, the system transmits pre-rendered visual representations that serve as a mediator between the caption content and the display. This intermediary approach simplifies the transmission process while maintaining flexibility for updates.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS9307292B2Overlay of visual representations of captions on video
Publication Date: 2016.04.05 HULU LLC
  • US9307292B2 patent drawing
  • US9307292B2 patent drawing
  • US9307292B2 patent drawing

AI summary

In one embodiment, a method receiving a request for a media program from a user device. The method then determines a set of visual representations of captions for the media program and determines video information for the media program. Visual representations from the set of visual representations of captions are sent with the video information over a network to the user device where text for the captions has been pre-rendered in the sent visual representations before sending of the visual representations to the user device. Also, the user device is configured to directly render and overlay a visual representation of a caption from the visual representations over a portion of the video information without rendering of the text for caption on the portion of the video information at the user device.