Text-Based Subtitle Reproduction Apparatus with Style Segmentation

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing multimedia storage systems face challenges in efficiently producing and editing subtitle or caption data due to its multiplexing with video, audio, and interactive graphic streams, and limited flexibility in changing output styles.

Innovation Solution

A storage medium and reproducing apparatus that separate text-based subtitle data from image data, using a subtitle decoder to convert presentation information into bitmap images based on style information, allowing for synchronized output and user-defined style changes, including support for multiple languages and effects like fade-in/out.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Area of moving object

If bitmap-based caption data is multiplexed with video, audio, and interactive graphic streams, then the subtitle data can be integrated into the main AV data stream, but the production and edition of subtitle data become very difficult

Engineering Contradiction:
Improveintegration of subtitle dataVSAvoidproduction and edition of subtitle data
Core Design Contradiction:
Area of moving objectVSEase of manufacture

Solution Approach 1:

The patent segments subtitle data from the multiplexed AV stream by extracting it as a separate component. The reproduction apparatus separates subtitle data from video, audio, and interactive graphic streams, allowing independent processing and editing while maintaining the ability to integrate them during playback. This segmentation enables easy production and edition of subtitle data without affecting other media components.

Inventive Principle:
Principle #1Segmentation

2Ease of operation

If bitmap-based caption data is used, then subtitles can be displayed on images, but the output style of the caption data cannot be changed in a variety of ways

Engineering Contradiction:
Improvedisplay of subtitlesVSAvoidoutput style changes
Core Design Contradiction:
Ease of operationVSAdaptability or versatility

Solution Approach 1:

The patent implements dynamic style switching for subtitles by introducing a style selection mechanism. The reproduction apparatus can change output styles of captions dynamically based on user selection or preset configurations. Multiple style parameters (font, size, color, position) are stored and can be switched without regenerating the subtitle data, enabling versatile output style changes while maintaining easy display operation.

Inventive Principle:
Principle #15Dynamics

3Reliability

If caption data is multiplexed with other data streams, then a complete AV data stream can be created, but the big size of bitmap-based caption data becomes a problem

Engineering Contradiction:
Improvecomplete AV data streamVSAvoidsize of caption data
Core Design Contradiction:
ReliabilityVSQuantity of substance

Solution Approach 1:

The patent extracts subtitle data from the multiplexed AV stream as a separate, independently manageable component. By taking out subtitle data from the integrated stream, the system allows for efficient compression and selective processing. The extraction mechanism enables the AV stream to maintain completeness for reliable playback while managing caption data size separately through efficient encoding and selective decompression only when needed for display.

Inventive Principle:
Principle #2Taking out (Extraction)

Data Source

PatentUS8437612B2Storage medium recording text-based subtitle stream, reproducing apparatus and reproducing method for reproducing text-based subtitle stream recorded on the storage medium
Publication Date: 2013.05.07 SAMSUNG ELECTRONICS CO LTD
  • US8437612B2 patent drawing
  • US8437612B2 patent drawing
  • US8437612B2 patent drawing

AI summary

A non-transitory computer readable storage medium and apparatus to reproduce from a storage medium are provided. The apparatus has a video decoder to decode audio-visual data and a subtitle decoder to receive text-based subtitle data having dialog presentation units and a dialog style unit defining a set of output styles to be applied to the dialog presentation units, converting the dialog presentation units into bitmap images based on the dialog style unit, and controlling an output of the converted dialog presentation units synchronized with decoded audio-visual data. Each dialog presentation unit has dialog text information, time information indicating a time for the dialog text information to be output, palette information defining colors to be applied to the dialog text information, and a color update flag indicating whether only the palette information has changed as compared with a graphical composition of a previous dialog presentation unit.