Mobile Terminal Text Extraction from Video Content

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional mobile terminals lack the capability to conveniently manipulate and control video content by converting images and sound into text, limiting user interaction and management of multimedia functions.

Innovation Solution

A mobile terminal with a controller that extracts and displays text from video content, allowing users to select and analyze specific sections, and edit video content based on extracted text, enhancing user interaction and management through touchscreen inputs and multimedia processing.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If video content is played back conventionally without text extraction, then the playback is simple and fast, but user interaction and content management capability are limited

Engineering Contradiction:
Improveuser interaction capabilityVSAvoidprocessing complexity
Core Design Contradiction:
Ease of operationVSDevice complexity

Solution Approach 1:

The controller extracts text from video content in advance during playback, creating a subtitle file that stores text information corresponding to time codes. This preliminary text extraction enables subsequent user interactions such as searching, selecting, and managing video content based on text, without requiring real-time processing during user operations.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system creates a text copy (subtitle file) of the video content's auditory and visual information. This text copy serves as an independent representation that can be searched, manipulated, and managed separately from the original video file, enabling enhanced user interaction without modifying the original video data.

Inventive Principle:
Principle #26Copying

2Adaptability or versatility

If text extraction and analysis functions are added to enable content manipulation, then user interaction improves, but processing time and complexity increase

Engineering Contradiction:
Improvecontent management capabilityVSAvoidprocessing time
Core Design Contradiction:
Adaptability or versatilityVSLoss of time

Solution Approach 1:

Text extraction is performed in advance during the initial playback, creating a subtitle file that contains all text information with corresponding time codes. This preliminary processing enables rapid subsequent operations such as searching for specific text, selecting content segments, and managing video files without requiring additional processing time during user interactions.

Inventive Principle:
Principle #10Preliminary action

Solution Approach 2:

The system automatically extracts text and generates subtitle files without requiring user intervention during the extraction process. The controller autonomously analyzes the video content, extracts text from audio and visual streams, and creates the subtitle file, freeing user time for subsequent content management tasks.

Inventive Principle:
Principle #25Self-service

3Loss of information

If the controller extracts and displays text from video content, then user interaction and content analysis improve, but the device requires additional processing resources

Engineering Contradiction:
Improveinformation accessibilityVSAvoidprocessing capability requirement
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The controller creates a text-based copy (subtitle file) of the video content's information. This text copy preserves all essential information from the video in a searchable and manipulable format, making information accessible without requiring the original video processing capabilities for subsequent operations.

Inventive Principle:
Principle #26Copying

Solution Approach 2:

The subtitle file acts as an intermediary between the video content and the user. Instead of directly processing the complex video data for text extraction during user operations, the system uses the pre-generated subtitle file as a mediator, enabling efficient information access and content management with reduced processing requirements.

Inventive Principle:
Principle #24Intermediary (Mediator)

Data Source

PatentUS10162489B2Multimedia segment analysis in a mobile terminal and control method thereof
Publication Date: 2018.12.25 LG ELECTRONICS INC
  • US10162489B2 patent drawing
  • US10162489B2 patent drawing
  • US10162489B2 patent drawing

AI summary

A mobile terminal and a method for controlling the same are disclosed. The mobile terminal includes a display and a controller configured to display at least one piece of video content on the display, to extract at least one text from at least one of an image and sound included in at least a portion of the video content and to display the at least one text on at least one specific position of the display, wherein the at least one specific position is related to at least one point of the video content, from which the at least one text is extracted. According to the present invention, video content can be manipulated more conveniently by displaying images and sound included in the video content as text.