Video Display Device Condensing Content via Audio Characteristics
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Current video display devices require user-initiated double speed play and manual thumbnail generation, leading to missed scenes and inefficient storage usage, especially with high-quality content recordings.
Innovation Solution
A video display device that automatically searches for important content portions based on audio characteristics, such as specific words, main characters, or sound tracks, to output a condensed version, reducing storage consumption and allowing users to quickly access desired scenes.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Speed
If double speed play function is initiated by user operation, then contents can be scanned more quickly, but important scenes may be missed and user attention is required continuously
Solution Approach 1:
The video display device automatically performs scene importance analysis and generates condensed versions without requiring continuous user intervention. The system self-evaluates content to identify important portions and presents them to users, eliminating the need for users to continuously monitor during fast playback.
Solution Approach 2:
The device performs preliminary analysis of content to pre-identify important scenes before user viewing. By analyzing audio characteristics, video characteristics, and metadata in advance, the system prepares condensed versions that highlight key moments, allowing users to review important content without missing anything during accelerated playback.
2Manufacturing precision
If thumbnails are manually generated by broadcasting companies, then thumbnail quality can be ensured, but additional effort and cost are required
Solution Approach 1:
The video display device automatically generates thumbnails by analyzing content characteristics without requiring manual intervention from broadcasting companies. The system self-evaluates video frames, audio segments, and metadata to identify representative moments and generate thumbnails autonomously, maintaining quality while eliminating manual effort.
Solution Approach 2:
The patent replaces manual thumbnail generation (mechanical human effort) with automated computer-based analysis. The system uses algorithmic evaluation of content characteristics to substitute human operators, achieving both quality and efficiency through automated processing.
3Ease of manufacture
If thumbnails are randomly generated at predetermined uniform intervals, then generation effort is reduced, but thumbnails may not reflect what users desire to see
Solution Approach 1:
The system changes the parameter of thumbnail selection from fixed uniform intervals to dynamic selection based on content characteristics. By evaluating audio intensity, video motion, scene transitions, and metadata, the system adapts thumbnail placement to reflect important moments, maintaining low effort while improving relevance to user interests.
4Manufacturing precision
If high quality contents are supplied and series recording function is used, then content quality is maintained, but storage space is consumed rapidly
Solution Approach 1:
The device extracts and stores only the essential characteristics and metadata needed for condensed version generation rather than duplicating entire high-quality content. By separating the analysis data from the original content, the system maintains content quality for viewing while minimizing additional storage requirements for the recording function.
Data Source
Figure 1
Figure 2(a)~2(c)
Figure 3
AI summary
According to one embodiment, a video display device configured to play contents including audio includes: a controller configured to receive a request for a condensed version of the contents and to search the contents based on audio characteristics information corresponding to a condensing criterion in order to output the condensed version; and a display configured to display the contents. The condensing criterion includes at least a specific word, a name of a main character, an original sound track, a sound effect or a voice print of an actor.