Lyric Structure Extraction Using Repeated Pattern Analysis
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional audio file players face challenges in efficiently searching and selecting audio files due to limited display size and increasing storage capacity, requiring users to remember detailed information and play audio files from the beginning to confirm identification, which is time-consuming and inconvenient.
Innovation Solution
An apparatus and method that extracts the structure of song lyrics using a repeated pattern to create a tree structure, allowing for the extraction of thematic portions of audio files, reducing the time required to select audio files by arranging interlude sections, character strings, and paragraphs in a tree structure for quicker retrieval.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If users search for audio files by playing from the beginning portion to confirm identification, then they can accurately identify the desired audio file, but it takes a lot of time
Solution Approach 1:
The patent extracts the prelude portion (introductory section) from the audio file and uses it to generate a fingerprint. This extracted portion contains unique characteristics that can identify the audio file without requiring playback of the entire file from the beginning, thus reducing identification time while maintaining accuracy
Solution Approach 2:
The system performs preliminary analysis of the prelude portion to generate a fingerprint before the user needs to identify the audio file. This pre-computed fingerprint is then used for rapid comparison and identification, eliminating the need to play through the beginning of each file during the selection process
2Volume of moving object
If the display window size is decreased for miniaturization, then the device size is reduced, but selecting a song title by manipulating buttons becomes inconvenient
Solution Approach 1:
The patent replaces the mechanical button manipulation system with an acoustic recognition system. Instead of requiring users to physically navigate through small display buttons, the system uses speech recognition to identify and select audio files, making operation convenient regardless of display size
Solution Approach 2:
The patent introduces speech recognition as an intermediary between the user and the audio file selection process. The speech recognition system acts as a mediator that translates user speech into file selection commands, eliminating the need for direct manipulation of the small display interface
3Quantity of substance
If the storage capacity is increased to store more audio files, then the data storage capability is improved, but it takes longer to retrieve desired audio files
Solution Approach 1:
The patent creates a fingerprint copy of the prelude portion for each audio file and stores this fingerprint in a database. During retrieval, users provide a reference fingerprint that is compared against the stored fingerprints, enabling rapid identification of matching files without scanning through the entire music library
Solution Approach 2:
The system performs preliminary processing to generate and store fingerprints for all audio files in advance. This pre-computation allows for rapid comparison and retrieval operations, as the system only needs to compare fingerprints rather than analyze entire files during the retrieval process
Data Source
AI summary
An apparatus, system, and method for extracting the structure of song lyrics using a repeated pattern thereof are provided. The apparatus includes a lyric extractor extracting lyric information from metadata related to an audio file, a character string information extractor extracting an interlude section and a repeated character string based on the extracted lyric information, a paragraph extractor extracting a paragraph based on the repeated character string and then a set of paragraphs having the same repeated pattern among the extracted paragraphs, and a lyric structure generator arranging an interlude section, a character string, and a paragraph related to the audio file in a tree structure.


