Indirect Character Train Estimation via Phoneme Comparison
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Conventional singing voice generation systems face difficulties in designating desired characters through simple operations, as they require complex character designation processes and offer limited flexibility in selecting characters that deviate from sequenced lyrics, making it challenging to perform ad lib performances with desired melodies.
Innovation Solution
The system allows indirect designation of a target character train using a limited set of particular phonemes, such as vowels and consonants, by comparing a target phoneme train to a reference phoneme train, enabling quick and accurate estimation of the corresponding character sequence in the reference character train.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Adaptability or versatility
If characters are designated one by one using conventional methods (vowel keys, consonant keys, voiced sound symbol keys), then character selection flexibility is improved, but operation complexity increases significantly
Solution Approach 1:
The patent segments the character designation process into two independent dimensions: phoneme selection (vowel/consonant) and position selection (lyric progression). This allows users to independently control character content and timing, reducing operational complexity while maintaining flexibility.
Solution Approach 2:
The patent introduces phoneme trains as an intermediary representation layer between user input and character output. By operating on phonemes rather than directly on characters, the system simplifies the designation process while enabling complex character selection through combination of phonemes.
2Extent of automation
If lyrics progress automatically in synchronism with music performance, then voice generation automation is improved, but ad lib performance capability deteriorates
Solution Approach 1:
The patent makes the lyric progression dynamic and user-controllable rather than fixed and automatic. Users can adjust the progression speed and timing of phoneme trains independently from the music playback, enabling both automatic synchronized performance and flexible ad lib performance.
Solution Approach 2:
The patent prepares multiple phoneme train patterns in advance that correspond to different lyric progressions. Users can select from pre-prepared phoneme sequences or modify them in real-time, combining the benefits of preparation with flexible performance.
3Adaptability or versatility
If all character designating choices are provided, then character selection versatility is improved, but selection operation difficulty increases
Solution Approach 1:
The patent divides the character set into phoneme components (vowels and consonants) that can be independently selected and combined. This segmentation reduces the selection space from thousands of characters to a manageable set of phonemes, making operations easier while maintaining versatility through combination.
Solution Approach 2:
The patent changes the selection parameter from character level to phoneme level. By selecting phonemes rather than complete characters, users reduce the number of selection operations needed while maintaining the ability to form any desired character through phoneme combination.
4Measurement precision
If character progression speed matches music performance speed, then synchronization accuracy is improved, but selection speed requirement increases
Solution Approach 1:
The patent segments character designation into phoneme-level operations that can be executed more quickly than full character selection. This segmentation allows users to progress through phoneme trains at speeds matching music performance while maintaining synchronization accuracy.
Solution Approach 2:
The patent uses phoneme trains as simplified copies or representations of the full character sequences. These phoneme trains can be processed and selected more rapidly than complete character strings, enabling selection speed to match music performance speed while maintaining synchronization.
Data Source
AI summary
A desired character train included in a predefined reference character train, such as lyrics, is set as a target character train, and a user designates a target phoneme train that is indirectly representative of the target character train by use of a limited plurality of kinds of particular phonemes, such as vowels and a particular consonants. A reference phoneme train indirectly representative of the reference character train by use of the particular phonemes is prepared in advance. Based on a comparison between the target phoneme train and the reference phoneme train, a sequence of the particular phonemes in the reference phoneme train that matches the target phoneme train is identified, and a character sequence in the reference character train that corresponds to the identified sequence of the particular phonemes is identified. The thus-identified character sequence estimates the target character train.


