Indirect Character Train Estimation via Phoneme Comparison

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Conventional singing voice generation systems face difficulties in designating desired characters through simple operations, as they require complex character designation processes and offer limited flexibility in selecting characters that deviate from sequenced lyrics, making it challenging to perform ad lib performances with desired melodies.

Innovation Solution

The system allows indirect designation of a target character train using a limited set of particular phonemes, such as vowels and consonants, by comparing a target phoneme train to a reference phoneme train, enabling quick and accurate estimation of the corresponding character sequence in the reference character train.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Adaptability or versatility

If characters are designated one by one using conventional methods (vowel keys, consonant keys, voiced sound symbol keys), then character selection flexibility is improved, but operation complexity increases significantly

Engineering Contradiction:
Improvecharacter selection flexibilityVSAvoidoperation complexity
Core Design Contradiction:
Adaptability or versatilityVSDevice complexity

Solution Approach 1:

The patent segments the character designation process into two independent dimensions: phoneme selection (vowel/consonant) and position selection (lyric progression). This allows users to independently control character content and timing, reducing operational complexity while maintaining flexibility.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent introduces phoneme trains as an intermediary representation layer between user input and character output. By operating on phonemes rather than directly on characters, the system simplifies the designation process while enabling complex character selection through combination of phonemes.

Inventive Principle:
Principle #24Intermediary (Mediator)

2Extent of automation

If lyrics progress automatically in synchronism with music performance, then voice generation automation is improved, but ad lib performance capability deteriorates

Engineering Contradiction:
Improvevoice generation automationVSAvoidad lib performance capability
Core Design Contradiction:
Extent of automationVSAdaptability or versatility

Solution Approach 1:

The patent makes the lyric progression dynamic and user-controllable rather than fixed and automatic. Users can adjust the progression speed and timing of phoneme trains independently from the music playback, enabling both automatic synchronized performance and flexible ad lib performance.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The patent prepares multiple phoneme train patterns in advance that correspond to different lyric progressions. Users can select from pre-prepared phoneme sequences or modify them in real-time, combining the benefits of preparation with flexible performance.

Inventive Principle:
Principle #10Preliminary action

3Adaptability or versatility

If all character designating choices are provided, then character selection versatility is improved, but selection operation difficulty increases

Engineering Contradiction:
Improvecharacter designating choicesVSAvoidselection operation difficulty
Core Design Contradiction:
Adaptability or versatilityVSEase of operation

Solution Approach 1:

The patent divides the character set into phoneme components (vowels and consonants) that can be independently selected and combined. This segmentation reduces the selection space from thousands of characters to a manageable set of phonemes, making operations easier while maintaining versatility through combination.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent changes the selection parameter from character level to phoneme level. By selecting phonemes rather than complete characters, users reduce the number of selection operations needed while maintaining the ability to form any desired character through phoneme combination.

Inventive Principle:
Principle #35Parameter changes

4Measurement precision

If character progression speed matches music performance speed, then synchronization accuracy is improved, but selection speed requirement increases

Engineering Contradiction:
Improvesynchronization accuracyVSAvoidselection speed
Core Design Contradiction:
Measurement precisionVSSpeed

Solution Approach 1:

The patent segments character designation into phoneme-level operations that can be executed more quickly than full character selection. This segmentation allows users to progress through phoneme trains at speeds matching music performance while maintaining synchronization accuracy.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent uses phoneme trains as simplified copies or representations of the full character sequences. These phoneme trains can be processed and selected more rapidly than complete character strings, enabling selection speed to match music performance speed while maintaining synchronization.

Inventive Principle:
Principle #26Copying

Data Source

PatentUS9711133B2Estimation of target character train
Publication Date: 2017.07.18 YAMAHA CORP
  • US9711133B2 patent drawing
  • US9711133B2 patent drawing
  • US9711133B2 patent drawing

AI summary

A desired character train included in a predefined reference character train, such as lyrics, is set as a target character train, and a user designates a target phoneme train that is indirectly representative of the target character train by use of a limited plurality of kinds of particular phonemes, such as vowels and a particular consonants. A reference phoneme train indirectly representative of the reference character train by use of the particular phonemes is prepared in advance. Based on a comparison between the target phoneme train and the reference phoneme train, a sequence of the particular phonemes in the reference phoneme train that matches the target phoneme train is identified, and a character sequence in the reference character train that corresponds to the identified sequence of the particular phonemes is identified. The thus-identified character sequence estimates the target character train.