Lyrics Rendering Synchronization for Polyphonic Words

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Existing methods for rendering lyrics fail to accurately synchronize the display of polyphonic words with their correct pronunciations during song playback, leading to mismatched lyrics and an unsatisfactory artistic experience, particularly in languages like Japanese where pronunciation variations are common.

Innovation Solution

A method and apparatus for rendering lyrics that acquire pronunciation and playback time information of polyphonic words, determine the number of furiganas or syllables, and segment pixels to correspond with each furigana or syllable, allowing for simultaneous and synchronized rendering of polyphonic words and their pronunciations on a display.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Measurement precision

If polyphonic words are marked with correct pronunciation in brackets, then pronunciation accuracy is improved, but lyrics synchronization with audio playback deteriorates

Engineering Contradiction:
Improvepronunciation accuracyVSAvoidlyrics synchronization
Core Design Contradiction:
Measurement precisionVSLoss of time

Solution Approach 1:

The patent segments the polyphonic word into multiple parts: the original word and separate furigana characters for each syllable of the correct pronunciation. Each furigana is assigned a specific rendering duration corresponding to its syllable timing, allowing the pronunciation to be displayed sequentially without blocking the lyrics flow. This segmentation resolves the contradiction by enabling both accurate pronunciation display and proper synchronization.

Inventive Principle:
Principle #1Segmentation

Solution Approach 2:

The patent displays the correct pronunciation in a different spatial dimension (above or below the polyphonic word) rather than replacing or interrupting the lyrics text. This dimensional separation allows the pronunciation information to coexist with the lyrics without causing synchronization issues, as the furigana are rendered in parallel to the main lyrics flow.

Inventive Principle:
Principle #17Another dimension (Dimensionality change)

2Loss of information

If pronunciation is displayed adjacent to polyphonic word, then pronunciation information is provided, but display complexity increases

Engineering Contradiction:
Improvepronunciation informationVSAvoiddisplay complexity
Core Design Contradiction:
Loss of informationVSDevice complexity

Solution Approach 1:

The patent implements dynamic rendering where the furigana characters are displayed sequentially based on their assigned rendering durations rather than all at once. The display system dynamically controls the visibility and timing of each furigana character to match the audio playback timing, reducing visual clutter while maintaining complete pronunciation information.

Inventive Principle:
Principle #15Dynamics

Solution Approach 2:

The patent pre-calculates and assigns rendering durations to each furigana character based on the audio timing information before playback begins. This preliminary preparation allows the display system to simply follow the pre-determined timing schedule during playback, reducing the complexity of real-time display control while ensuring accurate synchronization.

Inventive Principle:
Principle #10Preliminary action

Data Source

PatentUS11604919B2Method and apparatus for rendering lyrics
Publication Date: 2023.03.14 TENCENT MUSIC ENTERTAINMENT TECH (SHENZHEN) CO LTD
  • US11604919B2 patent drawing
  • US11604919B2 patent drawing
  • US11604919B2 patent drawing

AI summary

A method for rendering lyrics is provided, including: acquiring pronunciation of a polyphonic word to be rendered in target lyrics, and acquiring playback time information of the pronunciation in the process of rendering the target lyrics; determining a first number of furiganas contained in the pronunciation; and word-by-word simultaneously rendering, according to the first number and the playback time information of the pronunciation of the polyphonic word to be rendered, the polyphonic word to be rendered and each furigana in the pronunciation of the polyphonic word to be rendered, wherein the pronunciation of the polyphonic word to be rendered is adjacent to and parallel to the polyphonic word to be rendered.