Lyrics Rendering Synchronization for Polyphonic Words
Find Innovative SolutionsGenerate Solutions
Solution Overview
Problem
Existing methods for rendering lyrics fail to accurately synchronize the display of polyphonic words with their correct pronunciations during song playback, leading to mismatched lyrics and an unsatisfactory artistic experience, particularly in languages like Japanese where pronunciation variations are common.
Innovation Solution
A method and apparatus for rendering lyrics that acquire pronunciation and playback time information of polyphonic words, determine the number of furiganas or syllables, and segment pixels to correspond with each furigana or syllable, allowing for simultaneous and synchronized rendering of polyphonic words and their pronunciations on a display.
Engineering Contradictions & Design Principles
Engineering Contradiction Analysis
1Measurement precision
If polyphonic words are marked with correct pronunciation in brackets, then pronunciation accuracy is improved, but lyrics synchronization with audio playback deteriorates
Solution Approach 1:
The patent segments the polyphonic word into multiple parts: the original word and separate furigana characters for each syllable of the correct pronunciation. Each furigana is assigned a specific rendering duration corresponding to its syllable timing, allowing the pronunciation to be displayed sequentially without blocking the lyrics flow. This segmentation resolves the contradiction by enabling both accurate pronunciation display and proper synchronization.
Solution Approach 2:
The patent displays the correct pronunciation in a different spatial dimension (above or below the polyphonic word) rather than replacing or interrupting the lyrics text. This dimensional separation allows the pronunciation information to coexist with the lyrics without causing synchronization issues, as the furigana are rendered in parallel to the main lyrics flow.
2Loss of information
If pronunciation is displayed adjacent to polyphonic word, then pronunciation information is provided, but display complexity increases
Solution Approach 1:
The patent implements dynamic rendering where the furigana characters are displayed sequentially based on their assigned rendering durations rather than all at once. The display system dynamically controls the visibility and timing of each furigana character to match the audio playback timing, reducing visual clutter while maintaining complete pronunciation information.
Solution Approach 2:
The patent pre-calculates and assigns rendering durations to each furigana character based on the audio timing information before playback begins. This preliminary preparation allows the display system to simply follow the pre-determined timing schedule during playback, reducing the complexity of real-time display control while ensuring accurate synchronization.
Data Source
AI summary
A method for rendering lyrics is provided, including: acquiring pronunciation of a polyphonic word to be rendered in target lyrics, and acquiring playback time information of the pronunciation in the process of rendering the target lyrics; determining a first number of furiganas contained in the pronunciation; and word-by-word simultaneously rendering, according to the first number and the playback time information of the pronunciation of the polyphonic word to be rendered, the polyphonic word to be rendered and each furigana in the pronunciation of the polyphonic word to be rendered, wherein the pronunciation of the polyphonic word to be rendered is adjacent to and parallel to the polyphonic word to be rendered.


