Audio Timing Adjustment for Rap Singing Synchronization

Resolve Bottlenecks,
Find Innovative Solutions
Generate Solutions

Solution Overview

Problem

Ordinary users face difficulties in singing rap music due to the need for music theory knowledge and singing skills, resulting in poor matching between user singing audio and original rap music audio.

Innovation Solution

An audio data processing method and apparatus that obtains song information, determines a predefined portion of the song and corresponding music score information, receives user-input audio data, and processes word time lengths based on time information and music score information to improve the matching between user singing audio and original rap music audio.

Engineering Contradictions & Design Principles

VSEngineering Contradiction Analysis

1Ease of operation

If conventional karaoke products are used to sing rap music, then users can freely sing with various sound effects, but the matching between user singing audio and original rap music audio is poor due to lack of music theory knowledge and singing skills

Engineering Contradiction:
Improveease of singing rap musicVSAvoidmatching precision between user singing audio and original rap music audio
Core Design Contradiction:
Ease of operationVSManufacturing precision

Solution Approach 1:

The patent introduces an automatic timing adjustment mechanism as an intermediary between the user's singing audio and the original rap music. The system automatically detects the timing of each word in the user's singing and adjusts it to match the music score timing, eliminating the need for users to have music theory knowledge while ensuring accurate synchronization with the original music

Inventive Principle:
Principle #24Intermediary (Mediator)

Solution Approach 2:

The patent replaces the manual music theory-based timing adjustment mechanism with an automated digital signal processing system. The system uses audio analysis algorithms to automatically detect word timing and applies time-stretching and pitch-shifting operations to synchronize user singing with the original music, substituting complex manual musical adjustment with automated computational processing

Inventive Principle:
Principle #28Mechanics substitution (Replace mechanical system)

2Manufacturing precision

If users rely on music theory knowledge and singing skills to sing rap music, then the matching between user singing audio and original rap music audio can be improved, but ordinary users face difficulties due to lack of these skills

Engineering Contradiction:
Improvematching precision between user singing audio and original rap music audioVSAvoidease of singing rap music
Core Design Contradiction:
Manufacturing precisionVSEase of operation

Solution Approach 1:

The patent enables the system to automatically perform timing adjustment and synchronization tasks that would normally require music theory knowledge. The automatic timing detection and adjustment mechanisms allow the system to self-correct timing discrepancies without user intervention, making rap singing accessible to ordinary users while maintaining high matching precision

Inventive Principle:
Principle #25Self-service

Solution Approach 2:

The patent dynamically adjusts audio parameters such as timing, tempo, and pitch of the user's singing to match the original rap music. By automatically modifying these parameters based on detected timing information and music score data, the system achieves professional-quality synchronization without requiring users to possess music theory expertise

Inventive Principle:
Principle #35Parameter changes

Data Source

PatentUS10789290B2Audio data processing method and apparatus, and computer storage medium
Publication Date: 2020.09.29 TENCENT TECHNOLOGY (SHENZHEN) CO LTD
  • US10789290B2 patent drawing
  • US10789290B2 patent drawing
  • US10789290B2 patent drawing

AI summary

The present disclosure discloses an audio data processing performed by a computing device. The computing device obtains song information of a song, the song information comprising an accompaniment file, a lyric file, and a music score file that correspond to the song and then determines a predefined portion of the song and music score information corresponding to the predefined portion according to the song information. After receiving audio data that is input by a user, the computing device determines time information of each word in the audio data and then processes the audio data according to the time information of each word in the audio data and the music score information of the predefined portion of the song. Finally, the computing device obtains mixed audio data by mixing the processed audio data and the accompaniment file.