A song practicing auxiliary method, device, medium and product
Patent Information
- Application Number
- CN202610913485.6
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2026-06-23
- Publication Date
- 2026-09-18
AI Technical Summary
[0002]在主流的在线唱歌应用中,如果用户想要对某首歌曲进行练习,一种方式是手动选择整首歌曲,这种情况下每次都需要从头到尾的练习整首歌曲,反复循环;另一种方式是手动选择某个歌曲段落进行练习,但该选择完全依赖于用户的主观判断,而用户往往难以准确选择自己真正演唱薄弱的歌曲段落,最终给用户造成假性掌握练唱歌曲的错觉
[0019] In this application, the following steps are taken: obtaining the singing audio of a target user after singing a target song; performing segment-level evaluation on the singing audio to obtain evaluation results for each song segment in the target song; determining the song segment to be practiced from each song segment based on the evaluation results for each song segment; dividing the song segment to be practiced into practice segments based on the song structure information corresponding to the practice segment; and executing a progressive practice process corresponding to the practice segment to guide the target user to complete the practice of the practice segment.
Smart Images

Figure CN122781237A_ABST
Abstract
Description
Technical Field
[0001] This invention relates to the field of singing practice assistance, and in particular to a method, device, medium, and product for assisting in singing practice. Background Technology
[0002] In mainstream online singing apps, if a user wants to practice a song, one way is to manually select the entire song, which requires practicing the whole song from beginning to end repeatedly. Another way is to manually select a specific section of the song, but this selection relies entirely on the user's subjective judgment, and users often struggle to accurately choose sections where they are truly weak, ultimately creating a false impression of mastery. Furthermore, both methods offer broad practice granularity, easily leading to repeated practice sessions and low efficiency. Summary of the Invention
[0003] In view of this, the purpose of this invention is to provide a song practice assistance method, device, medium, and product that can accurately locate the user's weakest singing sections from a target song, and divide the song sections into practice segments based on the song's structural information to execute a progressive practice process, thereby improving the efficiency of song practice and helping users truly master the song. The specific solution is as follows: Firstly, this application discloses a method for assisting in song practice, including: Obtain the audio recording of the target user singing the target song; The singing audio is evaluated at the segment level to obtain the evaluation results for each segment of the target song. Based on the evaluation results corresponding to each of the aforementioned song segments, the song segments to be practiced are determined from each of the aforementioned song segments; Based on the song structure information corresponding to the song segment to be practiced, the song segment to be practiced is divided to generate practice segments; The progressive practice process corresponding to the practice segment is executed to guide the target user to complete the practice of the practice segment.
[0004] Optionally, the evaluation results for each song segment include the evaluation score for each song segment; Accordingly, determining the song segment to be practiced from each of the song segments based on the evaluation results corresponding to each of the song segments includes: Based on the evaluation scores corresponding to each of the song segments, the song segments with evaluation scores lower than the preset scores are identified from the song segments to be practiced.
[0005] Optionally, the step of dividing the song segment to be practiced into practice segments based on the song structure information corresponding to the song segment to be practiced includes: Based on the phrase boundaries in the song segment to be practiced, the song segment to be practiced is divided into phrases to obtain the target phrases; Merge the musical phrases in the target musical phrase that meet the preset merging conditions to obtain the merged musical phrase; the preset merging conditions include that the musical phrases to be merged are adjacent in the target song and the sum of their singing durations is not greater than a preset duration; A practice segment is generated based on the merged musical phrase and the unmerged musical phrase in the target musical phrase.
[0006] Optionally, the step of dividing the song segment to be practiced into musical phrases to obtain the target musical phrase includes: Each musical phrase obtained by dividing the song segment to be practiced into musical phrases is determined as the target musical phrase; Alternatively, based on the evaluation results corresponding to each musical phrase, a target musical phrase can be determined from each musical phrase; the evaluation results corresponding to each musical phrase are the results obtained after performing a musical phrase-level evaluation on the singing audio.
[0007] Optionally, the progressive singing practice process includes singing practice levels with increasing difficulty, and the singing practice levels with increasing difficulty correspond to singing practice assistance modes with decreasing assistance levels. Accordingly, the process of executing the progressive practice singing process corresponding to the practice segment includes: Based on the target singing practice assistance mode corresponding to the current singing practice level, the target user is guided to complete the singing practice of the singing practice segment in order to obtain the singing practice audio; The practice audio is evaluated to determine the evaluation result corresponding to the practice audio. Based on the evaluation results corresponding to the practice audio, determine whether the current practice level has been passed.
[0008] Optionally, the step of guiding the target user to complete the practice singing of the practice segment based on the target practice singing assistance mode corresponding to the current practice singing level, in order to obtain the practice singing audio, includes: Obtain the start and end timestamps corresponding to the practice segment; Position the playback pointer to the playback start position in the target audio track corresponding to the start timestamp; the target audio track is the auxiliary audio track that needs to be played in the target singing practice auxiliary mode, and the auxiliary audio track includes the original vocal track and / or accompaniment track of the target song; The target audio track is played from the playback start position to guide the target user to start practicing the practice segment and to start the audio recording function. When the playback pointer is detected to have reached the playback termination position, the audio recording function is terminated to obtain the practice audio; the playback termination position is the playback position in the target audio track corresponding to the termination timestamp.
[0009] Optionally, after determining whether the current singing practice level has been passed based on the evaluation results corresponding to the singing practice audio, the method further includes: If the current singing practice level is not passed, the playback pointer will be reset to the playback start position to guide the target user to practice the current singing practice level again for the practice segment.
[0010] Optionally, the starting timestamp is a first timestamp, or a timestamp obtained by subtracting a preset time from the first timestamp; the first timestamp is the earliest timestamp of the practice segment in the target song; The termination timestamp is the second timestamp, or the timestamp obtained by adding the preset time to the second timestamp; the second timestamp is the latest timestamp of the practice segment in the target song.
[0011] Optionally, the evaluation results corresponding to the singing practice audio include the challenge score corresponding to the singing practice audio; Accordingly, determining whether to pass the current singing practice stage based on the evaluation results corresponding to the singing practice audio includes: Determine whether the challenge score corresponding to the singing practice audio is not less than the preset score corresponding to the current singing practice level; If it is not less than, then the current singing practice level is considered passed; If the value is less than the required value, then the current singing practice level has not been passed.
[0012] Optionally, after determining whether the current singing practice level has been passed based on the evaluation results corresponding to the singing practice audio, the method further includes: If the current singing practice level is not passed, the cumulative number of failures for the current singing practice level is determined, and when the cumulative number of failures reaches a preset number, the preset score is lowered.
[0013] Optionally, after determining whether the current singing practice level has been passed based on the evaluation results corresponding to the singing practice audio, the method further includes: Update the current progress status based on the judgment result of whether the current singing practice level has been passed; The current challenge status includes the current level identifier, the level status of all practice levels corresponding to the practice segment and the highest historical challenge score, the total number of practice sessions for the practice segment, and the practice completion status.
[0014] Optionally, the song practice aid method further includes: When a preset event is detected, the display operation corresponding to the preset event is triggered; The preset events include completing a practice session, passing all practice sessions corresponding to the practice segment, and completing practice sessions for all the practice segments; the display operations include displaying practice results and / or displaying special effects.
[0015] Optionally, the song practice aid method further includes: If an exit event is detected during the execution of the progressive singing practice process, the current singing practice progress is recorded so that a breakpoint recovery operation can be performed based on the current singing practice progress.
[0016] Secondly, this application discloses an electronic device, comprising: Memory, used to store computer programs; A processor is used to execute the computer program to implement the song practice assistance method as described above.
[0017] Thirdly, this application discloses a computer-readable storage medium for storing a computer program, which, when executed by a processor, implements the aforementioned song practice assistance method.
[0018] Fourthly, this application discloses a computer program product, including a computer program / instruction, which, when executed by a processor, implements the aforementioned song practice assistance method.
[0019] In this application, the following steps are taken: obtaining the singing audio of a target user after singing a target song; performing segment-level evaluation on the singing audio to obtain evaluation results for each song segment in the target song; determining the song segment to be practiced from each song segment based on the evaluation results for each song segment; dividing the song segment to be practiced into practice segments based on the song structure information corresponding to the practice segment; and executing a progressive practice process corresponding to the practice segment to guide the target user to complete the practice of the practice segment.
[0020] Therefore, this application, by conducting segment-level evaluations of the singing audio, objectively and accurately identifies the user's truly weak singing segments within the target song based on the evaluation results. This helps the user focus their attention on practicing the segments that truly need improvement, thus better mastering the target song. Furthermore, this application, based on song structure information, divides song segments into fine-grained practice segments to execute a progressive practice process. This not only avoids disrupting the musical structure during the segment division process but also keeps the length of each practice session within a reasonable range, avoiding the need for repeated practice cycles when the song is long. This improves the efficiency of song practice and helps users more accurately master each practice segment. Attached Figure Description
[0021] To more clearly illustrate the technical solutions in the embodiments of the present invention or the prior art, the drawings used in the description of the embodiments or the prior art will be briefly introduced below. Obviously, the drawings described below are only embodiments of the present invention. For those skilled in the art, other drawings can be obtained based on the provided drawings without creative effort.
[0022] Figure 1 A system framework diagram provided for an embodiment of this application; Figure 2 A flowchart of a song practice assistance method provided in this application embodiment; Figure 3 A specific flowchart of a challenge-based practice exercise is provided for an embodiment of this application; Figure 4 An overall architecture diagram of a singing practice system provided in this application embodiment; Figure 5 A specific flowchart for assisting in song practice is provided in the embodiments of this application; Figure 6 This is a structural diagram of an electronic device provided in an embodiment of this application. Detailed Implementation
[0023] The technical solutions of the embodiments of the present invention will be clearly and completely described below with reference to the accompanying drawings. Obviously, the described embodiments are only some embodiments of the present invention, and not all embodiments. Based on the embodiments of the present invention, all other embodiments obtained by those skilled in the art without creative effort are within the scope of protection of the present invention.
[0024] In mainstream online singing applications, users typically manually select the entire song or a specific section for practice. This reliance on subjective selection makes it difficult to accurately pinpoint the user's weakest areas, ultimately creating a false impression of mastery. Furthermore, the practice granularity is broad, whether for the entire song or a section, often requiring repeated practice to achieve proficiency, resulting in low efficiency. Therefore, this application provides a song practice assistance method that precisely identifies the user's weakest sections within a target song and, based on song structure information, finely divides these sections into practice segments, executing a progressive practice process to improve efficiency and help users truly master the song.
[0025] The system framework used in the method disclosed in this application can be found in [reference needed]. Figure 1 As shown, it may specifically include: a backend server 01 and one or more user terminals 02 connected to the backend server 01. The user terminal 02 is a terminal device with audio playback and audio recording functions, such as a smartphone, tablet, or laptop.
[0026] Furthermore, the backend server 01 in this application is mainly used to implement song practice assistance. Specifically, when the backend server 01 implements song practice assistance, it executes the following steps: after the target user selects a target song through the user terminal 02 and sings it, it obtains the singing audio obtained by the target user after singing the target song; it performs segment-level evaluation on the singing audio to obtain the evaluation results corresponding to each song segment in the target song; based on the evaluation results corresponding to each song segment, it determines the song segment to be practiced from each song segment; based on the song structure information corresponding to the song segment to be practiced, it divides the song segment to be practiced to generate practice segments; and it executes the progressive practice process corresponding to the practice segment to guide the target user to complete the practice of the practice segment.
[0027] See Figure 2 As shown, this embodiment of the invention discloses a method for assisting in song practice, including: Step S11: Obtain the singing audio of the target user after singing the target song.
[0028] Step S12: Perform segment-level evaluation on the singing audio to obtain the evaluation results corresponding to each song segment in the target song.
[0029] In this embodiment of the application, the singing audio obtained by the target user after singing the target song is first obtained, and the singing audio is evaluated at the segment level on the preset evaluation dimensions using a preset audio evaluation system to determine the evaluation results corresponding to each song segment in the target song.
[0030] The preset audio evaluation system is based on artificial intelligence (AI). The preset evaluation dimensions include any one or a combination of pitch accuracy, rhythm, and breath stability. Pitch accuracy is reflected by the deviation between the standard pitch of the target song and the user's actual pitch in the singing audio; rhythm is reflected by the degree of consistency between the standard duration of the target song and the user's actual duration in the singing audio; and breath stability is reflected by the stability and continuity of the user's voice.
[0031] Furthermore, the evaluation results for each song segment include the evaluation score and / or evaluation tag for each song segment. The evaluation tag for each song segment is specifically determined based on its evaluation score, and can be categorized as Excellent, Good, Pass, and Fail. Specifically, if the evaluation score for a song segment is 90-100, the evaluation tag for that song segment is Excellent; if the evaluation score for a song segment is 80-89, the evaluation tag for that song segment is Good; if the evaluation score for a song segment is 60-79, the evaluation tag for that song segment is Pass; and if the evaluation score for a song segment is 0-59, the evaluation tag for that song segment is Fail.
[0032] According to one example, the evaluation score corresponding to any song segment can be determined based on the dimensional scores of any song segment on each preset evaluation dimension. For example, the dimensional scores of any song segment on each preset evaluation dimension are weighted and calculated to obtain the evaluation score corresponding to any song segment.
[0033] Step S13: Based on the evaluation results corresponding to each of the song segments, determine the song segments to be practiced from each of the song segments.
[0034] In this embodiment of the application, after determining the evaluation results corresponding to each song segment in the target song, the song segment to be practiced is determined from each song segment based on the evaluation results corresponding to each song segment. It should be noted that the song segment to be practiced is the song segment in the target song where the user is actually weak in singing.
[0035] In one example, when the evaluation results for each song segment include the evaluation score for each song segment, the song segments to be practiced are determined based on the evaluation scores for each song segment, where the evaluation score is lower than the preset score. For example, if the preset score is 80 points, the song segments to be practiced are those with evaluation scores lower than 80 points.
[0036] In another example, when the evaluation results for each song segment include evaluation tags for that song segment, the song segments to be practiced are determined from among the song segments based on the evaluation tags for each song segment, with the evaluation tags being preset tags. For example, if the preset tags are "pass" and "fail," then the song segments to be practiced are those with evaluation tags of "pass" or "fail."
[0037] It should be noted that the number of song segments to be practiced may be one or more. This embodiment of the application can display a practice suggestion panel on the user's end, showing the song segments to be practiced in ascending order of their evaluation scores. This allows the user to obtain instructions through the practice suggestion panel and determine the corresponding song segment from the song segments to be practiced, executing the subsequent practice segment division and progressive practice process. Specifically, if the user instruction is a selection instruction, the corresponding song segment is selected from the song segments to be practiced based on the selection instruction, executing the subsequent practice segment division and progressive practice process. If the user instruction is a practice start instruction, the corresponding song segment is selected from the song segments to be practiced based on a default rule, which can be selecting the song segment with the lowest evaluation score from the song segments to be practiced.
[0038] The practice suggestion panel can specifically display the lyrics, evaluation score, and weakness type of the song segment to be practiced. The weakness type of the song segment to be practiced can be determined based on the dimension score of the song segment in each preset evaluation dimension. For example, the weakness type can be low pitch, unstable rhythm, unstable breath, etc.
[0039] In this way, the embodiments of this application utilize a preset audio evaluation system to evaluate the singing audio on preset evaluation dimensions, so as to objectively and accurately locate the song sections in the target song where the user's singing is truly weak, thereby helping the user to focus their attention on the song sections that truly need improvement for practice, so as to better master the target song.
[0040] Step S14: Based on the song structure information corresponding to the song segment to be practiced, divide the song segment to be practiced to generate practice segments.
[0041] In order to automatically divide the song segment to be practiced into independent practice segments of appropriate length and reasonable boundaries, this application embodiment introduces a song segment division based on song structure information. Specifically, based on the song structure information corresponding to the song segment to be practiced, the song segment to be practiced is divided to generate several practice segments.
[0042] Based on the song structure information corresponding to the segment of the song to be practiced, the segment is divided to generate practice segments. Specifically, this may include: dividing the segment into phrases based on the phrase boundaries to obtain target phrases; merging phrases within the target phrases that meet preset merging conditions to obtain merged phrases; wherein the preset merging conditions include that the phrases to be merged are adjacent in the target song and that the sum of their performance times does not exceed a preset duration; and generating practice segments based on the merged phrases and the unmerged phrases in the target phrases. For example, both the merged phrases and the unmerged phrases in the target phrases are designated as independent practice segments.
[0043] According to one example, the song segment to be practiced is divided into phrases to obtain target phrases. Specifically, this may include: each phrase obtained after dividing the song segment to be practiced into phrases is determined as a target phrase.
[0044] According to another example, the song segment to be practiced is divided into musical phrases to obtain the target musical phrase. Specifically, this may include: dividing the song segment to be practiced into musical phrases to obtain each musical phrase, and determining the target musical phrase from each musical phrase based on the evaluation results corresponding to each musical phrase; wherein, the evaluation results corresponding to each musical phrase are the results obtained after using a preset audio evaluation system to perform a phrase-level evaluation on the singing audio in a preset evaluation dimension.
[0045] Specifically, when the evaluation results for each musical phrase include the evaluation score for each phrase, the target musical phrases with evaluation scores lower than the preset score are identified based on the evaluation scores for each phrase. Taking a preset score of 80 points as an example, the target musical phrases are those with evaluation scores lower than 80 points.
[0046] Furthermore, this application pre-stores the timeline data of each musical phrase in the target song in the data storage layer. The timeline data is stored in the form of a JSON (JavaScript Object Notation) array, which specifically includes the phrase identifier, lyrics text, start timestamp (milliseconds), and end timestamp (milliseconds). Based on the start and end timestamps in the timeline data of any musical phrase, the performance duration of that phrase can be determined; that is, the performance duration of any musical phrase equals the end timestamp minus the start timestamp.
[0047] According to one example, in determining the practice segment, each target phrase is scanned sequentially according to the singing order, and a merging judgment is performed on the currently scanned target phrases: First, it is determined whether there are adjacent phrases among the target phrases. The adjacent phrases are adjacent to the currently scanned target phrases in the target song and are located after the currently scanned target phrases. If adjacent phrases exist, it is determined whether the sum of the singing duration of the currently scanned target phrases and the adjacent phrases is not greater than a preset duration. If it is not greater than the preset duration, the currently scanned target phrases and the adjacent phrases are merged, and the merged phrase is determined as the practice segment. If it is greater than the preset duration or there are no adjacent phrases, the currently scanned target phrase is directly determined as the practice segment. The sum of the singing duration of the currently scanned target phrases and the adjacent phrases is equal to the end timestamp of the adjacent phrases minus the start timestamp of the currently scanned target phrases.
[0048] It should be noted that each practice segment corresponds to a structured data object, which includes the segment identifier (segmentId), start timestamp (startPos), end timestamp (endPos), total performance duration (duration), lyrics, original score (originalScore), weaknessType (weaknessType), and difficulty level (difficulty).
[0049] The start timestamp of a practice segment is the earliest timestamp of the musical phrase involved in the practice segment within the target song, or the timestamp obtained by subtracting a preset time from the earliest timestamp. The end timestamp of a practice segment is the latest timestamp of the musical phrase involved in the practice segment within the target song, or the timestamp obtained by adding a preset time to the latest timestamp. In this way, this application generates buffered start and end times by adding a preset time (e.g., 500 milliseconds) buffer before and after the start and end times of each practice segment, thus avoiding abrupt start and end times for the practice segments.
[0050] The total performance time of a practice segment is equal to the end timestamp of the practice segment minus the start timestamp. The evaluation score of a practice segment is determined based on the evaluation scores of the musical phrases involved in the segment; the evaluation score for each phrase is obtained after performing a phrase-level evaluation of the performance audio. The weakness type of a practice segment is determined based on the dimensional scores of the practice segment across various preset evaluation dimensions. For example, a weakness type could be low pitch, unstable rhythm, or unstable breath control. The difficulty level of a practice segment is determined based on its evaluation score. For example, a score of 0-59 corresponds to a high difficulty level, a score of 60-79 to a medium difficulty level, and a score of 80-100 to a low difficulty level.
[0051] In this way, this application divides the song into segments based on the boundaries of musical phrases, and merges and judges the musical phrases obtained after the segmentation to finally obtain each practice segment. This not only ensures that each practice segment is complete in terms of musical structure (without being cut off in the middle of the musical phrase), but also ensures that the singing length of each practice segment is moderate (e.g., 3-10 seconds, suitable for focused and repeated practice) and that the boundaries are smooth (by adding buffers to avoid abrupt beginnings and endings).
[0052] Step S15: Execute the progressive singing practice process corresponding to the practice segment to guide the target user to complete the singing practice for the practice segment.
[0053] This application divides the song segment to be practiced into practice segments, and then executes a progressive practice process corresponding to each practice segment to guide the target user in practicing the corresponding practice segment. It should be noted that the progressive practice process can be a staged, increasingly difficult challenge-based practice, that is, the progressive practice process includes practice levels of increasing difficulty, and each increasingly difficult practice level corresponds to a practice assistance mode with decreasing levels of assistance.
[0054] When there is more than one practice segment, the current practice segment is determined from the practice segments according to the singing order. The progressive practice process corresponding to the current practice segment is unlocked and executed (i.e., entering the challenge practice corresponding to the current practice segment) to guide the target user to complete the practice of the current practice segment. After passing the challenge practice corresponding to the current practice segment, the next practice segment is determined from the practice segments according to the singing order. The progressive practice process corresponding to the next practice segment is unlocked and executed (i.e., entering the challenge practice corresponding to the next practice segment) to guide the target user to complete the practice of the next practice segment. This process continues until all practice segments are completed and the challenge practice corresponding to each practice segment is passed, thus completing the practice of the song segment to be practiced.
[0055] It should be noted that the progressively more difficult practice levels corresponding to the practice segment are initially locked. The next practice level is only unlocked after the previous one is completed. The first practice level is automatically unlocked when the progressive practice process corresponding to the practice segment begins. That is, in executing the progressive practice process corresponding to the practice segment, this application first automatically unlocks the first practice level and guides the target user to practice the first practice level for the practice segment. After completing the first practice level, the next practice level is automatically unlocked, and the target user is guided to practice the next practice level for the practice segment, and so on, until all practice levels are completed, thus completing the practice of the practice segment and helping the user gradually master the practice segment from easy to difficult.
[0056] When guiding target users to practice the practice segment for the current practice level, this application guides target users to complete the practice of the practice segment based on the target practice assistance mode corresponding to the current practice level, so as to obtain the practice audio; uses a preset audio evaluation system to evaluate the practice audio in preset evaluation dimensions to determine the evaluation result corresponding to the practice audio; and determines whether the current practice level is passed based on the evaluation result corresponding to the practice audio.
[0057] It should be noted that the practice levels with increasing difficulty correspond to practice assistance modes with decreasing levels of assistance. That is, as the difficulty of the practice level increases, the level of assistance in the corresponding practice assistance mode will decrease. The purpose is to achieve a gradual transition mechanism from assistance to no assistance, thereby simulating the song practice process that conforms to the cognitive law of skill acquisition (from familiarity to reinforcement to mastery).
[0058] Taking a singing practice course with increasing difficulty, consisting of three levels, as an example, the first level's practice assistance mode can play both the original vocal track and the instrumental track of the target song, and also highlight and display the lyrics of the practice segment word by word. The second level's practice assistance mode can play only the instrumental track of the target song (excluding the original vocal track), while simultaneously displaying the standard pitch curve of the practice segment and the real-time pitch curve of the target user singing the practice segment, guiding the target user to adjust their singing based on pitch feedback. The third level's practice assistance mode can play only the instrumental track of the target song, without providing any other assistance information; the target user sings independently based solely on the instrumental track.
[0059] Specifically, acquiring the practice audio can include: obtaining the start and end timestamps corresponding to the practice segment; positioning the playback pointer to the start position in the target audio track corresponding to the start timestamp; the target audio track is the auxiliary audio track to be played in the target practice mode, which includes the original vocal track and / or accompaniment track of the target song; starting playback of the target audio track from the start position to guide the target user to begin practicing the practice segment, and starting the audio recording function; when the playback pointer reaches the end position, ending the audio recording function to obtain the practice audio obtained by the target user after completing the practice of the practice segment; the end position is the playback position in the target audio track corresponding to the end timestamp.
[0060] It should be noted that the start timestamp of the practice segment is the first timestamp, or the timestamp obtained by subtracting the preset time from the first timestamp; the first timestamp is the earliest timestamp of the musical phrase involved in the practice segment in the target song. The end timestamp of the practice segment is the second timestamp, or the timestamp obtained by adding the preset time to the second timestamp; the second timestamp is the latest timestamp of the musical phrase involved in the practice segment in the target song.
[0061] Furthermore, after determining whether the current singing practice level has been passed based on the evaluation results corresponding to the singing practice audio, if the current singing practice level has not been passed, the playback pointer will be reset to the playback start position in the target audio track to guide the target user to practice the current singing practice level again for the singing practice segment.
[0062] Specifically, during the progressive practice process corresponding to the practice segment (i.e., entering the challenge practice corresponding to the practice segment), the player instance is first initialized, the complete audio resources of the target song (including the accompaniment track and the original vocal track) are preloaded, and an audio decoding buffer is established to store the decoded accompaniment track and the original vocal track. The first practice level is then unlocked. In the corresponding practice assistance mode, the start and end timestamps of the practice segment are obtained. The underlying audio engine's seek interface is used to position the playback pointer to the starting position on the target track corresponding to the start timestamp. The target track is the auxiliary track to be played in the practice assistance mode for the first practice level. Playback of the target track begins from the starting position to guide the user to practice the practice segment. Audio recording is initiated, and the playback pointer's position is monitored in real-time. When the playback pointer reaches the ending position on the target track corresponding to the end timestamp, audio recording ends to obtain the practice audio after the user completes the practice segment. A preset audio evaluation system is used to evaluate the practice audio across preset dimensions to determine the evaluation result. If the evaluation result indicates failure to pass the first practice level, the playback pointer is repositioned back to the starting position on the target track, achieving seamless loop playback to guide the user to practice the first practice level again. If the evaluation results based on the practice audio indicate that the first practice level has been passed, and there is a next practice level that does not require switching to a new practice segment, then the next practice level is unlocked, and the playback pointer is reset to the starting position on the target audio track to guide the user to practice the next practice level for that practice segment. If the evaluation results based on the practice audio indicate that the current practice level has been passed, and a switch to a new practice segment is required, then the progressive practice process corresponding to the new continuous singing segment is unlocked and executed (i.e., entering the challenge practice of the new continuous singing segment).
[0063] During the progressive practice process corresponding to the practice segment, if a preset retry event for the current practice level is detected, the playback and audio recording functions of the target audio track are immediately interrupted, and the playback pointer is repositioned back to the playback start position on the target audio track. This achieves seamless positioning and jump, guiding the target user to practice the current practice level again for the practice segment. The preset retry event can be the event where the target user clicks "try again."
[0064] In audio playback, seek technology refers to the technique of jumping to a specific time point or position in an audio file by controlling the playback position. This application is based on precise seek, which can achieve millisecond-level audio positioning and jumping, and supports segment-level audio loop playback, manual retry and segment switching, providing underlying playback capabilities to support segmented singing practice.
[0065] Furthermore, the seek accuracy of this application does not exceed a deviation of 50 milliseconds. When dividing practice segments, this application adds a preset time buffer (e.g., 500 milliseconds) before the start time and after the end time for each practice segment, thereby avoiding seeking into the middle of the practice segment and avoiding abrupt start and end times. Moreover, when repositioning the playback pointer back to the playback start position in the target audio track, the loop interval is 0 milliseconds, achieving seamless audio looping. When switching to a new practice segment, the seek operation is performed based on the start and end timestamps corresponding to the new practice segment, with a track switching delay of no more than 100 milliseconds, exhibiting low time latency and improving the user's practice experience.
[0066] In one example, the evaluation result corresponding to the practice audio includes a challenge score for the practice audio. Accordingly, based on the evaluation result corresponding to the practice audio, determining whether the current practice level has been passed includes: judging whether the challenge score corresponding to the practice audio is not less than the preset score corresponding to the current practice level; if it is not less than, the current practice level is passed; if it is less than, the current practice level is not passed.
[0067] The practice levels, with increasing difficulty, correspond to different preset scores. For example, if there are three practice levels with increasing difficulty, the preset score for the first practice level can be 70 points, the preset score for the second practice level can be 80 points, and the preset score for the third practice level can be 90 points.
[0068] In one scenario, if the evaluation results corresponding to the practice audio indicate that the current practice level has been passed, then it is determined whether the current practice level is the last practice level. If so, it signifies the completion of the progressive practice process corresponding to the practice segment, i.e., passing the challenge practice of the practice segment. If not, then the target user is guided to practice the next practice level for the practice segment.
[0069] In another scenario, if the evaluation results based on the practice audio indicate that the user has failed the current practice level, the user is guided to practice the current practice level again for the practice segment. Furthermore, this application can determine the cumulative number of failures for the current practice level and, when the cumulative number of failures reaches a preset number, lower the preset score for the current practice level. The lowering operation involves reducing the preset score by a preset value, such as 5.
[0070] It should be noted that after lowering the preset score, you can either apply the lowered preset score directly, or send the lowered preset score to the user for confirmation and then apply it after confirmation.
[0071] It should also be noted that the cumulative number of failures in the current practice level will be reset to zero and the count will start again each time the preset number of failures is reached.
[0072] Furthermore, after determining whether the current singing practice level has been passed based on the evaluation results corresponding to the singing practice audio, this application can also update the current progress status based on the judgment result representing whether the current singing practice level has been passed. The current progress status includes the current level identifier, the level status of all singing practice levels corresponding to the singing practice segment and the highest historical progress score, the total number of practice sessions for the singing practice segment, and the singing practice completion status; the current level identifier is the identifier of the current singing practice level; level status includes unlocked, in progress, and passed; singing practice completion status includes incomplete singing practice and completed singing practice.
[0073] Specifically, such as Figure 3As shown, the singing practice system running on the backend server, during the progressive singing practice process corresponding to the singing practice segment, first initializes the current challenge state of the singing practice segment through the state manager. This includes the current level identifier (initially the first singing practice level, at which point the current level identifier is 1), the level status of all singing practice levels corresponding to the singing practice segment (initially all are locked, later becoming active if they become active, and passed if they become passed), the highest historical challenge score of all singing practice levels corresponding to the singing practice segment (initially all are 0), the total number of practice attempts for the singing practice segment (initially 0), and the completion status of the singing practice segment (initially false if the singing practice is not completed, later becoming true if all singing practice levels are passed). Then, the initialized current challenge state is persistently stored in the storage service. If the storage is successful, the practice interface of the current singing practice level (initially the first singing practice level) is displayed on the user's end based on the current challenge state to guide the target user to practice the current singing practice level of the singing practice segment.
[0074] During the practice of the current singing practice level, the singing practice system obtains the singing practice audio obtained by the target user after singing the practice segment. It uses a preset audio evaluation system to evaluate the singing practice audio in preset evaluation dimensions to determine the challenge score corresponding to the singing practice audio. The status manager determines whether the challenge score corresponding to the singing practice audio is not less than the preset score corresponding to the current singing practice level. If it is not less than, the current singing practice level is considered passed; if it is less than, the current singing practice level is considered not passed.
[0075] If the current practice level is passed and it is the last practice level, the current level status is updated through the status manager. This includes updating the level status from "active" to "passed", determining whether to update the highest historical score of the current practice level based on the level score corresponding to the practice audio, incrementing the total number of practice attempts by one, and updating the practice completion status from "false" (never completed) to "true" (completed). Everything else remains unchanged. The updated current level status is then persistently stored in the storage service, and the progressive practice process corresponding to the new practice segment is unlocked and executed.
[0076] If the current practice level is passed and is not the last practice level, the current level status is updated through the status manager. This includes updating the current level identifier to the identifier of the next practice level, updating the current practice level status from "active" to "passed", updating the next practice level status from "locked" to "active", and determining whether to update the highest historical score of the current practice level based on the level score corresponding to the practice audio. At the same time, the total number of practice attempts is incremented by one, while other aspects remain unchanged. The updated current level status is then persistently stored in the storage service, and the practice interface of the next practice level is displayed on the user's device based on the current level status to guide the target user to practice the next practice level for the practice audio segment.
[0077] If the current practice level is not passed, the current challenge status is updated through the status manager. This includes determining whether to update the highest historical challenge score for the current practice level based on the challenge score corresponding to the practice audio, incrementing the total number of practice attempts by one, keeping other settings unchanged, and then persistently storing the updated current challenge status in the storage service. The target user is then guided to practice the current practice level again for the practice audio segment.
[0078] For example, the data structure for the current level state is defined as follows: segmentId (the segment identifier of the practice segment currently being practiced), currentLevel (an integer, the identifier of the current level), levelStates (an object, the level state of all practice levels corresponding to the practice segment, with the key being level and the value being locked / active / passed), levelBestScores (an object, the highest historical level score of all practice levels corresponding to the practice segment, with the key being level and the value being the highest historical level score), totalAttempts (an integer, the total number of practice attempts for the practice segment), isCompleted (a boolean value, indicating whether the practice for the practice segment has been completed, false indicating that the practice has not been completed, true indicating that the practice has been completed), and consequentFailures (an integer, the cumulative number of failures for the current practice level).
[0079] It should be noted that this application can display the progress bar of the current practice level on the practice interface of the current practice level of the practice segment, and can also display the overall progress bar of each practice segment. For example, different colored progress bars can be used to indicate the progress status of each practice segment. A gray progress bar indicates that the practice segment has not yet been unlocked, that is, it has not yet been attempted. A blue progress bar indicates that the practice segment is in progress. A green progress bar indicates that the practice segment has been completed.
[0080] It should also be noted that the current progress status is persistently stored in the data storage layer. Furthermore, this application adopts a hierarchical segmented scoring storage structure to fully record the total number of practice sessions and the completion status of each practice segment, as well as the historical progress scores, level status, and number of practice sessions for each practice segment in each practice level, making the entire practice process quantifiable and traceable.
[0081] Furthermore, if an exit event is detected during the execution of the progressive singing practice process, this application records the current singing practice progress so as to perform a breakpoint recovery operation based on the current singing practice progress.
[0082] In other words, if the target user is detected to have quit midway through the progressive singing practice process, the current singing practice progress is recorded. When the target user resumes the progressive singing practice process for the corresponding singing practice segment later, the breakpoint recovery operation is performed based on the current singing practice progress. This allows the target user to continue practicing the singing practice segment based on the previous singing practice progress, avoiding the target user having to start over from the beginning and improving the user's singing practice experience.
[0083] In addition, this application triggers the corresponding display operation when a preset event is detected. The preset events include completing a practice level, passing all practice levels corresponding to a practice segment, and completing practice of all practice segments; the display operation includes displaying practice results and / or displaying special effects.
[0084] For example, after completing each practice session, a star rating can be displayed based on the evaluation result corresponding to the practice audio, such as three stars, two stars, one star, or no stars. Completing a practice session and passing the level can display a star flying in effect; passing all practice sessions corresponding to the practice segment can display a star exploding effect; and completing all practice segments can display a fireworks celebration effect, thereby motivating the user.
[0085] It's worth noting that the introduction of gamification elements such as level-based mechanics, star ratings, and level-clearing effects transforms the originally tedious singing practice into a game experience with clear goals and immediate feedback. Each practice session has a specific goal (reaching a designated score), and each attempt provides instant scoring and star ratings, with visual effects rewards upon completion. This design leverages flow theory and variable reward mechanisms from game psychology, effectively enhancing users' motivation to practice.
[0086] The practice results display includes, but is not limited to, showing the number of practice sessions, practice time, best performance score, progress curve, and weaknesses improved for each practice level. The progress curve and weakness improvement information are based on the performance score for each practice session. The progress curve shows the trend of the performance score increasing with the number of practice sessions, while the weakness improvement information shows a comparison of the performance score before and after practice. In this way, the practice results display helps users clearly see their progress trajectory in each weak area. This quantitative positive feedback mechanism further strengthens user motivation and product stickiness.
[0087] It should be noted that the method in this application is implemented through a singing practice system running on a backend server. The overall architecture of the singing practice system is as follows: Figure 4 As shown, the overall architecture of the singing practice system adopts a layered design, divided into four layers from top to bottom: user interaction layer, application logic layer, core engine layer, and data storage layer. Each layer has a clear responsibility and is loosely coupled and collaborative.
[0088] The user interaction layer is responsible for presenting users with UI (User Interface) interactive interfaces such as the practice suggestion panel (for displaying the song segments to be practiced), the challenge practice interface, the progress display panel (for displaying the challenge progress), and the results report page (for displaying the practice results). It can also receive user operations and pass them to the application logic layer.
[0089] The application logic layer is the core control center of the business process, including the practice process controller (coordinating the overall practice flow, such as the flow of practice segments and the flow of practice segments in various practice levels), the status manager (managing the progress status and pass determination of practice segments), the progress tracking management (managing breakpoint recovery and progress synchronization), and the UI animation engine (managing star rating and triggering and rendering of pass effects).
[0090] The core engine layer provides underlying technical capabilities, including a preset audio evaluation system (such as an AI evaluation engine that provides multi-dimensional singing scores), an intelligent slicing engine (which executes an adaptive slicing algorithm based on musical phrase boundaries), a precise Seek engine (which provides millisecond-level audio positioning and loop playback), and a pitch analysis engine (which provides real-time pitch detection and pitch curve comparison).
[0091] The data storage layer is responsible for persistently storing scoring history (the progress of each practice segment), progress, song timeline data (the timeline data of each musical phrase in the song, including the start and end timestamps of each phrase) and user configuration data, and supports local caching and cloud synchronization.
[0092] In addition to being applicable to song practice, the method of this application can also be transferred to other practice areas, such as instrument practice and language practice. This application transfer utilizes the innovative combination of evaluation, slicing, and progressive challenge techniques in this application.
[0093] For example, in the field of piano practice, an AI-based evaluation system can evaluate a user's audio recordings measure by measure, automatically identifying weak measures such as wrong notes and rhythmic deviations. These measures are then divided into independent practice segments to guide the user through challenge-based practice. The challenge-based practice assistance modes can be adjusted as follows: Level 1: Follow-along mode (plays demonstration audio with highlighted keys); Level 2: Score reading mode (displays sheet music but does not play demonstration audio); Level 3: Blindfolded playing mode (plays entirely from memory without assistance).
[0094] For example, in the field of language practice, an AI-based assessment system can evaluate a user's spoken reading, identify words or sentences with inaccurate pronunciation, and automatically segment them into independent practice segments to guide the user through challenge-based practice for each segment. The challenge-based practice assistance modes can be adjusted as follows: Level 1: Follow-along mode (plays standard pronunciation with phonetic prompts); Level 2: Text-reading mode (displays text but does not play standard pronunciation); Level 3: Dictation mode (reads aloud entirely from memory without assistance).
[0095] In addition, in order to reduce the reliance on the accuracy of artificial intelligence evaluation and reduce resource consumption, this application can appropriately sacrifice automation and objectivity, that is, change the identification of weak segments from automatic identification by artificial intelligence to user self-marking. For example, after completing the singing of the target song, the user can manually mark the song segments that they think need to be practiced. Then, based on the song segments marked by the user, intelligent segmentation and challenge practice are carried out, thereby realizing lightweight song practice.
[0096] Therefore, this application utilizes a preset audio evaluation system to evaluate the singing audio across preset evaluation dimensions. This allows for the objective and accurate identification of the user's weakest singing sections within the target song, helping the user focus their attention on those sections that truly need improvement and thus better master the target song. Furthermore, this application divides song sections into fine-grained practice segments based on musical phrase boundaries for challenging practice. This not only avoids disrupting the musical structure during segment division but also keeps the length of each practice session within a reasonable range, preventing the need for repeated loops when the song is long. Compared to traditional full-song loop practice, this application shortens the effective practice time from 3-5 minutes for the entire song to 3-10 seconds for a single segment, potentially reducing the total practice time required to achieve the same effect by more than 50%. This precise and targeted practice method avoids wasting time on already mastered sections, significantly improving the efficiency of song practice and helping users more accurately master each practice segment.
[0097] by Figure 5 Taking an example, a song practice assistance method provided by an embodiment of the present invention will be described in detail, specifically including: The system obtains the singing audio of the target user after they sing the target song, and uses an AI-based evaluation system to evaluate the singing audio in multiple dimensions and segment by segment to determine the evaluation results for each segment of the target song. Based on the evaluation results for each segment, it determines whether there are any weak segments in each segment that need to be practiced. If not, the practice for the target song ends.
[0098] If there are sections of the song that need to be practiced but the singers are weak, then the sections of the song need to be divided based on the boundaries of the musical phrases in those sections, so as to obtain several practice sections.
[0099] The system selects the current practice segment from the practice segments according to the singing order and unlocks the challenge practice for that segment. This guides the target user to practice the current singing challenge for that segment and determines whether they have passed it. If they have not passed, they are guided to practice the current singing challenge for that segment again. If they have passed, the system further determines whether they have passed all singing challenges for that segment. If they have not passed all singing challenges for that segment, they are guided to practice the next singing challenge for that segment. If they have passed all singing challenges for that segment, the system determines whether they have passed the challenge practice for all singing challenges for that segment. If they have not passed the challenge practice for all singing challenges for that segment, the system selects the next practice segment from the practice segments according to the singing order and unlocks the challenge practice for that segment. If they have passed the challenge practice for all singing challenges for that segment, the practice for the target song ends.
[0100] Therefore, this application utilizes an AI-based evaluation system to assess the singing audio across preset evaluation dimensions. This objectively and accurately identifies the user's weakest singing sections within the target song, helping the user focus their attention on those sections that truly need improvement and thus better master the target song. Furthermore, this application divides song sections into fine-grained practice segments based on musical phrase boundaries for challenging practice. This not only avoids disrupting the musical structure during segment division but also keeps the length of each practice session within a reasonable range. This avoids the need for repeated practice sessions when the song is long, improving the efficiency of song practice and helping users more accurately master each practice segment.
[0101] Furthermore, embodiments of this application also disclose an electronic device, Figure 6 This is a structural diagram of an electronic device 10 according to an exemplary embodiment. The content of the diagram should not be construed as limiting the scope of this application.
[0102] Figure 6 This is a schematic diagram of the structure of an electronic device 10 provided in an embodiment of this application. Specifically, the electronic device 10 may include: at least one processor 11, at least one memory 12, a power supply 13, a communication interface 14, an input / output interface 15, and a communication bus 16. The memory 12 stores a computer program, which is loaded and executed by the processor 11 to implement the relevant steps in the song practice assistance method disclosed in any of the foregoing embodiments. Furthermore, the electronic device 10 in this embodiment may specifically be an electronic computer.
[0103] In this embodiment, the power supply 13 is used to provide operating voltage for each hardware device on the electronic device 10; the communication interface 14 can create a data transmission channel between the electronic device 10 and external devices, and the communication protocol it follows can be any communication protocol applicable to the technical solution of this application, and is not specifically limited here; the input / output interface 15 is used to acquire external input data or output data to the outside world, and its specific interface type can be selected according to specific application needs, and is not specifically limited here.
[0104] In addition, the memory 12, as a carrier for resource storage, can be a read-only memory, random access memory, disk or optical disk, etc. The resources stored thereon can include operating system 121, computer program 122, etc., and the storage method can be temporary storage or permanent storage.
[0105] The operating system 121 is used to manage and control the various hardware devices on the electronic device 10 and the computer program 122, which may be Windows Server, Netware, Unix, Linux, etc. In addition to including a computer program capable of performing the song practice assistance method executed by the electronic device 10 as disclosed in any of the foregoing embodiments, the computer program 122 may further include a computer program capable of performing other specific tasks.
[0106] Furthermore, this application also discloses a computer-readable storage medium for storing a computer program; wherein, when the computer program is executed by a processor, it implements the aforementioned song practice assistance method. Specific steps of this method can be found in the corresponding content disclosed in the foregoing embodiments, and will not be repeated here.
[0107] Furthermore, this application also discloses a computer program product, including a computer program / instructions, wherein the computer program / instructions, when executed by a processor, implement the aforementioned song practice assistance method. Specific steps of this method can be found in the corresponding content disclosed in the foregoing embodiments, and will not be repeated here.
[0108] The various embodiments in this specification are described in a progressive manner, with each embodiment focusing on its differences from other embodiments. Similar or identical parts between embodiments can be referred to interchangeably. For the apparatus disclosed in the embodiments, since it corresponds to the method disclosed in the embodiments, the description is relatively simple; relevant parts can be referred to in the method section.
[0109] Those skilled in the art will further recognize that the units and algorithm steps of the various examples described in conjunction with the embodiments disclosed herein can be implemented in electronic hardware, computer software, or a combination of both. To clearly illustrate the interchangeability of hardware and software, the components and steps of the various examples have been generally described in terms of functionality in the foregoing description. Whether these functions are implemented in hardware or software depends on the specific application and design constraints of the technical solution. Those skilled in the art can use different methods to implement the described functions for each specific application, but such implementation should not be considered beyond the scope of this application.
[0110] The steps of the methods or algorithms described in conjunction with the embodiments disclosed herein can be implemented directly by hardware, a software module executed by a processor, or a combination of both. The software module can be located in random access memory (RAM), main memory, read-only memory (ROM), electrically programmable ROM, electrically erasable programmable ROM, registers, hard disk, removable disk, CD-ROM, or any other form of storage medium known in the art.
[0111] Finally, it should be noted that in this document, relational terms such as "first" and "second" are used only to distinguish one entity or operation from another, and do not necessarily require or imply any such actual relationship or order between these entities or operations. Furthermore, the terms "comprising," "including," or any other variations thereof are intended to cover non-exclusive inclusion, such that a process, method, article, or apparatus that comprises a list of elements includes not only those elements but also other elements not expressly listed, or elements inherent to such a process, method, article, or apparatus. Without further limitations, an element defined by the phrase "comprising one..." does not exclude the presence of other identical elements in the process, method, article, or apparatus that includes said element.
[0112] The technical solutions provided in this application have been described in detail above. Specific examples have been used to illustrate the principles and implementation methods of this application. The descriptions of the above embodiments are only for the purpose of helping to understand the methods and core ideas of this application. At the same time, for those skilled in the art, there will be changes in the specific implementation methods and application scope based on the ideas of this application. Therefore, the content of this specification should not be construed as a limitation of this application.
Claims
1. A method for assisting in song practice, characterized in that, include: Obtain the audio recording of the target user singing the target song; The singing audio is evaluated at the segment level to obtain the evaluation results for each segment of the target song. Based on the evaluation results corresponding to each of the aforementioned song segments, the song segments to be practiced are determined from each of the aforementioned song segments; Based on the song structure information corresponding to the song segment to be practiced, the song segment to be practiced is divided to generate practice segments; The progressive practice process corresponding to the practice segment is executed to guide the target user to complete the practice of the practice segment.
2. The song practice aid method according to claim 1, characterized in that, The evaluation results for each song segment include the evaluation score for each song segment. Accordingly, determining the song segment to be practiced from each of the song segments based on the evaluation results corresponding to each of the song segments includes: Based on the evaluation scores corresponding to each of the song segments, the song segments with evaluation scores lower than the preset scores are identified from the song segments to be practiced.
3. The song practice aid method according to claim 1, characterized in that, The step of dividing the song segment to be practiced into practice segments based on the song structure information corresponding to the segment to be practiced includes: Based on the phrase boundaries in the song segment to be practiced, the song segment to be practiced is divided into phrases to obtain the target phrases; Merge the musical phrases in the target musical phrase that meet the preset merging conditions to obtain the merged musical phrase; the preset merging conditions include that the musical phrases to be merged are adjacent in the target song and the sum of their singing durations is not greater than a preset duration; A practice segment is generated based on the merged musical phrase and the unmerged musical phrase in the target musical phrase.
4. The song practice aid method according to claim 3, characterized in that, The process of dividing the song segment to be practiced into musical phrases to obtain the target musical phrase includes: Each musical phrase obtained by dividing the song segment to be practiced into musical phrases is determined as the target musical phrase; Alternatively, based on the evaluation results corresponding to each musical phrase, a target musical phrase can be determined from each musical phrase; the evaluation results corresponding to each musical phrase are the results obtained after performing a musical phrase-level evaluation on the singing audio.
5. The song practice aid method according to claim 1, characterized in that, The progressive singing practice process includes singing practice levels with increasing difficulty, and each singing practice level with increasing difficulty corresponds to a singing practice assistance mode with decreasing assistance level. Accordingly, the process of executing the progressive practice singing process corresponding to the practice segment includes: Based on the target singing practice assistance mode corresponding to the current singing practice level, the target user is guided to complete the singing practice of the singing practice segment in order to obtain the singing practice audio; The practice audio is evaluated to determine the evaluation result corresponding to the practice audio. Based on the evaluation results corresponding to the practice audio, determine whether the current practice level has been passed.
6. The song practice aid method according to claim 5, characterized in that, The target singing assistance mode, based on the current singing practice level, guides the target user to complete the singing practice segment to obtain the singing practice audio, including: Obtain the start and end timestamps corresponding to the practice segment; Position the playback pointer to the playback start position in the target audio track corresponding to the start timestamp; the target audio track is the auxiliary audio track that needs to be played in the target singing practice auxiliary mode, and the auxiliary audio track includes the original vocal track and / or accompaniment track of the target song; The target audio track is played from the playback start position to guide the target user to start practicing the practice segment and to start the audio recording function. When the playback pointer is detected to have reached the playback termination position, the audio recording function is terminated to obtain the practice audio; the playback termination position is the playback position in the target audio track corresponding to the termination timestamp.
7. The song practice aid method according to claim 6, characterized in that, After determining whether the current singing practice level has been passed based on the evaluation results corresponding to the singing practice audio, the process further includes: If the current singing practice level is not passed, the playback pointer will be reset to the playback start position to guide the target user to practice the current singing practice level again for the practice segment.
8. The song practice aid method according to claim 6, characterized in that, The starting timestamp is a first timestamp, or a timestamp obtained by subtracting a preset time from the first timestamp; the first timestamp is the earliest timestamp of the practice segment in the target song; The termination timestamp is the second timestamp, or the timestamp obtained by adding the preset time to the second timestamp; The second timestamp is the latest timestamp of the practice segment in the target song.
9. The song practice aid method according to claim 5, characterized in that, The evaluation results corresponding to the practice audio include the challenge score corresponding to the practice audio. Accordingly, determining whether to pass the current singing practice stage based on the evaluation results corresponding to the singing practice audio includes: Determine whether the challenge score corresponding to the singing practice audio is not less than the preset score corresponding to the current singing practice level; If it is not less than, then the current singing practice level is considered passed; If the value is less than the threshold, then the current singing practice level has not been passed.
10. The song practice aid method according to claim 9, characterized in that, After determining whether the current singing practice level has been passed based on the evaluation results corresponding to the singing practice audio, the process further includes: If the current singing practice level is not passed, the cumulative number of failures for the current singing practice level is determined, and when the cumulative number of failures reaches a preset number, the preset score is lowered.
11. The song practice aid method according to claim 9, characterized in that, After determining whether the current singing practice level has been passed based on the evaluation results corresponding to the singing practice audio, the process further includes: Update the current progress status based on the judgment result of whether the current singing practice level has been passed; The current challenge status includes the current level identifier, the level status of all practice levels corresponding to the practice segment and the highest historical challenge score, the total number of practice sessions for the practice segment, and the practice completion status.
12. The song practice aid method according to claim 9, characterized in that, Also includes: When a preset event is detected, the display operation corresponding to the preset event is triggered; The preset events include completing a practice session, passing all practice sessions corresponding to the practice segment, and completing practice sessions for all the practice segments; the display operations include displaying practice results and / or displaying special effects.
13. The song practice aid method according to any one of claims 1 to 12, characterized in that, Also includes: If an exit event is detected during the execution of the progressive singing practice process, the current singing practice progress is recorded so that a breakpoint recovery operation can be performed based on the current singing practice progress.
14. An electronic device, characterized in that, include: Memory is used to store computer programs; A processor for executing the computer program to implement the song practice assistance method as described in any one of claims 1 to 13.
15. A computer-readable storage medium, characterized in that, Used to store a computer program, which, when executed by a processor, implements the song practice assistance method as described in any one of claims 1 to 13.
16. A computer program product comprising a computer program / instructions, characterized in that, When the computer program / instructions are executed by the processor, they implement the song practice assistance method as described in any one of claims 1 to 13.