Lyrics labeling method, device and equipment and storage medium

By identifying the target background music and its style, and generating and adjusting the lyric annotation information, the problem of inaccurate and inefficient lyric annotation in existing technologies is solved, realizing intelligent lyric annotation and improving the accuracy and efficiency of lyric creation and recording.

CN114282045BActive Publication Date: 2026-01-02BEIJING BAIDU NETCOM SCI & TECH CO LTD
View PDF 1 Cites 0 Cited by

Patent Information

Application Number
CN202111581515.1
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2021-12-22
Publication Date
2026-01-02
Estimated Expiration
2041-12-22

AI Technical Summary

Technical Problem

In existing technologies, users cannot guarantee the accuracy and efficiency of annotation when editing lyrics, and lyrics annotation cannot be completed intelligently.

Method used

By determining the target background music and its style, initial annotation information for preset lyrics is generated, and target annotation information is generated based on the user's adjustment operations. Finally, the target lyrics are generated and intelligently annotated in combination with the target background music and its style.

Benefits of technology

It improves the accuracy and efficiency of lyric annotation, allowing users to intuitively and synchronously create lyrics, thus enhancing the creative experience and the efficiency of song recording.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN114282045B_ABST
    Figure CN114282045B_ABST
Patent Text Reader

Abstract

The present disclosure provides a lyrics labeling method and device, equipment and storage medium, relates to the technical field of data processing, and particularly relates to the field of intelligent recommendation and multimedia application. The specific implementation scheme is as follows: a target background music and a corresponding target style thereof are determined; initial labeling information corresponding to preset lyrics is generated based on the target background music and the target style; target labeling information of the preset lyrics is determined based on the initial labeling information corresponding to the preset lyrics; and target lyrics of the target background music are generated based on the preset lyrics and the corresponding target labeling information. According to the technical scheme of the present disclosure, the accuracy and efficiency of labeling can be ensured.
Need to check novelty before this filing date? Find Prior Art

Description

TECHNICAL FIELD

[0001] The present disclosure relates to the technical field of data processing, in particular to the technical fields of intelligent recommendation and multimedia application. BACKGROUND

[0002] With the continuous development of Internet technology, more and more users like to create through the network, such as writing and singing songs. Moreover, in order to facilitate users to record songs, the user will also be shown the lyrics written by the user. In related technologies, the user only has simple editing and line changing operations when editing the lyrics, and the accuracy of annotating the lyrics cannot be guaranteed, and the annotation efficiency is low. SUMMARY

[0003] The present disclosure provides a lyrics annotation method, device, equipment, storage medium and computer program product.

[0004] According to a first aspect of the present disclosure, a lyrics annotation method is provided, comprising:

[0005] determining a target background music and a target style corresponding thereto;

[0006] generating initial annotation information corresponding to a preset lyric based on the target background music and the target style;

[0007] determining target annotation information of the preset lyric based on the initial annotation information corresponding to the preset lyric;

[0008] generating a target lyric of the target background music based on the preset lyric and the target annotation information corresponding thereto.

[0009] According to a second aspect of the present disclosure, a lyrics annotation device is provided, comprising:

[0010] a first determining unit configured to determine a target background music and a target style corresponding thereto;

[0011] a first generating unit configured to generate initial annotation information corresponding to a preset lyric based on the target background music and the target style;

[0012] a second determining unit configured to determine target annotation information of the preset lyric based on the initial annotation information corresponding to the preset lyric;

[0013] a second generating unit configured to generate a target lyric of the target background music based on the preset lyric and the target annotation information corresponding thereto.

[0014] According to a third aspect of the present disclosure, an electronic device is provided, comprising:

[0015] at least one processor; and

[0016] a memory in communication with the at least one processor; wherein

[0017] The memory stores instructions executable by the at least one processor, the instructions being executed by the at least one processor to enable the at least one processor to perform the method described above.

[0018] According to a fourth aspect of the present disclosure, a non-transitory computer readable storage medium storing computer instructions for causing a computer to perform the method described above is provided.

[0019] According to a fifth aspect of the present disclosure, a computer program product comprising a computer program which, when executed by a processor, implements the method described above is provided.

[0020] According to the technical solution of the present disclosure, the accuracy and efficiency of labeling can be ensured.

[0021] It should be understood that the content described in this part is not intended to identify key or important features of the embodiments of the present disclosure, nor to limit the scope of the present disclosure. Other features of the present disclosure will become apparent from the following description. BRIEF DESCRIPTION OF DRAWINGS

[0022] The accompanying drawings are used to better understand the present solution and do not limit the present disclosure. Among them:

[0023] Figure 1 is an implementation flowchart of a song lyrics labeling method according to an embodiment of the present disclosure;

[0024] Figure 2 is a selection interface diagram of target background music and target style according to an embodiment of the present disclosure;

[0025] Figure 3 is a relationship diagram between preset lyrics and initial labeling information under a song lyrics creation interface according to an embodiment of the present disclosure;

[0026] Figure 4 is a diagram of changing initial labeling information of preset lyrics to target labeling information of target lyrics under a song lyrics creation interface according to an embodiment of the present disclosure Figure 1 ;

[0027] Figure 5 is a diagram of changing initial labeling information of preset lyrics to target labeling information of target lyrics under a song lyrics creation interface according to an embodiment of the present disclosure Figure 2 ;

[0028] Figure 6 is a diagram of changing initial labeling information of preset lyrics to target labeling information of target lyrics under a song lyrics creation interface according to an embodiment of the present disclosure Figure 3 ;

[0029] Figure 7 is a schematic diagram of a display effect of a target lyric under a recording interface according to an embodiment of the present disclosure;

[0030] Figure 8 is a structural schematic diagram of a lyric labeling device according to an embodiment of the present disclosure;

[0031] Figure 9 is a block diagram of an electronic device for implementing a lyric labeling method according to an embodiment of the present disclosure. DETAILED DESCRIPTION

[0032] Exemplary embodiments of the present disclosure are described below with reference to the accompanying drawings, which include various details of the embodiments of the present disclosure to assist in understanding, and should be considered as merely exemplary. Thus, those of ordinary skill in the art will recognize various changes and modifications of the embodiments described herein, without departing from the scope and spirit of the present disclosure. Also, descriptions of known functions and constructions are omitted in the following description for clarity and conciseness.

[0033] The terms "first", "second", and "third" and the like in the description of the specification embodiments and claims of the present disclosure and the above-described drawings are used to distinguish similar objects, and do not necessarily have to be used to describe a particular order or sequence. In addition, the terms "include" and "have" and any variations thereof are intended to cover non-exclusive inclusion, for example, inclusion of a series of steps or units. The method, system, product, or device is not necessarily limited to those steps or units clearly listed, but can include other steps or units not clearly listed or inherent to the processes, methods, products, or devices.

[0034] The embodiments of the present disclosure provide a lyric labeling method, which can be applied to an electronic device, including but not limited to fixed devices and / or mobile devices. For example, the fixed devices include but are not limited to servers, which can be cloud servers or ordinary servers. For example, the mobile devices include but are not limited to one or more terminals of a mobile phone or a tablet computer. As shown in the figure, the lyric labeling method includes: Figure 1

[0035] S101: determining a target background music and a corresponding target style thereof;

[0036] S102: generating initial labeling information corresponding to a preset lyric based on the target background music and the target style;

[0037] S103: determining target labeling information of the preset lyric based on the initial labeling information corresponding to the preset lyric;

[0038] ​S104: generating target lyrics of the target background music based on the preset lyrics and the corresponding target annotation information.

[0039] In the present disclosure, the target background music can be music without vocals, which can include pure music or music filtered from a singer's voice in a song. The target background music can be background music selected by a user from a background music library, or background music recommended by the system for the user, and the present disclosure does not limit how to obtain the background music.

[0040] In the present disclosure, one target background music can correspond to multiple preset styles, wherein the preset style is a music style corresponding to the target background music. For example, in the case of rap music, the preset style includes but is not limited to pop rap, Latin rap, hardcore rap, comedy rap, etc. For example, in the case of pop music, the preset style includes but is not limited to rock, jazz, country music, etc.

[0041] In the present disclosure, the target style is one preset style determined from at least one preset style corresponding to the target background music.

[0042] Figure 2 The selection interface of the target background music and the target style is shown in the following figures. Figure 2 As shown in the left figure, there are multiple background musics on the song creation interface, and after background music 1 is selected as the target background music, it jumps to the selection interface of the preset style corresponding to the target background music 1 as shown in the right figure. Figure 2 As shown in the right figure, there are multiple preset styles corresponding to the target background music 1, and after preset 2 is selected as the target style, the target background music is determined as background music 1, and the target style is preset 2 corresponding to background music 1.

[0043] It should be understood that Figure 2 The selection interface shown in the figures is one optional specific implementation, and those skilled in the art can make various obvious changes and / or replacements based on the examples Figure 2 The resulting technical solutions still belong to the disclosure of the present disclosure.

[0044] In the present disclosure, the preset lyrics are lyrics edited by a user. For example, the lyrics are input in a voice manner. For another example, the lyrics are manually input through an input device such as a keyboard or a soft keyboard. The present disclosure does not limit the input manner when the user edits the lyrics.

[0045] In the present disclosure, the initial annotation information is annotation content of at least one character in the preset lyrics. The annotation content includes at least one of the following: connected reading, single reading, accent, bass. It should be noted that different annotation contents can be represented by different representation forms. The types of representation forms include but are not limited to: color, line, interval, symbol. It should be noted that the specific correspondence between the annotation content and the representation form can be set or adjusted according to design requirements or user requirements. For example, the upper horizontal line and the space between the lyrics both represent pause or rhythm, the short line represents single reading, the long line represents connected reading, the colored line represents bass, and the colored font represents accent. The same type of representation form can be further subdivided to represent different annotation contents. For example, a thin line represents bass, and a thick line represents accent. For example, red represents emphasis, and black represents non-emphasis. For example, a large interval between lines represents a long pause time, and a small interval between lines represents a short pause time. Since the annotation content is bound to the representation form, the target lyrics displayed on the creation lyrics interface and the recording interface have the same identity, so that the user's inspiration when synchronously creating lyrics can be improved. The user's creation experience can be improved.

[0046] In the present disclosure, the target annotation information is target annotation content of at least one character in the preset lyrics. The target annotation information can be understood as the determined target annotation content of at least one character in the preset lyrics.

[0047] The technical solutions described in the embodiments of the present disclosure determine the target background music and its corresponding target style; generate initial annotation information corresponding to the preset lyrics based on the target background music and the target style; determine the target annotation information of the preset lyrics based on the initial annotation information corresponding to the preset lyrics; and generate the target lyrics of the target background music based on the preset lyrics and the corresponding target annotation information. In this way, the annotation information can be generated in combination with the target background music and its corresponding target style, and then the target lyrics can be generated based on the annotation information, which can more intelligently complete the annotation operation of the lyrics and ensure the accuracy and efficiency of the annotation.

[0048] In some embodiments, determining the target background music includes: obtaining and displaying a plurality of background musics available for editing; and determining a selected background music as the target background music based on a selection operation of the user on any background music in the plurality of background musics. In this way, a plurality of background musics are provided in advance, which reduces the time for the user to select the target background music and helps to improve the annotation efficiency.

[0049] In some embodiments, determining the target background music includes: determining a background music name according to the input information of the user; and taking the background music corresponding to the background music name as the target background music. In this way, the target background music is directly determined according to the input information of the user, which can quickly lock the target background music and help to improve the annotation efficiency.

[0050] It can be understood that the manner of determining the target background music is not limited to the above two manners, and the present disclosure does not limit the manner of determining the target background music, and an exhaustive list is not made here.

[0051] In some embodiments, determining the target background music and the corresponding target style includes:

[0052] In the case where the target background music corresponds to multiple preset styles, the preset style with the highest heat value among the multiple preset styles is taken as the target style according to the heat values of the multiple preset styles.

[0053] For example, the target background music 1 corresponds to five preset styles, which are respectively preset style 1, preset style 2, preset style 3, preset style 4, and preset style 5. The heat values of the five preset styles are ranked as: the heat value of preset style 2 > the heat value of preset style 3 > the heat value of preset style 1 > the heat value of preset style 4 > the heat value of preset style 5. Therefore, preset style 2 is taken as the target style of the target background music 1.

[0054] In this way, by recommending the preset style with the highest heat value to the user and labeling based on the preset style, the accuracy of labeling is improved, and the time for the user to select the preset style is reduced, and the efficiency of labeling is improved.

[0055] In some embodiments, in the case where the target background music corresponds to multiple preset styles, the preset style with the highest priority among the multiple preset styles is taken as the target style according to the priorities of the multiple preset styles.

[0056] Here, the priority ranking can be set according to factors such as user individual characteristics, or user mood, or user-specified priority ranking.

[0057] For example, the target background music 1 corresponds to five preset styles, which are respectively preset style 1, preset style 2, preset style 3, preset style 4, and preset style 5. The priority ranking of the five preset styles is: preset style 3 > preset style 4 > preset style 1 > preset style 5 > preset style 2. Therefore, preset style 3 is taken as the target style of the target background music 1.

[0058] In this way, by recommending the preset style with the highest priority to the user and labeling based on the preset style, the accuracy of labeling is improved, and the time for the user to select the preset style is reduced, and the efficiency of labeling is improved.

[0059] In some embodiments, the initial annotation information corresponding to the preset lyrics is generated according to the target background music and the target style, including: determining at least one candidate annotation content according to the target style; splitting the preset lyrics according to the target background music to obtain at least one splitting result of the preset lyrics; determining annotation content corresponding to each splitting result based on the at least one candidate annotation content; and generating the initial annotation information corresponding to the preset lyrics based on the annotation content corresponding to each splitting result.

[0060] The candidate annotation content at least includes one of the following: bass, accent, connected reading, and single reading.

[0061] The splitting result is a splitting of each character in the preset lyrics. For example, some characters are connected reading, and some lyrics are single reading. For another example, some characters are accent characters, and some characters are bass characters.

[0062] Figure 3 A preset lyrics and initial annotation information relationship diagram under a creative lyrics interface is shown, as shown in Figure 3 The current preset lyrics are "take out the innovative attitude and let our product experience exceed expectations. Friends consciously offer shouts, and exclaim that this special effect is not only five dollars but at least seven dollars". The splitting result is "take out the innovative attitude, let our product experience exceed expectations, friends consciously offer shouts, and exclaim that this special effect is not only five dollars but at least seven dollars". Each character has initial annotation information represented by a line of different lengths above it. Each character has a different line above it. Some characters have long lines, and some characters have short lines.

[0063] In this way, the initial annotation information corresponding to the preset lyrics can be automatically generated based on the target background music and the target style, improving the annotation efficiency and the intelligent degree of annotation.

[0064] In some embodiments, the preset lyrics are split according to the target background music to obtain at least one splitting result of the preset lyrics, including at least one of the following: performing connected reading splitting on the preset lyrics according to the target background music and the meaning of the preset lyrics to obtain at least one splitting result of the preset lyrics; performing single reading splitting on the preset lyrics according to the target background music and the meaning of the preset lyrics to obtain at least one splitting result of the preset lyrics; performing accent splitting on the preset lyrics according to the target background music and the meaning of the preset lyrics to obtain at least one splitting result of the preset lyrics; and performing bass splitting on the preset lyrics according to the target background music and the meaning of the preset lyrics to obtain at least one splitting result of the preset lyrics.

[0065] Taking the preset lyrics "Take the innovative attitude, let our product experience exceed expectations, friends all consciously offer a shout, exclaim that the special effect is not only five dollars but at least seven dollars" as an example, the preset lyrics are split according to the meaning of the target background music and the preset lyrics, and the split result includes "Take the innovative attitude, let our product experience exceed expectations, friends all consciously offer a shout, exclaim that the special effect is not only five dollars but at least seven dollars", wherein "innovation", "product", "all consciously", and "this special effect" are connected readings; "take", "out", "pose", "attitude", "let", "we", "our", "body", "experience", "exceed", "expectations", "expectations", "friends", "friends", "offer", "offer", "shout", "shout", "exclaim", "exclaim", "not", "not", "five", "five", "at least", "at least", and "seven" are single readings; "take out", "pose", "exceed expectations", "exclaim", and "seven dollars" are stressed, and other characters are unstressed.

[0066] In this way, the preset lyrics are split according to the meaning of the target background music and the preset lyrics, so that the split result not only meets the rhythm of the target background music but also fits the meaning of the lyrics, thereby improving the accuracy and efficiency of labeling.

[0067] In some embodiments, based on the preset lyrics and the initial labeling information corresponding thereto, the target labeling information of the preset lyrics is determined, including: in response to an adjustment operation on at least one labeling content of the initial labeling information of the preset lyrics, determining the target labeling content; and based on the target labeling content, determining the target labeling information of the preset lyrics.

[0068] For example, after the system automatically matches the initial labeling information for the preset lyrics, the target labeling information of part of the characters is obtained according to the adjustment operation of the user on the initial labeling information of the part of the characters.

[0069] For another example, after the initial labeling information is set for the preset lyrics according to the editing operation of the user, the target labeling information of part of the characters is obtained according to the adjustment operation of the user on the initial labeling information of the part of the characters.

[0070] Figure 4 The process of changing the initial labeling information of the preset lyrics in the songwriting interface to the target labeling information of the target lyrics is shown Figure 1 As Figure 4 As shown in the left drawing, compared with the initial labeling information in Figure 3 , the font color of "take out" and "pose" is changed, that is, the initial labeling information of "take out" and "pose" is adjusted, and then the initial labeling information of part of other characters is adjusted to finally obtain the target labeling information of the target lyrics as shown in the right drawing of Figure 4 .

[0071] Thus, based on the adjustment of at least one annotation content of the initial annotation information of the preset lyrics, the target annotation information of the preset lyrics is finally obtained, and the annotation accuracy and efficiency are improved.

[0072] To better split the preset lyrics, in some embodiments, before splitting the preset lyrics according to the target background music, the preset lyrics are further divided into sentences to obtain each sentence of the divided lyrics, so as to split each sentence of the divided lyrics to obtain at least one split result of the preset lyrics.

[0073] Thus, the preset lyrics are first divided into sentences, and then each sentence of the divided lyrics is split, which not only reduces the difficulty of splitting, but also makes the splitting targeted and improves the accuracy of splitting.

[0074] In some embodiments, the initial annotation information corresponding to the preset lyrics is generated according to the target background music and the target style, including: in response to an annotation editing operation on the preset lyrics, the initial annotation information corresponding to the preset lyrics is generated based on the annotation editing operation according to the target background music and the target style.

[0075] The annotation editing operation can be understood as an editing operation of the annotation content of each character in the preset lyrics.

[0076] Figure 5 The process of changing the initial annotation information of the preset lyrics on the songwriting interface to the target annotation information is shown Figure 2 As can be seen from the left graph in Figure 5 , the initial annotation information is displayed on the songwriting interface, and the lyrics can be filled based on the initial annotation information, or the initial annotation information can be adaptively modified according to the demand, and finally the target annotation information corresponding to the preset lyrics as shown in Figure 5 the right graph is obtained.

[0077] Figure 6 The process of changing the initial annotation information of the preset lyrics on the songwriting interface to the target annotation information is shown Figure 3 , from Figure 6 As can be seen from the left graph in Figure 6 , after editing the preset lyrics, the annotation content corresponding to each character is edited one by one, and finally the target annotation information corresponding to the preset lyrics as shown in the right graph is obtained.

[0078] Thus, the initial annotation information can be set according to the user's demand, so that the entire creation process is more original and the annotation accuracy is improved.

[0079] Further, in the above scheme, the lyrics annotation method can further include:

[0080] store the target lyrics of the target background music;

[0081] In response to the start of recording, display the target lyrics on the recording interface.

[0082] The "start recording" operation is used to indicate entry into the recording interface, and this disclosure does not limit the method of issuing the "start recording" operation. For example, upon receiving a user click... Figure 5 or Figure 6 The user's action of clicking the touch-sensitive button to "Put on headphones and enter the recording studio" confirms the user's intention to start recording and enters the recording interface. Alternatively, by listening to the user's voice message indicating the start of recording, the user's intention to start recording is confirmed, and the recording interface is entered.

[0083] Figure 7 This shows a schematic diagram illustrating the display effect of the target lyrics on the recording interface. Figure 7 As can be seen, in addition to the lyrics, there are also annotations such as short lines, long lines, spacing between characters, and font color. Figure 7 The target lyrics and Figure 4 The right image and Figure 6 The lyrics saved in the lyrics creation interface shown on the right are the same as the target lyrics. Thus, because the recording interface displays the target lyrics with target annotation information, compared to simply displaying the text, it enriches the display dimensions of the target lyrics, allowing users to clearly and intuitively synchronize their inspiration when creating lyrics. This helps to stimulate users' creative enthusiasm and inspiration, and further improves the efficiency of song recording.

[0084] The lyrics annotation method disclosed herein can be used in projects such as lyrics annotation, intelligent recommendation, or multimedia applications. For example, the subject executing the method can be an electronic device, such as a mobile phone, tablet, or desktop computer; or it can be a server, such as a regular server or a cloud server.

[0085] This disclosure also provides a lyrics annotation device, such as... Figure 8 As shown, the lyrics annotation device includes:

[0086] The first determining unit 810 is used to determine the target background music and its corresponding target style;

[0087] The first generation unit 820 is used to generate initial annotation information corresponding to preset lyrics based on the target background music and target style;

[0088] The second determining unit 830 is used to determine the target annotation information of the preset lyrics based on the initial annotation information corresponding to the preset lyrics;

[0089] The second generation unit 840 is used to generate target lyrics for target background music based on preset lyrics and their corresponding target annotation information.

[0090] In some embodiments, the first determining unit 810 is specifically configured to:

[0091] In the case where the target background music corresponds to multiple preset styles, the preset style with the highest heat value among the multiple preset styles is taken as the target style according to the heat values of the multiple preset styles.

[0092] Alternatively, in the case where the target background music corresponds to multiple preset styles, the preset style with the highest priority among the multiple preset styles is taken as the target style according to the priorities of the multiple preset styles.

[0093] In some embodiments, the first generating unit 820 includes:

[0094] A first determining sub-unit, configured to determine at least one candidate annotation content according to the target style.

[0095] A splitting sub-unit, configured to split the preset lyrics according to the target background music to obtain at least one splitting result of the preset lyrics.

[0096] A second determining sub-unit, configured to determine annotation content corresponding to each of the at least one splitting result based on the at least one candidate annotation content.

[0097] A generating sub-unit, configured to generate initial annotation information corresponding to the preset lyrics based on the annotation content corresponding to each of the at least one splitting result.

[0098] In some embodiments, the splitting sub-unit is specifically configured to:

[0099] perform continuous reading splitting on the preset lyrics according to the target background music and the meaning of the preset lyrics to obtain at least one splitting result of the preset lyrics.

[0100] perform single reading splitting on the preset lyrics according to the target background music and the meaning of the preset lyrics to obtain at least one splitting result of the preset lyrics.

[0101] perform accent splitting on the preset lyrics according to the target background music and the meaning of the preset lyrics to obtain at least one splitting result of the preset lyrics.

[0102] perform bass splitting on the preset lyrics according to the target background music and the meaning of the preset lyrics to obtain at least one splitting result of the preset lyrics.

[0103] In some embodiments, the second determining unit 830 is configured to:

[0104] determine the target annotation content in response to an adjustment operation on at least one annotation content of the initial annotation information of the preset lyrics.

[0105] Based on the target annotation content, target annotation information of the preset lyrics is determined.

[0106] Those skilled in the art should understand that the functions of the processing modules in the lyric annotation device of the embodiments of the present disclosure can be understood with reference to the foregoing description of the lyric annotation method, and the processing modules in the lyric annotation device of the embodiments of the present disclosure can be implemented by an analog circuit that implements the functions of the embodiments of the present disclosure, or can be implemented by the running of software that implements the functions of the embodiments of the present disclosure on an electronic device.

[0107] The lyric annotation device of the embodiments of the present disclosure can guarantee the accuracy and efficiency of annotation.

[0108] In the technical solutions of the present disclosure, the acquisition, storage and application of user personal information comply with relevant laws and regulations and do not violate public order and good customs.

[0109] According to the embodiments of the present disclosure, the present disclosure further provides an electronic device, a readable storage medium and a computer program product.

[0110] Figure 9 A schematic block diagram of an example electronic device 900 that can be used to implement embodiments of the present disclosure is shown. The electronic device is intended to represent various forms of digital computers, such as laptops, desktops, tablets, personal digital assistants, servers, blade servers, mainframes, and other appropriate computers. The electronic device can also represent various forms of mobile devices, such as personal digital processors, cellular telephones, smart phones, wearable devices, and other similar computing devices. The components shown here, their connections and relationships, and their functions, are meant to be examples only, and are not intended to limit implementations of the present disclosure described and / or claimed in this document.

[0111] As shown in Figure 9 The device 900 includes a computing unit 901 that can perform various appropriate actions and processes in accordance with a computer program stored in a Read-Only Memory (ROM) 902 or a computer program loaded from a storage unit 908 into a Random Access Memory (RAM) 903. Various programs and data required for the operation of the device 900 can also be stored in the RAM 903. The computing unit 901, the ROM 902, and the RAM 903 are connected to each other through a bus 904. An Input / Output (I / O) interface 905 is also connected to the bus 904.

[0112] A plurality of components in the device 900 are connected to the I / O interface 905, including: an input unit 906, such as a keyboard, a mouse, etc.; an output unit 907, such as various types of displays, speakers, etc.; a storage unit 908, such as a magnetic disk, an optical disk, etc.; and a communication unit 909, such as a network card, a modem, a wireless communication transceiver, etc. The communication unit 909 allows the device 900 to exchange information / data with other devices through a computer network, such as the Internet, and / or various telecommunication networks.

[0113] The computing unit 901 can be various general-purpose and / or special-purpose processing components with processing and computing capabilities. Some examples of the computing unit 901 include, but are not limited to, a Central Processing Unit (CPU), a Graphics Processing Unit (GPU), various specialized Artificial Intelligence (AI) computing chips, various computing units running machine learning model algorithms, a Digital Signal Processor (DSP), and any appropriate processor, controller, microcontroller, etc. The computing unit 901 performs various methods and processes described above, such as the lyrics annotation method. For example, in some embodiments, the lyrics annotation method can be implemented as a computer software program, which is tangibly contained in a machine-readable medium, such as the storage unit 908. In some embodiments, part or all of the computer program can be loaded and / or installed on the device 900 via the ROM 902 and / or the communication unit 909. When the computer program is loaded to the RAM 903 and executed by the computing unit 901, one or more steps of the lyrics annotation method described above can be performed. Alternatively, in other embodiments, the computing unit 901 can be configured to perform the lyrics annotation method by any other appropriate means, such as by means of firmware.

[0114] The various embodiments of the systems and techniques described above can be implemented in digital electronic circuitry, integrated circuitry, a field programmable gate array (FPGA), an application specific integrated circuit (ASIC), an application-specific standard product (ASSP), a system on chip (SOC), a complex programmable logic device (CPLD), computer hardware, firmware, software, and / or combinations thereof. These various embodiments can include implementation in one or more computer programs that are executable and / or interpretable on a programmable system including at least one programmable processor, which can be special or general purpose, coupled to receive data and instructions from, and to transmit data and instructions to, a storage system, at least one input device, and at least one output device.

[0115] Program code for carrying out methods of the present disclosure can be written in any combination of one or more programming languages. This program code can be provided to a processor or controller of a general or special purpose computer, special purpose computer, or other programmable data processing apparatus to produce a machine, such that the program code, when executed by the processor or controller, produces a means for implementing the functions / operations specified in the flowcharts and / or block diagrams. The program code can execute entirely on a machine, partly on the machine, as a stand-alone software package, partly on the machine and partly on a remote machine or entirely on the remote machine or server.

[0116] In the context of this disclosure, a machine-readable medium can be a tangible medium that contains or stores a program for use by or in connection with an instruction execution system, apparatus, or device. The machine-readable medium can be a machine-readable signal medium or a machine-readable storage medium. A machine-readable medium can include but is not limited to an electronic, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any suitable combination of the foregoing. More specific examples of the machine-readable storage medium would include an electrical connection based on one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or Flash memory), an optical fiber, a portable compact disc read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the foregoing.

[0117] To provide for interaction with a user, the systems and techniques described here can be implemented on a computer having a display device (e.g., a Cathode Ray Tube (CRT) or Liquid Crystal Display (LCD) monitor) for displaying information to the user and a keyboard and a pointing device (e.g., a mouse or a trackball) by which the user can provide input to the computer. Other kinds of devices can be used to provide for interaction with a user as well; for example, feedback provided to the user can be any form of sensory feedback (e.g., visual feedback, auditory feedback, or tactile feedback); and input from the user can be received in any form, including acoustic, speech, or tactile input.

[0118] The systems and techniques described here can be implemented in a computing system that includes a back end component (e.g., as a data server), or that includes a middleware component (e.g., an application server), or that includes a front end component (e.g., a user computer having a graphical user interface or a Web browser through which a user can interact with an implementation of the systems and techniques described here), or any combination of such back end, middleware, or front end components. The components of the system can be interconnected by any form or medium of digital data communication (e.g., a communication network). Examples of communication networks include a local area network (LAN), a wide area network (WAN), and the Internet.

[0119] The computer system can include clients and servers. A client and server are generally remote from each other and typically interact through a communication network. The relationship of client and server arises by virtue of computer programs running on the respective computers and having a client-server relationship to each other. The server can be a cloud server, a server of a distributed system, or a server combined with a blockchain.

[0120] It should be understood that the various forms of flow shown above can be re-ordered, added to, or have steps deleted, using the steps described above. For example, the steps described in the present disclosure can be performed in parallel, in series, or in a different order, as long as the desired results of the technology disclosed in the present disclosure can be achieved, which is not limited herein.

[0121] The specific implementation described above does not constitute a limitation on the protection scope of the present disclosure. Those skilled in the art should understand that various modifications, combinations, sub-combinations, and substitutions can be made according to design requirements and other factors. Any modifications, equivalent replacements, and improvements made within the spirit and principles of the present disclosure shall be included in the protection scope of the present disclosure.

Claims

1. A lyrics labeling method, comprising: determining a target background music and a target style corresponding thereto; determining at least one candidate labeling content according to the target style; based on the target background music and the meaning of a preset lyric, performing a connected reading, a single reading, a stress or a bass split on the preset lyric to obtain at least one split result; generating initial labeling information corresponding to the preset lyric based on the at least one candidate labeling content and the split result; determining target labeling information of the preset lyric in response to a user's adjustment operation on the initial labeling information; generating target lyrics of the target background music based on the preset lyric and the target labeling information corresponding thereto, wherein the target labeling information is mapped to a lyric display of a recording interface through a visual identifier.

2. The method of claim 1, wherein, The determination of the target background music and the target style corresponding thereto comprises at least one of the following: In the case that the target background music corresponds to a plurality of preset styles, the preset style with the highest heat value among the plurality of preset styles is taken as the target style according to the heat values of the plurality of preset styles; In the case that the target background music corresponds to a plurality of preset styles, the preset style with the highest priority among the plurality of preset styles is taken as the target style according to the priorities of the plurality of preset styles.

3. The method of claim 1, wherein, The adjustment operation comprises a modification of a connected reading or a single reading or a stress or a bass annotation.

4. The method of claim 1, wherein, The visual identifier comprises a color, a line interval or a symbol mark.

5. The method of claim 1, wherein, The method further comprises: synchronously mapping the target lyrics and the target labeling information to a song recording interface.

6. A lyrics labeling apparatus, comprising: a first determination unit configured to determine a target background music and a target style corresponding thereto; a first generation unit configured to determine at least one candidate labeling content according to the target style; based on the target background music and the meaning of a preset lyric, performing a connected reading, a single reading, a stress or a bass split on the preset lyric to obtain at least one split result; generating initial labeling information corresponding to the preset lyric based on the at least one candidate labeling content and the split result; a second determination unit configured to determine target labeling information of the preset lyric in response to a user's adjustment operation on the initial labeling information; a second generation unit configured to generate target lyrics of the target background music based on the preset lyric and the target labeling information corresponding thereto, wherein the target labeling information is mapped to a lyric display of a recording interface through a visual identifier.

7. The apparatus of claim 6, wherein, The first determination unit is specifically configured to: In the case that the target background music corresponds to a plurality of preset styles, the preset style with the highest heat value among the plurality of preset styles is taken as the target style according to the heat values of the plurality of preset styles; Or, specifically configured to, in the case that the target background music corresponds to a plurality of preset styles, the preset style with the highest priority among the plurality of preset styles is taken as the target style according to the priorities of the plurality of preset styles.

8. An electronic device, comprising: at least one processor; and a memory connected in communication with the at least one processor; wherein, The memory stores instructions executable by the at least one processor, the instructions being executed by the at least one processor to enable the at least one processor to perform the method of any one of claims 1-5.

9. A non-transitory computer readable storage medium having stored thereon computer instructions, wherein, The computer instructions are for causing the computer to perform the method of any one of claims 1-5.

10. A computer program product comprising a computer program which, when executed by a processor, implements the method of any one of claims 1-5.

Citation Information

Patent Citations

  • Lyric timestamp generation method and storage medium

    CN112786020A