Music generation method and apparatus, device, and storage medium
By generating candidate lyrics in the editing interface and providing music based on user selections, the problem of ordinary users struggling to create high-quality lyrics is solved, thus improving the efficiency and quality of music generation.
Patent Information
- Application Number
- PCT/CN2025/091595
- Authority / Receiving Office
- WO · WO
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2024-06-04
- Filing Date
- 2025-04-27
- Publication Date
- 2025-12-11
AI Technical Summary
Music generation is highly dependent on lyrics, and ordinary users find it difficult to provide high-quality lyric input, which affects the quality of music content generation.
By presenting an editing interface, including lyrics editing controls, candidate lyrics are generated based on contextual information. Users select target lyrics, and music content based on the target lyrics is provided in response to music generation requests.
This improved the efficiency and quality of lyric editing, thereby enhancing the quality of the generated music content.
Smart Images

Figure CN2025091595_11122025_PF_FP_ABST
Abstract
Description
Method, apparatus, device and storage medium for generating music
[0001] The present application claims priority to the Chinese patent application No. 202410718366.6, filed on June 4, 2024, entitled "Method, apparatus, device and storage medium for generating music", the whole content of which is incorporated herein by reference. TECHNICAL FIELD
[0002] Example embodiments of the present disclosure generally relate to the field of Internet, and in particular, to a method, apparatus, device, storage medium and product for generating music. BACKGROUND
[0003] In recent years, with the rapid development of the Internet, generative artificial intelligence technology has been gradually applied to the creation of various types of media content. Compared with media content such as video content and picture content, the generation quality of music content is more dependent on lyrics, which requires people to have strong knowledge of music theory. SUMMARY
[0004] In a first aspect of the present disclosure, a method for generating music is provided. The method comprises: presenting an editing interface, the editing interface comprising a lyrics editing control; in response to a selection of first lyrics content in the lyrics editing control, presenting a set of candidate lyrics content, the set of candidate lyrics content being generated based on context information associated with the lyrics editing control; determining a target lyrics based on a selection of second lyrics content in the set of candidate lyrics content; and in response to receiving a music generation request, providing music content generated based on the target lyrics.
[0005] In a second aspect of the present disclosure, an apparatus for generating music is provided. The apparatus comprises: an interface presentation module configured to present an editing interface, the editing interface comprising a lyrics editing control; a lyrics presentation module configured to, in response to a selection of first lyrics content in the lyrics editing control, present a set of candidate lyrics content, the set of candidate lyrics content being generated based on context information associated with the lyrics editing control; a lyrics determination module configured to determine a target lyrics based on a selection of second lyrics content in the set of candidate lyrics content; and a music providing module configured to, in response to receiving a music generation request, provide music content generated based on the target lyrics.
[0006] In a third aspect of the present disclosure, an electronic device is provided. The device comprises at least one processing unit; and at least one memory coupled to the at least one processing unit and storing instructions for execution by the at least one processing unit. The instructions, when executed by the at least one processing unit, cause the device to perform the method of the first aspect.
[0007] In a fourth aspect of the present disclosure, a computer readable storage medium is provided. The computer readable storage medium has stored thereon a computer program, the computer program being executable by a processor to implement the method of the first aspect.
[0008] In a fifth aspect of the present disclosure, a computer program product is provided. The computer program product comprises a computer program executable by a processing unit, the computer program comprising instructions for performing the method of the first aspect.
[0009] It should be understood that nothing in the Summary is to be construed as a limitation on the scope of the embodiments of the present disclosure. Other features, aspects, and advantages of the present disclosure will become apparent from the following detailed description, figures and claims. BRIEF DESCRIPTION OF DRAWINGS
[0010] The above and other features, aspects and advantages of embodiments of the present disclosure will become more apparent from the following detailed description when taken in conjunction with the accompanying drawings. In the drawings, like reference numerals refer to like elements, in which:
[0011] FIG. 1 shows a schematic diagram of an example environment in which embodiments according to the present disclosure can be implemented;
[0012] FIGS. 2A to 2E show example interfaces according to some embodiments of the present disclosure;
[0013] FIG. 3 shows a flowchart of an example process of generating music according to some embodiments of the present disclosure;
[0014] FIG. 4 shows a schematic structural block diagram of an example apparatus for generating music according to some embodiments of the present disclosure; and
[0015] FIG. 5 shows a block diagram of an electronic device capable of implementing various embodiments of the present disclosure. DETAILED DESCRIPTION
[0016] Embodiments of the present disclosure will be described in more detail with reference to the drawings. While certain embodiments of the present disclosure are shown in the drawings, it is understood that the present disclosure can be embodied in various forms and should not be construed as being limited to the embodiments set forth herein; rather, these embodiments are provided so that the present disclosure will be more thoroughly and completely understood. It should be understood that the drawings and embodiments of the present disclosure are only for illustrative purposes and should not be construed as limiting the scope of protection of the present disclosure.
[0017] It should be noted that the titles of any sections / sub-sections provided herein are not limiting. Various embodiments are described throughout this document and any type of embodiment can be included under any section / sub-section. Furthermore, embodiments described in any section / sub-section can be combined with any other embodiments described in the same section / sub-section and / or different section / sub-section in any manner.
[0018] In the description of embodiments of the disclosure, the term "includes" and its conjugates are open-ended, meaning "including but not limited to". The term "based on" is intended to mean "based, at least in part, on" The term "one embodiment" or "an embodiment" means "at least one embodiment". The term "some embodiments" means "at least some embodiments". Other explicit or implicit definitions can also be included below. The terms "first", "second", etc. can refer to different or the same objects. Other explicit and implicit definitions can also be included below.
[0019] Data of users, acquisition and / or use of data, etc. can be involved in embodiments of the disclosure. These aspects all comply with corresponding laws and regulations and relevant provisions. In embodiments of the disclosure, all data collection, acquisition, processing, processing, forwarding, use, etc. are carried out on the premise that the user is aware of and confirms. Accordingly, when implementing embodiments of the disclosure, the type of data or information that can be involved, the use range, the use scenario, etc. should be notified to the user and the authorization of the user should be obtained according to relevant laws and regulations through appropriate means. The specific notification and / or authorization method can vary according to the actual situation and application scenario, and the scope of the disclosure is not limited in this respect.
[0020] In the specification and embodiments of the present disclosure, if personal information processing is involved, it will be processed on the premise of legality (for example, obtaining the consent of the subject of personal information, or being necessary for the performance of a contract, etc.), and only within the prescribed or agreed range. Users refuse to process personal information other than the necessary information required for basic functions, which will not affect the user's use of basic functions.
[0021] As mentioned above, compared with media content such as video content and picture content, the generation quality of music content has a strong dependence on lyrics, which requires people to have strong music theory knowledge. Due to the professionalism of lyric creation, ordinary users are difficult to provide high-quality lyric input, which greatly affects the quality of the generated music content.
[0022] Embodiments of the present disclosure propose a scheme for generating music. According to the scheme, an editing interface is presented, the editing interface including a lyrics editing control; in response to a selection of first lyrics content in the lyrics editing control, a set of candidate lyrics content is presented, the set of candidate lyrics content being generated based on context information associated with the lyrics editing control; based on a selection of second lyrics content in the set of candidate lyrics content, a target lyrics is determined; and in response to receiving a music generation request, music content generated based on the target lyrics is provided.
[0023] In this way, embodiments of the present disclosure can help users to edit lyrics content based on context information, thereby improving editing efficiency and quality of input lyrics, and further improving quality of generated music content.
[0024] Various example implementations of the scheme are described in further detail below in conjunction with the accompanying drawings.
[0025] Example Environment
[0026] FIG. 1 illustrates a schematic diagram of an example environment 100 in which embodiments of the present disclosure can be implemented. As shown in FIG. 1, the example environment 100 can include an electronic device 110.
[0027] In this example environment 100, the electronic device 110 can run an application 120 that supports interface interaction. The application 120 can be any suitable type of application for interface interaction, examples of which can include, but are not limited to, a music application or other suitable application. A user 140 can interact with the application 120 via the electronic device 110 and / or its attached devices.
[0028] In the environment 100 of FIG. 1, if the application 120 is in an active state, the electronic device 110 can present, through the application 120, an interface 150 for supporting interface interaction.
[0029] In some embodiments, the electronic device 110 communicates with the server 130 to enable provisioning of services of the application 120. The electronic device 110 can be any type of mobile terminal, fixed terminal, or portable terminal including a mobile handset, a tablet computer, a laptop computer, a notebook computer, a netbook computer, a tablet computer, a media computer, a multimedia tablet, a palmtop computer, a portable gaming terminal, a VR / AR device, a Personal Communication System (PCS) terminal, a personal navigation device, a Personal Digital Assistant (PDA), an audio / video player, a digital camera / camcorder, a positioning device, a television receiver, a radio broadcast receiver, an electronic book device, a game device, or any combination thereof, including an accessory or peripheral device for the foregoing, or any combination thereof. In some embodiments, the electronic device 110 can also support any type of interface to a user (such as “wearable” circuitry, etc.).
[0030] The server 130 can be a standalone physical server, a server cluster or distributed system composed of multiple physical servers, or a cloud server providing cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communications, middleware services, domain name services, security services, content distribution networks, and basic cloud computing services such as big data and artificial intelligence platforms. The server 130 may, for example, include a computing system / server, such as a mainframe, an edge computing node, a computing device in a cloud environment, and the like. The server 130 can provide background services for the application 120 in the electronic device 110 that supports a virtual scene.
[0031] A communication connection can be established between the server 130 and the electronic device 110. The communication connection can be established by wired or wireless means. The communication connection can include, but is not limited to, a Bluetooth connection, a mobile network connection, a Universal Serial Bus (USB) connection, a Wireless Fidelity (WiFi) connection, and the like, and embodiments of the present disclosure are not limited in this regard. In embodiments of the present disclosure, the server 130 and the electronic device 110 can implement signaling interaction through the communication connection therebetween.
[0032] It should be understood that the structure and function of the various elements in the environment 100 are described for illustrative purposes only, and do not imply any limitation on the scope of the present disclosure.
[0033] Some example embodiments of the present disclosure will be described below with continued reference to the accompanying drawings.
[0034] Example process
[0035] FIGS. 2A-2E illustrate example interfaces 200A-200E, in accordance with some embodiments of the present disclosure. The interfaces 200A-200E can be provided by, for example, the electronic device 110 illustrated in FIG. 1.
[0036] FIG. 2A illustrates an editing interface 200A, in accordance with some embodiments of the present disclosure. As shown in FIG. 2A, the editing interface 200A can include one or more controls for editing lyrics creation parameters.
[0037] In particular, the editing interface 200A can include a lyrics editing control 202. The lyrics editing control 202 can support a user to directly input lyrics content, or can also generate lyrics content by inputting prompt items.
[0038] Taking FIG. 2A as an example, the electronic device 110 can acquire prompt items 204, e.g., text prompt items, via the lyrics editing control 202. In some embodiments, the prompt items 204 can be text prompt items corresponding to preset templates. For example, upon receiving a preset request from a user, the electronic device 110 can display text prompt items corresponding to preset templates in the lyrics editing control 202.
[0039] In some embodiments, the electronic device 110 can also support a user to quickly edit one or more contents in the prompt items 204. In particular, as shown in FIG. 2A, the electronic device 110 can provide, for example, editing entrances corresponding to the contents “history” and “ancient style”. Upon receiving a selection of an editing entrance for a content 206, e.g., “history”, the electronic device 110 can associatively provide a set of candidate contents 208, e.g., “affection” and “friendship”. Further, the electronic device 110 can replace the content 206 in the prompt items 204 based on, for example, a user’s selection of the candidate contents 208.
[0040] Based on such a manner, embodiments of the present disclosure can provide users with templated prompt items, thereby improving the quality of prompt items input by users and improving editing efficiency.
[0041] In some embodiments, a user can also edit the prompt items 204 by, for example, inserting, modifying, deleting, etc. In yet some embodiments, the electronic device 110 can also acquire text prompt items directly input by a user in the lyrics editing control 202.
[0042] Furthermore, upon receiving a selection for button 210, electronic device 110 may, for example, trigger the generation of lyrics corresponding to prompt 204. As an example, electronic device 110 may, for instance, provide prompt 204 to a language model to obtain the corresponding lyrics. It should be understood that such a language model may be implemented based on appropriate machine learning model techniques, and this disclosure is not intended to limit it.
[0043] Furthermore, as shown in Figure 2B, the electronic device 110 may display lyrics 220 in the lyrics editing control 202, for example. As an example, the lyrics 220 may be generated based on the prompt 204 mentioned above. As another example, the lyrics 220 may also be lyrics directly entered by the user using the lyrics editing control 202.
[0044] As another example, the electronic device 110 can also receive preset operations from the user on the completed musical content (also known as the target musical content) and obtain the user's secondary creation request. Accordingly, the electronic device 110 can present the editing interface 200B as shown in FIG2B based on the preset operations. In this case, the lyrics 220 displayed by the lyrics editing control 202 can be, for example, the lyrics of the selected target musical content, to support the user to further edit the lyrics.
[0045] In the secondary creation scenario based on the target music content, one or more of the attribute editing controls 212, 214, and 216 in the editing interface 200B can display the attributes corresponding to the target music content by default, such as style, timbre, and name.
[0046] Referring again to Figure 2B, the electronic device 110 can receive the user's selection of the first lyric content 222 in the lyrics 220 (e.g., "the obsession in the eyes"), and can accordingly present a set of candidate lyric content 224.
[0047] In some embodiments, the set of candidate lyrics 224 may be generated based on context information associated with the lyrics editing control 202. In some embodiments, such context information may include, for example, existing lyrics in the lyrics editing control 202, so that the provided candidate lyrics 224 can be adapted to existing lyrics.
[0048] In some embodiments, such context information may include at least one lyric parameter. In some embodiments, such lyric parameters may include rhyme parameters, theme parameters, etc. In some embodiments, such lyric parameters may be automatically determined based on existing lyric content. Alternatively, the electronic device 110 may also provide controls for editing lyric parameters in the lyric editing control 202, such as controls 228 and 230.
[0049] In this way, the set of candidate lyric content 224 provided by the electronic device 110 can match the lyric parameters associated with the lyric editing control 202, e.g., with matching rhymes and matching themes, etc.
[0050] In some embodiments, the electronic device 110 may, for example, receive a user selection of a second lyric content (e.g., “In the eyes of the missing”) of the set of candidate lyric content 224, and may, with the second lyric content, replace the first lyric content 222 in the lyrics 220, thereby completing the quick editing of the lyrics.
[0051] In this way, embodiments of the present disclosure can support users to efficiently optimize the content of the lyrics, thereby improving the efficiency of the editing of the lyrics. Further, the provided candidate lyric content can match the contextual information, which enables users to simply create high-quality lyrics.
[0052] In some embodiments, the lyric editing control 202 may, further, receive a user selection of the control 228 or the control 230, and may, accordingly, update the lyrics 220. For example, upon a user changing a new rhyme via the control 228, the electronic device 110 may, provide updated lyrics matching the new rhyme. Or, upon a user specifying a new theme via the control 230, the electronic device 110 may, provide updated lyrics matching the new theme.
[0053] In some embodiments, the lyric editing control 202 may, further, include a continuation control 226. Further, upon receiving a selection of the continuation control 226, the electronic device 110 may, in the lyric editing control 202, provide additional lyric content created based on the existing lyrics 220.
[0054] In some embodiments, such additional lyric content may, for example, be generated based on the contextual information associated with the lyric editing control. Such contextual information may, for example, be the same as or different from the contextual information used to generate the candidate lyric content 224.
[0055] For example, a user may, through the lyric editing control 202, input a few lines of lyrics, and the electronic device 110 may, with a language model, automatically continue the lyrics based on the existing lyric content. To ensure the matching of the lyrics, the continued lyric content may, for example, match the rhymes and / or the theme of the existing lyrics.
[0056] In some embodiments, as shown in FIG. 2C, the electronic device 110 can further provide a segmentation control 232 in the lyrics editing control 202. Upon receiving a trigger for the segmentation control 232, the electronic device 110 can present a set of indication elements associated with the lyrics 220 in the lyrics editing control accordingly based on the received segmentation request, e.g., indication element 234-1 and indication element 234-2 (individually or collectively referred to as indication elements 234).
[0057] In some embodiments, such indication elements 234 can be used to indicate the structure type of the corresponding lyrics segment in the lyrics 220, e.g., a verse part, a chorus part, etc. In some embodiments, such indication elements 234 can represent the structure information of the lyrics 220. As an example, such structure information can be provided for the generation of the corresponding music content.
[0058] In some embodiments, the user can edit the indication elements 234. For example, the user can insert a new indication element in a specific position of the lyrics to indicate the structure type of the corresponding lyrics segment. Alternatively, the user can also delete one or more of the existing indication elements 234. Alternatively, the user can also modify the type of the indication elements 234 to modify the corresponding lyrics segment from a first structure type to a second structure type, for example.
[0059] In some embodiments, as shown in FIG. 2D, the electronic device 110 can further support the user to edit the indication elements freely. Specifically, upon receiving a preset identifier 236 (e.g., “ / ”) input by the user in the lyrics editing control 202, the electronic device 110 can accordingly provide a set of candidate indication elements 238.
[0060] Further, the electronic device 110 can receive a selection of a target indication element from the set of candidate indication elements 238 by the user, and can insert the target indication element into a target position corresponding to the preset identifier 236 to indicate the structure type of the corresponding lyrics segment.
[0061] In this way, the embodiments of the present disclosure can edit the structure information of the lyrics more effectively, thereby further improving the quality of the generated music content.
[0062] Further, as shown in FIGS. 2B-2D, the electronic device 110 can obtain the final lyrics in the lyrics editing control 202 and the parameter information input in the other attribute editing controls 212-216, and can trigger the generation process of the corresponding music content based on the selection of the input control 218 by the user.
[0063] As an example, the electronic device 110 can provide such creative parameters to the music generation model to generate corresponding music content. Such creative parameters can include, but are not limited to, the lyrics, the musical style, the tone color, the song name, the structural information of the lyrics, etc. mentioned above. In addition, the music generation model can be implemented by using appropriate machine learning models available or available in the future, and the present disclosure is not intended to be limited thereto.
[0064] Exemplarily, after the music content is generated, the electronic device 110 can present the interface 200E as shown in FIG. 2E accordingly. As shown, the interface 200E can include a first area 240 for displaying creative parameters, a second area 250 for displaying a creative history, and a third area 260 for displaying description information of the currently played music content.
[0065] As an example, the electronic device 110 can present a content item 255 corresponding to the generated music content in the second area 250, and can trigger the display of the description information (e.g., cover, title, lyrics, etc.) of the music content in the third area 260 based on the selection of the content item 255. In addition, the electronic device 110 can also trigger the playback of the music content accordingly, and can provide a playback control 270 to control the playback process of the music content.
[0066] In this way, the embodiments of the present disclosure can improve the efficiency of music creation and improve the quality of the generated music content.
[0067] Example process
[0068] FIG. 3 shows a flowchart of an example process 300 of generating music, according to some embodiments of the present disclosure. The process 300 can be implemented at the electronic device 110. The process 300 is described below with reference to FIG. 1.
[0069] As shown in FIG. 3, at block 310, the electronic device 110 presents an editing interface, the editing interface including a lyrics editing control.
[0070] At block 320, the electronic device 110 presents a set of candidate lyrics content in response to a selection of first lyrics content in the lyrics editing control, the set of candidate lyrics content being generated based on context information associated with the lyrics editing control.
[0071] At block 330, the electronic device 110 determines a target lyrics based on a selection of second lyrics content in the set of candidate lyrics content.
[0072] At block 340, the electronic device 110 provides music content generated based on the target lyrics in response to receiving a music generation request.
[0073] In some embodiments, the context information comprises at least one lyric parameter, the at least one lyric parameter comprising a foot parameter and / or a theme parameter; and / or existing lyric content in the lyric editing control.
[0074] In some embodiments, the lyric editing control comprises a first configuration control for configuring the foot parameter; and / or a second configuration control for configuring the theme parameter.
[0075] In some embodiments, the method further comprises: in response to obtaining an updated foot parameter via the first configuration control, providing updated lyric content corresponding to the updated foot parameter in the lyric editing control; or in response to obtaining an updated theme parameter via the second configuration control, providing updated lyric content corresponding to the updated theme parameter in the lyric editing control.
[0076] In some embodiments, the context information is first context information, the method further comprising: in response to the received continuation request, presenting additional lyric content in the lyric editing control, the additional lyric content being generated based on second context information associated with the lyric editing control.
[0077] In some embodiments, the method further comprises: in response to the received segmentation request, presenting a set of indication elements associated with the existing lyric content in the lyric editing control, the set of indication elements being used to represent a structure type of a corresponding lyric segment.
[0078] In some embodiments, the music content is further generated based on structure information of the target lyric, the structure information representing a structure type of a set of lyric segments of the target lyric.
[0079] In some embodiments, the method further comprises: modifying the structure information of the existing lyric content based on an editing operation on the set of indication elements.
[0080] In some embodiments, the editing operation comprises at least one of: modifying a type of an existing indication element to modify a corresponding lyric segment from a first structure type to a second structure type; adding a new indication element to indicate a structure type of a corresponding lyric segment; and deleting at least one indication element.
[0081] In some embodiments, adding the new indication element comprises: in response to inputting a preset identifier in the lyric editing control, presenting a set of candidate indication elements; and based on a selection of a target indication element from the set of candidate indication elements, inserting the target indication element at a target position in the lyric editing control.
[0082] In some embodiments, the first lyric content is a portion of a reference lyric, the method further comprising: presenting a text prompt in the lyric editing control; and based on a received lyric generation request, providing the reference lyric generated based on the text prompt in the lyric editing control.
[0083] In some embodiments, presenting the text prompt in the lyrics editing control includes: presenting the text prompt corresponding to the preset template in the lyrics editing control; and based on selection of target content in the text prompt, presenting a set of candidate contents for replacing the target part.
[0084] In some embodiments, presenting the editing interface includes: based on the preset operation on the target music content, presenting the editing interface; and presenting lyrics of the target music content in a lyrics editing control of the editing interface.
[0085] In some embodiments, the editing interface further includes a property editing control for configuring at least one property of the music content to be generated.
[0086] Example apparatus and device
[0087] Embodiments of the present disclosure also provide a corresponding apparatus for implementing the above method or process. FIG. 4 shows a schematic structural block diagram of an example apparatus 400 for generating music, according to certain embodiments of the present disclosure. The apparatus 400 can be implemented as or included in the electronic device 110. Various modules / components in the apparatus 400 can be implemented by hardware, software, firmware, or any combination thereof.
[0088] As shown in FIG. 4, the apparatus 400 includes an interface presenting module 410 configured to present an editing interface, the editing interface including a lyrics editing control; a lyrics presenting module 420 configured to, in response to selection of first lyrics content in the lyrics editing control, present a set of candidate lyrics content, the set of candidate lyrics content being generated based on context information associated with the lyrics editing control; a lyrics determining module 430 configured to, based on selection of second lyrics content in the set of candidate lyrics content, determine target lyrics; and a music providing module 440 configured to, in response to receiving a music generation request, provide music content generated based on the target lyrics.
[0089] In some embodiments, the context information includes: at least one lyrics parameter, the at least one lyrics parameter including a rhyme parameter and / or a theme parameter; and / or existing lyrics content in the lyrics editing control.
[0090] In some embodiments, the lyrics editing control includes: a first configuration control for configuring the rhyme parameter; and / or a second configuration control for configuring the theme parameter.
[0091] In some embodiments, the music providing module 440 is further configured to: in response to obtaining the updated rhyme parameter via the first configuration control, provide updated lyric content corresponding to the updated rhyme parameter in the lyric editing control; or in response to obtaining the updated theme parameter via the second configuration control, provide updated lyric content corresponding to the updated theme parameter in the lyric editing control.
[0092] In some embodiments, the context information is first context information, and the method further includes: in response to the received continuation request, presenting additional lyric content in the lyric editing control, the additional lyric content being generated based on second context information associated with the lyric editing control.
[0093] In some embodiments, the interface presenting module 410 is further configured to: in response to the received segmentation request, present a set of indication elements associated with the existing lyric content in the lyric editing control, the set of indication elements being used to represent a structure type of a corresponding lyric segment.
[0094] In some embodiments, the music content is further generated based on structure information of the target lyric, the structure information representing a structure type of a set of lyric segments of the target lyric.
[0095] In some embodiments, the lyric determining module 430 is further configured to: based on an editing operation on the set of indication elements, modify the structure information of the existing lyric content.
[0096] In some embodiments, the editing operation includes at least one of: modifying a type of an existing indication element to modify a corresponding lyric segment from a first structure type to a second structure type; adding a new indication element to indicate a structure type of a corresponding lyric segment; and deleting at least one indication element.
[0097] In some embodiments, adding the new indication element includes: in response to inputting a preset identifier in the lyric editing control, presenting a set of candidate indication elements; and based on a selection of a target indication element in the set of candidate indication elements, inserting the target indication element at a target position in the lyric editing control.
[0098] In some embodiments, the first lyric content is a portion of a reference lyric, and the method further includes: presenting a text prompt in the lyric editing control; and based on a received lyric generation request, providing the reference lyric generated based on the text prompt in the lyric editing control.
[0099] In some embodiments, presenting the text prompt in the lyric editing control includes: presenting a text prompt corresponding to a preset template in the lyric editing control; and based on a selection of target content in the text prompt, presenting a set of candidate content for replacing the target portion.
[0100] In some embodiments, presenting the editing interface includes: presenting the editing interface based on the preset operation on the target music content; and presenting lyrics of the target music content in a lyrics editing control of the editing interface.
[0101] In some embodiments, the editing interface further includes an attribute editing control for configuring at least one attribute of the music content to be generated.
[0102] FIG. 5 illustrates a block diagram of an electronic device 500 in which one or more embodiments of the disclosure can be implemented. It should be understood that the electronic device 500 illustrated in FIG. 5 is merely exemplary and should not be construed as limiting the functionality and scope of the embodiments described herein. The electronic device 500 illustrated in FIG. 5 can be used to implement the electronic device 110 of FIG. 1.
[0103] As shown in FIG. 5, the electronic device 500 is in the form of a general electronic device. The components of the electronic device 500 can include, but are not limited to, one or more processors or processing units 510, a memory 520, a storage device 530, one or more communication units 540, one or more input devices 550, and one or more output devices 560. The processing unit 510 can be a real or virtual processor and is capable of performing various processing according to programs stored in the memory 520. In a multi-processor system, multiple processing units perform computer-executable instructions in parallel to improve the parallel processing capability of the electronic device 500.
[0104] The electronic device 500 typically includes a number of computer storage media. Such media can be any available media that is accessible by the electronic device 500 and includes both volatile and non-volatile media, removable and non-removable media. The memory 520 can be a volatile memory (e.g., registers, cache, random access memory (RAM)), a non-volatile memory (e.g., read-only memory (ROM), electrically erasable programmable read-only memory (EEPROM), flash memory), or some combination thereof. The storage device 530 can be a removable or non-removable media and can include machine-readable media such as a flash drive, a magnetic disk drive, or any other media that can be used to store information and / or data and that can be accessed by the electronic device 500.
[0105] The electronic device 500 can further include additional detachable / non-detachable, volatile / non-volatile storage media. Although not shown in FIG. 5, a disk drive for reading from or writing to a detachable, non-volatile magnetic disk (e.g., a "floppy disk"), and an optical disk drive for reading from or writing to a detachable, non-volatile optical disk (e.g., a CD-ROM) can be provided. In these cases, each drive can be connected to the bus (not shown) by one or more data media interfaces. The memory 520 can include a computer program product 525 having one or more program modules configured to carry out the various methods or acts of the various embodiments of the present disclosure.
[0106] The communication unit 540 enables communication with other electronic devices through communication media. Additionally, the functionality of the components of the electronic device 500 can be implemented in a single computing cluster or a plurality of computer machines capable of communicating with one another through a communication connection. As such, the electronic device 500 can operate in a networked environment using logical connections to one or more other servers, network personal computers (PCs), or another network nodes in the networking environment.
[0107] The input device 550 can be one or more input devices, such as a mouse, a keyboard, a trackball, etc. The output device 560 can be one or more output devices, such as a display, a speaker, a printer, etc. The electronic device 500 can also communicate with one or more external devices (not shown) such as a storage device, a display device, etc., one or more devices that enable a user to interact with the electronic device 500, or any devices (e.g., a network card, a modem, etc.) that enable the electronic device 500 to communicate with one or more other electronic devices, through the communication unit 540, as needed. Such communication can be carried out via an input / output (I / O) interface (not shown).
[0108] According to an example implementation of the present disclosure, a computer readable storage medium having computer executable instructions stored thereon is provided, where the computer executable instructions are executed by a processor to implement the method described above. According to an example implementation of the present disclosure, a computer program product is also provided, which is tangibly stored on a non-transitory computer readable medium and includes computer executable instructions, where the computer executable instructions are executed by a processor to implement the method described above.
[0109] Various aspects of the disclosure are now described with reference to the drawings. In general, the drawings described below are diagrammatic and schematic representations of actual or conceptual structures and processes, and are not limiting of the scope of the present disclosure. In the drawings, the same reference numerals are used to represent similar or like items. The embodiments of the present disclosure will be described with reference to the drawings, wherein like designations denote like elements, and wherein:
[0110] The computer readable program instructions can also be loaded onto a computer, other programmable data processing apparatus, or other device to cause a series of operational steps to be performed on the computer, other programmable apparatus or other device to produce a computer implemented process such that the instructions which execute on the computer or other programmable apparatus provide processes for implementing the functions / acts specified in the flowchart and / or block diagram block or blocks.
[0111] The computer readable program instructions can also be loaded onto a computer, other programmable data processing apparatus, or other device to cause a series of operational steps to be performed on the computer, other programmable apparatus or other device to produce a computer implemented process such that the instructions which execute on the computer or other programmable apparatus provide processes for implementing the functions / acts specified in the flowchart and / or block diagram block or blocks.
[0112] The computer readable program instructions can also be loaded onto a computer, other programmable data processing apparatus, or other device to cause a series of operational steps to be performed on the computer, other programmable apparatus or other device to produce a computer implemented process such that the instructions which execute on the computer or other programmable apparatus provide processes for implementing the functions / acts specified in the flowchart and / or block diagram block or blocks.
[0113] The implementations of the disclosure have been described above with the intent to be illustrative rather than limiting. Although being shown in only a few of the various implementations, the principle of each implementation can be extended to any other implementation. Some of the present implementations have also been described with the intent to be illustrative rather than restrictive. Many modifications and variations of the described implementations are possible in light of the above teachings. It is therefore contemplated that the application can encompass modifications and variations provided they come within the scope of the following claims. It is also contemplated that the implementing specific electric circuitry such as, for example, application specific integrated circuits (ASICs) can implement one or more of the described implementations. It is therefore accurate that the claims are framed to encompass both the specific examples disclosed and also the generic implementations so disclosed.
Claims
1.A method of generating music, comprising: presenting an editing interface, the editing interface comprising a lyrics editing control; in response to a selection of first lyrics content in the lyrics editing control, presenting a set of candidate lyrics content, the set of candidate lyrics content being generated based on contextual information associated with the lyrics editing control; based on a selection of second lyrics content in the set of candidate lyrics content, determining a target lyrics; and in response to receiving a music generation request, providing music content generated based on the target lyrics. 2.The method of claim 1, wherein the contextual information comprises: at least one lyrics parameter, the at least one lyrics parameter comprising a rhyme parameter and / or a theme parameter; and / or existing lyrics content in the lyrics editing control. 3.The method of claim 2, wherein the lyrics editing control comprises: a first configuration control for configuring the rhyme parameter; and / or a second configuration control for configuring the theme parameter. 4.The method of claim 3, further comprising: in response to obtaining an updated rhyme parameter via the first configuration control, providing updated lyrics content corresponding to the updated rhyme parameter in the lyrics editing control; or in response to obtaining an updated theme parameter via the second configuration control, providing updated lyrics content corresponding to the updated theme parameter in the lyrics editing control. 5.The method of any one of claims 1-4, wherein the contextual information is first contextual information, the method further comprising: in response to receiving a continuation request, presenting additional lyrics content in the lyrics editing control, the additional lyrics content being generated based on second contextual information associated with the lyrics editing control. 6.The method of any one of claims 1-4, further comprising: in response to receiving a segmentation request, presenting a set of indication elements associated with existing lyrics content in the lyrics editing control, the set of indication elements being used to represent a structure type of a corresponding lyrics segment. 7.The method of claim 6, wherein the music content is further generated based on structure information of the target lyrics, the structure information representing the structure type of a set of lyrics segments of the target lyrics. 8.The method of claim 6, further comprising: based on an editing operation on the set of indication elements, modifying structure information of the existing lyrics content. 9.The method of claim 8, wherein the editing operation comprises at least one of: modifying a type of an existing indication element to modify a corresponding lyrics segment from a first structure type to a second structure type; adding a new indication element to indicate a structure type of a corresponding lyrics segment; and deleting at least one indication element. 10.The method of claim 9, wherein adding a new indication element comprises: in response to inputting a preset identifier in the lyrics editing control, presenting a set of candidate indication elements; and based on a selection of a target indication element in the set of candidate indication elements, inserting the target indication element at a target position in the lyrics editing control. 11.The method of any of claims 1-4 and 7-10, wherein the first lyric content is a part of a reference lyric, the method further comprising: presenting a text prompt in the lyric editing control; and providing, in the lyric editing control, the reference lyric generated based on the text prompt based on a received lyric generation request. 12.The method of claim 11, wherein presenting a text prompt in the lyric editing control comprises: presenting the text prompt corresponding to a preset template in the lyric editing control; and presenting a set of candidate contents for replacing a target part based on a selection of a target content in the text prompt. 13.The method of any of claims 1-4, 7-10, and 12, wherein presenting an editing interface comprises: presenting the editing interface based on a preset operation on a target music content; and presenting a lyric of the target music content in the lyric editing control of the editing interface. 14.The method of any of claims 1-4, 7-10, and 12, wherein the editing interface further comprises an attribute editing control for configuring at least one attribute of the music content to be generated. 15.An apparatus for generating music, comprising: an interface presenting module configured to present an editing interface, the editing interface comprising a lyric editing control; a lyric presenting module configured to present a set of candidate lyric contents in response to a selection of a first lyric content in the lyric editing control, the set of candidate lyric contents being generated based on context information associated with the lyric editing control; a lyric determining module configured to determine a target lyric based on a selection of a second lyric content in the set of candidate lyric contents; and a music providing module configured to provide a music content generated based on the target lyric in response to a received music generation request. 16.An electronic device, comprising: at least one processing unit; and at least one memory coupled to the at least one processing unit and storing instructions for execution by the at least one processing unit, the instructions when executed by the at least one processing unit cause the electronic device to perform the method according to any of claims 1-14. 17.A computer-readable storage medium having stored thereon a computer program, the computer program being executable by a processor to implement the method according to any of claims 1-14. 18.A computer program product tangibly stored in a computer storage medium and comprising computer-executable instructions that, when executed by a device, cause the device to perform the method according to any of claims 1-14.
Citation Information
Patent Citations
Lyric generation method and device, electronic equipment and computer readable storage medium
CN112632906A
Music generation method, device and system and storage medium
CN117012170A
Music generation method and device, equipment and storage medium
CN118486282A
Method and system for interactive song generation
US20210312897A1
Methods and systems for interactive lyric generation
US20210335334A1