Method, apparatus, electronic device, and storage medium for generating works

By performing a process-based method of selecting audio, editing text and recording audio on different pages, the music creation process is simplified, the work generation efficiency is improved, and the creation threshold is lowered.

CN113963674BActive Publication Date: 2025-07-04BEIJING BAIDU NETCOM SCI & TECH CO LTD
View PDF 1 Cites 0 Cited by

Patent Information

Application Number
CN202111161807.X
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2021-09-30
Publication Date
2025-07-04
Estimated Expiration
2041-09-30

AI Technical Summary

Technical Problem

The existing music creation software is complex in operation, resulting in low efficiency in production of works.

Method used

Provides a work generation method, which simplifies the creative process and realizes process through performing different types of creative links on different pages, including selecting audio, editing text and recording audio.

Benefits of technology

It improves the efficiency of music creation, enables music lovers to generate works more conveniently, and lowers the threshold for creation.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN113963674B_ABST
    Figure CN113963674B_ABST
Patent Text Reader

Abstract

The present disclosure provides a method, apparatus, electronic device, and storage medium for generating a work, relating to the technical field of audio processing, to at least solve the technical problem in the related art that the work generation efficiency is relatively low due to the complex creation process. The specific implementation solution is as follows: in response to an operation instruction for a first control on a first page, obtain a first audio and jump to a second page, where the first page includes at least one candidate audio index, and the second page is used for editing text; obtain a first text on the second page, and in response to an operation instruction for a second control on the second page, jump to a third page, where the third page is used to provide an audio recording function; obtain a second audio on the third page, and in response to an operation instruction for a third control on the third page, synthesize the first audio, the first text, and the second audio to generate a first work.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present disclosure relates to the field of audio processing, and particularly to a method, apparatus, electronic device, and storage medium for generating a work. Background Art

[0002] Currently, music creation has a certain threshold. If music lovers want to have a good work, they not only need to have good singing skills, but also need good equipment and the ability to proficiently use music production software to produce music. However, the operation process of current music production software is complex, and music production software usually needs to be used in cooperation with professional equipment. Summary of the Invention

[0003] The present disclosure provides a method, apparatus, electronic device, and storage medium for generating a work, so as to at least solve the technical problem in the related art that the work creation efficiency is too low due to the complex creation process.

[0004] According to one aspect of the present disclosure, there is provided a method for generating a work, including: in response to an operation instruction for a first control on a first page, obtaining a first audio and jumping to a second page, where the first page includes at least one candidate audio index, and the second page is used for editing text; obtaining a first text on the second page, and in response to an operation instruction for a second control on the second page, jumping to a third page, where the third page is used for providing a function of recording an audio; obtaining a second audio on the third page, and in response to an operation instruction for a third control on the third page, synthesizing the first audio, the first text, and the second audio to generate a first work.

[0005] According to another aspect of the present disclosure, there is further provided a generation module for a work, including: an obtaining module, configured to obtain a first audio in response to an operation instruction for a first control on a first page and jump to a second page, where the first page includes at least one candidate audio index, and the second page is used for editing text; a jumping module, configured to obtain a first text on the second page and jump to a third page in response to an operation instruction for a second control on the second page, where the third page is used for providing a function of recording an audio; a synthesizing module, configured to obtain a second audio on the third page and synthesize the first audio, the first text, and the second audio to generate a first work in response to an operation instruction for a third control on the third page.

[0006] According to another aspect of the present disclosure, there is provided an electronic device, including: at least one processor; and a memory communicatively connected to the at least one processor; where the memory stores instructions executable by the at least one processor, and the instructions are executed by the at least one processor to enable the at least one processor to execute the method for generating a work proposed by the present disclosure.

[0007] In accordance with another aspect of the present disclosure, there is provided a non-transitory computer-readable storage medium storing computer instructions, wherein the computer instructions are for causing a computer to execute the method for generating a work proposed by the present disclosure.

[0008] In accordance with another aspect of the present disclosure, there is provided a computer program product including a computer program, and the computer program, when executed by a processor, performs the method for generating a work proposed by the present disclosure.

[0009] In the present disclosure, in response to an operation instruction for a first control in a first page, a first audio can be obtained and a jump can be made to a second page, where the first page includes at least one candidate audio index, the second page is for editing text, a first text is obtained in the second page, in response to an operation instruction for a second control in the second page, a jump is made to a third page, where the third page is for providing a function of recording audio, a second audio is obtained in the third page, and in response to an operation instruction for a third control in the third page, the first audio, the first text, and the second audio are synthesized to generate a first work, achieving the purpose of creating a work. By performing different types of creation steps on different pages, the steps of the creation process can be simplified, the process of the creation process can be made flow-based, thereby improving the creation efficiency, and further solving the technical problem in the related art that the creation efficiency of works is too low due to the complex creation process.

[0010] It should be understood that the content described in this part is not intended to identify the key or important features of the embodiments of the present disclosure, nor is it used to limit the scope of the present disclosure. Other features of the present disclosure will become easily understandable through the following description. BRIEF DESCRIPTION OF THE DRAWINGS

[0011] The drawings are used to better understand the solution and do not constitute a limitation to the present disclosure. Among them:

[0012] Figure 1 is a hardware structure block diagram of a computer terminal (or mobile device) for implementing the method for generating a work according to an embodiment of the present disclosure;

[0013] Figure 2 is a flowchart of a method for generating a work according to an embodiment of the present disclosure;

[0014] Figure 3 is a flowchart of another method for generating a work according to an embodiment of the present disclosure;

[0015] Figure 4 is a schematic diagram of a device for generating a work according to an embodiment of the present disclosure. DETAILED DESCRIPTION OF THE EMBODIMENTS

[0016] The exemplary embodiments of the present disclosure will be described below with reference to the accompanying drawings. Various details of the embodiments of the present disclosure are included to facilitate understanding, and they should be considered merely exemplary. Therefore, those of ordinary skill in the art should recognize that various changes and modifications can be made to the embodiments described herein without departing from the scope and spirit of the present disclosure. Similarly, descriptions of well-known functions and structures are omitted in the following description for clarity and conciseness.

[0017] It should be noted that the terms "first", "second", etc. in the specification and claims of the present disclosure and the above-mentioned drawings are used to distinguish similar objects and do not necessarily describe a specific order or sequence. It should be understood that such data can be interchanged under appropriate circumstances so that the embodiments of the present disclosure described herein can be implemented in an order different from those illustrated or described herein. In addition, the terms "comprising" and "having" and any variations thereof are intended to cover non-exclusive inclusion. For example, a process, method, system, product, or device that includes a series of steps or units does not necessarily limit to those steps or units clearly listed, but may include other steps or units not clearly listed or inherent to these processes, methods, products, or devices.

[0018] According to an embodiment of the present disclosure, a method for generating a work is provided. It should be noted that the steps shown in the flowchart of the accompanying drawings can be executed in a computer system such as a set of computer-executable instructions. And although the logical order is shown in the flowchart, in some cases, the steps shown or described can be executed in a different order from that herein.

[0019] The method embodiments provided by the embodiments of the present disclosure can be executed in a mobile terminal, a computer terminal, or a similar electronic device. The electronic device is intended to represent various forms of digital computers, such as a laptop computer, a desktop computer, a workbench, a personal digital assistant, a server, a blade server, a mainframe computer, and other suitable computers. The electronic device can also represent various forms of mobile devices, such as a personal digital processor, a cellular phone, a smart phone, a wearable device, and other similar computing devices. The components shown herein, their connections and relationships, and their functions are merely examples and are not intended to limit the implementation of the present disclosure described and / or claimed herein. Figure 1 A hardware structure block diagram of a computer terminal (or mobile device) for implementing the method for generating a work is shown.

[0020] As Figure 1As shown, the computer terminal 100 includes a computing unit 101, which can perform various appropriate actions and processes according to a computer program stored in a read-only memory (ROM) 102 or a computer program loaded from a storage unit 108 into a random access memory (RAM) 103. In the RAM 103, various programs and data required for the operation of the computer terminal 100 can also be stored. The computing unit 101, the ROM 102, and the RAM 103 are connected to each other via a bus 104. An input / output (I / O) interface 105 is also connected to the bus 104.

[0021] Multiple components in the computer terminal 100 are connected to the I / O interface 105, including: an input unit 106, such as a keyboard, a mouse, etc.; an output unit 107, such as various types of displays, speakers, etc.; a storage unit 108, such as a magnetic disk, an optical disc, etc.; and a communication unit 109, such as a network card, a modem, a wireless communication transceiver, etc. The communication unit 109 allows the computer terminal 100 to exchange information / data with other devices via a computer network such as the Internet and / or various telecommunication networks.

[0022] The computing unit 101 can be various general-purpose and / or special-purpose processing components with processing and computing capabilities. Some examples of the computing unit 101 include but are not limited to a central processing unit (CPU), a graphics processing unit (GPU), various dedicated artificial intelligence (AI) computing chips, various computing units running machine learning model algorithms, a digital signal processor (DSP), and any appropriate processor, controller, microcontroller, etc. The computing unit 101 executes the method for generating the works described herein. For example, in some embodiments, the method for generating the works can be implemented as a computer software program, which is tangibly contained in a machine-readable medium, such as the storage unit 108. In some embodiments, part or all of the computer program can be loaded and / or installed onto the computer terminal 100 via the ROM 102 and / or the communication unit 109. When the computer program is loaded into the RAM 103 and executed by the computing unit 101, one or more steps of the method for generating the works described herein can be executed. Alternatively, in other embodiments, the computing unit 101 can be configured to execute the method for generating the works in any other appropriate manner (e.g., by means of firmware).

[0023] The various embodiments of the systems and techniques described herein can be implemented in digital electronic circuitry, integrated circuit systems, field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), systems on a chip (SOCs), complex programmable logic devices (CPLDs), computer hardware, firmware, software, and / or combinations thereof. These various embodiments can include: being implemented in one or more computer programs that can be executed and / or interpreted on a programmable system including at least one programmable processor, which can be a special or general programmable processor that can receive data and instructions from a storage system, at least one input device, and at least one output device, and transmit the data and instructions to the storage system, the at least one input device, and the at least one output device.

[0024] It should be noted here that in some alternative embodiments, the above Figure 1 The electronic device shown may include hardware elements (including circuits), software elements (including computer code stored on a computer-readable medium), or a combination of both hardware elements and software elements. It should be pointed out that Figure 1 is only an example of a specific specific instance and is intended to illustrate the types of components that may exist in the above electronic device.

[0025] Under the above operating environment, the present disclosure provides a method for generating a work as Figure 2 shown, and this method can be executed by a Figure 1 computer terminal or a similar electronic device shown. Figure 2 is a flowchart of a method for generating a work provided according to an embodiment of the present disclosure. As Figure 2 shown, this method may include the following steps:

[0026] Step S202, in response to an operation instruction for a first control on a first page, obtain a first audio and jump to a second page.

[0027] Wherein, the first page includes at least one candidate audio index, and the second page is used for editing text.

[0028] The above first audio may be an accompaniment audio.

[0029] The above first page may be a page for selecting an accompaniment audio, which may contain multiple candidate audios, each candidate audio corresponding to a first control, or the index of each candidate audio corresponding to a first control.

[0030] The above first control may be a confirmation button.

[0031] In an alternative embodiment, if a user wants to record a song, the user can enter the first page. The first page contains indexes of multiple candidate audio files. The user can select the desired audio file as the accompaniment audio according to the indexes of the multiple candidate audio files and click the first control corresponding to the index of the accompaniment audio file, so as to retrieve the accompaniment audio file from the database according to the index of the accompaniment audio file.

[0032] Further, after obtaining the accompaniment audio file, the user can be redirected to the second page to edit the lyrics of the work.

[0033] In an alternative embodiment, the second page can be used to edit the lyrics of the work. After obtaining the accompaniment audio file, the user can enter the page for editing lyrics to edit the lyrics.

[0034] Step S204: Obtain the first text on the second page. In response to the operation instruction for the second control on the second page, redirect to the third page.

[0035] Among them, the third page is used to provide the function of recording audio.

[0036] The above-mentioned second page may include a text box for editing text. The user can enter the first text in this text box.

[0037] The above-mentioned second control can be a button for completing text input.

[0038] The above-mentioned third page can be used to record the user's voice.

[0039] In an alternative embodiment, the user-edited lyrics can be obtained on the second page. After editing the lyrics, the user can press the button for completing text input, that is, the above-mentioned second control, to redirect to the page for recording the user's voice.

[0040] In another alternative embodiment, the lower limit and upper limit of the number of characters of the first text can be set. It can be set that when the number of characters exceeds the preset number of characters, the second control is displayed, so that the user can click the second control to perform the next step of audio recording.

[0041] In another alternative embodiment, after the user edits the text, the user can also process the edited text through the text processing control on the third page. For example, the unrhymed text can be marked so that the user can improve the edited text. After the user clicks the text processing control, related phrases associated with the text being edited can also be displayed in the text processing control for the user to select, giving the user inspiration for creating text.

[0042] Step S206, obtain the second audio in the third page. In response to the operation instruction for the third control in the third page, synthesize the first audio, the first text, and the second audio to generate a first work.

[0043] The above-mentioned third page may have a recording button, and the user can press the recording button to record sound.

[0044] The above-mentioned third control may be a control for synthesis.

[0045] In an alternative embodiment, after the user records the audio in the third page, the user can press the synthesis control in the third page to synthesize the accompaniment audio, the edited lyrics, and the recorded audio to generate a first work.

[0046] In another alternative embodiment, the user can wear headphones to record sound. After recording, the user can click the audition button on the third page. During the audition, the user can click to play the accompaniment audio. If the user feels that the recorded sound and the accompaniment audio are not aligned or the rhythm is not right, the user can manually splice and cut the recorded audio, and generate a first work based on the spliced and cut audio.

[0047] The above work generation method can be applied to the generation scenario of rap works.

[0048] Through the above steps, in response to the operation instruction for the first control in the first page, obtain the first audio, and jump to the second page. The first page includes at least one candidate audio index. The second page is used for editing text. Obtain the first text in the second page. In response to the operation instruction for the second control in the second page, jump to the third page. The third page is used to provide the function of recording audio. Obtain the second audio in the third page. In response to the operation instruction for the third control in the third page, synthesize the first audio, the first text, and the second audio to generate a first work, achieving the purpose of creating a work. By performing different types of creation links on different pages, the steps of the creation process can be simplified, the process of the creation process can be made flow-based, thereby improving the creation efficiency, and further solving the technical problem in the related technology that the creation efficiency of works is too low due to the complex creation process.

[0049] Optionally, after generating the first work based on the first audio, the first text, and the second audio, the method further includes: jump to the fourth page, obtain adjustment parameters in the fourth page, and adjust at least one sound effect of the first work based on the adjustment parameters to obtain the adjusted first work.

[0050] The above-mentioned fourth page is used to adjust the sound effects of the generated first work.

[0051] The above sound effects can be equalization, reverb, electro - sound, breathy voice, harmony, delay, etc.

[0052] The above - mentioned adjustment parameters can be obtained by the user dragging the adjustment box corresponding to at least one sound effect on the fourth page.

[0053] The above - mentioned fourth page can be a post - production room, where the post - production room is used to optimize the first work to obtain a better first work.

[0054] In an alternative embodiment, after jumping to the fourth page for sound effect adjustment, the user can adjust the sound effects in the adjustment box corresponding to each sound effect on the fourth page to obtain adjustment parameters. After obtaining the adjustment parameters, at least one sound effect of the first work can be adjusted according to the adjustment parameters to obtain an adjusted first work.

[0055] Optionally, after jumping to the fourth page, the method further includes: in response to an operation instruction for a fourth control on the fourth page, jumping to a fifth page, where the fifth page includes at least one candidate video index; in response to an operation instruction for a fifth control on the fifth page, obtaining a first video and jumping back to the fourth page; in response to an operation instruction for a sixth control on the fourth page, synthesizing the adjusted first work and the first video to generate a first work.

[0056] The above - mentioned fourth control can be a control for adding a video.

[0057] The above - mentioned fifth page can include multiple video indexes, and each video index corresponds to a fourth control.

[0058] In an alternative embodiment, after jumping to the post - production room, in response to an operation instruction for the video - adding control on the fourth page, jumping to the fifth page, the user can select a first video according to at least one video index on the fifth page. Specifically, the user can click on the fifth control corresponding to the video index to obtain the first video corresponding to the video index.

[0059] The above - mentioned sixth control can be a synthesis control on the fourth page.

[0060] The above - mentioned first video can be used to set off the atmosphere of the work, so that the work can be presented better.

[0061] In another alternative embodiment, after obtaining the first video, it can jump back to the fourth page, click on the sixth control on the fourth page, and synthesize the adjusted first work and the first video to generate a first work, so that the first work is more complete.

[0062] In another optional embodiment, after the first video is acquired, the first video may be processed, for example, special effects processing may be performed on the first video. Specifically, the user may select a video special effects template to add to the first video.

[0063] Optionally, after jumping back to the fourth page, the method also includes: synthesizing the first video and the first text in the fourth page to obtain a second video; in response to an operation instruction for a sixth control in the fourth page, synthesizing the adjusted first work and the second video to generate the first work.

[0064] The first text mentioned above may be edited lyrics.

[0065] In an optional embodiment, when jumping back to the fourth page, the first video and the first text can be synthesized on the fourth page to obtain a second video, wherein the second video is a video containing lyrics. Since the second video contains the lyrics of the recorded song, the user can see the lyrics while watching the second video, thereby improving the viewing experience of the generated first work.

[0066] In another optional embodiment, after synthesizing the first video and the first text to obtain the second video, the user can adjust the first text in the second video so that the display of the first text can be aligned with the sound of the recorded song.

[0067] In another optional embodiment, the font, color, etc. of the lyrics in the second video may also be set to make the second video more beautiful.

[0068] Optionally, the second audio is obtained in the third page, and in response to the operation instruction for the third control in the third page, the first audio, the first text and the second audio are synthesized to generate the first work, including: obtaining clipping parameters in the third page, and clipping the second audio based on the clipping parameters to obtain the third audio; in response to the operation instruction for the third control in the third page, the first audio, the first text and the third audio are synthesized to generate the first work.

[0069] In an optional embodiment, after obtaining the second audio on the third page, the recorded audio can be edited on the third page according to the editing parameters, and unnecessary sounds can be removed, or sounds to be retained can be selected to obtain the third audio. The editing parameters can be obtained by the user manually editing the obtained second audio. After obtaining the third audio, the user can click the synthesis control on the third page, that is, the third control mentioned above, to synthesize the first audio, the first text and the third audio to generate the first work.

[0070] In another alternative embodiment, the user can wear headphones to record sound. After the recording is completed, the user can click the audition button on the third page. During the audition process, the user can click to play the accompaniment audio. If the user feels that the recorded sound is not aligned with the accompaniment audio or the rhythm is not right, the user can manually splice and cut the recorded audio to obtain the above-mentioned editing parameters, and edit the second audio according to the editing parameters to obtain the third audio. Based on the generated third audio, the first work is obtained, making the obtained first work more complete.

[0071] Optionally, obtain the second audio on the third page, and in response to the operation instruction for the third control on the third page, synthesize the first audio, the first text, and the second audio to generate the first work, including: obtaining the timestamp of the second audio on the third page, and using the timestamp to annotate the first text to obtain the second text; in response to the operation instruction for the third control on the third page, synthesize the first audio, the second text, and the third audio to generate the first work.

[0072] In one alternative embodiment, after obtaining the second audio on the third page, the corresponding timestamp of the second audio can be obtained, where the timestamp is used to represent the playback time of each segment of sound in the second audio. Exemplarily, if there are 10 segments of sound in the second audio, the time when the first sentence of sound starts to be recorded can be determined as the start playback time.

[0073] In another alternative embodiment, the audio can be segmented according to the pauses in the second audio.

[0074] Since the second audio corresponds to the first text, the second audio and the first text can be matched using the timestamp to determine the lyrics corresponding to each segment of sound in the second audio.

[0075] In another alternative embodiment, after obtaining the timestamp of the second audio, the first text can be annotated using the timestamp to obtain the second text, where the second text corresponds to the playback time of the second audio. After obtaining the second text, the user can click the synthesis control on the third page, that is, the above-mentioned third control, and then synthesize the first audio, the second text, and the third audio based on the synthesis control to generate the first work.

[0076] The following combines Figure 3 A detailed description will be given of an embodiment of a method for generating a work according to the present disclosure. This step includes:

[0077] Step S301, select the accompaniment audio;

[0078] Step S302, edit the lyrics;

[0079] Step S303, record audio;

[0080] Step S304, determine whether the recording time of the audio is greater than 3S. If so, execute Step S306; if not, execute Step S305;

[0081] Step S305, click to record again;

[0082] Step S306, edit the recorded audio to obtain the edited audio;

[0083] Step S307, determine whether it is necessary to adjust the human voice in the edited audio. If so, execute Step S308; if not, execute Step S309;

[0084] Step S308, slide the progress bar to change the audio time to be adjusted and adjust the audio;

[0085] Step S309, audition the audio and determine that the audio production is completed;

[0086] Optionally, after the audio is produced, the produced audio can be uploaded to the local.

[0087] Step S310, annotate the lyrics for the produced audio;

[0088] Step S311, jump to the post-production room and select the video corresponding to the audio;

[0089] Step S312, decorate the video, lyrics, and audio;

[0090] Optionally, a video effect template can be selected to decorate the video, the font and color of the lyrics can be selected to decorate the lyrics, and the audio can be decorated with at least one sound effect.

[0091] The above sound effects can be equalization, reverb, electro - sound, breathy voice, harmony, delay, etc.

[0092] Step S313, synthesize the decorated video, lyrics, and audio to generate a work;

[0093] Step S314, adjust the generated work to generate a complete work.

[0094] Furthermore, a video file corresponding to the complete work can also be generated.

[0095] Optionally, the work can be adjusted by setting filters, subtitles, width, height, etc. An audio track set, a rap audio track, an accompaniment audio track, and an empty audio track can also be set for the work to ensure the audio noise reduction and gain effects of the above work. Subtitles, stickers, etc. can also be synthesized for the work.

[0096] Through the above steps, complex audio production can be simplified, enabling music lovers to easily and with low barriers complete their own works, bringing good enhancement effects to the works, attracting more music lovers to create, and increasing the usage rate of music production software.

[0097] According to an embodiment of the present disclosure, there is also provided a work generation device, which can execute the work generation method in the above embodiment. The specific implementation manner and preferred application scenario are the same as those in the above embodiment and will not be elaborated here.

[0098] Figure 4 is a schematic diagram of a work generation device according to an embodiment of the present disclosure, as Figure 4 shown, the device includes:

[0099] An acquisition module 402, in response to an operation instruction for a first control on a first page, is configured to acquire a first audio and jump to a second page, where the first page includes at least one candidate audio index, and the second page is for editing text;

[0100] A jump module 404, configured to acquire a first text on the second page and, in response to an operation instruction for a second control on the second page, jump to a third page, where the third page is for providing a function of recording audio;

[0101] A synthesis module 406, configured to acquire a second audio on the third page and, in response to an operation instruction for a third control on the third page, synthesize the first audio, the first text, and the second audio to generate a first work.

[0102] Optionally, the device further includes: an adjustment module, configured to jump to a fourth page, acquire adjustment parameters on the fourth page, and adjust at least one sound effect of the first work based on the adjustment parameters to obtain an adjusted first work.

[0103] Optionally, the jump module is further configured to, in response to an operation instruction for a fourth control on the fourth page, jump to a fifth page, where the fifth page includes at least one candidate video index; the acquisition module is further configured to, in response to an operation instruction for a fifth control on the fifth page, acquire a first video and jump back to the fourth page; the synthesis module is further configured to, in response to an operation instruction for a sixth control on the fourth page, synthesize the adjusted first work and the first video to generate a first work.

[0104] Optionally, the device further includes: the synthesis module is further configured to synthesize the first video and the first text on the fourth page to obtain a second video; a generation module, configured to, in response to an operation instruction for a sixth control on the fourth page, synthesize the adjusted first work and the second video to generate a first work.

[0105] Optionally, the synthesis module includes: a clipping unit, configured to obtain clipping parameters in the third page, clip the second audio based on the clipping parameters to obtain a third audio; and a synthesis unit, configured to synthesize the first audio, the first text, and the third audio in response to an operation instruction for a third control in the third page to generate a first work.

[0106] Optionally, the synthesis unit includes: an annotation subunit, configured to obtain the timestamp of the second audio in the third page, annotate the first text using the timestamp to obtain a second text; and a synthesis subunit, configured to synthesize the first audio, the second text, and the third audio in response to an operation instruction for a third control in the third page to generate a first work.

[0107] It should be noted that the above-mentioned various modules can be implemented by software or hardware. For the latter, it can be implemented in the following ways, but not limited thereto: the above-mentioned modules are all located in the same processor; or, the above-mentioned various modules are separately located in different processors in any combination form.

[0108] According to an embodiment of the present disclosure, there is also provided an electronic device, including: at least one processor; and a memory communicatively connected to the at least one processor; wherein, the memory stores instructions executable by the at least one processor, and the instructions are executed by the at least one processor to enable the at least one processor to execute the method for generating any one of the above-mentioned works.

[0109] According to an embodiment of the present disclosure, there is also provided a non-transitory computer-readable storage medium storing computer instructions, wherein the computer instructions are used to cause a computer to execute the method for generating any one of the above-mentioned works.

[0110] Optionally, in this embodiment, the above-mentioned non-transitory computer-readable storage medium may include, but is not limited to, electronic, magnetic, optical, electromagnetic, infrared, or semiconductor systems, devices, or equipment, or any suitable combination of the above. More specific examples of the readable storage medium would include electrical connections based on one or more wires, portable computer disks, hard disks, random access memory (RAM), read-only memory (ROM), erasable programmable read-only memory (EPROM or flash memory), optical fibers, portable compact disk read-only memory (CD-ROM), optical storage devices, magnetic storage devices, or any suitable combination of the above.

[0111] According to an embodiment of the present disclosure, there is also provided a computer program product including a computer program which, when executed by a processor, implements the generation method of any of the above works. The program code for implementing the generation method of the works of the present disclosure can be written in any combination of one or more programming languages. These program codes can be provided to a processor or a controller of a general-purpose computer, a special-purpose computer, or other programmable data processing devices, so that when the program codes are executed by the processor or the controller, the functions / operations specified in the flowchart and / or the block diagram are implemented. The program codes can be executed entirely on the machine, partially on the machine, executed partially on the machine as an independent software package and partially on a remote machine, or executed entirely on a remote machine or a server.

[0112] In the above embodiments of the present disclosure, the descriptions of the respective embodiments have their own focuses. For parts not detailed in a certain embodiment, reference can be made to the relevant descriptions of other embodiments.

[0113] In several embodiments provided by the present disclosure, it should be understood that the disclosed technical content can be implemented in other ways. Among them, the device embodiments described above are merely illustrative. For example, the division of the units can be a logical function division, and there can be other division methods in actual implementation. For example, multiple units or components can be combined or integrated into another system, or some features can be ignored or not executed. Another point is that the displayed or discussed couplings or direct couplings or communication connections to each other can be through some interfaces, and the indirect couplings or communication connections of the units or modules can be in electrical or other forms.

[0114] The units described as separate components may or may not be physically separated, and the components displayed as units may or may not be physical units, that is, they can be located in one place or distributed to multiple units. Some or all of the units can be selected according to actual needs to achieve the purpose of the solution of this embodiment.

[0115] In addition, the functional units in various embodiments of the present disclosure can be integrated in a processing unit, or each unit can exist physically alone, or two or more units can be integrated in one unit. The above integrated units can be implemented in the form of hardware or in the form of software functional units.

[0116] When the integrated unit is implemented in the form of a software functional unit and sold or used as an independent product, it can be stored in a computer-readable storage medium. Based on such an understanding, the technical solution of the present disclosure, in essence, or the part that contributes to the prior art, or all or part of the technical solution, can be embodied in the form of a software product. The computer software product is stored in a storage medium and includes several instructions for causing a computer device (which may be a personal computer, a server, or a network device, etc.) to execute all or part of the steps of the methods described in the various embodiments of the present disclosure. The foregoing storage medium includes: various media that can store program codes, such as USB flash drives, read-only memories (ROMs), random access memories (RAMs), mobile hard disks, magnetic disks, or optical discs.

[0117] The foregoing are only the preferred embodiments of the present disclosure. It should be noted that for those of ordinary skill in the art, without departing from the principle of the present disclosure, several improvements and refinements can be made, and these improvements and refinements should also be regarded as the protection scope of the present disclosure.

Claims

1. A method for generating a work, comprising: In response to an operation instruction for a first control on a first page, obtaining a first audio and jumping to a second page, where the first page includes at least one candidate audio index, and the second page is used for editing text; Obtaining a first text on the second page, and in response to an operation instruction for a second control on the second page, jumping to a third page, where the third page is used to provide an audio recording function; Obtaining a second audio on the third page, and in response to an operation instruction for a third control on the third page, synthesizing the first audio, the first text, and the second audio to generate a first work; Wherein, the method further includes: Jumping to a fourth page, obtaining adjustment parameters on the fourth page, and adjusting at least one sound effect of the first work based on the adjustment parameters to obtain an adjusted first work; Wherein, after jumping to the fourth page, the method further includes: In response to an operation instruction for a fourth control on the fourth page, jumping to a fifth page, where the fifth page includes at least one candidate video index; In response to an operation instruction for a fifth control on the fifth page, obtaining a first video and jumping back to the fourth page; After jumping back to the fourth page, in response to an operation instruction for a sixth control on the fourth page, synthesizing the adjusted first work and the first video to generate the first work; Or, after jumping back to the fourth page, synthesizing the first video and the first text on the fourth page to obtain a second video; in response to an operation instruction for the sixth control on the fourth page, synthesizing the adjusted first work and the second video to generate the first work.

2. The method according to claim 1, wherein Obtaining a second audio on the third page, and in response to an operation instruction for a third control on the third page, synthesizing the first audio, the first text, and the second audio to generate a first work, including: Obtaining clip parameters on the third page, and clipping the second audio based on the clip parameters to obtain a third audio; In response to an operation instruction for the third control on the third page, synthesizing the first audio, the first text, and the third audio to generate the first work.

3. The method according to claim 2, wherein Obtaining a second audio on the third page, and in response to an operation instruction for a third control on the third page, synthesizing the first audio, the first text, and the second audio to generate a first work, including: Obtaining the timestamp of the second audio on the third page, and annotating the first text with the timestamp to obtain a second text; In response to an operation instruction for the third control on the third page, synthesizing the first audio, the second text, and the third audio to generate the first work.

4. A device for generating a work, comprising: An acquisition module, configured to obtain a first audio in response to an operation instruction for a first control in a first page, and jump to a second page, where the first page includes at least one candidate audio index, and the second page is used for editing text; A jump module, configured to obtain a first text in the second page, and jump to a third page in response to an operation instruction for a second control in the second page, where the third page is used to provide an audio recording function; A synthesis module, configured to obtain a second audio in the third page, and synthesize the first audio, the first text, and the second audio in response to an operation instruction for a third control in the third page to generate a first work; Wherein, the device is further configured to jump to a fourth page, obtain adjustment parameters in the fourth page, and adjust at least one sound effect of the first work based on the adjustment parameters to obtain an adjusted first work; Wherein, after jumping to the fourth page, the device is further configured to jump to a fifth page in response to an operation instruction for a fourth control in the fourth page, where the fifth page includes at least one candidate video index; obtain a first video in response to an operation instruction for a fifth control in the fifth page, and jump back to the fourth page; after jumping back to the fourth page, synthesize the adjusted first work and the first video in response to an operation instruction for a sixth control in the fourth page to generate the first work; Or, after jumping back to the fourth page, synthesize the first video and the first text in the fourth page to obtain a second video; synthesize the adjusted first work and the second video in response to an operation instruction for the sixth control in the fourth page to generate the first work.

5. An electronic device, comprising: At least one processor; And A memory communicatively connected to the at least one processor; Wherein, the memory stores instructions executable by the at least one processor, and the instructions are executed by the at least one processor to enable the at least one processor to execute the method according to any one of claims 1-3.

6. A non-transitory computer-readable storage medium storing computer instructions, wherein, The computer instructions are used to cause the computer to execute the method according to any one of claims 1-3.

7. A computer program product, comprising a computer program, where the computer program, when executed by a processor, implements the method according to any one of claims 1-3.

Citation Information

Patent Citations

  • Audio synthesis method and device, storage medium and computer device

    CN110189741A