Audio processing method, device, electronic device, and storage medium
The audio processing method enhances user interaction and editing efficiency by automatically modifying audio in video applications, improving audio quality and enabling concurrent editing operations.
Patent Information
- Application Number
- JP2023578707
- Authority / Receiving Office
- JP · JP
- Patent Type
- Patents
- Current Assignee / Owner
- Priority Date
- 2021-09-28
- Filing Date
- 2022-09-19
- Publication Date
- 2025-08-26
- Estimated Expiration
- 2042-09-19
AI Technical Summary
The quality of user-created works in video applications is low due to the low quality of audio recordings, leading to a small number of users posting their creations.
An audio processing method that automatically performs sound modification on the original audio when a video editing page is accessed, displaying a target control indicating the process, allowing simultaneous execution of other editing operations on the video.
Improves audio quality, enhances user interaction, and increases editing efficiency by allowing users to enjoy a good audio playback effect while performing other editing operations concurrently.
Smart Images

Figure 0007729927000001 
Figure 0007729927000002 
Figure 0007729927000003
Abstract
Description
[Technical Field]
[0001] CROSS-REFERENCE TO RELATED APPLICATIONS This disclosure claims priority to an application filed on September 28, 2021, entitled "Audio Processing Method, Apparatus, Electronic Device, and Storage Medium," with Chinese Patent Application Number 202111145023.8, the entire contents of which are incorporated herein by reference.
[0002] Technical Field The present disclosure relates to the field of information technology, and in particular to audio processing methods, devices, electronic equipment, and storage media. [Background technology]
[0003] With the rapid development of terminal and network technologies, current video applications generally have functions such as uploading works. Users can create works through video applications, for example, by recording audio and video.
[0004] However, the relevant statistical data shows that although the number of users who create works using video applications is large, the number of users who post works is small, which may be due to the low quality of the works created by users using video applications.
[0005] Therefore, how to improve the quality of users' creations is a major problem to be solved. Summary of the Invention
[0006] To solve or at least partially solve the above technical problems, embodiments of the present disclosure provide an audio processing method, device, electronic device, and storage medium that can improve the quality of the original audio and achieve a good audio playback effect.
[0007] According to a first aspect, an embodiment of the present disclosure comprises: displaying a video editing page in response to a trigger operation that first accesses the video editing page for the target video; When the original audio in the target video satisfies a predetermined condition, the video editing page is displayed, and simultaneously, an audio correction process is performed on the original audio, and a target control in a first state is displayed on the video editing page, where the target control in the first state is for indicating that an audio correction process is being performed on the original audio; An audio processing method is provided, which includes, when performing sound modification processing on the original audio, performing an editing operation corresponding to the editing control on the target video in response to a trigger operation acting on the editing control, and displaying at least one editing control on the video editing page.
[0008] According to a second aspect, an embodiment of the present disclosure comprises: a first display module for displaying the video editing page in response to a trigger operation for first accessing the video editing page for the target video; a processing module for displaying the video editing page and simultaneously performing sound modification processing on the original audio when the original audio in the target video satisfies a predetermined condition; the first display module is further used to display a target control in a first state on the video editing page, the target control in the first state being for indicating that a sound modification process is being performed on the original audio; The audio processing device further includes an editing module for executing an editing operation corresponding to the editing control on the target video in response to a trigger operation acting on the editing control when performing sound modification processing on the original audio, the editing module displaying at least one editing control on the video editing page.
[0009] According to a third aspect, an embodiment of the present disclosure comprises: one or more processors; a storage device for storing one or more programs; The present invention further provides an electronic device that, when the one or more programs are executed by the one or more processors, causes the one or more processors to implement the audio processing method described above.
[0010] According to a fourth aspect, an embodiment of the present disclosure further provides a computer-readable storage medium having stored thereon a computer program that, when executed by a processor, causes the above audio processing method to be implemented.
[0011] According to a fifth aspect, embodiments of the present disclosure further provide a computer program product including a computer program or instructions that, when executed by a processor, causes the computer program or instructions to implement the above audio processing method.
[0012] The technical solution according to the embodiment of the present disclosure has at least the following advantages over the related art: When a video editing page is accessed, the audio processing method according to the embodiment of the present disclosure automatically performs audio modification processing on the original audio of a video, and displays a target control in a first state on the video editing page to indicate that the original audio is being modified, thereby realizing intelligent audio modification of the original audio, improving the quality of the original audio, and achieving a good audio playback effect. Furthermore, the target control in the first state can inform the user that audio modification processing is currently being performed on the original audio, allowing the user to enjoy a good interaction experience. Furthermore, while the original audio is being modified, other editing operations performed by the user on the target video are not affected, further improving the user experience and improving editing efficiency. [Brief explanation of the drawings]
[0013] These and other features, advantages, and aspects of the embodiments of the present disclosure will become more apparent from the following detailed description taken in conjunction with the accompanying drawings, in which the same or similar elements are designated by the same or similar reference numerals throughout the drawings. It should be understood that the drawings are schematic and that the materials and elements are not necessarily drawn to scale. [Figure 1] 1 is a flowchart of an audio processing method according to an embodiment of the present disclosure. [Figure 2] 1 is a schematic diagram of a video recording page according to an embodiment of the present disclosure. [Figure 3] FIG. 1 is a schematic diagram of a video editing page according to an embodiment of the present disclosure. [Figure 4] FIG. 1 is a schematic diagram of a video editing page according to an embodiment of the present disclosure. [Figure 5] FIG. 1 is a schematic diagram of a video editing page according to an embodiment of the present disclosure. [Figure 6] 1 is a schematic diagram illustrating the configuration of a video processing device according to an embodiment of the present disclosure. [Figure 7] 1 is a schematic diagram illustrating the configuration of an electronic device according to an embodiment of the present disclosure. DETAILED DESCRIPTION OF THE INVENTION
[0014] The following describes in more detail the embodiments of the present disclosure with reference to the accompanying drawings. Although several embodiments of the present disclosure are shown in the accompanying drawings, it should be understood that the present disclosure can be realized in various forms and should not be construed as being limited to the embodiments described herein. On the contrary, these embodiments are provided to provide a more thorough and complete understanding of the present disclosure. It should be understood that the accompanying drawings and embodiments of the present disclosure are merely illustrative and do not limit the scope of protection of the present disclosure.
[0015] It should be understood that the steps described in the method embodiments of the present disclosure may be performed in a different order and / or in parallel. Furthermore, method embodiments may include additional steps and / or omit performing steps as shown. The scope of the present disclosure is not limited in this respect.
[0016] As used herein, the term "comprises" and variations thereof are open-ended, meaning "including, but not limited to." The term "based on" means "based at least in part on." The term "in one embodiment" means "at least one embodiment," the term "another embodiment" means "at least one other embodiment," and the term "some embodiments" means "at least some embodiments." Relevant definitions of other terms are provided in the description below.
[0017] It should be noted that the concepts of "first," "second," etc. described in this disclosure are merely intended to distinguish between different devices, modules, or units, and are not intended to limit the order or interdependence of functions performed by these devices, modules, or units.
[0018] It should be noted that the modifications of "one" and "multiple" described in the present disclosure are exemplary and not limiting. It is obvious to those skilled in the art that unless otherwise specified, it should be understood as "one or multiple."
[0019] The names of messages or information exchanged between devices in the embodiments of the present disclosure are merely descriptive and do not limit the scope of these messages or information.
[0020] 1 is a flowchart of an audio processing method according to an embodiment of the present disclosure. This embodiment is applied to a case where a video client performs video recording, typically a case where a karaoke performance is recorded, to improve the user's singing effect and beautify the user's singing audio. The method may be performed by an audio processing device, which may be implemented in software and / or hardware, and which may be located in an electronic device, such as a terminal, including, but not limited to, a smartphone, a palmtop, a tablet, a wearable device with a display, a desktop, a laptop, an all-in-one device, a smart home device, etc.
[0021] As shown in FIG. 1, the method may specifically include the following steps:
[0022] Step 110: In response to a trigger operation that accesses the video editing page for the target video for the first time, the video editing page is displayed.
[0023] Here, the target video may be recorded in real time by a user. For example, the user records, or in other words, shoots, the target video through a shooting page of a video application, and then jumps to a video editing page for the target video based on the shooting page of the video application. This process is the first time the user accesses the video editing page for the target video. The target video may also be a video selected from the user's album. The video may also be a video previously recorded or downloaded by the user.
[0024] Specifically, the target video may be a video containing audio of a user singing karaoke, recorded in the case of singing karaoke. In the case of singing karaoke, the user can select a song as a predetermined reference audio through the video application. The lyrics of the predetermined reference audio are displayed on the recording page of the video application, and when the camera of the video application is on, the recording page may also display an image of the user. When the camera of the video application is off (the user can customize the camera on and off), the recording page may further display a predetermined screen, such as a music video screen of the song itself.
[0025] In some embodiments, the recording page for the karaoke singing scenario further displays target presentation information. The target presentation information is determined based on the attribute features of the song selected by the user and the user's recording behavior. Specifically, if a sound modification resource (e.g., a MIDI file) corresponding to the song selected by the user exists in the video application or in a server associated with the video application, i.e., if the song selected by the user supports intelligent sound modification, the target presentation information 1 may be "The current song supports intelligent sound modification. Please wear wired earphones at all times." The target presentation information is intended to remind the user to always wear wired earphones when recording. Audio recorded using wired earphones has a high effect, is of good quality, and can provide an original resource with excellent intelligent sound modification, ensuring that the user can record effective audio. If the video application or a server associated with the video application does not have a sound modification resource (e.g., a MIDI file) corresponding to the song selected by the user, in other words, if the song selected by the user does not support intelligent sound modification and the user is not wearing wired earphones but is using wireless earphones but the performance of the earphones is poor, then target presentation information 2 may be "activate the microphone to record according to the earphone performance." If the video application or a server associated with the video application does not have a sound modification resource (e.g., a MIDI file) corresponding to the song selected by the user, in other words, if the song selected by the user does not support intelligent sound modification and the user is not using earphones but is using a microphone to record, then target presentation information 3 may be "wearing earphones will be more effective."
[0026] In some embodiments, the above target presentation information 1, target presentation information 2, and target presentation information 3 may be displayed alternately in order regardless of whether the song selected by the user supports intelligent sound modification. Displaying the target presentation information on the recording page can accurately guide the user's recording behavior, guide the user to record effective singing audio, and improve the user's user experience.
[0027] Generally, target presentation information is displayed on the recording page of the target video based on the attribute features of a predetermined reference audio (i.e., the audio of a song selected by the user) and / or the user's recording behavior. The audio in the target video is recorded based on the predetermined reference audio. For example, see the interface schematic diagram of a video recording page for singing karaoke shown in FIG. 2 . The lyrics of the song selected by the user, a predetermined image (a background image below the lyrics), and target presentation information 210 are displayed. When the user triggers the recording control 220, recording begins. By triggering the camera control 230, the camera state can be switched, for example, to turn the camera on or off. Note that, due to limited display area, if the target presentation information 210 contains a large number of characters, it may be displayed in a scrolling manner, like a "kaleidoscope." In some embodiments, the target presentation information 210 is displayed only before recording, and the display of the target presentation information 210 is canceled when the user triggers the recording control 220 to start recording. The user may manually turn off the display of the target presentation information 210 by triggering the off control “x” 211 .
[0028] Step 120: If the original audio in the target video satisfies a predetermined condition, the video editing page is displayed, and at the same time, sound correction processing is performed on the original audio, and a target control in a first state is displayed on the video editing page to indicate that sound correction processing is being performed on the original audio.
[0029] Here, the predetermined condition may be that the video application or a server associated with the video application has an audio modification resource corresponding to the song selected by the user, and that the user always wears wired earphones during recording. It may also be said that a corresponding audio modification resource exists for the original audio of the target video, and that the original audio was recorded with wired earphones always worn. If the original audio of the target video satisfies the predetermined condition, audio modification processing is performed on the original audio simultaneously with accessing the video editing page. This eliminates the need for the user to perform audio modification processing on the original video after triggering the relevant audio modification processing control. This realizes automated processing of the original audio, improving the processing efficiency and user experience of the original audio. To allow the user to understand the audio modification processing performed on the original audio, a target control in a first state is displayed on the video editing page to indicate that audio modification processing is being performed on the original audio. This allows the user to easily and timely understand the related processing performed on the original audio.
[0030] For example, refer to the schematic diagram of the video editing page shown in Figure 3. In this figure, a target control 310 in a first state is displayed, which indicates that the user is currently performing sound modification processing on the original audio of the target video.
[0031] Step 130: When performing sound modification processing on the original audio, in response to a trigger operation acting on an editing control, an editing operation corresponding to the editing control is performed on the target video, and at least one editing control is displayed on the video editing page.
[0032] For example, refer to the schematic diagram of the video editing page shown in FIG. 3 . The video editing page further displays at least one editing control, such as a “Text” control 320, a “Sticker” control 330, a “Filter” control 340, an “Effect” control 350, or a “Enhance” control 360. When performing sound correction processing on the original audio, if the user accepts an operation triggered by the “Text” control 320, an editing page for adding text is displayed. If the user accepts an operation triggered by the “Filter” control 340, a filter is applied to the target video. In other words, while performing sound correction processing on the original audio, other editing operations performed by the user on the target video are not affected. For example, the user can add editing operations such as filters, text, stickers, or enhancements to the target video. In other words, while performing automatic sound correction processing on the original audio, the user can manually perform other editing operations on the target video, such as adding text, stickers, or filters, according to their needs. In this way, when the sound correction process for the original audio is completed, other editing operations performed manually by the user are also basically completed, and rather than waiting for the sound correction process to be completed and then performing other editing operations manually, the former can further improve the editing efficiency for the target video, reduce the user's effort, and achieve the purpose of improving the user experience.
[0033] In some embodiments, the original audio is played back simultaneously with the video editing page being displayed, and upon detecting completion of the audio modification process, the modified audio continues to be played back from the playback progress position of the original audio at the time of completion of the audio modification process. For example, if the audio modification process takes 3 seconds, when the audio modification process is completed, the playback progress of the original audio reaches the 3-second position. Instead of starting playback from the beginning of the original audio, i.e., the playback progress does not start from 0 seconds, but starts from the 3-second position.
[0034] In some embodiments, the target control 410 is displayed in a second state to indicate that the sound modification process on the original audio is complete, with reference to the schematic diagram of a video editing page shown in Figure 4. Generally, upon detecting completion of the sound modification process, the state of the target control is controlled to switch to the second state, and the target control in the second state is used to indicate that the sound modification process on the original audio is complete.
[0035] In some embodiments, when a user triggers the target control 410, the state of the target control 410 is controlled to switch to a third state, for example, the target control 510 in the third state shown in FIG. 5 . The target control 510 in the third state indicates that no sound correction processing is performed on the original audio. After the state of the target control 510 is switched to the third state, the original audio continues to be played from the current playback progress position on the video editing page, and the sound-corrected audio is not played. Generally, in response to an operation to switch the state of the target control from the second state to the third state, the original audio continues to be played from the current playback progress position, and when the target control is in the second state, the sound-corrected audio is played on the video editing page. For example, when the video editing page is accessed and playback of the original audio is started, sound correction processing is automatically performed on the original audio, and the state of the target control is in the first state. If the sound correction process is completed at 3s and the state of the target control is switched from the first state to the second state, the playback progress of the original audio reaches 3s. The playback progress plays the audio after sound correction from 4s, and when the user switches the state of the target control from the second state to the third state at 6s, the playback progress continues to play the original audio from 7s.
[0036] In some embodiments, in response to a trigger operation acting on a posting control (e.g., posting control 420 in FIG. 4 or posting control 520 in FIG. 5), if the target control is in the second state, a target video including audio after sound modification processing is posted, and if the target control is in the third state, a target video including the original audio is posted. That is, when a user posts a target video, the user can choose whether to post the audio after sound modification or the original audio before sound modification. By providing users with multiple choices, the personalization needs of different users can be individually met. For example, some users may think they are better singers and do not like sound effects after sound modification. Therefore, the user may switch the state of the target control to the third state before posting the target video. On the other hand, some users may think they are tone-deaf and cannot sing well, and prefer sound effects after sound modification. Therefore, such users may switch the state of the target control to the second state before posting the target video.
[0037] In some embodiments, a video editing page including the target control is displayed in response to an operation to return to the video editing page after exiting the video editing page. Here, if the original audio has not changed, the state of the target control is controlled to maintain the same state as the state of the target control at the time of exiting the video editing page. That is, if a user exits the video editing page but does not change the original audio, when the user accesses the video editing page again, the state of the target control at the time of exit is maintained, and sound correction processing is not performed again on the original audio. For example, if the state of the target control is in a first state, i.e., "sound correction in progress," when the user exits the video editing page but does not change the original audio, sound correction processing continues on the original audio when the user returns to the video editing page, and the state of the target control remains in the first state, i.e., "sound correction in progress." If there is a change in the original audio, sound correction processing is automatically performed on the changed original audio when the user returns to the video editing page, and the state of the target control becomes the first state, i.e., "sound correction in progress." When the user exits the video editing page, if the state of the target control is in the second state, i.e., "audio correction completed", if the user exits the video editing page but does not change the original audio, when the user returns to the video editing page, the audio correction process will not be performed again on the original audio, but the audio for which the previous audio correction process has been completed will be directly called, and the state of the target control will maintain the second state, i.e., "audio correction completed". If there is a change in the original audio, when the user returns to the video editing page, the audio correction process will automatically be performed on the changed original audio, and the state of the target control will return to the first state, i.e., "audio correction in progress".When the user exits the video editing page, if the state of the target control is in the third state, i.e., "no audio correction," and if the user exits the video editing page but does not change the original audio, when the user returns to the video editing page, the state of the target control will maintain the third state, i.e., "no audio correction." If the user switches the state of the target control to the second state, i.e., "audio correction completed," the audio will not be corrected again for the original audio, and the audio for which the previous audio correction process has been completed will be directly called. If there is a change in the original audio, when the user returns to the video editing page, the state of the target control will maintain the third state, i.e., "no audio correction." If the user switches the state of the target control to the second state, i.e., "audio correction completed," the audio will automatically be corrected for the changed original audio, and the state of the target control will be controlled to the first state, i.e., "audio correction in progress."
[0038] In some embodiments, displaying the video editing page including the target control in response to an operation of returning to the video editing page after exiting the video editing page includes: in response to a trigger operation of accessing a video recording page after exiting the video editing page and returning to the video editing page after exiting the video recording page, if the state of the target control was in a first state or a second state when the video editing page was exited, displaying the video editing page while performing an audio correction process on the changed original audio and displaying the target control in the first state on the video editing page; or in response to a trigger operation of accessing a video recording page after exiting the video editing page and returning to the video editing page after exiting the video recording page, if the state of the target control was in a third state when the video editing page was exited, displaying the video editing page and in response to an operation of switching the state of the target control from the third state to the second state, performing an audio correction process on the changed original audio and displaying the target control in the first state on the video editing page.
[0039] In some embodiments, displaying the video editing page including the target control in response to an operation of returning to the video editing page after exiting the video editing page as described above includes accessing a video posting page after exiting the video editing page, and displaying the video editing page including the target control in response to a trigger operation of returning to the video editing page after exiting the video posting page, and controlling the state of the target control to remain consistent with the state of the target control at the time the video editing page was exited.
[0040] Specifically, whether the original audio has changed may be determined based on the target page accessed after the user exits the video editing page. For example, if the user accesses a video recording page after exiting the video editing page, the original audio may be considered to have changed when the user returns from the video recording page to the video editing page. If the user accesses a video posting page after exiting the video editing page and then returns from the video posting page to the video editing page, the original audio may be considered not to have changed.
[0041] In an audio processing method according to an embodiment of the present disclosure, when a video editing page is accessed for the first time, audio modification processing is automatically performed on the original audio of the video, and a target control in a first state is displayed on the video editing page to indicate that audio modification processing is being performed on the original audio, thereby realizing intelligent audio modification of the original audio, improving the quality of the original audio, and achieving a good audio playback effect. Furthermore, the target control in the first state notifies the user that audio modification processing is currently being performed on the original audio, allowing the user to enjoy a good interaction experience. Furthermore, while the audio modification processing on the original audio is being performed, other editing operations performed by the user on the target video are not affected, further improving the user experience and improving editing efficiency.
[0042] 6 is a schematic diagram of an audio processing device according to an embodiment of the present disclosure. The audio processing device according to the embodiment of the present disclosure may be disposed in a client. The audio processing device 60 specifically includes a first display module 610, a processing module 620, and an editing module 630.
[0043] Wherein, the first display module 610 is used to display the video editing page in response to a trigger operation to access the video editing page for the target video for the first time; the processing module 620 is used to display the video editing page and perform sound correction processing on the original audio when the original audio in the target video satisfies a predetermined condition; the first display module 610 is further used to display a target control in a first state on the video editing page to indicate that sound correction processing is being performed on the original audio; and the editing module 630 is used to perform an editing operation corresponding to the editing control on the target video in response to a trigger operation acting on the editing control when performing sound correction processing on the original audio, and display at least one editing control on the video editing page.
[0044] In some embodiments, the audio processing device further includes a playback module for playing the original audio while displaying the video editing page, and for, upon detecting completion of the sound modification process, continuing to play the modified audio from the playback progress position of the original audio at the time of completion of the sound modification process.
[0045] In some embodiments, the audio processing device further includes a control module for controlling the state of the target control to switch to a second state upon detecting completion of the sound modification process, the target control being in the second state being used to indicate that the sound modification process on the original audio has been completed.
[0046] In some embodiments, the playback module further continues to play the original audio from the current playback progress position in response to switching the state of the target control from the second state to a third state, and when the target control is in the second state, is used to play the audio after the sound modification process on the video editing page.
[0047] In some embodiments, the audio processing device further includes a posting module for, in response to a trigger operation acting on a posting control, posting a target video including audio after sound modification processing when the target control is in the second state, and posting a target video including the original audio when the target control is in the third state.
[0048] In some embodiments, the first display module is further used to display the video editing page in response to an operation of exiting the video editing page, accessing a specified page, and then returning to the video editing page from the specified page, and to control the state of the target control to remain consistent with the state of the target control when the video editing page was exited.
[0049] In some embodiments, the first display module is specifically used to: access a video recording page after exiting from a video editing page, and in response to a trigger operation of returning to the video editing page after exiting from the video recording page, if the state of the target control was in a first state or a second state when the video editing page was exited, display the video editing page while performing an audio correction process on the changed original audio and display the target control in the first state on the video editing page; or to: access a video recording page after exiting from the video editing page, and in response to a trigger operation of returning to the video editing page after exiting from the video recording page, if the state of the target control was in a third state when the video editing page was exited, display the video editing page, and in response to an operation of switching the state of the target control from the third state to the second state, perform an audio correction process on the changed original audio and display the target control in the first state on the video editing page.
[0050] In some embodiments, the first display module is specifically used to display a video editing page including the target control in response to a trigger operation of accessing a video posting page after exiting a video editing page and returning to the video editing page after exiting the video posting page, and to control the state of the target control to remain consistent with the state of the target control when the video editing page was exited.
[0051] In some embodiments, the recording page of the target video further includes a second display module for displaying target presentation information based on attribute features of a predetermined reference audio and / or a user's recording behavior, and the audio in the target video is recorded based on the predetermined reference audio.
[0052] In some examples, the original audio in the target video meeting certain conditions includes that the original audio in the target video has a corresponding sound modification resource and the original audio was recorded while wearing wired earphones at all times.
[0053] The audio processing device according to the embodiments of the present disclosure can perform the steps performed by the client in the audio processing method according to the method embodiments of the present disclosure, and the performing steps and beneficial effects it comprises will not be further described here.
[0054] FIG. 7 is a structural schematic diagram of an electronic device according to an embodiment of the present disclosure. Hereinafter, specific reference will be made to FIG. 7 , which illustrates a structural schematic diagram suitable for implementing an electronic device 500 according to an embodiment of the present disclosure. The electronic device 500 according to an embodiment of the present disclosure may include, but is not limited to, mobile terminals such as mobile phones, laptops, digital broadcast receivers, personal digital assistants (PDAs), tablets, portable multimedia players (PMPs), in-vehicle terminals (e.g., in-vehicle navigation terminals), and wearable devices, as well as fixed terminals such as digital TVs, desktop computers, and smart homes. The electronic device illustrated in FIG. 7 is merely an example and does not impose any limitations on the functionality and scope of use of the embodiment of the present disclosure.
[0055] 7, electronic device 500 may include a processing unit (e.g., a central processing unit, a graphics processor, etc.) 501, which can perform various appropriate operations and processes according to a program stored in read-only memory (ROM) 502 or a program loaded from storage device 508 into random access memory (RAM) 503 to realize an audio processing method according to an embodiment of the present disclosure. RAM 503 further stores various programs and data necessary for the operation of electronic device 500. Processing unit 501, ROM 502, and RAM 503 are interconnected via bus 504. Input / output (I / O) interface 505 is also connected to bus 504.
[0056] Typically, input devices 506, including, for example, a touch screen, touch panel, keyboard, mouse, camera head, microphone, accelerometer, gyroscope, etc.; output devices 507, including, for example, a liquid crystal display (LCD), speaker, oscillator, etc.; storage devices 508, including, for example, a magnetic tape, hard disk, etc.; and communication devices 509 may be connected to the I / O interface 505. The communication devices 509 enable the electronic device 500 to communicate wirelessly or via wires with other devices to exchange data. While FIG. 7 illustrates the electronic device 500 with various devices, it should be understood that it is not intended to require the implementation or inclusion of all of the devices shown. More or fewer devices may alternatively be implemented or included.
[0057] In particular, according to embodiments of the present disclosure, the processes described above with reference to the flowcharts can be implemented as a computer software program. For example, embodiments of the present disclosure include a computer program product, which includes a computer program embodied in a non-transitory computer-readable medium, the computer program including program code for performing the methods shown in the flowcharts, thereby implementing the audio processing methods described above. In such embodiments, the computer program can be downloaded and installed from a network via the communication device 509, or installed from the storage device 508, or installed from the ROM 502. When executed by the processing device 501, the computer program performs the functions defined in the methods of the embodiments of the present disclosure.
[0058] It should be noted that the computer-readable medium in this disclosure may be a computer-readable signal medium, a computer-readable storage medium, or any combination thereof. The computer-readable storage medium may be, for example, but is not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination thereof. More specific examples of the computer-readable storage medium include, but are not limited to, an electrical connection having one or more conductors, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a compact disc read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination thereof. In this disclosure, the computer-readable storage medium may be any tangible medium that contains or stores a program, and the program can be used in or in combination with a command execution system, apparatus, or device. In this disclosure, the computer-readable signal medium may include a data signal propagated in baseband or as part of a carrier wave, carrying computer-readable program code therein. Such propagated data signals may take a variety of forms, including, but not limited to, electromagnetic signals, optical signals, or any suitable combination thereof. A computer-readable signal medium may also be any computer-readable medium other than a computer-readable storage medium, which is capable of transmitting, propagating, or transmitting a program for use in or in connection with a command execution system, apparatus, or device. The program code contained in the computer-readable medium may be transmitted over any suitable medium, including, but not limited to, electrical wire, optical cable, RF (radio frequency), or the like, or any suitable combination thereof.
[0059] In some embodiments, clients and servers may communicate using any now known or later developed network protocol, such as HTTP (HyperText Transfer Protocol), and may interconnect via any form or medium of digital data communication (e.g., a communication network). Examples of communication networks include a local area network ("LAN"), a wide area network ("WAN"), the Internet (e.g., internet), and an end-to-end network (e.g., an ad hoc end-to-end network), as well as any now known or later developed network.
[0060] The computer-readable medium may be included in the electronic device, or may be a standalone medium not mounted on the electronic device.
[0061] The computer-readable medium is equipped with one or more programs, and when the one or more programs are executed by the electronic device, the electronic device is caused to perform the following operations: display a video editing page for a target video in response to a trigger operation that accesses the video editing page for the first time; when the original audio in the target video satisfies a predetermined condition, perform sound correction processing on the original audio while displaying the video editing page, and display a target control in a first state on the video editing page to indicate that sound correction processing is being performed on the original audio; and when performing sound correction processing on the original audio, perform an editing operation corresponding to the editing control on the target video in response to a trigger operation acting on the editing control, and display at least one editing control on the video editing page.
[0062] In some embodiments, when the one or more programs are executed by the electronic device, the electronic device may perform other steps described in the above embodiments.
[0063] Computer program code for carrying out the operations of the present disclosure can be written using one or more programming languages, or a combination thereof, including, but not limited to, object-oriented programming languages such as Java, Smalltalk, and C++, as well as general procedural programming languages such as "C" or similar programming languages. The program code can execute entirely on the user's computer, partially on the user's computer, as a separate software package, partially on the user's computer and partially on a remote computer, or entirely on a remote computer or server. If a remote computer, the remote computer can be connected to the user's computer by any network, including a local area network (LAN) or a wide area network (WAN), or can be connected to an external computer (e.g., via the Internet using an Internet Service Provider).
[0064] The flowcharts and block diagrams in the accompanying drawings illustrate possible system architectures, functions, and operations of systems, methods, and computer program products according to various embodiments of the present disclosure. In this regard, each block in the flowcharts or block diagrams may represent a module, program segment, or portion of code, which includes one or more executable instructions for implementing a specified logical function. It should be noted that in some alternative implementations, the functions depicted in the blocks may be implemented in a different order than depicted in the drawings. For example, two consecutively shown blocks may be essentially executed in parallel, or, depending on the functionality, they may be executed in the reverse order. It should be noted that each block in the block diagrams and / or flowcharts, and combinations of blocks in the block diagrams and / or flowcharts, may be implemented by a dedicated hardware-based system that performs the specified functions or operations, or by a combination of dedicated hardware and computer instructions.
[0065] The units according to the embodiments of the present disclosure may be realized by software or hardware, and the names of the units may not necessarily limit the units themselves.
[0066] The functions described herein above may be performed, at least in part, by one or more hardware logic components. For example, exemplary hardware logic components that may be used include, but are not limited to, field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), systems on chips (SOCs), complex programmable logic devices (CPLDs), etc.
[0067] In the present disclosure, a machine-readable medium may be a tangible medium and may contain or store a program used by or in combination with a command execution system, apparatus, or device. The machine-readable medium may be a machine-readable signal medium or a machine-readable storage medium. The machine-readable medium may include, but is not limited to, an electronic, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any suitable combination of the above. More specific examples of machine-readable storage media include one or more wire-based electrical connections, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), optical fiber, a compact disc read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination thereof.
[0068] According to one or more embodiments of the present disclosure, the present disclosure provides an audio processing method, including: displaying a video editing page in response to a trigger operation that first accesses a video editing page for a target video; when original audio in the target video satisfies a predetermined condition, displaying the video editing page while simultaneously performing sound correction processing on the original audio and displaying a target control in a first state on the video editing page to indicate that sound correction processing is being performed on the original audio; and when performing sound correction processing on the original audio, in response to a trigger operation acting on an editing control, performing an editing operation corresponding to the editing control on the target video and displaying at least one editing control on the video editing page.
[0069] According to one or more embodiments of the present disclosure, the audio processing method of the present disclosure further includes playing the original audio while displaying the video editing page, and when completion of the sound modification process is detected, continuing to play the modified audio from the playback progress position of the original audio at the time of completion of the sound modification process.
[0070] According to one or more embodiments of the present disclosure, in an audio processing method according to the present disclosure, when completion of the sound modification process is detected, the state of the target control is controlled to switch to a second state, and the target control in the second state is used to indicate that the sound modification process on the original audio has been completed.
[0071] According to one or more embodiments of the present disclosure, the audio processing method according to the present disclosure further includes continuing to play the original audio from a current playback progress position in response to an operation of switching a state of the target control from the second state to a third state; When the target control is in the second state, the audio after the sound modification process is played on the video editing page.
[0072] According to one or more embodiments of the present disclosure, the audio processing method of the present disclosure further includes, in response to a trigger operation acting on a posting control, posting a target video including audio after sound modification processing if the target control is in the second state, and posting a target video including the original audio if the target control is in the third state.
[0073] According to one or more embodiments of the present disclosure, the audio processing method of the present disclosure further includes, in response to an operation of exiting the video editing page, accessing a specified page, and then returning to the video editing page from the specified page, displaying the video editing page and controlling the state of the target control to remain consistent with the state of the target control when the video editing page was exited.
[0074] According to one or more embodiments of the present disclosure, in the audio processing method according to the present disclosure, displaying the video editing page including the target control in response to an operation of returning to the video editing page after exiting the video editing page as described above includes: In response to a trigger operation of accessing a video recording page after exiting a video editing page and returning to the video editing page after exiting the video recording page, if the state of the target control is in a first state or a second state when the video editing page is exited, displaying the video editing page and simultaneously performing a sound correction process on the original audio after the change, and displaying the target control in the first state on the video editing page; or accessing a video recording page after exiting the video editing page and, in response to a trigger operation of returning to the video editing page after exiting the video recording page, displaying the video editing page if the state of the target control was in a third state when the video editing page was exited; In response to an operation of switching the state of the target control from the third state to the second state, a sound correction process is performed on the original audio after the change, and the target control in the first state is displayed on the video editing page.
[0075] According to one or more embodiments of the present disclosure, in the audio processing method according to the present disclosure, displaying the video editing page including the target control in response to an operation of returning to the video editing page after exiting the video editing page as described above includes: The method includes displaying a video editing page including the target control in response to a trigger operation of accessing a video posting page after exiting the video editing page and returning to the video editing page after exiting the video posting page, and controlling the state of the target control to be consistent with the state of the target control when the user exited the video editing page.
[0076] According to one or more embodiments of the present disclosure, the audio processing method of the present disclosure further includes displaying target presentation information on the recording page of the target video based on attribute features of a predetermined reference audio and / or the user's recording behavior, and the audio in the target video is recorded based on the predetermined reference audio.
[0077] According to one or more embodiments of the present disclosure, in the audio processing method of the present disclosure, the original audio in the target video satisfies a predetermined condition by: The original audio in the target video includes a corresponding sound modification resource, and the original audio is recorded by constantly wearing wired earphones.
[0078] According to one or more embodiments of the present disclosure, the present disclosure provides a method for manufacturing a semiconductor device, comprising: a first display module for displaying the video editing page in response to a trigger operation for first accessing the video editing page for the target video; a processing module for displaying the video editing page and simultaneously performing sound modification processing on the original audio when the original audio in the target video satisfies a predetermined condition; the first display module is further used to display a target control in a first state on the video editing page, the target control in the first state being for indicating that a sound modification process is being performed on the original audio; An audio processing device is provided, which includes an editing module for executing an editing operation corresponding to an editing control on the target video in response to a trigger operation acting on the editing control when performing sound modification processing on the original audio, the editing module displaying at least one editing control on the video editing page.
[0079] According to one or more embodiments of the present disclosure, the audio processing device of the present disclosure further includes a playback module that plays the original audio while displaying the video editing page, and when completion of the sound modification process is detected, continues playing the modified audio from the playback progress position of the original audio at the time of completion of the sound modification process.
[0080] According to one or more embodiments of the present disclosure, an audio processing device according to the present disclosure further includes a control module that controls the state of the target control to switch to a second state when completion of the sound modification process is detected, and the target control in the second state is used to indicate that the sound modification process on the original audio has been completed.
[0081] According to one or more embodiments of the present disclosure, in the audio processing device of the present disclosure, the playback module further continues to play the original audio from the current playback progress position in response to an operation of switching the state of the target control from the second state to a third state, wherein when the target control is in the second state, it is used to play the audio after the sound modification process on the video editing page.
[0082] According to one or more embodiments of the present disclosure, an audio processing device according to the present disclosure further includes a posting module for posting a target video including audio after sound modification processing when the target control is in the second state, and posting a target video including the original audio when the target control is in the third state, in response to a trigger operation acting on a posting control.
[0083] According to one or more embodiments of the present disclosure, in the audio processing device of the present disclosure, the first display module is further used to display the video editing page in response to an operation of exiting the video editing page, accessing a specified page, and then returning to the video editing page from the specified page, and to control the state of the target control to remain consistent with the state of the target control when the video editing page was exited.
[0084] According to one or more embodiments of the present disclosure, in the audio processing device according to the present disclosure, the first display module is specifically used to: access a video recording page after exiting from a video editing page, and in response to a trigger operation of returning to the video editing page after exiting from the video recording page, if the state of the target control when the video editing page was in a first state or a second state when the video editing page was exited, display the video editing page while performing sound correction processing on the changed original audio and display the target control in the first state on the video editing page; or to: access a video recording page after exiting from the video editing page, and in response to a trigger operation of returning to the video editing page after exiting from the video recording page, if the state of the target control when the video editing page was in a third state when the video editing page was exited, display the video editing page, and in response to an operation of switching the state of the target control from the third state to the second state, perform sound correction processing on the changed original audio and display the target control in the first state on the video editing page.
[0085] According to one or more embodiments of the present disclosure, in the audio processing device of the present disclosure, the first display module is specifically used in that there is an audio modification resource corresponding to the original audio in the target video, and the original audio is recorded by wearing wired earphones at all times.
[0086] According to one or more embodiments of the present disclosure, in the audio processing device of the present disclosure, the recording page of the target video further includes a second display module for displaying target presentation information based on attribute features of a predetermined reference audio and / or a user's recording behavior, and the audio in the target video is recorded based on the predetermined reference audio.
[0087] According to one or more embodiments of the present disclosure, in an audio processing device according to the present disclosure, the original audio in the target video satisfying a predetermined condition includes that a sound modification resource corresponding to the original audio in the target video exists, and the original audio is recorded by wearing wired earphones at all times.
[0088] According to one or more embodiments of the present disclosure, the present disclosure provides a method for manufacturing a semiconductor device, comprising: one or more processors; a memory for storing one or more programs; The present invention also provides an electronic device that, when the one or more programs are executed by the one or more processors, causes the one or more processors to implement any of the audio processing methods according to the present disclosure.
[0089] According to one or more embodiments of the present disclosure, the present disclosure provides a computer-readable storage medium having stored thereon a computer program that, when executed by a processor, causes any of the audio processing methods according to the present disclosure to be implemented.
[0090] An embodiment of the present disclosure further provides a computer program product including a computer program or instructions that, when executed by a processor, causes the above audio processing method to be implemented.
[0091] The above is merely a description of preferred embodiments and the technical principles applied in the present disclosure. It is obvious to those skilled in the art that the scope of the present disclosure is not limited to the technical solution based on the specific combination of the above technical features, but should also include other technical solutions formed by any combination of the above technical features or equivalent features without departing from the concept of the above disclosure. For example, it also includes technical solutions formed by mutually replacing the above features with technical features having similar functions disclosed in the present disclosure (including but not limited to those).
[0092] Also, although operations are described in a particular order, this should not be understood as requiring such operations to be performed in the particular order shown, or in a sequential order. In certain environments, multitasking and parallel processing may be advantageous. Similarly, although certain specific implementation details are included in the above description, they should not be construed as limiting the scope of the present disclosure. Certain features that are described in the context of a single embodiment can also be implemented in a single embodiment in combination. Conversely, various features that are described in the context of a single embodiment can also be implemented in multiple embodiments separately or in any suitable subcombination.
[0093] Although the present subject matter has been described in language specific to structural features and / or methodological operations, it is to be understood that the subject matter defined in the appended claims is not necessarily limited to the specific features or operations described above. Rather, the specific features and operations described above are merely example forms of implementing the claims.
Claims
1. displaying a video editing page in response to a trigger operation that first accesses the video editing page for the target video; If there is a sound modification resource corresponding to the original audio in the target video and the audio effect of the original audio satisfies a condition, the video editing page is displayed while sound modification processing is performed on the original audio, and a target control in a first state is displayed on the video editing page, where the target control in the first state is for indicating that sound modification processing is being performed on the original audio; When performing sound modification processing on the original audio, in response to a trigger operation acting on an edit control, execute an edit operation corresponding to the edit control on the target video, wherein at least one edit control is displayed on the video editing page; switching the state of the target control to a second state in response to completion of the sound modification process; switching a state of the target control from the second state to a third state in response to a trigger operation on the target control; and, in response to a trigger operation acting on a posting control, if the target control is in the second state, posting a target video including the audio after sound modification processing, and if the target control is in the third state, posting a target video including the original audio. Audio processing methods.
2. displaying the video editing page and simultaneously playing the original audio; 2. The method of claim 1, further comprising upon detecting completion of the sound modification process, continuing to play the sound-modified audio from a position in the playback progress of the original audio at the time the sound modification process was completed.
3. The method described in claim 1, wherein the target control in the second state is used to indicate that sound modification processing on the original audio has been completed.
4. and continuing to play the original audio from a current playback progress position in response to an operation of switching the state of the target control from the second state to a third state; The method of claim 3 , wherein when the target control is in the second state, the video editing page plays audio after sound modification processing.
5. 2. The method of claim 1, further comprising: displaying a video editing page including the target control in response to an operation of returning to the video editing page after exiting the video editing page; and controlling the state of the target control to remain consistent with the state of the target control when the video editing page was exited if the original audio has not changed.
6. Displaying the video editing page including the target control in response to an operation of returning to the video editing page after exiting the video editing page, as described above, In response to a trigger operation of accessing a video recording page after exiting a video editing page and returning to the video editing page after exiting the video recording page, if the state of the target control is in a first state or a second state when the video editing page is exited, displaying the video editing page while simultaneously performing sound correction processing on the original audio after the change, and displaying the target control in the first state on the video editing page; or accessing a video recording page after exiting the video editing page, and in response to a trigger operation of returning to the video editing page after exiting the video recording page, if the state of the target control was in a third state when the video editing page was exited, displaying the video editing page; The method of claim 5, further comprising: in response to an operation of switching the state of the target control from the third state to the second state, performing sound correction processing on the original audio after the change; and displaying the target control in the first state on the video editing page.
7. Displaying the video editing page including the target control in response to an operation of returning to the video editing page after exiting the video editing page, as described above, 6. The method of claim 5, further comprising: in response to a trigger operation of accessing a video posting page after exiting a video editing page and returning to the video editing page after exiting the video posting page, displaying a video editing page including the target control, and controlling the state of the target control to remain consistent with the state of the target control when the video editing page was exited.
8. The method of claim 1, further comprising displaying target presentation information on the recording page of the target video based on attribute features of a specified reference audio and / or the user's recording behavior, wherein the audio in the target video is recorded based on the specified reference audio.
9. The method described in claim 1, wherein the original audio is recorded while wearing wired earphones at all times so that the audio effect of the original audio satisfies a condition.
10. a first display module for displaying the video editing page in response to a trigger operation for first accessing the video editing page for the target video; a processing module for displaying the video editing page and performing sound modification processing on the original audio when the original audio in the target video has a corresponding sound modification resource and the audio effect of the original audio satisfies a condition; the first display module is further used to display a target control in a first state on the video editing page, the target control in the first state being for indicating that a sound modification process is being performed on the original audio; an editing module for executing an editing operation corresponding to the editing control on the target video in response to a trigger operation acting on the editing control when performing a sound modification process on the original audio, the editing module displaying at least one editing control on the video editing page; a control module for switching a state of the target control to a second state in response to completion of the sound modification process, and further for switching a state of the target control from the second state to a third state in response to a trigger operation on the target control; and a posting module that, in response to a trigger operation acting on a posting control, posts a target video including audio after sound modification processing if the target control is in the second state, and posts a target video including the original audio if the target control is in the third state.
11. one or more processors; a storage device for storing one or more programs; An electronic device, wherein the one or more programs, when executed by the one or more processors, cause the one or more processors to implement the method of any one of claims 1 to 9.
12. A computer-readable storage medium having stored thereon a computer program that, when executed by a processor, implements the method according to any one of claims 1 to 9.
Citation Information
Patent Citations
Song recording method, tone modifying method and electronic device
CN110010162A
Media-Editing Application with Media Clips Grouping Capabilities
US20120210231A1
Song recording method, sound correction method and electronic device
WO2020173391A1