Audio sharing method, device, equipment and medium

The audio sharing method generates a preset playback interface with target audio as background music, addressing limitations in current sharing methods by providing diverse and convenient audio sharing through electronic devices, enhancing user experience.

JP7732001B2Active Publication Date: 2025-09-01BEIJING ZITIAO NETWORK TECH CO LTD
View PDF 7 Cites 0 Cited by

Patent Information

Application Number
JP2023574227
Authority / Receiving Office
JP · JP
Patent Type
Patents
Current Assignee / Owner
Priority Date
2021-06-02
Filing Date
2022-05-30
Publication Date
2025-09-01
Estimated Expiration
2042-05-30

Smart Images

  • Figure 0007732001000001
    Figure 0007732001000001
  • Figure 0007732001000002
    Figure 0007732001000002
  • Figure 0007732001000003
    Figure 0007732001000003
Patent Text Reader

Abstract

The present disclosure relates to an audio sharing method, device, equipment, and medium. The audio sharing method includes displaying a target object including an original video with a target audio waiting to be shared as background music, which is a published audio, and / or a sharing control of the target audio on a target interaction interface, and displaying a preset playback interface including a visualization material generated based on the target audio and displaying a target video with the target audio as background music, for sharing the target audio, when a first trigger operation on the target object is detected. According to the embodiment of the present disclosure, it is possible to meet individual requirements of users and lower the barrier to video creation.
Need to check novelty before this filing date? Find Prior Art

Description

[Technical Field]

[0001] [CROSS-REFERENCE TO RELATED APPLICATIONS] This application is based on and claims priority from Chinese Patent Application No. 202110615705.4 filed on June 2, 2021. The entire text of the underlying Chinese patent application is incorporated herein by reference.

[0002] [Technical field] The present disclosure relates to the field of audio processing technology, and in particular to an audio sharing method, device, apparatus and medium. [Background technology]

[0003] 2. Description of the Related Art With the rapid development of computer technology and mobile communication technology, various video platforms via electronic devices have become widely used, greatly enriching people's daily lives.

[0004] Currently, when users listen to audio content that interests them on a video platform, they have the intention to share the audio content, but currently there is only one way to share the content. Summary of the Invention

[0005] To solve or at least partially solve the above technical problems, the present disclosure provides an audio sharing method, device, apparatus, and medium.

[0006] According to a first aspect, the present disclosure provides a method for manufacturing a semiconductor device, comprising: Displaying a target object including an original video with the target audio waiting to be shared, which is the published audio, as background music and / or a sharing control of the target audio in a target interaction interface; When a first trigger operation on a target object is detected, a preset playback interface for sharing the target audio is displayed, the preset playback interface including visualization material generated based on the target audio and displaying a target video with the target audio as background music.

[0007] According to a second aspect, the present disclosure provides a method for manufacturing a semiconductor device, comprising: a first display means configured to display a target object including an original video with the target audio waiting to be shared, which is the published audio, as background music and / or a shared control of the target audio in a target interaction interface; and a second display means configured to display a preset playback interface for sharing the target audio, which includes visualization material generated based on the target audio and displays a target video with the target audio as background music, when a first trigger operation on a target object is detected.

[0008] According to a third aspect, the present disclosure provides a method for manufacturing a semiconductor device comprising: a processor; a memory for storing executable instructions; An electronic device is provided in which a processor is configured to read executable instructions from a memory and execute the executable instructions to implement the audio sharing method according to the first aspect.

[0009] According to a fourth aspect, the present disclosure provides a computer-readable storage medium having stored thereon a computer program that, when executed by a processor, causes the processor to implement the audio sharing method described in the first aspect.

[0010] No. 5 According to another aspect, the present disclosure provides a computer program comprising instructions that, when executed by a processor, cause the computer program to implement the audio sharing method described in the first aspect.

[0011] No. 6 According to another aspect, the present disclosure provides a computer program product including a computer program or instructions that, when executed by a processor, causes the audio sharing method according to the first aspect to be implemented.

[0012] The technical solution according to the embodiments of the present disclosure has the following advantages over the related art:

[0013] According to the audio sharing method, device, equipment, and medium of the embodiments of the present disclosure, when a user triggers sharing of a target audio, a preset playback interface can be directly displayed, displaying a target video automatically generated from the target audio, with the target video using the target audio as background music. Sharing the target audio using the target video not only diversifies the shared content and meets the individual needs of users, but also lowers the barrier to video creation, and allows users to conveniently share audio with video content without having to shoot or upload videos. [Brief explanation of the drawings]

[0014] These and other features, advantages, and aspects of the embodiments of the present disclosure will become more apparent from the following detailed description, taken in conjunction with the accompanying drawings, in which the same or similar elements are designated by the same or similar reference numerals throughout the drawings. It should be understood that the drawings are schematic and that components and elements are not necessarily drawn to scale. [Figure 1] 1 is a schematic flow chart illustrating an audio sharing method according to an embodiment of the present disclosure. [Figure 2] FIG. 10 is a schematic diagram illustrating interactions that trigger audio sharing according to an embodiment of the present disclosure. [Figure 3] FIG. 10 is a schematic diagram illustrating another interaction that triggers audio sharing according to an embodiment of the present disclosure. [Figure 4]FIG. 1 is an interface schematic diagram illustrating a preset playback interface according to an embodiment of the present disclosure. [Figure 5] FIG. 10 is an interface schematic diagram illustrating another preset playback interface according to an embodiment of the present disclosure. [Figure 6] FIG. 10 is an interface schematic diagram illustrating yet another preset playback interface according to an embodiment of the present disclosure. [Figure 7] FIG. 1 is a schematic diagram illustrating video cropping interactions according to an embodiment of the present disclosure. [Figure 8] FIG. 1 is a schematic diagram illustrating material modification interactions according to an embodiment of the present disclosure. [Figure 9] FIG. 10 is a schematic diagram illustrating another material modification interaction according to an embodiment of the present disclosure. [Figure 10] FIG. 10 is a schematic diagram illustrating yet another material modification interaction according to an embodiment of the present disclosure. [Figure 11] 1 is a schematic flow chart illustrating another audio sharing method according to an embodiment of the present disclosure. [Figure 12] FIG. 1 is a schematic diagram illustrating a playback interface of a target video according to an embodiment of the present disclosure. [Figure 13] 1 is a schematic diagram illustrating the configuration of an audio sharing device according to an embodiment of the present disclosure. [Figure 14] FIG. 1 is a schematic diagram illustrating the configuration of an electronic device according to an embodiment of the present disclosure. DETAILED DESCRIPTION OF THE INVENTION

[0015] The following describes in more detail the embodiments of the present disclosure with reference to the accompanying drawings. Although several embodiments of the present disclosure are shown in the accompanying drawings, it should be understood that the present disclosure can be realized in various forms and should not be construed as being limited to the embodiments described herein. On the contrary, these embodiments are provided to provide a more thorough and complete understanding of the present disclosure. It should be understood that the accompanying drawings and embodiments of the present disclosure are merely illustrative and do not limit the scope of protection of the present disclosure.

[0016] It should be understood that the steps described in the method embodiments of the present disclosure may be performed in a different order and / or in parallel. Furthermore, method embodiments may include additional steps and / or omit performing steps as shown. The scope of the present disclosure is not limited in this respect.

[0017] As used herein, the term "comprises" and variations thereof are open inclusive, i.e., meaning "including, but not limited to." The term "based on" means "based at least in part on." The term "in one embodiment" means "at least one embodiment," the term "another embodiment" means "at least one other embodiment," and the term "some embodiments" means "at least some embodiments." Relevant definitions of other terms are provided below.

[0018] It should be noted that the concepts of "first," "second," etc. described in this disclosure are merely intended to distinguish between different devices, modules, or units, and are not intended to limit the order or interdependence of functions performed by these devices, modules, or units.

[0019] It should be noted that the modifications of "one" and "multiple" described in the present disclosure are general and not limiting. It is obvious to those skilled in the art that unless otherwise specified, it should be understood as "one or multiple."

[0020] The names of messages or information exchanged between devices in the embodiments of the present disclosure are merely descriptive and do not limit the scope of these messages or information.

[0021] Currently, when a user listens to audio content that interests him on an online video platform and wants to share the audio content, he generally achieves this by directly transferring the video content or directly sending music to other users via private messages, which means that the content sharing method is single and cannot meet the individual needs of users.

[0022] To solve the above problems, the embodiments of the present disclosure provide an audio sharing method, device, equipment, and medium that can intelligently generate videos and share music.

[0023] Hereinafter, an audio sharing method according to an embodiment of the present disclosure will be described first with reference to FIGS.

[0024] In an embodiment of the present disclosure, the audio sharing method may be performed by an electronic device, which may include, but is not limited to, mobile terminals such as mobile phones, laptops, digital broadcast receivers, PDAs (personal digital assistants), PADs (tablets), PMPs (portable multimedia players), in-vehicle terminals (e.g., in-vehicle navigation terminals), and wearable devices, and fixed terminals such as digital TVs, desktop computers, and smart homes.

[0025] FIG. 1 is a schematic flow chart illustrating an audio sharing method according to an embodiment of the present disclosure.

[0026] As shown in FIG. 1, the audio sharing method may include the following steps:

[0027] S110: A target object including an original video with the target audio waiting to be shared as background music and / or a shared control of the target audio is displayed on a target interaction interface.

[0028] In an embodiment of the present disclosure, the target interaction interface may be an interface for exchanging information between a user and an electronic device, and may provide information to a user, for example, display a target object, and may also accept information or operations input by a user, for example, accept an operation by a user on a target object.

[0029] In the embodiment of the present disclosure, the target audio may be recorded audio recorded by the presenter of the original video, or may be music audio, and is not limited thereto.

[0030] In some embodiments, the target object may include an original video with the target audio waiting to be shared as background music.

[0031] Here, the original video may be a public video that the user has already distributed to other users via a server, or a private video that the user has published and stored on a server. The original video may also be a public video that the user has viewed and that other users have distributed via a server.

[0032] In some other embodiments, the target object may include shared control of target audio.

[0033] Here, the sharing control of the target audio may be a control for triggering sharing of the target audio. Specifically, the control may be an object that can be triggered by a user, such as a button or an icon.

[0034] S120: When a first trigger operation on a target object is detected, a preset playback interface for sharing the target audio is displayed, which includes visualization material generated based on the target audio and displays a target video with the target audio as background music.

[0035] In an embodiment of the present disclosure, when a user wants to share target audio related to a target object, the user may input a first trigger operation for the target object into the electronic device. Here, the first trigger operation may be an operation for triggering sharing of the target audio related to the target object. When the electronic device detects the first trigger operation for the target object, the electronic device may display a preset playback interface including a target video with the target audio as background music. Here, since the background music of the target video is the target audio, the target video is used to share the target audio. That is, the user can share the target audio by sharing the target video.

[0036] In an embodiment of the present disclosure, the target video may be automatically generated from the target audio.

[0037] Additionally, the target video may include visual materials automatically generated from the target audio, where the visual materials are visible video elements in the target video.

[0038] Optionally, the visualization material may include images and / or text generated based on relevant information of the target audio, as will be described in more detail below.

[0039] In the embodiments of the present disclosure, when a user triggers audio sharing for a target audio, a preset playback interface can be directly displayed, displaying a target video automatically generated from the target audio, with the target video using the target audio as background music. Sharing the target audio using the target video not only diversifies the content to be shared and meets the individual needs of users, but also lowers the barrier to video creation, allowing users to conveniently share audio with video content without having to shoot or upload videos.

[0040] In another embodiment of the present disclosure, in order to improve the convenience of users in sharing target audio, multiple trigger methods for sharing target audio may be provided to users.

[0041] In some embodiments of the present disclosure, the target interaction interface may include an audio exhibit interface, and the target object may include shared control of the target audio.

[0042] Here, the audio display interface may be an interface for displaying information related to the target audio, for example, an interface of a music detail page of the target audio.

[0043] Optionally, the related information of the target audio may include at least one of an audio jacket image of the target audio, a presenter icon of the target audio, audio style information of the target audio, lyrics of the target audio, performer names of the target audio, audio name of the target audio, and a video jacket image of a released video whose background music is the target audio.

[0044] Furthermore, a target object may be displayed in the audio display interface of the target audio, where the target object may be a sharing control for the target audio. For example, the sharing control for the target audio may be a sharing button for the target audio or a sharing icon for the target audio.

[0045] In these embodiments, the first trigger operation may optionally include a gesture control operation (e.g., tap, long press, double tap, etc.), a voice control operation, or a facial expression control operation on the shared control of the target audio.

[0046] FIG. 2 is a schematic diagram illustrating interactions that trigger audio sharing according to an embodiment of the present disclosure.

[0047] As shown in FIG. 2 , electronic device 201 displays interface 202 of a music detail page for a song titled "Hula Dance XXX" sung by person A. Interface 202 of the music detail page displays audio jacket image 203 of "Hula Dance XXX," audio name "Hula Dance XXX," performer name "Mr. A," and video jacket image 204 of a publicly released video using "Hula Dance XXX" as background music. Interface 202 of the music detail page also displays a "Share" button 205, which is a share button for "Hula Dance XXX." A user can share the song "Hula Dance XXX" by tapping "Share" button 205 in interface 202 of the music detail page. At this time, electronic device 201 can automatically generate a video using "Hula Dance XXX" as background music.

[0048] In these embodiments, optionally, before S110, the audio sharing method includes: The method may further include displaying an audio display control of the original video with the target audio as background music on the video playback interface and the target audio.

[0049] Accordingly, S110 specifically: When a second trigger operation on the audio exhibit control is detected, the method may include displaying an audio exhibit interface in which the shared control for the target audio is displayed.

[0050] Specifically, the electronic device may play the original video in a video playback interface and display an audio exhibition control for the target audio, which is set as background music for the original video, in the video playback interface. If the user is interested in the target audio, the user may input a second trigger operation on the audio exhibition control for the target audio into the electronic device. Here, the second trigger operation may be an operation for triggering access to the audio exhibition interface for the target audio. When the electronic device detects the second trigger operation on the audio exhibition control, the electronic device may display an audio exhibition interface including a sharing control for the target audio. This allows the user to trigger sharing of the target audio by inputting a first trigger operation on the sharing control for the target audio in the audio exhibition interface.

[0051] Here, the audio exhibit control may be used to display at least one of the lyrics of the target audio, the performer name of the target audio, and the audio name of the target audio.

[0052] In some other embodiments of the present disclosure, the target interaction interface may include a video playback interface, the target object may include an original video, and the target object may include an original video with the target audio as background music.

[0053] Here, the video playback interface may be an interface for playing the original video, and a target object may be displayed in the video playback interface of the original video, where the target object may include the original video.

[0054] In these embodiments, the first trigger operation may include a trigger operation on the original video.

[0055] In some embodiments, the first trigger operation may include a function pop-up window trigger operation for triggering the popup of a function pop-up window, such as a gesture control operation (e.g., tap, long press, double tap, etc.), a voice control operation, or a facial expression control operation on the original video, and a share trigger operation for triggering the sharing of the target audio, such as a gesture control operation (e.g., tap, long press, double tap, etc.), a voice control operation, or a facial expression control operation on a share button in the function pop-up window. That is, the user needs to first trigger the display of the function pop-up window in the video playback interface, and then trigger the sharing of the target audio based on the share button in the function pop-up window.

[0056] In some other embodiments, the first trigger operation may include a sharing trigger operation for triggering sharing of the target audio, such as a gesture control operation (e.g., tap, long press, double tap, etc.), a voice control operation, or a facial expression control operation on the original video, that is, a user can directly trigger sharing of the target audio in the video playback interface.

[0057] In yet some further embodiments of the present disclosure, the target interaction interface may include a video playback interface, the target object may include an original video, and the original video may include audio controls for the target audio, e.g., an audio sticker.

[0058] Here, the video playback interface may be an interface for playing the original video. Furthermore, a target object may be displayed on the video playback interface of the original video. In this case, the target object may include the original video. Furthermore, an audio sticker of the target audio may be displayed on the video screen of the original video played on the video playback interface.

[0059] Here, the audio sticker may be used to display at least one of the lyrics of the target audio, the performer name of the target audio, and the audio name of the target audio.

[0060] In these examples, the first triggering operation may include a triggering operation for an audio sticker.

[0061] In some embodiments, the first trigger operation may include a function pop-up window trigger operation for triggering the popup display of a function pop-up window, such as a gesture control operation (e.g., tap, long press, double tap, etc.), a voice control operation, or a facial expression control operation on an audio sticker, and a share trigger operation for triggering the sharing of the target audio, such as a gesture control operation (e.g., tap, long press, double tap, etc.), a voice control operation, or a facial expression control operation on a share button in the function pop-up window. That is, a user needs to first trigger the display of a function pop-up window in the video playback interface, and then trigger the sharing of the target audio based on the share button in the function pop-up window.

[0062] FIG. 3 is a schematic diagram illustrating another interaction that triggers audio sharing according to an embodiment of the present disclosure.

[0063] As shown in FIG. 3 , electronic device 301 displays video playback interface 302 for a video posted by person B, and the background music for the video is the song "Hula Dance XXX." If a user is interested in the song "Hula Dance XXX," they may long press the song control 303 for the song "Hula Dance XXX." At this time, electronic device 301 may display function pop-up window 304, and a "Share" button may be displayed in function pop-up window 304. The "Share" button is a share button for "Hula Dance XXX." The user can share the song "Hula Dance XXX" by tapping the "Share" button. At this time, electronic device 301 can automatically generate a video with "Hula Dance XXX" as background music.

[0064] In some other embodiments, the first trigger operation may include a share trigger operation for triggering sharing of the target audio, such as a gesture control operation (e.g., tap, long press, double tap, etc.) on the audio sticker, a voice control operation, or a facial expression control operation, etc. That is, a user can directly trigger sharing of the target audio in the video playback interface.

[0065] In still other embodiments of the present disclosure, the target interaction interface may include a video playback interface, and the target object may include shared control of the original video and the target audio.

[0066] Here, the video playback interface may be an interface for playing the original video. Furthermore, the video playback interface of the original video may display the original video and a target object. In this case, the target object may include a shared control for the target audio. Furthermore, the shared control for the target audio may be displayed at any position within the video playback interface.

[0067] In these examples, the first trigger operation may include a trigger operation on a shared control.

[0068] Specifically, the first trigger operation may include a sharing trigger operation for triggering sharing of the target audio, such as a gesture control operation (e.g., tap, long press, double tap, etc.), a voice control operation, or a facial expression control operation on the sharing control of the target audio, that is, the user can directly trigger sharing of the target audio in the video playback interface.

[0069] In yet another embodiment of the present disclosure, the visualization material may include images and / or text generated based on relevant information of the target audio.

[0070] Here, the related information of the target audio may include at least one of an audio jacket image of the target audio, a presenter icon of the target audio, audio style information of the target audio, audio content information of the target audio, lyrics of the target audio, performer names of the target audio, audio name of the target audio, and musical characteristics of the target audio.

[0071] In some embodiments of the present disclosure, the visualization material may include a first visualization material. In this case, the first visualization material may include an image generated based on related information. The related information may include a related image of the target audio. The related image may include at least one of an audio jacket image and a presenter icon.

[0072] In some embodiments, if the target audio is song audio, the associated images may include audio jacket images.

[0073] In some other implementations, if the target audio is recorded audio recorded by a presenting user of the original video, the associated image may include a presenter icon.

[0074] Optionally, the first visualization material may include at least one of the following:

[0075] 1. Background material: The background material includes images generated based on the image features of related images. Here, the background material is a dynamic or static background image that has the image characteristics of the related image.

[0076] Optionally, the image features may include at least one of a color feature, a brightness feature, and a saturation feature.

[0077] In some examples, for example, when the image features include color features, when the electronic device detects a first trigger operation on the target object, it may first extract the color with the most pixels appearing or the color with the number of pixels appearing exceeding a predetermined threshold value from the related image of the target audio, then select a color that falls within a predetermined color gamut from the extracted colors, and generate a background image with a solid color or a gradient color corresponding to the background material based on the selected color, and further generate a target video based on the background material.

[0078] The preset number threshold and the preset color gamut may be set according to the needs of the user, and are not limited here.

[0079] In some other examples, taking the case where the image features include saturation features as an example, when the electronic device detects a first trigger operation on the target object, it performs saturation detection on the related images of the target audio and determines that the related images only contain Morandi colors, that is, all of the related images are determined to be gray colors with low saturation, and the electronic device may generate a background image of a solid color or a gradient color corresponding to the background material based on the colors of the Morandi colors, and further generate a target video based on the background material.

[0080] Additionally, background material is displayed in the first screen area of ​​the target video.

[0081] Optionally, the first screen area of ​​the target video may be at least a partial screen area of ​​the target video, and is not limited herein.

[0082] When the first screen area is the entire screen area of ​​the target video, the background material may be a background image that covers the entire screen area of ​​the target video.

[0083] When the first screen area is a partial screen area of ​​the target video, the background material may be a background image that covers the partial screen area of ​​the target video.

[0084] 2. Foreground Material: Foreground material includes related images or images generated based on image features of related images. Here, foreground material is a dynamic or static foreground image that sits on top of background material.

[0085] In some embodiments, the foreground material may be an associated image of the target audio.

[0086] Specifically, when the electronic device detects a first trigger operation on a target object, the electronic device may directly set the related image of the target audio as a foreground image corresponding to the foreground material, and further generate a target video based on the foreground material.

[0087] In some other embodiments, the foreground material may be a foreground image generated based on image features of the associated image of the target audio.

[0088] Optionally, the image features may include at least one of a color feature, a brightness feature, and a saturation feature.

[0089] In some examples, for example, when the image features include color features, when the electronic device detects a first trigger operation on the target object, it may first extract a color with the most pixels appearing or a color with a number of pixels appearing that exceeds a predetermined threshold value from the related image of the target audio, then select a color that falls within a predetermined color gamut from the extracted colors, and generate a foreground image with gradation colors corresponding to the foreground material based on the selected color, and further generate a target video based on the foreground material.

[0090] The preset number threshold and the preset color gamut may be set according to the needs of the user, and are not limited here.

[0091] Additionally, foreground material is displayed in a second screen area of ​​the target video.

[0092] Optionally, the second screen area of ​​the target video may be at least a partial screen area of ​​the target video, and is not limited herein.

[0093] If the second screen area is the full screen area of ​​the target video, the foreground material may be a foreground image that covers the full screen area of ​​the target video.

[0094] If the second screen area is a partial screen area of ​​the target video, the foreground material may be a foreground image that covers the partial screen area of ​​the target video.

[0095] In the embodiments of the present disclosure, the electronic device may generate a target video based on only the foreground material or the background material, or may generate a target video based on the foreground material and the background material.

[0096] When the electronic device generates a target video based only on foreground material, the first screen area may be a background display area of ​​the target video, at least a portion of the second screen area may be included within the first screen area, and the target video may further include canvas-style content formed from a foreground image corresponding to the foreground material.

[0097] When the electronic device generates a target video based on foreground material and background material, at least a portion of the second screen area may be included within the first screen area, so that at least a portion of the foreground material can cover the background material, and the target video includes canvas-style content consisting of a foreground image corresponding to the foreground material and a background image corresponding to the background material.

[0098] FIG. 4 is an interface schematic diagram illustrating a preset playback interface according to an embodiment of the present disclosure.

[0099] 4, a preset playback interface is displayed on electronic device 401, and a target video automatically generated based on the song "Hula Dance XXX" may be displayed on the preset playback interface. Here, the target video displays a foreground picture and a background picture, where the foreground picture is a song jacket image 402 of the song "Hula Dance XXX," and the background picture is a solid-color background image 403 generated based on the color features of song jacket image 402.

[0100] In some other embodiments of the present disclosure, the visualizable material may include a second visualizable material, and the second visualizable material may include a background image that matches related information selected from a material library, where the related information may include an audio tag for the target audio.

[0101] Here, the second visualization material may be an image that matches the audio tag of the target audio.

[0102] Optionally, the second visualization material may include a static or dynamic background image selected from a material library that matches the audio tag of the target audio.

[0103] Specifically, when detecting a first trigger operation on the target object, the electronic device may acquire an audio tag for the target audio. The audio tag may be a tag representing an audio style, an audio type, or a jacket image type. Images pre-stored in a material library are then searched for images having the audio tag, and at least one of the searched images is randomly selected as a background image corresponding to the second visualized material. Furthermore, a target video is generated based on the second visualized material so that the background image corresponding to the second visualized material in the target video matches the style, such as the atmosphere, type, and emotion, of the target audio and the jacket image of the target audio.

[0104] Furthermore, the audio tags of the target audio may include pre-labeled tags. The audio tags of the target audio may include tags obtained by detecting the audio content of the target audio and / or the jacket image of the target audio. Here, the tags obtained by detecting based on the audio content are used to represent the audio type or audio style, and the tags obtained by detecting based on the jacket image of the target audio are used to represent the jacket image style.

[0105] Furthermore, the second visualization material may be displayed in a first screen area of ​​the target video, which has been described above and will not be further described here.

[0106] FIG. 5 is an interface schematic diagram illustrating another preset playback interface according to an embodiment of the present disclosure.

[0107] 5, a preset playback interface is displayed on electronic device 501, and a target video automatically generated based on the song "Hula Dance XXX" may be displayed on the preset playback interface. Here, the target video displays a background image 502 that matches the music style of the song "Hula Dance XXX."

[0108] In still other embodiments of the present disclosure, the visualization material may include a third visualization material, which may have an animation effect generated based on the related information.

[0109] Here, the related information may include musical characteristics of the target audio, such as rhythmic characteristics, drum beat characteristics, tone characteristics, and timbre characteristics of the target audio.

[0110] Furthermore, the animation effects may include dynamic element shapes, dynamic element colors, dynamic element transformation methods, background colors, etc., that match the musical characteristics of the target audio.

[0111] Specifically, when the electronic device detects a first trigger operation on a target object, the electronic device may first detect musical characteristics of the target audio, and then input the detected musical characteristics of the target audio into a pre-trained animation effect generation model to obtain a background image corresponding to a third visualizable material. The background image corresponding to the third visualizable material may include an animation effect generated based on the musical characteristics.

[0112] Furthermore, the third visualization material may be displayed in a first screen area of ​​the target video, which has been described above and will not be further described here.

[0113] FIG. 6 is an interface schematic diagram illustrating yet another preset playback interface according to an embodiment of the present disclosure.

[0114] 6, a preset playback interface is displayed on an electronic device 601, and a target video automatically generated based on the song "Hula Dance XXX" may be displayed on the preset playback interface. Here, the target video displays a background dynamic effect 602 that matches the musical characteristics of the song "Hula Dance XXX."

[0115] In these embodiments, the electronic device may optionally first perform timbre detection and rhythm detection on the target audio to obtain timbre features and rhythm features of the target audio, identify a playing instrument corresponding to the timbre features, and then input the rhythm features into a pre-trained animation effect generation model corresponding to the playing instrument to obtain a background image corresponding to the third visualization material. The background image corresponding to the third visualization material may include an animation effect in which the playing instrument plays based on the rhythm features.

[0116] For example, after the electronic device detects the timbre of the target audio, if the target audio is identified as a piano accompaniment, the animation effect generated based on the musical characteristics may be a piano playing, and the rising and falling rhythm of the piano keys is the same as the rhythm of the target audio.

[0117] In still other embodiments of the present disclosure, the visualization material may include a fourth visualization material, and the fourth visualization material may include related information.

[0118] Here, the related information may include audio text of the target audio, and the audio text may include at least one of first text information related to the target audio and second text information obtained by speech recognition of the target audio.

[0119] Specifically, the fourth visual material may include audio text displayed in the form of a sticker.

[0120] In some embodiments, the first text information may include at least one of a performer name and an audio name.

[0121] If lyrics are not added to the target audio, when the electronic device detects a first trigger operation on the target object, it may obtain the performer name and audio name, generate a lyric sticker corresponding to the fourth visualized material based on the performer name and audio name, and further generate a target video based on the fourth visualized material.

[0122] In some other embodiments, if lyrics are added to the target audio, the first text information may further include the lyrics.

[0123] If lyrics are added to the target audio, when the electronic device detects a first trigger operation on the target object, it may directly obtain the lyrics, performer names, and audio name, generate a lyrics sticker corresponding to the fourth visualized material based on the lyrics, performer names, and audio name, and further generate a target video based on the fourth visualized material.

[0124] In still other embodiments, if the target audio does not have lyrics attached, the second text information may include lyrics obtained by speech recognition of the audio content of the target audio.

[0125] If the target audio does not have lyrics, when the electronic device detects a first trigger operation on the target object, it may obtain the performer name and audio name, and use audio-to-text conversion technology to automatically recognize the lyrics of the target audio, and then generate a lyric sticker corresponding to the fourth visualized material based on the lyrics, performer name, and audio name, and further generate a target video based on the fourth visualized material.

[0126] Continuing to refer to Fig. 4, the target audio may further display an audio sticker 404. Continuing to refer to Fig. 5, the target audio may further display an audio sticker 503.

[0127] In yet another embodiment of the present disclosure, the visualization material may further include a plurality of sequential images that are displayed in turn at preset time intervals.

[0128] In some embodiments, the preset time interval may be determined based on the audio duration of the target audio and the number of images in the sequence.

[0129] Specifically, after the electronic device selects a plurality of sequence images, it divides the audio time length of the target audio by the number of sequence images to obtain a predetermined time interval between two sequence images, and then generates a target video based on the plurality of sequence images, so that the sequence images in the target video can be displayed sequentially in sequence order, and there may be the predetermined time interval between every two sequence images.

[0130] Here, the sequence images may be a plurality of background materials, a plurality of foreground materials, a plurality of second visualization materials, or a plurality of third visualization materials, and are not limited thereto.

[0131] In some other embodiments, the preset time interval may be determined based on the audio rhythm of the target audio.

[0132] Specifically, the electronic device selects a plurality of sequence images, determines the number of pacings according to the number of sequence images, selects the number of rhythmic strong beats according to the audio rhythm of the target audio, and then sets the time intervals between every two rhythmic strong beats as preset time intervals, and generates a target video based on the plurality of sequence images, so that the sequence images in the target video can be displayed in sequence order, and there may be the preset time interval between every two sequence images.

[0133] Here, the sequence images may be a plurality of background materials, a plurality of foreground materials, a plurality of second visualization materials, or a plurality of third visualization materials, and are not limited thereto.

[0134] In yet another embodiment of the present disclosure, a simpler and more convenient video editing function may be provided to users to further lower the barrier to video creation.

[0135] In these embodiments, a preset playback interface may be used to edit the target video.

[0136] Optionally, the preset playback interface may be a video editing interface.

[0137] In some embodiments of the present disclosure, after S120, the audio sharing method includes: The method may further include, when a video cropping operation on the target video is detected, setting a video segment selected from the target video by the video cropping operation as a cropped target video.

[0138] Here, the video clipping operation may include a trigger operation for the video clipping mode, a selection operation for a video segment, and a confirmation operation for the selection result.

[0139] Specifically, the trigger operation for the video crop mode may include a gesture control operation (e.g., tap, long press, double tap, etc.), a voice control operation, or a facial expression control operation on the video crop control to trigger entering the video crop mode. The selection operation for the video segment includes a gesture control operation (e.g., drag, tap, etc.), a voice control operation, or a facial expression control operation on at least one of a start time node and an end time node on a time length selection control to select the start time and end time of the video segment. The confirmation operation for the selection result may include a gesture control operation (e.g., tap, long press, double tap, etc.), a voice control operation, or a facial expression control operation on the confirmation control to trigger video cropping.

[0140] Continuing to refer to FIG. 4, a pull-down button 409 may be displayed in the preset playback interface. The pull-down button 409 may be used to display a function button that is not currently displayed, such as a video crop button. The user may tap the pull-down button 409 to cause the electronic device 401 to display the video crop button, and then tap the video crop button to cause the electronic device to display the video crop interface shown in FIG. 7.

[0141] FIG. 7 is a schematic diagram illustrating video cropping interactions according to an embodiment of the present disclosure.

[0142] As shown in FIG. 7 , the video cropping interface 701 displays a target video preview window 702, a duration selection panel 703, a confirmation icon 704, and a cancel icon 705. The user can select a video frame corresponding to the start time of a video segment by dragging a start time node 706 on the duration selection panel 703. While the user is dragging the start time node 706, the video frames displayed in the preview window 702 change according to the timestamp corresponding to the start time node 706. Once the user confirms that the video cropping is complete, the user can tap the confirmation icon 704 to have the electronic device take the video segment selected by the user as the cropped target video and display the cropped target video in the preset playback interface shown in FIG. 4. If the user does not want to crop the video, the user can tap the cancel icon 705 to return the display of the electronic device to the preset playback interface shown in FIG. 4 and retain the target video before entering the video cropping interface 701.

[0143] In some embodiments of the present disclosure, after S120, the audio sharing method includes: The method may further include, when a material modification operation on the visualized material is detected, performing material modification on the visualized material in accordance with a material modification method corresponding to the material modification operation.

[0144] In the embodiments of the present disclosure, the expression form and content of various visualization materials can all be modified.

[0145] Here, the material modification method may include at least one of the following:

[0146] First, modify the content of the visualization material. Here, modifying the material content of the visualized material may include replacing images in the visualized material, changing text in the visualized material, adding new visualized material (e.g., adding new stickers, text), and deleting existing visualized material (e.g., deleting any existing images, stickers, text).

[0147] Here, the material modification operation may include a selection operation on an image to be replaced, an editing operation on text to be changed, an addition operation on the visualized material, and a deletion operation on the visualized material.

[0148] Continuing to refer to FIG. 4, the preset playback interface may display a text button 405, a sticker button 406, an effect button 407, and a filter button 408. Here, the text button 405 may be used to add new text, the sticker button 406 may be used to add new stickers, the effect button 407 may be used to add effects to the target video, and the filter button 408 may be used to add filters to the target video.

[0149] FIG. 8 is a schematic diagram illustrating material modification interactions according to an embodiment of the present disclosure.

[0150] 8, electronic device 801 may display a preset playback interface 802. A target video may be displayed in preset playback interface 802. The target video may be displayed with a lyric sticker 803. A user may tap lyric sticker 803 to display a delete icon 804. A user may tap delete icon 804 to delete lyric sticker 803.

[0151] FIG. 9 is a schematic diagram illustrating another material modification interaction according to an embodiment of the present disclosure.

[0152] As shown in Fig. 9, the electronic device 901 may display a preset playback interface 902. A target video may be displayed in the preset playback interface 902. A text button 903 may also be displayed in the preset playback interface 902. A user may tap the text button 903 to cause the electronic device 901 to display the preset playback interface shown in Fig. 10.

[0153] FIG. 10 is a schematic diagram illustrating yet another material modification interaction according to an embodiment of the present disclosure.

[0154] 10, an electronic device 1001 can display a preset playback interface 1002. A target video may be displayed in the preset playback interface 1002. The target video may have an added text box 1003. A user can edit the text they want to add in the text box 1003.

[0155] Second, modify the material style of the visualization material. Here, modifying the material style of the visualized material may include modifying the display format of the visualized material.

[0156] For example, if the visualization material is a sticker, the sticker may have different display forms, for example, with a border, without a border, a wide border, a narrow border, etc. For example, if the visualization material is an image, the image may have different display forms, for example, with a border, without a border, a wide border, a narrow border, etc. For example, if the visualization material is a lyric sticker, the lyric sticker may have different display forms, for example, different lyric scrolling forms, different player shapes, etc.

[0157] Continuing to refer to FIG. 8 , the electronic device 801 may display a preset playback interface 802. A target video may be displayed on the preset playback interface 802. The target video may have a lyrics sticker 803. A user can change the expression form of the lyrics sticker 803 by long pressing the lyrics sticker 803. Here, each time the user long presses the lyrics sticker 803, the lyrics sticker 803 changes its expression form once according to a preset sequence of changing the expression form.

[0158] Continuing to refer to FIG. 8 , electronic device 801 may display a preset playback interface 802. A target video may be displayed on preset playback interface 802. The target video may have a lyrics sticker 803. A user may also tap lyrics sticker 803 to display a refresh icon 806. A user may tap refresh icon 806 to change the presentation format of lyrics sticker 803. Here, each time a user taps refresh icon 806, lyrics sticker 803 changes its presentation format once according to a preset presentation format change sequence.

[0159] 3. Modify the display size of the visualization material. Specifically, the user can change the display size of the visualized material according to the gesture operation method by performing a gesture operation to enlarge or reduce the visualized material that the user wants to modify.

[0160] 4. Modify the display position of the visualization material. Specifically, by dragging the visualized material that the user wants to move, the display position of the visualized material can be changed in accordance with the user's dragging operation, and the visualized material can ultimately be displayed at the position where the user stops the dragging operation.

[0161] 5. Modify the display angle of the visualization material. Here, the display angle is the rotation angle of the visualized material.

[0162] Continuing to refer to FIG. 8 , electronic device 801 may display preset playback interface 802. A target video may be displayed in preset playback interface 802. The target video may have a lyrics sticker 803. A user may also tap lyrics sticker 803 to display a rotation icon 805. A user may change the rotation angle of lyrics sticker 803 by tapping rotation icon 805.

[0163] To allow users to interact with other users within the target video, the embodiments of the present disclosure further provide another audio sharing method.

[0164] FIG. 11 is a schematic flow chart diagram illustrating another audio sharing method according to an embodiment of the present disclosure.

[0165] As shown in FIG. 11, the audio sharing method may include the following steps:

[0166] S1110: A target object including a shared control of the original video and / or target audio with the target audio waiting to be shared as background music is displayed in a target interaction interface.

[0167] Among them, the target audio is the audio that has already been released.

[0168] S1120: When a first trigger operation on a target object is detected, a preset playback interface for sharing the target audio is displayed, which includes visualization material generated based on the target audio and displays a target video with the target audio as background music.

[0169] Here, S1110 to S1120 are similar to S110 to S120 in the above embodiment, and therefore will not be further explained here.

[0170] S1130: When a third trigger operation on the target video is detected, the target video is shared.

[0171] In an embodiment of the present disclosure, the user can select whether the target video is Hope If it is determined that the effect is satisfied, a third trigger operation may be input to the electronic device. Here, the third trigger operation may be an operation for triggering sharing of the target video. When the electronic device detects the third trigger operation for the target video, it may share the target video.

[0172] In some embodiments of the present disclosure, the third trigger operation may include a gesture control operation (eg, tap, long press, double tap, etc.), a voice control operation, or a facial expression control operation on the target video.

[0173] In some other embodiments of the present disclosure, the third trigger operation may further include a gesture control operation (e.g., tap, long press, double tap, etc.), a voice control operation, or a facial expression control operation on the sharing control of the target video. Here, the sharing control of the target video may be a control for triggering sharing of the target video. Specifically, the control may be an object that can be triggered by a user, such as a button, an icon, etc.

[0174] In an embodiment of the present disclosure, sharing the target video may include at least one of the following:

[0175] First, a target video is presented in the first application program to which the target interaction interface belongs. Here, the first application program may be any type of application program.

[0176] For example, the first application program may be a short video application program to which the target interaction interface belongs. Sharing the target video may specifically be publishing the target video in the short video application program to which the target interaction interface belongs, so that the target video is distributed to other users who use the short video application program, or the target video is stored as a private video on the server of the short video application program.

[0177] Continuing to refer to FIG. 4, the preset playback interface may display a "Publish" button 410. When the user finally wants to share the edited target video, the user can tap the "Publish" button 410 to publish the target video as a daily video.

[0178] Second, the target video is presented in a second application program other than the first application program. Here, the second application program may be any type of other application program other than the first application program to which the target interaction interface belongs.

[0179] For example, the second application program may be a social application program other than the short video application program to which the target interaction interface belongs. Sharing the target video may specifically mean publishing the target video within the social platform of the social application program.

[0180] Third, send the target video to at least one target user. Here, sharing the target video may specifically mean sending the target video to a chat interface between the user and at least one target user in a first application program, a chat interface between the user and at least one target user in a second application program, or a communication account of at least one target user via an instant messaging tool.

[0181] As a result, the embodiments of the present disclosure may realize sharing of the target video in various forms, allowing the user to present the target video as a legitimate work and receive positive feedback such as views and interactions from others.

[0182] FIG. 12 is a schematic diagram illustrating a playback interface of a target video according to an embodiment of the present disclosure.

[0183] As shown in FIG. 12, electronic device 1201 displays a video playback interface, which may display target video 1202 that is automatically generated based on the final edited version of the song "Hula Dance XXX" released by Mr. C.

[0184] In some other embodiments of the present disclosure, after S1130, the audio sharing method includes: When interaction information for the target video is received, the method may further include overlaying an interaction information display control generated based on the interaction information on the target video.

[0185] Specifically, a user who has viewed a target video can publish interaction information for the target video on the video playback interface of the target video. When the server receives the interaction information for the target video, it transmits the interaction information for the target video to the electronic device of the user who published the target video, and causes the electronic device to generate an interaction information display control for the target video based on the interaction information. The interaction information display control can display the interaction information for the target video, so that the user who published the target video can see the interaction information within the target video that users who have viewed the target video interact with.

[0186] Optionally, the interaction information display control may be used to interact with the information sender of the interaction information.

[0187] Specifically, the user who posted the target video can interact with the sender of the interaction information about the interaction information that interests them, for example, by 'liking' or commenting, through the interaction information display control.

[0188] As described above, the audio sharing method according to the embodiment of the present disclosure can intelligently generate target videos with various expression forms based on the target audio, and provides a simple and easy-to-operate video editing method, thereby lowering the barrier to video creation, so that even sharers lacking creative ability can create high-quality videos, and the created videos can achieve the sharing effect desired by the user. Furthermore, after a user shares the target audio using the intelligently generated target video, the target audio in the target video can be used as a public expression to receive views and interactions, thereby improving the user experience.

[0189] An embodiment of the present disclosure further provides an audio sharing device, which will be described below in conjunction with FIG.

[0190] In an embodiment of the present disclosure, the audio sharing device may be an electronic device, which may include, but is not limited to, mobile terminals such as mobile phones, laptops, digital broadcast receivers, PDAs, PADs, PMPs, in-vehicle terminals (e.g., in-vehicle navigation terminals), and wearable devices, and fixed terminals such as digital TVs, desktop computers, and smart homes.

[0191] FIG. 13 is a schematic diagram illustrating an audio sharing device according to an embodiment of the present disclosure.

[0192] As shown in FIG. 13, an audio sharing device 1300 may include a first display means 1310 and a second display means 1320.

[0193] The first display means 1310 may be configured to display a target object in a target interaction interface, including an original video with the target audio, which is published audio waiting to be shared, as background music and / or a shared control of the target audio.

[0194] The second display means 1320 may be configured to display, when a first trigger operation on the target object is detected, a preset playback interface for sharing the target audio, including visualization material generated based on the target audio and displaying a target video with the target audio as background music.

[0195] In the embodiments of the present disclosure, when a user triggers audio sharing for a target audio, a preset playback interface can be directly displayed, displaying a target video automatically generated from the target audio, with the target video using the target audio as background music. Sharing the target audio using the target video not only diversifies the shared content and meets the individual needs of users, but also lowers the barrier to video creation, allowing users to conveniently share audio with video content without having to shoot or upload videos.

[0196] In some embodiments of the present disclosure, the visualization material may include images and / or text generated based on relevant information of the target audio.

[0197] In some embodiments of the present disclosure, the visualization material may include a first visualization material. The first visualization material may include an image generated based on related information. The related information may include a related image of the target audio. The related image may include at least one of an audio jacket image and a presenter icon.

[0198] In some embodiments of the present disclosure, the first visualization material may include at least one of the following:

[0199] Background material. The background material includes an image generated based on image features of the related image. The background material is displayed in a first screen area of ​​the target video.

[0200] Foreground material. The foreground material includes related images or images generated based on image features of the related images. The foreground material is displayed in a second screen area of ​​the target video.

[0201] Here, at least a part of the second screen area is included within the first screen area, and the first screen area is a background display area of ​​the target moving image.

[0202] In some embodiments of the present disclosure, the visual material may include a second visual material, and the second visual material may include a background image that matches related information selected from a material library. The related information may include an audio tag for the target audio.

[0203] In some embodiments of the present disclosure, the visualization material may include a third visualization material, which may have animation effects generated based on the related information, and the related information may include musical characteristics of the target audio.

[0204] In some embodiments of the present disclosure, the visualization material may include a fourth visualization material, the fourth visualization material may include related information, and the related information may include audio text of the target audio.

[0205] Here, the audio text may include at least one of first text information related to the target audio and second text information obtained by speech recognition of the target audio.

[0206] In some embodiments of the present disclosure, the visualization material may include a plurality of sequential images that are displayed in sequence at preset time intervals.

[0207] Here, the preset time interval may be determined based on the audio time length of the target audio and the number of sequence images, or may be determined based on the audio rhythm of the target audio.

[0208] In some embodiments of the present disclosure, a preset playback interface may be used to edit the target video.

[0209] Here, the audio sharing device 1300 may further include a first processing means that may be configured to, after displaying the preset playback interface, when a video cropping operation on the target video is detected, set a video segment selected from the target video by the video cropping operation as a cropped target video.

[0210] In some embodiments of the present disclosure, a preset playback interface may be used to edit the target video.

[0211] Here, the audio sharing device 1300 may further include a second processing means that may be configured to, after displaying the preset playback interface, when a material modification operation on the visualized material is detected, perform material modification on the visualized material according to a material modification method corresponding to the material modification operation.

[0212] Here, the material modification method may include at least one of modifying the material content of the visualized material, modifying the material style of the visualized material, modifying the display size of the visualized material, modifying the display position of the visualized material, and modifying the display angle of the visualized material.

[0213] In some embodiments of the present disclosure, the target interaction interface may include an audio exhibit interface, and the target object may include shared control of the target audio.

[0214] Here, the audio sharing device 1300 may further include a third processing means that may be configured to display an original video with the target audio as background music and an audio display control of the target audio in the video playback interface before displaying the target object.

[0215] Here, the first display means may be further configured to display the audio exhibit interface in which the shared control is displayed when a second trigger operation on the audio exhibit control is detected.

[0216] In some embodiments of the present disclosure, the target interaction interface may include a video playback interface, the target object may include an original video with the target audio as background music, and the first trigger operation may include a trigger operation on the original video.

[0217] Alternatively, the target interaction interface may include a video playback interface, the target object may include an original video, the original video may include an audio control for the target audio, and the first trigger operation may include a trigger operation for the audio control.

[0218] Alternatively, the target interaction interface may include a video playback interface, the target object may include a shared control of the original video and the target audio, and the first trigger operation may include a trigger operation on the shared control.

[0219] In some embodiments of the present disclosure, the video sharing device 1300 may further include a video sharing unit. The video sharing unit may be configured to share the target video when a third trigger operation on the target video is detected after displaying the preset playback interface.

[0220] Here, sharing the target video may include at least one of publishing the target video within a first application program to which the target interaction interface belongs, publishing the target video within a second application program other than the first application program, and sending the target video to at least one target user.

[0221] It should be noted that the audio sharing device 1300 shown in FIG. 13 can execute each step in the method embodiments shown in FIGS. 1 to 12 and realize each process and effect in the method embodiments shown in FIGS. 1 to 12, and will not be further described here.

[0222] An embodiment of the present disclosure further provides an electronic device including a processor and a memory storing executable instructions, wherein the processor may be used to read the executable instructions from the memory and execute the executable instructions to implement the audio sharing method in the embodiment.

[0223] 14 is a block diagram of an electronic device according to an embodiment of the present disclosure. Reference is now made specifically to FIG. 14, which illustrates a block diagram of an electronic device 1400 suitable for implementing an embodiment of the present disclosure.

[0224] The electronic device 1400 in the embodiment of the present disclosure may be an electronic device, which may include, but is not limited to, mobile terminals such as mobile phones, notebook computers, digital broadcast receivers, PDAs, PADs, PMPs, in-vehicle terminals (e.g., in-vehicle navigation terminals), and wearable devices, and fixed terminals such as digital TVs, desktop computers, and smart homes.

[0225] It should be noted that the electronic device 1400 shown in FIG. 14 is merely an example and does not impose any limitations on the functions and scope of use of the embodiments of the present disclosure.

[0226] 14, the electronic device 1400 may include a processing unit (e.g., a central processing unit, a graphics processor, etc.) 1401, which can perform various appropriate operations and processes according to programs stored in a read-only memory (ROM) 1402 or programs loaded from a storage device 1408 into a random access memory (RAM) 1403. The RAM 1403 further stores various programs and data necessary for the operation of the electronic device 1400. The processing unit 1401, the ROM 1402, and the RAM 1403 are interconnected via a bus 1404. An input / output (I / O) interface 1405 is also connected to the bus 1404.

[0227] Typically, input devices 1406, including, for example, a touch screen, touch panel, keyboard, mouse, camera head, microphone, accelerometer, gyroscope, etc.; output devices 1407, including, for example, a liquid crystal display (LCD), speaker, oscillator, etc.; storage devices 1408, including, for example, a magnetic tape, hard disk, etc.; and communication devices 1409 may be connected to the I / O interface 1405. The communication devices 1409 enable the electronic device 1400 to communicate wirelessly or via wires with other devices to exchange data. While FIG. 14 illustrates the electronic device 1400 with various devices, it should be understood that it is not intended to require the implementation or inclusion of all of the devices shown. Alternative implementations and / or inclusion of more or fewer devices are possible.

[0228] An embodiment of the present disclosure further provides a computer-readable storage medium having stored thereon a computer program that, when executed by a processor, causes the processor to implement the audio sharing method in the above embodiment.

[0229] In particular, according to embodiments of the present disclosure, the processes described with reference to the flowcharts above can be implemented as a computer software program. For example, embodiments of the present disclosure include a computer program product, which includes a computer program stored on a non-transitory computer-readable medium, including program code for performing the methods illustrated in the flowcharts. In such embodiments, the computer program may be downloaded and installed from a network via the communication device 1409, or installed from the storage device 1408, or installed from the ROM 1402. When the computer program is executed by the processing device 1401, it performs the functions defined above in the audio sharing method of the embodiments of the present disclosure.

[0230] In the present disclosure, the computer-readable medium may be a computer-readable signal medium, a computer-readable storage medium, or any combination thereof. The computer-readable storage medium may be, for example, but is not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination thereof. More specific examples of the computer-readable storage medium include, but are not limited to, an electrical connection having one or more conductors, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination thereof. In the present disclosure, the computer-readable storage medium may be any tangible medium that contains or stores a program, and the program can be used in or in combination with a command execution system, apparatus, or device. In the present disclosure, the computer-readable signal medium may include a data signal propagated in baseband or a data signal propagated as part of a carrier wave, with computer-readable program code embodied therein. Such propagated data signals may take a variety of forms, including, but not limited to, electromagnetic signals, optical signals, or any suitable combination thereof. A computer-readable signal medium may also be any computer-readable medium other than a computer-readable storage medium, which may transmit, propagate, or transmit a program for use in or in connection with a command execution system, apparatus, or device. The program code contained in the computer-readable medium may be transmitted over any suitable medium, including, but not limited to, electrical wire, optical cable, RF (radio frequency), or the like, or any suitable combination thereof.

[0231] In some embodiments, clients and servers may communicate using any now known or future developed network protocol, such as HTTP, and may interconnect with any form or medium of digital data communication (e.g., a communication network). Examples of communication networks include a local area network ("LAN"), a wide area network ("WAN"), the Internet, and an end-to-end network (e.g., an ad hoc end-to-end network), as well as any now known or future developed network.

[0232] The computer-readable medium may be included in the electronic device, or may be a standalone medium not mounted on the electronic device.

[0233] The computer-readable medium stores one or more programs, and when the one or more programs are executed by the electronic device, the electronic device: The target object including an original video with the target audio, which is published audio waiting to be shared, as background music and / or a sharing control for the target audio is displayed on a target interaction interface, and when a first trigger operation on the target object is detected, a preset playback interface is displayed in which a target video with the target audio as background music is displayed and which includes visualization material generated based on the target audio for sharing the target audio.

[0234] In embodiments of the present disclosure, computer program code for carrying out the operations of the present disclosure can be written using one or more programming languages, or a combination thereof, including, but not limited to, object-oriented programming languages ​​such as Java, Smalltalk, and C++, as well as general procedural programming languages ​​such as "C" or similar programming languages. The program code can execute entirely on the user computer, partially on the user computer, as a separate software package, partially on the user computer and partially on a remote computer, or entirely on a remote computer or server. In the case of a remote computer, the remote computer can be connected to the user computer by any network, including a local area network (LAN) or a wide area network (WAN), or can be connected to an external computer (e.g., via the Internet using an Internet Service Provider).

[0235] The flowcharts and block diagrams in the accompanying drawings illustrate possible system architectures, functions, and operations of systems, methods, and computer program products according to various embodiments of the present disclosure. In this regard, each block in the flowcharts or block diagrams may represent a module, program segment, or portion of code, which includes one or more executable instructions for implementing a specified logical function. It should be noted that in some alternative implementations, the functions depicted in the blocks may be implemented in a different order than depicted in the drawings. For example, two successively shown blocks may be executed substantially simultaneously, or, depending on the functionality, they may be executed in the reverse order. It should be noted that each block in the block diagrams and / or flowcharts, and combinations of blocks in the block diagrams and / or flowcharts, may be implemented by a dedicated hardware-based system that performs the specified functions or operations, or by a combination of dedicated hardware and computer instructions.

[0236] The units according to the embodiments of the present disclosure may be realized by software or hardware, and the names of the units may not necessarily limit the units themselves.

[0237] The functions described herein above may be performed, at least in part, by one or more hardware logic components. For example, exemplary hardware logic components that may be used include, but are not limited to, field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), systems on chips (SOCs), complex programmable logic devices (CPLDs), etc.

[0238] In the present disclosure, a machine-readable medium may be a tangible medium and may contain or store a program used by or in combination with a command execution system, apparatus, or device. The machine-readable medium may be a machine-readable signal medium or a machine-readable storage medium. The machine-readable medium may include, but is not limited to, an electronic, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any suitable combination of the above. More specific examples of machine-readable storage media include one or more wire-based electrical connections, portable computer disks, hard disks, random access memory (RAM), read-only memory (ROM), erasable programmable read-only memory (EPROM or flash memory), optical fiber, portable compact disc read-only memory (CD-ROM), optical storage device, magnetic storage device, or any suitable combination thereof.

[0239] An embodiment of the present disclosure further provides a computer program including instructions that, when executed by a processor, implement the audio sharing method in the above embodiment.

[0240] An embodiment of the present disclosure further provides a computer program product including a computer program or instructions that, when executed by a processor, implements the audio sharing method in the above embodiment.

[0241] The above is merely a description of the preferred embodiments and the technical principles applied in the present disclosure. It is obvious to those skilled in the art that the scope of the disclosure is not limited to the technical solution based on the specific combination of the above technical features, but should also include other technical solutions formed by any combination of the above technical features or equivalent features without departing from the concept of the above disclosure. For example, it also includes technical solutions formed by replacing the above features with technical features having similar functions (including but not limited to) disclosed in the present disclosure.

[0242] Also, although operations are described in a particular order, this should not be understood as requiring such operations to be performed in the particular order shown, or in a sequential order. In certain environments, multitasking and parallel processing may be advantageous. Similarly, although certain specific implementation details are included in the above description, they should not be construed as limiting the scope of the present disclosure. Certain features described in the context of a single embodiment can also be implemented in combination in a single embodiment. Conversely, various features described in the context of a single embodiment can be implemented in multiple embodiments separately or in any suitable subcombination.

[0243] Although the present subject matter has been described in language specific to structural features and / or methodological operations, it is to be understood that the subject matter defined in the appended claims is not necessarily limited to the specific features or operations described above. Rather, the specific features and operations described above are merely example forms of implementing the claims.

Claims

1. Displaying an original video having a target audio and / or a target object including a shared control of the target audio in a target interaction interface, the target audio being background music of the original video, and the target audio being published audio; When a first trigger operation on the target object is detected, a visualization material automatically generated based on the target audio is obtained, and a preset playback interface is displayed, in which a target video with the target audio as background music is displayed, the target video is for sharing the target audio, and the target video includes the visualization material; The audio sharing method, wherein the visualization material includes an image generated based on a related image of the target audio.

2. The method of claim 1 , wherein the visualization material includes images and / or text generated based on related information of the target audio.

3. The method of claim 2 , wherein the visualization material includes first visualization material, the first visualization material including an image generated based on the related information.

4. The method of claim 3 , wherein the associated images include at least one of an audio jacket image and a presenter icon.

5. The first visualization material includes: background material displayed in a first screen area of ​​the target video, the background material including an image generated based on image features of the related image; a foreground material displayed in a second screen area of ​​the target video, the foreground material including the related image or an image generated based on image features of the related image; and The method of claim 3 , wherein at least a portion of the second screen area is contained within a first screen area, the first screen area being a background display area of ​​the target video.

6. 3. The method of claim 2, wherein the visualizable material includes a second visualizable material, the second visualizable material including a background image that matches the related information selected from a material library, and the related information including an audio tag of the target audio.

7. 3. The method of claim 2, wherein the visualization material includes a third visualization material, the third visualization material having an animation effect generated based on the related information, and the related information including musical characteristics of the target audio.

8. The method of claim 2 , wherein the visualization material includes a fourth visualization material, the fourth visualization material includes the related information, and the related information includes an audio text of the target audio.

9. The method of claim 8 , wherein the audio text includes at least one of first text information related to the target audio and second text information obtained by speech recognition of the target audio.

10. The method of claim 1 , wherein the visualization material includes a plurality of sequential images that are displayed in sequence at preset time intervals.

11. The method of claim 10 , wherein the predetermined time interval is determined based on an audio time length of the target audio and the number of the sequence images, or based on an audio rhythm of the target audio.

12. The preset playback interface is further used to edit the target video; After displaying the preset playback interface, the method includes: The method of claim 1 , further comprising: when a video cropping operation on the target video is detected, setting a video segment selected from the target video by the video cropping operation as a cropped target video.

13. The preset playback interface is further used to edit the target video; After displaying the preset playback interface, the method includes: The method of claim 1 , further comprising: when a material modification operation on the visualized material is detected, performing material modification on the visualized material in accordance with a material modification method corresponding to the material modification operation.

14. The method of claim 13 , wherein the material modification method includes at least one of modifying the material content of the visualized material, modifying the material style of the visualized material, modifying the display size of the visualized material, modifying the display position of the visualized material, and modifying the display angle of the visualized material.

15. The target interaction interface includes an audio display interface, and the target object includes shared control of the target audio; Before displaying the target object in the target interaction interface, the method includes: further comprising displaying an original video with the target audio as background music and an audio display control of the target audio on a video playback interface; Displaying the target object in the target interaction interface includes: The method of claim 1 , further comprising displaying the audio exhibit interface with the shared control displayed when a second triggering action on the audio exhibit control is detected.

16. The target interaction interface includes a video playback interface, the target object includes an original video with the target audio as background music, and the first trigger operation includes a trigger operation on the original video; Alternatively, the target interaction interface includes a video playback interface, the target object includes the original video, the original video includes an audio control for the target audio, and the first trigger operation includes a trigger operation for the audio control; Alternatively, the method of claim 1 , wherein the target interaction interface includes a video playback interface, the target object includes a shared control of the original video and the target audio, and the first trigger operation includes a trigger operation on the shared control.

17. After displaying the preset playback interface, the method includes: The method of claim 1 , further comprising: sharing the target video when a third trigger operation on the target video is detected.

18. 18. The method of claim 17, wherein sharing the target video includes at least one of publishing the target video within a first application program to which the target interaction interface belongs, publishing the target video within a second application program other than the first application program, and sending the target video to at least one target user.

19. A first display means configured to display an original video having a target audio and / or a target object including a shared control of the target audio in a target interaction interface, wherein the target audio is background music of the original video and the target audio is published audio; a second display means configured to, when a first trigger operation on the target object is detected, acquire a visual material automatically generated based on the target audio and display a preset playback interface, wherein a target video with the target audio as background music is displayed in the preset playback interface, the target video being for sharing the target audio, and the target video including the visual material; The audio sharing device, wherein the visualization material includes an image generated based on an image related to the target audio.

20. The preset playback interface is further used to edit the target video; 20. The audio sharing device of claim 19, further comprising: a first processing means configured to, when a video cropping operation on the target video is detected after displaying the preset playback interface, set a video segment selected from the target video by the video cropping operation as a cropped target video.

21. The preset playback interface is further used to edit the target video; 20. The audio sharing device of claim 19, further comprising: second processing means configured to, when a material modification operation on the visualized material is detected after displaying the preset playback interface, perform material modification on the visualized material in accordance with a material modification method corresponding to the material modification operation.

22. The audio sharing device of any one of claims 19 to 21, further comprising a video sharing means configured to share the target video when a third trigger operation on the target video is detected after the preset playback interface is displayed.

23. a processor; a memory for storing executable instructions; The processor is configured to read the executable instructions from the memory and execute the executable instructions to realize the audio sharing method described in any one of claims 1 to 18.

24. A computer-readable storage medium having stored thereon a computer program that, when executed by a processor, causes the processor to implement the audio sharing method according to any one of claims 1 to 18.

25. A computer program comprising instructions which, when executed by a processor, cause the computer program to implement the audio sharing method of any one of claims 1 to 18.

26. A computer program product comprising a computer program or instructions which, when executed by a processor, causes the computer program or instructions to implement the audio sharing method according to any one of claims 1 to 18.

Citation Information

Patent Citations

  • Song sharing method, device and storage medium

    CN109144346A

  • Music poster generation method and device, electronic equipment and medium

    CN112069360A

  • AUDIO COVER DISPLAY METHOD AND APPARATUS

    JP2017532582A

  • Method and apparatus for playing music segments

    JP2019506065A

  • System and method for automatically generating media

    US20180374461A1