Interaction method and apparatus, device, medium, and product

By generating and presenting candidate multimedia content in parallel on a content sharing platform, the problem of long waiting times during the generation process is solved, achieving efficient content creation and interactive experience.

CN122331797APending Publication Date: 2026-07-03DOUYIN VISION CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
DOUYIN VISION CO LTD
Filing Date
2026-04-10
Publication Date
2026-07-03

Smart Images

  • Figure CN122331797A_ABST
    Figure CN122331797A_ABST
Patent Text Reader

Abstract

This document provides one or more scenarios for an interactive method, apparatus, device, medium, and product. The interactive method includes: receiving a first operation, the first operation being used to generate candidate multimedia content for first shared content, the first operation being associated with first text; during the generation of at least one candidate multimedia content, receiving a second operation, the at least one candidate multimedia content being generated based on the first text, the second operation being independent of the generation process of the at least one candidate multimedia content; and presenting at least one candidate multimedia content, the at least one candidate multimedia content being used to add to the first shared content. Addressing the issue of low content sharing efficiency, during the generation of candidate multimedia content, users can simultaneously perform other operations related to content sharing. The aforementioned asynchronous generation of candidate multimedia content can effectively improve the efficiency of creating and sharing content and enhance the interactive experience.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] One or more of the situations described herein relate to an interaction method, interaction device, electronic device, computer-readable storage medium, and computer program product. Background Technology

[0002] Applications related to content sharing (such as content sharing platforms) allow users to share content, which can include multimodal information, such as text content and multimedia content (such as pictures, videos, etc.).

[0003] Some content-sharing applications can automatically generate and recommend candidate multimedia content for users to share. Therefore, it is particularly important to improve the efficiency of the above-mentioned interaction process and achieve efficient content sharing. Summary of the Invention

[0004] This summary section is provided to briefly introduce the concepts, which will be described in detail in the detailed description section below. This summary section is not intended to identify key or essential features of the claimed technical solution, nor is it intended to limit the scope of the claimed technical solution.

[0005] This document provides an interaction method in at least one scenario, comprising: receiving a first operation, wherein the first operation is used to generate candidate multimedia content for a first shareable content, the first operation being associated with a first text; receiving a second operation during the generation of at least one candidate multimedia content, wherein the at least one candidate multimedia content is generated based on the first text, and the second operation is independent of the generation process of the at least one candidate multimedia content; and presenting the at least one candidate multimedia content, wherein the at least one candidate multimedia content is used to add to the first shareable content.

[0006] This document provides at least one interactive device, comprising: a first receiving module configured to receive a first operation, wherein the first operation is used to generate candidate multimedia content for a first shareable content, and the first operation is associated with a first text; a second receiving module configured to receive a second operation during the generation of at least one candidate multimedia content, wherein the at least one candidate multimedia content is generated based on the first text, and the second operation is independent of the generation process of the at least one candidate multimedia content; and a presentation module configured to present the at least one candidate multimedia content, wherein the at least one candidate multimedia content is used to add to the first shareable content.

[0007] At least one aspect of this document provides an electronic device, including: at least one processor; and at least one memory, including one or more computer program instructions; wherein the one or more computer program instructions are executed by the processor to perform the interactive method provided by at least one aspect of this document.

[0008] At least one aspect of this document provides a computer-readable storage medium that non-transitory stores computer-readable instructions, wherein the interaction method provided by at least one aspect of this document is implemented when the computer-readable instructions are executed by a processor.

[0009] At least one aspect of this document provides a computer program product, including a computer program that, when executed by a processor, implements the interactive method provided by at least one aspect of this document.

[0010] In one of the interaction methods provided in at least one scenario of this paper, during the overall interaction process of content sharing, candidate multimedia content (such as candidate cover images) for sharing is automatically generated based on the text associated with the user's operation. Furthermore, during the generation of candidate multimedia content, the user can simultaneously perform other operations related to content sharing without having to wait for the candidate multimedia content to be generated. The above-mentioned asynchronous generation of candidate multimedia content can effectively improve the efficiency of creating and sharing content and enhance the interactive experience during the content sharing process. Attached Figure Description

[0011] The above and other features, advantages, and aspects of the various scenarios described herein will become more apparent when taken in conjunction with the accompanying drawings and the following detailed description. Throughout the drawings, the same or similar reference numerals denote the same or similar elements. It should be understood that the drawings are schematic, and the originals and elements are not necessarily drawn to scale.

[0012] Figure 1 This illustration shows an application scenario diagram of at least one of the interaction methods provided in this paper;

[0013] Figure 2 The schematic diagram illustrates a flowchart of an interaction method provided in at least one scenario of this paper;

[0014] Figures 3A to 3M This paper schematically illustrates a page diagram provided in at least one of the scenarios described herein;

[0015] Figure 4 The schematic diagram illustrates the structure of an interactive device provided in at least one of the present invention; and

[0016] Figure 5 A schematic diagram of the structure of an electronic device suitable for implementing at least one of the situations described herein is shown. Detailed Implementation

[0017] One or more scenarios described herein will now be described in more detail with reference to the accompanying drawings. While some scenarios are shown in the drawings, it should be understood that this document can be implemented in various forms and should not be construed as limited to the scenarios set forth herein; rather, these scenarios are provided to provide a more thorough and complete understanding of this document. It should be understood that the accompanying drawings and scenarios are for illustrative purposes only and are not intended to limit the scope of this document.

[0018] It should be understood that the steps described in the method embodiments herein may be performed in different orders and / or in parallel. Furthermore, the method embodiments may include additional steps and / or omit the steps shown. The scope of this document is not limited in this respect.

[0019] The term "comprising" and its variations as used herein are open-ended inclusions, meaning "including but not limited to". The term "based on" means "at least partially based on". The term "one situation" means "at least one situation"; the term "another situation" means "at least one additional situation"; the term "some situations" means "at least some situations". Definitions of other terms will be given in the following description.

[0020] It should be noted that the concepts of "first" and "second" mentioned in this article are only used to distinguish different devices, modules or units, and are not used to limit the order of the functions performed by these devices, modules or units or their interdependencies.

[0021] It should be noted that the terms "one" and "more" used in this document are illustrative rather than restrictive, and those skilled in the art should understand that, unless otherwise expressly indicated in the context, they should be understood as "one or more".

[0022] The names of the messages or information exchanged between the various devices in the embodiments herein are for illustrative purposes only and are not intended to limit the scope of these messages or information.

[0023] It is understood that the data involved in this technical solution (including but not limited to the data itself, the acquisition, use, storage or deletion of the data) shall comply with the requirements of relevant laws, regulations and related provisions.

[0024] It is understood that before using the technical solutions disclosed in each scenario in this article, relevant users should be informed of the type, scope of use, and usage scenarios of the information involved in this article and their authorization should be obtained through appropriate means in accordance with relevant laws and regulations. Relevant users may include any type of rights holder, such as individuals, enterprises, or groups.

[0025] For example, in response to receiving an active request from a user, a prompt message is sent to the relevant user to clearly inform the user that the requested operation will require obtaining and using the user's information, thereby enabling the relevant user to choose whether to provide information to the software or hardware such as electronic devices, applications, servers, or storage media that perform the operation of any of the technical solutions described herein.

[0026] As an optional but non-restrictive implementation, in response to a user's active request, a prompt message can be sent to the user, such as a pop-up window, where the prompt message can be presented in text format. Furthermore, the pop-up window can also include a selection control allowing the user to choose "agree" or "disagree" to provide information to the electronic device.

[0027] It is understood that the above notification and user authorization process are merely illustrative and do not constitute a limitation on the implementation method described in this article. Other methods that comply with relevant laws and regulations may also be applied to the implementation method described in this article.

[0028] A content sharing platform can be understood as an internet platform that provides content sharing functionality. For example, users can publish and share content on a content sharing platform, and users can also view and interact with content shared by others on a content sharing platform.

[0029] Content sharing platforms can be of different types. For example, in social, knowledge-based, and creative applications, users can post and interact with posts published by others; therefore, social, knowledge-based, and creative applications can be understood as a type of content sharing platform. Similarly, in video applications (including both long-form and short-form video applications), users can post videos and interact with other posted videos; therefore, video applications can be understood as a type of content sharing platform. Furthermore, in shopping and lifestyle service applications, users can post reviews and interact with reviews published by others; therefore, shopping and lifestyle service applications can be understood as a type of content sharing platform. Finally, in novel and podcast applications, users can post discussion threads related to the content of the novel or podcast and interact with discussion threads published by others; therefore, novel and podcast applications can be understood as a type of content sharing platform.

[0030] Shared content can include multimodal information, such as text content and multimedia content (e.g., images, videos). During the process of users creating and sharing content on the content sharing platform, the platform can automatically generate and recommend candidate multimedia content for the shared content. For example, a content sharing voucher can automatically generate and recommend candidate cover images for the shared content, eliminating the need for users to manually create cover images.

[0031] Typically, the process of generating candidate multimedia content is executed sequentially with other operations for creating and sharing content. For example, a user can first input a portion of text content, and the content generation platform generates candidate multimedia content based on the user's input. After the candidate multimedia content is generated, the user can select the multimedia content they need from the candidate multimedia content and then continue to edit other parts of the shared content (such as other text content).

[0032] For example, some content sharing platforms provide controls for composing post text. After a user triggers these controls, they enter the post text in the text writing interface, triggering the next control provided by the text writing interface. Then, a window above the writing interface informs the user that candidate cover images are being generated, or the cover image generation interface informs the user that candidate cover images are being generated. After the content sharing platform has generated all the candidate cover images, the candidate cover images are displayed in the cover image viewing interface. The user selects the desired cover image from the candidate cover images and enters the post writing interface (for example, an interface different from the text writing interface) to continue editing the post content.

[0033] In the above-mentioned process of generating candidate multimedia content in sequence, users need to wait for the candidate multimedia content to be generated. The above-mentioned waiting time is a "bottleneck" in the overall interactive process of content sharing, resulting in low efficiency of the user's content sharing interaction process.

[0034] To at least partially solve the above-mentioned technical problem, this paper provides an interaction method in at least one scenario, comprising: receiving a first operation, the first operation being used to generate candidate multimedia content of a first shareable content, the first operation being associated with a first text; during the generation process of at least one candidate multimedia content, receiving a second operation, the at least one candidate multimedia content being generated based on the first text, the second operation being unrelated to the generation process of the at least one candidate multimedia content; presenting at least one candidate multimedia content, the at least one candidate multimedia content being used to add to the first shareable content.

[0035] Based on the interaction method provided in at least one of the embodiments described herein, at least one of the embodiments described herein also provides an interaction device, an electronic device, a computer-readable storage medium, and a computer program product.

[0036] In one of the interaction methods provided in at least one scenario of this paper, during the overall interaction process of content sharing, candidate multimedia content (such as candidate cover images) for sharing is automatically generated based on the text associated with the user's operation. Furthermore, during the generation of candidate multimedia content, the user can simultaneously perform other operations related to content sharing without having to wait for the candidate multimedia content to be generated. The above-mentioned asynchronous generation of candidate multimedia content can effectively improve the efficiency of creating and sharing content and enhance the interactive experience during the content sharing process.

[0037] The following detailed description, with reference to the accompanying drawings, illustrates one or more scenarios and some examples thereof.

[0038] Figure 1 The illustration shows an application scenario diagram of an interaction method provided in at least one of the cases described in this paper.

[0039] like Figure 1 As shown, the application scenario provided in this case may include user 101, terminal device 102, and content sharing platform 103. Terminal device 102 can be various electronic devices capable of providing interactive pages, such as smart wearable devices, smart appliances, smart cars, mobile phones, tablets, laptops, or desktop computers.

[0040] The terminal device 102 may have a client installed. This client may be a client of the content sharing platform 103. The content sharing platform 103 may be a server that supports the operation of the client installed on the terminal device 102. The client in the terminal device 102 may interact with the content sharing platform 103 in different ways. For example, the client in the terminal device 102 may interact with the content sharing platform 103 through a browser, that is, the client in the terminal device 102 may be a web page; or, for another example, the client in the terminal device 102 may interact with the content sharing platform 103 through an application (APP), that is, the client in the terminal device 102 may be a mobile device.

[0041] The content sharing platform 103 can be deployed in different ways. For example, the content sharing platform 103 can be deployed in the cloud, that is, the content sharing platform 103 can be a cloud server.

[0042] User 101 can be a user of a client installed on terminal device 102. For example, user 101 can be a user who publishes, views, or interacts with shared content on the client in terminal device 102.

[0043] The content sharing platform 103 can communicate with the terminal device 102. For example, the content sharing platform 103 can provide the client installed on the terminal device 102 with relevant data required by the terminal device to run the client (such as page data of interactive pages); or, for example, the content sharing platform 103 can also receive relevant data returned by the terminal device 102 during the running of the client (such as data related to the shared content).

[0044] The interaction methods provided in one or more scenarios described herein can be implemented in software, hardware, firmware, or any combination thereof.

[0045] For example, the interaction methods provided in one or more scenarios of this document are applicable to a client installed in terminal device 102, and the client installed in terminal device 102 can load and execute the interaction methods. For example, the client installed in terminal device 102 may include a central processing unit (CPU), graphics processing unit (GPU), digital signal processor (DSP), neural network processing unit (NPU), or other forms of processing units with data processing capabilities and / or instruction execution capabilities, storage units, etc. The client installed in terminal device 102 may also have an operating system and various types of application programming interfaces (APIs) installed on it, and implement the interaction methods provided in one or more scenarios of this document by running code or instructions.

[0046] The following will combine Figure 2 , Figures 3A to 3M This paper provides a detailed description of an interaction method for at least one scenario.

[0047] Figure 2 The diagram illustrates a flowchart of an interaction method provided in at least one scenario of this paper.

[0048] like Figure 2 As shown, the interaction method in this scenario includes steps S201 to S203, and the steps included in this interaction method are described below:

[0049] Step S201: Receive the first operation.

[0050] In one or more scenarios described herein, the first operation can be understood as an operation triggered during the process of a user sharing content on a content sharing platform, and the first operation can be used to generate candidate multimedia content for the first shared content.

[0051] The first share content can be understood as the content that the user wants to share on the content sharing platform. For example, the first share content can be the content that the user is currently editing.

[0052] The first shared content may include multimedia content. For example, the first shared content may include text content and multimedia content, and there may be a content association between the text content and the multimedia content. The multimedia content in the first shared content may be at least one image, or the multimedia content in the first shared content may also be at least one video, or the multimedia content in the first shared content may also be at least one image and at least one video.

[0053] On some content sharing platforms, when displaying shared content to users, a piece of shared content is presented as a cover image and a piece of text. In this case, multimedia content can also be presented as a cover image.

[0054] The candidate multimedia content of the first shared content can be understood as multimedia content associated with the first shared content that is available for the user to choose from. That is, the user can select one or more of the candidate multimedia content of the first shared content as the multimedia content in the first shared content.

[0055] For example, the candidate multimedia content for the first shared content can be multiple candidate cover images, and users can choose one from the candidate multimedia content as the cover image for the first shared content.

[0056] Candidate multimedia content for the first shared content can be generated automatically. For example, the candidate multimedia content for the first shared content can be automatically generated by the content sharing platform. That is, when a user shares content using the content sharing platform, the platform can automatically generate candidate multimedia content and reduce the workload of the user in editing the multimedia content in the first shared content by recommending candidate multimedia content to the user.

[0057] In other words, the first operation can be understood as the operation that initiates the generation process of candidate multimedia content for the first shared content, that is, in response to the first operation, the generation of candidate multimedia content for the first shared content begins.

[0058] In one or more cases in this paper, the first operation can be associated with the first text, that is, the first operation can be understood as an operation associated with the first text, and the first operation can carry the first text.

[0059] This article does not restrict the way the first operation is received in one or more scenarios. For example, the first operation can be received in any interface provided by the content sharing platform.

[0060] For example, the first shared content may not be document content. That is to say, the interaction method provided in at least one scenario in this article is not aimed at the scenario of editing document content in a document application (such as a cloud document service), but at the scenario of editing shared content (such as editing a post) on a content sharing platform.

[0061] Step S202: During the generation of at least one candidate multimedia content, receive the second operation.

[0062] In one or more cases described herein, at least one candidate multimedia content is generated based on a first text, that is, the first text can be understood as the basis for generating at least one candidate multimedia content, and at least one candidate multimedia content is related to the first text.

[0063] This paper does not restrict the method of generating candidate multimedia content in one or more scenarios. In some scenarios, at least one candidate multimedia content can be generated locally (i.e., on the client side). For example, a multimedia content template can be pre-configured locally, and the first text can be added to the multimedia content template to generate candidate multimedia content. In other scenarios, at least one candidate multimedia content can also be generated in the cloud (i.e., on the server side). For example, a multimedia content template can be pre-configured in the cloud, and the first text can be added to the multimedia content template to generate candidate multimedia content. Alternatively, the cloud can connect to an artificial intelligence model (e.g., a multimodal large model) and send the first text to the artificial intelligence model to generate candidate multimedia content based on the first text.

[0064] The generation process of at least one candidate multimedia content can be understood as the process from the start of generation of at least one candidate multimedia content to the completion of generation. For example, when at least one candidate multimedia content includes multiple candidate multimedia content, the generation process of at least one candidate multimedia content can start from the first candidate multimedia content and continue until the last candidate multimedia content is generated.

[0065] In one or more of the cases described herein, the second operation may be unrelated to the generation process of at least one candidate multimedia content; that is, the second operation is not used to control the generation process of at least one candidate multimedia content. For example, the second operation is not an operation used to stop or cancel the generation of at least one candidate multimedia content.

[0066] This article does not restrict the way the second operation is received in one or more scenarios. For example, the second operation can be received in any interface provided by the content sharing platform. The interface for receiving the first operation and the interface for receiving the second operation can be the same or different.

[0067] Step S203: Present at least one candidate multimedia content.

[0068] In one or more scenarios described herein, at least one candidate multimedia content can be added to the first shared content, that is, at least one candidate multimedia content can be used as multimedia content in the first shared content, such as as the cover image of the first shared content; the first shared content can be published to a content sharing platform, for example, the first shared content can be published on the content sharing platform in the form of a post, discussion thread, etc., and other users can view the first shared content on the content sharing platform and interact with the first shared content, such as by commenting, liking, or collecting.

[0069] The present invention does not limit the manner in which at least one candidate multimedia content is presented. For example, at least one candidate multimedia content may be presented in any interface provided by the content sharing platform; or, for example, at least one candidate multimedia content may be presented in a region within any interface provided by the content sharing platform. Furthermore, presenting at least one candidate multimedia content may involve presenting all of the candidate multimedia content or presenting only a portion of the candidate multimedia content.

[0070] For example, at least one candidate multimedia content can be presented in the first display interface.

[0071] In this way, during the editing and sharing of content, the generation of candidate multimedia content is executed in parallel with other operations. Users can perform other operations while the candidate multimedia content is being generated, without having to wait for the candidate multimedia content to be generated before performing other operations. This reduces the time users spend waiting for the candidate multimedia content to be generated and improves the efficiency of content sharing and interaction.

[0072] Furthermore, in response to a triggering operation on the second multimedia content among at least one candidate multimedia content, the second multimedia content is presented in the first shared content.

[0073] The second multimedia content can be any one of the at least one candidate multimedia content, and the second multimedia content can be understood as the candidate multimedia content selected from at least one candidate multimedia content.

[0074] In other words, by presenting at least one candidate multimedia content, users can select a second multimedia content that meets their needs from the at least one candidate multimedia content and add the second multimedia content to the first shared content as multimedia content in the first shared content (e.g., the cover image in the first shared content).

[0075] In this way, at least one candidate multimedia content is automatically generated based on the first text associated with the first operation, and the user can directly add the second multimedia content that meets the requirements to the first shared content through a simple trigger operation, without the user having to manually edit the multimedia content in the first shared content, thus improving the efficiency of editing shared content.

[0076] Furthermore, in response to the publishing operation of the first shared content, the first displayed content is presented on the shared content presentation interface.

[0077] The publishing operation of the first shared content can be understood as the operation of publishing the first shared content to the content sharing platform. For example, if a publishing control is provided, the publishing operation of the first shared content can be the operation of triggering the publishing control.

[0078] The content sharing interface can be used to display content published by multiple publishers. In other words, the content sharing interface can be understood as an interface provided by the content sharing platform for viewing content shared by different users.

[0079] After the first share content is published, the first display content can be presented on the share content presentation interface. The first display content can be understood as the content related to the first share content displayed on the share content presentation interface. The first display content may include: the second multimedia content and some text content in the first share content.

[0080] For example, the content sharing interface can display the first content in a rectangular area, with the upper part of the rectangular area used to display the second multimedia content and the lower part of the rectangular area used to display a portion of the text content in the first content to be shared.

[0081] In this way, by selecting the most informative content from the first shared content to display, other users can quickly understand the message conveyed by the first shared content when viewing it on the content display interface, thus improving information acquisition efficiency.

[0082] In some cases, receiving a first operation includes receiving a first editing operation, which can be used to edit a first text, and the first shared content includes the first text.

[0083] In other words, the first operation can be the operation of editing the text content in the first shared content. For example, the content sharing platform can provide a first editing interface, which can be used to edit the first shared content, and the first editing operation can be triggered in the first editing interface; or, for example, the content sharing platform can also provide a separate text editing interface, which can be used only to edit the text content in the first shared content, and the first editing operation can be triggered in the text editing interface.

[0084] Since the first editing operation can be used to edit the text content in the first shared content (e.g., including the first text), and the candidate multimedia content of the first shared content can be related to the text content in the first shared content, the first editing operation can trigger the generation of candidate multimedia content of the first shared content. That is, in response to the triggered first editing operation, the generation of candidate multimedia content of the first shared content begins based on the first text.

[0085] Thus, after a user triggers the first editing operation and edits the text content in the first shared content, the process of generating candidate multimedia content is automatically started. Based on the first text carried by the first editing operation, the content sharing platform automatically generates candidate multimedia content for the first shared content. During the editing of the shared content, candidate multimedia content is generated and recommended to the user in a timely manner, thereby improving the efficiency of content sharing.

[0086] In some cases, the first shared content can be published to a content sharing platform, which can provide the first content. In this case, receiving the first operation includes receiving a selection operation, which can be used to select the first text from the first content.

[0087] For example, when the content sharing platform is a novel application, the first content can be a novel, and the first text can be a passage from that novel; as another example, when the content sharing platform is a knowledge application, the first content can be a knowledge article, and the first text can be a passage from that knowledge article.

[0088] In other words, the first operation can be the operation of selecting text from existing content provided by the content sharing platform. For example, when the first shared content edited by the user is related to the first text in the first content, the selection operation can be triggered to select the first text from the first content, indicating that the first shared content is related to the first text in the first content.

[0089] Since the first operation can be used to select the first text from the first content, the candidate multimedia content of the first shared content can be related to the first text selected in the first content. Therefore, the selection operation can generate candidate multimedia content of the first shared content. That is, in response to the selection operation, the generation of candidate multimedia content of the first shared content based on the first text begins.

[0090] For example, the candidate multimedia content of the first shared content can also be related to the first content. When the first content is a novel, the candidate multimedia content of the first shared content can be related to the attribute information of the first content, such as the novel category and the novel plot, to further enhance the relevance between the candidate multimedia content of the first shared content and the first text.

[0091] In this way, the user-triggered "select first text from first content" operation can be used to initiate the generation of candidate multimedia content. After the user triggers the selection operation and selects the first text from the first content, the generation process of candidate multimedia content is automatically started. Based on the first text indicated by the selection operation, the content sharing platform automatically generates candidate multimedia content for the first shared content. That is, the content sharing platform supports the quick operation of "selecting content to generate candidate multimedia content", without requiring the user to manually take a screenshot or copy the first text in the first content, improving the efficiency of users editing shared content related to the content provided by the content sharing platform, especially improving the efficiency of users editing multimedia content in shared content related to the content provided by the content sharing platform.

[0092] In some possible implementations, receiving a selection operation includes: in response to an add operation, presenting first content and receiving a selection operation.

[0093] In other words, users can first trigger the add operation to present the first content to the user, and then trigger the select operation to select the first text from the first content.

[0094] For example, presenting the first content includes presenting the first content in units of paragraphs. In this case, the selection operation can be a selection operation on one or more paragraphs in the first content.

[0095] For example, a content sharing platform can provide a content display interface, which can be used to display the content provided by the content sharing platform. In response to the triggered add operation, the content display interface presents the first content, and the user can also trigger the selection operation in the content display interface.

[0096] In this way, by providing users with primary content, users can intuitively and clearly understand the primary content, making it easier for them to select the primary text associated with the primary shared content. This improves the efficiency of users selecting the primary text (i.e., triggering the selection operation), and in turn, improves the efficiency of triggering the generation of candidate multimedia content for the primary shared content.

[0097] One or more scenarios described herein support triggering the add operation in different ways. For example, in response to the add operation, the first content is presented, including at least one of the following: in response to a first add operation in the first editing interface, the first content is presented; in response to a second add operation in the first editing interface, the first content is presented; in response to a third add operation in the first display interface, the first content is presented.

[0098] The first editing interface can be used to edit the first shared content. For example, the first editing interface can be a post editor interface. Users can edit various aspects of information related to the first shared content in the first editing interface, such as adding multimedia content, editing the title, editing the text content, adding tags, etc.

[0099] The first display interface can be used to present candidate multimedia content for the first sharing content. For example, the first display interface can be a candidate cover image display interface. Users can view candidate multimedia content in the first display interface, and users can also select the desired candidate multimedia content in the first display interface and add the selected candidate multimedia content to the sharing content, so that the selected candidate multimedia content becomes the multimedia content in the sharing content.

[0100] The first add operation can be used to associate the first shared content with the first content. The first add operation can be understood as establishing a relationship between the first shared content and the first content. When a user triggers the first add operation, it indicates that the first shared content is related to the first content. After being published, the first shared content can be categorized and presented in the corresponding section of the first content. For example, when the content sharing platform is a novel application, the first editing interface can provide a first add control (such as an "Add Original Text" control). The first add operation can be the operation of triggering the first add control. By triggering the first add control, the user associates the first shared content with the first content (such as a novel).

[0101] The second add operation can be used to edit the multimedia content in the first shared content. The second add operation can be understood as the operation of starting to edit the multimedia content in the first shared content and adding multimedia content to the first shared content. For example, the first editing interface can provide a second add control (such as an "add image or video" control). The second add operation can be the operation of triggering the second add control. By triggering the second add control, the user can edit the multimedia content in the first shared content, such as adding images or videos to the first shared content.

[0102] The third add operation can be used to edit the multimedia content in the first shared content. The third add operation can be understood as the operation of opening the editing of the multimedia content in the first shared content and adding multimedia content to the first shared content. For example, the first display interface can provide a third add control (such as the "Add Original Text to Image" control). The third add operation can be the operation of triggering the third add control. By triggering the third add control, the user can open the editing of the multimedia content in the first shared content again when viewing the candidate multimedia content, such as regenerating the candidate multimedia content of the first shared content.

[0103] In other words, from the perspective of the interface that triggers the add operation, the operation can be triggered in the first editing interface or the first display interface. From the perspective of the entry point for triggering the add operation, the operation can be triggered from the entry point of the content shared with the content sharing platform, or from the entry point of the multimedia content in the shared content.

[0104] In this way, during the process of editing and sharing content, different entry points are provided on different interfaces for users to trigger the addition operation, so as to present the first content to the user and then trigger the selection operation to generate candidate multimedia content of the first sharing content. Users can conveniently enable the "select content to generate candidate multimedia content" function at different stages of editing and sharing content, reducing the operation chain of the "select content to generate candidate multimedia content" function and improving the efficiency of content sharing.

[0105] In some cases, first information associated with at least one candidate multimedia content may also be presented, which may be used to indicate the generation progress of at least one candidate multimedia content.

[0106] For example, during the generation of at least one candidate multimedia content, first information associated with at least one candidate multimedia content is presented in the same interface as when the second operation is triggered. That is, the user can view the generation progress of at least one candidate multimedia content through the first information while triggering the second operation.

[0107] Since users can trigger other operations in parallel during the generation of at least one candidate multimedia content, in order to inform users of the generation progress of at least one candidate multimedia content, first information associated with at least one candidate multimedia content is provided to users so that users can perceive the generation progress of at least one candidate multimedia content in real time during the triggering of the second operation.

[0108] In this way, providing users with real-time progress updates on the generation of candidate multimedia content allows them to understand the current status of automatically generated candidate multimedia content while triggering the second operation, facilitating timely viewing of the candidate multimedia content and improving content sharing efficiency.

[0109] In some cases, presenting first information associated with at least one candidate multimedia content includes at least one of the following: presenting first progress information and presenting second progress information.

[0110] For example, the first progress information can be used to indicate the current percentage of completion of at least one candidate multimedia content, and the second progress information can be used to indicate the generation progress of each candidate multimedia content among at least one candidate multimedia content.

[0111] In other words, the first progress information can be used to describe the overall generation progress of at least one candidate multimedia content. For example, when the number of at least one candidate multimedia content is N, the first progress information can be "X / N pieces", where N is an integer greater than or equal to 1, and X is the number of candidate multimedia content that has been generated and is less than or equal to N. The second progress information can be used to describe the individual generation progress of each candidate multimedia content in the at least one candidate multimedia content. For example, the second progress information can be "60%", indicating that the generation progress of the candidate multimedia content is 60%.

[0112] This document does not restrict the manner in which the first progress information and the second progress information are presented. In some cases, the first progress information and the second progress information can be presented on the same interface, that is, the first progress information and the second progress information are provided to the user at the same time. In other cases, the first progress information and the second progress information can also be presented on different interfaces, for example, the first progress information is presented on the first editing interface and the second progress information is presented on the first display interface. In still other cases, the first progress information and the second progress information can also be presented in a switchable manner, for example, the first progress information is presented on the first editing interface and the second progress information is presented on the first display interface in response to a triggered switching operation (e.g., a triggering operation on the first progress information).

[0113] In this way, users are provided with two different levels of progress information, enabling them to understand the generation progress of at least one candidate multimedia content from different dimensions and comprehensively, making it easier for users to grasp the generation progress of at least one candidate multimedia content more accurately.

[0114] In some cases, during the generation of at least one candidate multimedia content, receiving a second operation includes: during the generation of at least one candidate multimedia content, receiving a second editing operation for editing the first shared content.

[0115] For example, the second editing operation can be one or more of the following: editing the text content in the first shared content, associating the first content with the first shared content, adding tags, or adding other multimedia content.

[0116] For example, in the generation of at least one candidate multimedia content, a second editing operation is received in the first editing interface. In this case, the first information associated with the at least one candidate multimedia content (e.g., first progress information) can also be presented in the first editing interface.

[0117] In other words, during the generation of at least one candidate multimedia content, the user can simultaneously edit the first shareable content, thus achieving "editing the shareable content while waiting for the candidate multimedia content to be generated".

[0118] In this way, during the generation of at least one candidate multimedia content, users do not need to wait for the candidate multimedia content to be generated. Instead, they can simultaneously trigger the second editing operation to edit the first sharing content, reducing unnecessary waiting time during the editing and sharing process and effectively improving the efficiency of editing and sharing content.

[0119] It should be noted that one or more embodiments in this document do not limit the second operation. That is, the second operation can be any operation triggered in the client of the content sharing platform that is unrelated to the generation process of at least one candidate multimedia content. For example, the second operation can also be an operation to view other shared content, an operation to interact with other shared content, etc.

[0120] Furthermore, if the second operation is a second editing operation, a third operation can also be received. In response to the third operation meeting the set conditions, a generation result is presented, which may include the candidate multimedia content generated when the third operation is received.

[0121] The third operation can be understood as an operation different from the second editing operation. For example, the third operation can be triggered during the generation of at least one candidate multimedia content, or it can be triggered after the generation of at least one candidate multimedia content is completed.

[0122] The third operation can be triggered on the same interface as the second editing operation. For example, if the second editing operation is received in the first editing interface, the third operation can also be received in the first editing interface.

[0123] Setting conditions can be understood as describing the conditions under which a third operation meets the requirements for providing a generated result. In other words, setting conditions can be used to determine whether to provide a generated result based on a third operation. For example, setting conditions can be used to describe operations that need to provide a generated result to the user.

[0124] The candidate multimedia content generated upon receiving the third operation can be understood as at least one candidate multimedia content that has been generated at the moment the third operation is triggered.

[0125] In other words, pre-configured conditions are set, and for the received third operation, it is determined whether the third operation meets the set conditions. When the third operation meets the set conditions, the generated result is provided to the user.

[0126] In this way, the current generation result is presented to the user at the appropriate time (i.e. when the third operation triggered by the user meets the set conditions), which intuitively informs the user of the candidate multimedia content that has been generated. The user can decide on the subsequent operation based on the candidate multimedia content that has been generated (such as selecting one or more of them as multimedia content in the first sharing content or continuing to trigger the second editing operation). This improves the flexibility of user interaction without affecting the generation process of at least one candidate multimedia content and the process of the user editing the first sharing content.

[0127] In some cases, the third operation that meets the set conditions includes at least one of the following: collapse operation and publish operation.

[0128] The collapse operation can be used to collapse the information input area, and the second editing operation is triggered in the information input area. For example, the information input area can be a keyboard area, a voice input area, a handwriting input area, etc.

[0129] In other words, when a user triggers a second editing operation in the information input area to edit the first shared content, and the user triggers a collapse operation to collapse the information input area, it indicates that the user may have completed editing the first shared content. In this case, the generated result is presented to the user, informing the user of the candidate multimedia content that has been generated so that the user can select and enrich the multimedia content in the first shared content.

[0130] The publish operation can be used to publish the first shared content to the content sharing platform. For example, the publish operation can be an operation that triggers the publish control. For example, when the second operation is a second editing operation triggered by the first editing interface, the first editing interface can provide a publish control, and the publish operation can be an operation that triggers the publish control of the first editing interface.

[0131] In other words, when a user triggers a second editing operation to edit the first shared content, and then triggers a publishing operation, it indicates that the user intends to publish the first shared content. In this case, the generated results are presented to the user, informing them of the candidate multimedia content that has been generated so that the user can choose to enrich the multimedia content in the first shared content.

[0132] In this way, by combining the actual operation configuration settings that users may trigger during the editing and sharing of content, the generated results can be presented to users at appropriate times. When users may need to view the generated candidate multimedia content, recommendations can be made to users in a timely manner, thereby improving the interactive experience of users in the process of editing and sharing content.

[0133] Furthermore, at least one candidate multimedia content includes the first multimedia content. Before the first multimedia content is generated, the generation result does not include the first multimedia content. In response to the completion of the generation of the first multimedia content, the first multimedia content is presented in the generation result.

[0134] The first multimedia content can be any candidate multimedia content among at least one multimedia content. The fact that the generated result does not include the first multimedia content can be understood as the first multimedia content not being fully generated when the third operation is triggered.

[0135] Since the first multimedia content has not yet been fully generated when the third operation is triggered, the generated result will not include the first multimedia content. When the first multimedia content is fully generated later, the generated result will be updated to include the first multimedia content, and the first multimedia content will be presented in the generated result.

[0136] In this way, by combining the actual generation of candidate multimedia content, the generation results are updated in real time, so that the generation results presented to the user match the actual generation situation, thereby improving the real-time performance and accuracy of the generation results.

[0137] In some cases, in response to a third operation satisfying a set condition, a generation result is presented, including: in response to a third operation satisfying a set condition, the generation result is presented on a first editing interface. In this case, at least one candidate multimedia content is presented, including: in response to a trigger operation on the generation result, at least one candidate multimedia content is presented on a first display interface.

[0138] In other words, users can jump to the first display interface by triggering the generated result of the first editing interface, and view at least one candidate multimedia content displayed in the first display interface.

[0139] For example, considering that at least one candidate multimedia content may not have been fully generated yet, that is, the generation result may only include some candidate multimedia content, in response to the triggering operation of the generation result, the candidate multimedia content included in the generation result is presented on the first display interface. For the candidate multimedia content that has not yet been fully generated, a second progress information can be presented on the first display interface to inform the generation progress of the candidate multimedia content that has not yet been fully generated.

[0140] In this way, users can jump directly from the generated results on the first editing interface to the first display interface. After informing the user of the currently generated candidate multimedia content, the user can jump to the first display interface to view the candidate multimedia content in detail, so as to select from the candidate multimedia content, simplifying the user's operation process and improving the efficiency and convenience of content sharing.

[0141] The following example uses a content sharing platform for novels, combined with... Figures 3A to 3M The interactive pages involved in the above interaction methods are illustrated below.

[0142] Figures 3A to 3M The illustration shows a page diagram provided in at least one of the scenarios described herein.

[0143] like Figure 3A As shown, Figure 3A The text describes a reading interface 30 provided by a content sharing platform. Users can read the first content (e.g., a novel) provided by the platform in the reading interface 30. The reading interface 30 displays the novel content and discussion posts, such as discussion post A and discussion post B. Discussion post A and discussion post B can be discussion posts related to the first content.

[0144] The reading interface 30 also presents a primary entry point 301, such as a "post a discussion" control. Users can trigger the primary entry point 301 to begin the editing process of the first shared content. For example, the first shared content can be related to the first topic. In addition, the reading interface 30 can also present a "request for updates" control and a "send a gift" control.

[0145] like Figure 3B As shown, Figure 3B The text editing interface 31 provided by the content sharing platform is shown. The text editing interface 31 can be used only to edit the text content in the first shared content, for example... Figure 3B The first text 311 in the text is “#Prove you've read this book in one sentence, highly recommended!” At the same time, the text editing interface 31 also provides a next step control 312. The first editing operation can be the operation that triggers the next step control 312. That is, triggering the next step control 312 can trigger the generation of candidate multimedia content for the first sharing content.

[0146] like Figure 3C As shown, Figure 3C The first editing interface 32 provided by the content sharing platform is shown. The first editing interface 32 can be used to edit the first shared content, for example, because... Figure 3B The first editing operation in the interface 32 is used to edit the first text 311. Therefore, the first shared content includes the first text 311. At the same time, users can add images / videos, enter titles, continue to edit text content based on the first text 311, add tags, associate with the first content, add books / shows, etc.

[0147] In addition, the first editing interface 32 can also present first progress information 321. For example, if at least one candidate multimedia content includes four candidate cover images, the first progress information 321 can be "Cover generation in progress 0 / 4 images", indicating the current percentage of completion of at least one candidate multimedia content.

[0148] like Figure 3D As shown, during the generation of at least one candidate multimedia content (e.g., the first progress information 321 is generated by...), Figure 3C The "Cover generation in progress 0 / 4 images" message has been updated to... Figure 3D In the "Cover generation in progress 1 / 4" section, a second operation is received. The second operation can be a second editing operation triggered by the user in the first editing interface 32. For example, the second editing operation can be used to edit the text content in the first shared content. The text content in the first shared content can be based on the first text, with the addition of "Character A's personality...".

[0149] The first editing interface 32 includes a publishing control 322 and an information input area 323. The information input area 323 provides a collapse control 3231. The user can trigger a second editing operation through the information input area 323 to input text content from the first shared content. In this case, a third operation that meets the set conditions can be either triggering the publishing control 322 (i.e., publishing operation) or triggering the collapse control 3231 (i.e., collapsing operation).

[0150] like Figure 3E As shown, a third operation is received, and in response to the third operation meeting the set conditions, the generation result 324 is presented on the first editing interface 32. Since the candidate multimedia content that has been generated at the time the third operation is triggered includes candidate cover image A and candidate cover image B, the generation result 324 includes candidate cover image A and candidate cover image B.

[0151] like Figure 3F As shown, in response to the first multimedia content (i.e. Figure 3F The candidate cover image C) is generated. The generation result 324 is updated. In the updated generation result 324, the first multimedia content (i.e. Figure 3F Candidate cover image C).

[0152] like Figure 3G As shown, Figure 3G The first display interface 33 provided by the content sharing platform is shown, for example, in response to... Figure 3C or Figure 3D The first progress information 321 in the process is triggered, and the first display interface 33 is presented; for example, in response to the... Figure 3E or Figure 3F The trigger operation of the generated result 324 in the middle will present the first display interface.

[0153] The first display interface 33 can present at least one candidate multimedia content (e.g., is...) Figure 3G The candidate cover images (A to C) are shown in the first display interface. Since the fourth candidate multimedia content has not yet been fully generated, a second progress information 331, such as "50%", can also be displayed on the first display interface 33 to indicate the generation progress of the fourth candidate multimedia content. In addition, the first display interface 33 also displays a multimedia content adding control 332.

[0154] like Figure 3H As shown, with the fourth candidate multimedia content (i.e. Figure 3H Once the candidate cover image (D) is generated, the first display interface 33 presents at least one candidate multimedia content (e.g., a...). Figure 3H (Candidate cover images A to D) In ​​the list, users can select a second multimedia content from at least one candidate multimedia content (i.e., Figure 3H The candidate cover image (B) triggers the multimedia content addition control 332 to add the second multimedia content to the first shared content.

[0155] like Figure 3I As shown, the first shared content displayed on the first editing interface 32 includes second multimedia content 325, that is, the second multimedia content 325 serves as multimedia content in the first shared content, such as the cover image of the first shared content.

[0156] like Figure 3J As shown, content sharing platforms, in addition to supporting [the following], [are also supported by] [other entities]. Figure 3B The text editing interface 31 can be accessed to enter the first editing interface 32, or it can be directly accessed to the first editing interface 32. Figure 3J In the first editing interface 32, the first shared content is empty. The first editing interface 32 provides a first add control 326 and a second add control 327. The user can trigger the first add control 326 to associate the first shared content with the first content. The user can trigger the second add control 327 to add multimedia content to the first shared content.

[0157] like Figure 3K As shown, the first display interface 33 provides a third add control 333. Users can trigger the third add control 333 to generate candidate multimedia content for the first shareable content and add multimedia content to the first shareable content.

[0158] The above Figure 3J The first added control 326 and the second added control 327 and Figure 3K The third add control 333 can trigger an add operation, and in response to the triggered add operation, the first content is displayed.

[0159] like Figure 3L As shown, Figure 3L The content display interface 34 provided by the content sharing platform is shown, for example, triggering... Figure 3J After adding the first control 326 in the middle, it is displayed Figure 3L The content display interface 34 presents the first content. It also provides controls 341 for adding the original text as text and 342 for adding the original text as an image. Users can select the first text (e.g., ...) on the content display interface 34. Figure 3L (From the novel content B), then trigger the control 342 to add the original text as an image, thereby triggering the selection operation and opening up the candidate multimedia content for generating the first shareable content.

[0160] like Figure 3M As shown, trigger Figure 3J The second added control 327 or Figure 3K After adding the third control 333, it will be displayed. Figure 3M The content display interface 34 in the middle, due to Figure 3J The second added control 327 and Figure 3K The third add control 333 is used to edit the multimedia content in the first shared content, therefore... Figure 3M The content display interface 34 can be an interface under the "Original Text" tag. Through the simultaneously provided tags "My Generation" and "Album," users can select multimedia content from different sources to add to the first shared content. In this case, the content display interface 34 only provides a control 342 for adding the original text as an image. Users can select the first text (e.g., ...) on the content display interface 34. Figure 3M (From the novel content B), then trigger the control 342 to add the original text as an image, thereby triggering the selection operation and opening up the candidate multimedia content for generating the first shareable content.

[0161] Based on the interaction method provided in at least one aspect of this paper, an interaction device is also provided in at least one aspect of this paper. The following will combine... Figure 4 Provide a detailed description of the interactive device.

[0162] Figure 4 The schematic diagram illustrates the structure of an interactive device provided in at least one of the present invention.

[0163] like Figure 4As shown, the interactive device 400 in this scenario includes a first receiving module 401, a second receiving module 402, and a presentation module 403. For example, the first receiving module 401, the second receiving module 402, and the presentation module 403 can be implemented using hardware (e.g., circuit) modules or software modules, as in the following cases, and will not be repeated here. For example, the first receiving module 401, the second receiving module 402, and the presentation module 403 can be implemented using a central processing unit (CPU), a general-purpose graphics processor (GPGPU), a graphics processing unit (GPU), a tensor processor (TPU), a field-programmable gate array (FPGA), or other processing units with data processing capabilities and / or instruction execution capabilities, along with corresponding computer instructions.

[0164] The first receiving module 401 is configured to receive a first operation, wherein the first operation is used to generate candidate multimedia content for the first shared content, and the first operation is associated with the first text. For example, the first receiving module 401 can be configured to execute step S201 described above; its specific implementation principle can be found in the relevant description of step S201, and will not be repeated here.

[0165] The second receiving module 402 is configured to receive a second operation during the generation process of at least one candidate multimedia content, wherein the at least one candidate multimedia content is generated based on the first text, and the second operation is unrelated to the generation process of the at least one candidate multimedia content. For example, the second receiving module 402 can be configured to execute step S202 described above; its specific implementation principle can be found in the relevant description of step S202, and will not be repeated here.

[0166] The presentation module 403 is configured to present the at least one candidate multimedia content, wherein the at least one candidate multimedia content is used to add to the first shared content. For example, the presentation module 403 can be configured to execute step S203 described above; its specific implementation principle can be found in the relevant description of step S203, and will not be repeated here.

[0167] In at least one of the embodiments described herein, the first receiving module 401 is further configured to: receive a first editing operation, wherein the first editing operation is used to edit the first text, and the first shared content includes the first text.

[0168] In at least one of the embodiments described herein, the first shared content is used to be published to a content sharing platform, the content sharing platform providing the first content, and the first receiving module 401 is further configured to: receive a selection operation, wherein the selection operation is used to select the first text from the first content.

[0169] In at least one of the embodiments described herein, the first receiving module 401 is further configured to: in response to an add operation, present the first content and receive the selection operation.

[0170] In at least one embodiment of this document, the first receiving module 401 is further configured to: present the first content in response to a first add operation in a first editing interface, wherein the first editing interface is used to edit the first shared content, and the first add operation is used to associate the first shared content with the first content; present the first content in response to a second add operation in the first editing interface, wherein the second add operation is used to edit multimedia content in the first shared content; and present the first content in response to a third add operation in a first display interface, wherein the first display interface is used to display candidate multimedia content of the first shared content, and the third add operation is used to edit multimedia content in the first shared content.

[0171] In at least one of the embodiments described herein, the presentation module 403 is further configured to present first information associated with the at least one candidate multimedia content, wherein the first information is used to indicate the generation progress of the at least one candidate multimedia content.

[0172] In at least one of the embodiments described herein, the presentation module 403 is further configured to: present first progress information, wherein the first progress information is used to indicate the current percentage of completion of the at least one candidate multimedia content; and present second progress information, wherein the second progress information is used to indicate the generation progress of each candidate multimedia content among the at least one candidate multimedia content.

[0173] In at least one of the embodiments described herein, the second receiving module 402 is further configured to: receive a second editing operation during the generation of the at least one candidate multimedia content, wherein the second editing operation is used to edit the first shared content.

[0174] In at least one of the embodiments described herein, the second receiving module 402 is further configured to receive a third operation; the presentation module 403 is further configured to present a generation result in response to the third operation satisfying a set condition, wherein the generation result includes the candidate multimedia content generated when the third operation is received.

[0175] In at least one of the scenarios described herein, the third operation satisfying the set conditions includes at least one of the following: a collapse operation, wherein the collapse operation is used to collapse the information input area, and the second editing operation is triggered in the information input area; and a publishing operation, wherein the publishing operation is used to publish the first shared content to a content sharing platform.

[0176] In at least one of the embodiments described herein, the at least one candidate multimedia content includes a first multimedia content, and the generation result does not include the first multimedia content before the first multimedia content is generated. The presentation module 403 is further configured to: in response to the completion of the generation of the first multimedia content, present the first multimedia content in the generation result.

[0177] In at least one of the embodiments described herein, the presentation module 403 is further configured to: in response to the third operation satisfying the set conditions, present the generation result on a first editing interface, wherein the first editing interface is used to edit the first shared content; and in response to a triggering operation on the generation result, present the at least one candidate multimedia content on a first display interface.

[0178] In at least one of the embodiments described herein, the presentation module 403 is further configured to: in response to a triggering operation on the second multimedia content among the at least one candidate multimedia content, present the second multimedia content in the first shared content.

[0179] In at least one of the embodiments described herein, the presentation module 403 is further configured to: in response to a publishing operation on the first shared content, present first display content on a shared content presentation interface; wherein the first display content includes: the second multimedia content and a portion of the text content in the first shared content, and the shared content presentation interface is used to present display content published by multiple publishers.

[0180] In at least one of the scenarios described herein, the first shared content is not document content.

[0181] It should be noted that, for clarity and brevity, not all components of the interactive device 400 are shown in at least one of the embodiments herein. To achieve the necessary functions of the interactive device 400, those skilled in the art may provide or configure other components (not shown) according to specific needs, and the embodiments herein do not impose any limitations on this.

[0182] The interactive device 400 provided in at least one aspect of this document and the interactive method provided in at least one aspect of this document are based on the same inventive concept and can achieve the same technical effect and the same technical purpose as the interactive method provided in at least one aspect of this document. For details, please refer to the relevant description above, which will not be repeated here.

[0183] This document also provides an electronic device, including a processing device and a storage device, the storage device including one or more computer program modules; wherein the one or more computer program modules are stored in the storage device and configured to be executed by the processing device, the one or more computer program modules being used to implement the interactive method provided in any of the present invention.

[0184] For example, the processing device may be a processor, such as a central processing unit (CPU), digital signal processor (DSP), image processor (GPU), general-purpose graphics processor (GPGPU), or other form of processing unit with data processing capabilities and / or instruction execution capabilities. It may be a general-purpose processor or a dedicated processor and may control other components in the electronic device to perform the desired functions.

[0185] For example, the storage device may be a memory, which may include one or more computer program products. These computer program products may include various forms of computer-readable storage media, such as volatile memory and / or non-volatile memory. The volatile memory may, for example, include random access memory (RAM) and / or cache memory. The non-volatile memory may, for example, include read-only memory (ROM), hard disk, flash memory, etc. One or more computer program instructions may be stored on the computer-readable storage medium, and a processing device may execute these program instructions to implement the functions described in at least one of the embodiments herein (implemented by the processing device) and / or other desired functions. Various application programs and various data may also be stored on the computer-readable storage medium, which is not limited in the embodiments described herein.

[0186] The following is for reference. Figure 5 The diagram illustrates a structural schematic of an electronic device (e.g., a terminal device or a server) 500 suitable for implementing at least one of the embodiments described herein. The terminal device in at least one embodiment may include, but is not limited to, mobile terminals such as mobile phones, laptops, digital radio receivers, personal digital assistants (PDAs), tablet computers (PADs), portable multimedia players (PMPs), in-vehicle terminals (e.g., in-vehicle navigation terminals), and fixed terminals such as digital televisions and desktop computers. Figure 5 The electronic device shown is merely an example and should not impose any limitation on the functionality and scope of use of at least one of the situations described herein.

[0187] like Figure 5As shown, electronic device 500 may include a processing unit (e.g., a central processing unit, a graphics processing unit, etc.) 501, which can perform various appropriate actions and processes according to a program stored in read-only memory (ROM) 502 or a program loaded from storage device 508 into random access memory (RAM) 503. RAM 503 also stores various programs and data required for the operation of electronic device 500. Processing unit 501, ROM 502, and RAM 503 are interconnected via bus 504. Input / output (I / O) interface 505 is also connected to bus 504.

[0188] Typically, the following devices can be connected to I / O interface 505: input devices 506 including, for example, touchscreens, touchpads, keyboards, mice, cameras, microphones, accelerometers, gyroscopes, etc.; output devices 507 including, for example, liquid crystal displays (LCDs), speakers, vibrators, etc.; storage devices 508 including, for example, magnetic tapes, hard disks, etc.; and communication devices 509. Communication device 509 allows electronic device 500 to communicate wirelessly or wiredly with other devices to exchange data. Although Figure 5 An electronic device 500 with various devices is shown; however, it should be understood that it is not required to implement or possess all of the devices shown. More or fewer devices may be implemented or possessed alternatively.

[0189] In particular, according to one or more embodiments herein, the processes described in the above-referenced flowcharts can be implemented as computer software programs. For example, one or more embodiments herein include a computer program product comprising a computer program carried on a non-transitory computer-readable medium, the computer program containing program code for performing the methods shown in the flowcharts. In such an embodiment, the computer program can be downloaded and installed from a network via communication device 509, or installed from storage device 508, or installed from ROM 502. When the computer program is executed by processing device 501, it performs the functions defined in the methods of at least one embodiment herein.

[0190] The electronic device 500 provided in at least one aspect of this article and the interaction method provided in at least one aspect of this article are based on the same inventive concept and can achieve the same technical effect and the same technical purpose as the interaction method provided in at least one aspect of this article. For details, please refer to the relevant description above, which will not be repeated here.

[0191] It should be noted that the computer-readable medium described above can be a computer-readable signal medium, a computer-readable storage medium, or any combination thereof. A computer-readable storage medium can be, for example,—but not limited to—an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination thereof. More specific examples of a computer-readable storage medium may include, but are not limited to: an electrical connection having one or more wires, a portable computer disk, a hard disk, random access memory (RAM), read-only memory (ROM), erasable programmable read-only memory (EPROM or flash memory), optical fiber, portable compact disk read-only memory (CD-ROM), optical storage device, magnetic storage device, or any suitable combination thereof. In this document, a computer-readable storage medium can be any tangible medium containing or storing a program that can be used by or in conjunction with an instruction execution system, apparatus, or device. In this document, a computer-readable signal medium can include a data signal propagated in baseband or as part of a carrier wave, carrying computer-readable program code. Such propagated data signals can take various forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination thereof. A computer-readable signal medium may be any computer-readable medium other than a computer-readable storage medium, which can send, propagate, or transmit a program for use by or in connection with an instruction execution system, apparatus, or device. The program code contained on the computer-readable medium can be transmitted using any suitable medium, including but not limited to: wires, optical fibers, radio frequency (RF), etc., or any suitable combination thereof.

[0192] The computer-readable storage medium provided in at least one aspect of this document and the interaction method provided in at least one aspect of this document are based on the same inventive concept and can achieve the same technical effect and the same technical purpose as the interaction method provided in at least one aspect of this document. For details, please refer to the relevant descriptions above, which will not be repeated here.

[0193] In some implementations, clients and servers can communicate using any currently known or future-developed network protocol, such as the Hypertext Transfer Protocol (HTTP), and can interconnect with digital data communication (e.g., communication networks) of any form or medium. Examples of communication networks include local area networks (LANs), wide area networks (WANs), the Internet (e.g., the Internet), and peer-to-peer networks (e.g., ad hoc peer-to-peer networks), as well as any currently known or future-developed networks.

[0194] The aforementioned computer-readable medium may be included in the aforementioned electronic device; or it may exist independently and not assembled into the electronic device.

[0195] The aforementioned computer-readable medium carries one or more programs, which, when executed by the electronic device, cause the electronic device to perform the aforementioned interactive method.

[0196] Computer program code for performing the operations described herein may be written in one or more programming languages ​​or a combination thereof, including but not limited to object-oriented programming languages ​​such as Java, Smalltalk, and C++, as well as conventional procedural programming languages ​​such as the "C" language or similar programming languages. The program code may execute entirely on the user's computer, partially on the user's computer, as a standalone software package, partially on the user's computer and partially on a remote computer, or entirely on a remote computer or server. In cases involving remote computers, the remote computer may be connected to the user's computer via any type of network—including a local area network (LAN) or a wide area network (WAN)—or may be connected to an external computer (e.g., via the Internet using an Internet service provider).

[0197] One or more embodiments of this document also provide a computer program product comprising one or more computer instructions. When these computer instructions are loaded and executed on a computing device, all or part of the processes or functions described in any of these embodiments are generated.

[0198] The computer instructions may be stored in a computer-readable storage medium or transmitted from one computer-readable storage medium to another. For example, the computer instructions may be transmitted from one website, computer, or data center to another website, computer, or data center via wired (e.g., coaxial cable, fiber optic, digital subscriber line (DSL)) or wireless (e.g., infrared, wireless, microwave, etc.) means.

[0199] When the computer program product is executed by a computer, the computer performs any of the aforementioned interactive methods. The computer program product can be a software installation package; when any of the aforementioned interactive methods is required, the computer program product can be downloaded and executed on the computer.

[0200] The computer program product provided in at least one of the embodiments described herein and the interaction method provided in at least one of the embodiments described herein are based on the same inventive concept and can achieve the same technical effect and the same technical purpose as the interaction method provided in at least one of the embodiments described herein. For details, please refer to the relevant descriptions above, which will not be repeated here.

[0201] The flowcharts and block diagrams in the accompanying figures illustrate the architecture, functionality, and operation of possible implementations of the systems, methods, and computer program products according to the various scenarios described herein. In this respect, each block in a flowchart or block diagram may represent a module, segment, or portion of code containing one or more executable instructions for implementing the specified logical function. It should also be noted that in some alternative implementations, the functions indicated in the blocks may occur in a different order than those indicated in the figures. For example, two consecutively indicated blocks may actually be executed substantially in parallel, and they may sometimes be executed in reverse order, depending on the functions involved. It should also be noted that each block in the block diagrams and / or flowcharts, and combinations of blocks in the block diagrams and / or flowcharts, may be implemented using a dedicated hardware-based system that performs the specified function or operation, or using a combination of dedicated hardware and computer instructions.

[0202] The units or modules described in at least one of the scenarios herein can be implemented in software or hardware. The names of the units or modules do not, in some cases, constitute a limitation on the unit or module itself.

[0203] The functions described above in this document can be performed at least in part by one or more hardware logic components. For example, exemplary types of hardware logic components that can be used, without limitation, include: field-programmable gate arrays (FPGAs), application-specific integrated circuits (ASICs), application-specific standard products (ASSPs), system-on-a-chip (SoCs), complex programmable logic devices (CPLDs), and so on.

[0204] In the context of this document, a machine-readable medium can be a tangible medium that may contain or store a program for use by or in conjunction with an instruction execution system, apparatus, or device. A machine-readable medium can be a machine-readable signal medium or a machine-readable storage medium. A machine-readable medium can be, but is not limited to, electronic, magnetic, optical, electromagnetic, infrared, or semiconductor systems, apparatus, or devices, or any suitable combination of the foregoing. More specific examples of machine-readable storage media include electrical connections based on one or more wires, portable computer disks, hard disks, random access memory (RAM), read-only memory (ROM), erasable programmable read-only memory (EPROM or flash memory), optical fiber, portable compact disk read-only memory (CD-ROM), optical storage devices, magnetic storage devices, or any suitable combination of the foregoing.

[0205] Based on one or more scenarios described in this article, Example 1 provides an interaction method, including:

[0206] Receive a first operation, wherein the first operation is used to generate candidate multimedia content for the first shared content, and the first operation is associated with the first text;

[0207] During the generation of at least one candidate multimedia content, a second operation is received, wherein the at least one candidate multimedia content is generated based on the first text, and the second operation is independent of the generation process of the at least one candidate multimedia content;

[0208] Present the at least one candidate multimedia content, wherein the at least one candidate multimedia content is used to add to the first shared content.

[0209] Based on one or more scenarios described in this article, Example 2 provides the receiving first operation from Example 1, including:

[0210] Receive a first editing operation, wherein the first editing operation is used to edit the first text, and the first shared content includes the first text.

[0211] According to one or more scenarios described herein, Example 3 provides the first shared content from Example 1 for publication to a content sharing platform. The content sharing platform provides the first content, and the receiving of the first operation includes:

[0212] A selection operation is received, wherein the selection operation is used to select the first text from the first content.

[0213] Based on one or more scenarios in this paper, Example 4 provides the receive selection operation from Example 3, including:

[0214] In response to the add operation, the first content is presented, and the selection operation is received.

[0215] Based on one or more scenarios described herein, Example 5 provides the response to the add operation in Example 4, presenting the first content, including at least one of the following:

[0216] In response to a first add operation in the first editing interface, the first content is presented, wherein the first editing interface is used to edit the first shared content, and the first add operation is used to associate the first shared content with the first content;

[0217] In response to the second add operation in the first editing interface, the first content is presented, wherein the second add operation is used to edit the multimedia content in the first shared content;

[0218] In response to a third add operation in the first display interface, the first content is presented, wherein the first display interface is used to present candidate multimedia content of the first shared content, and the third add operation is used to edit the multimedia content in the first shared content.

[0219] Depending on one or more scenarios described herein, Example 6 provides a method from any of the examples in Examples 1 through 5, and also includes:

[0220] Present first information associated with the at least one candidate multimedia content, wherein the first information is used to indicate the generation progress of the at least one candidate multimedia content.

[0221] According to one or more scenarios described herein, Example 7 provides first information associated with the presentation in Example 6 and the at least one candidate multimedia content, including at least one of the following:

[0222] Presenting first progress information, wherein the first progress information is used to indicate the current percentage of completion of the generation of the at least one candidate multimedia content;

[0223] A second progress information is presented, wherein the second progress information is used to indicate the generation progress of each candidate multimedia content among the at least one candidate multimedia content.

[0224] According to one or more scenarios described herein, Example 8 provides a second operation received during the generation of at least one candidate multimedia content in any of Examples 1 to 5, as described in Example 1 through Example 5, including:

[0225] During the generation of the at least one candidate multimedia content, a second editing operation is received, wherein the second editing operation is used to edit the first shared content.

[0226] Depending on one or more scenarios described in this article, Example 9 provides the method from Example 8, and also includes:

[0227] Receive third operation;

[0228] In response to the third operation satisfying the set conditions, a generation result is presented, wherein the generation result includes the candidate multimedia content generated when the third operation is received.

[0229] According to one or more scenarios described herein, Example 10 provides that the third operation satisfying the set conditions in Example 9 includes at least one of the following:

[0230] A collapse operation, wherein the collapse operation is used to collapse the information input area, and the second editing operation is triggered in the information input area;

[0231] The publishing operation is used to publish the first shared content to the content sharing platform.

[0232] According to one or more scenarios described herein, Example 11 provides at least one candidate multimedia content from Example 9 that includes first multimedia content. Before the first multimedia content is generated, the generation result does not include the first multimedia content. The method further includes:

[0233] In response to the completion of the generation of the first multimedia content, the first multimedia content is presented in the generation result.

[0234] Based on one or more scenarios described herein, Example Twelve provides the response from Example Nine where the third operation satisfies a set condition, presenting a generated result, including:

[0235] In response to the third operation satisfying the set conditions, the generation result is presented on the first editing interface, wherein the first editing interface is used to edit the first shared content; and

[0236] Presenting the at least one candidate multimedia content includes:

[0237] In response to a triggering operation on the generated result, the at least one candidate multimedia content is presented on the first display interface.

[0238] Depending on one or more scenarios described herein, Example Thirteen provides a method from any of the examples in Examples One through Five, and also includes:

[0239] In response to a triggering operation on the second multimedia content among the at least one candidate multimedia content, the second multimedia content is presented in the first shared content.

[0240] Depending on one or more scenarios described in this article, Example Fifteen provides the method from Example Thirteen, and also includes:

[0241] In response to the publishing operation of the first shared content, the first display content is presented on the shared content presentation interface;

[0242] The first display content includes: the second multimedia content and a portion of the text content in the first shared content, and the shared content presentation interface is used to present display content published by multiple publishers.

[0243] According to one or more scenarios in this article, Example 15 provides that the first shared content in any of the examples from Example 1 to Example 5 is not document content.

[0244] According to one or more of the scenarios described herein, Example Sixteen provides an interactive device comprising:

[0245] The first receiving module is configured to receive a first operation, wherein the first operation is used to generate candidate multimedia content for the first shared content, and the first operation is associated with the first text.

[0246] The second receiving module is configured to receive a second operation during the generation process of at least one candidate multimedia content, wherein the at least one candidate multimedia content is generated based on the first text, and the second operation is unrelated to the generation process of the at least one candidate multimedia content.

[0247] The presentation module is configured to present the at least one candidate multimedia content, wherein the at least one candidate multimedia content is used to add to the first shared content.

[0248] According to one or more of the provisions of this article, Example Seventeen provides an electronic device comprising:

[0249] At least one processor; and

[0250] At least one memory, including one or more computer program instructions;

[0251] The one or more computer program instructions are executed by the processor at least one of the interactive methods provided herein.

[0252] According to one or more of the present invention, Example 18 provides a computer-readable storage medium that non-transitory stores computer-readable instructions, wherein the interaction method provided by at least one of the present invention is implemented when the computer-readable instructions are executed by a processor.

[0253] According to one or more of the embodiments described herein, Example Nineteen provides a computer program product including a computer program that, when executed by a processor, implements the interaction method provided in at least one of the embodiments described herein.

[0254] The above description is merely a preferred embodiment and an explanation of the technical principles employed. Those skilled in the art should understand that the scope of disclosure herein is not limited to technical solutions formed by specific combinations of the above-described technical features, but also includes other technical solutions formed by arbitrary combinations of the above-described technical features or their equivalents without departing from the above-disclosed concept. For example, technical solutions formed by substituting the above features with (but not limited to) technical features disclosed herein that have similar functions.

[0255] Furthermore, while the operations are described in a specific order, this should not be construed as requiring these operations to be performed in the specific order shown or in a sequential order. In certain contexts, multitasking and parallel processing may be advantageous. Similarly, while some specific implementation details are included in the above discussion, these should not be interpreted as limiting the scope of this paper. Certain features described in the context of a single case can also be implemented in combination within that single case. Conversely, various features described in the context of a single case can also be implemented individually or in any suitable sub-combination in multiple cases.

[0256] Although the subject matter has been described using language specific to structural features and / or methodological logic, it should be understood that the subject matter defined in the appended claims is not necessarily limited to the specific features or actions described above. Rather, the specific features and actions described above are merely illustrative examples of implementing the claims.

Claims

1. An interaction method, comprising: Receive a first operation, wherein the first operation is used to generate candidate multimedia content for the first shared content, and the first operation is associated with the first text; During the generation of at least one candidate multimedia content, a second operation is received, wherein the at least one candidate multimedia content is generated based on the first text, and the second operation is independent of the generation process of the at least one candidate multimedia content; Present the at least one candidate multimedia content, wherein the at least one candidate multimedia content is used to add to the first shared content.

2. The method according to claim 1, wherein, The receiving of the first operation includes: Receive a first editing operation, wherein the first editing operation is used to edit the first text, and the first shared content includes the first text.

3. The method according to claim 1, wherein, The first shared content is used to publish to a content sharing platform, the content sharing platform provides the first content, and the receiving of the first operation includes: A selection operation is received, wherein the selection operation is used to select the first text from the first content.

4. The method according to claim 3, wherein, The receive selection operation includes: In response to the add operation, the first content is presented, and the selection operation is received.

5. The method according to claim 4, wherein, The response to the add operation, displaying the first content, includes at least one of the following: In response to a first add operation in the first editing interface, the first content is presented, wherein the first editing interface is used to edit the first shared content, and the first add operation is used to associate the first shared content with the first content; In response to the second add operation in the first editing interface, the first content is presented, wherein the second add operation is used to edit the multimedia content in the first shared content; In response to a third add operation in the first display interface, the first content is presented, wherein the first display interface is used to present candidate multimedia content of the first shared content, and the third add operation is used to edit the multimedia content in the first shared content.

6. The method according to any one of claims 1 to 5, further comprising: Present first information associated with the at least one candidate multimedia content, wherein the first information is used to indicate the generation progress of the at least one candidate multimedia content.

7. The method according to claim 6, wherein, The presentation of first information associated with the at least one candidate multimedia content includes at least one of the following: Presenting first progress information, wherein the first progress information is used to indicate the current percentage of completion of the generation of the at least one candidate multimedia content; A second progress information is presented, wherein the second progress information is used to indicate the generation progress of each candidate multimedia content among the at least one candidate multimedia content.

8. The method according to any one of claims 1 to 5, wherein, The step of receiving a second operation during the generation of at least one candidate multimedia content includes: During the generation of the at least one candidate multimedia content, a second editing operation is received, wherein the second editing operation is used to edit the first shared content.

9. The method according to claim 8, further comprising: Receive third operation; In response to the third operation satisfying the set conditions, a generation result is presented, wherein the generation result includes the candidate multimedia content generated when the third operation is received.

10. The method according to claim 9, wherein, The third operation that satisfies the set conditions includes at least one of the following: A collapse operation, wherein the collapse operation is used to collapse the information input area, and the second editing operation is triggered in the information input area; The publishing operation is used to publish the first shared content to the content sharing platform.

11. The method according to claim 9, wherein, The at least one candidate multimedia content includes a first multimedia content. Before the first multimedia content is generated, the generation result does not include the first multimedia content. The method further includes: In response to the completion of the generation of the first multimedia content, the first multimedia content is presented in the generation result.

12. The method according to claim 9, wherein, The response to the third operation satisfying the set conditions, presenting the generated result, includes: In response to the third operation satisfying the set conditions, the generation result is presented on the first editing interface, wherein the first editing interface is used to edit the first shared content; and Presenting the at least one candidate multimedia content includes: In response to a triggering operation on the generated result, the at least one candidate multimedia content is presented on the first display interface.

13. The method according to any one of claims 1 to 5, further comprising: In response to a triggering operation on the second multimedia content among the at least one candidate multimedia content, the second multimedia content is presented in the first shared content.

14. The method of claim 13, further comprising: In response to the publishing operation of the first shared content, the first display content is presented on the shared content presentation interface; The first display content includes: the second multimedia content and a portion of the text content in the first shared content, and the shared content presentation interface is used to present display content published by multiple publishers.

15. The method according to any one of claims 1 to 5, wherein, The first shared content is not document content.

16. An interactive device, comprising: The first receiving module is configured to receive a first operation, wherein the first operation is used to generate candidate multimedia content for the first shared content, and the first operation is associated with the first text. The second receiving module is configured to receive a second operation during the generation process of at least one candidate multimedia content, wherein the at least one candidate multimedia content is generated based on the first text, and the second operation is unrelated to the generation process of the at least one candidate multimedia content. The presentation module is configured to present the at least one candidate multimedia content, wherein the at least one candidate multimedia content is used to add to the first shared content.

17. An electronic device comprising: At least one processor; as well as At least one memory, including one or more computer program instructions; The one or more computer program instructions are executed by the processor to perform the method according to any one of claims 1 to 15.

18. A computer-readable storage medium for non-transitory storage of computer-readable instructions, wherein, The method of any one of claims 1 to 15 is implemented when the computer-readable instructions are executed by a processor.

19. A computer program product comprising a computer program that, when executed by a processor, implements the method of any one of claims 1 to 15.