Method, apparatus, device and storage medium for interaction

By presenting guiding information and receiving user-triggered actions in the video interaction system, and using machine learning to identify narrative-based comments, the problem of limited video interaction methods and creative difficulties has been solved, thereby increasing video exposure and user engagement.

CN122340302APending Publication Date: 2026-07-03DOUYIN VISION CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
DOUYIN VISION CO LTD
Filing Date
2026-04-10
Publication Date
2026-07-03

AI Technical Summary

Technical Problem

Existing video interaction methods are simplistic, leaving users lacking inspiration and direction when creating videos. The operation process is cumbersome, and the video exposure is insufficient, resulting in low user interest and efficiency in creation.

Method used

By presenting guidance information related to video comment content, receiving user-triggered actions, and providing an interface for users to view or create videos related to the comment content, the system utilizes machine learning models to identify narrative-driven comment content and provide personalized guidance information.

Benefits of technology

It increased video exposure and user engagement, simplified the creation process, and improved interaction efficiency and user experience.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN122340302A_ABST
    Figure CN122340302A_ABST
Patent Text Reader

Abstract

A method, apparatus, device, medium, and program product for interaction are provided. The proposed method includes: presenting first comment content and first guidance information, wherein the first comment content is issued to a first video, the first video displays at least one first event related to at least one object, the first comment content indicates a second event related to the first video, and the first guidance information indicates content creation related to the second event; receiving a first trigger operation, the first trigger operation being issued to the first guidance information; and presenting a first interface for interaction related to the second video, the second video being used to display the second event. In this manner, based on comment content belonging to the narrative category, guidance information can be provided to the user, thereby facilitating the user's video-related interactive behavior.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The various example implementations generally relate to the field of computers, and in particular to methods, apparatuses, devices, computer-readable storage media, and computer program products for interaction. Background Technology

[0002] With the development of information technology, various terminal devices can provide people with a variety of services in work and life. Applications providing these services can be deployed on these terminal devices. Terminal devices or applications can provide users with intelligent system functions to assist them in using the devices or applications. For example, users can view videos generated by intelligent systems by operating various applications installed on their terminal devices. However, existing video interaction methods are usually relatively simple. Summary of the Invention

[0003] In a first aspect, an interactive method is provided. The method includes: presenting first comment content and first guidance information, the first comment content being issued to a first video, the first video displaying at least one first event related to at least one object, the first comment content indicating a second event related to the first video, and the first guidance information indicating content creation related to the second event; receiving a first trigger operation, the first trigger operation being issued to the first guidance information; and presenting a first interface for interaction related to the second video, the second video being used to display the second event.

[0004] In a second aspect, an apparatus for interaction is provided. The apparatus includes: a guidance information presentation module configured to present first comment content and first guidance information, the first comment content being issued to a first video, the first video displaying at least one first event related to at least one object, the first comment content indicating a second event related to the first video, and the first guidance information indicating content creation related to the second event; a trigger operation receiving module configured to receive a first trigger operation, the first trigger operation being issued to the first guidance information; and an interface presentation module configured to present a first interface for interaction related to the second video, the second video being used to display the second event.

[0005] In a third aspect, an electronic device is provided. The device includes at least one processor; and at least one memory coupled to the at least one processor and storing instructions for execution by the at least one processor. When executed by the at least one processor, the instructions cause the device to perform the method of the first aspect.

[0006] In a fourth aspect, a computer-readable storage medium is provided. The computer-readable storage medium stores computer-executable instructions that can be executed by a processor to implement the method of the first aspect.

[0007] In a fifth aspect, a computer program product is provided, including computer-executable instructions, wherein the computer-executable instructions, when executed by a processor, implement the method of the first aspect.

[0008] In this way, based on narrative-driven comments, users can be guided to perform video-related interactive actions. For example, users can easily view or create videos related to the comments.

[0009] It should be understood that the content described in this section is not intended to limit the key or important features to be protected, nor is it intended to restrict the scope of protection. Other features will become readily apparent from the following description. Attached Figure Description

[0010] The above and other features, advantages, and aspects of the various implementations will become more apparent from the accompanying drawings and the following detailed description. In the drawings, the same or similar reference numerals denote the same or similar elements, wherein: Figure 1 A schematic diagram of an example environment in which one or more scenarios can be implemented is shown; Figure 2 A flowchart illustrating an example process of interaction based on several scenarios is shown; Figures 3A to 3F A schematic diagram of an example interface for interaction based on several scenarios is shown; Figure 4 A schematic structural block diagram of an example device for interaction under certain conditions is shown; and Figure 5 A block diagram of an electronic device capable of implementing multiple illustrative scenarios is shown. Detailed Implementation

[0011] The examples in this document will now be described in more detail with reference to the accompanying drawings. While some examples are shown in the drawings, it should be understood that solutions can be implemented in various forms and should not be construed as limited to the examples presented herein. Rather, these examples are provided to provide a more thorough and complete understanding of the solutions. It should be understood that the drawings and examples in this document are for illustrative purposes only and are not intended to limit the scope of protection of the solutions.

[0012] It should be noted that the headings of any section / subsection provided herein are not restrictive. Various examples are described throughout this document, and examples of any type may be included under any section / subsection. Furthermore, examples described in any section / subsection may be combined in any way with any other examples described in the same section / subsection and / or different sections / subsections.

[0013] In the examples herein, the term "including" and similar expressions should be understood as open inclusion, i.e., "including but not limited to". The term "based on" should be understood as "at least partially based on". The terms "an example" or "the example" should be understood as "at least one example". The term "some examples" should be understood as "at least some examples". Other explicit and implicit definitions may also be included below. The terms "first", "second", etc., may refer to different or the same objects. Other explicit and implicit definitions may also be included below.

[0014] The examples in this document may involve user data, data acquisition, and / or use. All of these aspects comply with relevant laws, regulations, and provisions. In the examples presented herein, all data collection, acquisition, processing, manipulation, forwarding, and use are conducted with the user's knowledge and confirmation. Accordingly, when implementing each example, the type, scope of use, and usage scenarios of any data or information that may be involved should be communicated to the user and their authorization obtained through appropriate means, in accordance with relevant laws and regulations. The specific methods of notification and / or authorization can vary depending on the actual situation and application scenario; the scope of the solution is not limited in this regard.

[0015] In this document and the examples provided, any processing of personal information will be conducted only on a legal basis (e.g., with the consent of the data subject, or as necessary for the performance of a contract) and will only be carried out within the scope stipulated or agreed upon. A user's refusal to process personal information beyond what is necessary for basic functions will not affect their use of those basic functions.

[0016] In this paper, an intelligent system refers to a system capable of autonomous control based on machine learning models. An intelligent system, for example, can make decisions and autonomously execute actions based on machine learning models to achieve preset goals or complete preset tasks. An intelligent system can be an automated program that understands user intent and can utilize models or invoke tools to complete various types of tasks. In some contexts, examples of intelligent systems may include, but are not limited to: agents, bots, chatbots, digital avatars, intelligent customer service, digital assistants, etc. Alternatively, an intelligent system can also be an intelligent role implemented based on machine learning models. An "intelligent system" can process user requests based on generative models (e.g., language models, multimodal models) to perform specified types of tasks.

[0017] In some scenarios, intelligent systems can be represented as virtual avatars or physical entities for interaction with users. The term "virtual object" can refer to a digital entity capable of interacting with a user. Intelligent systems can therefore also be called virtual objects, digital assistants, intelligent assistants, AI assistants, chatbots, virtual agents, etc. Virtual objects can possess intelligent dialogue and information processing capabilities, responding to user queries and providing appropriate answers. In some examples, virtual objects can be presented in a graphical form within the interactive interface, such as as avatars, animated characters, or other visual representations.

[0018] Figure 1 A schematic diagram of example environment 100 is shown. (e.g.) Figure 1 As shown, example environment 100 may include terminal device 110. Application 120 is installed on terminal device 110. User 140 can interact with application 120 via terminal device 110 and / or an attached device of terminal device 110.

[0019] In some cases, application 120 can be any appropriate application that can provide video. Figure 1 In environment 100, if application 120 is active, terminal device 110 can display page 150 of application 120. Page 150 may include various types of pages provided by application 120, such as interactive pages, video viewing interfaces, video creation pages, query pages, search pages, search result display pages, etc.

[0020] In some cases, terminal device 110 communicates with server 130 to provide services to application 120. Terminal device 110 can be any type of mobile terminal, fixed terminal, or portable terminal, including mobile phones, desktop computers, laptop computers, notebook computers, netbook computers, tablet computers, media computers, multimedia tablets, handheld computers, portable gaming terminals, virtual reality (VR) devices, augmented reality (AR) devices, personal communication system (PCS) devices, personal navigation devices, personal digital assistants (PDAs), audio / video players, digital cameras / camcorders, positioning devices, television receivers, radio receivers, e-book devices, gaming devices, or any combination of the foregoing, including accessories and peripherals of these devices or any combination thereof. In some cases, terminal device 110 may also support any type of user-facing interface (such as "wearable" circuitry).

[0021] Server 130 can be a standalone physical server, a server cluster or distributed system composed of multiple physical servers, or a cloud server providing basic cloud computing services such as cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communication, middleware services, domain name services, security services, content delivery networks, and big data and artificial intelligence platforms. Server 130 may include, for example, computing systems / servers such as mainframes, edge computing nodes, computing devices in a cloud environment, etc. Server 130 can provide backend services for interactive applications 120 in terminal device 110.

[0022] A communication connection can be established between server 130 and terminal device 110. This communication connection can be established via wired or wireless means. The communication connection can include, but is not limited to, Bluetooth, mobile network, Universal Serial Bus (USB), and Wireless Fidelity (WiFi) connections. In some cases, server 130 and terminal device 110 can exchange signaling information through their communication connection.

[0023] It should be understood that the structure and function of the various elements in environment 100A are described for illustrative purposes only and do not imply any limitation on the scope of the scheme.

[0024] In some cases, application 120 can provide interaction capabilities with the intelligent system. Application 120 can be dedicated to providing services from the intelligent system, or it can be an application integrated with the intelligent system (meaning it can provide functions or services other than those of the intelligent system). Although Figure 1 The image shows a single application, but in reality, multiple applications can be installed on the terminal device 110.

[0025] In this paper, the intelligent system can be deployed locally on terminal device 110 or remotely. In the case of remote deployment, terminal device 110 can directly call the intelligent system, or it can call the intelligent system via server 130.

[0026] In some scenarios, during interaction with user 140, the intelligent system can respond to user 140's requests and handle tasks instructed by the user. In some scenarios, the intelligent system may have intelligent dialogue and task processing capabilities. Terminal device 110 can provide an interface 150 for interacting with the intelligent system. In interface 150, user 140 can initiate task requests to the intelligent system by inputting natural language (e.g., text input or voice input). Alternatively or additionally, user 140 can upload local files or specify online files to instruct the intelligent system to assist in completing various tasks.

[0027] As briefly described above, users can view videos by operating various applications installed on their terminal devices. For example, users can view videos generated by the intelligent system based on narrative-driven videos. Typically, users can browse reviews of narrative-driven videos. For instance, users might create videos based on reviews of narrative-driven content. In this scenario, when the current user or other users browse these reviews, it's difficult to find videos related to those reviews, resulting in insufficient distribution and exposure for videos created based on those reviews.

[0028] Furthermore, users need to create videos through the creation portal on the homepage. On the one hand, this method often lacks a clear direction of inspiration when users have the intention to create videos, and the operation process is cumbersome, reducing users' interest in creating videos. On the other hand, the conventional creation method requires users to manually fill in the event description information used to create videos, which has problems with standardization and low operation efficiency.

[0029] In view of this, an improved interaction scheme is proposed. In this scheme, first comment content and first guidance information are presented. The first comment content is issued to a first video, which displays at least one first event related to at least one object. The first comment content indicates a second event related to the first video, and the first guidance information indicates content creation related to the second event. Further, a first trigger operation is received, which is issued in response to the first guidance information. Correspondingly, a first interface is presented for interaction related to the second video, which displays the second event.

[0030] This approach allows for the presentation of guiding information related to the first comment, facilitating user interactions related to the video. This guides users to view or create videos connected to the first comment, improving interaction efficiency. Furthermore, guiding users to view videos increases video exposure and satisfies users' desire to extend the original storyline into videos. Correspondingly, guiding users to create videos helps them quickly clarify their creative direction, increasing their motivation to create content.

[0031] The following description of the examples will continue with reference to the accompanying drawings. It should be understood that the pages shown in the drawings are merely examples, and various page designs are possible in practice. The various graphic elements on the page may have different arrangements and different visual representations, one or more elements may be omitted or replaced, and one or more other elements may also be present. This document is not limited in this respect. Furthermore, the examples will be described primarily with respect to terminal device 110 in the following text. It should be understood that the actions described with respect to terminal device 110 can also be performed by application 120 on terminal device 110, or by application in conjunction with its server (e.g., server 130).

[0032] For ease of understanding, the following text will refer to Figure 2 To describe an example process used for interaction. Figure 2 A schematic diagram of an example process 200 for interaction based on some scenarios is shown.

[0033] In the discussion of process 200, for better understanding, we will combine... Figures 3A to 3F Let's explain some examples used for interaction. Figures 3A to 3F Schematic diagrams of example interfaces 300A to 300F for interactions under certain conditions are shown.

[0034] In some cases, terminal device 110 may provide commentary content related to the first video. The first video is used to display at least one first event related to at least one object. In this document, the first video may be an original video created and released by a creator (such as an individual or organization). As an example, the first video may be a film or television work, such as a TV series, short drama, movie, short video, documentary, variety show, etc. At least one object may refer to an object in the first video, such as including but not limited to people (e.g., characters), animals, plants, virtual objects (e.g., cartoon characters or cartoon animals), etc. At least one event related to at least one object may refer to any appropriate event included in the first video, such as a comedic event, a family event, or a historical event.

[0035] In some cases, the first video may include a narrative video related to, for example, a person, animal, or virtual object, where at least one object includes objects appearing in the narrative video, such as a person, animal, or virtual object, and the first event may include a storyline related to that person, animal, or virtual object. In other cases, the first video may include a documentary about an animal or plant, and the first event may describe the life or growth process of that animal or plant, etc. Of course, the above-described first video, object, and first event are merely examples. The object and first event may differ depending on the type or subject matter of the first video. This document does not impose any limitations on this.

[0036] In some examples, comments related to the first video may include, but are not limited to, comments on the video's picture quality, the actors' performances, and the plot. These comments may be posted by multiple users of the first video. In some examples, plot-related comments (e.g., the first comment content) may be comments on the video's plot, storyline, character settings, story logic, foreshadowing and ending, character relationships, etc. In some cases, the first comment content (i.e., plot-related comments) may indicate a second event related to the first video. For example, plot speculations, plot associations, etc., posted by a user (e.g., user 140 or any other suitable user) regarding the plot of the first video. Alternatively / additionally, the first comment content may also indicate analysis, criticism, interpretation, etc., posted by user 140 regarding the plot of the first video.

[0037] The first comment can be any suitable form of comment. In some examples, the first comment can be a comment on the video details page. In some examples, the first comment can be a comment about the first video in a community or forum. In some examples, the first comment can include bullet comments. For example, terminal device 110 can display bullet comments on the video frame used to present the first video.

[0038] In some cases, refer to Figure 2 In box 210, terminal device 110 can display first guiding information for the first comment content. In some examples, terminal device 110 can display the first guiding information for the first comment content belonging to the narrative category. The first guiding information can instruct content creation related to the second event. For example, the first guiding information can instruct the user (e.g., the user who posted the first comment content) on the narrative associated with the first video.

[0039] refer to Figure 3A Terminal device 110 can display comment content 312 (an example of the first comment content) and comment content 313 (an example of the first comment content) via interface 300A. It should be understood that terminal device 110 can display other comments via interface 300A, such as comment content 311 and comment content 314 (for example, used to comment on the picture quality of the first video). As an example, terminal device 110 can display guidance information 315 (an example of the first guidance information) for comment content 312. For example, assuming comment content 312 is "I really want to see the subsequent life of character 1 and character 2," then guidance information 315 can be associated with "the subsequent life of character 1 and character 2." Guidance information 315 can guide user 140 to view a second video (also known as a "secondary creation video") themed around "the subsequent life of character 1 and character 2."

[0040] As an example, refer to Figure 3E Terminal device 110 can display guidance information 351 (an example of the first guidance information) in response to comment content 313. For example, assuming comment content 313 is "Has character 3 gone offline so quickly?", then guidance information 316 can be associated with "the sequel to character 3's storyline". Guidance information 316 can guide user 140 to create a second video on the theme of "the sequel to character 3's storyline". It should be understood that although they are respectively in Figure 3A and Figure 3E The example shown is of guide information 315 and guide information 351, but this is merely an example. For instance, terminal device 110 may simultaneously display guide information 315 and guide information 351 on the same interface. For example, terminal device 110 may display guide information 315 and guide information 316 in association with comment content 312 and comment content 313 via interface 300A or 300E.

[0041] The following describes how terminal device 110 determines that the first comment content belongs to the narrative category. In some cases, terminal device 110 can determine that the first comment content is narrative-related if it determines that the first comment content includes keywords related to the narrative of the first video. Keywords can be pre-configured based on the first video and may include, but are not limited to, the name of at least one object, key events (e.g., key plot points), plot turning points, descriptions of object relationships, etc.

[0042] In some cases, terminal device 110 can determine that the first comment content belongs to the narrative category if it determines that the semantics of the first comment content matches the semantics of at least one first event. As an example, terminal device 110 can invoke a first machine learning model to perform semantic parsing on the first comment content to determine its intent. If the intent of the first comment content indicates at least one of discussing the plot of the first video, commenting on plot development, speculating on plot direction, or mentioning plot-related details, then the first comment content can be identified as narrative-related comment content. In some examples, the first machine learning model may, for example, include a Natural Language Processing (NLP) model.

[0043] Additionally / alternatively, terminal device 110 may utilize a second machine learning model to determine that the first comment content belongs to the drama category. In some examples, the second machine learning model may be a trained classification model used to identify drama-type comment content. The classification model is configured to output a confidence level that the first comment content belongs to the drama category based on the first comment content. As an example, terminal device 110 inputs the first comment content into the classification model to obtain a confidence level that the first comment content belongs to the drama category. If this confidence level is greater than a confidence threshold (e.g., 80% or any other appropriate threshold), the first comment content can be identified as drama-type comment content.

[0044] In box 220, terminal device 110 receives a first trigger operation, which indicates confirmation of guidance information. In some cases, terminal device 110 may receive a first trigger operation from user 140 on the first guidance information. After receiving the first trigger operation on the first guidance information, terminal device 110 may proceed to box 230. In box 230, terminal device 110 may display a first interface. The first interface is used for interaction related to the second video, such as user 140 watching the second video. Alternatively, user 140 may create a second video.

[0045] In some cases, the second video is used to showcase a second event. As an example, terminal device 110 can generate a second video associated with the first comment content. For instance, continuing with the example above, assuming comment content 312 is "I really want to see what happens to Character 1 and Character 2 later," the main storyline of the second video could be "what happens to Character 1 and Character 2 later." For example, the second video could instruct terminal device 110 to generate a continuation of the story based on comment content 312, building upon the first video. In some cases, terminal device 110 can utilize an intelligent system to generate the second video. The following description of how to create a second video will detail how this is done.

[0046] Reference Figures 3A to 3B If terminal device 110 receives a trigger on guidance information 315, it can present interface 300B (which is an example interface of the first interface). Terminal device 110 can play a second video in interface 300B for user 140 to watch. In some cases, terminal device 110 can provide multiple second videos to user 140 in relation to the first comment content. Terminal device 110 can sort these second videos based on their relevance to the first comment content. In some examples, terminal device 110 can play these second videos to user 140 in sequence. For example, terminal device 110 can prioritize playing second videos with a relevance threshold.

[0047] In some examples, refer to Figure 3B Terminal device 110 can present second guidance information (e.g., guidance information 321) in association with the second video. Guidance information 321 can be used to guide user 140 to create a third video. Terminal device 110 can create the third video based on the creation information corresponding to the second video. As an example, if terminal device 110 receives a trigger from user 140 on guidance information 321, it can present a first creation interface. In this scenario, the first creation interface can display the creation information corresponding to the second video. Terminal device 110 can create the third video based on this creation information. As an example, user 140 can update the creation information via the first creation interface. Terminal device 110 can generate the third video using the updated creation information. The process of creating the third video can be referred to the process of creating the second video, which will not be repeated here. The following will describe in detail how to create the second video.

[0048] In some examples, terminal device 110 may present third guidance information on the first interface. For example, terminal device 110 may present third guidance information while playing a second or third video. This third guidance information is used to guide the user to watch the original video, and it may include information in one or more modalities, such as text, images, etc. For example, terminal device 110 may present the third guidance information "Watch the original video" on the first interface while playing a second video. If terminal device 110 receives a fourth trigger operation on the third guidance information, terminal device 110 may present the first video. In this way, an entry point can be provided to switch back from the secondary creation video to the original video, enabling the original video and the secondary creation video to form a "content ecosystem." After watching the secondary creation video, the user can jump to the playback interface of the original video through this entry point, which helps to improve the completion rate of the original video and the user retention rate of the content platform.

[0049] Reference Figures 3E to 3F If terminal device 110 receives a trigger on guidance information 351, it can display interface 300F (which is an example interface of the second creation interface). The second creation interface is used to receive creation information; for example, it can be referred to as a "creation workbench." Terminal device 110 can receive creation information via the second creation interface. Terminal device 110 can create a second video based on the creation information received via interface 300F. The following will refer to... Figure 3E and Figure 3F This section describes in detail how to create a second video based on the guidance information 351.

[0050] In some cases, if terminal device 110 determines that the first comment content belongs to the narrative category, it may present guidance information. In other cases, if terminal device 110 receives a second trigger operation related to the first comment content and the second trigger operation is a predetermined type of operation, it may present the first guidance information. In this way, by providing guidance information to user 140 under certain operational conditions, a second video associated with the current comment content can be accurately pushed to user 140.

[0051] In some scenarios, if terminal device 110 determines that the second triggering operation is an interaction performed on the first comment content, it may present the first guiding information. The interaction may include, but is not limited to, liking, replying, or mentions of user 140 by other users. In some scenarios, if terminal device 110 receives a comment posted by user 140 and determines that the comment content is a narrative-based comment, it may present the first guiding information. Alternatively / additionally, if terminal device 110 determines that the first comment content includes mentions of user 140 (e.g., @user140), it may present the first guiding information. In some scenarios, if the first comment content is a bullet comment (danmu), and if user 140 likes the bullet comment and the bullet comment is a narrative-based comment, the first guiding information may be presented.

[0052] This approach, which presents guidance information based on user behavior and comment content, enables personalized guidance and avoids ineffective push notifications. It allows users to access or create comments while browsing them, thus improving the user experience.

[0053] The above describes the conditions under which the first guidance message can be presented. The following will refer to... Figure 3A and Figures 3C to 3D This describes how the first guidance message should be presented. For ease of understanding, guidance message 315 will be used as an example. Of course, other first guidance messages can still be presented in the same way.

[0054] Reference Figure 3A Terminal device 110 can present guidance information 315 in the style of message 316. Message 316 can be configured to trigger the presentation or create an entry point for a second video. In some examples, terminal device 110 can present message 316 in the area associated with comment content 312. See reference. Figure 3C Terminal device 110 can present guidance information 315 in the form of card 331. Card 331 can be configured to display a thumbnail of the second video and the name of the second video. When the guidance information is used to guide user 140 to create the second video, an entry point for creating the second video (not shown in the figure) can be presented via card 331.

[0055] Additional / alternative sites, see Figure 3D Terminal device 110 can display guidance information 315 in window 341. Terminal device 110 can also display confirmation control 342 in window 341. Confirmation control 342 can be configured to confirm viewing or create a second video.

[0056] In this way, when users browse comments, they can more easily and quickly find out which comments they can view or create a second video for. This also improves the readability of the guidance information for user 140.

[0057] As mentioned above, in some situations, the first comment can be a bullet screen (danmu). In such scenarios, terminal device 110 can present first guiding information associated with the bullet screen. For example, terminal device 110 can present the first guiding information in the video frame as a message associated with the first bullet screen. This message can be configured as an entry point for viewing or creating a second video. In this scenario, when scrolling through the bullet screen and the associated first guiding information, the scrolling speed of the first bullet screen can be slower than the scrolling speed of other bullet screens. This makes it easier for users to trigger the first guiding information for the first bullet screen, thereby facilitating the creation or viewing of a second video.

[0058] As another example, terminal device 110 can present the first guidance information in a card format within the video frame. This card can be configured to display a thumbnail of the second video and the name of the second video. Alternatively / additionally, terminal device 110 can present the first guidance information in a window within the video frame. For example, terminal device 110 can present a confirmation control in the window, which can be configured to confirm viewing or creating the second video.

[0059] In this way, by presenting guiding information associated with bullet comments, users can obtain an entry point to view or create a second video while watching the first video, thereby improving the user experience.

[0060] The above describes how to present guiding information. The following text will continue to refer to... Figures 3E to 3F This describes how to create a second video based on the guidance information.

[0061] In some cases, refer to Figures 3E to 3FIn a scenario where guidance information 351 guides user 140 to create a second video, terminal device 110 can present creation information for creating the second video in interface 300F (which is an example interface of the second creation interface). The creation information describes a second event and is used to generate the second video. The creation information can describe a second event related to first object information in at least one object. As an example, the first object information may include one or more objects in the first video, and the second event may be the same as or different from the first event. For instance, the first object information may include some objects in the original video, and the plot indicated by the second event may be different from the plot indicated by the first event.

[0062] If terminal device 110 receives a trigger on guidance information 351, it can present interface 300F. Terminal device 110 can present object selection information and event description information in interface 300F. In some cases, the object selection information indicates the selection of one or more objects from at least one set of objects. Terminal device 110 can present one or more objects in a selected state in interface 300F, such as role 1 and role 3. Alternatively / additionally, terminal device 110 can also present other objects from at least one set of objects in interface 300F for user 140 to select for creating a second video.

[0063] In some cases, terminal device 110 may present event description information 361 in input component 362 to describe the second event. For example, the second event may be the plot to be shown in a derivative video (i.e., a second video), and the event description information may be used to describe the plot. In some cases, the event description information may include information in one or more modalities, such as text, voice, or images. In other words, the user may describe the plot of the derivative video using at least one of text, voice, or images. Terminal device 110 may include the event description information as at least part of the creation information.

[0064] In some cases, the event description information 361 may be generated based on the comment content 313 and the plot information of the first video. In some examples, the terminal device 110 may extract the plot information corresponding to the comment content 313. For example, the terminal device 110 may perform semantic analysis on the comment content 313 to obtain the plot information. The plot information may include, but is not limited to, the object, object relationships, plot fragments (e.g., the second event), conflicts, etc., corresponding to the comment content 313. Furthermore, the terminal device 110 may obtain the plot information of the first video. For example, the terminal device 110 may obtain the plot information of the first video from the database corresponding to the first video. The plot information corresponding to the first video may include, but is not limited to, the complete plot, object relationship graph, object attribute settings (e.g., personality or any other appropriate attribute), plot background, etc. Furthermore, the terminal device 110 may generate the event description information 361 based on the plot information corresponding to the comment content 313 and the plot information of the first video.

[0065] In some cases, terminal device 110 can receive creation information matching the current interaction based on an interaction action issued for at least one of object selection information and event description information. As an example, terminal device 110 can create a second video based on the creation information currently presented in interface 300F. As an example, terminal device 110 can create a second video based on updated creation information. How a second video is created will be described in detail below.

[0066] The following first describes how a second video is created based on the creation information currently presented in interface 300F. In some cases, if terminal device 110 determines that the interactive operation indicates confirmation of object selection information and event description information, it receives the object selection information and event description information as creation information. As an example, if terminal device 110 receives a trigger on control 365, it can use one or more currently presented objects (e.g., character 1, character 3) and event description information 361 as creation information to create a second video. In this way, by automatically generating event description information 361 when user 140 triggers guidance information 351, user 140 can quickly clarify the creative direction, thereby reducing manual input costs and increasing user motivation for creation.

[0067] The following section describes how to create a second video based on the updated creation information.

[0068] In some cases, regarding the creation information currently displayed on the interface 300F by the terminal device 110, the user 140 can update this creation information, allowing the terminal device 110 to create a second video based on the updated creation information. This makes the created second video more tailored to the user's needs. Specifically, refer to... Figure 3F If the terminal device 110 receives an editing operation on the event description information 361 via the input component 362, it can obtain the updated event description information as part of the creation information. Editing operations include, but are not limited to, modifying content, adding content, deleting content, and other editing operations.

[0069] Furthermore, if the terminal device 110 receives a selection of one or more other objects (e.g., character 2, character 4) from at least one object via the selection component, it can obtain updated object selection information as part of the creation information. Then, the terminal device 110 creates a second video based on the updated event description information and the other one or more objects (e.g., character 2, character 4). In this way, the created second video can better meet the user's needs, thereby enriching the content of the second video.

[0070] Additionally / alternatively, terminal device 110 may update event description information based on labels presented in interface 300F. Specifically, if terminal device 110 receives a selection for at least one label (e.g., label 363, or label 364, etc.), it may update event description information 361 based on that label.

[0071] In some examples, terminal device 110 may present at least one label (e.g., also referred to as "inspiration labels"), each indicating at least one event summary. The event summary may indicate a concise description of content creation. In some cases, the event summary may indicate a creative idea, inspiration, or intention. A user may select one or more labels from the at least one label; for example, a user may select a first label from the at least one label. Terminal device 110 may receive a first selection operation instructing the selection of a label from the at least one label, which may indicate an event summary of a second event. Subsequently, terminal device 110 may present updated event description information that matches the event summary of the second event.

[0072] In some examples, terminal device 110 may display labels 363, 364, etc., in interface 300F (e.g., input component 362). For example, label 363 may include "A Little Romantic," indicating an event summary of "generating a video with a romantic and heartwarming theme." Label 364 may also include "Embracing While Watching the Sea," indicating an event summary of "the male and female protagonists embracing while watching the sea."

[0073] In some cases, terminal device 110 can acquire event information from the original video (i.e., the first video), which describes at least one first event. In some cases, the event information may also be referred to as plot information, which describes the plot of the original video. Terminal device 110 can use a machine learning model to generate the at least one tag based on the plot information. In some cases, the event summary indicated by the at least one tag may differ from the plot of the original video.

[0074] In some examples, user 140 can select one or more tags from a range of options, such as tag 363. Terminal device 110 can then obtain updated event description information based on tag 363. For example, if user 140 selects character 3 and tag 363, terminal device 110 can use a machine learning model to generate updated event description information based on the plot information of the first video, character 3's name, and tag 363. Terminal device 110 can then add the updated event description information to input component 362.

[0075] In some cases, if terminal device 110 determines that the number of created second videos corresponding to the first comment content is less than a number threshold (e.g., 5 or any other suitable number threshold), it can present guidance information to encourage the user to create a second video. Conversely, if terminal device 110 determines that the number of created second videos corresponding to the first comment content exceeds the number threshold, it can present guidance information to encourage the user to view the second video. Additionally / alternatively, terminal device 110 can simultaneously present guidance information for both creating and viewing second videos for the first comment content. This enriches the variety of second videos and improves interaction efficiency.

[0076] In summary, by presenting guiding information related to the first comment, users can be directed to view or create videos associated with that comment, thereby improving interaction efficiency. Furthermore, guiding users to view videos increases their exposure and satisfies users' desire to create videos that extend the original storyline. Correspondingly, guiding users to create videos helps them quickly clarify their creative direction, increasing their motivation to create content.

[0077] In some cases, corresponding apparatus for implementing the above methods or processes is also provided. Figure 4 A schematic structural block diagram of an example device 400 based on interactions in some scenarios is shown. Device 400 may be implemented as or included in terminal device 110. The various modules / components in device 400 may be implemented by hardware, software, firmware, or any combination thereof.

[0078] As shown in the figure, the device 400 guide information presentation module 410 is configured to present first comment content and first guide information. The first comment content is issued to a first video, the first video displays at least one first event related to at least one object, the first comment content indicates a second event related to the first video, and the first guide information indicates content creation related to the second event. The device 400 also includes a trigger operation receiving module 420, configured to receive a first trigger operation, which is issued to the first guide information. The device 400 also includes an interface presentation module 430, configured to present a first interface for interaction related to the second video, and the second video for displaying the second event.

[0079] In some cases, the guidance information presentation module 410 is also configured to receive a second trigger operation related to the content of the first comment; and to present the first guidance information in response to the second trigger operation being a predetermined type of operation.

[0080] In some cases, the guidance information presentation module 410 is also configured to present first guidance information in response to a second triggering operation that sends a first comment; or to present first guidance information in response to an interaction performed on the first comment.

[0081] In some cases, the first guidance information is used to guide the viewing of the second video, and the first interface is used to play the second video.

[0082] In some cases, the first interface includes second guidance information that instructs the creation of a third video, and the device 400 also includes a creation interface presentation module configured to receive a third trigger operation that is issued in response to the second guidance information; and in response to the third trigger operation, to present a first creation interface that displays creation information used to generate the third video, the creation information being related to the second video.

[0083] In some cases, device 400 also includes a video presentation module configured to present third guidance information on a first interface, the third guidance information instructing the viewer to watch a first video; receive a fourth trigger operation, the fourth trigger operation being issued in response to the third guidance information; and present the first video in response to the fourth trigger operation.

[0084] In some cases, the first guidance information instructs the creation of a second video, and the first interface includes a second creation interface for receiving creation information that describes a second event and is used to generate the second video.

[0085] In some cases, device 400 also includes a creation information receiving module configured to present object selection information and event description information on a second creation interface, wherein the object selection information indicates the selection of one or more objects from at least one object, and the event description information describes a second event; receive an interactive operation, wherein the interactive operation is issued in response to at least one of the object selection information and the event description information; and receive creation information that matches the interactive operation.

[0086] In some cases, the creation information receiving module is also configured to receive object selection information and event description information as creation information in response to interactive operation instructions to confirm object selection information and event description information.

[0087] In some cases, the creation information receiving module is also configured to: receive updated object selection information as part of creation information in response to an interactive operation instruction to select one or more other objects from at least one object; the updated object selection information indicates the selection of one or more other objects; and receive updated event description information as part of creation information in response to an interactive operation instruction to update event description information.

[0088] In some cases, the second creation interface presents at least one label, each label indicating at least one event summary, and the device 400 further includes an interactive operation receiving module configured to receive a selection operation, the selection operation instructing the selection of a label from at least one label, the selected label indicating a first event summary of the second event, wherein the updated event description information matches the first event summary.

[0089] The modules included in device 400 can be implemented in various ways, including software, hardware, firmware, or any combination thereof. In some cases, one or more modules can be implemented using software and / or firmware, such as machine-executable instructions stored on a storage medium. In addition to or as an alternative to machine-executable instructions, some or all of the units in device 400 can be implemented at least partially by one or more hardware logic components. By way of example, and not limitation, exemplary types of hardware logic components that can be used include field-programmable gate arrays (FPGAs), application-specific integrated circuits (ASICs), application-specific standard parts (ASSPs), systems on a chip (SOCs), complex programmable logic devices (CPLDs), and so on.

[0090] Figure 5 A block diagram of an electronic device 500 in which one or more examples may be implemented is shown. It should be understood that... Figure 5 The electronic device 500 shown is merely exemplary and should not be construed as limiting the functionality and scope of the examples described herein. Figure 5 The electronic device 500 shown can be used to implement the electronic device 110 discussed above.

[0091] like Figure 5 As shown, electronic device 500 is in the form of a general-purpose electronic device. Components of electronic device 500 may include, but are not limited to, one or more processing units or processors 510, memory 520, storage devices 530, one or more communication units 540, one or more input devices 550, and one or more output devices 560. Processor 510 may be a physical or virtual processor and is capable of performing various processes according to programs stored in memory 520. In a multiprocessor system, multiple processors execute computer-executable instructions in parallel to improve the parallel processing capability of electronic device 500.

[0092] Electronic device 500 typically includes multiple computer storage media. Such media can be any accessible media that is accessible to electronic device 500, including but not limited to volatile and non-volatile media, removable and non-removable media. Memory 520 can be volatile memory (e.g., registers, cache, random access memory (RAM)), non-volatile memory (e.g., read-only memory (ROM), electrically erasable programmable read-only memory (EEPROM), flash memory), or some combination thereof. Storage device 530 can be removable or non-removable media and can include machine-readable media, such as flash drives, disks, or any other media that can be used to store information and / or data and can be accessed within electronic device 500.

[0093] Electronic device 500 may further include additional removable / non-removable, volatile / non-volatile storage media. Although not explicitly stated... Figure 5As shown, disk drives for reading from or writing to removable, non-volatile disks (e.g., "floppy disks") and optical disk drives for reading from or writing to removable, non-volatile optical disks can be provided. In these cases, each drive can be connected to a bus (not shown) via one or more data media interfaces. Memory 520 may include computer program product 525 having one or more program modules configured to perform various methods or actions of various examples.

[0094] The communication unit 540 enables communication with other electronic devices via a communication medium. Additionally, the functionality of the components of the electronic device 500 can be implemented using a single computing cluster or multiple computing machines capable of communicating via communication connections. Therefore, the electronic device 500 can operate in a networked environment using logical connections to one or more other servers, networked personal computers, or another network node.

[0095] Input device 550 can be one or more input devices, such as a mouse, keyboard, trackball, etc. Output device 560 can be one or more output devices, such as a monitor, speaker, printer, etc. Electronic device 500 can also communicate with one or more external devices (not shown) via communication unit 540 as needed. These external devices include storage devices, display devices, etc., and can communicate with one or more devices that enable user interaction with electronic device 500, or with any device that enables electronic device 500 to communicate with one or more other electronic devices (e.g., network card, modem, etc.). Such communication can be performed via an input / output (I / O) interface (not shown).

[0096] A computer-readable storage medium is provided that stores computer-executable instructions thereon, wherein the computer-executable instructions are executed by a processor to implement the methods described above. A computer program product is also provided, which is tangibly stored on a non-transitory computer-readable medium and includes computer-executable instructions, which are executed by a processor to implement the methods described above.

[0097] The flowcharts and / or block diagrams of the methods, apparatus, devices, and computer program products referred to herein describe various aspects. It should be understood that each block of the flowcharts and / or block diagrams, as well as combinations of blocks in the flowcharts and / or block diagrams, can be implemented by computer-readable program instructions.

[0098] These computer-readable program instructions can be provided to a processor of a general-purpose computer, a special-purpose computer, or other programmable data processing apparatus to produce a machine such that, when executed by the processor of the computer or other programmable data processing apparatus, they create means for implementing the functions / actions specified in one or more blocks of the flowchart and / or block diagram. These computer-readable program instructions can also be stored in a computer-readable storage medium that causes a computer, programmable data processing apparatus, and / or other device to operate in a particular manner; thus, the computer-readable medium storing the instructions comprises an article of manufacture that includes instructions for implementing aspects of the functions / actions specified in one or more blocks of the flowchart and / or block diagram.

[0099] Computer-readable program instructions can be loaded onto a computer, other programmable data processing apparatus, or other device to cause a series of operational steps to be performed on the computer, other programmable data processing apparatus, or other device to produce a computer-implemented process, thereby causing the instructions that execute on the computer, other programmable data processing apparatus, or other device to perform the functions / actions specified in one or more boxes of a flowchart and / or block diagram.

[0100] The flowcharts and block diagrams in the accompanying figures illustrate the architecture, functionality, and operation of possible implementations of systems, methods, and computer program products under various scenarios. In this respect, each block in a flowchart or block diagram may represent a module, segment, or portion of an instruction, which contains one or more executable instructions for implementing the specified logical function. In some alternative implementations, the functions marked in the blocks may occur in a different order than those shown in the figures. For example, two consecutive blocks may actually be executed substantially in parallel, and they may sometimes be executed in reverse order, depending on the functions involved. It should also be noted that each block in the block diagrams and / or flowcharts, and combinations of blocks in the block diagrams and / or flowcharts, can be implemented using a dedicated hardware-based system that performs the specified function or action, or using a combination of dedicated hardware and computer instructions.

[0101] Various examples have been described above. The foregoing descriptions are exemplary and not exhaustive, nor are they limited to the disclosed implementations. Many modifications and variations will be apparent to those skilled in the art without departing from the scope and spirit of the described implementations. The terminology used herein is chosen to best explain the principles, practical applications, or improvements to technology in the market, or to enable others skilled in the art to understand the various implementations disclosed herein.

Claims

1. An interaction method, comprising: Presenting a first comment and first guidance information, the first comment being made to a first video, the first video showing at least one first event related to at least one object, the first comment indicating a second event related to the first video, and the first guidance information indicating content creation related to the second event; Receive a first trigger operation, wherein the first trigger operation is issued in response to the first guidance information; as well as A first interface is presented, which is used for interaction with the second video, and the second video is used to display the second event.

2. The method of claim 1, wherein presenting the first guidance information comprises: Receive a second trigger operation, which is related to the content of the first comment; as well as In response to the second triggering operation being a predetermined type of operation, the first guidance information is presented.

3. The method of claim 2, wherein in response to the second triggering operation being a predetermined type operation, presenting the guidance information includes at least one of the following: In response to the second triggering operation of issuing the first comment content, the first guidance information is presented; or In response to the second triggering operation being an interaction performed on the first comment content, the first guidance information is presented.

4. The method according to claim 1, wherein the first guidance information is used to guide the viewing of the second video, and the first interface is used to play the second video.

5. The method of claim 4, wherein the first interface includes second guidance information, the second guidance information instructing the creation of a third video, and the method further includes: Receive a third trigger operation, wherein the third trigger operation is issued in response to the second guidance information; as well as In response to the third triggering operation, a first creation interface is presented, which displays creation information used to generate the third video. The creation information is related to the second video.

6. The method according to claim 4, further comprising: On the first interface, a third guide message is presented, which instructs the user to watch the first video. Receive a fourth trigger operation, which is issued in response to the third guidance information; as well as In response to the fourth triggering operation, the first video is presented.

7. The method of claim 1, wherein the first guidance information instructs the creation of the second video, and the first interface includes a second creation interface for receiving creation information, the creation information describing the second event and for generating the second video.

8. The method according to claim 7, further comprising: In the second creation interface, object selection information and event description information are presented. The object selection information indicates the selection of one or more objects from the at least one object, and the event description information is used to describe the second event. Receive an interaction operation, wherein the interaction operation is issued in response to at least one of the object selection information and the event description information; as well as The creation information is received and matched with the interactive operation.

9. The method according to claim 8, wherein receiving the creation information includes: In response to the interactive operation instruction to confirm the object selection information and the event description information, the object selection information and the event description information are received as the creation information.

10. The method of claim 8, wherein receiving the creation information includes at least one of the following: In response to the interactive operation instruction to select one or more additional objects from the at least one object, updated object selection information is received as part of the creation information, the updated object selection information indicating the selection of the one or more additional objects. In response to the interactive operation instruction to update the event description information, the updated event description information is received as part of the creation information.

11. The method of claim 8, wherein the second authoring interface presents at least one label, the at least one label respectively indicating at least one event summary, and receiving the interactive operation includes: A selection operation is received, the selection operation instructing the selection of a tag from the at least one tag, the selected tag indicating a first event summary of the second event. The updated event description information matches the first event summary.

12. A device for interaction, comprising: The guidance information presentation module is configured to present first comment content and first guidance information, wherein the first comment content is issued to a first video, the first video displays at least one first event related to at least one object, the first comment content indicates a second event related to the first video, and the first guidance information indicates content creation related to the second event; The trigger operation receiving module is configured to receive a first trigger operation, which is issued in response to the first guidance information; as well as The interface presentation module is configured to present a first interface, which is used for interaction related to the second video, and the second video is used to display the second event.

13. An electronic device, comprising: At least one processor; as well as At least one memory coupled to the at least one processor and storing instructions for execution by the at least one processor, the instructions causing the electronic device to perform the method according to any one of claims 1 to 11 when executed by the at least one processor.

14. A computer-readable storage medium having stored thereon computer-executable instructions that can be executed by a processor to implement the method according to any one of claims 1 to 11.

15. A computer program product comprising computer-executable instructions, wherein the computer-executable instructions, when executed by a processor, implement the method according to any one of claims 1 to 11.