Interaction method and device, storage medium and program product
By displaying graphic and text content on the session interface between virtual objects and users, the problem of low information interaction efficiency in the prior art is solved, more efficient and intuitive information transmission is achieved, and user experience is improved.
Patent Information
- Application Number
- CN202510065337.9
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2025-01-15
- Publication Date
- 2025-05-06
AI Technical Summary
In the prior art, the information interaction between virtual objects and users is relatively low, and users need to read a large amount of text or watch a complete video to understand the reply content, which has a poor intuitive experience.
By displaying the graphic content related to the user input message on the session interface, including at least one image group and at least one text block, the graphic content is part of the content in the media work published by the user corresponding to the virtual object.
It improves the efficiency and intuitiveness of information transmission, enables users to understand the reply content of virtual objects more quickly, and improves the interactive experience.
Smart Images

Figure CN119937855A_ABST
Abstract
Description
Technical Field
[0001] The present disclosure relates to the field of terminals, and in particular to an interaction method, device, storage medium and program product. Background Art
[0002] In the related art, user A can create a virtual object on the Internet platform, user B can communicate or interact with the virtual object through the network, and the virtual object can answer corresponding questions based on user B's questions. Summary of the invention
[0003] According to some embodiments of the present disclosure, an interaction method is provided, including: displaying a conversation interface between a first user and a virtual object; receiving an input message on the conversation interface; and displaying graphic content related to the input message in a message body on the conversation interface, the graphic content including at least one image group and at least one text block, the graphic content being part of a media work published by a second user corresponding to the virtual object.
[0004] According to some other embodiments of the present disclosure, an interaction device is provided, comprising: a first display module, configured to display a conversation interface between a first user and a virtual object; a receiving module, configured to receive an input message on the conversation interface; and a second display module, configured to display graphic content related to the input message in a message body on the conversation interface, the graphic content comprising at least one image group and at least one text block, the graphic content being part of a media work published by a second user corresponding to the virtual object.
[0005] According to some further embodiments of the present disclosure, an interaction device is provided, comprising: a processor; and a memory coupled to the processor, for storing instructions, wherein when the instructions are executed by the processor, the processor executes the interaction method of any embodiment of the present disclosure.
[0006] According to some further embodiments of the present disclosure, a computer-readable storage medium is provided, on which a computer program is stored, wherein when the program is executed by a processor, the processor implements the interaction method of any embodiment of the present disclosure.
[0007] According to some further embodiments of the present disclosure, a computer program product is provided, including a computer program or instructions, wherein the computer program or instructions are executed by a processor to implement the interaction method according to any embodiment of the present disclosure.
[0008] Other features, aspects and advantages of the present disclosure will become apparent from the following detailed description of exemplary embodiments of the present disclosure with reference to the attached drawings. BRIEF DESCRIPTION OF THE DRAWINGS
[0009] The following is an explanation of the embodiments of the present disclosure with reference to the accompanying drawings. It should be understood that the drawings described below only relate to some embodiments of the present disclosure and do not constitute a limitation to the present disclosure. In the accompanying drawings:
[0010] Figure 1 A schematic diagram showing a flow chart of an interaction method according to some embodiments of the present disclosure;
[0011] Figure 2A A schematic diagram showing displaying a text and an image in a message body according to some embodiments of the present disclosure;
[0012] Figure 2B Schematic diagram showing displaying a text and an image in a message body in some other embodiments of the present disclosure;
[0013] Figure 2C A schematic diagram showing displaying a text and an image in a message body according to yet other embodiments of the present disclosure;
[0014] Figure 2D A schematic diagram showing displaying a text and an image in a message body in some further embodiments of the present disclosure;
[0015] Figure 3A A schematic diagram showing displaying two texts and two images in a message body according to some embodiments of the present disclosure;
[0016] Figure 3B A schematic diagram showing displaying two texts and two images in a message body in some other embodiments of the present disclosure;
[0017] Figure 4 A schematic diagram showing displaying a text and two images in a message body according to some embodiments of the present disclosure;
[0018] Figure 5A A schematic diagram showing two texts and four images displayed in a message body according to some embodiments of the present disclosure;
[0019] Figure 5B A schematic diagram showing displaying two texts and four images in a message body in some other embodiments of the present disclosure;
[0020] Figure 6 A schematic diagram showing construction of a knowledge base in some embodiments of the present disclosure;
[0021] Figure 7 A block diagram showing an interactive device according to some embodiments of the present disclosure;
[0022] Figure 8 A block diagram showing an interaction device according to some other embodiments of the present disclosure;
[0023] Fig. 9A block diagram of an electronic device according to some embodiments of the present disclosure is shown.
[0024] It should be understood that, for ease of description, the sizes of the various parts shown in the drawings are not necessarily drawn according to the actual proportional relationship. The same or similar reference numerals are used in the various drawings to represent the same or similar parts. Therefore, once an item is defined in one drawing, it may not be further discussed in subsequent drawings. DETAILED DESCRIPTION
[0025] The technical solutions in the embodiments of the present disclosure will be described clearly and completely below in conjunction with the accompanying drawings in the embodiments of the present disclosure. It should be understood that the present disclosure can be implemented in various forms and should not be construed as being limited to the embodiments described herein.
[0026] It should be understood that the various steps recorded in the method embodiments of the present disclosure can be performed in different orders, and / or performed in parallel. In addition, the method embodiments may include additional steps and / or omit the steps shown in the execution. The scope of the present disclosure is not limited in this respect. Unless otherwise specifically stated, the relative arrangement, numerical expressions and numerical values of the parts and steps set forth in these embodiments should be interpreted as being merely exemplary and do not limit the scope of the present disclosure.
[0027] The term “including” and its variations used in the present disclosure are intended to be open terms that include at least the following elements / features but do not exclude other elements / features, that is, “including but not limited to.” The term “based on” means “at least partly based on.”
[0028] It should be noted that the concepts of "first", "second", etc. mentioned in the present disclosure are only used to distinguish different devices, modules or units, and are not used to limit the order or interdependence of the functions performed by these devices, modules or units. Unless otherwise specified, the concepts of "first", "second", etc. are not intended to imply that the objects described in this way must be in a given order in time, space, ranking, or any other manner.
[0029] It should be noted that the modifications of "one" and "plurality" mentioned in the present disclosure are illustrative rather than restrictive, and those skilled in the art should understand that unless otherwise clearly indicated in the context, it should be understood as "one or more".
[0030] The names of the messages or information exchanged between multiple devices in the embodiments of the present disclosure are only used for illustrative purposes and are not used to limit the scope of these messages or information.
[0031] The user information (including but not limited to user device information, user personal information, etc.) and data (including but not limited to data used for analysis, stored data, displayed data, etc.) involved in this disclosure are all information and data authorized by the user or fully authorized by all parties, and the collection, use and processing of relevant data must comply with the relevant laws, regulations and standards of relevant countries and regions, and provide corresponding operation entrances for users to choose to authorize or refuse.
[0032] The embodiments of the present disclosure are described in detail below in conjunction with the accompanying drawings, but the present disclosure is not limited to these specific embodiments. The following specific embodiments can be combined with each other, and the same or similar concepts or processes may not be repeated in some embodiments. In addition, in one or more embodiments, specific features, structures or characteristics can be combined in any suitable manner that will be clear from the present disclosure by a person of ordinary skill in the art.
[0033] It should be understood that the present disclosure does not limit how to obtain the image to be applied / processed. In some embodiments of the present disclosure, it can be obtained from a storage device, such as an internal memory or an external storage device. In other embodiments of the present disclosure, a photographic component can be mobilized to shoot. It should be noted that the acquired image can be a captured image or a frame of an image in a captured video, and is not particularly limited to this.
[0034] In the context of the present disclosure, an image may refer to any of a variety of images, such as a color image, a grayscale image, etc. It should be noted that in the context of the present specification, the type of image is not specifically limited. In addition, the image may be any appropriate image, such as an original image obtained by a camera device, or an image that has been subjected to specific processing, such as preliminary filtering, anti-aliasing, color adjustment, contrast adjustment, normalization, etc. It should be noted that the preprocessing operation may also include other types of preprocessing operations known in the art, which will not be described in detail here.
[0035] Virtual objects created by users on the Internet platform, such as virtual avatars, virtual characters, can also be called AI (Artificial Intelligence) avatars. AI avatars can, to a certain extent, serve as auxiliary intelligent entities of the creators to complete certain tasks, such as answering questions in the comment area and having simple conversations with other users.
[0036] Users can interact with virtual objects corresponding to other users through the Internet platform. In related technologies, after users ask questions, virtual objects can respond and answer text content. This information transmission efficiency is low, and users need to read a lot of text in a short period of time, and the intuitive experience may not be good. If the virtual object replies with a complete video or picture, it is not enough to express the reply content clearly and efficiently, and users need a lot of time to understand the content in the video. How virtual objects can transmit information more efficiently to improve interaction efficiency is a focus worthy of attention.
[0037] Figure 1 A flowchart illustrating an interaction method according to some embodiments of the present disclosure is provided.
[0038] like Figure 1 As shown, the interaction method includes: step S11, displaying a conversation interface between a first user and a virtual object; step S12, receiving an input message on the conversation interface; step S13, displaying graphic content related to the input message in a message body on the conversation interface, the graphic content including at least one image group and at least one text block, and the graphic content is part of the media work published by the second user corresponding to the virtual object.
[0039] The first user can be understood as a user of an Internet platform or a current application, generally referring to a user who interacts with a virtual object. The virtual object corresponds to, for example, a virtual avatar of a second user, and the second user is a user who publishes a virtual avatar and a media work. The present disclosure does not limit how to create a virtual avatar of the second user.
[0040] The conversation interface is an instant messaging interface, such as a chat interface. The conversation interface includes an input area and an interactive information display area. The interactive information display area can be used to display interactive information sent between the first user and the virtual object. The input area can be used to input interactive information and display interactive information to be sent. For example, the first user can input a message in the input area on the conversation interface, and the input message is, for example, an input question. The virtual object can reply to the question input by the first user in the interactive information display area. For example, on the conversation interface, the reply content is displayed in the form of pictures and texts.
[0041] In some embodiments, a machine learning model is used to analyze and understand the input message, and then determine whether to trigger the display of graphics and text. For example, if the user only says hello, the subsequent display of graphics and text will not be triggered. The machine learning model is used to call the media work knowledge base to obtain the graphics and text content related to the input message, and the graphics and text content is displayed in a message body through the conversation interface. The present disclosure does not limit the algorithm for triggering the display of graphics and text, and how to call the knowledge base.
[0042] Exemplarily, a message body corresponds to a message bubble, which is a visual element for presenting information. For example, a message bubble is a frame similar to a bubble shape. A message body describes the expression of a message, and a message body corresponds to a conversation container, which is, for example, a multi-image and text mixed container that meets the conversation consumption experience.
[0043] In the related art, an image is first displayed through a message body, and then the extended reading information corresponding to the image is displayed through another message body. In this embodiment, graphic content is displayed in a message body, and the graphic content includes both images and texts, which can also be called multimedia content. The graphic content comes from the media work published by the second user. The image can be a static image, a dynamic image, or a video clip. For example, at least one image group and at least one text block are displayed in a message body. Each image group in at least one image group can include one or more images. Each text block in at least one text block can include one or more texts.
[0044] The media work is the work content published by the second user, for example, it can be a video or a graphic work. The work content can also be the live broadcast content of the second user. The second user can authorize the video, graphic work, text content, etc. that are allowed to be used. There are many ways to authorize, and this disclosure does not limit the authorization method.
[0045] If the media work is a video, in some embodiments, the graphic content is at least one video frame in the video. For example, the graphic content includes one frame of image, multiple frames of image, one video clip, or multiple video clips in the video. In other embodiments, the graphic content is at least one text associated with the audio or text information such as subtitles, descriptions, etc. in the video. For example, the graphic content includes one text or multiple texts associated with the audio in the video.
[0046] If the media work is a graphic work, the graphic content may be part of the pictures and part of the text in the graphic work.
[0047] Part of the content in the media work is determined based on the input message. For example, the server first parses the work content and converts the structured fields into work knowledge, and retrieves the work knowledge corresponding to the question according to the first user's question, thereby displaying the retrieved work knowledge in a message body.
[0048] The number of images and the number of texts in the graphic content displayed in a message body are determined based on at least one of the duration and the amount of information corresponding to the partial content in the media work.
[0049] For example, if the duration of a part of the content in a media work is longer, the number of images and the number of texts will be greater. If the duration of a part of the content in a media work is shorter, the number of images and the number of texts will be less. Alternatively, if the amount of information corresponding to a part of the content in a media work is greater, the number of images and the number of texts will be greater. If the amount of information corresponding to a part of the content in a media work is smaller, the number of images and the number of texts will be less.
[0050] When the duration of some content in the media work is longer or the amount of information is larger, the number of image groups and the number of text blocks may be increased accordingly. When the duration of some content in the media work is shorter or the amount of information is smaller, the number of image groups and the number of text blocks may be reduced accordingly.
[0051] In the above embodiment, based on the message input by the first user, the corresponding graphic content is displayed in a message body on the conversation interface, so that the interactive content can be presented more efficiently, intuitively and vividly, thereby improving the interaction efficiency. For example, the questions asked by the first user can be answered more efficiently, thereby improving the user experience.
[0052] Next, the graphic and text contents related to the input message displayed in a message body on the conversation interface are introduced.
[0053] In some embodiments, at least one image group is presented in the message body, the at least one image group including a plurality of images arranged based on a playback order of the media work.
[0054] For example, it supports presenting multiple images in a message bubble, and the multiple images are arranged in the order of video playback, that is, multiple images are displayed in a message bubble according to the contextual relationship of the multiple images. In some specific examples, for example, the graphic content includes image A, image B, and image C, and image A, image B, and image C are respectively a frame of image or a video clip in the video, and when the video is played, image A is displayed first, image B is displayed after image A, and image C is displayed after image B, then image A, image B, and image C are arranged in order in a message bubble. Image A, image B, and image C can be arranged left and right in a message bubble, or up and down in a message bubble.
[0055] In this embodiment, arranging the multiple images according to the order in which the media works are played enables the first user to more intuitively understand the content and logic of the images.
[0056] In other embodiments, at least one text block is presented in the message body, and the at least one text block includes a plurality of texts arranged based on the playback order of the media work.
[0057] For example, it supports the presentation of multiple texts in a message bubble, and the multiple texts are arranged in the order of playing the audio in the video, that is, multiple texts are displayed in a message bubble according to the contextual relationship between the multiple texts. In some specific examples, the graphic content includes text A', text B' and text C', and text A', text B' and text C' are the texts corresponding to a section of audio in the video. When the video is playing, the audio corresponding to text A' is played first, the audio corresponding to text B' is played after the audio corresponding to text A', and the audio corresponding to text C' is played after the audio corresponding to text B', then text A', text B' and text C' are arranged in order in a message bubble. Text A', text B' and text C' can be arranged left and right in a message bubble, or they can be arranged up and down in a message bubble.
[0058] In this embodiment, multiple texts are arranged in the order in which the media works are played, so that the first user can understand the content and logic of the texts more clearly.
[0059] In some further embodiments, the image group includes a plurality of images, and the display position of each of the plurality of images in the message body is associated with the playback order of the media works.
[0060] For example, at least one image group includes one or more image groups, and each image group may include one or more images. If an image group includes multiple images, the display positions of the multiple images in a message body are associated with the playback order of the media works. In some specific examples, for example, an image group includes image A, image B, and image C, and image A, image B, and image C are respectively a frame of image or a video clip in a video, and when the video is played, image A is displayed first, image B is displayed after image A, and image C is displayed after image B, then image A, image B, and image C are arranged in order in a message bubble. Image A, image B, and image C can be arranged left and right in a message bubble, or up and down in a message bubble.
[0061] In this embodiment, multiple images in the image group are arranged and displayed according to the playback order of the media works, so that the first user can understand the content and logic of the images more intuitively.
[0062] In other embodiments, at least one image group includes multiple image groups, each image group includes at least one image, and the display position of each image group in the multiple image groups in the message body is associated with the playback order of the media works.
[0063] For example, multiple image groups are displayed in a message body, and one image group includes one or more associated images. The display positions of multiple image groups in a message body are associated with the playback order of the media works, that is, multiple image groups are displayed in a message bubble according to the contextual relationship between the multiple image groups. In some specific examples, for example, the graphic content includes image group 1, image group 2, and image group 3, and image group 1, image group 2, and image group 3 are respectively a group of images in the video, and when the video is played, the images in image group 1 are displayed first, the images in image group 2 are displayed after image group 1, and the images in image group 3 are displayed after image group 2, then image group 1, image group 2, and image group 3 are arranged in order in a message bubble. Image group 1, image group 2, and image group 3 can be arranged left and right in a message bubble, or can be arranged up and down in a message bubble.
[0064] In this embodiment, multiple image groups in the graphic content are arranged and displayed according to the playback order of the media works, so that the first user can understand the graphic content more clearly.
[0065] In other embodiments, the text block includes multiple texts, and the display position of each of the multiple texts in the message body is associated with the playback order of the media works.
[0066] For example, at least one text block includes one or more text blocks, and each text block includes one or more texts. If a text block includes multiple texts, the display positions of the multiple texts in a message body are associated with the playback order of the media works. In some specific examples, for example, a text block includes text A', text B' and text C', and text A', text B' and text C' are the texts corresponding to a segment of audio in the video, and when the video is played, the audio corresponding to text A' is played first, the audio corresponding to text B' is played after the audio corresponding to text A', and the audio corresponding to text C' is played after the audio corresponding to text B', then text A', text B' and text C' are arranged in order in a message bubble. Text A', text B' and text C' can be arranged left and right in a message bubble, or up and down in a message bubble.
[0067] In this embodiment, multiple texts in the text block are arranged and displayed according to the playback order of the media works, so that the first user can understand the content and logic of the text more clearly.
[0068] In other embodiments, at least one text block includes multiple text blocks, each text block includes at least one text, and the display position of each text block in the message body is associated with the playback order of the media works.
[0069] For example, multiple text blocks are displayed in some message bodies, and one text block includes one or more associated texts. The display positions of multiple text blocks in a message body are associated with the playback order of the media works, that is, multiple text blocks are displayed in a message bubble according to the contextual relationship between the multiple text blocks. In some specific examples, for example, the graphic content includes text block 1', text block 2' and text block 3', and text block 1', text block 2' and text block 3' are respectively a group of texts corresponding to the audio in the video, and when the audio is played, the audio corresponding to text block 1' is played first, the audio corresponding to text block 2' is played after the audio corresponding to text block 1', and the audio corresponding to text block 3' is played after the audio corresponding to text block 2', then text block 1', text block 2' and text block 3' are arranged in order in a message bubble. Text block 1', text block 2' and text block 3' can be arranged left and right in a message bubble, or can be arranged up and down in a message bubble.
[0070] In this embodiment, multiple text blocks in the graphic content are arranged and displayed according to the playback order of the media works, so that the first user can understand the content and logic of the text more clearly.
[0071] In the above embodiments, it has been described that multiple images, multiple image groups, multiple images in an image group, multiple texts, multiple text blocks, and multiple texts in a text block in the graphic content are displayed in the conversation interface according to the order in which the media works are played. The following will continue to describe how to display graphic content related to the input message in the message body on the conversation interface.
[0072] In some embodiments, at least one image group and at least one text block are mixed and displayed in the message body, and the mixed display means that the image and the text are not superimposed together. For example, in a message bubble, from top to bottom, the first row is the text block, the second row is the image group, the third row is the text block, the fourth row is the image group, and so on.
[0073] The display position of each of the at least one image group in the message body and the display position of each of the at least one text block in the message body are determined based on the content relevance between the at least one image group and the at least one text block.
[0074] For example, the contents of image group 1 and text block 1' are correlated, the contents of image group 2 and text block 2' are correlated, and the contents of image group 3 and text block 3' are correlated, then they can be arranged in the order of text block 1', image group 1, text block 2', image group 2, text block 3', image group 3, that is, image group 1 and text block 1' are displayed next to each other, image group 2 and text block 2' are displayed next to each other, and text block 3' and image group 3 are displayed next to each other.
[0075] If an image group includes an image and a text block having content relevance to the image group includes a text, then the image in the image group and the text in the text block having content relevance to the image group are also displayed adjacent to each other.
[0076] If an image group includes multiple images, and a text block with content relevance to the image group includes a text, the multiple images in the image group and the text in the text block with content relevance to the image group are also displayed adjacent to each other. The multiple images in the image group can be displayed in the order in which the media work is played.
[0077] If an image group includes an image, and a text block with content relevance to the image group includes multiple texts, the image in the image group and the multiple texts in the text block with content relevance to the image group are also displayed adjacent to each other. The multiple texts in the text block can be displayed in the order in which the media work is played.
[0078] If an image group includes multiple images, and a text block with content relevance to the image group includes multiple texts, each image and the text with content relevance to the image are displayed in close proximity. The multiple images in the image group and the multiple texts in the text block can be displayed in the order in which the media work is played.
[0079] Through the above settings, the image content and the corresponding text content are set according to the relevance, which can improve the readability of the first user, improve the interaction efficiency, and further improve the experience of the first user.
[0080] Mixed presentation of at least one image group and at least one text block in the message body includes: displaying text in the text block on at least one side of at least one image corresponding to the text block.
[0081] For example, the corresponding text is displayed above the image. Alternatively, the corresponding text is displayed below, on the left, or on the right of the image. Of course, the corresponding text can also be displayed around the image.
[0082] In some embodiments, at least one image group and at least one text block are superimposed and displayed in a message body, and the display position of each image group in the at least one image group and the display position of each text block in the message body are determined based on the content relevance between the at least one image group and the at least one text block.
[0083] Overlay display means that the image contains text content and the text is located within the image.
[0084] For example, if the contents of image group 1 and text block 1' are correlated, the contents of image group 2 and text block 2' are correlated, and the contents of image group 3 and text block 3' are correlated, then text block 1' is located in image group 1, text block 2' is located in image group 2, and text block 3' is located in image group 3'.
[0085] Superimposing and displaying at least one image group and at least one text block in a message body includes: displaying text in the text block in at least one image corresponding to the text block.
[0086] If an image group includes an image and a text block having content relevance to the image group includes a text, the text is displayed in the image.
[0087] If an image group includes multiple images, and a text block having content relevance to the image group includes a text, a portion of the text can be displayed in each image, and each image has content relevance to a corresponding portion of the text. The multiple images in the image group can be displayed in an arranged manner according to the playback order of the media work.
[0088] If an image group includes an image, and a text block having content relevance to the image group includes multiple texts, then the multiple texts are displayed in the image. The multiple texts in the text block can be displayed in the order in which the media work is played.
[0089] If an image group includes multiple images, and a text block having content relevance to the image group includes multiple texts, each text is displayed in a corresponding image.
[0090] Through the above settings, the image content and the corresponding text content are set according to the relevance, which can improve the readability of the first user, improve the interaction efficiency, and further improve the experience of the first user.
[0091] The above describes displaying graphic and text content related to an input message in a message body on a conversation interface, and the style of the message body is determined based on at least one of the following: the amount of content information in the graphic and text content; the number of image groups and the number of text blocks in the graphic and text content; the number of images in each image group of at least one image group, and the number of texts in each text block of at least one text block; the amount of information of each image in each image group, and the amount of information of each text in each text block.
[0092] For example, the greater the amount of information in the graphic content, the more complex the style of the message body is, and the more images and texts the message body can carry.
[0093] For another example, the number of image groups and text blocks that can be displayed in the message body matches the number of image groups and text blocks in the graphic content. For example, the number of image groups and text blocks displayed in the message body in the vertical arrangement direction of the conversation interface is related to the number of image groups and text blocks in the graphic content.
[0094] For another example, the number of images displayed in the message body in the left-right direction of the conversation interface matches the number of images in each image group, and the number of texts displayed in the message body in the left-right direction of the conversation interface matches the number of texts in each text block.
[0095] For another example, by comparing the information volume of the image with the information volume of the text, it is determined whether to display multiple images or multiple texts in the message body. For another example, whether to display dynamic images, static images or video images in the message body is determined according to the amount of information that the image can express.
[0096] Next, we will combine FIG. 2A to FIG. 5B Describe the conversation interface that contains graphic and text content.
[0097] Figure 2A A schematic diagram showing a text and an image in a message body in some embodiments of the present disclosure is shown. Figure 2A As shown, the dialogue interface of this embodiment is the dialogue interface between the first user AA and the AI avatar corresponding to the second user BB. In the dialogue, the first user AA asks the AI avatar corresponding to the second user BB "I want to go hiking recently", and the input message 201 "I want to go hiking recently" will be displayed on the dialogue interface.
[0098] After analyzing the question of the first user AA, the machine learning model determines the image A and the text 202 "I recently went hiking in ** Mountain, ** Country! It was pretty good" in the knowledge base corresponding to the media work published by the second user BB. The AI avatar corresponding to the second user BB displays the image A and the corresponding text 202 in a message bubble 4. The text 202 can be located above the image A, so that the first user AA can understand the content of the reply of the AI avatar corresponding to the second user BB.
[0099] In other embodiments, Figure 2B As shown, Figure 2B A schematic diagram showing displaying a text and an image in a message body in some other embodiments of the present disclosure, Figure 2B The Chinese text 202 is located in the image A, so that the first user AA can read the text while looking at the image, thereby improving the interaction efficiency. Figure 2B and Figure 2A The differences are shown below, and the similarities will not be repeated here.
[0100] In response to the first user AA repeatedly inputting the same information on the conversation interface, the style of the displayed graphic content or message body is adjusted. For example, the first user AA has asked a question to the AI avatar corresponding to the second user BB once, and the AI avatar corresponding to the second user BB has also given an answer, but the first user AA may not be satisfied with the answer of the AI avatar corresponding to the second user BB, so the first user AA inputs message 204 again on the conversation interface, and the input message 204 includes the same question "I want to go hiking recently." After the machine learning model analyzes the question of the first user AA, it re-determines the image B and the text 203 "I recently went hiking in ** Mountain in ** Country! It was not bad, and there were antelopes in the Gobi Desert." Figure 2C As shown, the AI clone corresponding to the second user BB displays the image B and the corresponding text 203 in a message bubble 4. Figure 2C and Figure 2A The differences are as follows, and the similarities are not repeated. Those skilled in the art should understand that the first user AA may ask the same question multiple times, and therefore, the AI avatar corresponding to the second user BB may answer the same question multiple times.
[0101] In some embodiments, Figure 2D As shown, an image E and corresponding text 202 can also be displayed in a message bubble. Image E can be a video clip or a dynamic image in a media work. Since video clips and dynamic images carry more information, the graphic content can transmit information more efficiently, thereby improving the user experience. Figure 2D and Figure 2A It is understandable that the above video clips or dynamic images can be played automatically or after being triggered by the user.
[0102] FIG. 2A to FIG. 2D The graphic content can be regarded as including a set of image groups and a text block, the image group includes an image, and the text block includes a text.
[0103] Figure 3A A schematic diagram showing some embodiments of the present disclosure showing two texts and two images in a message body, such as Figure 3A As shown, the dialogue interface of this embodiment is an interface for a dialogue between the first user AA and the AI avatar corresponding to the second user BB. In the dialogue, the first user AA asks the AI avatar corresponding to the second user BB "I want to go hiking recently", and the input message 301 "I want to go hiking recently" is displayed on the dialogue interface.
[0104] After the machine learning model analyzes the question of the first user AA, it identifies image A, image B, and text 302 "I recently went hiking in ** Mountain, ** Country! It was pretty good", and text 303 "There are antelopes in the Gobi Desert" in the knowledge base corresponding to the media work published by the second user BB. Image A and text 302 are content-related, and image B and text 303 are content-related. The order in which images and texts are arranged is associated with the order in which the corresponding media works are played. When the media work is played, image A is in front and image B is in the back; text 302 is in front and text 303 is in the back. Therefore, if Figure 3A As shown, in a message bubble 4, the order of text 302, image A, text 303, and image B is mixed and displayed according to the user reading order. This facilitates the first user AA to understand the content of the reply of the AI avatar corresponding to the second user BB.
[0105] The image displayed in the message body may be any frame of the media work corresponding to the question of the first user AA, for example, Figure 3B As shown, the image corresponding to the text 302 is image D. The display order of the images is arranged according to the order in which the media works are played. Figure 3B and Figure 3A The differences are shown below, and the similarities will not be repeated here.
[0106] FIG. 3A to FIG. 3B The graphic content can be regarded as including two image groups and two text blocks, each image group includes an image, and each text block includes a text.
[0107] Figure 4 A schematic diagram showing some embodiments of the present disclosure showing a text and two images in a message body, such as Figure 4 As shown, the dialogue interface of this embodiment is an interface for a dialogue between the first user AA and the AI avatar corresponding to the second user BB. In the dialogue, the first user AA asks the AI avatar corresponding to the second user BB "Have you been to ** country?", and the conversation interface will display the input message 401 "Have you been to ** country?"
[0108] After the machine learning model analyzes the question of the first user AA, it determines the image F and the image G and the text 402 "Of course I've been there! ** country is very beautiful, the * waterfall and the food city are very good" in the knowledge base corresponding to the media work published by the second user BB. When the media work is played, the image F is in front and the image G is in the back. The text can also be displayed in the image, for example, "** country" is displayed in the image F, and "beautiful" is displayed in the image G. Therefore, in a message bubble 4, the text 402 is displayed above the image F and the image G, and the image F is on the left side of the image G. In addition, the AI avatar corresponding to the second user BB can also actively ask questions, for example, "What are you interested in in ** country?" In addition, there can also be prompt options below the message bubble, such as "Route planning to the * waterfall", "Food city check-in guide", etc. In response to the first user AA triggering the prompt option, the prompt message is automatically input and sent in the input area of the conversation interface to improve the interaction efficiency. The first user AA can also edit the prompt message in the input area and then send the edited message.
[0109] Figure 4 The graphic content can be considered as including an image group and a text block, and the image group includes two images.
[0110] Figure 5A A schematic diagram showing two texts and four images in a message body in some embodiments of the present disclosure is shown. Figure 5A As shown, the first user AA asks the AI avatar corresponding to the second user BB "*Waterfall", and the input message 501 "*Waterfall" is displayed on the conversation interface. The question can be a question entered by the first user AA after triggering the prompt option, or it can be a question entered directly by the first user AA in the input area.
[0111] After the machine learning model analyzes the question of the first user AA, it determines the image H, image I, image J and image K, as well as the text 502 "* Waterfall is a group of waterfalls formed by multiple waterfalls, 2.7 kilometers wide", and the text 503 "Even if you take out one waterfall here, it's pretty good!" in the knowledge base corresponding to the media work published by the second user BB. The image H and image I have content relevance with the text 502, and the image J and image K have content relevance with the text 503. When the media work is played, the text 502 is played first, and then the text 503, and the image H, image I, image J and image K are played in sequence. Therefore, in a message bubble 4, the text 502 is displayed above the image H and image I, and the text 503 is displayed above the image J and image K. The image H and image I can be regarded as an image group, and the image J and image K can be regarded as an image group. The image group containing the image H and image I is located before the image group containing the image J and image K. In addition, the text "* Waterfall" can be inserted into the image H, and the text "2.7 kilometers wide" can be inserted into the image I. In addition, the AI clone corresponding to the second user BB can continue to actively ask questions, for example Figure 5A As shown in the video, "If you want to know anything else, just ask me! Hehehe."
[0112] Figure 5A The graphic content can be regarded as including two text blocks and two image groups, each text block includes a text, and each image group includes two images. For example, the first image group includes image H and image I, and the second image group includes image J and image K.
[0113] In other embodiments, Figure 5B As shown, the image group including image H and image I and the corresponding text "*The waterfall is a waterfall group formed by multiple waterfalls, 2.7 kilometers wide" are located on the left side of the message bubble, and the image group including image J and image K and the corresponding text "Even if you take out one waterfall here, it's pretty good!" are located on the right side of the message bubble. This embodiment only describes Figure 5B and Figure 5A The differences are shown below, and the similarities will not be repeated here.
[0114] Those skilled in the art should understand that the above description is for example only, and the arrangement of graphic content displayed in the message body can be expanded to more styles, as long as the graphic content can clearly reflect the content of a media work corresponding to the user input message.
[0115] The image content in the embodiment of the present disclosure is part of the content in the media work published by the second user corresponding to the virtual object. For example, the image content is obtained by calling the knowledge base constructed by the media work. Figure 6 As shown, the construction of the disclosed knowledge base is introduced.
[0116] Figure 6 A schematic diagram of building a knowledge base according to some embodiments of the present disclosure is shown. The embodiment includes steps S61 to S63 and is executed by a server.
[0117] In step S61, chapter contents in the media work are extracted.
[0118] For example, a chapter extraction model is used to extract chapter content from a media work.
[0119] For example, the media work is divided into multiple segments through algorithmic asynchronous parsing, and the highlight content of each segment is extracted from different angles, such as the picture, theme, description, etc.
[0120] In step S62, chapter contents and videos are structured into fields.
[0121] For example, according to the entry identifier, the full text summary, chapter title, chapter description, chapter start and end timestamps, chapter fragment start and end timestamps, and content of the fragments within the chapter are obtained. After the video is converted into pictures and texts, frame images are extracted according to the chapter start and end timestamps and stored in the fragments within the chapter.
[0122] In step S63, various fields are used to construct a knowledge base corresponding to the media work.
[0123] In the above embodiment, the media works are parsed and structured into fields to form a knowledge base. If the image display is triggered according to the user input content, the large model can be used to call the corresponding graphic knowledge in the knowledge base according to the user input message, and call the multi-graphic mixed conversation container to present the graphic knowledge in the form of a message body on the client's conversation interface.
[0124] Those skilled in the art will appreciate that, in the above method of a specific embodiment, the order in which the steps are written does not imply a strict execution order and does not constitute any limitation on the implementation process. The specific execution order of the steps should be determined by their functions and possible internal logic.
[0125] The above are some interactive methods provided by some embodiments of the present disclosure. Figure 7 Describe the interaction device in some embodiments of the present disclosure.
[0126] Figure 7 A block diagram showing an interactive device according to some embodiments of the present disclosure is shown. Figure 7 As shown, the interaction device 7 includes a first display module 71 , a receiving module 72 and a second display module 73 .
[0127] The first display module 71 is configured to display a conversation interface between the first user and the virtual object; the receiving module 72 is configured to receive an input message on the conversation interface; and the second display module 73 is configured to display graphic content related to the input message in the message body on the conversation interface, the graphic content including at least one image group and at least one text block, and the graphic content is part of the media work published by the second user corresponding to the virtual object.
[0128] The interactive device 7 can be used to perform Figure 1 Step S11 to step S13.
[0129] In some embodiments, the media work includes a video, and the portion of the content includes at least one of the following: at least one video frame in the video; at least one text associated with audio in the video.
[0130] In some embodiments, graphic and text content related to the input message is displayed in a message body on a conversation interface, including at least one of the following: presenting at least one image group in the message body, at least one image group including multiple images arranged based on the playback order of the media works; presenting at least one text block in the message body, at least one text block including multiple texts arranged based on the playback order of the media works.
[0131] In some embodiments, the image group includes multiple images, and the display position of each of the multiple images in the message body is associated with the playback order of the media works.
[0132] In some embodiments, at least one image group includes multiple image groups, each image group includes at least one image, and the display position of each image group in the multiple image groups in the message body is associated with the playback order of the media works.
[0133] In some embodiments, the text block includes multiple texts, and the display position of each of the multiple texts in the message body is associated with the playback order of the media works.
[0134] In some embodiments, at least one text block includes multiple text blocks, each text block includes at least one text, and the display position of each text block in the message body is associated with the playback order of the media works.
[0135] In some embodiments, graphic and text content related to the input message is displayed in the message body on the conversation interface, including: mixed display of at least one image group and at least one text block in the message body, and / or superimposed display of at least one image group and at least one text block in the message body, the display position of each image group in the at least one image group in the message body, and the display position of each text block in the at least one text block in the message body are determined based on the content relevance between the at least one image group and the at least one text block.
[0136] In some embodiments, displaying at least one image group and at least one text block in a mixed manner in the message body includes: displaying text in the text block on at least one side of at least one image corresponding to the text block.
[0137] In some embodiments, superimposing and displaying at least one image group and at least one text block in a message body includes: displaying text in the text block in at least one image corresponding to the text block.
[0138] In some embodiments, a portion of the content in the media work is determined based on the input message.
[0139] In some embodiments, the number of images and the number of texts in the graphic content are determined based on at least one of the duration and the amount of information corresponding to the partial content in the media work.
[0140] In some embodiments, the images in at least one of the image groups include at least one of static images, video clips, and dynamic images.
[0141] In some embodiments, the style of the message body is determined based on at least one of the following: the amount of content information in the graphic content; the number of images and the number of texts in the graphic content; the number of images in each image group of at least one image group, and the number of texts in each text block of at least one text block; the amount of information of each image in each image group, and the amount of information of each text in each text block.
[0142] It should be noted that the above-mentioned units are only logical modules divided according to the specific functions they implement, and are not used to limit the specific implementation methods. For example, they can be implemented in software, hardware, or a combination of software and hardware. In actual implementation, the above-mentioned units can be implemented as independent physical entities, or can also be implemented by a single entity (for example, a processor (CPU or DSP, etc.), an integrated circuit, etc.). In addition, the above-mentioned units are shown with dotted lines in the accompanying drawings to indicate that these units may not actually exist, and the operations / functions they implement can be implemented by the processing circuit itself.
[0143] The above description of various embodiments tends to emphasize the differences between the various embodiments. The same or similar parts can be referenced to each other, and for the sake of brevity, the present disclosure will not repeat them.
[0144] Figure 8 A block diagram of an interaction device according to another embodiment of the present disclosure is shown. Figure 8 As shown, the interactive device 8 includes a memory 81 and a processor 82 coupled to the memory. The processor 82 is configured to execute the interactive method of any of the above embodiments based on the instructions stored in the memory.
[0145] The memory 81 is used to store one or more computer-readable instructions. The memory 81 may include any combination of various forms of computer-readable storage media, such as volatile memory and / or non-volatile memory, including but not limited to random access memory (RAM), dynamic random access memory (DRAM), static random access memory (SRAM), read-only memory (ROM), flash memory. The memory 81 may store, for example, an operating system, an application, a boot loader (BootLoader), a database, and other programs, and may also store various applications and various data.
[0146] The processor 82 is used to run computer-readable instructions to implement the interaction method described in any of the above embodiments or the interaction method described in any of the above embodiments. The specific implementation of each step of the interaction method can refer to the above embodiments, and the repeated parts are not repeated here.
[0147] The processor 82 may be configured to execute Figure 1 The processor 82 may be embodied as various processing devices, such as a central processing unit (CPU), a network processor (NP), etc.; it may also be a digital signal processor (DSP), an application specific integrated circuit (ASIC), a field programmable gate array (FPGA) or other programmable logic devices, discrete gate or transistor logic devices, discrete hardware components. The central processing unit (CPU) may be an X86 or ARM architecture, etc.
[0148] The processor 82 and the memory 81 may communicate with each other directly or indirectly. For example, the processor 82 and the memory 81 may communicate with each other via a network. The network may include a wireless network, a wired network, and / or any combination of a wireless network and a wired network. The processor 82 and the memory 81 may also communicate with each other via a system bus, which is not limited in the present disclosure.
[0149] It should be noted that Figure 8 The components of the interactive device 8 shown are only exemplary and non-restrictive. The interactive device 8 may also have other components according to actual application requirements. The processor 82 may control other components in the interactive device 8 to perform desired functions.
[0150] The interactive device may be implemented by software, firmware and / or hardware, and may be integrated into a device installed with a related application program.
[0151] Fig. 9 A block diagram of an electronic device of some embodiments of the present disclosure is shown. In some embodiments, the interactive device is presented in the form of an electronic device.
[0152] Fig. 9The electronic device 9 shown may be a computer system with a dedicated hardware structure, which can execute corresponding functions when a relevant application program is installed.
[0153] Electronic devices include, but are not limited to, mobile terminals such as smart phones, laptops, personal digital assistants (PDA), tablet personal computers (Tablet PC), PMPs (portable multimedia players), vehicle-mounted terminals (such as vehicle-mounted navigation terminals), wearable devices, etc., and fixed terminals such as digital televisions, desktop computers, etc.
[0154] like Fig. 9 As shown, a central processing unit (CPU) 91 performs various processes according to a program stored in a read-only memory (ROM) 92 or a program loaded from a storage section 98 to a random access memory (RAM) 93. In the RAM 93, data required when the CPU 91 performs various processes, etc., is stored as needed. The central processing unit is merely exemplary, and it may also be other types of processors, such as the various processors described above. The ROM 92, the RAM 93, and the storage section 98 may be various forms of computer-readable storage media. It should be noted that although Fig. 9 ROM 92, RAM 93 and storage section 98 are shown separately in the figure, but one or more of them may be combined or located in the same or different memory or storage modules.
[0155] The CPU 91, the ROM 92, and the RAM 93 are connected to one another via a bus 94. To the bus 94, an input / output interface 95 is also connected.
[0156] The following components are connected to the input / output interface 95: an input section 96, such as a touch screen, a touch pad, a keyboard, a mouse, an image sensor, a microphone, an accelerometer, a gyroscope, etc.; an output section 97, including a display, such as a cathode ray tube (CRT), a liquid crystal display (LCD), a speaker, a vibrator, etc.; a storage section 98, including a hard disk, a magnetic tape, etc.; and a communication section 99, including a network interface card such as a LAN card, a modem, etc. The communication section 99 allows communication processing to be performed via a network such as the Internet. It is easy to understand that although Fig. 9 Some of the electronic devices 9 are shown to communicate via a bus 94, but they may also communicate via a network or other means, wherein the network may include a wireless network, a wired network, and / or any combination of a wireless network and a wired network.
[0157] A drive 910 is also connected to the input / output interface 95 as needed. A removable medium 911 such as a magnetic disk, an optical disk, a magneto-optical disk, a semiconductor memory or the like is mounted on the drive 910 as needed so that a computer program read therefrom is installed into the storage section 98 as needed.
[0158] When the above-described series of processing is realized by software, a program constituting the software can be installed from a network such as the Internet or a storage medium such as the removable medium 911 .
[0159] According to an embodiment of the present disclosure, the process described above with reference to the flowchart can be implemented as a computer software program. For example, some embodiments of the present disclosure include a computer program product, which, when running on a computer, enables the computer to implement the interactive method described in any of the aforementioned embodiments. The computer program product includes a computer instruction carried on a computer-readable medium, containing a program code for executing the interactive method shown in the flowchart. In such an embodiment, the computer instruction can be downloaded and installed from a network through a communication part 99, or installed from a storage part 98, or installed from a ROM 92. When the computer program is executed by a CPU 91, the interactive method of an embodiment of the present disclosure is executed.
[0160] It should be noted that, in the context of the present disclosure, a computer-readable medium may be a tangible medium that may contain or store a program for use by an instruction execution system, apparatus, or device or for use in conjunction with an instruction execution system, apparatus, or device.
[0161] The computer readable medium may be a computer readable storage medium, or a computer readable signal medium, or any combination of the two.
[0162] Computer-readable storage media include, but are not limited to, electrical, magnetic, optical, electromagnetic, infrared, or semiconductor systems, devices or components, or any combination thereof. More specific examples of computer-readable storage media may include, but are not limited to, electrical connections with one or more wires, portable computer disks, hard disks, random access memories (RAM), read-only memories (ROM), erasable programmable read-only memories (EPROM or flash memory), optical fibers, portable compact disk read-only memories (CD-ROMs), optical storage devices, magnetic storage devices, or any suitable combination thereof. In the present disclosure, a computer-readable storage medium may be any tangible medium containing or storing a program that may be used by or in conjunction with an instruction execution system, device, or device. Computer instructions are stored on a computer-readable storage medium, and when the instructions are executed by a processor, the interactive method described in any of the foregoing embodiments is implemented.
[0163] Computer readable signal media may include data signals propagated in baseband or as part of a carrier wave, which carry computer readable program codes. Such propagated data signals may take a variety of forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination of the above. Computer readable signal media may also be any computer readable medium other than a computer readable storage medium, which may send, propagate, or transmit a program for use by or in conjunction with an instruction execution system, device, or device. The program code contained on the computer readable medium may be transmitted using any suitable medium, including but not limited to: wires, optical cables, RF (radio frequency), etc., or any suitable combination of the above.
[0164] The computer-readable medium may be included in the electronic device, or may exist independently without being incorporated into the electronic device.
[0165] In some embodiments, a computer program is further provided, comprising: instructions, which, when executed by a processor, cause the processor to execute the interaction method described in any of the above embodiments. For example, the instructions may be embodied as computer program codes.
[0166] In embodiments of the present disclosure, computer program codes for performing the operations of the present disclosure may be written in one or more programming languages or combinations thereof, including but not limited to object-oriented programming languages, such as Java, Smalltalk, C++, and conventional procedural programming languages, such as "C" language or similar programming languages. The program code may be executed entirely on a user's computer, partially on a user's computer, as an independent software package, partially on a user's computer, partially on a remote computer, or entirely on a remote computer or server. In situations involving a remote computer, the remote computer may be connected to the user's computer via any type of network (including a local area network (LAN) or a wide area network (WAN)), or may be connected to an external computer (e.g., using an Internet service provider to connect via the Internet).
[0167] The flow chart and block diagram in the accompanying drawings illustrate the possible architecture, function and operation of the system, interactive method and computer program product according to various embodiments of the present disclosure. In this regard, each box in the flow chart or block diagram can represent a module, a program segment or a part of a code, and the module, the program segment or a part of the code contains one or more executable instructions for realizing the specified logical function. It should also be noted that in some alternative implementations, the functions marked in the box can also occur in a different order from the order marked in the accompanying drawings. For example, two boxes represented in succession can actually be executed substantially in parallel, and they can sometimes be executed in the opposite order, depending on the functions involved. It should also be noted that each box in the block diagram and / or flow chart, and the combination of the boxes in the block diagram and / or flow chart can be implemented with a dedicated hardware-based system that performs the specified function or operation, or can be implemented with a combination of dedicated hardware and computer instructions.
[0168] The functions described above may be performed at least in part by one or more hardware logic components. For example, without limitation, exemplary hardware logic components that may be used include: field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), systems on chips (SOCs), complex programmable logic devices (CPLDs), and the like.
[0169] Although some specific embodiments of the present disclosure have been described in detail by way of example, it should be understood by those skilled in the art that the above examples are for illustration only and are not intended to limit the scope of the present disclosure. It should be understood by those skilled in the art that the above embodiments may be modified without departing from the scope and spirit of the present disclosure. The scope of the present disclosure is defined by the appended claims.
Claims
1. An interactive method, comprising: Displaying a conversation interface between the first user and the virtual object; Receiving an input message on the conversation interface; as well as Graphical content related to the input message is displayed in a message body on the conversation interface, wherein the graphic content includes at least one image group and at least one text block, and the graphic content is part of the media work published by the second user corresponding to the virtual object.
2. The interactive method according to claim 1, wherein: The media work includes a video, and the partial content includes at least one of the following: at least one video frame in the video; At least one text associated with the audio in the video.
3. The interactive method according to claim 1, wherein: The displaying of graphic and text content related to the input message in a message body on the conversation interface includes at least one of the following: Presenting the at least one image group in the message body, the at least one image group comprising a plurality of images arranged based on a playback order of the media work; The at least one text block is presented in the message body, and the at least one text block includes a plurality of texts arranged based on the playback order of the media work.
4. The interactive method according to claim 3, wherein: The image group includes a plurality of images, and the display position of each of the plurality of images in the message body is associated with the playback order of the media works.
5. The interactive method according to claim 3, wherein: The at least one image group includes a plurality of image groups, each of which includes at least one image, and the display position of each of the plurality of image groups in the message body is associated with the playback order of the media works.
6. The interactive method according to claim 3, wherein: The text block includes multiple texts, and the display position of each of the multiple texts in the message body is associated with the playback order of the media works.
7. The interactive method according to claim 3, wherein: The at least one text block includes multiple text blocks, each of which includes at least one text, and the display position of each of the multiple text blocks in the message body is associated with the playback order of the media works.
8. The interactive method according to claim 1, wherein: The displaying of graphic and text content related to the input message in the message body on the conversation interface includes: The at least one image group and the at least one text block are mixed and displayed in the message body, and / or the at least one image group and the at least one text block are superimposed and displayed in the message body, The display position of each of the at least one image group in the message body and the display position of each of the at least one text block in the message body are determined based on the content relevance between the at least one image group and the at least one text block.
9. The interactive method according to claim 8, wherein: The mixed display of the at least one image group and the at least one text block in the message body comprises: The text in the text block is displayed on at least one side of at least one image corresponding to the text block.
10. The interactive method according to claim 8, wherein: The at least one image group and the at least one text block are displayed in a superimposed manner in the message body, comprising: The text in the text block is displayed in at least one image corresponding to the text block.
11. The interactive method according to claim 1, wherein: Part of the content in the media work is determined based on the input message.
12. The interactive method according to claim 11, wherein: The number of images and the number of texts in the graphic content are determined based on at least one of the duration and the amount of information corresponding to the partial content in the media work.
13. The interactive method according to any one of claims 1 to 12, wherein: The images in the at least one image group include at least one of static images, video clips, and dynamic images.
14. The interactive method according to any one of claims 1 to 12, wherein: The format of the message body is determined according to at least one of the following: The amount of content information in the graphic content; The number of image groups and text blocks in the graphic content; the number of images in each of the at least one image group, and the number of texts in each of the at least one text block; The amount of information of each image in each image group, and the amount of information of each text in each text block.
15. An interactive device, comprising: A first display module is configured to display a conversation interface between the first user and the virtual object; A receiving module, configured to receive an input message on the conversation interface; as well as The second display module is configured to display graphic content related to the input message in the message body on the conversation interface, wherein the graphic content includes at least one image group and at least one text block, and the graphic content is part of the media work published by the second user corresponding to the virtual object.
16. An interactive device, comprising: processor; as well as A memory coupled to the processor, for storing instructions, wherein when the instructions are executed by the processor, the processor executes the interaction method according to any one of claims 1 to 14.
17. A computer-readable storage medium having a computer program stored thereon, wherein: When the program is executed by a processor, the interactive method described in any one of claims 1 to 14 is implemented.
18. A computer program product comprising: The method comprises a computer program or an instruction, which, when executed by a processor, implements the interactive method according to any one of claims 1 to 14.
Citation Information
Patent Citations
Method and device of carrying out mixed arrangement on multimedia information in chat dialog box
CN104159203A
Message processing method and device
CN108712322A
Media content recommendation through chatbots
CN109844708A
Intelligent interaction method and device, computer equipment and storage medium
CN110995569A
Message processing method and device and electronic device
CN114067027A
Cited By
Interface interaction method and device, equipment and storage medium
CN120631216A
Content display method, device and equipment, computer readable storage medium and product
CN121907809A
Interaction method and apparatus, and storage medium and program product
WO2026153358A1