Session processing method and apparatus, and device and storage medium
By identifying the target object in the multimedia content and displaying the corresponding conversation content in the instant messaging software, the conversation interaction method is enriched, the problem of single conversation method is solved, and the user experience is improved.
Patent Information
- Application Number
- PCT/CN2024/077248
- Authority / Receiving Office
- WO · WO
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2023-02-17
- Filing Date
- 2024-02-15
- Publication Date
- 2025-09-25
AI Technical Summary
The existing instant messaging software has a relatively simple conversation mode and cannot meet the users' growing demand for diversified conversation interaction.
By determining the target object in the multimedia content corresponding to the first user and displaying the conversation content based on the target object on the conversation page, including audio, emoticons, virtual images and other interactive methods, the conversation interaction function is enriched.
It has realized diversified conversation interaction methods, improved the user's conversation experience and immersion, and met the user's growing demand for diversified interaction.
Smart Images

Figure CN2024077248_25092025_PF_FP_ABST
Abstract
Description
A session processing method, device, equipment and storage medium
[0001] This application claims priority to the Chinese invention patent application entitled “A session processing method, apparatus, device and storage medium” and application number 202310154269.4, filed on February 17, 2023. The entire contents of that application are incorporated by reference into this application. Technical Field
[0002] The present disclosure relates to the field of data processing, and in particular to a session processing method, apparatus, device, and storage medium. Background Art
[0003] With the continuous development of Internet technology, the interaction between users relying on the network is becoming increasingly close. Instant Messaging (IM) has become a more popular communication method on the Internet. Various instant messaging software have emerged one after another, bringing greater convenience to people's interactions in work, study, life, entertainment and other aspects.
[0004] However, the conversation modes supported by current instant messaging software are relatively simple and cannot meet users' growing demand for diversified conversation interaction.
[0005] Summary of the Invention
[0006] In order to solve the above technical problems, an embodiment of the present disclosure provides a session processing method.
[0007] In a first aspect, the present disclosure provides a session processing method, the method comprising:
[0008] Determining a first target object corresponding to the first user; wherein the first target object is an object in the multimedia content;
[0009] In response to a sending operation on the first conversation content, a second conversation content corresponding to the first conversation content is displayed on the conversation page; wherein the second conversation content is determined based on the first conversation content and the first target object.
[0010] Optionally, before determining the first target object corresponding to the first user, the method further includes:
[0011] On the object selection page, multiple object identifiers to be selected are displayed according to the preset classification method;
[0012] Accordingly, determining the first target object corresponding to the first user includes:
[0013] In response to a selection operation on a target object identifier among the multiple candidate object identifiers, the object corresponding to the target object identifier is determined as a first target object corresponding to the first user.
[0014] Optionally, after displaying multiple identifiers of objects to be selected according to a preset classification method on the object selection page, the method further includes:
[0015] In response to a preset triggering operation for a first object identifier among the to-be-selected object identifiers, displaying an object details page corresponding to the first object identifier;
[0016] The object information corresponding to the first object identifier is displayed on the object details page.
[0017] Optionally, the object information includes audio information, and after displaying the object information corresponding to the first object identifier on the object details page, the method further includes:
[0018] In response to an input operation for first audio data, second audio data corresponding to the first audio data is played; wherein the second audio data is determined based on audio information corresponding to the first object identifier and the first audio data.
[0019] Optionally, the first conversation content is user input content, and the second conversation content is audio conversation content determined based on the user input content and audio information of the first target object.
[0020] Optionally, in response to the sending operation on the first conversation content, before displaying the second conversation content corresponding to the first conversation content on the conversation page, the method further includes:
[0021] Receiving text content entered into an input box on the conversation page;
[0022] Displaying at least one optional conversation content corresponding to the text content on the conversation page; wherein the optional conversation content is determined from a set of sentences associated with the first target object based on the text content;
[0023] In response to a selection operation on a target conversation content among the at least one selectable conversation content, the target conversation content is determined as the first conversation content; wherein the target conversation content is audio conversation content or text conversation content.
[0024] Optionally, a session content generation control is provided on the session page, and in response to the sending operation on the first session content, before displaying the second session content corresponding to the first session content on the session page, the processing further includes:
[0025] In response to a triggering operation on the session content generation control, a first session content is generated, and a sending operation on the first session content is performed; wherein the first session content is determined from a statement set associated with the first target object based on the historical content of this session.
[0026] Optionally, after determining the first target object corresponding to the first user, the method further includes:
[0027] In response to a selection operation on a first target object corresponding to the first user, the first image identifier of the first user on the conversation page is updated and displayed as a second image identifier related to the first target object.
[0028] Optionally, the method further includes: displaying a set of emoticon images corresponding to the first target object on the conversation page;
[0029] In response to a sending operation on a target emoticon image in the emoticon image set corresponding to the first target object, the target emoticon image is displayed on the conversation page.
[0030] Optionally, the method further includes: displaying a scene conversation page in response to a conversation trigger operation for a three-dimensional virtual image on the conversation page; wherein an image display window and a conversation window are provided on the scene conversation page, the image display window is used to display the three-dimensional virtual images corresponding to the first user and the second user respectively in a preset image display scene, and the conversation window is used to display the conversation content between the first user and the second user.
[0031] Optionally, the method further includes: displaying the latest conversation contents corresponding to the first user and the second user respectively in the image display window.
[0032] Optionally, the method further includes: in response to a sending operation on the third conversation content, sending the third conversation content to a target server; wherein the target server is configured to determine an interactive action corresponding to the third conversation content;
[0033] receiving an interactive action corresponding to the third conversation content returned by the target server, and controlling the three-dimensional virtual image corresponding to the first user in the image display window to execute the interactive action;
[0034] And / or, receiving the interactive action corresponding to the second user from the target server, and controlling the three-dimensional virtual image corresponding to the second user in the image display window to execute the interactive action corresponding to the second user.
[0035] Optionally, the method further comprises: displaying an interactive action panel on the scene conversation page; wherein at least one interactive action icon is provided on the interactive action panel;
[0036] In response to a triggering operation on a target interactive action icon among the at least one interactive action icon, the three-dimensional virtual image corresponding to the first user in the image display window is controlled to perform an interactive action corresponding to the target interactive action icon.
[0037] Optionally, the method further includes: sending a communication request to establish a voice connection or a video connection with the second user based on the scenario conversation page;
[0038] When a voice connection or a video connection is successfully established between the first user and the second user, the conversation content input based on the voice connection or the video connection is displayed in the image display window.
[0039] Optionally, the method further includes: in response to a video generation operation for a target scene session triggered on a scene session page, generating a scene session video based on the content displayed in the image display window in the target scene session.
[0040] In a second aspect, an embodiment of the present disclosure provides a session processing device, the device comprising:
[0041] A first determining module is configured to determine a first target object corresponding to a first user; wherein the first target object is an object in multimedia content;
[0042] The first display module is configured to display, in response to a sending operation on the first conversation content, a second conversation content corresponding to the first conversation content on the conversation page; wherein the second conversation content is determined based on the first conversation content and the first target object.
[0043] In a third aspect, the present disclosure provides a computer-readable storage medium, wherein instructions are stored in the computer-readable storage medium. When the instructions are executed on a terminal device, the terminal device implements the above method.
[0044] In a fourth aspect, the present disclosure provides a session processing device, comprising: a memory, a processor, and a computer program stored in the memory and executable on the processor, wherein the processor implements the above-mentioned method when executing the computer program.
[0045] In a fifth aspect, the present disclosure provides a computer program product, which includes a computer program / instructions, and the computer program / instructions implement the above method when executed by a processor.
[0046] The technical solution provided by the embodiments of the present disclosure has at least the following advantages compared with the prior art:
[0047] The disclosed embodiments provide a conversation processing method. Specifically, a first target object corresponding to a first user is first determined, where the first target object is an object within multimedia content. Then, upon receiving a send operation for the first conversation content, second conversation content corresponding to the first conversation content is displayed on a conversation page. The second conversation content is determined based on the first conversation content and the first target object. Thus, the disclosed embodiments implement conversation interaction functionality on a conversation page based on the characteristics of objects within multimedia content, enriching instant messaging conversation methods and thereby meeting users' growing demand for diverse conversation interactions.
[0048] It should be understood that the content described in this summary section is not intended to limit the key features or important features of the embodiments of the present disclosure, nor is it intended to limit the scope of the present disclosure. Other features of the present disclosure will become easily understood through the following description. BRIEF DESCRIPTION OF THE DRAWINGS
[0049] The above and other features, advantages and aspects of the embodiments of the present disclosure will become more apparent with reference to the following detailed description in conjunction with the accompanying drawings. In the accompanying drawings, the same or similar reference numerals represent the same or similar elements, wherein:
[0050] FIG1 is a flow chart of a session processing method provided by an embodiment of the present disclosure;
[0051] FIG2 is a schematic diagram of an object selection page provided by an embodiment of the present disclosure;
[0052] FIG3 is a schematic diagram of an object information card provided by an embodiment of the present disclosure;
[0053] FIG4 is a schematic diagram of an object screening page provided by an embodiment of the present disclosure;
[0054] FIG5 is a schematic diagram of an object details page provided by an embodiment of the present disclosure;
[0055] FIG6 is a schematic diagram of a conversation page provided by an embodiment of the present disclosure;
[0056] FIG7 is a schematic diagram of another session page provided by an embodiment of the present disclosure;
[0057] FIG8 is a schematic diagram of another session page provided by an embodiment of the present disclosure;
[0058] FIG9 is a schematic diagram of a scene loading page provided by an embodiment of the present disclosure;
[0059] FIG10 is a schematic diagram of a scenario conversation page provided by an embodiment of the present disclosure;
[0060] FIG11 is a schematic diagram of another scenario conversation page provided by an embodiment of the present disclosure;
[0061] FIG12 is a schematic diagram of a video sharing page provided by an embodiment of the present disclosure;
[0062] FIG13 is a schematic diagram of an object matching page provided by an embodiment of the present disclosure;
[0063] FIG14 is a schematic diagram of a matching result page provided by an embodiment of the present disclosure;
[0064] FIG15 is a structural diagram of a session processing device provided by an embodiment of the present disclosure;
[0065] FIG16 is a structural diagram of a session processing device provided in an embodiment of the present disclosure. DETAILED DESCRIPTION
[0066] In order to more clearly understand the above-mentioned objectives, features and advantages of the present disclosure, the scheme of the present disclosure will be further described below. It should be noted that the embodiments of the present disclosure and the features therein can be combined with each other in the absence of conflict.
[0067] In the following description, many specific details are set forth to facilitate a full understanding of the present disclosure, but the present disclosure may also be implemented in other ways different from those described herein; it is obvious that the embodiments in the specification are only part of the embodiments of the present disclosure, rather than all of the embodiments.
[0068] In the related art, the conversation mode supported by instant messaging is relatively simple and cannot meet the users' growing demand for diversified conversation interaction. To this end, the embodiment of the present disclosure provides a conversation processing method. First, a first target object corresponding to a first user is determined, wherein the first target object is an object in multimedia content; then, when a sending operation for the first conversation content is received, the second conversation content corresponding to the first conversation content is displayed on the conversation page; wherein the second conversation content is determined based on the first conversation content and the first target object. It can be seen that the embodiment of the present disclosure implements the conversation interaction function on the conversation page based on the characteristics of the objects in the multimedia content, enriches the conversation mode of instant messaging, and thus meets the users' growing demand for diversified conversation interaction.
[0069] Based on this, an embodiment of the present disclosure further provides a session processing method. Referring to FIG1 , which is a flowchart of a session processing method provided by an embodiment of the present disclosure, the method includes:
[0070] S101: Determine a first target object corresponding to a first user.
[0071] The first target object is an object in multimedia content.
[0072] In the disclosed embodiments, multimedia content may include film and television dramas, documentaries, cartoons, short videos, online games, etc., and objects in the multimedia content may include characters in the multimedia content, such as film and television drama characters, animal characters in animal-related documentaries, cartoon characters, game characters, etc.
[0073] The present disclosure provides the following three methods for determining the first target object corresponding to the first user:
[0074] First, the first user can actively select the first target object through the identifier of the candidate object displayed on the object selection page.
[0075] Specifically, first, on the object selection page, multiple object identifiers for selection are displayed according to a preset classification method; wherein the preset classification method may include a classification method based on at least one of the following: object recommendation degree, the genre of the multimedia content to which the object belongs, and object usage time information. Specifically, the object recommendation degree refers to the degree to which each object is recommended; the genre of the multimedia content to which the object belongs may include film and television genres, game genres, animal world documentary genres, animation genres, etc.; and the object usage time information refers to the time dimension of the first user's usage of each object.
[0076] Figure 2 shows a schematic diagram of an object selection page provided in an embodiment of the present disclosure. The object identifiers displayed on the object selection page include the identifiers of objects recently used by the first user, such as object identifier 201. These identifiers also include identifiers of highly recommended objects and identifiers of objects categorized by the genre of multimedia content to which they belong. The identifiers of the objects to be selected may include object icons, object names, and other items. The present embodiment does not limit the manner in which the identifiers of the objects to be selected are displayed.
[0077] In the disclosed embodiment, based on the candidate object identifiers displayed on the object selection page, a first user can select a first target object based on their needs. Specifically, upon receiving a selection operation for a target object identifier, the object corresponding to the target object identifier is determined as the first target object corresponding to the first user, enabling the first user to actively select the desired object, thereby satisfying the user's object selection needs.
[0078] Second, through the random selection control set on the object selection page, the first user can trigger the random selection of the first target object.
[0079] Specifically, as shown in FIG2 , a random selection control 202 is provided on the object selection page. When a trigger operation for the random selection control 202 is received, the first target object corresponding to the first user can be randomly determined from the objects corresponding to the multiple object identifiers to be selected.
[0080] In one optional embodiment, upon receiving a trigger operation for the random selection control 202, a first target object corresponding to the first user is randomly determined from among the objects corresponding to the multiple candidate object identifiers, and an object information card corresponding to the first target object is displayed. FIG3 is a schematic diagram of an object information card provided in an embodiment of the present disclosure. The object information card may include a "Change" control 301 for triggering a random object switch if the first user is dissatisfied with the object currently displayed on the object information card. Furthermore, the object information card may also include a "Role Details" control 302 for triggering the display of an object details page.
[0081] Third: The first user can filter out the first target object that meets his needs by inputting object screening information.
[0082] Specifically, based on the object filtering information input by the first user, the first target object corresponding to the first user is determined; wherein the object filtering information is used to describe the characteristics of the object required by the first user. As shown in Figure 4, a schematic diagram of an object filtering page provided in an embodiment of the present disclosure is provided, wherein the user can input object filtering information based on a slider (AH shown in Figure 4 are respectively used to represent different filtering information, corresponding to different filtering conditions). Based on the object filtering information input by the user, at least one optional object identifier, such as optional object identifier 401, can be displayed on the object filtering page, and the first user can select the first target object from the at least one optional object identifier. In addition, a "change group" control 402 can also be provided on the object filtering page to trigger the re-determination of the next group of optional object identifiers based on the object filtering information input by the first user when the first user is not satisfied with the optional object identifier currently displayed.
[0083] It should be noted that the above three methods are merely examples of methods for determining the first target object and cannot be used as limitations on the methods for determining the first target object.
[0084] In an optional embodiment, the display of a corresponding object details page can be triggered for the selected object identifier displayed on the object selection page. As shown in Figure 2, upon receiving a preset trigger operation for a first object identifier 201, the object details page corresponding to the first object identifier 201 is displayed. The first object identifier can be any object identifier on the object selection page. Figure 5 is a schematic diagram of an object details page provided in an embodiment of the present disclosure, in which object information corresponding to the first object identifier 201 is displayed on the object details page.
[0085] In another optional embodiment, upon receiving a trigger operation for any optional object identifier (such as optional object identifier 401) displayed on the object filtering page as shown in Figure 4, the object details page corresponding to the optional object identifier can also be displayed, and the object information corresponding to the optional object identifier can be displayed on the object details page.
[0086] The object information displayed on the object details page in the embodiment of the present disclosure may include object introduction information, wonderful video clips associated with the object, a set of sentences associated with the object, object audio information, etc. The first user can gain a deeper understanding of the object through the object information and decide whether to select it as the first target object.
[0087] In an optional embodiment, the first user can understand the audio information of the object by inputting audio data. Specifically, after receiving the first audio data input by the first user, based on the object audio information of the currently displayed object on the object details page, the first audio data is converted into the second audio data corresponding to the object, and the second audio data is played; wherein, the conversion of the first audio data into the second audio data corresponding to the object can be executed on the client side or on the server side. The second audio data is audio data played based on the audio information of the object. By playing the second audio data, the first user can understand the audio characteristics of the currently displayed object, and thus decide whether to select it as the first target object.
[0088] In addition, as shown in FIG5 , a “play this role” control 501 may be provided on the object details page to trigger determination of the current display object on the object details page as the first target object corresponding to the first user.
[0089] S102: In response to a sending operation on the first conversation content, display a second conversation content corresponding to the first conversation content on the conversation page.
[0090] The second conversation content is determined based on the first conversation content and the first target object.
[0091] In the disclosed embodiment, after determining the first target object corresponding to the first user, a conversation page may be displayed, and on the conversation page, based on the characteristics of the first target object, the conversation content may be sent to the second user. The first user and the second user may then interact with each other based on the conversation page. It is worth noting that the conversation page may also be a group conversation page, i.e., a conversation page that supports conversation interaction between multiple users (e.g., at least three users).
[0092] The features of the first target object may include at least one of audio features, language features, action features, etc. The disclosed embodiments do not limit the type of features, and any feature that can express any dimensional characteristics of the first target object can be used as the feature of the first target object.
[0093] In an optional embodiment, when a sending operation for the first session content is received, the first session content can be sent to the server. The server can convert the first session content into second session content based on the characteristics of the first target object corresponding to the first user, and return the second session content so that the second session content can be displayed on the session page of the first user's client.
[0094] On the conversation page, the first user sends conversation content to the second user based on the characteristics of the first target object, that is, the characteristics of the first target object are integrated into the conversation content sent by the first user to the second user, which not only expresses the first user's conversation content to the second user, but also improves the second user's conversation experience when browsing the conversation content from an entertainment perspective.
[0095] In one optional embodiment, after determining the first target object corresponding to the first user, a resource card corresponding to the first target object is first displayed on the conversation page to display the object resources of the first target object so that the second user can understand the first target object corresponding to the first user through the resource card. As shown in Figure 6, a schematic diagram of a conversation page provided in an embodiment of the present disclosure is shown, in which a resource card 600 is displayed on the conversation page. The resource card 600 displays the object resources of the first target object corresponding to the first user. The object resources may include an object image, an object image animation, an object name, and other information that can describe the object.
[0096] In addition, the resource card can also be used to initiate an object session invitation to a second user. The second user can click the resource card on the session page to enter the object selection page, thereby determining a second target object for the second user. When the first user's client receives an object selection message triggered by the second user in response to the object session invitation, an object selection indicator, such as object selection indicator 610 shown in Figure 6, is displayed on the resource card to indicate that the second user has selected the object. If the first user's client does not receive an object selection message triggered by the second user in response to the object session invitation, a prompt message "Waiting for the other party to select a role" is displayed on the object card to indicate that the second user has not yet selected an object.
[0097] In an optional embodiment, as shown in FIG6 , a schematic diagram of a conversation page provided in an embodiment of the present disclosure is provided. On the conversation page, a first user can input voice content through a voice input control 601 or input text content in a text input box 602. After receiving the user input content, the user input content is converted into audio conversation content for the first target object, i.e., the second conversation content, based on the audio information of the first target object corresponding to the first user. The audio conversation content is then sent to the second user on the conversation page, as shown in audio conversation content 603 in FIG6 . The audio conversation content carries the user input content of the first user and expresses the first user's input content using the audio features of the first target object. The user input content can be text input content or voice input content.
[0098] In response to a send operation for user input content, the user input content is sent to a server. The server then converts the user input content into audio conversation content corresponding to the first target object, i.e., the second conversation content, based on the audio information of the first target object corresponding to the first user. A specific implementation method can be employed, and the specific implementation method is not limited in this embodiment.
[0099] In another optional embodiment, a conversation content generation control, such as conversation content generation control 604 shown in FIG6 , may be provided on the conversation page. Upon receiving a trigger operation for conversation content generation control 604 , the second conversation content may be determined from a set of sentences associated with the first target object based on the current conversation history, and the second conversation content may be sent to the second user on the conversation page. Each object is associated with a set of sentences, which may include frequently expressed utterances by the object in the multimedia content (e.g., a film or TV series) to which it belongs. The current conversation history includes the historical conversation content between the first user and the second user, displayed directly or hidden on the current conversation page (i.e., re-displayed by dragging the page).
[0100] In an optional embodiment, after receiving a message from a user triggering a session content generation control, the server determines the intent of the first user's session content by analyzing the historical content of the session. Then, based on the analyzed intent of the session content, the server determines the second session content from a set of sentences associated with the first target object, i.e., the sentence that best expresses the intent of the session content. Then, based on the session page, the server sends the second session content to the second user. The second session content can be text session content or audio session content. The second session content not only expresses the first user's reply intent, but also combines the characteristics of the first target object, thereby improving the session experience of the second user when browsing the second session content.
[0101] In another optional embodiment, as shown in FIG7 , which is a schematic diagram of another conversation page provided by an embodiment of the present disclosure, a first user can enter text content in an input box 701 on the conversation page. After receiving the first text content entered by the first user, at least one optional conversation content is first determined from a set of sentences associated with the first target object based on the first text content. The optional conversation content is a sentence associated with the first target object determined based on keywords in the text content entered by the first user.
[0102] After determining at least one optional conversation content from the set of sentences associated with the first target object, each optional conversation content is displayed on the conversation page for selection by the second user. Upon receiving a selection operation for a target conversation content among the optional conversation content, the selected target conversation content, i.e., the first conversation content, is sent to the second user on the conversation page. The first conversation content can be text conversation content or audio conversation content.
[0103] To enhance the interactive immersion of both parties, the disclosed embodiment may also update the first user's first image identifier on the conversation page to a second image identifier associated with the first target object. The second image identifier may be, for example, an avatar of the first target object. As shown in FIG6 , icon 605 displays the avatar of the first target object corresponding to the first user.
[0104] In addition, in order to further enrich the functional diversity of conversation interaction, the embodiment of the present disclosure can use the expression image set of the first target object to conduct a conversation on the conversation page. Specifically, the expression image set of the first target object is displayed on the conversation page. The expression image set of the first target object is displayed on the expression package panel 606 as shown in Figure 6, which may specifically include multiple expression images, expression animations, etc. On the conversation page, the first user can use the expression image set of the first target object to send the conversation content to the second user. Specifically, the first user can select an expression image or expression animation from the expression package panel 606 and send it to the second user on the conversation page to realize the function of conversation interaction.
[0105] In the conversation processing method provided by the disclosed embodiments, a first target object corresponding to a first user is first determined, where the first target object is an object within multimedia content. Then, upon receiving a send operation for the first conversation content, second conversation content corresponding to the first conversation content is displayed on a conversation page. The second conversation content is determined based on the first conversation content and the first target object. Thus, the disclosed embodiments implement conversation interaction functionality on a conversation page based on the characteristics of objects within multimedia content, enriching instant messaging conversation methods and meeting users' growing demand for diverse conversation interactions.
[0106] Based on the above embodiment, the disclosed embodiment can also conduct conversational interaction based on a three-dimensional virtual image. Specifically, when a conversation trigger operation for a three-dimensional virtual image is received on a conversation page, a scene conversation page is displayed. The scene conversation page is provided with an image display window and a conversation window. The image display window is used to display the three-dimensional virtual images corresponding to the first user and the second user, respectively, in a preset image display scene, and the conversation window is used to display the conversation content between the first user and the second user.
[0107] In an optional embodiment, the conversation triggering operation for the three-dimensional virtual image may include: when the first user and the second user both select an object for setting a three-dimensional virtual image, triggering entry into the scene conversation page. As shown in FIG5 , a three-dimensional virtual image control 500 of an object is displayed on the object details page. When it is determined that the first user and the second user both select an object for a three-dimensional virtual image based on the object selection page, the three-dimensional virtual resource cards corresponding to the two are displayed on the conversation page, and triggering entry into the scene conversation page. As shown in FIG8 , it is a schematic diagram of another conversation page provided in an embodiment of the present disclosure, wherein the conversation page displays a three-dimensional virtual resource card 801 corresponding to the first user and a three-dimensional virtual resource card 802 corresponding to the second user.
[0108] Before entering the scene conversation page, the scene loading page can be displayed first. As shown in Figure 9, it is a schematic diagram of a scene loading page provided in an embodiment of the present disclosure. The scene loading page displays three-dimensional virtual images of objects corresponding to the first user and the second user respectively, which are used for display during the process of loading the scene conversation page.
[0109] In actual applications, after entering the scene conversation page, a preset image display scene can be displayed in the image display window on the scene conversation page, and the three-dimensional virtual images corresponding to the first user and the second user can be displayed in the preset image display scene. As shown in Figure 10, a schematic diagram of a scene conversation page provided by an embodiment of the present disclosure is shown, in which the image display window 1001 on the scene conversation page displays the preset image display scene, and the three-dimensional virtual images corresponding to the first user and the second user can be displayed in the preset image display scene, as well as the latest input conversation content of the first user and the second user in a dialogue format. In addition, the conversation window 1002 on the scene conversation page is used to display the conversation content.
[0110] To further enrich the diversity of conversational interaction functions, embodiments of the present disclosure can also control the first user's corresponding three-dimensional virtual avatar in the avatar display window to perform an interactive action corresponding to the current input content based on the first user's current input content. Specifically, a correspondence between input keywords and interactive actions is pre-set for the first user's corresponding three-dimensional virtual avatar. Based on the input keyword carried in the first user's current input content, the corresponding interactive action is determined, and the first user's corresponding three-dimensional virtual avatar is controlled to perform the interactive action, thereby enriching the diversity of conversational interaction functions.
[0111] In addition, based on the current input content of the second user, the three-dimensional virtual image corresponding to the second user in the image display window on the first user's scene conversation page is controlled to perform the interactive action corresponding to the current input content. The specific implementation method can be understood by referring to the above content.
[0112] In an optional embodiment, an interactive action panel may also be displayed on the scene conversation page, such as interactive action panel 1003 shown in FIG10 . Interactive action panel 1003 may include at least one interactive action icon. Upon receiving a trigger operation for a target interactive action icon, the 3D virtual avatar corresponding to the first user in the avatar display window may be controlled to perform the interactive action corresponding to the target interactive action icon. For example, when the first user clicks the "Sit Down" interactive action icon on interactive action panel 1003, the 3D virtual avatar corresponding to the first user in the avatar display window may be controlled to perform the "Sit Down" interactive action.
[0113] To further enrich the diversity of conversation interaction functions, embodiments of the present disclosure may also establish a voice or video connection between a first user and a second user on the scene conversation page, and subsequently conduct conversation interaction based on the voice or video connection. The specific method for establishing the voice or video connection is not limited in embodiments of the present disclosure.
[0114] In one optional implementation, a voice or video connection is first established between a first user and a second user based on a scenario conversation page. The conversation content entered via the voice or video connection is then displayed in a visual display window. Figure 11 illustrates another scenario conversation page provided by an embodiment of the present disclosure. Within the visual display window on the scenario conversation page, the conversation content entered by the first user and the second user is displayed in a dialog format. The conversation content is entered by both users based on the voice or video connection.
[0115] In actual applications, after receiving the conversation content input by the user based on the voice connection or video connection, voice recognition is performed to obtain the text conversation content, and the text conversation content is displayed in the form of a dialogue in the image display window on the scene conversation page.
[0116] The conversation processing method provided by the embodiment of the present disclosure can further enrich the conversation interaction functions implemented on the conversation page by conducting conversation interaction through a three-dimensional virtual image, thereby enriching the conversation mode of instant messaging and meeting the users' growing demand for diversified conversation interaction.
[0117] Based on the above embodiment, materials during the user conversation process, such as conversation content, content displayed in the image display window, etc., can be obtained and used as materials for generating a video.
[0118] In an optional implementation, when a video generation operation for a target scenario session is received that is triggered on a session scenario page, a scenario session video can be generated based on the content displayed in the image display window in the target scenario session, thereby enriching the generated video types.
[0119] In another optional implementation, the first user can also share or publish the generated scene conversation video. Figure 12 shows a schematic diagram of a video sharing page provided by an embodiment of the present disclosure. In preview window 1201 on the video sharing page, the generated scene conversation video can be previewed and played. The scene conversation video is generated based on the voice input content and the content displayed in the image display window. In conversation record window 1202 on the video sharing page, the conversation content is displayed in real time based on the preview playback time information.
[0120] In addition, the video sharing page is provided with a "Send to Friends" control 1203 and a "Publish" control 1204. The "Send to Friends" control 1203 is used to share the scene conversation video with designated friends, and the "Publish" control 1204 is used to upload the scene conversation video to the server, which is then published by the server, thereby enriching the video-related functions.
[0121] In addition, in each of the above embodiments, before entering the conversation page, the first user first determines the user who is to have a conversation with the first user, that is, the second user.
[0122] In an optional implementation, the second user may be determined for the first user based on a random matching method. Specifically, when a session user matching operation is received, the second user is determined based on a preset object matching condition.
[0123] The preset object matching condition may include selecting a user with an object or selecting a user with an object belonging to the same multimedia content as the first target object.
[0124] As shown in Figure 13, it is a schematic diagram of an object matching page provided in an embodiment of the present disclosure, wherein, on the object matching page, the first user can trigger the random determination of the first target user from all users who have selected an object by selecting the "All Roles" control 1301. The first user can also trigger the random determination of the second user from users who have set an object belonging to the same multimedia content as the first target object by selecting the "Same Drama Role" control 1302. For example, a user who has selected an object in the same TV series as the first target object of the first user is selected as the second user.
[0125] After the second user is randomly determined, the first target object corresponding to the first user and the second target object corresponding to the second user are displayed on the matching result page. FIG14 is a schematic diagram of a matching result page provided in an embodiment of the present disclosure. The matching result page displays a first target object 1401 corresponding to the first user, a second target object 1402 corresponding to the second user, and a start conversation control 1403. After receiving a trigger operation for the start conversation control 1403, the conversation page between the first user and the second user is entered, such as the conversation page shown in FIG6 and FIG7 .
[0126] The conversation content between the first user and the second user is displayed on the conversation page, and a "drama value" can be generated based on the conversation content and the duration of the conversation between the two to represent the excitement of the conversation content.
[0127] In addition, in each of the above embodiments, there may be multiple paths for triggering the execution of the session processing method of the embodiment of the present disclosure.
[0128] In an optional implementation, the object conversation mode can be entered through the message center page. In the object conversation mode, after selecting the second user, the conversation page can be entered, and object conversation interaction can be implemented based on the conversation page. The specific implementation method can be referred to the description of the above embodiments and will not be repeated here.
[0129] In another optional implementation, an entrance to the object conversation mode may be provided on the conversation page. On the conversation page between the first user and the second user, the object conversation mode may be entered through the entrance, and object conversation interaction may be realized based on the conversation page. For the specific implementation method, please refer to the description of the above-mentioned embodiments and will not be repeated here.
[0130] It is worth noting that the embodiment of the present disclosure does not limit the path for triggering the execution of the session processing method of the embodiment of the present disclosure, and the above two paths are only used as examples.
[0131] Based on the above method embodiment, the present disclosure further provides a session processing device. Referring to FIG15 , which is a schematic diagram of the structure of a session processing device provided in an embodiment of the present disclosure, the device includes:
[0132] A first determining module 1501 is configured to determine a first target object corresponding to a first user; wherein the first target object is an object in multimedia content;
[0133] The first display module 1502 is configured to display, in response to a sending operation on the first conversation content, a second conversation content corresponding to the first conversation content on the conversation page; wherein the second conversation content is determined based on the first conversation content and the first target object.
[0134] In an optional embodiment, the device further includes:
[0135] The second display module is used to display multiple identifications of objects to be selected according to a preset classification method on the object selection page;
[0136] Accordingly, the first determining module is specifically configured to:
[0137] In response to a selection operation on a target object identifier among the multiple candidate object identifiers, the object corresponding to the target object identifier is determined as a first target object corresponding to the first user.
[0138] In an optional embodiment, the device further includes:
[0139] a third display module, configured to display an object details page corresponding to a first object identifier among the to-be-selected object identifiers in response to a preset triggering operation on the first object identifier;
[0140] A fourth display module is configured to display object information corresponding to the first object identifier on the object details page.
[0141] In an optional implementation, the object information includes audio information, and the apparatus further includes:
[0142] A playing module is used to play second audio data corresponding to the first audio data in response to an input operation on the first audio data; wherein the second audio data is determined based on the audio information corresponding to the first object identifier and the first audio data.
[0143] In an optional implementation, the first conversation content is user input content, and the second conversation content is audio conversation content determined based on the user input content and audio information of the first target object.
[0144] In an optional embodiment, the device further includes:
[0145] A first receiving module, configured to receive text content input into an input box on the conversation page;
[0146] a fifth display module, configured to display at least one optional conversation content corresponding to the text content on the conversation page; wherein the optional conversation content is determined from a set of sentences associated with the first target object based on the text content;
[0147] The second determining module is configured to, in response to a selection operation on a target conversation content in the at least one optional conversation content, determine the target conversation content as the first conversation content; wherein the target conversation content is audio conversation content or text conversation content.
[0148] In an optional implementation manner, a conversation content generation control is provided on the conversation page, and the apparatus further includes:
[0149] A content generation module is configured to generate first conversation content in response to a trigger operation on the conversation content generation control, and to perform a sending operation on the first conversation content; wherein the first conversation content is determined from a set of statements associated with the first target object based on the historical content of this conversation.
[0150] In an optional embodiment, the device further includes:
[0151] The update display module is configured to update and display the first image identifier of the first user on the conversation page to a second image identifier related to the first target object in response to a selection operation on the first target object corresponding to the first user.
[0152] In an optional embodiment, the device further includes:
[0153] A sixth display module, configured to display a set of expression images corresponding to the first target object on the conversation page;
[0154] The seventh display module is configured to display the target emoticon image on the conversation page in response to a sending operation on the target emoticon image in the emoticon image set corresponding to the first target object.
[0155] In an optional embodiment, the device further includes:
[0156] A scene display module is used to display a scene conversation page in response to a conversation trigger operation for a three-dimensional virtual image on the conversation page; wherein, an image display window and a conversation window are provided on the scene conversation page, the image display window is used to display the three-dimensional virtual images corresponding to the first user and the second user respectively in a preset image display scene, and the conversation window is used to display the conversation content between the first user and the second user.
[0157] In an optional embodiment, the device further includes:
[0158] An eighth display module is configured to display the latest conversation contents corresponding to the first user and the second user respectively in the image display window.
[0159] In an optional embodiment, the device further includes:
[0160] a content sending module, configured to send the third conversation content to a target server in response to a sending operation on the third conversation content; wherein the target server is configured to determine an interactive action corresponding to the third conversation content;
[0161] a second receiving module, configured to receive the interactive action corresponding to the third conversation content returned by the target server, and control the three-dimensional virtual image corresponding to the first user in the image display window to execute the interactive action;
[0162] And / or, a third receiving module is used to receive the interactive action corresponding to the second user from the target server, and control the three-dimensional virtual image corresponding to the second user in the image display window to execute the interactive action corresponding to the second user.
[0163] In an optional embodiment, the device further includes:
[0164] a ninth display module, configured to display an interactive action panel on the scene conversation page; wherein the interactive action panel is provided with at least one interactive action icon;
[0165] The action control module is used to control the three-dimensional virtual image corresponding to the first user in the image display window in response to the trigger operation of the target interactive action icon in the at least one interactive action icon, and execute the interactive action corresponding to the target interactive action icon.
[0166] In an optional embodiment, the device further includes:
[0167] a request sending module, configured to send a communication request for establishing a voice connection or a video connection with the second user based on the scenario conversation page;
[0168] The tenth display module is used to display the conversation content input based on the voice connection or the video connection in the image display window when a voice connection or a video connection is successfully established between the first user and the second user.
[0169] In an optional embodiment, the device further includes:
[0170] The video generation module is used to generate a scene session video based on the content displayed in the image display window in the target scene session in response to a video generation operation for the target scene session triggered on the scene session page.
[0171] In the conversation processing device provided by the disclosed embodiment, a first target object corresponding to a first user is first determined, where the first target object is an object within multimedia content. Then, upon receiving a send operation for the first conversation content, second conversation content corresponding to the first conversation content is displayed on a conversation page. The second conversation content is determined based on the first conversation content and the first target object. This demonstrates that the disclosed embodiment implements conversation interaction functionality on a conversation page based on the characteristics of objects within multimedia content, enriching instant messaging conversation methods and meeting users' growing demand for diverse conversation interactions.
[0172] In addition to the above-mentioned method and apparatus, the embodiment of the present disclosure further provides a computer-readable storage medium, which stores instructions. When the instructions are executed on a terminal device, the terminal device implements the session processing method described in the embodiment of the present disclosure.
[0173] The embodiment of the present disclosure further provides a computer program product, which includes a computer program / instructions. When the computer program / instructions are executed by a processor, the session processing method described in the embodiment of the present disclosure is implemented.
[0174] In addition, the embodiment of the present disclosure further provides a session processing device, as shown in FIG16 , which may include:
[0175] Processor 1601, memory 1602, input device 1603, and output device 1604. The session processing device may include one or more processors 1601; FIG16 illustrates one processor as an example. In some embodiments of the present disclosure, processor 1601, memory 1602, input device 1603, and output device 1604 may be connected via a bus or other means; FIG16 illustrates a bus connection as an example.
[0176] Memory 1602 can be used to store software programs and modules. Processor 1601 executes the various functional applications and data processing of the session processing device by running the software programs and modules stored in memory 1602. Memory 1602 may primarily include a program storage area and a data storage area. The program storage area may store an operating system, at least one application required for a function, and the like. Furthermore, memory 1602 may include high-speed random access memory and non-volatile memory, such as at least one disk storage device, flash memory device, or other volatile solid-state storage device. Input device 1603 may be used to receive input digital or character information and generate signal input related to user settings and function control of the session processing device.
[0177] Specifically in this embodiment, the processor 1601 will load the executable files corresponding to the processes of one or more applications into the memory 1602 according to the following instructions, and the processor 1601 will run the applications stored in the memory 1602, thereby realizing the various functions of the above-mentioned session processing device.
[0178] It should be noted that, in this document, relational terms such as "first" and "second" are used only to distinguish one entity or operation from another entity or operation, and do not necessarily require or imply any actual relationship or order between these entities or operations. Moreover, the terms "comprises," "comprising," or any other variations thereof are intended to cover non-exclusive inclusion, so that a process, method, article, or device comprising a series of elements includes not only those elements, but also other elements not explicitly listed, or elements inherent to such process, method, article, or device. In the absence of further limitations, an element defined by the phrase "comprising a ..." does not exclude the presence of other identical elements in the process, method, article, or device comprising the element.
[0179] The foregoing description is intended only to provide specific embodiments of the present disclosure, intended to enable those skilled in the art to understand and implement the present disclosure. Various modifications to these embodiments will be readily apparent to those skilled in the art, and the general principles defined herein may be implemented in other embodiments without departing from the spirit or scope of the present disclosure. Therefore, the present disclosure is not intended to be limited to the embodiments described herein, but rather to be construed in the broadest manner consistent with the principles and novel features disclosed herein.
Claims
1. A session processing method, the method comprising: Determining a first target object corresponding to the first user; wherein the first target object is an object in the multimedia content; In response to a sending operation on the first conversation content, a second conversation content corresponding to the first conversation content is displayed on a conversation page; wherein the second conversation content is determined based on the first conversation content and the first target object.
2. The method according to claim 1, before determining the first target object corresponding to the first user, further comprising: On the object selection page, multiple object identifiers to be selected are displayed according to the preset classification method; Accordingly, determining the first target object corresponding to the first user includes: In response to a selection operation on a target object identifier among the multiple candidate object identifiers, the object corresponding to the target object identifier is determined as a first target object corresponding to the first user.
3. The method according to claim 2, further comprising: after displaying multiple identifiers of objects to be selected according to a preset classification method on the object selection page; In response to a preset triggering operation for a first object identifier among the to-be-selected object identifiers, displaying an object details page corresponding to the first object identifier; The object information corresponding to the first object identifier is displayed on the object details page.
4. The method according to claim 3, wherein the object information includes audio information, and after displaying the object information corresponding to the first object identifier on the object details page, further comprising: In response to an input operation for first audio data, second audio data corresponding to the first audio data is played; wherein the second audio data is determined based on audio information corresponding to the first object identifier and the first audio data. 5 . The method according to claim 1 , wherein the first conversation content is user input content, and the second conversation content is audio conversation content determined based on the user input content and audio information of the first target object.
6. The method according to claim 1, wherein, in response to the sending operation on the first conversation content, before displaying the second conversation content corresponding to the first conversation content on the conversation page, further comprises: Receiving text content entered into an input box on the conversation page; Displaying at least one optional conversation content corresponding to the text content on the conversation page; wherein the optional conversation content is determined from a set of sentences associated with the first target object based on the text content; In response to a selection operation on a target conversation content among the at least one selectable conversation content, the target conversation content is determined as the first conversation content; wherein the target conversation content is audio conversation content or text conversation content.
7. The method according to claim 1, wherein a session content generation control is provided on the session page, and the method further comprises: In response to a triggering operation on the session content generation control, a first session content is generated, and a sending operation on the first session content is performed; wherein the first session content is determined from a set of statements associated with the first target object based on the historical content of this session.
8. The method according to claim 1, after determining the first target object corresponding to the first user, further comprising: In response to a selection operation on a first target object corresponding to the first user, the first image identifier of the first user on the conversation page is updated and displayed as a second image identifier related to the first target object.
9. The method according to claim 1, further comprising: Displaying a set of emoticon images corresponding to the first target object on the conversation page; In response to a sending operation on a target emoticon image in the emoticon image set corresponding to the first target object, the target emoticon image is displayed on the conversation page.
10. The method according to claim 1, further comprising: In response to the conversation triggering operation for the virtual image on the conversation page, the scene conversation page is displayed; wherein the scene conversation page is provided with an image display window and a conversation window. The image display window is used to display the virtual images corresponding to the first user and the second user respectively in a preset image display scene, and the conversation window is used to display the conversation content between the first user and the second user.
11. The method according to claim 10, further comprising: In the image display window, the latest conversation contents corresponding to the first user and the second user are displayed respectively.
12. The method according to claim 10, further comprising: In response to the sending operation for the third conversation content, the third conversation content is sent to a target server; wherein the target server is used to determine an interactive action corresponding to the third conversation content; receiving an interactive action corresponding to the third conversation content returned by the target server, and controlling the three-dimensional virtual image corresponding to the first user in the image display window to execute the interactive action; And / or, receiving the interactive action corresponding to the second user from the target server, and controlling the three-dimensional virtual image corresponding to the second user in the image display window to execute the interactive action corresponding to the second user.
13. The method according to claim 10, further comprising: Displaying an interactive action panel on the scene conversation page; wherein at least one interactive action icon is provided on the interactive action panel; In response to a triggering operation on a target interactive action icon among the at least one interactive action icon, the three-dimensional virtual image corresponding to the first user in the image display window is controlled to perform an interactive action corresponding to the target interactive action icon.
14. The method according to claim 10, further comprising: Based on the scenario conversation page, sending a communication request to establish a voice connection or a video connection with the second user; When a voice connection or a video connection is successfully established between the first user and the second user, the conversation content input based on the voice connection or the video connection is displayed in the image display window.
15. The method according to any one of claims 10 to 14, further comprising: In response to a video generation operation for a target scenario session triggered on a scenario session page, a scenario session video is generated based on the content displayed in the image display window in the target scenario session.
16. A conversation processing device, comprising: A first determining module is configured to determine a first target object corresponding to a first user; wherein the first target object is an object in multimedia content; The first display module is configured to display, in response to a sending operation on the first conversation content, a second conversation content corresponding to the first conversation content on a conversation page; wherein the second conversation content is determined based on the first conversation content and the first target object.
17. A computer-readable storage medium, wherein instructions are stored in the computer-readable storage medium. When the instructions are executed on a terminal device, the terminal device implements the method according to any one of claims 1 to 15.
18. A session processing device, comprising: A memory, a processor, and a computer program stored in the memory and executable on the processor, wherein when the processor executes the computer program, the method according to any one of claims 1 to 15 is implemented.