Information interaction method and device, equipment and storage medium

By introducing scene-based interaction methods in digital assistants, users can select suitable scenarios and plug-ins to perform complex tasks, solving the problem of inflexible interaction functions in the prior art and achieving a more efficient and flexible user interaction experience.

CN119916980APending Publication Date: 2025-05-02BEIJING ZITIAO NETWORK TECH CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Application Number
CN202311436108.0
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2023-10-31
Publication Date
2025-05-02

AI Technical Summary

Technical Problem

In the prior art, the interaction function of digital assistants is not flexible enough, and users can only complete limited tasks in one scenario and find it difficult to handle complex tasks.

Method used

By providing a scene-based interaction method, the user can select different scenarios, each scenario is configured with corresponding configuration information, including scene setting information and plug-in information, for performing specific types of tasks. In response to user's message, the system will present scene switching guidance information in the interactive window, automatically switching to a scene suitable for performing tasks.

Benefits of technology

It lowers the threshold for users to use digital assistants, simplifies user operations, realizes diversified interaction with digital assistants, can complete complex tasks naturally and smoothly, and improves user experience.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN119916980A_ABST
    Figure CN119916980A_ABST
Patent Text Reader

Abstract

The embodiment of the invention provides an information interaction method and device, equipment and a storage medium. In the method for information interaction, in response to the fact that a first scene in a set of scenes is selected, interaction with a user is executed in an interaction window of the user and a digital assistant based on the first scene; and in response to receiving a first message of a user in the first scene, presenting scene switching guide information in the interaction window based on the first message and at least one part of the configuration information of each scene in the group of scenes, the scene switching guide information indicating switching from the first scene to the second scene, and executing the task indicated by the first message. In this way, scenes can be automatically prompted and switched based on user requirements, and diversified interactive operation with the digital assistant is achieved.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] Example embodiments of the present disclosure generally relate to the field of computers, and more particularly, to methods, devices, apparatuses, and computer-readable storage media for information interaction. Background Art

[0002] With the rapid development of Internet technology, the Internet has become an important platform for people to obtain and share content. Users can access the Internet through terminal devices to enjoy various Internet services. The terminal device presents the corresponding content through the user interface of the application and interacts with the user and provides services to the user. Therefore, the colorful interactive interface of the application is an important means to improve the user experience. With the development of information technology, various terminal devices can provide various services to people in work and life. For example, applications that provide services can be deployed in terminal devices. Terminal devices or applications can provide users with digital assistant functions to assist users in using terminal devices or applications. How to improve the flexibility of interaction between users and digital assistants is a technical problem to be explored at present. Summary of the invention

[0003] In a first aspect of the present disclosure, a method for information interaction is provided. The method includes: in response to a first scene in a group of scenes being selected, performing interaction with a user in an interaction window between the user and the digital assistant based on the first scene, wherein at least one scene in the group of scenes is configured with configuration information for performing tasks of the corresponding type, and the configuration information includes at least one of the following: scene setting information, plug-in information, wherein the scene setting information is used to describe information related to the corresponding scene, and the plug-in information indicates at least one plug-in for performing tasks in the corresponding scene; and in response to receiving a first message from a user in the first scene, based on the first message and at least a part of the configuration information of each scene in the group of scenes, presenting scene switching guide information in the interaction window, the scene switching guide information indicating switching from the first scene to the second scene to perform the task indicated by the first message.

[0004] In the second aspect of the present disclosure, a device for information interaction is provided. The device includes: an interaction execution module, configured to respond to the first scene in a group of scenes being selected, and to perform interaction with the user in the interaction window between the user and the digital assistant based on the first scene, wherein at least one scene in the group of scenes is configured with configuration information for performing tasks of the corresponding type, and the configuration information includes at least one of the following: scene setting information, plug-in information, wherein the scene setting information is used to describe information related to the corresponding scene, and the plug-in information indicates at least one plug-in for performing tasks in the corresponding scene; and an information presentation module, configured to respond to receiving a first message from the user in the first scene, based on the first message and at least part of the configuration information of each scene in the group of scenes, present scene switching guide information in the interaction window, and the scene switching guide information indicates switching from the first scene to the second scene to perform the task indicated by the first message.

[0005] In a third aspect of the present disclosure, an electronic device is provided. The device includes at least one processing unit; and at least one memory, the at least one memory is coupled to the at least one processing unit and stores instructions for execution by the at least one processing unit. When the instructions are executed by the at least one processing unit, the device executes the method of the first aspect.

[0006] In a fourth aspect of the present disclosure, a computer-readable storage medium is provided, wherein a computer program is stored on the computer-readable storage medium, and the computer program can be executed by a processor to implement the method of the first aspect.

[0007] It should be understood that the content described in this section is not intended to limit the key features or important features of the embodiments of the present disclosure, nor is it intended to limit the scope of the present disclosure. Other features of the present disclosure will become easily understood through the following description. BRIEF DESCRIPTION OF THE DRAWINGS

[0008] The above and other features, advantages and aspects of the embodiments of the present disclosure will become more apparent with reference to the following detailed description in conjunction with the accompanying drawings. In the accompanying drawings, the same or similar reference numerals represent the same or similar elements, wherein:

[0009] Figure 1 A schematic diagram showing an example environment in which embodiments of the present disclosure can be implemented;

[0010] FIG. 2A to FIG. 2D A schematic diagram showing an example client interface of an interactive window according to some embodiments of the present disclosure;

[0011] Figure 3 A flowchart showing a process of information interaction according to some embodiments of the present disclosure is shown;

[0012] Figure 4 A block diagram showing an apparatus for information interaction according to some embodiments of the present disclosure; and

[0013] Figure 5 A block diagram of an electronic device in which one or more embodiments of the present disclosure may be implemented is shown. DETAILED DESCRIPTION

[0014] Embodiments of the present disclosure will be described in more detail below with reference to the accompanying drawings. Although certain embodiments of the present disclosure are shown in the accompanying drawings, it should be understood that the present disclosure can be implemented in various forms and should not be construed as being limited to the embodiments set forth herein. On the contrary, these embodiments are provided to provide a more thorough and complete understanding of the present disclosure. It should be understood that the drawings and embodiments of the present disclosure are only for exemplary purposes and are not intended to limit the scope of protection of the present disclosure.

[0015] In the description of the embodiments of the present disclosure, the term "including" and similar terms should be understood as open inclusion, that is, "including but not limited to". The term "based on" should be understood as "based at least in part on". The term "one embodiment" or "the embodiment" should be understood as "at least one embodiment". The term "some embodiments" should be understood as "at least some embodiments". Other explicit and implicit definitions may also be included below.

[0016] Herein, unless explicitly stated, executing a step “in response to A” does not mean executing the step immediately after “A” but may include one or more intermediate steps.

[0017] It is understandable that the data involved in this technical solution (including but not limited to the data itself, the acquisition, use, storage or deletion of the data) shall comply with the requirements of relevant laws, regulations and relevant provisions.

[0018] It is understandable that before using the technical solutions disclosed in the various embodiments of the present disclosure, the types, scopes of use, usage scenarios, etc. of the information involved in the present disclosure should be informed to relevant users and their authorization should be obtained in an appropriate manner in accordance with relevant laws and regulations. The relevant users may include any type of right holders, such as individuals, enterprises, and groups.

[0019] For example, in response to receiving an active request from a user, a prompt message is sent to the relevant user to clearly prompt the relevant user that the operation requested to be performed will require obtaining and using the information of the relevant user, so that the relevant user can independently choose whether to provide information to software or hardware such as an electronic device, application, server or storage medium that executes the operation of the technical solution of the present disclosure based on the prompt message.

[0020] As an optional but non-limiting implementation, in response to receiving an active request from a relevant user, a prompt message is sent to the relevant user, for example, in the form of a pop-up window, in which the prompt message may be presented in text form. In addition, the pop-up window may also carry a selection control for the user to select "agree" or "disagree" to provide information to the electronic device.

[0021] It is understandable that the above notification and the process of obtaining user authorization are merely illustrative and do not constitute a limitation on the implementation of the present disclosure. Other methods that meet relevant laws and regulations may also be applied to the implementation of the present disclosure.

[0022] Figure 1 A schematic diagram of an example environment 100 in which embodiments of the present disclosure can be implemented is shown. In the example environment 100, a digital assistant 120 and an application 125 are installed in a terminal device 110. A user 140 can interact with the digital assistant 120 and the application 125 via the terminal device 110 and / or an attached device of the terminal device 110.

[0023] In some embodiments, the digital assistant 120 and the application 125 can be downloaded and installed on the terminal device 110. In some embodiments, the digital assistant 120 and the application 125 can also be accessed through other means, such as through a web page. Figure 1 In the environment 100 , in response to the application 125 being started, the terminal device 110 can present the interface 150 of the digital assistant 120 and the application 125 .

[0024] Applications 125 include, but are not limited to, one or more of the following: chat applications (also known as instant messaging applications), document applications, audio and video conferencing applications, email applications, task applications, calendar applications, objectives and key results (OKR) applications, etc. Figure 1 A single application is shown in the figure, but in fact, multiple applications can be installed on the terminal device 110. In some embodiments, the application 125 may include a multi-functional collaboration platform, such as an office collaboration platform (also called an office suite) that can provide integration of multiple types of applications or components to facilitate people to carry out office work, communication and other activities. In the multi-functional collaboration platform, people can start different applications or components as needed to complete corresponding information processing, sharing, communication, etc.

[0025] The application 125 may provide a content entity 126. The content entity 126 may be a content instance created by the user 140 or other users on the application 125. For example, depending on the type of the application 125, the content entity 126 may be a document (e.g., a word document, a pdf document, a presentation, a spreadsheet document, etc.), an email, a message (e.g., a conversation message on an instant messaging application), a calendar, a schedule, a task, an audio, a video, an image, etc.

[0026] In some embodiments, the digital assistant 120 can be provided by a separate application, or can be integrated into an application 120 that can provide content entities. The application for providing the client interface of the digital assistant can correspond to a single-function application or a multi-function collaboration platform, such as an office suite or other collaboration platform that can integrate multiple components. In some embodiments, the digital assistant 120 supports the use of plug-ins. Each plug-in can provide one or more functions of the application. Such plug-ins include, but are not limited to, one or more of the following: search plug-ins, contact plug-ins, message plug-ins, document plug-ins, table plug-ins, mail plug-ins, calendar plug-ins, schedule plug-ins, task plug-ins, and the like.

[0027] The digital assistant 120 is an intelligent assistant of the user, which has intelligent dialogue and information processing capabilities. In an embodiment of the present disclosure, the digital assistant 120 is used to interact with the user 140 to assist the user 140 in using a terminal device or application. An interaction window with the digital assistant 120 may be presented in the client interface. In the interaction window, the user 140 can communicate with the digital assistant 120 by inputting natural language to instruct the digital assistant to assist in completing various tasks, including operations on the content entity 126.

[0028] In some embodiments, the digital assistant 120 can be included in the contact list of the current user 140 in the office suite as a contact of the user 140, or included in the information flow of the chat component. In some embodiments, the user 140 has a corresponding relationship with the digital assistant 120. For example, the first digital assistant corresponds to the first user, the second digital assistant corresponds to the second user, and so on. In some embodiments, the first digital assistant can correspond uniquely to the first user, the second digital assistant can correspond uniquely to the second user, and so on. That is, the first digital assistant of the first user can be specific to or exclusive to the first user. For example, in the process of the first digital assistant providing assistance or service to the first user, the first digital assistant can use its historical interaction information with the first user, the data authorized by the first user that it can access, and the current interaction context with the first user. If the first user is an individual or an individual, the first digital assistant can be regarded as a personal digital assistant. It can be understood that in the disclosed embodiments, the first digital assistant is based on the authorized access to the data to which the permission is granted by the first user. It should be understood that "uniquely corresponding" or similar expressions in the present disclosure are not intended to limit the first digital assistant to be updated accordingly based on the interaction process between the first user and the first digital assistant. Of course, depending on actual application needs, the digital assistant 120 does not have to be specific to the current user 140, but can be a general digital assistant.

[0029] In some embodiments, multiple interaction modes between the user 140 and the digital assistant 120 can be provided, and the multiple interaction modes can be flexibly switched. When a certain interaction mode is triggered, the corresponding interaction area is presented to facilitate the interaction between the user 140 and the digital assistant 120. The user 140 and the digital assistant 120 interact in different ways in different interaction modes, so that it can be flexibly adapted to the interaction needs in different application scenarios.

[0030] In some embodiments, information processing services specific to user 140 can be provided based on historical interaction information between user 140 and digital assistant 120 and / or a data range specific to user 140. In some embodiments, historical interaction information of user 140 interacting with digital assistant 120 in multiple interaction modes can be stored in association with user 140. In this way, in one interaction mode (any one or a designated one of the interaction modes) among the multiple interaction modes, digital assistant 120 can provide services to user 140 based on historical interaction information stored in association with user 140.

[0031] The digital assistant 120 can be called or awakened by an appropriate method (e.g., shortcut keys, buttons, or voice) to present an interactive window with the user 140. By selecting the digital assistant 1201, the interactive window with the digital assistant 120 can be opened. The interactive window may include interface elements for information interaction, such as an input box, a message list, a message bubble, and the like. In other embodiments, the digital assistant 120 can be awakened through an entry control or menu provided in a page, or by inputting a preset instruction.

[0032] The interactive window between the digital assistant 120 and the user 140 may include a conversation window, such as a conversation window in an instant messaging application or an instant messaging module of a target application. In some embodiments, the interactive window between the digital assistant 120 and the user 140 may include a floating window corresponding to the digital assistant.

[0033] In some embodiments, the digital assistant 120 may support an interactive mode of a conversation window, also referred to as a conversation mode. In this interactive mode, a conversation window between the user 140 and the digital assistant 120 is presented, in which the user 140 and the digital assistant 120 interact through conversation messages. In the conversation mode, the digital assistant 120 may perform tasks according to the conversation messages in the conversation window.

[0034] In some embodiments, the conversation mode between the user 140 and the digital assistant 120 can be called or awakened by an appropriate method (e.g., a shortcut key, a button, or voice) to present a conversation window. By selecting the digital assistant 120, a conversation window with the digital assistant 120 can be opened. The conversation window may include interface elements for information interaction, such as an input box, a message list, a message bubble, and the like.

[0035] In some embodiments, the digital assistant 120 may support an interactive mode of a floating window (or floating window), also referred to as a floating window mode. When the floating window mode is triggered, an operation panel (also referred to as a floating window) corresponding to the digital assistant 120 is presented, and the user 140 may issue instructions to the digital assistant 120 based on the operation panel. In some embodiments, the operation panel may include at least one candidate shortcut instruction. Alternatively or additionally, the operation panel may include an input control for receiving instructions. In floating window mode, the digital assistant 120 may perform tasks according to instructions issued by the user 140 through the operation panel.

[0036] In some embodiments, the floating window mode of the user 140 and the digital assistant 120 can also be called or awakened by an appropriate method (e.g., a shortcut key, a button, or voice) to present the corresponding operation panel. In some embodiments, the awakening of the digital assistant 120 can be supported in a specific application, such as a document application, to provide interaction in the floating window mode. In some embodiments, in order to trigger the floating window mode to present the operation panel corresponding to the digital assistant 120, an entry control for the digital assistant 120 can be presented in the application interface. In response to detecting a trigger operation for the entry control, it can be determined that the floating window mode is triggered, and the operation panel corresponding to the digital assistant 120 is presented in the target interface area.

[0037] In some embodiments described below, for ease of discussion, the interaction window between the user and the digital assistant is mainly taken as an example of a conversation window.

[0038] In some embodiments, the terminal device 110 communicates with the server 130 to provide services for the digital assistant 120 and the application 125. The terminal device 110 can be any type of mobile terminal, fixed terminal or portable terminal, including a mobile phone, a desktop computer, a laptop computer, a notebook computer, a netbook computer, a tablet computer, a media computer, a multimedia tablet, a personal communication system (PCS) device, a personal navigation device, a personal digital assistant (PDA), an audio / video player, a digital camera / camcorder, a television receiver, a radio broadcast receiver, an e-book device, a gaming device, or any combination of the foregoing, including accessories and peripherals of these devices or any combination thereof. In some embodiments, the terminal device 110 can also support any type of interface for the user (such as a "wearable" circuit, etc.). The application 130 can be various types of computing systems / servers that can provide computing capabilities, including but not limited to mainframes, edge computing nodes, computing devices in cloud environments, and the like.

[0039] It should be understood that the structure and function of the various elements in the environment 100 are described for exemplary purposes only and do not imply any limitation on the scope of the present disclosure.

[0040] As briefly mentioned above, digital assistants can assist users in using terminal devices or applications. Some applications can provide integrated functions of different plug-ins. In addition to being able to have free conversations with digital assistants, users can also use natural language instructions to enable digital assistants to use different plug-ins to complete some more complex business-related operations related to the application, such as creating documents, inviting schedules, creating tasks, etc. However, since the plug-ins that users can support or the things they can do in a scenario are very limited, if the task that the user wants to handle is very complex, the user cannot complete the task completely in a previously set scenario. This makes the interactive function of the digital assistant not flexible enough.

[0041] According to some embodiments of the present disclosure, an improved scheme for information interaction is proposed. In an embodiment of the present disclosure, in response to the first scene in a group of scenes being selected, interaction with the user is performed in a conversation window between the user and the digital assistant based on the first scene. At least one scene in a group of scenes is configured with configuration information for performing tasks of the corresponding type, and the configuration information includes at least one of the following: scene setting information, plug-in information. The scene setting information is used to describe information related to the corresponding scene, and the plug-in information indicates at least one plug-in for performing tasks in the corresponding scene. In response to receiving a first conversation message from the user in the first scene, based on the first conversation message and at least a part of the configuration information of each scene in a group of scenes, scene switching guide information is presented in the conversation window. The scene switching guide information indicates switching from the first scene to the second scene to perform the task indicated by the first conversation message.

[0042] In general solutions, users can only communicate with digital assistants by selecting plug-ins, but this requires users to have a certain level of cognitive ability, know which plug-ins should be selected in what kind of task scenarios, and need to manually select the plug-ins one by one. According to an embodiment of the present disclosure, by providing scenario-based interaction, the user's entry threshold for using digital assistants is lowered, and the user's operation is simplified. Furthermore, in an embodiment of the present disclosure, in scenario-based interaction, automatic prompts and scene switching based on user needs can also be implemented, allowing users to interact with digital assistants in more matching scenarios. This can achieve diversified interactive operations with digital assistants.

[0043] Some example embodiments of the present disclosure will be described in detail below with reference to examples of the accompanying drawings.

[0044] As described above, in an embodiment of the present disclosure, a digital assistant is used to interact with a user. An interaction window between a user and a digital assistant may be presented in a client interface. The interaction window between a user and a digital assistant may include a conversation window, in which the interaction between the user and the digital assistant may be presented in the form of a conversation message. Alternatively or additionally, the interaction window between a user and a digital assistant may also include other types of windows, such as a window in a floating window mode, in which a user may trigger the digital assistant to perform a corresponding operation by inputting instructions, selecting shortcut instructions, and the like. As an intelligent assistant, a digital assistant has intelligent dialogue and information processing capabilities. In the interaction window, a user inputs an interactive message, and the digital assistant provides a reply message in response to the user input. The client interface for providing a digital assistant may correspond to a single-function application or a multi-function collaboration platform, such as an office suite or other collaboration platform capable of integrating multiple components.

[0045] In some embodiments, the terminal device 110 may present a conversation window between a user and one or more other users. Here, the conversation window between the user and one other user may be, for example, a single chat window between user A and user B, and the conversation window between the user and multiple other users may be, for example, a group chat window between user A and a group of other users. The user may interact with at least one other user by sending and / or receiving messages in the conversation window. The terminal device 110 may, for example, determine the content of the message to be sent by the user in response to detecting the text entered by the user in the input box. The terminal device 110 may also, for example, collect the user audio in response to detecting a trigger operation (such as a click operation, a long press operation, etc.) of the user on the audio control. The terminal device 110 may determine the audio or the text content corresponding to the audio as the content of the message to be sent by the user. It should be noted that the operations performed by the aforementioned terminal device 110 and the operations performed by the terminal device 110 described later may be specifically performed by the relevant application installed on the terminal device 110.

[0046] In the conversation window, the user can input a message through an input box or through other appropriate means (e.g., voice), and the digital assistant can provide a reply message based on the input message and in combination with relevant knowledge. The message in the conversation window is usually a conversation message. Such a conversation message can be regarded as part of a topic.

[0047] FIG. 2A to FIG. 2C A schematic diagram of an example client interface 201 of an interactive window according to some embodiments of the present disclosure is shown. The client interface 201 may be implemented at the terminal device 110. Figure 1 describe FIG. 2A to FIG. 2C .

[0048] In a conversation window between a user and at least one other user, the terminal device 110 (e.g., an instant messaging application installed on the terminal device 110, or a suite application installed on the terminal device 110 that integrates an instant messaging application) may present a trigger control for a digital assistant (e.g., a first digital assistant) in response to detecting a preset input in an input box of the conversation window. Figure 2A As shown, the launch entry of the instant messaging application is presented in area 230 of the interface 201, and the information stream (feed) of the instant messaging application is presented in area 220 of the interface 201. In response to the user selecting any contact or any group in the information stream of the instant messaging application, the terminal device 110 may present a corresponding conversation window in area 210 of the interface 201. For example, in response to the user waking up the digital assistant through a preset operation (such as selecting a digital assistant from a contact list), the terminal device 110 may present a conversation window (also referred to as a main conversation window) for the user to interact with the digital assistant in area 210 of the interface 201.

[0049] In some embodiments, the terminal device 110 may include multiple scenes (for example, it may also be referred to as a group of scenes). At least one scene in a group of scenes is configured with configuration information for performing tasks of the corresponding type. The scene here refers to a collection of tasks of the same type, that is, one scene corresponds to multiple tasks of the same type. The configuration information of the scene includes at least one of the following: scene setting information, plug-in information. The scene setting information is used to describe information related to the corresponding scene. The plug-in information indicates at least one plug-in used to perform the task in the corresponding scene. As will be discussed below, the configuration information of the scene may also include, for example, an indication of the selected model (the model here is called to determine the reply to the user in the corresponding scene), scene guidance information (the scene guidance information is presented to the user after the corresponding scene is selected), at least one recommended question for the digital assistant (at least one recommended question is presented to the user for selection after the corresponding scene is selected), and so on. In some embodiments, the scene setting information and configuration information of the scene can be completed, for example, in a natural language manner, so that the scene creator can easily constrain the output of the model and configure a variety of scenes.

[0050] The scene setting information of the scene can affect the digital assistant's reply to the user to a certain extent, or be used to determine the digital assistant's reply to the user. In some embodiments, the scene setting information is used to construct a prompt word (prompt) input to provide to the model used in the corresponding scene. The digital assistant's reply to the user is based on the output of the model. The scene setting information of the scene may, for example, include a description of the corresponding type of task, the reply style of the digital assistant in the scene, the definition of the workflow to be executed in the corresponding scene, the definition of the reply format of the digital assistant in the corresponding scene, and so on. In some embodiments, the digital assistant will understand the user input with the help of the model and provide a reply to the user based on the output of the model. The model used by the digital assistant can be run locally on the terminal device 110 or on a remote server. By using the scene setting information to construct a part of the prompt word input of the model, the model can be guided to complete the task to be achieved in the corresponding scene. In some embodiments, the model can be a machine learning model, a deep learning model, a learning model, a neural network, etc. In some embodiments, the model can be based on a language model (LM). The language model can have question-answering capabilities by learning from a large amount of corpus. The model can also be based on other appropriate models.

[0051] Through the plug-in information of the scene, the plug-in to be used in the corresponding scene can be configured. In some embodiments, in the corresponding scene, during the operation of the plug-in, the plug-in can also call the model to complete the corresponding task. In some embodiments, a plug-in can also call the open interface provided by other applications (for example, documents, calendars, meetings, etc.) to complete the corresponding task, such as modifying documents, creating schedules, summarizing meetings, etc.

[0052] In some embodiments, the configuration information of the scene may also include a scene name, description information of the scene, and the like. In some embodiments, the terminal device 110 may provide a message card to the user in the conversation window, and at least part of a set of scenes may be presented in the message card. The terminal device 110 may present the scene name and / or description information of the scene of the corresponding scene in association with the scene in the message card. For example, the user may select a scene that meets his or her needs based on the scene name and / or description information of the scene presented in the message card.

[0053] In some embodiments, the configuration information of the scene may also include, for example, the selected model (the model here is called to determine the response to the user in the corresponding scene), scene guidance information (the scene guidance information is presented to the user after the corresponding scene is selected), at least one recommended question for the digital assistant (at least one recommended question is presented to the user for selection after the corresponding scene is selected), etc. The scene guidance information may be, for example, descriptive information of the task instance that can be performed in the scene. In some embodiments, the scene setting information and configuration information of the scene may be configured, for example, in a natural language manner, so that the scene creator can easily constrain the output of the model and configure a variety of scenes.

[0054] For easier understanding, refer to Figure 2D In response to the digital assistant being called up, the terminal device 110 may open a new topic (e.g., the first topic) in the conversation window by default, and present a dividing line corresponding to the first topic in the conversation window. In response to the digital assistant being called up, in response to receiving a trigger on the new topic space 251, a new topic is opened between the user and the digital assistant. The terminal device 110 may present a dividing line 261 corresponding to the new topic and a message card 262 associated with the new topic in the interface 250 by default. In some embodiments, the message card 262 may present a set of scenes that the user can select. In some embodiments, the selection of a scene can be triggered by triggering the scene dialogue control 252.

[0055] In some embodiments, the terminal device 110 may provide a page for creating a target scene in response to receiving a scene creation operation, so that the user can customize the required scene. The terminal device 110 may obtain scene creation information of the target scene via the first page, and the scene creation information includes at least target configuration information of the target scene. The terminal device 110 may then create the target scene based on the acquired target configuration information in response to receiving a creation confirmation operation. Thus, editing of the created scene can be achieved.

[0056] In an embodiment of the present disclosure, if the first scene in a group of scenes is selected, the terminal device 110 performs interaction with the user in the conversation window between the user and the digital assistant based on the first scene. Specifically, the digital assistant in the terminal device 110 can analyze the conversation message from the user in the conversation window based on the configuration information of the first scene, and present a reply message for the conversation message from the user in the conversation window. In some embodiments, the digital assistant can understand the user input with the help of a model and provide a reply to the user based on the output of the model. The model used by the digital assistant can run locally on the terminal device 110 or on a remote server. In some embodiments, the scene setting information of the scene is used to construct a prompt word (prompt) input to provide to the model used in the corresponding scene. The scene setting information of the scene can include, for example, a description of the corresponding type of task, the reply style of the digital assistant in the scene, the definition of the workflow to be executed in the corresponding scene, the definition of the reply format of the digital assistant in the corresponding scene, and so on. By using the scene setting information to construct a part of the prompt word input of the model, the model can be guided to complete the task to be implemented in the corresponding scene. In some embodiments, the model can be a machine learning model, a deep learning model, a learning model, a neural network, etc. In some embodiments, the model may be based on a language model (LM). The language model can have question-answering capabilities by learning from a large amount of corpus. The model may also be based on other appropriate models.

[0057] In some embodiments, the terminal device 110 may also determine whether the type of task corresponding to the current scene matches the task indicated by the conversation message from the user (it may also be referred to as determining whether the current scene matches the conversation message). Taking the first conversation message of the user received in the first scene as an example, the terminal device 110 may, for example, determine the first task type corresponding to the first scene based on the configuration information of the first scene, and then determine whether the task indicated by the first conversation message matches the first task type. In some embodiments, the terminal device 110 may directly determine whether the task indicated by the first conversation message matches the first task type based on part of the information in the configuration information of the first scene. Exemplarily, the terminal device 110 may, for example, determine whether the task indicated by the first conversation message matches the first task type based on the scene name of the first scene and the description information of the scene. In some embodiments, the terminal device 110 may also determine whether the task indicated by the first conversation message matches the first task type with the help of a model. For example, the terminal device 110 may provide the first conversation message and the configuration information of the first scene to the model, and obtain a model output indicating whether the task indicated by the first conversation message matches the first task type from the model.

[0058] If the task indicated by the first conversation message matches the type of task corresponding to the first scene (it can also be referred to as the first conversation message matching the first scene), the terminal device 110 can present the digital assistant's first reply to the first conversation message in the conversation area corresponding to the first scene in the conversation window. It can be understood that the first reply here is determined based on the first scene. In some embodiments, the terminal device 110 can construct a prompt word input for the model based on the first conversation message and at least a part of the configuration information of the first scene. In some embodiments, the part of the configuration information relied on can include at least the scene name of the first scene, description information of the scene, etc. Based on this part of information, it can be determined whether the user's reply to the first conversation message is within the capability of the current scene (i.e., the first scene). In some embodiments, the part of the configuration information relied on can also include other information in the configuration information, such as scene guidance information, recommended questions, models used, and so on.

[0059] In some embodiments, the constructed prompt word input will be provided to the model. The terminal device 110 and / or the digital assistant can receive the model output from the model and determine the first reply to the first conversation message based on the model output. Figure 2AAs shown, in the interface 201, for the conversation message "What is A" from the user, the terminal device 110 can determine whether the task indicated by the conversation message matches the task type indicated by the current scenario, and if it is determined that the two match, obtain and present the reply message "A is X XXXXXXXXXXXXXXXXXXXXXXXX" for the conversation message from the model. Similarly, for the conversation message "Summarize the core ideas" from the user, the terminal device 110 can also determine whether the task indicated by the conversation message matches the task type indicated by the current scenario, and if it is determined that the two match, obtain and present the reply message "The core idea is XXXXXXXXXXXXXXXXXXXXX" for the conversation message from the model.

[0060] When it is determined that the task indicated by the first conversation message is not within the capability range of the current first scene, the configuration information based on other scenes and the first conversation message of the user can be further used to perform intent recognition to determine a suitable scene. Specifically, if the task indicated by the first conversation message does not match the type of the task corresponding to the first scene (it can also be referred to as the first conversation message does not match the first scene), the terminal device 110 can determine that the second scene matches the task indicated by the first conversation message based on at least a part of the configuration information of the optional scene. The terminal device 110 presents the scene switching guidance information based on the selected second scene. The terminal device 110 can obtain at least a part of the configuration information of other optional scenes except the first scene, and select a second scene that matches the task indicated by the first conversation message based on at least a part of the configuration information of the other scenes. In some embodiments, the part of the configuration information relied on can include at least the scene name of each other scene, the description information of the scene, etc. Based on this part of information, it can be determined whether the reply to the first conversation message of the user is within the capability range of the corresponding scene.

[0061] In some embodiments, the part of the configuration information relied on may also include other information in the configuration information, such as scenario guidance information, recommended questions, the model used, and the like. For example, the terminal device 110 may determine the task types corresponding to the other scenarios based on the configuration information of the other scenarios, and then determine the task type that matches the task from the task types corresponding to the other scenarios based on the task indicated by the first conversation message, and the scenario corresponding to the task type is the second scenario that matches the task indicated by the first conversation message. For example, the terminal device 110 may also use the model to determine the second scenario that matches the task indicated by the first conversation message from the other scenarios. For example, the terminal device 110 may construct a prompt word input based on the first conversation message and at least a part of the configuration information of the other scenarios to provide to the model used, and obtain a model output indicating the second scenario to be switched from the model.

[0062] In this case, the terminal device 110 can present the scene switching guidance information in the conversation window. The scene switching guidance information indicates switching from the first scene to the second scene to perform the task indicated by the first conversation message. Regarding the scene switching guidance information, it is similar to determining the reply message. In some embodiments, the terminal device 110 can construct a prompt word input for the model based on the first conversation message and at least a portion of the configuration information of each scene in a set of scenes. The terminal device 110 can provide the prompt word input to the model, receive the model output from the model, and determine the scene switching guidance information based on the model output. In this way, when recommending a new scene to the user, a message card can be presented in the conversation of the current scene as scene switching guidance information to ask the user whether to switch. This can avoid the user's experience of incorrect switching from being affected.

[0063] In some embodiments, in response to the task indicated by the first conversation message not matching the type of task corresponding to the first scene, the terminal device 110 may directly present scene switching guidance information in the conversation window to instruct switching from the first scene to the second scene to perform the task indicated by the first conversation message. Alternatively or additionally, in order to enhance the user's interactive experience, in some embodiments, the terminal device 110 may also present the first reply to the first conversation message determined by the digital assistant based on the first scene in the conversation area corresponding to the first scene in the conversation window while presenting the scene switching guidance information. This can prevent the user from still being able to obtain a reply in the current first scene when the scene switching judgment is inaccurate. Figure 2B As shown, in the interface 201, it is assumed that the scenario selected under the current topic is "content understanding". In response to receiving a conversation message 211 in the first scenario, and the task indicated by the conversation message 211 does not match the type of the task corresponding to the first scenario, the terminal device 110 can present a reply message 212 and scene switching guide information 213 in the interface 201. The reply message 212 is a reply to the conversation message 211 determined by the digital assistant based on the current scenario "content understanding".

[0064] In some embodiments, the terminal device 110 may switch directly to the second scene in response to the first conversation message not matching the first scene, after determining the second scene matching the first conversation message. Then, based on the second scene, interaction with the user may be performed in the conversation window between the user and the digital assistant. In some embodiments, since there is still a situation where the user desires to continue to maintain the first scene, the terminal device 110 may also present a switching entry for the second scene. In response to detecting a triggering operation on the switching entry, the terminal device 110 determines that the second scene is selected. At this point, the terminal device 110 may switch from the first scene to the second scene, and based on the second scene, interaction with the user may be performed in the conversation window between the user and the digital assistant. This switching entry may be presented at any appropriate location in the conversation window. For example, it may be presented in the conversation window in the form of a quick-connect instruction in association with an input box. In some embodiments, the switching entry may be presented in the scene switching guide information. That is, the scene switching guide information may include a switching entry for the second scene. As Figure 2B As shown, a switching entry 214 of the scene "content creation" (the scene "content creation" here is the second scene) may be presented in the scene switching guide information 213. The terminal device 110 may determine that the scene "content creation" is selected in response to receiving a trigger operation on the switching entry 214. The terminal device 110 may then switch to the scene "content creation" and perform interaction with the user in the conversation window based on the scene "content creation".

[0065] In each topic that is launched, a specific scene can be selected for the interaction between the user and the digital assistant. In this article, a "topic" corresponds to a specific context of the interaction. During the interaction process of each topic, the interaction information between the user and the digital assistant will be regarded as contextual information to assist the digital assistant in determining subsequent conversation messages. In some embodiments, topics are sometimes also referred to as or presented as topics. In some embodiments, the terminal device 110 can present a dividing line between different topics in the conversation window (for example, presenting a dividing line between the current topic and the previous topic before the current topic), and present a topic identifier of the current topic at the dividing line. The topic identifier may include a scene identifier. In some embodiments, in response to a scene in the current topic being selected, the terminal device 110 may present a topic identifier including a scene identifier at the dividing line. Exemplarily, in Figure 2B In the interaction between the user and the digital assistant shown, in response to the scene "content understanding" being selected, the terminal device 110 presents a topic identifier including the scene identifier of the scene at the dividing line 205 between the topic corresponding to the scene and the previous topic.

[0066] The topic identifier may also include a task identifier corresponding to the task instance. In some embodiments, the terminal device 110 may also determine the task instance corresponding to the task indicated by the first conversation message in the selected scene. As described above, a scene may include multiple task instances of the same type. For example, the terminal device 110 may analyze the first conversation message received in the scene to determine the task instance corresponding to the first conversation message in the scene. The terminal device 110 may then present the task identifier of the task instance in the dividing line of the topic to prompt the user to perform a specific task under the topic. Figure 2C As shown, in response to receiving a trigger operation on the switching entry 214, the terminal device 110 determines that the scene "content creation" is selected and switches to the scene "content creation". The terminal device 110 can also determine the task instance corresponding to the task indicated by the session message 211 in the scene "content creation". For example, the terminal device 110 can analyze the session message 211 to determine that the task indicated by the session message corresponds to the task instance "generate a plan for summer camp activities" in the scene "content creation". The terminal device 110 then presents the scene identifier "content creation" and the task identifier "generate a plan for summer camp activities" in the dividing line 215.

[0067] In some embodiments, in response to determining that the second scenario is selected, the terminal device 110 may also present the first conversation message and the second reply to the first conversation message determined by the digital assistant based on the second scenario in the conversation area corresponding to the second scenario in the conversation window. Figure 2C As shown, the terminal device 110 may also present a reply message 217 in the interface 201. The reply message 217 is a reply message to the conversation message 211 determined by the terminal device 110 based on the scenario "content creation".

[0068] In the above embodiments, the scene recommendation and scene selection in the conversation window between the user and the digital assistant are described. Although not shown in the accompanying drawings, scene recommendation and scene selection can also be provided in the floating window in the floating window mode of the digital assistant. In some embodiments, the floating window mode of the digital assistant may be evoked in a specific application, for example, when the user edits the target document through the document application. In this case, one or more scenes may also be provided for the user to select. In the floating window mode, the one or more scenes provided may be related to the context or specific application in which the floating window mode of the digital assistant is evoked, so a finer-grained scene can be defined for the context or application scenario, and each scene is configured to perform a corresponding finer-grained task. For example, if the digital assistant is evoked in the document application, then for the document application, finer-grained scenes such as content continuation, content editing, and content understanding can be configured. Further, based on the interaction process between the user and the digital assistant, the scene switching guidance information can be prompted to the user in a similar process to guide the user to switch from the current scene to other more suitable scenes.

[0069] According to various embodiments of the present disclosure, in the process of interaction between a user and a digital assistant, the problem that a user can only complete limited tasks in a scene due to scene configuration restrictions can be solved. By automatically judging user needs and automatically switching scenes, the user can naturally and smoothly complete complex tasks, thereby improving the interaction experience between the user and the digital assistant.

[0070] It should be understood that some embodiments of the present disclosure are described above in conjunction with the specific examples in the drawings, but these specific examples do not limit the scope of the embodiments of the present disclosure. The described embodiments can also be implemented in various other variations.

[0071] Figure 3 FIG. 3 is a flowchart of a process 300 of information interaction according to some embodiments of the present disclosure. The process 300 may be implemented at the terminal device 110. Figure 1 Process 300 is described.

[0072] In block 310, in response to a first scene in a set of scenes being selected, the terminal device 110 performs interaction with the user in an interaction window between the user and the digital assistant based on the first scene, wherein at least one scene in the set of scenes is configured with configuration information for performing a task of a corresponding type, and the configuration information includes at least one of the following: scene setting information and plug-in information. The scene setting information is used to describe information related to the corresponding scene, and the plug-in information indicates at least one plug-in for performing a task in the corresponding scene.

[0073] In box 320, in response to receiving a first message from a user in a first scene, the terminal device 110 presents scene switching guide information in an interactive window based on the first message and at least a portion of the configuration information of each scene in a set of scenes, and the scene switching guide information indicates switching from the first scene to the second scene to perform the task indicated by the first message.

[0074] In some embodiments, process 300 also includes: while presenting the scene switching guidance information, presenting the digital assistant's first reply to the first message in an interaction area corresponding to the first scene within the interaction window, wherein the first reply is determined based on the first scene.

[0075] In some embodiments, the scene switching guide information includes a switching entry for a second scene, and process 300 further includes: in response to receiving a trigger operation on the switching entry, determining that the second scene is selected; and performing interaction with the user in the interactive window based on the second scene.

[0076] In some embodiments, process 300 also includes: in response to determining that the second scene is selected, presenting the first message and the second reply of the digital assistant to the first message in an interaction area corresponding to the second scene within the interaction window, and the second reply is determined based on the second scene.

[0077] In some embodiments, a first scene is selected for interaction of a first topic in an interaction window, and determining that a second scene is selected includes: in response to receiving a trigger operation on a switching entry, starting a second topic in the interaction window, wherein the second scene is selected for interaction of the second topic.

[0078] In some embodiments, at least a portion of the configuration information of each scene in a group of scenes includes at least one of the following: a scene name, or description information of the scene.

[0079] In some embodiments, the configuration information also includes at least one of the following: a selected model, which is called to determine a response to the user in a corresponding scenario; scene guidance information, which is presented to the user after the corresponding scene is selected; or at least one recommended question for the digital assistant, which is presented to the user for selection after the corresponding scene is selected.

[0080] In some embodiments, presenting scene switching guidance information includes: determining, based on at least a portion of the configuration information of the first message and the first scene, that the task indicated by the first message does not match the type of task corresponding to the first scene; and selecting a second scene from a group of scenes based on at least a portion of the configuration information of other scenes in a group of scenes; and presenting the scene switching guidance information based on the selected second scene.

[0081] In some embodiments, process 300 also includes: constructing a prompt word input for the model based on the first message and at least a portion of the configuration information of each scene in a set of scenes; receiving a model output from the model by providing the prompt word input to the model; and determining scene switching guidance information based on the model output.

[0082] In some embodiments, the interaction window includes one or more of the following: a conversation window, a floating window.

[0083] Figure 4 A block diagram of an apparatus 400 for information interaction according to some embodiments of the present disclosure is shown. The apparatus 400 may be implemented in or included in the terminal device 110, for example. Each module / component in the apparatus 400 may be implemented by hardware, software, firmware, or any combination thereof.

[0084] As shown in the figure, the device 400 includes an interaction execution module 410, which is configured to perform interaction with the user in the interaction window between the user and the digital assistant based on the first scene in response to the first scene in the group of scenes being selected, wherein at least one scene in the group of scenes is configured with configuration information for performing tasks of the corresponding type, and the configuration information includes at least one of the following: scene setting information, plug-in information. The scene setting information is used to describe information related to the corresponding scene, and the plug-in information indicates at least one plug-in used to perform the task in the corresponding scene.

[0085] The device 400 also includes an information presentation module 420, which is configured to present scene switching guide information in an interactive window in response to receiving a first message from a user in a first scene, based on the first message and at least a portion of the configuration information of each scene in a group of scenes, wherein the scene switching guide information indicates switching from the first scene to the second scene to perform the task indicated by the first message.

[0086] In some embodiments, the device 400 also includes: a first reply presentation module, configured to present the digital assistant's first reply to the first message in an interaction area corresponding to the first scene in the interaction window while presenting scene switching guidance information, wherein the first reply is determined based on the first scene.

[0087] In some embodiments, the scene switching guidance information includes a switching entry for a second scene, and the device 400 also includes: a second execution module, configured to determine that the second scene is selected in response to receiving a trigger operation on the switching entry; and perform interaction with the user in the interactive window based on the second scene.

[0088] In some embodiments, the device 400 also includes: a second reply presentation module, configured to present the first message and the second reply of the digital assistant to the first message in an interaction area corresponding to the second scene within the interaction window in response to determining that the second scene is selected, wherein the second reply is determined based on the second scene.

[0089] In some embodiments, the first scene is selected for interaction of a first topic in the interaction window, and the second execution module is further configured to: in response to receiving a trigger operation on a switching entry, start a second topic in the interaction window, wherein the second scene is selected for interaction of the second topic.

[0090] In some embodiments, at least a portion of the configuration information of each scene in a group of scenes includes at least one of the following: a scene name, or description information of the scene.

[0091] In some embodiments, the configuration information also includes at least one of the following: a selected model, which is called to determine a response to the user in a corresponding scenario; scene guidance information, which is presented to the user after the corresponding scene is selected; or at least one recommended question for the digital assistant, which is presented to the user for selection after the corresponding scene is selected.

[0092] In some embodiments, the information presentation module 420 is further configured to: determine, based on at least a portion of the configuration information of the first message and the first scene, that the task indicated by the first message does not match the type of task corresponding to the first scene; and select a second scene from a group of scenes based on at least a portion of the configuration information of other scenes in a group of scenes; and present scene switching guide information based on the selected second scene.

[0093] In some embodiments, the device 400 also includes: a guidance determination module, configured to construct a prompt word input for the model based on the first message and at least a portion of the configuration information of each scene in a set of scenes; receive a model output from the model by providing the prompt word input to the model; and determine the scene switching guidance information based on the model output.

[0094] In some embodiments, the interaction window includes one or more of the following: a conversation window, a floating window.

[0095] It should be understood that one or more steps in the above method can be performed by a suitable electronic device or a combination of electronic devices. Such an electronic device or a combination of electronic devices may include, for example, Figure 1 The server 130, the terminal device 110 and / or the combination of the server 130 and the terminal device 110.

[0096] Figure 51 shows a block diagram of an electronic device 500 in which one or more embodiments of the present disclosure may be implemented. It should be understood that Figure 5 The electronic device 500 shown is merely exemplary and should not constitute any limitation on the functionality and scope of the embodiments described herein. Figure 5 The electronic device 500 shown can be used to implement Figure 1 The terminal device 110 and / or Figure 4 The device 400 is shown.

[0097] like Figure 5 As shown, the electronic device 500 is in the form of a general electronic device. The components of the electronic device 500 may include, but are not limited to, one or more processors or processing units 510, a memory 520, a storage device 530, one or more communication units 540, one or more input devices 550, and one or more output devices 560. The processing unit 510 may be an actual or virtual processor and is capable of performing various processes according to a program stored in the memory 520. In a multi-processor system, multiple processing units execute computer executable instructions in parallel to improve the parallel processing capability of the electronic device 500.

[0098] The electronic device 500 typically includes a plurality of computer storage media. Such media may be any available media accessible to the electronic device 500, including but not limited to volatile and non-volatile media, removable and non-removable media. The memory 520 may be a volatile memory (e.g., registers, caches, random access memory (RAM)), a non-volatile memory (e.g., a read-only memory (ROM), an electrically erasable programmable read-only memory (EEPROM), flash memory), or some combination thereof. The storage device 530 may be a removable or non-removable medium, and may include a machine-readable medium, such as a flash drive, a disk, or any other medium, which may be capable of being used to store information and / or data and may be accessed within the electronic device 500.

[0099] The electronic device 500 may further include additional removable / non-removable, volatile / non-volatile storage media. Figure 5 As shown in , a disk drive for reading or writing from a removable, non-volatile disk (e.g., a "floppy disk") and an optical drive for reading or writing from a removable, non-volatile optical disk may be provided. In these cases, each drive may be connected to the bus (not shown) by one or more data media interfaces. The memory 520 may include a computer program product 525 having one or more program modules that are configured to perform various methods or actions of various embodiments of the present disclosure.

[0100] The communication unit 540 implements communication with other electronic devices through a communication medium. Additionally, the functions of the components of the electronic device 500 can be implemented in a single computing cluster or multiple computing machines that can communicate through a communication connection. Therefore, the electronic device 500 can operate in a networked environment using a logical connection with one or more other servers, a network personal computer (PC), or another network node.

[0101] The input device 550 may be one or more input devices, such as a mouse, a keyboard, a tracking ball, etc. The output device 560 may be one or more output devices, such as a display, a speaker, a printer, etc. The electronic device 500 may also communicate with one or more external devices (not shown) through the communication unit 540 as needed, such as a storage device, a display device, etc., communicate with one or more devices that allow a user to interact with the electronic device 500, or communicate with any device that allows the electronic device 500 to communicate with one or more other electronic devices (e.g., a network card, a modem, etc.). Such communication may be performed via an input / output (I / O) interface (not shown).

[0102] According to an exemplary implementation of the present disclosure, a computer-readable storage medium is provided, on which computer-executable instructions are stored, wherein the computer-executable instructions are executed by a processor to implement the method described above. According to an exemplary implementation of the present disclosure, a computer program product is also provided, which is tangibly stored on a non-transitory computer-readable medium and includes computer-executable instructions, and the computer-executable instructions are executed by a processor to implement the method described above.

[0103] Various aspects of the present disclosure are described herein with reference to the flowcharts and / or block diagrams of the methods, devices, equipment, and computer program products implemented according to the present disclosure. It should be understood that each box in the flowchart and / or block diagram and the combination of each box in the flowchart and / or block diagram can be implemented by computer-readable program instructions.

[0104] These computer-readable program instructions can be provided to a processing unit of a general-purpose computer, a special-purpose computer, or other programmable data processing device, thereby producing a machine, so that when these instructions are executed by the processing unit of the computer or other programmable data processing device, a device that implements the functions / actions specified in one or more boxes in the flowchart and / or block diagram is generated. These computer-readable program instructions can also be stored in a computer-readable storage medium, and these instructions cause the computer, programmable data processing device, and / or other equipment to work in a specific manner, so that the computer-readable medium storing the instructions includes a manufactured product, which includes instructions for implementing various aspects of the functions / actions specified in one or more boxes in the flowchart and / or block diagram.

[0105] Computer-readable program instructions can be loaded onto a computer, other programmable data processing apparatus, or other device so that a series of operational steps are performed on the computer, other programmable data processing apparatus, or other device to produce a computer-implemented process, so that the instructions executed on the computer, other programmable data processing apparatus, or other device implement the functions / actions specified in one or more boxes in the flowchart and / or block diagram.

[0106] The flow chart and block diagram in the accompanying drawings show the possible architecture, function and operation of the system, method and computer program product according to multiple implementations of the present disclosure. In this regard, each square box in the flow chart or block diagram can represent a part of a module, program segment or instruction, and a part of a module, program segment or instruction includes one or more executable instructions for realizing the logical function of the specification. In some implementations as replacements, the function marked in the square box can also occur in a sequence different from that marked in the accompanying drawings. For example, two continuous square boxes can actually be executed substantially in parallel, and they can sometimes be executed in reverse order, depending on the functions involved. It should also be noted that each square box in the block diagram and / or flow chart, and the combination of the square boxes in the block diagram and / or flow chart can be realized by a special hardware-based system that performs the function or action of the specification, or can be realized by a combination of special hardware and computer instructions.

[0107] The above descriptions of various implementations of the present disclosure are exemplary, non-exhaustive, and not limited to the disclosed implementations. Many modifications and variations will be apparent to those of ordinary skill in the art without departing from the scope and spirit of the described implementations. The selection of terms used herein is intended to best explain the principles of the implementations, practical applications, or improvements to the technology in the market, or to enable other persons of ordinary skill in the art to understand the various implementations disclosed herein.

Claims

1. A method for conversation interaction, comprising: In response to a first scene in a group of scenes being selected, performing interaction with the user in an interaction window between the user and the digital assistant based on the first scene, wherein at least one scene in the group of scenes is configured with configuration information, and the configuration information includes at least one of the following: scene setting information, plug-in information, wherein the scene setting information is used to describe information related to the corresponding scene, and the plug-in information indicates at least one plug-in for performing a task in the corresponding scene; and In response to receiving a first message from the user in the first scene, scene switching guide information is presented in the interactive window based on the first message and at least a portion of the configuration information of each scene, and the scene switching guide information indicates switching from the first scene to the second scene to perform the task indicated by the first message.

2. The method according to claim 1, further comprising: While presenting the scene switching guidance information, a first reply of the digital assistant to the first message is presented in an interaction area corresponding to the first scene within the interaction window, wherein the first reply is determined based on the first scene.

3. The method according to claim 1, wherein the scene switching guide information includes a switching entry of the second scene, and the method further comprises: In response to receiving a trigger operation on the switching entry, determining that the second scene is selected; as well as Perform interaction with the user in the interaction window based on the second scenario.

4. The method according to claim 3, further comprising: In response to determining that the second scene is selected, the first message and the second reply of the digital assistant to the first message are presented in an interaction area corresponding to the second scene within the interaction window, and the second reply is determined based on the second scene.

5. The method according to claim 3, wherein the first scene is selected for interaction of the first topic of the interaction window, and determining that the second scene is selected comprises: In response to receiving a trigger operation on the switching entry, a second topic is started in the interaction window, wherein the second scene is selected for interaction of the second topic.

6. The method according to claim 1, wherein at least a portion of the configuration information includes at least one of the following: The scene name, or Description of the scene.

7. The method according to claim 1, wherein the configuration information further comprises at least one of the following: The selected model is called to determine a response to the user in a corresponding scenario; Scene guidance information, which is presented to the user after the corresponding scene is selected; or For at least one recommended question of the digital assistant, after the corresponding scenario is selected, the at least one recommended question is presented to the user for selection.

8. The method according to claim 1, wherein presenting the scene switching guide information comprises: Based on the first message and at least a portion of the configuration information of the first scenario, determining that the task indicated by the first message does not match a type of task corresponding to the first scenario; as well as Based on at least a portion of the configuration information of the second scenario, determining that the task indicated by the first message does not match a type of task corresponding to the second scenario; as well as The scene switching guide information is presented based on the second scene.

9. The method according to claim 1, further comprising: constructing a prompt word input for a model based on the first message and at least a portion of the configuration information of each scenario in the set of scenarios; receiving a model output from the model by providing the cue word input to the model; as well as The scene switching guidance information is determined based on the model output.

10. The method according to claim 1, wherein the interactive window comprises one or more of the following: a conversation window, a floating window.

11. A device for conversation interaction, comprising: an interaction execution module, configured to, in response to a first scene in a group of scenes being selected, execute interaction with the user in an interaction window between the user and the digital assistant based on the first scene, wherein at least one scene in the group of scenes is configured with configuration information for executing a task of a corresponding type, the configuration information comprising at least one of the following: scene setting information, plug-in information, wherein the scene setting information is used to describe information related to the corresponding scene, and the plug-in information indicates at least one plug-in for executing a task in the corresponding scene; and An information presentation module is configured to, in response to receiving a first message from the user in the first scene, present scene switching guide information in the interactive window based on the first message and at least a portion of the configuration information of each scene in the set of scenes, wherein the scene switching guide information indicates switching from the first scene to a second scene to perform the task indicated by the first message.

12. An electronic device comprising: at least one processing unit; as well as At least one memory, the at least one memory being coupled to the at least one processing unit and storing instructions for execution by the at least one processing unit, the instructions causing the electronic device to perform the method according to any one of claims 1 to 10 when executed by the at least one processing unit.

13. A computer-readable storage medium having a computer program stored thereon, wherein the computer program can be executed by a processor to implement the method according to any one of claims 1 to 10.