Wake-up-free voice control method for terminal device, terminal device and server

By realizing wake-free voice control in the terminal device, users can continuously interact through voice commands during cooking, solving the cumbersome operation problems caused by frequent wake-ups, and improving user experience and operation efficiency.

CN115240674BActive Publication Date: 2025-06-06HISENSE VISUAL TECH CO LTD
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
CN202210873362.6
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2022-07-21
Publication Date
2025-06-06
Estimated Expiration
2042-07-21

AI Technical Summary

Technical Problem

During the cooking process, users need to frequently wake up the operation interface and cooking steps, resulting in cumbersome operations and reducing the user experience.

Method used

By obtaining the voice wake-up command input by the user, detecting the current interface scene and entering the wake-up-free mode, providing a list of wake-up-free word prompts. The user can perform voice control based on the list and perform corresponding operations.

Benefits of technology

This enables users to control voice without frequent wake-up during cooking, improves user experience, reduces operation steps, and saves user time.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN115240674B_ABST
    Figure CN115240674B_ABST
Patent Text Reader

Abstract

Some embodiments of the present application provide a wake-up-free voice control method for a terminal device, a terminal device, and a server, which obtains a voice wake-up instruction input by a user, responds to the voice wake-up instruction, detects the current interface scene according to the voice wake-up instruction, and controls the terminal device to enter a wake-up-free mode. Obtain a wake-up-free word prompt list according to the current interface scene, the wake-up-free word prompt list includes several wake-up-free words, and control the display to display the wake-up-free words. Obtain a voice control instruction input by the user based on the wake-up-free word prompt list. If the voice control instruction contains a wake-up-free word in the wake-up-free word prompt list, perform the operation corresponding to the wake-up-free word according to the current interface scene. The present application performs voice control through a wake-up-free voice response method, adds a wake-up-free word prompt to the display interface, and scrolls all supported wake-up-free word instructions, which is not only convenient for users to learn and use, but also can improve the user experience.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The present application relates to the field of smart electrical appliance technology, and in particular to a wake-up-free voice control method for a terminal device, a terminal device, and a server. Background Art

[0002] With the rapid development of the Internet era, people have an increasing demand for intelligence and diversification in various daily environments. As the kitchen is an indispensable living environment in daily life, cooking itself is a relatively tedious and time-consuming task. Therefore, users naturally want to get a simpler, more comfortable and smarter experience in the process.

[0003] At present, in order to facilitate users to view recipe information during cooking, the smart range hood screen has integrated a smart recipe broadcasting function. Recipe information can be viewed through voice control, freeing the user's hands. When viewing recipe information or performing other operations, users can use multiple voice wake-ups to broadcast recipes or view recipe related information, such as "Hello xx" to view the recipe information of fried dough sticks; "Hello xx" to jump to the next step; "Hello xx" to exit the recipe interface, etc.

[0004] However, the above method requires the user to say the wake-up word "Hello xx" to pull up the voice reception interface before using voice each time, and then say their own control commands, such as: next step, previous step, etc. After the voice command is recognized, the recipe application is notified to respond. This causes the user to wake up frequently during the cooking process, which is cumbersome and reduces the user experience. Summary of the invention

[0005] The present invention provides a wake-up-free voice control method for a terminal device, a terminal device and a server, so as to solve the problem that the user needs to frequently wake up the operation interface and cooking steps during the cooking process, resulting in cumbersome operations.

[0006] In a first aspect, some embodiments of the present application provide a terminal device, the terminal device comprising:

[0007] a display configured to display a user interface;

[0008] The controller is configured as:

[0009] Get the voice wake-up command input by the user;

[0010] In response to the voice wake-up instruction, detecting the current interface scene according to the voice wake-up instruction, and controlling the terminal device to enter a wake-up-free mode, wherein the wake-up-free mode is used to enable continuous voice interaction between the user and the terminal device;

[0011] Acquire a wake-up-free word prompt list according to the current interface scene, the wake-up-free word prompt list including a plurality of wake-up-free words, and control the display to display the wake-up-free words;

[0012] Obtain a voice control instruction input by the user based on the wake-up-free word prompt list. If the voice control instruction contains a wake-up-free word in the wake-up-free word prompt list, perform an operation corresponding to the wake-up-free word according to the current interface scene.

[0013] In a second aspect, some embodiments of the present application provide a server, the server comprising:

[0014] A communication device configured to establish a communication connection with a user and a terminal device;

[0015] A control device, configured to receive a voice wake-up instruction input by a user, determine an interface scene according to the voice wake-up instruction, and determine a wake-up-free word prompt list according to the interface scene, wherein the wake-up-free word prompt list includes a plurality of wake-up-free words;

[0016] Send the wake-up-free word prompt list to the terminal device so that the terminal device displays the wake-up-free word; and obtain the voice control instruction input by the user based on the wake-up-free word prompt list. If the voice control instruction contains the wake-up-free word in the wake-up-free word prompt list, perform the operation corresponding to the wake-up-free word according to the current interface scene.

[0017] In a third aspect, some embodiments of the present application provide a wake-up-free voice control method for a terminal device, which is applied to a terminal device, wherein the terminal device includes a display and a controller, and the method includes:

[0018] Get the voice wake-up command input by the user;

[0019] In response to the voice wake-up instruction, detecting the current interface scene according to the voice wake-up instruction, and controlling the terminal device to enter a wake-up-free mode, wherein the wake-up-free mode is used to enable continuous voice interaction between the user and the terminal device;

[0020] Acquire a free wake-up word prompt list according to the current interface scene, where the free wake-up word prompt list includes a plurality of free wake-up words;

[0021] Obtain a voice control instruction input by the user based on the wake-up-free word prompt list. If the voice control instruction contains a wake-up-free word in the wake-up-free word prompt list, perform an operation corresponding to the wake-up-free word according to the current interface scene.

[0022] It can be seen from the above technical solutions that some embodiments of the present application provide a wake-up-free voice control method, terminal device and server for a terminal device, which obtains the voice wake-up command input by the user, responds to the voice wake-up command, detects the current interface scene according to the voice wake-up command, and controls the terminal device to enter the wake-up-free mode, and the wake-up-free mode is used to enable continuous voice interaction between the user and the terminal device. According to the current interface scene, a wake-up-free word prompt list is obtained, and the wake-up-free word prompt list includes several wake-up-free words, and the display is controlled to display the wake-up-free words. The voice control command input by the user based on the wake-up-free word prompt list is obtained. If the voice control command contains a wake-up-free word in the wake-up-free word prompt list, the operation corresponding to the wake-up-free word is performed according to the current interface scene. The present application enables the user to perform voice control through the wake-up-free voice response method during the cooking process, and adds a wake-up-free word prompt to the broadcast interface, and scrolls to play all supported wake-up-free word instructions, which can not only improve the user experience but also facilitate user learning and use. BRIEF DESCRIPTION OF THE DRAWINGS

[0023] In order to more clearly illustrate the technical solution of the present application, the drawings required for use in the embodiments are briefly introduced below. Obviously, for ordinary technicians in this field, other drawings can be obtained based on these drawings without any creative work.

[0024] Figure 1 A schematic diagram of the system architecture of a wake-up-free voice control method, a terminal device 200 and a server 10 according to some embodiments is shown;

[0025] Figure 2 shows a hardware configuration block diagram of a terminal device 200 according to some embodiments;

[0026] Figure 3 shows a software configuration diagram in a terminal device 200 according to some embodiments;

[0027] Figure 4 A schematic diagram of a voice interaction network architecture according to some embodiments is shown;

[0028] Figure 5 A schematic diagram of interaction of wake-up-free voice control of a terminal device according to some embodiments is shown;

[0029] Figure 6 A schematic diagram of a wake-up word-free adaptive display process of a terminal device according to some embodiments is shown;

[0030] Figure 7 A schematic diagram of a recipe default scene interface of a terminal device according to some embodiments is shown;

[0031] Figure 8 A schematic diagram of a recipe details scene interface of a terminal device according to some embodiments is shown;

[0032] Fig. 9 A schematic diagram of a recipe broadcasting scene interface of a terminal device according to some embodiments is shown;

[0033] Fig.10 A schematic diagram of a process of switching scenes in an interface scene of a terminal device according to some embodiments is shown;

[0034] Fig.11 A schematic diagram of a recipe search scene interface of a terminal device according to some embodiments is shown;

[0035] Fig.12 A schematic diagram of a recipe classification scenario interface of a terminal device according to some embodiments is shown. DETAILED DESCRIPTION

[0036] In order to make the purpose, implementation mode and advantages of the present application clearer, the exemplary implementation mode of the present application will be clearly and completely described below in conjunction with the drawings in the exemplary embodiments of the present application. Obviously, the described exemplary embodiments are only part of the embodiments of the present application, rather than all the embodiments.

[0037] Based on the exemplary embodiments described in this application, all other embodiments obtained by ordinary technicians in this field without making creative work are within the scope of protection of the claims attached to this application. In addition, although the disclosure in this application is introduced according to one or several exemplary examples, it should be understood that each aspect of these disclosures can also constitute a complete implementation method separately. It should be noted that the brief description of the terms in this application is only for the convenience of understanding the implementation methods described below, and is not intended to limit the implementation methods of this application. Unless otherwise specified, these terms should be understood according to their common and usual meanings.

[0038] The terms "first", "second", "third", etc. in the specification and claims of this application and the above drawings are used to distinguish similar or similar objects or entities, and do not necessarily mean to limit a specific order or sequence, unless otherwise noted. It should be understood that the terms used in this way can be interchangeable under appropriate circumstances.

[0039] The terms "comprises," "comprising," and "having," and any variations thereof, are intended to cover but not exclude inclusion, for example, a product or device comprising a list of components is not necessarily limited to all the components expressly listed but may include other components not expressly listed or inherent to such product or device.

[0040] The term "module" refers to any known or later developed hardware, software, firmware, artificial intelligence, fuzzy logic, or combination of hardware and / or software code that is capable of performing the functions associated with that element.

[0041] Figure 1 The following is an exemplary system architecture to which the wake-up-free voice control method, server, and terminal device of the present application can be applied. Figure 1 As shown, 10 is a server and 200 is a terminal device, exemplarily including (smart TV 200a, mobile device 200b, smart speaker 200c). In this application, the server 10 and the terminal device 200 communicate data through various communication methods. The terminal device 200 may be allowed to communicate and connect through a local area network (LAN), a wireless local area network (WLAN) and other networks. The server 10 can provide various content and interactions to the terminal device 200. Exemplarily, the terminal device 200 and the server 10 can send and receive information, and receive software program updates.

[0042] The server 10 may be a server that provides various services, such as a background server that provides support for audio data collected by the terminal device 200. The background server may analyze and process the received audio data, and feed back the processing results (such as endpoint information) to the terminal device 200. The server 10 may be a server cluster or multiple server clusters, and may include one or more types of servers.

[0043] The terminal device 200 can be hardware or software. When the terminal device 200 is hardware, it can be various electronic devices with sound collection functions, including but not limited to smart smoke machines, smart phones, televisions, tablet computers, e-book readers, smart watches, players, computers, AI devices, robots, smart vehicles, etc. When the terminal device 200 is software, it can be installed in the electronic devices listed above. It can be implemented as multiple software or software modules (for example, to provide sound collection services), or it can be implemented as a single software or software module. No specific limitation is made here.

[0044] It should be noted that the method for wake-up-free voice control provided in some embodiments of the present application can be executed by the server 10, or by the terminal device 200, or by both the server 10 and the terminal device 200, and the present application does not limit this.

[0045] Figure 2 FIG. 2 shows a hardware configuration block diagram of a terminal device 200 according to an exemplary embodiment. Figure 2The terminal device 200 shown includes at least one of a communicator 220, a detector 230, an external device interface 240, a controller 250, a display 260, an audio output interface 270, a memory, a power supply, and a user interface 280. The controller 250 includes a central processing unit, an audio processor, a RAM, a ROM, and a first interface to an nth interface for input / output.

[0046] The display 260 includes a display screen component for presenting images, a driving component for driving image display, a component for receiving image signals output from the controller, and a component for displaying video content, image content, menu control interface, and user control UI interface. The display 260 can be a liquid crystal display or an OLED display.

[0047] The communicator 220 is a component for communicating with an external device or server according to various communication protocol types. For example, the communicator 220 may include at least one of a Wifi module, a Bluetooth module, a wired Ethernet module, and other network communication protocol chips or a near field communication protocol chip, and an infrared receiver. The terminal device 200 can establish the transmission and reception of control signals and data signals with the server 10 through the communicator 220.

[0048] The user interface can be used to receive external control signals.

[0049] The detector 230 is used to collect signals from the external environment or the external interaction. For example, the detector 230 includes a light receiver, a sensor for collecting the intensity of ambient light; or, the detector 230 includes an image collector, such as a camera, which can be used to collect external environment scenes, user attributes or user interaction gestures; or, the detector 230 includes a sound collector, such as a microphone, etc., for receiving external sounds.

[0050] The sound collector can be a microphone, also called a "microphone" or "microphone", which can be used to receive the user's voice and convert the sound signal into an electrical signal. The terminal device 200 can be provided with at least one microphone. In other embodiments, the terminal device 200 can be provided with two microphones, which can not only collect sound signals but also realize the noise reduction function. In other embodiments, the terminal device 200 can also be provided with three, four or more microphones to realize the collection of sound signals, noise reduction, identification of the sound source, and realization of directional recording functions, etc.

[0051] In addition, the microphone may be built into the terminal device 200, or the microphone may be connected to the terminal device 200 by wire or wireless means. Of course, some embodiments of the present application do not limit the position of the microphone on the terminal device 200. Alternatively, the terminal device 200 may not include a microphone, that is, the microphone is not provided in the terminal device 200. The terminal device 200 may be connected to an external microphone (also referred to as a microphone) through an interface (such as a USB interface 130). The external microphone may be fixed to the terminal device 200 by an external fixing (such as a camera holder with a clip).

[0052] The controller 250 controls the operation of the terminal device 200 and responds to the user's operation through various software control programs stored in the memory. The controller 250 controls the overall operation of the terminal device 200.

[0053] Exemplarily, the controller 250 includes at least one of a central processing unit (CPU), an audio processor, a graphics processing unit (GPU), RAM Random Access Memory (RAM), ROM (Read-Only Memory, ROM), a first interface to an nth interface for input / output, a communication bus (Bus), etc.

[0054] In some examples, the operating system of the terminal device 200 is an Android system, for example, Figure 3 As shown, the terminal device 200 can be logically divided into an application layer (“application layer” for short) 21 , a kernel layer 22 and a hardware layer 23 .

[0055] Among them, Figure 3 As shown, the hardware layer 23 may include Figure 2 The controller 250, the communicator 220, the detector 230, etc. are shown. The application layer 21 includes one or more applications. The application can be a system application or a third-party application. For example, the application layer 21 includes a speech recognition application, which can provide a speech interaction interface and service for realizing the connection between the terminal device 200 and the server 10.

[0056] The kernel layer 22 serves as software middleware between the hardware layer 23 and the application layer 21 and is used to manage and control hardware and software resources.

[0057] In some examples, the kernel layer 22 includes a detector driver, which is used to send the voice data collected by the detector 230 to the voice recognition application. Exemplarily, when the voice recognition application in the terminal device 200 is started and the terminal device 200 establishes a communication connection with the server 10, the detector driver is used to send the voice data of the user input collected by the detector 230 to the voice recognition application. Afterwards, the voice recognition application sends the query information containing the voice data to the intent recognition module 102 in the server 10. The intent recognition module 102 is used to input the voice data sent by the terminal device 200 into the intent recognition model.

[0058] To clearly illustrate the embodiments of the present application, Figure 4 A voice interaction network architecture provided in some embodiments of the present application is described.

[0059] See also Figure 4 , Figure 4 A schematic diagram of a voice interaction network architecture provided for some embodiments of the present application. Figure 4 In the example, the terminal device 200 is used to receive input information and output the processing result of the information. The speech recognition module is deployed with a speech recognition service for recognizing audio as text; the semantic understanding module is deployed with a semantic understanding service for semantically parsing text; the business management module is deployed with a business instruction management service for providing business instructions; the language generation module is deployed with a language generation service (Natural Language Generation, NLG) for converting the instructions for the terminal device to execute into text language; the speech synthesis module is deployed with a speech synthesis (Text To Speech, TTS) service for processing the text language corresponding to the instruction and sending it to the speaker for broadcast. In one embodiment, Figure 4 In the illustrated architecture, there may be multiple physical service devices deployed with different business services, or one or more functional services may be integrated into one or more physical service devices.

[0060] In some embodiments, the following Figure 4 The process of processing the information input into the terminal device by the architecture shown is described by way of example, taking the information input into the terminal device 200 as a query statement input by voice as an example:

[0061] [Speech Recognition]

[0062] After receiving the query statement input by voice, the terminal device 200 may perform noise reduction processing and feature extraction on the audio of the query statement. The noise reduction processing here may include steps such as removing echoes and environmental noise.

[0063] [Semantic understanding]

[0064] Using acoustic models and language models, natural language understanding is performed on the identified candidate texts and associated context information, and the texts are parsed into structured, machine-readable information, business domain, intent, word slots and other information to express semantics, etc. The executable intent is obtained to determine the intent confidence score, and the semantic understanding module selects one or more candidate executable intents based on the determined intent confidence score.

[0065] [Business Management]

[0066] Based on the semantic analysis results of the query text, the semantic understanding module sends query instructions to the corresponding business management module to obtain the query results given by the business service, execute the actions required to "complete" the user's final request, and feedback the device execution instructions corresponding to the query results.

[0067] It should be noted that Figure 4 The architecture shown is only an example and does not limit the protection scope of the present application. In the embodiments of the present application, other architectures may also be used to implement similar functions. For example, all or part of the above process may be completed by a smart terminal, which will not be described in detail here.

[0068] The above embodiments introduce the hardware / software architecture of the terminal device 200 and the server 10 as well as the functional implementation. Among them, voice wake-up is a branch of the voice recognition task, which requires detecting a limited number of pre-defined activation words or keywords from a string of voice streams without recognizing all voices. This type of technology can be applied to various fields, such as mobile phones, smart range hoods, robots, smart homes, vehicle-mounted devices, and wearable terminal devices. In some scenarios, the user can perform relevant operations on the terminal device 200 through voice control, but the user needs to say the wake-up word before each relevant operation. For example, the voice wake-up word of the smart range hood is "Hello, Harry". During the cooking process, every time the user views the recipe information or other cooking information content through voice control, the user needs to say the "Hello, Harry" wake-up word to pull up the voice radio interface, and then say his own voice control command, that is, "Hello, Harry", to view the recipe information; "Hello, Harry" next step, previous step, etc. By recognizing the voice control command and notifying the recipe application to respond, users need to frequently perform related operations when using the smart range hood, and need to wake up frequently during the operation, which makes the operation cumbersome and reduces the user experience.

[0069] In order to improve the user's operating efficiency and reduce unnecessary voice wake-up, the user does not need to wake up the terminal device again to perform related operations after waking up once, and realizes the wake-up-free voice control of the terminal device 200. Figure 5As shown, an interactive schematic diagram of wake-up-free voice control of a terminal device provided by an embodiment of the present invention is shown.

[0070] Step S501: Acquire a voice wake-up instruction input by a user.

[0071] In some embodiments, when the user is in a cooking scenario, the terminal device 200 obtains a voice wake-up command input by the user. The voice wake-up command means that the user wakes up the terminal device 200 by voice and performs a corresponding operation. The voice wake-up command also includes the relevant operation that the user wants to perform or the interface information that the user wants to view. For example, when the user is cooking, he wants to view the details of the recipe. At this time, the user can say, "Hello, Harry" to pull up the voice reception interface, and then the user can give voice commands according to the information prompts on the display interface, and continue to say, "Read the recipe" to enter the interface such as Figure 8 The recipe details interface shown.

[0072] Step S502: In response to the voice wake-up instruction, the current interface scene is detected according to the voice wake-up instruction, and the terminal device is controlled to enter the wake-up-free mode.

[0073] In some embodiments, the terminal device 200 responds to the voice wake-up command input by the user, and detects the current interface scene according to the voice wake-up command said by the user. When the user is in a certain interface scene, the terminal device 200 will be automatically controlled to enter the wake-up-free mode. The wake-up-free mode is used to enable continuous voice interaction between the user and the terminal device 200. The wake-up-free mode will remain in the receiving state until the user exits the terminal device 200. For example, during the cooking process, the user may say: "Hello, Harry, I want to make donuts." At this time, the display shows the details of the methods of making a variety of donuts. The user can choose one of the methods, such as "Super delicious donuts without oven", and enter the following method according to the voice command. Figure 8 The recipe details interface shown in the figure automatically controls the terminal device 200 to enter the wake-up-free mode after the interface is started, so that the user no longer needs to use the wake-up word "Hello, Harry" to wake up the terminal device 200 to perform related operations, and can directly say "voice recipe" to perform related cooking operations related to the voice recipe. It can reduce the user's frequent wake-up operations and improve the user experience.

[0074] Step S5021: Register the scene attributes and the wake-up-free words of the interface scene into the voice service. The scene attributes in the voice service have a one-to-one mapping relationship with the wake-up-free words.

[0075] In some embodiments, the terminal device 200 defines different scene attributes for each different interface scene. For example, as shown in Table 1, when entering the recipe details interface scene, the scene attribute is first defined as "recipe_detail", and the corresponding scene attribute means the recipe details interface scene. The wake-up-free words corresponding to the scene attribute are registered to the voice service. As shown in Table 1, the wake-up-free words corresponding to the scene attribute "recipe_detail" include "voice recipe", "recipe announcement", "start announcement" and "announce recipe", etc., that is, the voice service can first obtain the scene attributes through the pre-registered scene attributes and wake-up-free words, and obtain the supported wake-up-free words in the scene according to the scene attributes. If the voice command issued by the user is among the supported wake-up-free words, the recipe application is notified to execute the corresponding instruction cooking processing operation.

[0076] In some embodiments, if the user issues a wake-up word supported by the interface scene in the recipe details interface, such as "recipe announcement", the terminal device 200 starts the recipe announcement interface scene after receiving the user's voice command. Before starting the recipe announcement interface scene, the scene attribute needs to be set to "recipe_default". As shown in Table 1, the "recipe_default" scene attribute means "default scene", that is, when starting the next interface scene, the current scene needs to be set to Figure 7 The recipe default scene shown is also the home page scene of the terminal device 200, which can effectively avoid problems such as interface acquisition errors caused by the failure of the terminal device 200 to start the scene due to other reasons.

[0077] Table 1: Definition of wake-up-free interface scene attributes

[0078]

[0079] Step S5022: If the scene attributes change, the wake-up word is replaced in real time according to the scene attributes.

[0080] In some embodiments, when the user enters the recipe broadcast interface scene from the recipe details interface scene through voice commands, the corresponding scene attributes also change. Since the scene attributes and the wake-up-free words are in a one-to-one mapping relationship, when the interface scene changes, the corresponding scene attributes and the wake-up-free words will be changed in real time, and replaced with the scene attributes and the wake-up-free words corresponding to the interface scene. For example, the user enters the recipe broadcast interface scene through the voice command of the wake-up-free word "recipe broadcast". As shown in Table 1, the scene attribute of the recipe broadcast interface scene is "recipe_broadcast", and the wake-up-free words corresponding to the scene attribute "recipe_broadcast" are "previous page", "next page", "exit broadcast" and "continue broadcast" and other wake-up-free words. The user no longer needs a wake-up word to wake up the corresponding cooking processing operation. Just say the wake-up-free word "previous page" or "next page" to implement the relevant operations corresponding to the wake-up-free word as shown in Table 2. If the user says the wake-up-free word "continue processing", the terminal device 200 will execute the corresponding response operation result "automatically slide to the next page", which is convenient for the user's use process.

[0081] Table 2: Response operation results corresponding to the wake-up word

[0082]

[0083] Step S503: Obtain a list of free wake-up words prompts according to the current interface scene, the list of free wake-up words prompts includes several free wake-up words, and control the display to display the free wake-up words.

[0084] In some embodiments, a corresponding wake-up-free word prompt list is obtained according to the current interface scene, and the wake-up-free word prompt list includes several wake-up-free words corresponding to the relevant interface scene. Therefore, according to the interface scene currently displayed by the user's terminal device 200, the wake-up-free word prompt list is obtained, and the wake-up-free word prompt list determines the wake-up-free word according to the scene attributes of the interface scene and controls the display to display the wake-up-free word used to prompt the user to perform related operations. For example, when the user is cooking, the interface scene displayed by the terminal device 200 is a recipe broadcast interface scene. The terminal device 200 obtains the corresponding wake-up-free word prompt list according to the recipe broadcast interface scene. As shown in Table 1, there are several wake-up-free words in the wake-up-free word prompt list. The terminal device 200 displays these wake-up-free words on the display in chronological order according to the wake-up-free words corresponding to the recipe broadcast interface scene, so that the user can perform voice control according to the prompted wake-up-free words. The user performs voice control according to the prompted wake-up-free words, which can reduce the probability of false triggering caused by the user's use of voice control.

[0085] When the wake-up-free word is displayed on the display, the terminal device 200 controls the display to display the wake-up-free word in the wake-up-free word prompt list one by one at a preset interval, and the wake-up-free word will be displayed on the terminal device 200 in a fixed format. For example, when the user enters a specific interface scene, in this interface scene, only the wake-up-free word voice command corresponding to the interface scene can be responded, such as Fig. 9 As shown, the wake-up-free word prompt in this application is composed of two sentences: No need to wake up, say "previous page" to me. It can also be composed of one sentence, such as only displaying the wake-up-free word "say me, previous page", and all supported wake-up-free words will be displayed in sentences composed of this format at the bottom of the display. The specific position of the display is not limited in this application. When the wake-up-free words are displayed in a carousel, each wake-up-free word can be displayed with an interval of 1.5 seconds, and the interval time can also be tailored according to the user's usage habits. If the user feels that the display is too fast, the duration can be appropriately increased. This application does not make specific limitations here. The present application can realize the function of adaptively displaying the wake-up-free word prompt according to the recipe interface scene index, effectively avoiding providing wrong voice control instructions to users and improving user experience.

[0086] like Figure 6 As shown, it is a schematic diagram of a wake-up word-free adaptive display process of a terminal device in some embodiments of the present application, and the specific steps include:

[0087] Step S5031: add index page numbers after some wake-up-free words to obtain an index list, and the index page numbers are used to hide some wake-up-free words under the index page numbers.

[0088] In some embodiments, not all wake-up-free words will be displayed on the display. You can add index pages after some wake-up-free words. The added index pages can be separated by the "#" character, and then the pages that cannot be displayed are added. By adding index pages, the wake-up-free words can be hidden under the index pages. For example, "No need to wake up, say to me: Previous page #1", it means that the wake-up-free word "Previous page" is not displayed on page 1. Add these wake-up-free words that do not need to be displayed to a list to obtain an index list, so that the terminal device 200 only needs to traverse the index list to know whether the wake-up-free words need to be hidden.

[0089] Step S5032: Obtain the wake-up-free word to be displayed and read the index list.

[0090] In some embodiments, the display displays the wake-up-free word at a preset interval, that is, a wake-up-free word is displayed every 1.5 seconds. After 1.5 seconds, the next wake-up-free word to be displayed is obtained, and the terminal device 200 reads the index list and judges whether the wake-up-free word to be displayed needs to be displayed. In some interface scenarios, there are no wake-up-free words that do not need to be displayed in the index list, and in some interface scenarios, there may be wake-up-free words that do not need to be displayed in the index list. For example, during use, the user is in the recipe details interface scene, and the corresponding wake-up-free words, such as "voice recipe", "recipe broadcast" and "start broadcast", are all wake-up-free words that need to be displayed, so these wake-up-free words do not need to be added to the index list. If the user is in the recipe broadcast interface scene, the corresponding wake-up-free words, such as "previous page", "next page" and "exit broadcast", if the current recipe broadcast interface is displaying the first page of content, then the wake-up-free word cannot display "previous page" at this time, so the wake-up-free word "previous page" needs to be added to the index list.

[0091] Step S5033: If the index list is empty, the wake-up word-free function is displayed.

[0092] In some embodiments, when obtaining the next wake-up-free word to be displayed, the terminal device 200 will traverse the index list to determine whether there is a wake-up-free word with an index page number in the index list. If there is no wake-up-free word with an index page number in the index list, the terminal device 200 will display the wake-up-free word.

[0093] Step S5034: If the index list contains a wake-up-free word with an index page number, and the current page number contains the index page number, the wake-up-free word is hidden.

[0094] In some embodiments, when the terminal device 200 traverses the index list, the index list is not empty and contains a wake-up-free word with an index page number, such as the wake-up-free word "previous page #1" with an index page number, where the index page number is 1 and the current page number is also 1, then the wake-up-free word "previous page" will be hidden and will not be displayed. There is another situation. If the index list contains the wake-up-free word with the index page number, and the current page number does not contain the index page number, such as the wake-up-free word "previous page #1" with the index page number, and the current recipe page number is 2, then the wake-up-free word will not be hidden, and the wake-up-free word "previous page" will be displayed. If the current page is already the last page, the terminal device 200 will prompt the user: "This is the last page of the recipe" to realize the control of the wake-up-free voice command. The present application can realize that the user does not need to wake up multiple times during use, and directly displays the recipe through the wake-up-free voice control command, and can flexibly view the recipe step information, reduce the user's operation steps, save the user's time, and improve the user's experience.

[0095] Step S504: Obtain the voice control instruction input by the user based on the free wake-up word prompt list. If the voice control instruction contains a free wake-up word in the free wake-up word prompt list, perform the operation corresponding to the free wake-up word according to the current interface scene.

[0096] In some embodiments, during the cooking process, the user can perform voice control through the wake-up word prompt given in the wake-up word prompt list. When the user speaks the voice control command, it is necessary to determine whether the voice control command contains the wake-up word. For example, if the user says the "read the recipe" voice control command, it happens to be a wake-up word in the wake-up word prompt list. Then, according to the current interface scene, the operation corresponding to the wake-up word "read the recipe" is performed, that is, the ingredients and materials required for the dish that the user wants to complete are broadcast. If the user says the voice control command "recipe content" during the cooking process, and the "recipe content" does not contain the wake-up word in the wake-up word prompt list, that is, there is no wake-up word "recipe content", the terminal device 200 will not respond to the "recipe content" voice control command to perform any related cooking operations. The terminal device 200 can inform the user that this is an incorrect command, or prompt the user to try other voice commands, or directly display the wake-up word-free voice command that can execute the operation command to the user.

[0097] Step S5041: parse the voice control instruction and determine the scene attributes of the current interface scene.

[0098] In some embodiments, the voice control command input by the user is obtained, and the voice control command spoken by the user is parsed to determine the scene attribute of the current interface scene. For example, the user speaks the voice control command "read the recipe", and the voice control command "read the recipe" is semantically parsed to obtain the scene attribute of "read the recipe" as "recipe_detail", that is, the current interface scene is the recipe details interface scene. If the parsing of the voice control command spoken by the user fails, the terminal device 200 will voice prompt the user to try other voice control commands.

[0099] Step S5042: Obtain the corresponding wake-up word prompt list according to the scene attributes.

[0100] In some embodiments, the scene attribute has been obtained in step S1041, and then the corresponding wake-up-free word prompt list is obtained according to the scene attribute. For example, the scene attribute obtained in step S1041 is "recipe_detail", and the wake-up-free word in the wake-up-free word prompt list is obtained according to this scene attribute "recipe_detail", such as "voice recipe", "start broadcast", etc.

[0101] like Fig.10As shown, it is a schematic diagram of a process of switching scenes in an interface scene of a terminal device provided in some embodiments of the present application, which may specifically include the following steps:

[0102] Step 110: Obtain a switching instruction for switching interface scenes input by the user through a wake-up-free word in the wake-up-free word prompt list.

[0103] In some embodiments, the microphone of the terminal device 200 can be used to collect the voice control instructions issued by the user. The voice control instructions issued by the user are based on the wake-up-free words in the wake-up-free word prompt list, and are voice switching instructions for switching interface scenes. For example, the interface scene being displayed on the display of the user terminal device 200 is the recipe details interface scene. At this time, the user says the voice switching instruction "Say to me, switch to the recipe broadcast interface scene". Since the voice switching instruction is based on the wake-up-free words in the wake-up-free word prompt list, the terminal device 200 is certain to respond to the voice switching instruction, that is, to switch the current recipe details interface scene to the recipe broadcast interface scene. It is convenient for users to see how many grams of a certain ingredient required on the recipe details interface when they are in the recipe broadcast interface scene, or to check whether some ingredients can be replaced, so that users can use this function flexibly during the cooking process.

[0104] Step 120: In response to the switching instruction, extract the wake-up-free word corresponding to the interface scene before switching.

[0105] In some embodiments, the terminal device 200 responds to the user's switching instruction and extracts the wake-up-free word corresponding to the interface scene before the switch. For example, the current interface scene where the user is located is the recipe broadcast interface scene. The user issues a switching instruction to switch the recipe broadcast interface scene to the recipe details interface scene. It is necessary to extract the wake-up-free word corresponding to the interface scene before the switch, that is, to extract the corresponding wake-up-free word in the recipe broadcast interface scene, that is, the wake-up-free word "previous page", "exit broadcast", "query ingredients", etc.

[0106] Step 130: After the terminal device 200 enters the switched interface scene, the extracted wake-up-free word is marked as an invalid wake-up-free word, where an invalid wake-up-free word is a voice control instruction that cannot execute a corresponding operation.

[0107] In some embodiments, for example, after the terminal device 200 enters the switched recipe details interface scene, the wake-up-free words "previous page", "exit broadcast", "query ingredients", etc. extracted in step 120 are marked as invalid wake-up-free words, and invalid wake-up-free words cannot perform corresponding operations in the switched recipe details interface scene, such as invalid wake-up-free words "previous page", "exit broadcast", "query ingredients", etc., cannot perform corresponding operations in the switched recipe details interface scene.

[0108] Step 140: Control the display not to display invalid wake-up-free words in the wake-up-free word prompt list.

[0109] In some embodiments, the display will not display invalid wake-up-free words in the wake-up-free word prompt list for the interface scene before switching, such as invalid wake-up-free words "previous page", "exit broadcast", "query materials", etc. This application can define the wake-up-free words corresponding to the interface scene before switching as invalid wake-up-free words when the user switches the interface scene, which can greatly reduce the probability of the user saying the wrong voice control command, and the wake-up-free words supported in each interface scene can be increased or decreased at any time according to the business and user experience, which improves the flexibility and ease of use of the wake-up-free mode.

[0110] In some embodiments, the interface scene in the present application includes multiple recipe interfaces, such as Fig.11 and 12 The recipe search scene interface and recipe classification scene interface schematic diagrams, as well as the recipe details interface scene, the recipe broadcast interface scene, etc., are shown. The present application can also record the number of times these recipe interfaces are awakened by the user and sort them according to the number of times these recipe interfaces are awakened. For example, the recipe broadcast interface scene is awakened by the user a high number of times. Each time the interface scene is started, the terminal device 200 can give priority to prompting the user whether it is necessary to enter the recipe broadcast interface scene, so that the user can directly enter the recipe broadcast interface scene to perform related operations, saving more time for the user. For example, in some cases, the user is experienced in cooking and does not need to check the recipe details. The cooking operation can be performed directly. The user only needs to know the cooking steps. At this time, the user can be prompted with the corresponding scene by sorting the interface scenes. Of course, there will be different interface scene sortings for different users' needs, all of which need to be based on the number of times the user usually wakes up the interface scene.

[0111] Some embodiments of the present application further provide a server, the server comprising:

[0112] The communication device 101 is configured to establish a communication connection with a user and a terminal device.

[0113] The control device 102 is configured to receive a voice wake-up instruction input by a user, determine an interface scene according to the voice wake-up instruction, and determine a wake-up-free word prompt list according to the interface scene, wherein the wake-up-free word prompt list includes several wake-up-free words.

[0114] Send the wake-up-free word prompt list to the terminal device so that the terminal device displays the wake-up-free word; and obtain the voice control instruction input by the user based on the wake-up-free word prompt list. If the voice control instruction contains the wake-up-free word in the wake-up-free word prompt list, perform the operation corresponding to the wake-up-free word according to the current interface scene.

[0115] Some embodiments of the present application also provide a wake-up-free voice control method for a terminal device, the method is applied to a terminal device 200, the terminal device 200 includes a display and a controller, the method includes:

[0116] Get the voice wake-up command input by the user.

[0117] In response to the voice wake-up instruction, the current interface scene is detected according to the voice wake-up instruction, and the terminal device is controlled to enter a wake-up-free mode, where the wake-up-free mode is used to enable continuous voice interaction between the user and the terminal device.

[0118] A wake-up-free word prompt list is obtained according to the current interface scene, and the wake-up-free word prompt list includes several wake-up-free words.

[0119] Obtain a voice control instruction input by the user based on the wake-up-free word prompt list. If the voice control instruction contains a wake-up-free word in the wake-up-free word prompt list, perform an operation corresponding to the wake-up-free word according to the current interface scene.

[0120] In summary, some embodiments of the present application provide a wake-up-free voice control method for a terminal device, a terminal device 200 and a server 10. The present application obtains a voice wake-up instruction input by a user, responds to the voice wake-up instruction, detects the current interface scene according to the voice wake-up instruction, and controls the terminal device to enter the wake-up-free mode. The wake-up-free mode is used to enable continuous voice interaction between the user and the terminal device. Obtain a wake-up-free word prompt list according to the current interface scene, the wake-up-free word prompt list includes several wake-up-free words, and control the display to display the wake-up-free word. Obtain a voice control instruction input by the user based on the wake-up-free word prompt list. If the voice control instruction contains a wake-up-free word in the wake-up-free word prompt list, perform a cooking processing operation corresponding to the wake-up-free word according to the current interface scene. The technical solution of the present application responds to different wake-up-free words in different interface scenes. Since the wake-up-free word in the present application is based on the interface scene and is a word strongly related to the user's current interface content, the scene wake-up-free word will only be responded to in the current scene. When the user switches scenes, the previous scene words will become invalid, which can greatly reduce the probability of false triggering. In addition, the wake-up-free scenarios and the wake-up-free words supported in each scenario can also be added or adjusted at any time according to the business of the terminal device 200 and the user experience, thereby improving the flexibility and ease of use of the wake-up-free mode.

[0121] The same and similar parts between the various embodiments in this specification can be referenced to each other and will not be described again here.

[0122] Finally, it should be noted that the above embodiments are only used to illustrate the technical solutions of the present application, rather than to limit it. Although the present application has been described in detail with reference to the aforementioned embodiments, those skilled in the art should understand that they can still modify the technical solutions described in the aforementioned embodiments, or replace some or all of the technical features therein with equivalents. However, these modifications or replacements do not cause the essence of the corresponding technical solutions to deviate from the scope of the technical solutions of the embodiments of the present application.

[0123] For ease of explanation, the above description has been made in conjunction with specific embodiments. However, the above exemplary discussion is not intended to be exhaustive or limit the embodiments to the specific forms disclosed above. Based on the above teachings, various modifications and variations can be obtained. The selection and description of the above embodiments are intended to better explain the principles and practical applications, so that those skilled in the art can better use the embodiments and various different variations of the embodiments suitable for specific use considerations.

Claims

1. A terminal device, It is characterized in that include: a display configured to display a user interface; The controller is configured as: Get the voice wake-up command input by the user; In response to the voice wake-up instruction, detecting the current interface scene according to the voice wake-up instruction, and controlling the terminal device to enter a wake-up-free mode, wherein the wake-up-free mode is used to enable continuous voice interaction between the user and the terminal device; Acquire a wake-up-free word prompt list according to the current interface scene, the wake-up-free word prompt list including a plurality of wake-up-free words, and control the display to display the wake-up-free words; Obtaining a voice control instruction input by a user based on the wake-up-free word prompt list, and if the voice control instruction contains a wake-up-free word in the wake-up-free word prompt list, switching the current interface scene to an interface scene after executing an operation corresponding to the wake-up-free word; When the scene attributes of the current interface scene are different from the scene attributes of the interface scene after executing the operation corresponding to the wake-up-free word, the wake-up-free word in the wake-up-free word prompt list is replaced according to the wake-up-free word corresponding to the changed scene attributes.

2. The terminal device according to claim 1, It is characterized in that Before the controller executes acquiring the free wake-up word prompt list according to the current interface scene, the controller is further configured as follows: The scene attributes of the interface scene and the wake-up-free word are registered in the voice service, and the scene attributes in the voice service and the wake-up-free word have a one-to-one mapping relationship.

3. The terminal device according to claim 2, It is characterized in that After the controller executes and obtains the voice control instruction input by the user based on the wake-up word prompt list, the controller is further configured to: Parsing the voice control command to determine the scene attributes of the current interface scene; Acquire the corresponding wake-up word prompt list according to the scene attributes; If the voice control instruction includes a wake-up-free word in the wake-up-free word prompt list, the operation corresponding to the wake-up-free word is performed according to the current interface scene.

4. The terminal device according to claim 3, It is characterized in that After the controller executes acquiring the corresponding wake-up word prompt list according to the scene attribute, the controller is further configured to: The display is controlled to display the wake-up-free words in the wake-up-free word prompt list one by one at a preset time interval, and the wake-up-free words are composed of sentences in a fixed format and displayed on the terminal device.

5. The terminal device according to claim 4, It is characterized in that The controller is further configured to: Adding an index page number after some of the wake-up-free words to obtain an index list, wherein the index page number is used to hide some of the wake-up-free words under the index page number; Obtain the wake-up-free word to be displayed and read the index list; If the index list is empty, displaying the wake-up-free word; If the index list contains the wake-up-free word with the index page number, and the current page number contains the index page number, the wake-up-free word is hidden.

6. The terminal device according to claim 5, It is characterized in that The controller is further configured to: Obtain the wake-up-free word to be displayed and read the index list; If the index list contains the wake-up-free word with the index page number, and the current page number does not contain the index page number, the wake-up-free word is displayed.

7. The terminal device according to claim 1, It is characterized in that The controller is further configured to: Obtaining a switching instruction for switching the interface scene input by the user through the free wake-up word in the free wake-up word prompt list; In response to the switching instruction, extracting the wake-up-free word corresponding to the interface scene before switching; After the terminal device enters the switched interface scene, marking the extracted wake-up-free word as an invalid wake-up-free word, where the invalid wake-up-free word is a voice control instruction that cannot perform a corresponding operation; The display is controlled not to display invalid wake-up-free words in the wake-up-free word prompt list.

8. The terminal device according to claim 2, It is characterized in that The interface scene includes multiple recipe interfaces, and the controller is further configured to: Recording the number of times the recipe interface is woken up by the user; Sort by the number of times the recipe interface is viewed.

9. A server, It is characterized in that The server comprises: A communication device configured to establish a communication connection with a user and a terminal device; A control device, configured to receive a voice wake-up instruction input by a user, determine an interface scene according to the voice wake-up instruction, and determine a wake-up-free word prompt list according to the interface scene, wherein the wake-up-free word prompt list includes a plurality of wake-up-free words; Sending the wake-up-free word prompt list to a terminal device so that the terminal device displays the wake-up-free word; and obtaining a voice control instruction input by a user based on the wake-up-free word prompt list, and if the voice control instruction contains a wake-up-free word in the wake-up-free word prompt list, switching the current interface scene to an interface scene after executing an operation corresponding to the wake-up-free word; When the scene attributes of the current interface scene are different from the scene attributes of the interface scene after executing the operation corresponding to the wake-up-free word, the wake-up-free word in the wake-up-free word prompt list is replaced according to the wake-up-free word corresponding to the changed scene attributes.

10. A wake-up-free voice control method for a terminal device, It is characterized in that Applied to a terminal device, the terminal device includes a display and a controller, and the method includes: Get the voice wake-up command input by the user; In response to the voice wake-up instruction, detecting the current interface scene according to the voice wake-up instruction, and controlling the terminal device to enter a wake-up-free mode, wherein the wake-up-free mode is used to enable continuous voice interaction between the user and the terminal device; Acquire a free wake-up word prompt list according to the current interface scene, where the free wake-up word prompt list includes a plurality of free wake-up words; Obtaining a voice control instruction input by a user based on the wake-up-free word prompt list, and if the voice control instruction contains a wake-up-free word in the wake-up-free word prompt list, switching the current interface scene to an interface scene after executing an operation corresponding to the wake-up-free word; When the scene attributes of the current interface scene are different from the scene attributes of the interface scene after executing the operation corresponding to the wake-up-free word, the wake-up-free word in the wake-up-free word prompt list is replaced according to the wake-up-free word corresponding to the changed scene attributes.

Citation Information

Patent Citations

  • Menu interaction method and system and storage medium

    CN112256230A

  • Interaction control method and device, intelligent voice equipment and storage medium

    CN114356275A