A control method based on dialect voice interaction, intelligent terminal and storage medium
By acquiring voice request information and determining dialect settings, the problem of lacking dialect voice interaction control in existing technologies has been solved, enabling dialect-based operation of smart home devices and improving the user experience.
Patent Information
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2020-12-22
- Publication Date
- 2026-04-14
AI Technical Summary
The lack of existing technologies for controlling smart home devices based on dialect voice interaction makes it difficult for users in urban and rural areas and dialect regions to use smart devices.
By acquiring voice request information, determining dialect setting information, determining dialect speech instructions based on dialect setting information, and executing corresponding dialect control operations, dialect-based voice interaction control is achieved.
It meets the needs of users in urban and rural areas and dialect regions, and provides a convenient smart home device operation experience, especially the dialect interaction method, which makes it easy for the elderly to use.
Smart Images

Figure CN114664298B_ABST
Abstract
Description
Technical Field
[0001] This invention relates to the field of artificial intelligence technology, and in particular to a control method, smart terminal, and storage medium based on dialect voice interaction. Background Technology
[0002] Voice interaction devices are network devices that allow users to control and operate equipment via voice input, replacing remote controls or keyboard input. These include various IoT devices such as smart TVs, smart refrigerators, smart speakers, and smart integrated stoves. With the development of the internet and artificial intelligence, intelligent voice interaction is becoming increasingly widespread. In voice interaction, semantic recognition and information processing are extremely important. Traditional intelligent voice interaction technologies typically use Mandarin for announcements. However, as smart devices enter homes and the number of users in urban and rural areas increases, dialect-based voice interaction for controlling smart home devices is becoming more crucial. Currently, there is no technology for controlling smart home devices based on dialect-based voice interaction.
[0003] Therefore, existing technologies still need improvement and development. Summary of the Invention
[0004] The technical problem to be solved by the present invention is to provide a control method based on dialect voice interaction, which addresses the above-mentioned deficiencies of the prior art and aims to solve the problem that there is no control method for smart home devices based on dialect voice interaction in the prior art.
[0005] The technical solution adopted by this invention to solve the problem is as follows:
[0006] In a first aspect, embodiments of the present invention provide a control method based on dialect voice interaction, wherein the method includes:
[0007] Obtain voice request information, and determine the dialect setting information corresponding to the voice request information based on the voice request information;
[0008] Based on the dialect setting information, determine the dialect speech instructions corresponding to the dialect setting information;
[0009] According to the dialect speech instructions, execute the dialect control operation corresponding to the dialect speech instructions.
[0010] In one implementation, obtaining the voice request information includes:
[0011] Collect voice information within a preset range;
[0012] Based on the voice information, determine the voice request information.
[0013] In one implementation, obtaining the voice request information further includes:
[0014] Detect network transmission status;
[0015] When network transmission is normal, the system enters standby mode.
[0016] In one implementation, determining the dialect setting information corresponding to the voice request information includes:
[0017] The voice request information is converted into text information corresponding to the voice request information;
[0018] Based on the text information, determine the dialect setting information corresponding to the text information.
[0019] In one implementation, determining the dialect setting information corresponding to the text information includes:
[0020] The text information is parsed to determine the behavioral information and key information corresponding to the text information; wherein, the behavioral information is the intent information expressed by the user;
[0021] When the key information contains dialect keywords, the behavioral information and the key information are used as dialect setting information.
[0022] In one implementation, determining the dialect speech instruction corresponding to the dialect setting information based on the dialect setting information includes:
[0023] Based on the key information, determine the dialect language field that matches the key information;
[0024] Based on the aforementioned dialect language domain, determine the dialect language instruction corresponding to the behavioral information.
[0025] In one implementation, the step of performing a dialect control operation corresponding to the dialect speech instruction includes:
[0026] Switch the voice mode to the dialect playback mode corresponding to the dialect speech command; wherein, the dialect playback mode is a dialect broadcast and polling mode;
[0027] According to the dialect playback mode, the dialect speech corresponding to the dialect speech command is played.
[0028] Secondly, embodiments of the present invention also provide a control device based on dialect voice interaction, wherein the device includes:
[0029] A dialect setting information determination unit is used to acquire voice request information and determine dialect setting information based on the voice request information;
[0030] The dialect speech instruction acquisition unit is used to determine the dialect speech instruction corresponding to the dialect setting information based on the dialect setting information.
[0031] The dialect control operation unit is used to execute dialect control operations corresponding to the dialect speech instructions.
[0032] Thirdly, embodiments of the present invention also provide a smart terminal, including a memory and one or more programs, wherein one or more programs are stored in the memory and configured to be executed by one or more processors, the one or more programs including a control method for performing dialect-based voice interaction as described in any of the above.
[0033] Fourthly, embodiments of the present invention also provide a non-transitory computer-readable storage medium, wherein when the instructions in the storage medium are executed by a processor of an electronic device, the electronic device is able to execute the control method based on dialect voice interaction as described in any of the above.
[0034] The beneficial effects of this invention are as follows: In this embodiment, the system first acquires a user's voice request information; then, based on the voice request information, it determines the dialect setting information corresponding to the voice request information. After acquiring the dialect setting information, the system can identify the content of the dialect setting, i.e., what dialect setting is being performed. Then, based on the dialect setting information, it determines the corresponding dialect speech commands; thus, it can output a series of speech commands in that dialect. With these speech commands, the system can execute corresponding dialect control operations, i.e., realize different dialect control operations. Therefore, in this embodiment, by switching the voice to a dialect, the system achieves dialect-based voice interaction control of the mini smart screen, meeting the requirements of users in urban and rural areas and dialect-speaking regions, and bringing convenience to users. Attached Figure Description
[0035] To more clearly illustrate the technical solutions in the embodiments of the present invention or the prior art, the drawings used in the description of the embodiments or the prior art will be briefly introduced below. Obviously, the drawings described below are only some embodiments recorded in the present invention. For those skilled in the art, other drawings can be obtained based on these drawings without creative effort.
[0036] Figure 1 This invention provides a schematic flowchart of a control method based on dialect voice interaction.
[0037] Figure 2 A schematic diagram of the control device based on dialect voice interaction provided in this embodiment of the invention.
[0038] Figure 3 The internal structure principle block diagram of the smart terminal provided in the embodiment of the present invention. Detailed Implementation
[0039] This invention discloses a control method, a smart terminal, and a storage medium based on dialect voice interaction. To make the objectives, technical solutions, and effects of this invention clearer and more explicit, the invention is further described in detail below with reference to the accompanying drawings and embodiments. It should be understood that the specific embodiments described herein are only for explaining the invention and are not intended to limit the invention.
[0040] Those skilled in the art will understand that, unless specifically stated otherwise, the singular forms “a,” “an,” “the,” and “the” used herein may also include the plural forms. It should be further understood that the term “comprising” as used in this specification means the presence of the stated features, integers, steps, operations, elements, and / or components, but does not exclude the presence or addition of one or more other features, integers, steps, operations, elements, components, and / or groups thereof. It should be understood that when we say an element is “connected” or “coupled” to another element, it can be directly connected or coupled to the other element, or there may be intermediate elements. Furthermore, “connected” or “coupled” as used herein can include wireless connections or wireless coupling. The term “and / or” as used herein includes all or any units and all combinations of one or more associated listed items.
[0041] It will be understood by those skilled in the art that, unless otherwise defined, all terms used herein (including technical and scientific terms) have the same meaning as commonly understood by one of ordinary skill in the art to which this invention pertains. It should also be understood that terms such as those defined in general dictionaries should be understood to have the same meaning as in the context of the prior art, and should not be interpreted in an idealized or overly formal sense unless specifically defined as herein.
[0042] In existing technologies, semantic recognition and information processing are extremely important in voice interaction. Traditional intelligent voice interaction technologies typically use Mandarin for announcements. However, as smart devices become more widespread and the number of users in urban and rural areas increases, dialect-based voice interaction for controlling smart home devices has become even more crucial. Currently, there is no technology for controlling smart home devices based on dialect-based voice interaction.
[0043] To address the problems of existing technologies, this embodiment provides a control method based on dialect voice interaction. This embodiment first acquires a user's voice request information; then, based on the voice request information, it determines the dialect setting information corresponding to the voice request information. After acquiring the dialect setting information, the MINI Smart Screen system can identify the content of the dialect setting, i.e., what dialect setting is being performed. Then, based on the dialect setting information, it determines the dialect speech commands corresponding to the dialect setting information; it can then output a series of speech commands corresponding to that dialect. With these speech commands, the MINI Smart Screen system can execute corresponding dialect control operations, i.e., implement different dialect control operations. In other words, by determining the dialect-related dialect speech command series for the set dialect, the MINI Smart Screen system can execute the dialect control operation corresponding to the dialect speech command. Therefore, this embodiment of the invention achieves dialect-based voice interaction control of the mini smart screen by switching the voice to a dialect, meeting the requirements of users in urban and rural areas and dialect-speaking regions, and bringing convenience to users.
[0044] For example
[0045] Even with the same smart home device, different families may encounter difficulties using it due to varying family members' origins and dialects. If the smart home device could only interact via Mandarin, it would pose a significant challenge for elderly family members. Therefore, this invention adds a database of selectable dialect phrases to the MINI smart screen within the smart home device. After purchasing the smart home product, each family can customize the voice settings according to their own dialect. For example, younger family members can configure the MINI smart screen to use the local dialect of the elderly family members. When the younger family members are at work, the elderly can then interact with the MINI smart screen using their dialect to control the smart home device. In this embodiment, the MINI Smart Screen in the smart home device first acquires voice request information and determines the dialect setting information corresponding to the voice request information. The function of the dialect setting information is to set different dialects for the received voice. For example, a young person in the family wakes up the smart home device with a wake-up phrase and then sends a voice request. After receiving the voice request information sent by the user, the smart home device determines the dialect setting information corresponding to the voice request information. Then, it determines the dialect speech command corresponding to the dialect setting information, and the voice of the smart home device is converted into the local dialect, such as Cantonese. In this way, when the young person goes to work, the smart home device will output the speech command corresponding to Cantonese. The smart home device executes the dialect control operation corresponding to the dialect speech command, so the elderly in the family can interact with the smart home device in Cantonese and clearly understand what kind of control operation the smart home device is performing, which is convenient for the elderly to use these smart home products correctly.
[0046] Exemplary methods
[0047] This embodiment provides a control method based on dialect voice interaction, which can be applied to smart terminals in smart homes. Specifically, as follows... Figure 1 As shown, the method includes the following steps:
[0048] Step S100: Obtain voice request information, and determine the dialect setting information corresponding to the voice request information based on the voice request information;
[0049] In this embodiment, voice request information is acquired through the MINI Smart Screen. A young person in the family wakes up the smart home device using a wake-up phrase and then sends a voice request. Upon receiving the voice request, the smart home device determines the corresponding dialect setting based on the voice request. For example, if a young person wakes up the smart home device using a wake-up phrase and then sends a voice request containing a dialect setting, the smart home device can receive the user's voice setting request and determine the included dialect setting, such as switching to Cantonese.
[0050] To obtain the voice request information, the process of obtaining the voice request information includes the following steps:
[0051] Step S101: Collect voice information within a preset range;
[0052] Step S102: Determine the voice request information based on the voice information.
[0053] Specifically, in the control system of the MINI Smart Screen in the smart home device, the voice assistant microphone detects the wake-up voice message, and the voice assistant then activates the voice acquisition module and starts the recording function. In one implementation, the wake-up voice message (e.g., "Xiao T Xiao T") is used, and then the recording function is activated. In another implementation, the wake-up method can also be triggered externally (pressing the record button) / deactivating (releasing the record button), with no specific restrictions. In yet another implementation, after the voice assistant microphone in the control system of the smart home device receives the wake-up voice message, a pop-up image of "Xiao T" during the wake-up process is displayed on the MINI Smart Screen of the smart home device. Once the recording function is activated, it means that the recording function of the MINI Smart Screen in the smart home device is in a standby state. When in the first listening state and if the voice assistant does not detect any voice input information within a first preset time, it enters a second listening state; when in the second listening state and if the voice assistant does not detect any voice input information within a second preset time, the voice assistant exits the MINI Smart Screen display interface. Specifically, in the first listening state, if the voice assistant does not detect voice input within 3 seconds, it enters the second listening state and displays a pop-up message saying, "I didn't hear you clearly, please say it again." In the second listening state, if the voice assistant does not detect voice input within 6 seconds, it exits and displays the MINI Smart Screen interface. In practice, voice information generated during conversations is not within the preset collection range. Only voice information from family members that begins with a wake word and whose volume is within a certain range will be acquired by the voice assistant, thus obtaining the voice request information. At this time, when a young person in the family sends a voice request information containing dialect settings, the MINI Smart Screen in the smart home device can receive the voice request information and then determine the dialect settings contained therein.
[0054] In practice, the MINI Smart Screen in smart home devices must be connected to the internet to achieve intelligent interaction. Therefore, it is also necessary to detect the network transmission status. When the network transmission status is normal, the system enters standby mode and displays the wake-up image and listening text information on the MINI Smart Screen display interface. For example, the image pop-up displays the small T image for wake-up listening and the text "I am listening", and announces "I am listening". If the network transmission is abnormal, the text "Network abnormal, please check the network" will be displayed on the image pop-up. After 3 seconds, the voice assistant will disappear.
[0055] To obtain dialect setting information, determining the dialect setting information corresponding to the voice request information includes the following steps:
[0056] Step S103: Convert the voice request information into text information corresponding to the voice request information;
[0057] Step S104: Determine the dialect setting information corresponding to the text information based on the text information.
[0058] In practice, when the MINI Smart Screen in a smart home device receives a voice request, for example, if voice input is detected within 6 seconds (e.g., "Xiao T Xiao T, switch to Sichuan dialect"), the voice assistant receives the user's voice request containing dialect setting information, calls the Automatic Speech Recognition (ASR) function, and converts the received voice request information into text information. In one implementation, the text information is saved to storage and displayed in a pop-up window. Then, based on the text information, the dialect setting information contained in the text information can be determined.
[0059] To obtain dialect setting information, determining the dialect setting information corresponding to the text information includes the following steps: parsing the text information to determine the behavioral information and key information corresponding to the text information; wherein, the behavioral information is the user's expressed intent; when the key information contains dialect keywords, the behavioral information and the key information are used as dialect setting information.
[0060] Specifically, the text information is parsed to determine the behavioral information and key information. Behavioral information refers to the user's expressed intent. For example, "I want to turn on the TV," the behavioral information is "turn on," and the key information is "TV." "Switch to Sichuan dialect," the behavioral information is "switch," and the key information is "Sichuan dialect." When the key information contains a dialect keyword, the behavioral information and the key information are considered dialect setting information. For instance, if the parsed text information is "switch to Sichuan dialect," and its key information contains the dialect keyword "Sichuan dialect," then the behavioral information "switch" and the key information for "switch to Sichuan dialect" are determined to be dialect setting information.
[0061] This embodiment provides a control method based on dialect voice interaction, which can be applied to smart terminals in smart homes. Specifically, as follows... Figure 1 As shown, the method includes the following steps:
[0062] Step S200: Determine the dialect speech instruction corresponding to the dialect setting information based on the dialect setting information;
[0063] In this embodiment, after the MINI Smart Screen in the smart home device obtains the dialect setting information, it can determine the dialect speech command corresponding to the dialect setting information through a dialect speech model or a speech database corresponding to the dialect speech service. In this embodiment, the dialect speech model can be obtained by training and iterating the dialect using a neural network method. The speech database corresponding to the dialect speech service collects a series of commonly used dialects of the MINI Smart Screen in the smart home device and stores them in the dialect speech database. When a new dialect speech is discovered during actual use, the new dialect speech is added to the dialect speech database.
[0064] To obtain dialect speech instructions, determining the dialect speech instructions corresponding to the dialect setting information based on the dialect setting information includes the following steps:
[0065] Step S201: Based on the key information, determine the dialect language field that matches the key information;
[0066] Step S202: Determine the dialect speech instruction corresponding to the behavioral information based on the dialect speech domain.
[0067] Specifically, based on the key information, a dialect language domain is determined. For example, if the dialect setting information is "switch to Sichuan dialect," the behavior information is "switch," and Sichuan dialect is the key information, then the Sichuan dialect language domain can be determined based on the key information "Sichuan dialect." In this embodiment, based on the key information, it is determined whether the language database corresponding to the dialect voice service has dialect keywords corresponding to the key information. For example, the smart screen control system obtains the intent, key information, and classification of user commands from the storage, converts the command intent and key information into query language, and connects to the language database corresponding to the voice service in the technology center for knowledge query and reasoning. Then, based on the dialect language domain, the dialect language command corresponding to the behavior information is determined. For example, if the language database corresponding to the voice service in the technology center finds the corresponding dialect language command, the dialect command matches and switches to Sichuan dialect.
[0068] This embodiment provides a control method based on dialect voice interaction, which can be applied to smart terminals in smart homes. Specifically, as follows... Figure 1 As shown, the method includes the following steps:
[0069] Step S300: Execute the dialect control operation corresponding to the dialect speech instruction according to the dialect speech instruction.
[0070] Specifically, the MINI Smart Screen in the smart home device executes dialect control operations corresponding to the acquired dialect commands. These operations can include broadcasting and polling, turning on the smart home device, or issuing prompts, among other things. Thus, when the younger generation goes to work, the smart home device will output commands in Cantonese, allowing elderly family members to interact with the device in Cantonese and clearly understand the control operations being performed, facilitating their correct use of the smart home products.
[0071] In order to obtain dialect control operations, the step of executing dialect control operations corresponding to the dialect speech instructions includes the following steps:
[0072] Step S301: Switch the voice mode to the dialect playback mode corresponding to the dialect speech command; wherein, the dialect playback mode is a dialect broadcast and polling mode;
[0073] Step S302: Play the dialect speech corresponding to the dialect speech command according to the dialect playback mode.
[0074] Upon receiving a dialect-based speech command, the command and the MINI Smart Screen display interface are encapsulated and then sent to the MINI Smart Screen's control system. In one implementation, if the speech database corresponding to the voice service in the technology center does not support the corresponding dialect-based speech command, an "unsupported" message is returned, and the voice assistant function is disabled. When the MINI Smart Screen in the smart home device receives the encapsulated dialect voice information, it switches the voice mode to the dialect broadcast and polling mode corresponding to the dialect-based speech command, and plays the corresponding speech through a text-to-speech (TTS) synthesizer. The TTS synthesizer is a type of speech synthesis application that converts files stored in smart devices, such as help files or web pages, into natural speech output. TTS can convert sound into text. TTS applications include voice-driven email and voice-sensitive systems, and are often used in conjunction with Automatic Speech Recognition (ASR). Real-time conversion of text files can be performed, with conversion times measured in seconds. With its unique intelligent voice controller, the text output has a smooth tone, making the listener feel natural when listening to information, without any of the coldness or awkwardness of machine voice output.
[0075] To provide a better user experience, images, videos, or text information corresponding to the dialect commands are displayed on the MINI Smart Screen interface. This allows users to more intuitively understand the meaning of the dialect commands through images and videos, and to confirm the dialect control operations performed by the smart home devices through text information.
[0076] Exemplary device
[0077] like Figure 2 As shown in the figure, this embodiment of the invention provides a control device based on dialect voice interaction. The device includes a dialect voice setting request information acquisition unit 401, a dialect speech instruction acquisition unit 402, and a dialect control operation unit 403, wherein:
[0078] The dialect setting information determination unit 401 is used to acquire voice request information and determine dialect setting information based on the voice request information;
[0079] The dialect speech instruction acquisition unit 402 is used to determine the dialect speech instruction corresponding to the dialect setting information based on the dialect setting information.
[0080] The dialect control operation unit 403 is used to execute dialect control operations corresponding to the dialect speech instructions.
[0081] Based on the above embodiments, the present invention also provides a smart terminal, the principle block diagram of which can be as follows: Figure 3 As shown, the smart terminal includes a processor, memory, network interface, display screen, and temperature sensor connected via a system bus. The processor provides computing and control capabilities. The memory includes non-volatile storage media and internal memory. The non-volatile storage media stores the operating system and computer programs. The internal memory provides an environment for the operation of the operating system and computer programs stored in the non-volatile storage media. The network interface is used to communicate with external terminals via a network connection. When the computer program is executed by the processor, it implements a control method based on dialect voice interaction. The display screen can be an LCD screen or an e-ink screen. The temperature sensor is pre-installed inside the smart terminal to detect the operating temperature of internal devices.
[0082] Those skilled in the art will understand that Figure 3 The schematic diagram in the figure is merely a block diagram of a part of the structure related to the present invention and does not constitute a limitation on the smart terminal on which the present invention is applied. The specific smart terminal may include more or fewer components than shown in the figure, or combine certain components, or have different component arrangements.
[0083] In one embodiment, a smart terminal is provided, including a memory and one or more programs, wherein the one or more programs are stored in the memory and configured to be executed by one or more processors. The one or more programs include instructions for performing the following operations:
[0084] Obtain voice request information; wherein, the voice request information includes wake-up voice information and dialect voice setting request information;
[0085] Based on the dialect voice setting request information, determine the dialect speech instruction corresponding to the dialect voice setting request information;
[0086] According to the dialect speech instructions, execute the dialect control operation corresponding to the dialect speech instructions.
[0087] Those skilled in the art will understand that all or part of the processes in the methods of the above embodiments can be implemented by a computer program instructing related hardware. The computer program can be stored in a non-volatile computer-readable storage medium. When executed, the computer program can include the processes of the embodiments of the above methods. Any references to memory, storage, databases, or other media used in the embodiments provided by this invention can include non-volatile and / or volatile memory. Non-volatile memory can include read-only memory (ROM), programmable ROM (PROM), electrically programmable ROM (EPROM), electrically erasable programmable ROM (EEPROM), or flash memory. Volatile memory can include random access memory (RAM) or external cache memory. By way of illustration and not limitation, RAM is available in various forms, such as static RAM (SRAM), dynamic RAM (DRAM), synchronous DRAM (SDRAM), dual data rate SDRAM (DDRSDRAM), enhanced SDRAM (ESDRAM), synchronous link DRAM (SLDRAM), RAMbus direct RAM (RDRAM), direct memory bus dynamic RAM (DRDRAM), and RAMbus dynamic RAM (RDRAM), etc.
[0088] In summary, this invention discloses a control method, smart terminal, and storage medium based on dialect voice interaction. The method includes: firstly, acquiring a user's voice request information; wherein the voice request information includes a wake-up voice message and a dialect voice setting request; after acquiring the dialect voice setting request information, the system can identify the content of the request, i.e., what dialect voice setting to perform, and then determine the dialect speech command corresponding to the dialect voice setting request information. The system can then output a series of speech commands in that dialect. With these speech commands, the system can execute corresponding dialect control operations, i.e., realize different dialect control operations. Therefore, this invention, by switching the voice to a dialect, achieves dialect-based voice interaction control of the mini smart screen, meeting the requirements of users in urban and rural areas and dialect-speaking regions, and bringing convenience to users.
[0089] It should be understood that the present invention discloses a control method based on dialect voice interaction. It should also be understood that the application of the present invention is not limited to the examples above. Those skilled in the art can make improvements or modifications based on the above description, and all such improvements and modifications should fall within the protection scope of the appended claims.
Claims
1. A control method based on dialect voice interaction, characterized in that, The method includes: Obtain voice request information, and determine the dialect setting information corresponding to the voice request information based on the voice request information; Based on the dialect setting information, determine the dialect speech instructions corresponding to the dialect setting information; According to the dialect speech instructions, execute the dialect control operation corresponding to the dialect speech instructions; When the voice assistant does not detect voice input within a first preset time in the first listening state, it enters the second listening state. When the voice assistant is in the second listening state and does not detect voice input information within a second preset time, the voice assistant exits the MINI Smart Screen display interface. The step of determining the dialect setting information corresponding to the voice request information includes: The voice request information is converted into text information corresponding to the voice request information; Based on the text information, determine the dialect setting information corresponding to the text information; The step of determining the dialect setting information corresponding to the text information includes: The text information is parsed to determine the behavioral information and key information corresponding to the text information; wherein, the behavioral information is the intent information expressed by the user; When the key information contains dialect keywords, the behavioral information and the key information are used as dialect setting information. The step of determining the dialect speech instruction corresponding to the dialect setting information based on the dialect setting information includes: Based on the key information, determine the dialect language domain that matches the key information; Based on the dialect language domain, determine the dialect language instruction corresponding to the behavioral information; Based on the aforementioned dialect language domain, determine the dialect language instructions corresponding to the behavioral information, including: The dialect speech instructions corresponding to the dialect setting information are determined by the speech database; The step of executing the dialect control operation corresponding to the dialect speech instruction includes: Switch the voice mode to the dialect playback mode corresponding to the dialect speech command; wherein, the dialect playback mode is a dialect broadcast and polling mode; According to the dialect playback mode, the dialect speech corresponding to the dialect speech command is played.
2. The control method based on dialect voice interaction according to claim 1, characterized in that, The acquisition of voice request information includes: Collect voice information within a preset range; Based on the voice information, determine the voice request information.
3. The control method based on dialect voice interaction according to claim 2, characterized in that, The process of obtaining voice request information also includes: Detect network transmission status; When network transmission is normal, the system enters standby mode.
4. A control device based on dialect voice interaction, characterized in that, include: A dialect setting information determination unit is used to acquire voice request information and determine dialect setting information corresponding to the voice request information based on the voice request information. The dialect speech instruction acquisition unit is used to determine the dialect speech instruction corresponding to the dialect setting information based on the dialect setting information. The dialect control operation unit is used to execute dialect control operations corresponding to the dialect speech instructions according to the dialect speech instructions; The step of determining the dialect speech instruction corresponding to the dialect setting information based on the dialect setting information includes: Based on the key information, determine the dialect language domain that matches the key information; Based on the aforementioned dialect language domain, determine the dialect language instructions corresponding to the behavioral information; Based on the aforementioned dialect language domain, determine the dialect language instructions corresponding to the behavioral information, including: The dialect speech instructions corresponding to the dialect setting information are determined by the speech database; The step of executing the dialect control operation corresponding to the dialect speech instruction includes: Switch the voice mode to the dialect playback mode corresponding to the dialect speech command; wherein, the dialect playback mode is a dialect broadcast and polling mode; According to the dialect playback mode, the dialect voice corresponding to the dialect speech command is played; When the voice assistant does not detect voice input within a first preset time in the first listening state, it enters the second listening state. When the voice assistant is in the second listening state and does not detect voice input information within a second preset time, the voice assistant exits the MINI Smart Screen display interface. The step of determining the dialect setting information corresponding to the voice request information includes: The voice request information is converted into text information corresponding to the voice request information; Based on the text information, determine the dialect setting information corresponding to the text information; The step of determining the dialect setting information corresponding to the text information includes: The text information is parsed to determine the behavioral information and key information corresponding to the text information; wherein, the behavioral information is the intent information expressed by the user; When the key information contains dialect keywords, the behavioral information and the key information are used as dialect setting information.
5. A smart terminal, characterized in that, It includes a memory and one or more programs, wherein one or more programs are stored in the memory and configured to be executed by one or more processors, wherein the one or more programs include methods for performing any one of claims 1-3.
6. A non-transitory computer-readable storage medium, characterized in that, When the instructions in the storage medium are executed by the processor of the electronic device, the electronic device is able to perform the method as described in any one of claims 1-3.
Citation Information
Patent Citations
Voice instruction identification method and device, server and computer readable storage medium
CN108172223A
Method and device for synthesizing speech
CN110197655A