Voice Command Execution Method, Device, Smart Terminal, and Storage Medium
By obtaining voice operation request information and parsing it into text information, determining field information and operation event types, and generating operation instructions, the problem of low recognition accuracy of voice assistants is solved, and more efficient voice command execution and simplifying user operations are achieved.
Patent Information
- Application Number
- CN202011423206.7
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2020-12-08
- Publication Date
- 2025-08-05
- Estimated Expiration
- 2040-12-08
AI Technical Summary
In the prior art, the voice assistant has a low accuracy in identifying user intentions, and in order to identify new user intentions, users need to frequently update the voice assistant, which is cumbersome to operate and affects the user experience.
By obtaining voice operation request information, parsing and converting it into text information, determining field information, matching operation event types based on field information and preset mapping files, generating operation instructions and executing, avoiding frequent updates to voice assistants.
It improves the execution accuracy of voice commands, simplifies user operations, reduces the need for voice assistants to update, and improves user experience.
Smart Images

Figure CN114664296B_ABST
Abstract
Description
Technical Field
[0001] The present invention relates to the technical field of voice command execution, and in particular to a voice command execution method, device, intelligent terminal and storage medium. Background Art
[0002] Today, with the increasing prosperity of natural language processing technology, voice interaction technology has become more and more mature. Voice assistants are now widely used in various IOT devices (Internet of Things devices) such as mobile phones, TVs, and computers, covering various fields such as movies and TV shows, weather queries, device control, shopping, and consumption. However, the fields involved in voice interaction are diverse. In the prior art, voice interaction is basically realized based on voice assistants, but the recognition accuracy of user intentions by voice assistants is not high. Moreover, in order to recognize new user intentions, users need to frequently update the voice assistants to adapt to new needs, which is cumbersome to operate and affects user experience.
[0003] Therefore, the prior art still needs to be improved and enhanced. Summary of the Invention
[0004] The technical problem to be solved by the present invention is to provide a voice command execution method, device, intelligent terminal and storage medium for the above-mentioned defects of the prior art, aiming to solve the problems that the recognition accuracy of user intentions by voice assistants in the prior art is not high and users need to frequently update the voice assistants to recognize new user intentions.
[0005] To solve the above technical problems, the technical solutions adopted by the present invention are as follows:
[0006] In a first aspect, the present invention provides a voice command execution method, where the method includes:
[0007] Obtain voice operation request information generated based on a voice command, and determine an operation event type corresponding to the voice command according to the voice operation request information;
[0008] Determine an operation command corresponding to the operation event type according to the operation event type;
[0009] Execute the operation corresponding to the operation command according to the operation command.
[0010] In an implementation manner, the obtaining voice operation request information generated based on a voice command and determining an operation event type corresponding to the voice command according to the voice operation request information includes
[0011] Parse the voice operation request information to obtain the voice information corresponding to the voice command in the voice operation request information;
[0012] Convert the voice information into text information, and determine the operation event type corresponding to the voice command according to the text information.
[0013] In one implementation, the determining the operation event type corresponding to the voice command according to the text information includes:
[0014] Parse the text information to obtain the field information in the text information;
[0015] Determine the operation event type corresponding to the field information according to the field information.
[0016] In one implementation, the determining the operation event type corresponding to the field information according to the field information includes:
[0017] According to the field information, obtain the candidate operation event types that match the field information;
[0018] Obtain the priority information of the candidate operation event types, and determine the operation event type corresponding to the field information according to the priority information.
[0019] In one implementation, the determining the operation event type corresponding to the field information according to the field information includes:
[0020] Rewrite the field information into specified field information;
[0021] Generate the specified operation event type corresponding to the specified field information according to the specified field information, and use the specified operation event type as the operation event type.
[0022] In one implementation, the determining the operation instruction corresponding to the operation event type according to the operation event type includes:
[0023] Determine the name information of the operation event type according to the operation event type;
[0024] Determine the operation instruction corresponding to the name information according to the name information.
[0025] In one implementation, the determining the operation instruction corresponding to the name information according to the name information includes:
[0026] Determine the instruction template corresponding to the name information according to the name information;
[0027] Obtain the application name corresponding to the operation event type;
[0028] Fill in the application name in the instruction template to generate the operation instruction, which is used to operate the application corresponding to the application name.
[0029] In a second aspect, an embodiment of the present invention further provides a voice instruction execution device, where the device includes:
[0030] A voice instruction analysis unit, configured to obtain voice operation request information generated based on a voice instruction, and determine an operation event type corresponding to the voice instruction according to the voice operation request information;
[0031] An operation instruction determination unit, configured to determine an operation instruction corresponding to the operation event type according to the operation event type;
[0032] An operation instruction execution unit, configured to execute an operation corresponding to the operation instruction according to the operation instruction.
[0033] In a third aspect, an embodiment of the present invention further provides an intelligent terminal, where the intelligent terminal includes a memory, a processor, and a voice instruction execution program stored on the memory and executable on the processor. When the voice instruction execution program is executed by the processor, the steps of the voice instruction execution method described in any one of the above solutions are implemented.
[0034] In a fourth aspect, an embodiment of the present invention further provides a computer-readable storage medium, on which a voice instruction execution program is stored. When the voice instruction execution program is executed by a processor, the steps of the voice instruction execution method described in any one of the above solutions are implemented.
[0035] Advantageous effects: Compared with the prior art, the present invention provides a voice instruction execution method. First, the present invention first obtains voice operation request information generated based on a voice instruction, and determines an operation event type corresponding to the voice instruction according to the voice operation request information. Then, an operation instruction corresponding to the operation event type is determined according to the operation event type, and finally, an operation corresponding to the operation instruction is executed according to the operation instruction. Since in the present invention, the corresponding operation event type is determined according to the voice instruction, and then the corresponding operation instruction is generated based on the operation event type, and the voice instruction is executed by executing the operation instruction, therefore, in the present invention, as long as a voice instruction is obtained, only the operation event type needs to be determined according to the voice instruction, and then the corresponding operation instruction can be determined according to the operation event type. Compared with a voice assistant in the prior art, the present invention can improve the execution accuracy of voice instructions. Description of the Drawings
[0036] Figure 1It is a flowchart of the specific implementation manner of the voice command execution method provided by the embodiment of the present invention.
[0037] Figure 2 It is a flowchart of determining the operation event type in the voice command execution method provided by the embodiment of the present invention.
[0038] Figure 3 It is a flowchart of determining the operation instruction in the voice command execution method provided by the embodiment of the present invention.
[0039] Figure 4 It is a principle block diagram of the video screen dynamic moving device provided by the embodiment of the present invention.
[0040] Figure 5 It is a principle block diagram of the internal structure of the intelligent terminal provided by the embodiment of the present invention. Specific implementation manner
[0041] To make the purpose, technical solution and effects of the present invention clearer and more definite, the following further elaborates on the present invention by way of examples with reference to the accompanying drawings. It should be understood that the specific examples described herein are only used to explain the present invention and are not used to limit the present invention.
[0042] Today, with the increasingly booming natural language processing technology, the technology of voice interaction is also becoming more and more mature. Voice assistants are now widely used in various IOT devices (Internet of Things devices) such as mobile phones, TVs, and computers, covering various fields such as movies and TV shows, weather inquiries, device control, shopping, and consumption. However, the fields involved in voice interaction are diverse. In the prior art, voice interaction is basically realized based on voice assistants, but the recognition accuracy of the user's intention by voice assistants is not high. Moreover, in order to recognize new user intentions, users need to frequently update the voice assistants to adapt to new requirements, which is cumbersome to operate and affects the user experience. For example, in the prior art, on intelligent fitness equipment in a gym, users can often start corresponding functions on the intelligent fitness equipment by sending voice messages, such as starting a certain fitness mode. However, if the voice message spoken by the user cannot be recognized by the intelligent fitness equipment, the intelligent fitness equipment cannot respond to the user's intention (that is, the voice assistant cannot determine the user's intention based on the user's voice message), which affects the user's operation. And in order to make the intelligent fitness equipment respond to the user's voice message, it is necessary to upgrade the voice assistant on the intelligent fitness equipment so that the voice assistant can recognize the user's voice message, which affects the user's operation.
[0043] For this reason, this embodiment provides a method for executing a voice command. Through the method of this embodiment, the intention of the voice information can be accurately determined, so as to accurately respond to the voice information without considering whether the voice assistant needs to be upgraded. Specifically, in implementation, this embodiment first obtains voice operation request information generated based on a voice command, and determines an operation event type corresponding to the voice command according to the voice operation request information. Then, an operation instruction corresponding to the operation event type is determined according to the operation event type, and finally, the operation corresponding to the operation instruction is executed according to the operation instruction. Since in this embodiment, the corresponding operation event type is determined according to the voice command, and then the corresponding operation instruction is generated based on the operation event type, and the voice command is executed by executing the operation instruction, therefore, in this embodiment, as long as the voice command is obtained, only the operation event type needs to be determined according to the voice command, and then the corresponding operation instruction is determined according to the operation event type. Compared with the voice assistant in the prior art, the execution accuracy of the voice command can be improved in this embodiment.
[0044] Illustrated by way of example, the method of this embodiment is applied to the scenario of a gym. When the intelligent fitness equipment obtains the voice operation request information, the corresponding operation event type can be determined. Since the voice operation request information is generated based on a voice command, if the user's voice command is "start the jogging function", the voice operation request information will include the request of "start the jogging function". The intelligent fitness equipment can determine that the event operation type is the event type of starting the "running" mode, so the operation instruction can be determined according to the determined operation event type. This operation instruction is the instruction to start the "running" function mode. Therefore, the intelligent fitness equipment can start the "running" function mode according to the determined operation instruction to meet the needs of the user.
[0045] Exemplary method
[0046] The voice command execution method of this embodiment can be applied to an intelligent terminal, specifically as Figure 1 shown in, the voice command execution method specifically includes the following steps:
[0047] Step S100, obtain voice operation request information generated based on a voice command, and determine an operation event type corresponding to the voice command according to the voice operation request information.
[0048] In this embodiment, when a user wants to complete a certain operation, a voice command is sent to the smart terminal. This voice command is the voice information spoken by the user. Then, after the smart terminal obtains this voice command, it generates a voice operation request information according to this voice command. This voice operation request information is the request information used to reflect that the user wants to execute this voice command. For example, when the user wants to watch a wuxia drama, the user will send a voice command of "Play wuxia drama" to the smart TV. After the smart TV receives this voice command of "Play wuxia drama", it will generate a voice operation request information that the user wants to play a wuxia drama. When the smart terminal receives this voice operation request information, this embodiment can determine the operation event type corresponding to this voice command according to this voice operation request information. The operation event type in this embodiment reflects what type of operation the user hopes the smart terminal to complete. For example, in the above example, the operation event type corresponding to the voice command of "Play wuxia drama" is a play event, so that the smart TV can determine that the operation that the smart TV needs to complete is a play operation.
[0049] In one implementation, as Figure 2 shown in, step S100 specifically includes the following steps:
[0050] Step S101: Analyze the voice operation request information to obtain the voice information corresponding to the voice command in the voice operation request information;
[0051] Step S102: Convert the voice information into text information, and determine the operation event type corresponding to the voice command according to the text information.
[0052] Specifically in implementation, a recording function is set in the smart terminal in this embodiment. When the user outputs the voice information of "Play wuxia drama", the smart terminal can receive this voice information of "Play wuxia drama" by using the recording function. This voice information is the voice command. After the smart terminal obtains this voice command of "Play wuxia drama", it can generate the corresponding voice operation request information, and this voice operation request information includes this voice command of "Play wuxia drama". Therefore, when the smart terminal in this embodiment determines the voice information according to this voice operation request information, it only needs to analyze this voice operation request information to obtain the voice information corresponding to this voice command - "Play wuxia drama". In order to determine the operation event type corresponding to the voice command in this embodiment, this embodiment can convert the voice information into text information after obtaining the voice information, and then determine the operation event type corresponding to the voice command according to the text information.
[0053] In one implementation, in this embodiment, converting voice information into text information can be achieved by means of speech recognition or speech translation. After obtaining the text information, this embodiment can parse the text information, and then according to the field information in the text information, where the field information is some words or terms in the text information. Since the text information is a sentence or some words recognized from voice information, and some words have no meaning. For example, when the recognized text information is "Please play the currently popular TV drama Empresses in the Palace". Only the two field information "play" and "Empresses in the Palace" play a role in the execution process of the voice instruction in this text information. Because as long as these two field information are obtained, the intelligent terminal can determine the user's intention. And the "Please" and "currently popular" in this text information have no meaning for the execution process of the voice instruction. It can be seen that when determining the field information according to the text information in this embodiment, it is necessary to screen out some useless or meaningless fields in the text information to achieve the purpose of more accurately determining the user's intention.
[0054] After determining the field information, this embodiment can determine the operation event type corresponding to the field information according to the field information. In this embodiment, the field information is obtained from the text information, and the text information is obtained from the voice information corresponding to the voice instruction. Therefore, the field information can reflect the user's intention. For example, in the above example, the determined field information is "play" and "Empresses in the Palace", and the corresponding user intention is that the smart TV plays the TV drama Empresses in the Palace. And the user intention can reflect the type of the operation event. In one implementation, this embodiment can pre-set a mapping file, in which the corresponding relationship between the field information and the operation event type is set. Therefore, after obtaining the field information, the field information can be matched with the mapping file to match the corresponding operation event type. For example, in the above example, the determined field information is "play" and "Empresses in the Palace". According to the "play" in this field information, this embodiment can determine that the operation event type matched with "play" is "play event". Furthermore, according to the "Empresses in the Palace" in the field information, it can be determined that the one matched with "Empresses in the Palace" is "TV drama". Then, by combining the determined play event with the TV drama, the final play event type can be determined as: play TV drama.
[0055] In one implementation, the mapping file in this embodiment can be set based on the user's historical usage records. For example, for a smart fitness equipment, the functions used by the user are basically: running, skipping rope counting / timing, and playing fitness teaching videos. Therefore, for this smart fitness equipment, the mapping file can set the corresponding relationship between the setting field information and the operation event type according to the user's previous historical records. In addition, this embodiment can also determine the priority of the corresponding relationship between each field information and the operation event type according to the user's usage frequency. If there may be multiple intentions in the voice command output by the user, there will also be multiple determined field information, and there will be multiple corresponding operation event types, but the user's true intention may be only one. Therefore, this embodiment can use the operation event type determined according to the field information as the candidate operation event type, and then select the final operation event type from the candidate operation event types. To better implement the user's intention, this embodiment can sort the determined operation event types according to the priority of the field information and the operation event type. When multiple candidate operation event types are obtained, the one with a higher priority can be used as the final operation event type based on the priority of the candidate operation event types. For example, when the determined field information includes "play", "Empresses in the Palace", "volume", and "turn up", the corresponding candidate operation event types are: playing a movie and turning up the volume. Then, according to the priority of these two candidate operation event types, it is determined that the priority of playing a movie is higher than that of turning up the volume. Therefore, playing a movie is used as the final operation event type to achieve the user's true intention. Of course, if the user's voice command does indeed want to achieve multiple user intentions, this embodiment can also sort the determined operation event types according to the priority of the field information and the operation event type, and then sort according to each operation event type, so as to execute the voice command in the order of priority in subsequent steps. For example, when the determined field information includes "play", "Empresses in the Palace", "volume", and "turn up", the corresponding operation event types are: playing a movie and turning up the volume. Then, according to the priority of these two operation event types, it is determined that the priority of playing a movie is higher than that of turning up the volume. Therefore, in subsequent steps, the smart terminal can first execute the event of playing a movie and then execute the event of turning up the volume.
[0056] In addition, to meet more application scenarios, in this embodiment, the field information can also be rewritten, that is, the field information is customized, so that the operation type event can be customized. Specifically, when rewriting the field information, this embodiment can rewrite the field information obtained from the voice information into the specified field information, that is, the intended information desired by the user. For example, in the above example, the field information obtained from the voice information is "play" and "Empresses in the Palace", and since the film and television resources of "Empresses in the Palace" have been taken off the shelves, the smart TV cannot play this TV series. To meet the user's movie-watching needs, this embodiment can perform adaptive rewriting according to the determined field information. The so-called adaptive rewriting is to rewrite one or more of the field information into information associated with the original field information, that is, the specified field information, and then determine the specified operation event type according to the specified field information, and use the specified operation event type as the operation event type. For example, this embodiment can rewrite "Empresses in the Palace" into "Story of Yanxi Palace" because on the Internet, the two TV series "Empresses in the Palace" and "Story of Yanxi Palace" have the most associated entries. Therefore, when the field information is changed to the specified field information "Story of Yanxi Palace", the obtained operation event type is to play the TV series Story of Yanxi Palace, so that the user can watch movies normally.
[0057] Step S200: Determine the operation instruction corresponding to the operation event type according to the operation event type.
[0058] In this embodiment, determining the operation event type is to more accurately execute the voice command. After obtaining the operation event type, this embodiment can determine the operation instruction corresponding to the operation event type according to the operation event type. The operation instruction is used to complete the user intention corresponding to the operation event type, that is, the user intention realized by the operation instruction is the user intention in the voice command.
[0059] In one implementation, as Figure 3 shown in, the step S200 specifically includes the following steps:
[0060] Step S201: Determine the name information of the operation event type according to the operation event type;
[0061] Step S202: Determine the operation instruction corresponding to the name information according to the name information.
[0062] In specific implementation, after determining the type of operation event, this embodiment can obtain the name information corresponding to the type of operation event, and then retrieve the corresponding instruction template according to the name information. The instruction template is a template for generating instructions. Therefore, this embodiment can be pre-set with multiple instruction templates to determine the corresponding instruction template according to the name information. Since the type of operation event in this embodiment is obtained from the user's voice instruction, the user intention reflected is that the intelligent terminal is expected to perform corresponding operations after receiving the voice instruction. Therefore, the operation instruction generated in this embodiment is also for controlling the intelligent terminal to perform the operations corresponding to the voice instruction. Therefore, after obtaining the instruction template, this embodiment obtains the application name corresponding to the type of operation event, that is, obtains the name of the application that the user's voice instruction is to operate. Then, this embodiment fills the application name into the instruction template to generate the operation instruction, and the operation instruction is used to operate the application corresponding to the application name. For example, in the above example, when the determined type of operation event is the event of playing a movie or TV drama, the obtained instruction template is the instruction template for playing a movie or TV drama. Then, the obtained name information of the corresponding application is Youku APP, and then it is filled into the instruction template for playing a movie or TV drama to generate an operation instruction. Through this operation instruction, Youku APP can be opened and the movie or TV drama "Empresses in the Palace" can be played.
[0063] Step S300: According to the operation instruction, perform the operation corresponding to the operation instruction.
[0064] After obtaining the operation instruction, the intelligent terminal can perform the corresponding operation according to the operation instruction. For example, in the above example, when the determined type of operation event is the event of playing a movie or TV drama, the obtained instruction template is the instruction template for playing a movie or TV drama. Then, the obtained name information of the corresponding application is Youku APP, and then it is filled into the instruction template for playing a movie or TV drama to generate an operation instruction. Through this operation instruction, Youku APP can be opened and the movie or TV drama "Empresses in the Palace" can be played, so as to meet the user's movie-watching needs. Of course, the voice instruction execution method of this embodiment can be applied to other fields, such as being used in intelligent fitness equipment in a gym for fitness mode recommendation, etc.
[0065] In summary, in this embodiment, first, the voice operation request information generated based on the voice instruction is obtained, and the operation event type corresponding to the voice instruction is determined according to the voice operation request information. Then, the operation instruction corresponding to the operation event type is determined according to the operation event type, and finally, according to the operation instruction, the operation corresponding to the operation instruction is executed. Since in this embodiment, the corresponding operation event type is determined according to the voice instruction, and then the corresponding operation instruction is generated based on the operation event type, and the voice instruction is executed by executing the operation instruction, therefore, in this embodiment, as long as the voice instruction is obtained, only the operation event type needs to be determined according to the voice instruction, and then the corresponding operation instruction is determined according to the operation event type. Compared with the voice assistant in the prior art, the execution accuracy of the voice instruction can be improved in this embodiment.
[0066] Exemplary device
[0067] As Figure 4 shown in the figure, an embodiment of the present invention provides a video picture dynamic adjustment device, and the device includes: a voice instruction analysis unit 10, an operation instruction determination unit 20, and an operation instruction execution unit 30. Specifically, the voice instruction analysis unit 10 is configured to obtain voice operation request information generated based on a voice instruction, and determine an operation event type corresponding to the voice instruction according to the voice operation request information. The operation instruction determination unit 20 is configured to determine an operation instruction corresponding to the operation event type according to the operation event type. The operation instruction execution unit 30 is configured to execute the operation corresponding to the operation instruction according to the operation instruction.
[0068] In one implementation manner, the voice instruction analysis unit 10 includes:
[0069] A voice information determination subunit, configured to parse the voice operation request information to obtain voice information corresponding to the voice instruction in the voice operation request information;
[0070] An operation event type determination subunit, configured to convert the voice information into text information, and determine an operation event type corresponding to the voice instruction according to the text information.
[0071] In one implementation manner, the operation instruction determination unit 20 includes:
[0072] A name information determination subunit, configured to determine name information of the operation event type according to the operation event type;
[0073] An operation instruction determination subunit, configured to determine the operation instruction corresponding to the name information according to the name information.
[0074] Based on the above embodiments, the present invention further provides an intelligent terminal, and its principle block diagram can be as Figure 5 shown. The intelligent terminal includes a processor, a memory, a network interface, a display screen, and a temperature sensor connected through a system bus. Among them, the processor of the intelligent terminal is used to provide computing and control capabilities. The memory of the intelligent terminal includes a non-volatile storage medium and an internal memory. The non-volatile storage medium stores an operating system and a computer program. The internal memory provides an environment for the operation of the operating system and the computer program in the non-volatile storage medium. The network interface of the intelligent terminal is used to communicate with an external terminal through a network connection. When the computer program is executed by the processor, it realizes a method for executing voice instructions. The display screen of the intelligent terminal can be a liquid crystal display screen or an electronic ink display screen. The temperature sensor of the intelligent terminal is pre-set inside the intelligent terminal and is used to detect the operating temperature of internal devices.
[0075] Those skilled in the art can understand that Figure 5 the principle block diagram shown in
[0076] merely shows the block diagram of some structures related to the solution of the present invention, and does not constitute a limitation on the intelligent terminal to which the solution of the present invention is applied. The specific intelligent terminal may include more or fewer components than those shown in the figure, or combine certain components, or have different component arrangements.
[0077] Obtain the voice operation request information generated based on the voice instruction, and determine the operation event type corresponding to the voice instruction according to the voice operation request information;
[0078] Determine the operation instruction corresponding to the operation event type according to the operation event type;
[0079] Execute the operation corresponding to the operation instruction according to the operation instruction.
[0080] Those of ordinary skill in the art can understand that all or part of the processes in the methods of the above embodiments can be completed by instructing relevant hardware through a computer program. The computer program can be stored in a non-volatile computer-readable storage medium. When the computer program is executed, it can include the processes of the embodiments of the above methods. Among them, any reference to a memory, storage, database, or other medium used in the various embodiments provided by the present invention can include non-volatile and / or volatile memories. Non-volatile memory can include read-only memory (ROM), programmable ROM (PROM), electrically programmable ROM (EPROM), electrically erasable programmable ROM (EEPROM), or flash memory. Volatile memory can include random access memory (RAM) or external cache memory. By way of illustration and not limitation, RAM is available in various forms, such as static RAM (SRAM), dynamic RAM (DRAM), synchronous DRAM (SDRAM), double data rate SDRAM (DDR SDRAM), enhanced SDRAM (ESDRAM), synchronous link (Synchlink) DRAM (SLDRAM), memory bus (Rambus) direct RAM (RDRAM), direct memory bus dynamic RAM (DRDRAM), and memory bus dynamic RAM (RDRAM), etc.
[0081] In summary, the present invention discloses a voice instruction execution method, device, intelligent terminal, and storage medium. The method includes: obtaining voice operation request information generated based on a voice instruction, and determining an operation event type corresponding to the voice instruction according to the voice operation request information; determining an operation instruction corresponding to the operation event type according to the operation event type; and executing the operation corresponding to the operation instruction according to the operation instruction. The present invention determines the corresponding operation event type according to the voice instruction, then generates an operation instruction according to the operation event type, and thus executes the operation instruction. Thereby, when executing a voice instruction, it is completed by obtaining the corresponding operation event type, which is convenient for more quickly and accurately executing the voice instruction, and there is no need to consider the update of voice information, providing convenience for users.
[0082] Finally, it should be noted that the above embodiments are only used to illustrate the technical solutions of the present invention and are not intended to limit them. Although the present invention has been described in detail with reference to the foregoing embodiments, those of ordinary skill in the art should understand that they can still modify the technical solutions described in the foregoing embodiments or equivalently replace some of the technical features. These modifications or replacements do not make the essence of the corresponding technical solutions deviate from the spirit and scope of the technical solutions of the various embodiments of the present invention.
Claims
1. A method for executing a voice command, characterized in that: The method comprises: Acquire voice operation request information generated based on the voice instruction, and determine the operation event type corresponding to the voice instruction according to the voice operation request information; Determining, according to the operation event type, an operation instruction corresponding to the operation event type; According to the operation instruction, perform the operation corresponding to the operation instruction; The acquiring of voice operation request information generated based on the voice instruction, and determining the type of operation event corresponding to the voice instruction according to the voice operation request information, includes: Parsing the voice operation request information to obtain voice information corresponding to the voice instruction in the voice operation request information; Converting the voice information into text information, parsing the text information, and obtaining field information in the text information; Acquire, according to the field information, a candidate operation event type that matches the field information; Acquire priority information of the candidate operation event type, and determine the operation event type corresponding to the field information based on the priority information; The acquiring, according to the field information, a candidate operation event type matching the field information includes: The field information is matched with a preset mapping file to obtain a candidate operation event type matching the field information, wherein the mapping file is set based on historical usage records and contains a correspondence between the field information and the candidate operation event type.
2. The method for executing a voice command according to claim 1, wherein: The determining, according to the operation event type, an operation instruction corresponding to the operation event type includes: Determining name information of the operation event type according to the operation event type; According to the name information, the operation instruction corresponding to the name information is determined.
3. The voice command execution method according to claim 2, characterized in that: The determining, based on the name information, the operation instruction corresponding to the name information includes: Determining, according to the name information, an instruction template corresponding to the name information; Obtain the name of the application corresponding to the operation event type; The application name is entered into the instruction template to generate the operation instruction, which is used to operate the application corresponding to the application name.
4. A voice command execution device, characterized in that: The device is used to implement the steps of the method for executing a sound instruction according to any one of claims 1 to 3, and the device includes: a voice instruction analysis unit, configured to obtain voice operation request information generated based on the voice instruction, and determine an operation event type corresponding to the voice instruction according to the voice operation request information; an operation instruction determining unit, configured to determine an operation instruction corresponding to the operation event type according to the operation event type; The operation instruction execution unit is used to execute the operation corresponding to the operation instruction according to the operation instruction.
5. An intelligent terminal, characterized in that: The smart terminal includes a memory, a processor, and a voice instruction execution program stored in the memory and executable on the processor. When the voice instruction execution program is executed by the processor, the steps of the voice instruction execution method according to any one of claims 1 to 3 are implemented.
6. A computer-readable storage medium, characterized in that A voice instruction execution program is stored thereon, and when the voice instruction execution program is executed by the processor, the steps of the voice instruction execution method according to any one of claims 1 to 3 are implemented.
Citation Information
Patent Citations
A control method and device for intelligent apparatuses
CN106406806A
Voice recognition method and device, electronic equipment and storage medium
CN110675870A
Voice instruction execution method and device, cloud server and storage medium
CN114446292A