In-vehicle voice interaction method, device, computer device and storage medium
Through the method of voice command analysis and control simulation operation, the problem of increasing risks for users in the car computer system when driving is solved, voice interaction is used instead of manual operation, and user experience and security are improved.
Patent Information
- Application Number
- CN202310329850.5
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2023-03-30
- Publication Date
- 2025-06-24
- Estimated Expiration
- 2043-03-30
AI Technical Summary
In the existing car and computer systems, manually operating the application interface while driving will increase driving risks and the user experience is poor.
Through steps such as voice command analysis, control determination, simulation operation and display response, voice interaction can be used instead of manual operation, reducing driving risks and improving user experience.
It effectively reduces the risk of manual operation while driving, improves the interactive experience between users and the on-board system, and enhances safety and convenience.
Smart Images

Figure CN116403581B_ABST
Abstract
Description
Technical Field
[0001] The present application relates to the technical field of automotive voice interaction, and particularly to an in-vehicle voice interaction method, system, computer device, and storage medium. Background Art
[0002] There are many applications on the current in-vehicle system, such as multimedia applications, vehicle control and vehicle settings applications, map applications, settings applications, scene engine applications, etc. with interface applications. Various view (the base class of all controls in the Android system) controls will be displayed on these application interfaces. For these controls on the interfaces, users often can only manually click on the touch screen of the in-vehicle system to operate. When the vehicle owner is driving, manually operating the application interface will pose a certain danger.
[0003] Therefore, how to use voice commands to replace manual operations, reduce driving risks, and improve the user experience is an urgent problem to be solved. Summary of the Invention
[0004] Based on this, in view of the above technical problems, it is necessary to provide an in-vehicle voice interaction method, system, computer device, and storage medium that can use voice commands to replace manual operations, reduce driving risks, and improve the user experience.
[0005] On the one hand, an in-vehicle voice interaction method is provided, and the method includes:
[0006] Parse the voice operation command to be processed to obtain a first registered word;
[0007] According to the first registered word, determine the control to be operated corresponding to the first registered word, and the type of the control to be operated;
[0008] Based on the type of the control to be operated, perform a simulated operation on the control to be operated on a first interface;
[0009] In response to the process of the simulated operation, generate a second interface to display the response effect corresponding to the voice operation command.
[0010] In one of the embodiments, it further includes: the determining the control to be operated corresponding to the first registered word according to the first registered word includes: based on the first registered word, matching and obtaining a second registered word from a first service; according to the second registered word, determining the control to be operated corresponding to the first registered word from a second service, and obtaining the type of the control to be operated based on the attribute label of the control to be operated.
[0011] In one embodiment, it further includes: when it is detected that the second registered word is not matched in the first service based on the first registered word, the method further includes: prompting for the unmatched result, defining the first registered word as the third registered word, and saving the third registered word to the first service.
[0012] In one embodiment, it further includes: when it is detected that the number of occurrences of the third registered word is greater than a first preset value, calculating the similarity between the third registered word and a target second registered word in the first service; when it is detected that the similarity calculation result is greater than a second preset value and the first callback event is not monitored within a preset time, associating the third registered word with the target second registered word and saving it to the first service, and associating the third registered word with the control corresponding to the target second registered word and saving it to the second service.
[0013] In one embodiment, it further includes: when it is detected that the similarity calculation result is greater than a second preset value and the first callback event is monitored within a preset time, the method further includes: obtaining a fourth registered word corresponding to the first callback event in the second service; calculating the similarity between the target second registered word and the fourth registered word; when it is detected that the similarity calculation result is greater than a third preset value, associating the third registered word with the control corresponding to the fourth registered word and saving it to the second service, and associating the third registered word with the target second registered word and saving it to the first service.
[0014] In one embodiment, it further includes: the method for generating the second registered word includes: when a second callback event is detected, polling and traversing the control tree to obtain the control with the set attribute label corresponding to the second callback event and the registered word corresponding to the control, where one control corresponds to one or more registered words; sending the registered word to the first service for registration, generating the second registered word in response to the registration success result and saving it; associating the control and the second registered word corresponding to the control and sending it to the second service for saving.
[0015] In one embodiment, it further includes: performing a simulation operation on the to-be-operated control based on the type of the to-be-operated control to control the vehicle to execute the functional requirement corresponding to the voice operation instruction.
[0016] On the other hand, a vehicle-mounted voice interaction device is provided, and the device includes:
[0017] A voice service module, configured to parse a to-be-processed voice operation instruction to obtain a first registered word;
[0018] A control determination module, configured to determine a control to be operated corresponding to the first registered word and the type of the control to be operated according to the first registered word;
[0019] An analog operation module, configured to perform an analog operation on the control to be operated on a first interface based on the type of the control to be operated;
[0020] A display module, configured to generate a second interface in response to the process of the analog operation to display a response effect corresponding to the voice operation instruction.
[0021] On the other hand, a computer device is provided, including a memory, a processor, and a computer program stored on the memory and executable on the processor. When the processor executes the computer program, the following steps are implemented:
[0022] Parse a voice operation instruction to be processed to obtain a first registered word;
[0023] Determine a control to be operated corresponding to the first registered word and the type of the control to be operated according to the first registered word;
[0024] Perform an analog operation on the control to be operated on a first interface based on the type of the control to be operated;
[0025] Generate a second interface in response to the process of the analog operation to display a response effect corresponding to the voice operation instruction.
[0026] On another hand, a computer-readable storage medium is provided, on which a computer program is stored. When the computer program is executed by a processor, the following steps are implemented:
[0027] Parse a voice operation instruction to be processed to obtain a first registered word;
[0028] Determine a control to be operated corresponding to the first registered word and the type of the control to be operated according to the first registered word;
[0029] Perform an analog operation on the control to be operated on a first interface based on the type of the control to be operated;
[0030] Generate a second interface in response to the process of the analog operation to display a response effect corresponding to the voice operation instruction.
[0031] The above vehicle-mounted voice interaction method, device, computer device and storage medium, the method comprising: parsing a voice operation instruction to be processed to obtain a first registered word; determining, according to the first registered word, a control to be operated corresponding to the first registered word and the type of the control to be operated; based on the type of the control to be operated, performing a simulated operation on the control to be operated on a first interface; and generating a second interface in response to the process of the simulated operation to display a response effect corresponding to the voice operation instruction. The present application uses voice instructions instead of manual operations, improving the interaction experience between the user and the vehicle-mounted system and reducing driving risks. BRIEF DESCRIPTION OF THE DRAWINGS
[0032] Figure 1 It is an application environment diagram of a vehicle-mounted voice interaction method in an embodiment;
[0033] Figure 2 It is a schematic flowchart of a vehicle-mounted voice interaction method in an embodiment;
[0034] Figure 3 It is a schematic flowchart of a vehicle-mounted voice interaction step in an embodiment;
[0035] Figure 4 It is a structural block diagram of a vehicle-mounted voice interaction device in an embodiment;
[0036] Figure 5 It is an internal structure diagram of a computer device in an embodiment. DETAILED DESCRIPTION OF THE EMBODIMENTS
[0037] To make the objectives, technical solutions and advantages of the present application clearer, the technical solutions in the embodiments of the present application will be clearly and completely described below with reference to the accompanying drawings in the embodiments of the present application. Apparently, the described embodiments are only some of the embodiments of the present application, rather than all of them. All other embodiments obtained by those of ordinary skill in the art based on the embodiments of the present application without creative efforts shall fall within the scope of protection of the present application.
[0038] It should be understood that in the description of the present application, unless clearly required by the context, the words such as "including" and "comprising" throughout the specification should be interpreted as the meaning of including rather than exclusive or exhaustive; that is, the meaning of "including but not limited to".
[0039] It should also be understood that the terms "first", "second", etc. are only used for descriptive purposes and cannot be construed as indicating or implying relative importance. In addition, in the description of the present application, unless otherwise stated, the meaning of "a plurality of" is two or more.
[0040] It should be noted that the terms "S1", "S2", etc. are only used for the purpose of describing steps, and do not specifically refer to the order or sequence, nor are they used to limit this application. They are merely for the convenience of describing the method of this application and should not be construed as indicating the order of steps. Additionally, the technical solutions between various embodiments can be combined with each other, but it must be based on the ability of those of ordinary skill in the art to implement. When the combination of technical solutions results in contradictions or is unable to be implemented, it should be considered that such a combination of technical solutions does not exist and is not within the scope of protection required by this application.
[0041] The in-vehicle voice interaction method provided by this disclosure can be applied to Figure 1 the vehicle 100 shown in the figure. The vehicle 100 may include an in-vehicle terminal 120. The in-vehicle terminal 120 includes at least one memory and at least one processor. A computer program is stored in the at least one memory. When the computer program is executed by the at least one processor, it executes the in-vehicle voice interaction method according to the exemplary embodiments of this disclosure. Here, the in-vehicle terminal 120 does not have to be a single electronic device, but can also be any assembly of devices or circuits that can execute the above computer program alone or jointly.
[0042] In the in-vehicle terminal 120, the processor may include a central processing unit (CPU), a graphics processing unit (GPU), a programmable logic device, a dedicated processor system, a microcontroller, or a microprocessor. By way of example and not limitation, the processor may also include an analog processor, a digital processor, a microprocessor, a multi-core processor, a processor array, a network processor, etc.; in the in-vehicle terminal 120, the processor may run the computer program stored in the memory. The computer program may be divided into one or more modules / units (such as computer program 1, computer program 2,...). The one or more modules / units are stored in the memory and executed by the processor to complete this invention. The one or more modules / units may be a series of computer program instruction segments capable of performing specific functions, and these instruction segments are used to describe the execution process of the computer program in the terminal device. The memory may be integrated with the processor. For example, RAM or flash memory may be arranged within an integrated circuit microprocessor, etc. In addition, the memory may include independent devices, such as an external disk drive, a storage array, or other storage devices that can be used by any database system. The memory and the processor may be operatively coupled or may communicate with each other, for example, through an I / O port, a network connection, etc., so that the processor can read the files stored in the memory.
[0043] In addition, the in-vehicle terminal 120 may further include a display device (such as a liquid crystal display, etc.) and a user interaction interface (such as a keyboard, a mouse, a touch input device, etc.). All components of the in-vehicle terminal 120 may be connected to each other via a bus and / or a network.
[0044] Please refer to Figure 2 , Figure 2 A vehicle-mounted voice interaction method is provided for the first embodiment of the present invention. Taking the terminal in Figure 1 as an example, the method includes the following steps:
[0045] S1: Analyze the voice operation instruction to be processed to obtain the first registered word.
[0046] The voice operation instruction to be processed refers to the voice operation instruction issued by the vehicle-mounted user according to any operable event seen on the vehicle-mounted application display interface, replacing the manual operation according to actual needs. The vehicle-mounted application display interface refers to the operation interface contacted when adjusting the vehicle-mounted application, such as a vehicle-mounted liquid crystal display, etc.
[0047] Analyzing the voice operation instruction to be processed means identifying the voice operation instruction issued by the user and converting the voice signal into corresponding text information. In this embodiment, a third-party voice service SDK (Software Development Kit) can be used to analyze the voice operation instruction, such as iFlytek, or the native Cloud Speech API of the Android system - cloud voice application can be used to analyze the voice operation instruction.
[0048] The first registered word refers to the text information corresponding to the voice operation instruction.
[0049] In the embodiment of the present invention, the vehicle-mounted terminal receives the voice operation instruction issued by the user, and after identifying and converting the received voice operation instruction, generates the first registered word.
[0050] S2: According to the first registered word, determine the control to be operated corresponding to the first registered word, and the type of the control to be operated.
[0051] A control refers to any view control (the base class of all controls in the Android system) that can respond to a voice operation instruction in the visible and speakable service applied to the vehicle-mounted terminal.
[0052] The control type refers to the control operation type obtained based on the accessibility attribute (the basic attribute providing full accessibility for visually impaired people) label information of the view, such as clicking the control, swiping the control, etc.
[0053] The control to be operated refers to the control corresponding to the voice operation instruction.
[0054] In an embodiment of the present invention, a view control corresponding to a voice operation instruction is matched according to a parsed first registered word, and the type of the view control can be obtained according to the attribute tag information of the view control.
[0055] S3: Based on the type of the control to be operated, perform a simulated operation on the control to be operated on the first interface.
[0056] The first interface refers to the current in-vehicle display application interface that the user sees, such as music playback, air conditioning adjustment, Bluetooth phone, etc.
[0057] The simulated operation refers to performing a corresponding operation on the application display interface according to the type of the view control. For example, if the view control is a click control, a simulated click event is executed.
[0058] In an embodiment of the present invention, according to the type of the control corresponding to the voice operation instruction, a corresponding operation is performed on the in-vehicle application display interface.
[0059] S4: In response to the process of the simulated operation, generate a second interface to display the response effect corresponding to the voice operation instruction.
[0060] The second interface refers to the display interface formed by the continuous change of the current in-vehicle display application interface during the execution of the voice operation instruction by the control. For example, when the music playback volume is increased, the volume bar will change accordingly and be displayed in the in-vehicle display application interface.
[0061] The response effect refers to the visual effect of the in-vehicle application display interface responding to the voice operation instruction, such as the change process and result of the volume bar.
[0062] In an embodiment of the present invention, during the execution of the voice operation instruction by the control, the current in-vehicle display application interface continuously changes to form a second interface to display the execution effect of the voice operation instruction.
[0063] In an embodiment of the invention, the in-vehicle terminal receives a voice operation instruction issued by the user. After identifying and converting the received voice operation instruction, a first registered word is generated. A view control corresponding to the voice operation instruction is matched according to the parsed first registered word. The type of the view control can be obtained according to the attribute tag information of the view control. According to the type of the control corresponding to the voice operation instruction, a corresponding operation is performed on the in-vehicle application display interface. During the execution of the operation corresponding to the voice operation instruction by the control, the current in-vehicle display application interface continuously changes to form a second interface to display the execution effect of the operation corresponding to the voice operation instruction. This application uses voice instructions to replace manual operations to avoid the safety risks brought by manual operations during driving and improve the interactivity between the user and the in-vehicle terminal.
[0064] Please refer to Figure 3 , Figure 3 which is the step flowchart of a vehicle-mounted voice interaction method provided in the second embodiment of the present invention.
[0065] A vehicle-mounted voice interaction method provided by the present invention includes:
[0066] S21: Parse the voice operation instruction to be processed to obtain a first registered word.
[0067] It should be noted that this step specifically sends the user's voice operation instruction to the voice service, and uses the third-party voice service SDK of the voice engine module to parse the voice operation instruction, such as iFlytek, or uses the native CloudSpeechAPI-cloud voice application of the Android system to parse the voice operation instruction to obtain the first registered word.
[0068] S22: Determine the control to be operated corresponding to the first registered word and the type of the control to be operated according to the first registered word.
[0069] S23: Based on the type of the control to be operated, perform a simulated operation on the control to be operated on the first interface.
[0070] S24: In response to the process of the simulated operation, generate a second interface to display the response effect corresponding to the voice operation instruction.
[0071] In some embodiments, before parsing the voice operation instruction to be processed to obtain the first registered word, the method further includes generating a second registered word, specifically:
[0072] S11: Package the visible and speakable SDK (Software Development Kit), connect the SDK to the vehicle-mounted application, start the display interface of the vehicle-mounted application, and listen for changes in the display interface of the vehicle-mounted application through the ViewTreeObserver.OnGlobalLayoutListener (an observer for registering to listen to the view tree) of the native API (Application Programming Interface) of the Android system. When the controls on the display interface are drawn, this listening will be triggered and a second callback event will be issued, that is, the OnGlobalLayout event will be called back.
[0073] S12: In response to detecting the second callback event, poll and traverse the control tree to obtain the control with the set attribute label corresponding to the second callback event and the registered word corresponding to the control, where one control corresponds to one or more registered words.
[0074] It should be noted that when the visible-and-speakable SDK receives the second callback event, it starts to poll and traverse the control tree (view tree) of the current window of the Android application interface, and extracts the view controls with the accessibility attribute tags set in the control tree and the corresponding registered words. Among them, one control corresponds to one or more registered words, and the registered words are separated by ";". Exemplarily, accessibility = "registered word 1; registered word 2;...", and by default, the text information displayed by the view is used as the registered word.
[0075] S13: Send the registered word to the first service for registration. In response to the registration success result, generate the second registered word and save it.
[0076] It should be noted that the first service is a voice service. When the visible-and-speakable SDK determines that the voice service has been connected, it sends the registered words extracted in the above steps to the voice service, and uses the voice engine in the voice service to register the registered words. In response to detecting the registration success result, generate the second registered word, save the second registered word to the registry, generate a registered word list, and save it to the voice service.
[0077] S14: Associate the control with the second registered word corresponding to the control, and send it to the second service for saving.
[0078] It should be noted that the second service is a visible-and-speakable service. Generate a map set (a two-column set, each element having two data) in the form of using the generated second registered word as the key and its corresponding view control as the value, and save the map set to the visible-and-speakable service.
[0079] In some embodiments, according to the first registered word, determine the control to be operated corresponding to the first registered word, and the types of the controls to be operated include:
[0080] S220: Based on the first registered word, match and obtain the second registered word from the first service.
[0081] It should be noted that this step is specifically as follows: The voice service compares the received first registered word with the second registered words in the registered word list one by one. When the first registered word is the same as the target second registered word, it is determined that the comparison is successful, and the target second registered word can be extracted for subsequent obtaining of the control corresponding to the voice operation instruction.
[0082] S221: According to the second registered word, determine the control to be operated corresponding to the first registered word from the second service, and based on the attribute tags of the control to be operated, obtain the type of the control to be operated.
[0083] It should be noted that the visible-and-speakable service receives the target second registration word sent by the voice service, extracts the control corresponding to the target second registration word from the map set, and this control is the control to be operated. The control type can be determined according to the attribute label of the control to be operated. Exemplarily, if the view control type is a click control such as Text (text) / button (button), then when performing subsequent simulation operations, the performClick() of the view is executed to simulate a click event. If the view control type is a list, etc., then the onScroll() event is executed to slide the interface. If it is a progress bar type control, then the setProgress(*) event is executed to set the progress value.
[0084] In the above steps, after the voice operation instruction is parsed in the voice service, it can be compared to determine whether the voice operation instruction has been registered. If it has been registered, there will be a corresponding control, and there is no need to send it to the visible-and-speakable service for comparison operation, and then the comparison failure result is returned to the voice service, greatly improving the efficiency of in-vehicle voice interaction.
[0085] In some embodiments, when it is detected that the second registration word is not matched from the first service based on the first registration word, the method further includes:
[0086] S51: Prompt for the unmatched result, define the first registration word as the third registration word, and save the third registration word to the first service.
[0087] It should be noted that when the registration word corresponding to the voice operation instruction is not registered, the in-vehicle terminal can give corresponding prompts, such as voice prompt "I don't know what you're talking about", text prompt "I don't know what you're talking about" on the in-vehicle display interface, etc., and save the third registration word to the preset new word database in the first service for subsequent dynamic registration of corresponding words.
[0088] S52: In response to detecting that the number of occurrences of the third registration word is greater than the first preset value, calculate the similarity between the third registration word and the target second registration word in the first service.
[0089] It should be noted that the first preset value can be set according to actual needs, such as 3 times or 5 times, etc.; when the number of occurrences of the third registration word is greater than this number, the third registration word is calculated for similarity one by one with the data in the registration word table. The similarity calculation methods adopted in this embodiment are all conventional means in the art and will not be elaborated here.
[0090] S53: Dynamically register the third registration word according to the similarity calculation result.
[0091] It should be noted that this step is specifically as follows:
[0092] S531: In response to detecting that the similarity calculation result is greater than the second preset value and no first callback event is monitored within the preset time, associate the third registered word with the target second registered word and save it to the first service, and associate the third registered word with the control corresponding to the target second registered word and save it to the second service.
[0093] Among them, the second preset value and the preset time can be set according to actual needs, such as 90%, 5s, etc. The first callback event refers to the first callback event generated by the callback when the view control of the in-vehicle application display interface is drawn. Exemplarily, when it is detected that the similarity calculation result is greater than 90% and within 5s, the user does not perform operations such as swiping or clicking on the in-vehicle application display interface, then associate the third registered word with the target second registered word in the registration word table and save it to the voice service, and associate the third registered word with the control corresponding to the target second registered word in the map set and save it to the visible and speakable service.
[0094] S532: In response to detecting that the similarity calculation result is greater than the second preset value and the first callback event is monitored within the preset time, the method further includes:
[0095] (1) Obtain the fourth registered word corresponding to the first callback event in the second service.
[0096] It should be noted that monitoring the first callback event within the preset time means monitoring that the user manually operates on the in-vehicle application interface within the preset time, such as clicking the volume button, etc. Then extract the registered word corresponding to the manually operated control in the visible and speakable service and define it as the fourth registered word. In this embodiment, the second registered word and the fourth registered word are only for distinction, and they may actually be the same registered word.
[0097] (2) Calculate the similarity between the target second registered word and the fourth registered word.
[0098] In some embodiments, the similarity can also be calculated between the fourth registered word and the third registered word.
[0099] (3) When it is detected that the similarity calculation result is greater than a third preset value, associate the third registered word with the control corresponding to the fourth registered word, and save it in the second service. Also, associate the third registered word with the target second registered word and save it in the first service. The third preset value can be set according to actual needs, such as 93%. Exemplarily, when it is detected that the similarity calculation result is greater than 93%, associate the third registered word with the control corresponding to the fourth registered word and save it in the visible-and-speakable service. Associate the second registered word corresponding to the fourth registered word in the registered word list with the third registered word and save it in the voice service.
[0100] S533: When it is detected that the similarity calculation result is less than or equal to a second preset value, remove the third registered word from the new word database;
[0101] When the similarity calculation result between the target second registered word and the fourth registered word is less than or equal to the third preset value, or the similarity calculation result between the third registered word and the fourth registered word is less than or equal to the third preset value, continue to save the third registered word and wait for the next match.
[0102] In the above steps, when the user's real-time voice operation instruction is not registered, a series of methods such as similarity calculation are used to achieve the dynamic registration of new voice instructions, improving the user experience. If, after the voice instruction similarity match is successful, the user manually operates the screen control within a short period of time, at this time, in order to further identify the user's intention, a secondary similarity calculation is performed to avoid unnecessary errors and further improve the accuracy of the dynamic registration of new voice instructions.
[0103] In some embodiments, when generating a second interface to display the response effect corresponding to the voice operation instruction in response to the simulated operation, it further includes:
[0104] Perform a simulated operation on the control to be operated based on the type of the control to be operated to control the vehicle to execute the functional requirements corresponding to the voice operation instruction.
[0105] Specifically, when performing a simulated operation on the control, the first interface displayed on the current in-vehicle application display interface will change along with the simulated operation process, thus presenting the changed interface, that is, the second interface. In addition to the interface display changing, the corresponding functional hardware of the in-vehicle terminal will also execute the functional requirements corresponding to the voice operation instruction. Exemplarily, if the voice operation instruction is "turn up the volume of the music", the display interface will show the process of the volume bar sliding. At the same time, the corresponding in-vehicle music player will also play a louder music volume to respond to the voice operation instruction.
[0106] In an embodiment of the invention, the in-vehicle terminal receives a voice operation instruction issued by a user. After identifying and converting the received voice operation instruction, a first registered word is generated. According to the first registered word obtained by parsing, the view control corresponding to the voice operation instruction is matched. According to the attribute tag information of the view control, the type of the view control can be obtained. According to the control type corresponding to the voice operation instruction, corresponding operations are performed on the in-vehicle application display interface. During the process of the control performing the operations corresponding to the voice operation instruction, the current in-vehicle display application interface continuously changes to form a second interface to display the execution effect of the operations corresponding to the voice operation instruction. In this application, voice instructions are used to replace manual operations to avoid the safety risks brought by manual operations during driving and improve the interactivity between the user and the in-vehicle terminal.
[0107] It should be understood that although Figure 2-3 the steps in the flowchart of Figure 2-3 are shown in sequence according to the arrows, these steps are not necessarily executed in the order indicated by the arrows. Unless otherwise clearly stated in this article, the execution of these steps has no strict order limit, and these steps can be executed in other orders. Moreover,
[0108] Please refer to Figure 4 , Figure 4 which is a vehicle-mounted voice interaction device provided in the second embodiment of the present invention, including: a voice service module, a control determination module, a simulation operation module, and a display module, where:
[0109] The voice service module is used to parse the voice operation instruction to be processed to obtain a first registered word;
[0110] The control determination module is used to determine the control to be operated corresponding to the first registered word and the type of the control to be operated according to the first registered word;
[0111] The simulation operation module is used to perform a simulation operation on the control to be operated on the first interface based on the type of the control to be operated;
[0112] The display module is used to generate a second interface during the process of responding to the simulation operation to display the response effect corresponding to the voice operation instruction.
[0113] As a preferred embodiment, in the embodiment of the present invention, the control determination module is specifically used for:
[0114] Match and obtain a second registered term from a first service based on the first registered term;
[0115] Determine a control to be operated corresponding to the first registered term from a second service according to the second registered term, and obtain the type of the control to be operated based on the attribute label of the control to be operated.
[0116] As a preferred implementation manner, in the embodiment of the present invention, the voice service module is further specifically configured to:
[0117] In response to detecting that the second registered term is not matched from the first service based on the first registered term, prompt for the unsuccessful matching result;
[0118] Define the first registered term as a third registered term, and save the third registered term to the first service.
[0119] As a preferred implementation manner, in the embodiment of the present invention, the device further includes a similarity calculation module, and the similarity calculation module is specifically configured to:
[0120] In response to detecting that the number of occurrences of the third registered term is greater than a first preset value, calculate the similarity between the third registered term and a target second registered term in the first service;
[0121] In response to detecting that the similarity calculation result is greater than a second preset value and the first callback event is not monitored within a preset time, associate the third registered term with the target second registered term and save it to the first service, and associate the third registered term with the control corresponding to the target second registered term and save it to the second service.
[0122] As a preferred implementation manner, in the embodiment of the present invention, the similarity calculation module is further specifically configured to:
[0123] In response to detecting that the similarity calculation result is greater than a second preset value and the first callback event is monitored within a preset time, obtain a fourth registered term corresponding to the first callback event in the second service;
[0124] Calculate the similarity between the target second registered term and the fourth registered term;
[0125] In response to detecting that the similarity calculation result is greater than a third preset value, associate the third registered term with the control corresponding to the fourth registered term and save it to the second service, and associate the third registered term with the target second registered term and save it to the first service.
[0126] As a preferred implementation manner, in the embodiment of the present invention, the device further includes a registration word generation module, and the registration word generation module is specifically configured to:
[0127] When detecting a second callback event, poll and traverse the control tree to obtain the control with the set attribute label corresponding to the second callback event and the registration word corresponding to the control, where one control corresponds to one or more registration words;
[0128] Send the registration word to the first service for registration, and in response to the registration success result, generate the second registration word and save it;
[0129] Associate the control and the second registration word corresponding to the control, and send them to the second service for saving.
[0130] As a preferred implementation manner, in the embodiment of the present invention, the device further includes a function response module, and the function response module is specifically configured to:
[0131] Perform a simulation operation on the to-be-operated control based on the type of the to-be-operated control to control the vehicle to execute the function requirements corresponding to the voice operation instruction.
[0132] For the specific limitations of the in-vehicle voice interaction device, reference can be made to the limitations on the in-vehicle voice interaction method in the foregoing text, which will not be elaborated here. Each module in the above in-vehicle voice interaction device can be implemented in whole or in part by software, hardware, and their combination. The above modules can be embedded in or independent of the processor in the computer device in the form of hardware, or stored in the memory of the computer device in the form of software, so as to facilitate the processor to call and execute the operations corresponding to the above modules.
[0133] Please refer to Figure 5 , Figure 5 A computer device provided in Embodiment 3 of the present invention. This computer device may be a terminal, and its internal structure diagram may be as shown in Figure 5As shown in the figure. The computer device includes a processor, a memory, a network interface, a display screen, and an input device connected through a system bus. Among them, the processor of the computer device is used to provide computing and control capabilities. The memory of the computer device includes a non-volatile storage medium and an internal memory. The non-volatile storage medium stores an operating system and a computer program. The internal memory provides an environment for the operation of the operating system and the computer program in the non-volatile storage medium. The network interface of the computer device is used to communicate with an external terminal through a network connection. When the computer program is executed by the processor, it implements a vehicle-mounted voice interaction method. The display screen of the computer device can be a liquid crystal display screen or an electronic ink display screen. The input device of the computer device can be a touch layer covering the display screen, or a button, a trackball, or a touchpad provided on the housing of the computer device, or an external keyboard, touchpad, or mouse, etc.
[0134] Those skilled in the art can understand that Figure 5 the structure shown in the figure is only a block diagram of some structures related to the solution of the present application, and does not constitute a limitation on the computer device to which the solution of the present application is applied. The specific computer device may include more or fewer components than those shown in the figure, or combine some components, or have different component arrangements.
[0135] In one embodiment, a computer device is provided, including a memory, a processor, and a computer program stored on the memory and executable on the processor. When the processor executes the computer program, the following steps are implemented:
[0136] Parse the voice operation instruction to be processed to obtain a first registered word;
[0137] According to the first registered word, determine the control to be operated corresponding to the first registered word and the type of the control to be operated;
[0138] Based on the type of the control to be operated, perform a simulated operation on the control to be operated on the first interface;
[0139] In response to the process of the simulated operation, generate a second interface to display the response effect corresponding to the voice operation instruction.
[0140] In one embodiment, when the processor executes the computer program, the following steps are also implemented:
[0141] Based on the first registered word, match and obtain a second registered word from a first service;
[0142] According to the second registered word, determine the control to be operated corresponding to the first registered word from a second service, and based on the attribute label of the control to be operated, obtain the type of the control to be operated.
[0143] In one embodiment, when the processor executes the computer program, the following steps are further implemented:
[0144] In response to detecting that the second registered word is not matched in the first service based on the first registered word, a prompt for the unmatched result is given, the first registered word is defined as the third registered word, and the third registered word is saved in the first service.
[0145] In one embodiment, when the processor executes the computer program, the following steps are further implemented:
[0146] In response to detecting that the occurrence times of the third registered word are greater than a first preset value, a similarity calculation is performed between the third registered word and a target second registered word in the first service;
[0147] In response to detecting that the similarity calculation result is greater than a second preset value and the first callback event is not monitored within a preset time, the third registered word is associated with the target second registered word and saved in the first service, and the third registered word is associated with the control corresponding to the target second registered word and saved in the second service.
[0148] In one embodiment, when the processor executes the computer program, the following steps are further implemented:
[0149] Obtain a fourth registered word corresponding to the first callback event in the second service;
[0150] Perform a similarity calculation between the target second registered word and the fourth registered word;
[0151] In response to detecting that the similarity calculation result is greater than a third preset value, the third registered word is associated with the control corresponding to the fourth registered word and saved in the second service, and the third registered word is associated with the target second registered word and saved in the first service.
[0152] In one embodiment, when the processor executes the computer program, the following steps are further implemented:
[0153] In response to detecting a second callback event, poll and traverse the control tree to obtain the control with the set attribute label corresponding to the second callback event, and the registered word corresponding to the control, where one control corresponds to one or more registered words;
[0154] Send the registered word to the first service for registration, and in response to the registration success result, generate and save the second registered word;
[0155] Associate the control and the corresponding second registered word of the control, and send them to the second service for storage.
[0156] In one embodiment, when the processor executes the computer program, the following steps are further implemented:
[0157] Perform a simulated operation on the to-be-operated control based on the type of the to-be-operated control to control the vehicle to execute the functional requirements corresponding to the voice operation instruction.
[0158] A computer-readable storage medium provided in the fourth embodiment of the present invention, on which a computer program is stored. When the computer program is executed by a processor, the following steps are implemented:
[0159] Parse the to-be-processed voice operation instruction to obtain a first registered word;
[0160] According to the first registered word, determine the to-be-operated control corresponding to the first registered word and the type of the to-be-operated control;
[0161] Based on the type of the to-be-operated control, perform a simulated operation on the to-be-operated control on the first interface;
[0162] In response to the process of the simulated operation, generate a second interface to display the response effect corresponding to the voice operation instruction.
[0163] In one embodiment, when the computer program is executed by a processor, the following steps are further implemented:
[0164] Based on the first registered word, match and obtain a second registered word from the first service;
[0165] According to the second registered word, determine the to-be-operated control corresponding to the first registered word from the second service, and based on the attribute label of the to-be-operated control, obtain the type of the to-be-operated control.
[0166] In one embodiment, when the computer program is executed by a processor, the following steps are further implemented:
[0167] In response to detecting that when the second registered word is not matched from the first service based on the first registered word, prompt the unmatched result, define the first registered word as a third registered word, and save the third registered word to the first service.
[0168] In one embodiment, when the computer program is executed by a processor, the following steps are further implemented:
[0169] In response to detecting that the number of occurrences of the third registered word is greater than a first preset value, calculate the similarity between the third registered word and the target second registered word in the first service;
[0170] In response to detecting that the similarity calculation result is greater than a second preset value and no first callback event is monitored within a preset time, associate the third registered word with the target second registered word and save it to the first service, and associate the third registered word with the control corresponding to the target second registered word and save it to the second service.
[0171] In one embodiment, when the computer program is executed by a processor, the following steps are further implemented:
[0172] Obtain the fourth registered word corresponding to the first callback event in the second service;
[0173] Calculate the similarity between the target second registered word and the fourth registered word;
[0174] In response to detecting that the similarity calculation result is greater than a third preset value, associate the third registered word with the control corresponding to the fourth registered word and save it to the second service, and associate the third registered word with the target second registered word and save it to the first service.
[0175] In one embodiment, when the computer program is executed by a processor, the following steps are further implemented:
[0176] In response to detecting a second callback event, poll and traverse the control tree to obtain the control with the set property label corresponding to the second callback event and the registered word corresponding to the control, where one control corresponds to one or more registered words;
[0177] Send the registered word to the first service for registration, and in response to the registration success result, generate and save the second registered word;
[0178] Associate the control and the second registered word corresponding to the control and send them to the second service for saving.
[0179] In one embodiment, when the computer program is executed by a processor, the following steps are further implemented:
[0180] Perform a simulated operation on the to-be-operated control based on the type of the to-be-operated control to control the vehicle to execute the functional requirements corresponding to the voice operation instruction.
[0181] Those of ordinary skill in the art can understand that all or part of the processes in the methods of the above embodiments can be completed by instructing relevant hardware through a computer program. The computer program can be stored in a non-volatile computer-readable storage medium. When the computer program is executed, it can include the processes of the embodiments of the above methods. Among them, any reference to a memory, storage, database, or other medium used in the embodiments provided in the present application can include non-volatile and / or volatile memories. Non-volatile memories can include read-only memory (ROM), programmable ROM (PROM), electrically programmable ROM (EPROM), electrically erasable programmable ROM (EEPROM), or flash memory. Volatile memories can include random access memory (RAM) or external cache memory. By way of illustration and not limitation, RAM is available in various forms, such as static RAM (SRAM), dynamic RAM (DRAM), synchronous DRAM (SDRAM), double data rate SDRAM (DDR SDRAM), enhanced SDRAM (ESDRAM), synchronous link (Synchlink) DRAM (SLDRAM), Rambus direct RAM (RDRAM), direct memory bus dynamic RAM (DRDRAM), and Rambus dynamic RAM (RDRAM), etc.
[0182] The technical features of the above embodiments can be combined arbitrarily. For the sake of brevity of description, not all possible combinations of the technical features in the above embodiments are described. However, as long as there is no contradiction in the combination of these technical features, it should be considered as the scope recorded in this specification.
[0183] The above-described embodiments merely represent several implementation manners of the present application. The description is relatively specific and detailed, but it should not be construed as a limitation on the scope of the invention patent. It should be noted that for those of ordinary skill in the art, without departing from the concept of the present application, several modifications and improvements can still be made, and these all belong to the protection scope of the present application.
Claims
1. A vehicle-mounted voice interaction method, characterized in that, The method includes: Parsing the voice operation instruction to be processed to obtain a first registered word; Determining, according to the first registered word, the control to be operated corresponding to the first registered word and the type of the control to be operated; Based on the type of the control to be operated, performing a simulated operation on the control to be operated on a first interface; In response to the process of the simulated operation, generating a second interface to display the response effect corresponding to the voice operation instruction; Wherein, the determining, according to the first registered word, the control to be operated corresponding to the first registered word includes: Based on the first registered word, matching and obtaining a second registered word from a first service; According to the second registered word, determining, from a second service, the control to be operated corresponding to the first registered word, and obtaining the type of the control to be operated based on the attribute label of the control to be operated; In response to detecting that the second registered word is not matched from the first service based on the first registered word, the method further includes: Prompting the unmatched result, defining the first registered word as a third registered word, and saving the third registered word to the first service; In response to detecting that the occurrence times of the third registered word are greater than a first preset value, calculating the similarity between the third registered word and a target second registered word in the first service; In response to detecting that the similarity calculation result is greater than a second preset value and no first callback event is monitored within a preset time, associating the third registered word with the target second registered word and saving it to the first service, and associating the third registered word with the control corresponding to the target second registered word and saving it to the second service.
2. The in-vehicle voice interaction method according to claim 1, wherein In response to detecting that the similarity calculation result is greater than a second preset value and a first callback event is monitored within a preset time, the method further includes: Obtaining a fourth registered word corresponding to the first callback event in the second service; Calculating the similarity between the target second registered word and the fourth registered word; In response to detecting that the similarity calculation result is greater than a third preset value, associating the third registered word with the control corresponding to the fourth registered word and saving it to the second service, and associating the third registered word with the target second registered word and saving it to the first service.
3. The in-vehicle voice interaction method according to claim 1, wherein The method for generating the second registered word includes: In response to detecting a second callback event, polling and traversing the control tree to obtain the control with the set attribute label corresponding to the second callback event and the registered word corresponding to the control, where one control corresponds to one or more registered words; Sending the registered word to the first service for registration, and in response to the registration success result, generating and saving the second registered word; Associating the control and the second registered word corresponding to the control, and sending it to the second service for saving.
4. The in-vehicle voice interaction method according to any one of claims 1-3, characterized in that, The method further includes: Based on the type of the control to be operated, performing a simulated operation on the control to be operated to control the vehicle to execute the functional requirement corresponding to the voice operation instruction.
5. An in-vehicle voice interaction device for implementing the in-vehicle voice interaction method according to any one of claims 1-4, characterized in that, The device includes: A voice service module for parsing a voice operation instruction to be processed to obtain a first registered word; A control determination module for determining a control to be operated corresponding to the first registered word and the type of the control to be operated according to the first registered word; A simulation operation module for performing a simulation operation on the control to be operated on a first interface based on the type of the control to be operated; A display module for generating a second interface in response to the process of the simulation operation to display a response effect corresponding to the voice operation instruction.
6. A computer device, comprising a memory, a processor, and a computer program stored on the memory and executable on the processor, characterized in that, When the processor executes the computer program, the steps of the method according to any one of claims 1 to 4 are implemented.
7. A computer-readable storage medium having a computer program stored thereon, characterized in that, When the computer program is executed by the processor, the steps of the method according to any one of claims 1 to 4 are implemented.
Citation Information
Patent Citations
Method and device for controlling intelligent terminal in analog mode and intelligent terminal
CN110795175A
Equipment control method and device, equipment and storage medium
CN115527531A