Voice interaction method, electronic equipment and storage medium
By receiving and analyzing the user's voice information, determining the identity reference information in the voice information and executing relevant instruction operations, the problem of not being able to identify the speaker's identity and executing instructions containing identity relationships in the prior art is solved, and a more efficient voice interaction system is realized.
Patent Information
- Application Number
- CN202311671433.5
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2023-12-07
- Publication Date
- 2025-06-10
AI Technical Summary
Existing smart products cannot identify the speaker's identity and cannot complete command operations that include social or family identity relationships.
By receiving the voice information sent by the user, the first identity information corresponding to the identity reference information in the voice information is determined, and the instruction operations related to the identity information are performed. The specific method includes using a preset relationship tree to determine the first identity information corresponding to the identity reference information, determining the second identity information of the user through the voiceprint information of the voice information, and confirming the first identity information in the preset relationship tree.
Responsive to voice information instructions containing identity reference information is realized, corresponding operations can be performed, and the identity identification and command execution capabilities of the voice interaction system are enhanced.
Smart Images

Figure CN120126464A_ABST
Abstract
Description
Technical Field
[0001] The present disclosure relates to the technical field of voice command interaction, and in particular, to a voice interaction method, an electronic device, and a storage medium. Background Art
[0002] With the development of artificial intelligence technology, voice interaction functions have gradually entered various fields of people's lives. People can use voice interaction functions to control intelligent electronic devices by voice, such as display devices, air conditioners, and washing machines. People can use voice interaction functions to perform a series of operations such as watching videos, listening to music, checking the weather, and device control.
[0003] In the process of implementing the voice interaction function, generally, a voice recognition module recognizes the voice command input by the user as text, and then a semantic analysis module analyzes the text in terms of morphology, syntax, and semantics to analyze the user's needs. Finally, the control end performs corresponding operations according to the user's needs. Summary of the Invention
[0004] To overcome the problems existing in the related art, the present disclosure provides a voice interaction method, an electronic device, and a storage medium.
[0005] According to a first aspect of an embodiment of the present disclosure, a voice interaction method is provided. The voice interaction method includes: receiving voice information sent by a user; determining first identity information corresponding to identity reference information included in the voice information; and performing an instruction operation related to the first identity information.
[0006] In some embodiments, determining the first identity information corresponding to the identity reference information included in the voice information includes: determining the first identity information corresponding to the identity reference information according to a preset relationship tree.
[0007] In some embodiments, the voice interaction method further includes: determining that the voice information includes the identity reference information according to the voice information sent by the user.
[0008] In some embodiments, determining the first identity information corresponding to the identity reference information according to the preset relationship tree includes: determining second identity information of the user; and confirming the first identity information in the preset relationship tree according to the second identity information and the identity reference information, where the second identity information of the user is determined by voiceprint information of the voice information.
[0009] In some embodiments, the preset relationship tree is determined in advance in the following manner: obtaining identity information of multiple users; and obtaining an identity relationship tree according to the relationship between the multiple identity information, where the identity relationship tree is the preset relationship tree.
[0010] In some embodiments, the obtaining of the identity information of multiple users includes: receiving first voice information sent by a first user; the identity information and corresponding voiceprint information included in the first voice information; receiving second voice information sent by a second user; determining the identity information and corresponding voiceprint information included in the second voice information.
[0011] In some embodiments, the identity information of the user is determined by the voiceprint information of the voice information.
[0012] According to a second aspect of the embodiments of the present disclosure, there is provided an electronic device for performing the voice interaction method described in the first aspect. The electronic device includes: a receiving unit for receiving voice information sent by a user; a determining unit for determining first identity information corresponding to the identity reference information included in the voice information; and an operating unit for performing an instruction operation related to the first identity information.
[0013] In some embodiments, the determining unit determines the first identity information corresponding to the identity reference information included in the voice information by the following method: determining the first identity information corresponding to the identity reference information according to a preset relationship tree.
[0014] In some embodiments, the determining unit is further configured to determine that the voice information includes the identity reference information according to the voice information sent by the received user.
[0015] In some embodiments, the determining unit determines the first identity information corresponding to the identity reference information according to the preset relationship tree by the following method: determining a second identity information of the user; and confirming the first identity information in the preset relationship tree according to the second identity information and the identity reference information, wherein the second identity information of the user is determined by the voiceprint information of the voice information.
[0016] In some embodiments, the operating unit pre-determines the preset relationship tree by the following method: obtaining the identity information of multiple users; and obtaining an identity relationship tree according to the relationship between the multiple identity information, and the identity relationship tree is the preset relationship tree.
[0017] In some embodiments, the operating unit obtains the identity information of multiple users by the following method: receiving first voice information sent by a first user; the identity information and corresponding voiceprint information included in the first voice information; receiving second voice information sent by a second user; determining the identity information and corresponding voiceprint information included in the second voice information.
[0018] In some embodiments, the operation unit determines the identity information of the user through the voiceprint information of the voice message.
[0019] According to a third aspect of the embodiments of the present disclosure, there is provided an electronic device, which includes: a processor; a memory for storing instructions executable by the processor; wherein, the processor is configured to: execute the voice interaction method described in the first aspect.
[0020] According to a fourth aspect of the embodiments of the present disclosure, there is provided a storage medium, when the instructions in the storage medium are executed by a processor of a mobile terminal, enabling the mobile terminal to execute a voice interaction method, the voice interaction method including: the voice interaction method described in any item of the first aspect.
[0021] The technical solutions provided by the embodiments of the present disclosure may include the following beneficial effects: By confirming the first identity information corresponding to the identity reference information in the voice message, the present disclosure can respond to the voice message instruction containing the identity reference information and perform corresponding operations.
[0022] It should be understood that the above general description and the following detailed description are only exemplary and explanatory, and cannot limit the present disclosure. BRIEF DESCRIPTION OF THE DRAWINGS
[0023] The accompanying drawings herein are incorporated into the specification and form a part of the specification, showing embodiments consistent with the present disclosure, and are used together with the specification to explain the principles of the present disclosure.
[0024] Figure 1 is a schematic flowchart of a voice interaction method shown according to an exemplary embodiment.
[0025] Figure 2 is a schematic flowchart of a voice interaction method shown according to an exemplary embodiment.
[0026] Figure 3 is a schematic flowchart of a voice interaction method shown according to an exemplary embodiment.
[0027] Figure 4 is a schematic flowchart of a voice interaction method shown according to an exemplary embodiment.
[0028] Figure 5 is a schematic flowchart of a voice interaction method shown according to an exemplary embodiment.
[0029] Figure 6 is a schematic flowchart of a voice interaction method shown according to an illustrative embodiment.
[0030] Figure 7It is a block diagram of an electronic device shown according to an exemplary embodiment.
[0031] Figure 8 It is a block diagram of a device shown according to an exemplary embodiment.
[0032] Figure 9 It is a block diagram of a device shown according to an exemplary embodiment. Detailed implementation manners
[0033] Here, the exemplary embodiments will be described in detail, and the examples are shown in the drawings. When the following description refers to the drawings, unless otherwise indicated, the same numbers in different drawings represent the same or similar elements. The embodiments described in the following exemplary embodiments do not represent all embodiments consistent with the present disclosure. On the contrary, they are merely examples of devices and methods consistent with some aspects of the present disclosure as detailed in the appended claims.
[0034] In the related art, current smart products cannot identify the identity of the speaker and cannot complete instruction operations including social or family identity relationships.
[0035] To solve the above technical problems, according to an embodiment of the present disclosure, a voice interaction method is provided. The voice interaction method includes: receiving voice information sent by a user; determining first identity information corresponding to the identity reference information included in the voice information; and performing an instruction operation related to the first identity information.
[0036] The present disclosure can respond to a voice information instruction including identity reference information and perform corresponding operations by confirming the first identity information corresponding to the identity reference information in the voice information.
[0037] It can be understood that the voice interaction method involved in the present disclosure can be applied to a voice interaction system or any of the following terminals.
[0038] It can be understood that the terminal involved in the present disclosure, which can also be referred to as a terminal device, user equipment (UE), mobile station (MS), mobile terminal (MT), etc., is a device that provides voice and / or data connectivity to users. For example, the terminal can be a handheld device with wireless connection function, a vehicle-mounted device, etc. Currently, some examples of terminals are: smart phones (Mobile Phone), pocket personal computers (PPC), palm computers, personal digital assistants (PDA), laptop computers, tablet computers, wearable devices, or vehicle-mounted devices, etc. In addition, when it is a vehicle-to-everything (V2X) communication system, the terminal device can also be a vehicle-mounted device. It should be understood that the specific technologies and specific device forms adopted by the terminal in the embodiments of the present disclosure are not limited.
[0039] Figure 1 It is a flowchart showing a voice interaction method according to an exemplary embodiment.
[0040] In some embodiments, as Figure 1 shown, the voice interaction method may include the following steps:
[0041] S10: Receive the voice information sent by the user;
[0042] S20: Determine the first identity information corresponding to the identity reference information included in the voice information;
[0043] S30: Execute the instruction operation related to the first identity information.
[0044] Among them, the identity reference information can be the reference object in the voice information. The first identity information can be the identity of the specific object corresponding to the identity reference information.
[0045] In the embodiments of the present disclosure, the voice interaction system can receive the voice information sent by the user, and determine the first identity information corresponding to the identity reference information, so as to execute the voice information instruction operation related to the first identity information.
[0046] Exemplarily, user A can send a voice message of "call my son", and the voice interaction system can determine that the identity corresponding to the word "son" included in the voice information is user B, and then can execute the instruction operation related to user B and call user B.
[0047] The present disclosure can respond to a voice message instruction containing an identity reference information and perform corresponding operations by confirming the first identity information corresponding to the identity reference information in the voice message.
[0048] Figure 2 It is a schematic flowchart of a voice interaction method shown according to an exemplary embodiment.
[0049] In some embodiments, as Figure 2 shown, the voice interaction method may include the following steps:
[0050] S10: Receive the voice message sent by the user;
[0051] S21: Determine the first identity information corresponding to the identity reference information according to the preset relationship tree.
[0052] S30: Execute the instruction operation related to the first identity information.
[0053] Among them, the identity reference information may be the reference object in the voice message. The first identity information may be the identity of the specific object corresponding to the identity reference information.
[0054] In the embodiments of the present disclosure, the voice interaction system can receive the voice message sent by the user and determine the first identity information corresponding to the identity reference information according to the preset relationship tree, so as to execute the voice message instruction operation related to the first identity information.
[0055] The preset relationship tree may include the identity information of multiple users and the relationships between multiple users. Exemplarily, the relationship may be a social relationship or a family relationship. According to the preset relationship tree, the first identity relationship corresponding to the identity reference information can be determined.
[0056] Exemplarily, user A may send a voice message of "call my son", and the voice interaction system can determine that the identity corresponding to the word "son" included in the voice message in the preset relationship tree is user B, and then can execute the instruction operation related to user B, that is, call user B.
[0057] The present disclosure can respond to a voice message instruction containing an identity reference information and perform corresponding operations by confirming the first identity information corresponding to the identity reference information in the voice message.
[0058] Figure 3 It is a schematic flowchart of a voice interaction method shown according to an exemplary embodiment.
[0059] In some embodiments, as Figure 3 shown, the voice interaction method may include the following steps:
[0060] S10: Receive the voice message sent by the user;
[0061] S11: Determine that the voice message includes identity reference information according to the received voice message sent by the user.
[0062] S20: Determine the first identity information corresponding to the identity reference information included in the voice message;
[0063] S30: Execute the instruction operation related to the first identity information.
[0064] Among them, the identity reference information may be a reference object in the voice message. The first identity information may be the identity of the specific object corresponding to the identity reference information.
[0065] In the embodiments of the present disclosure, the voice interaction system may receive the voice message sent by the user. The voice interaction system may determine that the voice message includes identity reference information according to the received voice message. When it is determined that the voice message includes identity reference information, the voice interaction system determines the first identity information corresponding to the identity reference information, so that the instruction operation related to the first identity information can be executed.
[0066] Exemplarily, user A may send a voice message of "call my son", and the voice interaction system may determine that the word "son" included in the voice message is identity reference information. After confirming that the identity corresponding to the identity reference information is user B, the instruction operation related to user B can be further executed to call user B.
[0067] The present disclosure can respond to the voice message instruction including the identity reference information and perform corresponding operations by confirming the first identity information corresponding to the identity reference information in the voice message.
[0068] Figure 4 It is a schematic flowchart of a voice interaction method shown according to an exemplary embodiment.
[0069] In some embodiments, as Figure 4 shown, the voice interaction method may include the following steps:
[0070] S10: Receive the voice message sent by the user;
[0071] S211: Determine the second identity information of the user;
[0072] S210: Confirm the first identity information in the preset relationship tree according to the second identity information and the identity reference information,
[0073] S30: Execute the instruction operation related to the first identity information.
[0074] Among them, the identity reference information can be the reference object in the voice information. The first identity information can be the identity of the specific object corresponding to the identity reference information. The second identity information can be the identity of the user who sends the voice information.
[0075] In the embodiments of the present disclosure, the voice interaction system can receive the voice information sent by the user, and the voice interaction system can determine the identity of the user who sends the voice information. The voice interaction system can confirm the position of the second identity information in the preset relationship tree, and determine the position relevance in the preset relationship tree according to the relationship corresponding to the identity reference information. Furthermore, the corresponding first identity information can be obtained based on the second identity information in the preset relationship tree, so that the voice information instruction operation related to the first identity information can be executed.
[0076] Exemplarily, user A can send a voice message of "call my son", and the voice interaction system can determine that the user who sends the voice information is user A, and confirm the position of the user in the preset relationship tree.
[0077] Then, according to the identity reference information "son" included in the voice information, in the preset relationship tree, find the user whose identity is the son of user A, confirm that the identity corresponding to the identity reference information is user B, and then the instruction operation related to user B can be executed, and a call can be made to user B.
[0078] Among them, the voice interaction system can determine the second identity information of the user through the voiceprint information of the voice information. A voiceprint is the general term for the voice features contained in the voice that can represent and identify the speaker, and the voice model established based on these features (parameters). Voiceprint recognition is performed according to the information contained in the voiceprint, so that the voice interaction system can identify the speaker corresponding to the voice segment according to the voiceprint features of the voice to be recognized.
[0079] Exemplarily, when user A sends a voice message, the voice interaction system can identify that it is user A who sends the voice message according to the voiceprint features of the voice message, and then obtain the second identity information of user A.
[0080] The present disclosure obtains the first identity information corresponding to the identity reference information in the preset relationship tree by confirming the second identity information in the voice information, so that the voice information instruction containing the identity reference information can be responded to and the corresponding operation can be performed.
[0081] Figure 5 It is a flowchart of a voice interaction method shown according to an exemplary embodiment.
[0082] In some embodiments, as Figure 5 shown, the preset relationship tree can be determined in advance by the following steps:
[0083] S41: Obtain the identity information of multiple users;
[0084] S42: Obtain an identity relationship tree based on the relationships between multiple pieces of identity information, where the identity relationship tree is a preset relationship tree.
[0085] By determining the identity information of multiple users and the relationships between multiple users, the identity information of multiple users and the relationships between multiple users can form an identity relationship tree, so that the voice interaction system can obtain the first identity information corresponding to the identity reference information in the preset relationship tree.
[0086] Figure 6 It is a schematic flowchart of a voice interaction method shown according to an exemplary embodiment.
[0087] In some embodiments, as Figure 6 shown, the preset relationship tree can be determined in advance by the following steps:
[0088] S411: Receive the first voice message sent by the first user;
[0089] S412: The identity information and the corresponding voiceprint information included in the first voice message;
[0090] S413: Receive the second voice message sent by the second user;
[0091] S414: Determine the identity information and the corresponding voiceprint information included in the second voice message;
[0092] S42: Obtain an identity relationship tree based on the relationships between multiple pieces of identity information, where the identity relationship tree is a preset relationship tree.
[0093] The voice interaction system can receive the first voice message sent by the first user. The first voice message can be a voice message including the identity information of the first user and the relationships with other users. The voice interaction system can match the first user with the voiceprint information according to the voiceprint information of the first voice message, so that the voice interaction system can identify the first user according to the voiceprint information.
[0094] The voice interaction system can receive the first voice message sent by the second user. The second voice message can be a voice message including the identity information of the second user and the relationships with other users. The voice interaction system can match the second user with the voiceprint information according to the voiceprint information of the second voice message, so that the voice interaction system can identify the second user according to the voiceprint information.
[0095] By recording the voice information of multiple users multiple times, the identity information of multiple users and the relationships between multiple users can be determined, enabling the identity information of multiple users and the relationships between multiple users to form an identity relationship tree, so that the voice interaction system can obtain the first identity information corresponding to the identity reference information in the preset relationship tree.
[0096] In some embodiments, the identity information of the user can be determined by the voiceprint information in the voice information.
[0097] A voiceprint is the general term for the voice features contained in the voice that can represent and identify the speaker, as well as the voice model established based on these features (parameters). Voiceprint recognition is performed based on the information contained in the voiceprint, enabling the voice interaction system to identify the speaker corresponding to the voice segment according to the voiceprint features of the voice to be recognized.
[0098] Exemplarily, after receiving the voice information of the user, the voice interaction system analyzes the voiceprint of the voice information through voiceprint analysis to determine the second identity information of the user who sent the voice information. For example, the user who sent the voice message can be User A. The voice interaction system then processes and analyzes the voice information to recognize the semantics, completing the preliminary semantic analysis of the voice information and confirming the identity reference information in the voice information.
[0099] The voice interaction system confirms the first identity information corresponding to the identity reference information through the second identity information, the identity reference information, and the preset relationship network. For example, the identity reference information in the voice information can correspond to User B.
[0100] The voice interaction system further processes the voice information in combination with the first identity information, completes the semantic analysis of the voice information, and transfers the result, completing the arrangement of the instructions. Furthermore, the voice interaction system can perform corresponding operations according to the instructions and provide corresponding execution feedback to the user.
[0101] Based on the same concept, the embodiments of the present disclosure also provide an electronic device.
[0102] Among them, the electronic device can be a laptop computer, a smart speaker, a desktop computer, a mobile phone, a digital broadcast terminal, a message receiving and sending device, a game console, a tablet device, a medical device, a fitness device, a personal digital assistant, a translator, and wearable devices such as watches and bracelets, and can be any electronic device with voice interaction function. In the following description, a smart speaker is taken as an example for illustration, but the present disclosure is not limited thereto.
[0103] Figure 7 It is a block diagram of an electronic device shown according to an exemplary embodiment.
[0104] In some embodiments, such as Figure 7As shown in the figure, the electronic device may include: a receiving unit 10 for receiving voice information issued by a user; a determining unit 20 for determining first identity information corresponding to identity reference information included in the voice information; and an operating unit 30 for performing an instruction operation related to the first identity information.
[0105] In some embodiments, the determining unit 20 determines the first identity information corresponding to the identity reference information included in the voice information by the following method: determining the first identity information corresponding to the identity reference information according to a preset relationship tree.
[0106] In some embodiments, the determining unit 20 is further configured to determine, according to the voice information received from the user, that the voice information includes identity reference information.
[0107] In some embodiments, the determining unit 20 determines the first identity information corresponding to the identity reference information according to the preset relationship tree by the following method: determining the second identity information of the user; and confirming the first identity information in the preset relationship tree according to the second identity information and the identity reference information, wherein the second identity information of the user is determined by the voiceprint information of the voice information.
[0108] In some embodiments, the operating unit 30 determines the preset relationship tree in advance by the following method: obtaining the identity information of multiple users; and obtaining an identity relationship tree according to the relationship between the multiple identity information, where the identity relationship tree is the preset relationship tree.
[0109] In some embodiments, the operating unit 30 obtains the identity information of multiple users by the following method: receiving first voice information issued by a first user; the identity information and the corresponding voiceprint information included in the first voice information; receiving second voice information issued by a second user; and determining the identity information and the corresponding voiceprint information included in the second voice information.
[0110] In some embodiments, the operating unit 30 determines the identity information of the user by the voiceprint information of the voice information.
[0111] Regarding the electronic device in the above embodiments, the specific manners in which each module performs operations have been described in detail in the embodiments related to the method, and will not be elaborated here.
[0112] Based on the same concept, an embodiment of the present disclosure further provides an electronic device.
[0113] In some embodiments, an electronic device may include: a processor; and a memory for storing processor-executable instructions; wherein the processor is configured to: execute a voice interaction method.
[0114] Based on the same concept, embodiments of the present disclosure also provide a non-transitory computer-readable storage medium.
[0115] In some embodiments, when the instructions in the storage medium are executed by a processor of a mobile terminal, the mobile terminal is enabled to execute a voice interaction method, and the voice interaction method includes: the voice interaction method according to any one of claims 1 to 7.
[0116] Figure 8 FIG. 7 is a block diagram of an apparatus 800 for performing a voice interaction method according to an exemplary embodiment. For example, the apparatus 800 may be a smart speaker, a mobile phone, a computer, a digital broadcast terminal, a messaging device, a game console, a tablet device, a medical device, a fitness device, a personal digital assistant, etc.
[0117] Referring to Figure 8 , the apparatus 800 may include one or more of the following components: a processing component 802, a memory 804, a power component 806, a multimedia component 808, an audio component 810, an input / output (I / O) interface 812, a sensor component 814, and a communication component 816.
[0118] The processing component 802 generally controls the overall operation of the apparatus 800, such as operations associated with display, telephone calls, data communication, camera operations, and recording operations. The processing component 802 may include one or more processors 820 to execute instructions to complete all or part of the steps of the above method. In addition, the processing component 802 may include one or more modules to facilitate the interaction between the processing component 802 and other components. For example, the processing component 802 may include a multimedia module to facilitate the interaction between the multimedia component 808 and the processing component 802.
[0119] The memory 804 is configured to store various types of data to support the operation of the apparatus 800. Examples of these data include instructions for any application or method operating on the apparatus 800, contact data, phone book data, messages, pictures, videos, etc. The memory 804 may be implemented by any type of volatile or non-volatile storage device or a combination thereof, such as static random access memory (SRAM), electrically erasable programmable read-only memory (EEPROM), erasable programmable read-only memory (EPROM), programmable read-only memory (PROM), read-only memory (ROM), magnetic memory, flash memory, a magnetic disk, or an optical disk.
[0120] The power component 806 provides power to various components of the apparatus 800. The power component 806 may include a power management system, one or more power supplies, and other components associated with generating, managing, and distributing power for the apparatus 800.
[0121] The multimedia component 808 includes a screen that provides an output interface between the device 800 and the user. In some embodiments, the screen may include a liquid crystal display (LCD) and a touch panel (TP). If the screen includes a touch panel, the screen can be implemented as a touch screen to receive input signals from the user. The touch panel includes one or more touch sensors to sense touches, swipes, and gestures on the touch panel. The touch sensors can sense not only the boundaries of the touch or swipe actions but also detect the duration and pressure associated with the touch or swipe operation. In some embodiments, the multimedia component 808 includes a front camera and / or a rear camera. When the device 800 is in an operating mode, such as a shooting mode or a video mode, the front camera and / or the rear camera can receive external multimedia data. Each of the front camera and the rear camera can be a fixed optical lens system or have a focal length and optical zoom capabilities.
[0122] The audio component 810 is configured to output and / or input audio signals. For example, the audio component 810 includes a microphone (MIC) that is configured to receive external audio signals when the device 800 is in an operating mode, such as a call mode, a recording mode, and a voice recognition mode. The received audio signals can be further stored in the memory 804 or transmitted via the communication component 816. In some embodiments, the audio component 810 further includes a speaker for outputting audio signals.
[0123] The I / O interface 812 provides an interface between the processing component 802 and a peripheral interface module, which can be a keyboard, a click wheel, buttons, etc. These buttons can include but are not limited to: a home button, a volume button, a power button, and a lock button.
[0124] The sensor component 814 includes one or more sensors for providing a status assessment of various aspects of the device 800. For example, the sensor component 814 can detect the on / off state of the device 800, the relative positioning of components, such as the display and the keypad of the device 800. The sensor component 814 can also detect a change in the position of the device 800 or a component of the device 800, the presence or absence of user contact with the device 800, the orientation or acceleration / deceleration of the device 800, and the temperature change of the device 800. The sensor component 814 can include a proximity sensor configured to detect the presence of nearby objects without any physical contact. The sensor component 814 can also include a light sensor, such as a CMOS or a CCD image sensor, for use in imaging applications. In some embodiments, the sensor component 814 can further include an acceleration sensor, a gyroscope sensor, a magnetic sensor, a pressure sensor, or a temperature sensor.
[0125] The communication component 816 is configured to facilitate communication, either wired or wirelessly, between the device 800 and other devices. The device 800 may access a wireless network based on communication standards, such as WiFi, 2G, or 3G, or a combination thereof. In an exemplary embodiment, the communication component 816 receives a broadcast signal or broadcast-related information from an external broadcast management system via a broadcast channel. In an exemplary embodiment, the communication component 816 further includes a Near Field Communication (NFC) module to facilitate short-range communication. For example, the NFC module may be implemented based on Radio Frequency Identification (RFID) technology, Infrared Data Association (IrDA) technology, Ultra Wideband (UWB) technology, Bluetooth (BT) technology, and other technologies.
[0126] In an exemplary embodiment, the device 800 may be implemented by one or more Application Specific Integrated Circuits (ASICs), Digital Signal Processors (DSPs), Digital Signal Processing Devices (DSPDs), Programmable Logic Devices (PLDs), Field Programmable Gate Arrays (FPGAs), controllers, microcontrollers, microprocessors, or other electronic components for performing the above method.
[0127] In an exemplary embodiment, a non-transitory computer-readable storage medium including instructions, such as the memory 804 including instructions, is also provided. The above instructions may be executed by the processor 820 of the device 800 to complete the above method. For example, the non-transitory computer-readable storage medium may be a ROM, Random Access Memory (RAM), CD-ROM, magnetic tape, floppy disk, and optical data storage device, among others.
[0128] Figure 9 is a block diagram of a device 1100 for performing a voice interaction method according to an exemplary embodiment. For example, the device 1100 may be provided as a server. Referring to Figure 9 , the device 1100 includes a processing component 1122, which further includes one or more processors, and memory resources represented by the memory 1132 for storing instructions executable by the processing component 1122, such as application programs. The application programs stored in the memory 1132 may include one or more modules each corresponding to a set of instructions. In addition, the processing component 1122 is configured to execute the instructions to perform the above voice interaction method.
[0129] The apparatus 1100 may further include a power supply component 1126 configured to perform power management of the apparatus 1100, a wired or wireless network interface 1150 configured to connect the apparatus 1100 to a network, and an input / output (I / O) interface 1158. The apparatus 1100 may operate based on an operating system stored in the memory 1132, such as Windows ServerTM, MacOS XTM, UnixTM, LinuxTM, FreeBSDTM or the like.
[0130] It can be understood that in the present disclosure, "a plurality of" means two or more, and other quantifiers are similar thereto. "And / or" describes the association relationship of associated objects, indicating that three relationships may exist. For example, A and / or B may represent: A exists alone, A and B exist simultaneously, and B exists alone. The character " / " generally represents an "or" relationship between the associated objects before and after. The singular forms of "a", "the" and "said" are also intended to include the plural forms unless the context clearly indicates otherwise.
[0131] It can be further understood that terms such as "first", "second", etc. are used to describe various information, but this information should not be limited to these terms. These terms are only used to distinguish information of the same type from each other, and do not represent a specific order or importance. In fact, expressions such as "first", "second", etc. can be used interchangeably completely. For example, without departing from the scope of the present disclosure, the first information can also be called the second information, and similarly, the second information can also be called the first information.
[0132] It can be further understood that the orientation or positional relationship indicated by terms such as "center", "longitudinal", "transverse", "front", "rear", "upper", "lower", "left", "right", "vertical", "horizontal", "top", "bottom", "inner", "outer", etc. is based on the orientation or positional relationship shown in the drawings. It is only for the convenience of describing the present embodiment and simplifying the description, rather than indicating or implying that the device or element referred to must have a specific orientation, be constructed and operated in a specific orientation.
[0133] It can be further understood that unless otherwise specified, "connection" includes direct connection without other components between the two, and also includes indirect connection with other elements between the two.
[0134] It can be further understood that although the operations are described in a specific order in the drawings in the embodiments of the present disclosure, it should not be understood as requiring these operations to be performed in the specific order shown or in a serial order, or requiring all the operations shown to obtain the desired result. In a specific environment, multitasking and parallel processing may be advantageous.
[0135] Other embodiments of the present disclosure will be readily apparent to those skilled in the art upon consideration of the specification and practice of the invention disclosed herein. This application is intended to cover any variations, uses, or adaptations of the present disclosure that follow the general principles of the present disclosure and include known or customary technical means in the art not disclosed herein. The specification and examples are to be considered exemplary only, and the true scope and spirit of the present disclosure are indicated by the following claims.
[0136] It should be understood that the present disclosure is not limited to the exact structures described above and shown in the drawings, and various modifications and changes can be made without departing from its scope. The scope of the present disclosure is only limited by the appended claims.
Claims
1. A voice interaction method, characterized in that, the voice interaction method includes: receiving voice information issued by a user; determining first identity information corresponding to identity reference information included in the voice information; performing an instruction operation related to the first identity information.
2. The voice interaction method according to claim 1, characterized in that, the determining the first identity information corresponding to the identity reference information included in the voice information includes: determining the first identity information corresponding to the identity reference information according to a preset relationship tree.
3. The voice interaction method according to claim 1, characterized in that, the voice interaction method further includes: determining that the voice information includes the identity reference information according to the received voice information issued by the user.
4. The voice interaction method according to claim 2, characterized in that, the determining the first identity information corresponding to the identity reference information according to the preset relationship tree includes: determining second identity information of the user; confirming the first identity information in the preset relationship tree according to the second identity information and the identity reference information, wherein, the second identity information of the user is determined through voiceprint information of the voice information.
5. The voice interaction method according to claim 2, characterized in that, the preset relationship tree is pre-determined in the following manner: obtaining identity information of multiple users; obtaining an identity relationship tree according to the relationship between the multiple identity information, and the identity relationship tree is the preset relationship tree.
6. The voice interaction method according to claim 5, characterized in that, the obtaining identity information of multiple users includes: receiving first voice information issued by a first user; the identity information included in the first voice information and the corresponding voiceprint information; receiving second voice information issued by a second user; determining the identity information included in the second voice information and the corresponding voiceprint information.
7. The voice interaction method according to claim 5, characterized in that, the identity information of the user is determined through voiceprint information of the voice information.
8. An electronic device for performing the voice interaction method according to claims 1-7, characterized in that, it includes: a receiving unit for receiving voice information issued by a user; a determining unit for determining first identity information corresponding to identity reference information included in the voice information; an operating unit for performing an instruction operation related to the first identity information.
9. The electronic device according to claim 8, characterized in that, the determining unit determines the first identity information corresponding to the identity reference information included in the voice information through the following method: determining the first identity information corresponding to the identity reference information according to a preset relationship tree.
10. The electronic device according to claim 8, characterized in that, the determining unit is further used for determining that the voice information includes the identity reference information according to the received voice information issued by the user.
11. The electronic device according to claim 9, characterized in that, The determining unit determines the first identity information corresponding to the identity reference information according to a preset relationship tree by the following method: Determine the second identity information of the user; Confirm the first identity information in the preset relationship tree according to the second identity information and the identity reference information, wherein the second identity information of the user is determined by the voiceprint information of the voice information.
12. The electronic device according to claim 11, wherein, The operation unit determines the preset relationship tree in advance by the following method: Obtain the identity information of multiple users; An identity relationship tree is obtained according to the relationships between the multiple pieces of identity information, and the identity relationship tree is the preset relationship tree.
13. The electronic device according to claim 12, wherein, The operation unit obtains the identity information of multiple users by the following method: Receive the first voice information sent by the first user; The identity information included in the first voice information and the corresponding voiceprint information; Receive the second voice information sent by the second user; Determine the identity information included in the second voice information and the corresponding voiceprint information.
14. The electronic device according to claim 12, wherein, The operation unit determines the identity information of the user by the voiceprint information of the voice information.
15. An electronic device, wherein, comprises: a processor; a memory for storing processor-executable instructions; wherein the processor is configured to: execute the voice interaction method according to claims 1 to 7.
16. A non-transitory computer-readable storage medium, when the instructions in the storage medium are executed by a processor of a mobile terminal, enabling the mobile terminal to execute a voice interaction method, the voice interaction method comprises: the voice interaction method according to any one of claims 1 to 7.