Interactive method based on intelligent dinosaur toy and related device

By setting touch points on the smart dinosaur toy and enabling interaction with the server, the problem of existing dinosaur toys not being able to interact with voice has been solved, realizing touch and voice interaction between the smart dinosaur toy and the user, thus enhancing its intelligence.

CN116943253BActive Publication Date: 2026-05-29SHENZHEN RENMA INTERACTIVE TECH CO LTD

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
SHENZHEN RENMA INTERACTIVE TECH CO LTD
Filing Date
2023-07-28
Publication Date
2026-05-29

AI Technical Summary

Technical Problem

Existing dinosaur toys cannot interact with users via voice and lack intelligence.

Method used

Touch points are set on smart dinosaur toys. By detecting the user's touch information, an interaction request is sent to the server. The toy receives and responds to the server's interactive voice information to perform voice interaction, thus realizing touch and voice interaction.

Benefits of technology

This enhances the intelligence of the smart dinosaur toy, enabling it to interact with users and improve the user experience.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN116943253B_ABST
    Figure CN116943253B_ABST
Patent Text Reader

Abstract

Embodiments of the present application provide an interactive method based on a smart dinosaur toy and related devices, the method is applied to a smart dinosaur toy in a voice interaction system, the voice interaction system comprises a server and the smart dinosaur toy, and at least one touch point is arranged on the smart dinosaur toy; the method comprises the following steps: detecting touch information; sending the touch information to the server; if interactive voice information sent by the server in response to the touch information is received, then voice interaction is performed with a target user according to the interactive voice information. In this way, the smart dinosaur toy can perform touch interaction and voice interaction with the user, and the intelligence of the smart dinosaur toy is improved.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] This application belongs to the field of general data processing technology of the Internet, and specifically relates to an interactive method and related device based on a smart dinosaur toy. Background Technology

[0002] Currently, there are many types of dinosaur toys on the market. Some of these toys have sound-producing mechanisms that can be triggered by pressing buttons to produce specific sounds, such as pre-stored roars, lines, or stories. However, existing dinosaur toys cannot interact with users via voice, resulting in a lack of intelligence. Summary of the Invention

[0003] This application provides an interactive method and related device based on intelligent dinosaur toys, aiming to improve the intelligence of intelligent dinosaur toys.

[0004] In a first aspect, this application provides an interactive method based on a smart dinosaur toy, applied to a smart dinosaur toy in a voice interaction system, wherein the voice interaction system includes a server and the smart dinosaur toy, and the smart dinosaur toy is provided with at least one touch point; the method includes:

[0005] Touch information is detected, the touch information is used to instruct the target user to perform a touch operation on the target touch point, the at least one touch point is set based on the characteristics of the smart dinosaur toy, the characteristics refer to the features corresponding to the dinosaur species indicated by the smart dinosaur toy, the features are associated with the corresponding interactive content, the interactive content refers to the relevant knowledge content and / or story content corresponding to the features, and the at least one touch point includes the target touch point;

[0006] Send the touch information to the server;

[0007] If the server receives interactive voice information in response to the touch information, then the server performs voice interaction with the target user based on the interactive voice information; wherein, the interactive voice information is the voice information corresponding to the target interaction content determined by the server based on the number of times the target touch point is touched and historical interaction records, and the target interaction content refers to the relevant knowledge content and / or story content of the first feature corresponding to the target touch point.

[0008] Secondly, this application provides an intelligent toy, including a processor, a memory, a communication interface, and one or more programs, said one or more programs being stored in the memory and configured to be executed by the processor, said programs including instructions for performing the steps of the method described in the first aspect.

[0009] Thirdly, this application provides an electronic device including a processor, a memory, a communication interface, and one or more programs, said one or more programs being stored in the memory and configured to be executed by the processor, said programs including instructions for performing the steps of the first or second aspect of this application.

[0010] Fourthly, this application provides a computer storage medium storing a computer program for electronic data interchange, wherein the computer program causes a computer to perform some or all of the steps described in the first or second aspect of this application.

[0011] Fifthly, this application provides a computer program product, wherein the computer program product includes a non-transitory computer-readable storage medium storing a computer program operable to cause a computer to perform some or all of the steps described in the first or second aspect of this application. The computer program product may be a software installation package.

[0012] As can be seen, in this application, touch information is first detected; the touch information is then sent to the server; if interactive voice information sent by the server in response to the touch information is received, then voice interaction is performed with the target user based on the interactive voice information. This enables the smart dinosaur toy to interact with the user through both touch and voice, improving the intelligence of the smart dinosaur toy. Attached Figure Description

[0013] To more clearly illustrate the technical solutions in the embodiments of this application or the prior art, the drawings used in the description of the embodiments or the prior art will be briefly introduced below. Obviously, the drawings described below are only some embodiments of this application. For those skilled in the art, other drawings can be obtained based on these drawings without creative effort.

[0014] Figure 1a This is a schematic diagram of the structure of a voice interaction system provided in an embodiment of this application;

[0015] Figure 1b This is a schematic diagram of the structure of an electronic device provided in an embodiment of this application;

[0016] Figure 2 This is a flowchart illustrating an interactive method based on a smart dinosaur toy provided in an embodiment of this application;

[0017] Figure 3 This is a schematic diagram of the structure of an interactive device based on a smart dinosaur toy provided in an embodiment of this application. Detailed Implementation

[0018] To enable those skilled in the art to better understand the present application, the technical solutions in the embodiments of the present application will be clearly and completely described below with reference to the accompanying drawings. Obviously, the described embodiments are only some embodiments of the present application, and not all embodiments. Based on the embodiments in the present application, all other embodiments obtained by those of ordinary skill in the art without creative effort are within the scope of protection of the present application.

[0019] The terms "first," "second," etc., in the specification, claims, and accompanying drawings of this application are used to distinguish different objects, not to describe a specific order. Furthermore, the terms "comprising" and "having," and any variations thereof, are intended to cover non-exclusive inclusion. For example, a process, method, system, product, or apparatus that includes a series of steps or units is not limited to the listed steps or units, but may optionally include steps or units not listed, or may optionally include other steps or units inherent to these processes, methods, systems, products, or apparatuses.

[0020] In this document, the term "embodiment" means that a particular feature, structure, or characteristic described in connection with an embodiment may be included in at least one embodiment of this application. The appearance of this phrase in various places throughout the specification does not necessarily refer to the same embodiment, nor is it a separate or alternative embodiment mutually exclusive with other embodiments. It will be explicitly and implicitly understood by those skilled in the art that the embodiments described herein can be combined with other embodiments.

[0021] Currently, there are many types of dinosaur toys on the market. Some of these toys have sound-producing mechanisms that can be triggered by pressing buttons to produce specific sounds, such as pre-stored roars, lines, or stories. However, existing dinosaur toys cannot interact with users via voice, resulting in a lack of intelligence.

[0022] To address the aforementioned issues, this application provides an interactive method based on a smart dinosaur toy. This method can be applied to scenarios where the smart dinosaur toy interacts with a user. Upon detecting touch information, the touch information is sent to a server. If the server responds with interactive voice information in response to the touch information, the toy engages in voice interaction with the target user based on the interactive voice information. This enables the smart dinosaur toy to interact with the user through both touch and voice, enhancing its intelligence. This solution is applicable to various scenarios, including but not limited to the applications mentioned above.

[0023] The system architecture involved in the embodiments of this application is described below.

[0024] This application provides a voice interaction system 100, such as... Figure 1aAs shown, the voice interaction system 100 includes a server 101 and a smart dinosaur toy 102. The smart dinosaur toy 102 has at least one touch point, which is configured based on the characteristics of the smart dinosaur toy 102. These characteristics refer to the features corresponding to the dinosaur species indicated by the smart dinosaur toy, and these features are associated with corresponding interactive content. The interactive content refers to relevant knowledge content and / or story content corresponding to the features. The at least one touch point includes a target touch point. Specifically, the user triggers corresponding voice interaction information by touching the corresponding touch point to achieve voice interaction between the user and the smart dinosaur toy 102.

[0025] This application also provides an electronic device 10, such as... Figure 1b As shown, it includes at least one processor 11, a display screen 12, and a memory 13, and may also include a communications interface 15 and a bus 14. The processor 11, display screen 12, memory 13, and communications interface 15 can communicate with each other via the bus 14. The display screen 12 is configured to display a preset user guide interface in the initial setup mode. The communications interface 15 can transmit information. The processor 11 can call logical instructions in the memory 13 to execute the methods described in the above embodiments.

[0026] Optionally, the electronic device 10 may be a mobile electronic device, an electronic device or other device, and is not limited to a single type.

[0027] Furthermore, the logic instructions in the aforementioned memory 13 can be implemented as software functional units and, when sold or used as independent products, can be stored in a computer-readable storage medium.

[0028] The memory 13, as a computer-readable storage medium, can be configured to store software programs, computer-executable programs, such as program instructions or modules corresponding to the methods in the embodiments of this disclosure. The processor 11 executes functional applications and data processing by running the software programs, instructions, or modules stored in the memory 13, thereby implementing the methods in the above embodiments.

[0029] The memory 13 may include a program storage area and a data storage area. The program storage area may store the operating system and application programs required for at least one function; the data storage area may store data created based on the use of the electronic device 10. Furthermore, the memory 13 may include high-speed random access memory (RAM) and may also include non-volatile memory. For example, various media capable of storing program code, such as USB flash drives, portable hard drives, read-only memory (ROM), random access memory (RAM), magnetic disks, or optical disks, may be used, or they may be transient storage media.

[0030] The specific methods will be described in detail below.

[0031] Please see Figure 2 This application also provides an interactive method based on a smart dinosaur toy, applied to a smart dinosaur toy in a voice interaction system, wherein the voice interaction system includes a server and the smart dinosaur toy, and the smart dinosaur toy is provided with at least one touch point; the method includes:

[0032] Step 201: Touch information detected.

[0033] Wherein, the touch information is used to instruct the target user to perform a touch operation on the target touch point, the at least one touch point is set based on the characteristics of the smart dinosaur toy, the characteristics refer to the features corresponding to the dinosaur species indicated by the smart dinosaur toy, the features are associated with the corresponding interactive content, the interactive content refers to the relevant knowledge content and / or story content corresponding to the features, and the at least one touch point includes the target touch point.

[0034] Specifically, each touch point corresponds to at least one voice interaction content, and these at least one voice interaction content have a progressive relationship in terms of dinosaur-related knowledge. Furthermore, at least one of the at least one voice interaction content is also associated with voice interaction content corresponding to other touch points besides the target touch point. The introductory or associated content of each voice interaction content is related to a part of the dinosaur corresponding to the touch point. For example, if the user touches the dinosaur's wing, then the introductory content of that voice interaction content will be related to the dinosaur's wing, or the voice interaction content will describe content related to the dinosaur's wing.

[0035] Furthermore, each touch point can be linked to trigger corresponding voice interaction content. For example, the dinosaur's mouth can have two touch points: the upper jaw and the lower jaw. When the target user touches both touch points simultaneously, voice interaction content related to the user's bite force or predatory characteristics is provided, such as the mouth characteristics of the current dinosaur species and information about other dinosaurs with similar characteristics, etc., without requiring uniqueness. Conversely, if only one touch point is touched, such as only the upper jaw of the mouth, the characteristics of the upper jaw itself are introduced, such as its structure and type, etc., without requiring uniqueness. Other parts of the dinosaur can also be linked to trigger content. For example, between the pterosaur's wings and feet, a linking trigger could explain why the wings and feet are connected by a webbed structure; for example, between the Tyrannosaurus Rex's forelimbs and hindlimbs, a linking trigger could explain why the forelimbs are weaker than the hindlimbs. Other linking triggers are also possible, but not all are listed here.

[0036] Furthermore, different smart dinosaur toys can also be linked and triggered. When two smart dinosaur toys of different species are connected to the same network, a master-slave relationship is established between the two smart dinosaur toys, which can trigger a linkage between them. That is, when specific touch points of two dinosaurs are touched simultaneously, the server can determine the association between the two dinosaurs as the target interaction content, and then send the target interaction content to the master device between the two smart dinosaur toys. The master device then interacts with the target user via voice.

[0037] Step 202: Send the touch information to the server.

[0038] Step 203: If the server receives interactive voice information in response to the touch information, then the server performs voice interaction with the target user based on the interactive voice information.

[0039] The interactive voice information is the voice information corresponding to the target interactive content determined by the server based on the number of times the target touch point is touched and historical interaction records. The target interactive content refers to the relevant knowledge content and / or story content related to the first feature corresponding to the target touch point.

[0040] In one possible embodiment, the interactive voice information includes first interactive voice information and second interactive voice information, and the target interactive content includes first target interactive content and second target interactive content;

[0041] The step of receiving interactive voice information sent by the server in response to the touch information, and then performing voice interaction with the target user based on the interactive voice information, includes: if the first interactive voice information is received, then performing voice interaction with the target user based on the first interactive voice information, wherein the first interactive voice information is the voice information corresponding to the first target interactive content determined by the server when the first interactive content is not completed, within the interactive content where the first interactive content is not completed; if the second interactive voice information is received, then performing voice interaction with the target user based on the second interactive voice information, wherein the second interactive voice information is the voice information corresponding to the second target interactive content determined by the server when the first interactive content is completed, and the second target interactive content is associated with the first interactive content.

[0042] Specifically, the second interactive voice information includes third and fourth interactive voice information, and the second target interactive content includes third and fourth target interactive content. If the second interactive voice information is received, then voice interaction is performed with the target user based on the second interactive voice information, including: if the third interactive voice information is received, then voice interaction is performed with the target user based on the third interactive voice information, wherein the third interactive voice information is the voice information corresponding to the third target interactive content determined by the server when the first interactive content has been completed, and the third target interactive content refers to the second interactive content that the target touch point has not been interacted with when the current first number of touches of the smart dinosaur toy is greater than the first preset value. The interaction content is related to the third interactive content that has not been interacted with by the first touch point. The first touch point is the second touch point among a plurality of second touch points whose number of touches is less than the first preset value. The second touch point is any touch point in the smart dinosaur toy other than the target touch point. If a fourth interactive voice message is received, then a voice interaction is performed with the target user based on the second interactive voice message. The fourth interactive voice message is the voice message corresponding to the fourth target interactive content determined by the server when the first interactive content has been completed. The fourth target interactive content refers to the interactive content determined according to the depth of knowledge related to the smart dinosaur toy when the current first number of touches of the smart dinosaur toy is less than the first preset value.

[0043] In specific implementation, after receiving the touch information sent by the smart dinosaur toy, the server determines the target touch point touched by the target user based on the touch information. Then, it queries the historical interaction record to obtain the first interaction record when the user last touched the target touch point. Based on the first interaction record, it determines whether the previous voice interaction was completed. If the previous voice interaction was completed, it indicates that the next voice interaction needs to be performed. If the previous voice interaction was not completed, it continues to interact with the user from the previous interaction node, that is, it determines the next interaction content of the previous interaction node as the first target interaction content.

[0044] When the previous voice interaction has been completed, the server determines the number of times the target touch point has been touched, including a first touch count and a second touch count. When the server determines that the first touch count is greater than a first preset value, it compares at least one second interactive content that the target touch point has not interacted with the target user with at least one third interactive content that the first touch point has not interacted with the target user. It then determines whether any of the at least one second interactive content is related to the at least one third interactive content. If so, the related interactive content is designated as the third target interactive content and sent to the smart dinosaur toy. If no related interactive content exists, the next interactive content is determined according to a preset order of the at least one second interactive content and designated as the third interactive content, then sent to the smart dinosaur toy.

[0045] For example, since the server sets a preset order for each touch point of the smart dinosaur toy, the preset order is based on the depth of knowledge related to the smart dinosaur toy, where depth refers to the progressive relationship of dinosaur-related knowledge. Therefore, when the server determines that the first touch count is less than a first preset value, it determines the current interaction content according to the preset order, identifies this current interaction content as the fourth target interaction content, and then sends the fourth target interaction content to the smart dinosaur toy.

[0046] Furthermore, since the target voice interaction content is associated with the voice interaction content of other touch points besides the target touch point among the at least one touch point, the server can guide the user to touch the first touch point during the voice interaction process. At this time, the voice interaction content corresponding to the first touch point will be interspersed during the interaction, so that the user can fully interact with the interaction story of each touch point.

[0047] In one possible embodiment, before detecting touch information, the method further includes: acquiring an interaction request, and then sending the interaction request to the server, wherein the interaction request includes first voice information, the first voice information includes request voice information, the first voice information is provided by at least one first user, the request voice information is provided by the target user, the at least one first user includes the target user, the first voice information is used to indicate the current number of users to the server, the current number of users is used to indicate the voice interaction strategy between the smart dinosaur toy and the target user, the voice interaction strategy is used to instruct the smart dinosaur toy to perform voice interaction with the target user, and the request voice information is used to instruct the server to determine the target interaction content.

[0048] Specifically, the interaction request includes a first interaction request and a second interaction request; the first voice information includes second voice information and third voice information; the request voice information includes first request voice information and second request voice information; the second voice information includes the first request voice information; the third voice information includes the second request voice information; the target users include a first target user and a second target user; and the voice interaction strategy includes a single-person voice interaction strategy and a multi-person voice interaction strategy. Receiving an interaction request and then sending the interaction request to the server includes: if a first interaction request is received, then sending the first interaction request to the server, wherein the first request voice information is provided by the first target user. The target user refers to a first user whose frequency of outputting interactive requests is greater than a second preset value. The second voice information is used to instruct the server to perform voice interaction with the first target user when determining a single-person voice interaction strategy. If a second interaction request is obtained, the second interaction request is sent to the server. The third voice information is provided by multiple first users. The second request voice information refers to the request voice information with the highest accuracy probability corresponding to semantic recognition among the multiple request voice information in the third voice information. The second voice information is used to instruct the server to perform voice interaction with the second target user when determining a multi-person voice interaction strategy. The second target user is the first user corresponding to the second request voice information.

[0049] Furthermore, after sending the second interaction request to the server, the method further includes: obtaining target user response information, wherein the target user response information is user response information obtained after interacting with the target user a preset number of times, and the target user response is used to instruct the server to perform semantic analysis based on the target user response information; the semantic analysis refers to determining whether second voice information of the second user appears after the second target user responds, and engaging in voice interaction with the second user when the second voice information is determined to appear, or continuing to engage in voice interaction with the second target user when the second voice information does not appear; and sending the target user response information to the server.

[0050] In its specific implementation, the smart dinosaur toy acquires first voice information from the current environment, performs semantic analysis on the first voice information to determine whether target voice information exists within it, and if so, sends the first voice information to the server. The target voice information can be a preset keyword, word, or phrase. Upon receiving the first voice information, the server analyzes it to determine at least one fourth voice information containing the target voice information, and then identifies the request voice information from the fourth voice information. The server then determines the target user corresponding to the request voice information. Simultaneously, the server determines the number of users currently interacting with the target user in the current environment based on the first voice information. When the number of first users corresponding to the at least one fourth voice information is 1, the server interacts with the target user using a single-user interaction strategy; when the number of first users corresponding to the at least one fourth voice information is greater than 1, the server interacts with the target user using a multi-user interaction strategy. Specifically, the single-user interaction strategy means that in this voice interaction, the smart dinosaur toy only responds to the voice information of the first target user before the current voice interaction is completed. The multi-user interaction strategy refers to the ability to interact with multiple first users during the current voice interaction. For example, after interacting with the second target user a preset number of times, the fifth voice information of the current environment is reacquired. The fifth voice information includes the target user's response information to the sixth interactive voice information currently sent by the server. The server performs semantic analysis based on the target user's response information to determine whether the second user's second voice information appears after the second target user's response. If the second voice information appears, the server interacts with the second user. Alternatively, if the second voice information does not appear, the server continues to interact with the second target user.

[0051] As can be seen, in this embodiment, the voice interaction strategy can be determined based on the obtained interaction request, enabling the smart dinosaur toy to perform different targeted interaction methods according to the number of users, thereby improving the intelligence of the smart dinosaur toy.

[0052] In one possible embodiment, after interacting with the target user via voice based on the interactive voice information, the method further includes:

[0053] If a fifth interactive voice message is received, then a voice interaction is performed with the target user based on the fifth interactive voice message. The fifth interactive voice message is used to prompt the user to change the touch point when the target user does not like the target interactive content.

[0054] In a specific implementation, during the voice interaction between the smart dinosaur toy and the user, the server can also determine the user's level of liking for the current story based on the user's enthusiasm for the response. If it is determined that the user's level of liking for the target interactive content is not high, the user is prompted to change the touch point to change the voice interaction content.

[0055] Specifically, the level of engagement in responding can be determined by the time interval between responses. For example, if the user's response time interval is less than a third preset value, it is determined that the user has a high level of liking for the target interactive content; if the user's response time interval is greater than the third preset value but less than a fourth preset value, it is determined that the user has a medium level of liking for the target interactive content; if the user's response time interval is greater than the fourth preset value, it is determined that the user has a low level of liking for the target interactive content. When the level of liking is low, the user is prompted to switch to voice interaction content at other touch points.

[0056] As can be seen, in this embodiment, the intelligent dinosaur toy can switch interactive content according to the user's preference, thereby improving the intelligence of the dinosaur toy and the user's interactive interest.

[0057] The above primarily describes the solutions of the embodiments of this application from the perspective of the method execution process. It is understood that, in order to achieve the above functions, mobile electronic devices include corresponding hardware structures and / or software modules for executing each function. Those skilled in the art should readily recognize that, in conjunction with the units and algorithm steps of the various examples described in the embodiments provided herein, this application can be implemented in hardware or a combination of hardware and computer software. Whether a function is executed by hardware or by computer software driving hardware depends on the specific application and design constraints of the technical solution. Those skilled in the art can use different methods to implement the described functions for each specific application, but such implementation should not be considered beyond the scope of this application.

[0058] This application embodiment can divide the electronic device into functional units according to the above method example. For example, each function can be divided into a separate functional unit, or two or more functions can be integrated into one processing unit. The integrated unit can be implemented in hardware or as a software functional unit. It should be noted that the unit division in this application embodiment is illustrative and only represents one logical functional division. In actual implementation, there may be other division methods.

[0059] Please see Figure 3 This application also provides an interactive device 30 based on a smart dinosaur toy, applied to a smart dinosaur toy in a voice interaction system. The voice interaction system includes a server and the smart dinosaur toy, and the smart dinosaur toy is provided with at least one touch point. The interactive device 30 based on the smart dinosaur toy includes:

[0060] The detection unit 31 is used to detect touch information, which is used to instruct the target user to perform a touch operation on the target touch point. The at least one touch point is set based on the characteristics of the smart dinosaur toy. The characteristics refer to the features corresponding to the dinosaur species indicated by the smart dinosaur toy. The characteristics are associated with the corresponding interactive content. The interactive content refers to the relevant knowledge content and / or story content corresponding to the characteristics. The at least one touch point includes the target touch point.

[0061] Sending unit 32 is used to send the touch information to the server;

[0062] The voice interaction unit 33 is configured to, if it receives interactive voice information sent by the server in response to the touch information, perform voice interaction with the target user based on the interactive voice information; wherein, the interactive voice information is voice information corresponding to the target interactive content determined by the server based on the number of times the target touch point is touched and historical interaction records, and the target interactive content refers to the relevant knowledge content and / or story content of the first feature corresponding to the target touch point.

[0063] As can be seen, in this embodiment, touch information is first detected; the touch information is then sent to the server; if interactive voice information is received from the server in response to the touch information, then voice interaction is performed with the target user based on the interactive voice information. This enables the smart dinosaur toy to interact with the user through both touch and voice, improving the intelligence of the smart dinosaur toy.

[0064] In one possible embodiment, the interactive voice information includes first interactive voice information and second interactive voice information, and the target interactive content includes first target interactive content and second target interactive content; regarding the aspect of receiving interactive voice information sent by the server in response to the touch information, and then performing voice interaction with the target user based on the interactive voice information, the voice interaction unit 33 is specifically configured to: if the first interactive voice information is received, then perform voice interaction with the target user based on the first interactive voice information, wherein the first interactive voice information is the voice information corresponding to the first target interactive content determined by the server when the first interactive content is not completed, in the interactive content where the first interactive content is not completed; if the second interactive voice information is received, then perform voice interaction with the target user based on the second interactive voice information, wherein the second interactive voice information is the voice information corresponding to the second target interactive content determined by the server when the first interactive content is completed, and the second target interactive content is associated with the first interactive content.

[0065] In one possible embodiment, the second interactive voice information includes third interactive voice information and fourth interactive voice information, and the second target interactive content includes third target interactive content and fourth target interactive content; if the second interactive voice information is received, then the voice interaction unit 33 is specifically used for: if the third interactive voice information is received, then the voice interaction unit 33 is ... The interaction content related to the second interactive content that has not been interacted with by the first touch point and the third interactive content that has not been interacted with by the first touch point, wherein the first touch point is the second touch point among a plurality of second touch points whose number of touches is less than the first preset value, and the second touch point is any touch point in the smart dinosaur toy other than the target touch point; if a fourth interactive voice information is received, then voice interaction is performed with the target user based on the second interactive voice information, wherein the fourth interactive voice information is the voice information corresponding to the fourth target interactive content determined by the server when the first interactive content has been completed, and the fourth target interactive content refers to the interactive content determined according to the depth of knowledge related to the smart dinosaur toy when the current first number of touches of the smart dinosaur toy is less than the first preset value.

[0066] In one possible embodiment, before detecting touch information, the method further includes: an acquisition unit, configured to acquire an interaction request and then send the interaction request to the server, wherein the interaction request includes first voice information, the first voice information includes request voice information, the first voice information is provided by at least one first user, the request voice information is provided by the target user, the at least one first user includes the target user, the first voice information is used to indicate the current number of users to the server, the current number of users is used to indicate the voice interaction strategy between the smart dinosaur toy and the target user, the voice interaction strategy is used to instruct the smart dinosaur toy to perform voice interaction with the target user, and the request voice information is used to instruct the server to determine the target interaction content.

[0067] In one possible embodiment, the interaction request includes a first interaction request and a second interaction request; the first voice information includes second voice information and third voice information; the request voice information includes first request voice information and second request voice information; the second voice information includes the first request voice information; the third voice information includes the second request voice information; the target users include a first target user and a second target user; and the voice interaction strategy includes a single-person voice interaction strategy and a multi-person voice interaction strategy. The aspect of obtaining an interaction request and then sending the interaction request to the server is specifically configured to: if a first interaction request is obtained, send the first interaction request to the server, wherein the first request voice information is provided by… The first target user provides the voice information, which refers to a user whose frequency of outputting interactive requests is greater than a second preset value. The second voice information is used to instruct the server to interact with the first target user when determining a single-person voice interaction strategy. If a second interaction request is obtained, the server sends the second interaction request. The third voice information is provided by multiple first users. The second request voice information refers to the request voice information with the highest accuracy probability corresponding to semantic recognition among the multiple request voice information in the third voice information. The second voice information is used to instruct the server to interact with the second target user when determining a multi-person voice interaction strategy. The second target user is the first user corresponding to the second request voice information.

[0068] In one possible embodiment, after sending the second interaction request to the server, the method further includes: the acquisition unit, which is configured to acquire target user response information, wherein the target user response information is user response information acquired after a preset number of interactions with the target user, and the target user response is used to instruct the server to perform semantic analysis based on the target user response information; the semantic analysis refers to determining whether second voice information of the second user appears after the second target user's response, and, if the second voice information is determined to appear, engaging in voice interaction with the second user, or, if the second voice information does not appear, continuing to engage in voice interaction with the second target user; and sending the target user response information to the server.

[0069] In one possible embodiment, after interacting with the target user via voice based on the interactive voice information, the method further includes: the voice interaction unit 33 is further configured to, if receiving fifth interactive voice information, interact with the target user via voice based on the fifth interactive voice information, wherein the fifth interactive voice information is used to prompt the user to change the touch point when the target user's liking for the target interactive content is not high. The above embodiments can be implemented entirely or partially by software, hardware, firmware, or any other combination thereof. When implemented using software, the above embodiments can be implemented entirely or partially in the form of a computer program product. The computer program product includes one or more computer instructions or computer programs. When the computer instructions or computer program are loaded or executed on a computer, all or part of the processes or functions described in the embodiments of this application are generated. The computer can be a general-purpose computer, a special-purpose computer, a computer network, or other programmable device. The computer instructions can be stored in a computer-readable storage medium or transmitted from one computer-readable storage medium to another, for example, the computer instructions can be transmitted from one website, computer, server, or data center to another website, computer, server, or data center via wired or wireless means. The computer-readable storage medium can be any available medium that a computer can access, or a data storage device such as a server or data center that includes one or more sets of available media. The available medium can be magnetic media (e.g., floppy disks, hard disks, magnetic tapes), optical media (e.g., DVDs), or semiconductor media. Semiconductor media can be solid-state drives (SSDs).

[0070] This application also provides a computer storage medium storing a computer program for electronic data interchange, which causes a computer to perform some or all of the steps of any of the methods described in the above method embodiments, wherein the computer includes an electronic device.

[0071] This application also provides a computer program product, which includes a non-transitory computer-readable storage medium storing a computer program operable to cause a computer to perform some or all of the steps of any of the methods described in the above method embodiments. The computer program product may be a software installation package, and the computer may include an electronic device.

[0072] It should be understood that in the various embodiments of this application, the order of the above-mentioned processes does not imply the order of execution. The execution order of each process should be determined by its function and internal logic, and should not constitute any limitation on the implementation process of the embodiments of this application.

[0073] In the several embodiments provided in this application, it should be understood that the disclosed methods, apparatuses, and systems can be implemented in other ways. For example, the apparatus embodiments described above are merely illustrative; for example, the division of units is merely a logical functional division, and other division methods may exist in actual implementation; for example, multiple units or components may be combined or integrated into another system, or some features may be ignored or not executed. Furthermore, the coupling or direct coupling or communication connection shown or discussed may be through some interfaces, and the indirect coupling or communication connection between devices or units may be electrical, mechanical, or other forms.

[0074] The units described as separate components may or may not be physically separate. The components shown as units may or may not be physical units; that is, they may be located in one place or distributed across multiple network units. Some or all of the units can be selected to achieve the purpose of this embodiment according to actual needs.

[0075] Furthermore, the functional units in the various embodiments of the present invention can be integrated into one processing unit, or each unit can be physically comprised separately, or two or more units can be integrated into one unit. The integrated unit described above can be implemented in hardware or in the form of hardware plus software functional units.

[0076] The integrated units implemented as software functional units described above can be stored in a computer-readable storage medium. These software functional units, stored in a storage medium, include several instructions to cause a computer device (which may be a personal computer, server, or network device, etc.) to execute some steps of the methods described in the various embodiments of the present invention. The aforementioned storage medium includes: a USB flash drive, a portable hard drive, a magnetic disk, an optical disk, volatile memory, or non-volatile memory. The non-volatile memory can be read-only memory (ROM), programmable read-only memory (PROM), erasable programmable read-only memory (EPROM), electrically erasable programmable read-only memory (EEPROM), or flash memory. The volatile memory can be random access memory (RAM), which is used as an external cache. By way of example, but not limitation, many forms of random access memory (RAM) are available, such as static RAM (SRAM), dynamic RAM (DRAM), synchronous DRAM (SDRAM), double data rate SDRAM (DDR SDRAM), enhanced SDRAM (ESDRAM), synchronous linked DRAM (SLDRAM), and direct rambus RAM (DR RAM), etc., various media capable of storing program code.

[0077] While the present invention has been disclosed above, it is not limited thereto. Any person skilled in the art can easily conceive of variations or substitutions without departing from the spirit and scope of the present invention, and various modifications and alterations can be made, including combinations of the different functions and implementation steps described above, as well as software and hardware implementation methods, all of which are within the protection scope of the present invention.

Claims

1. An interactive method based on a smart dinosaur toy, characterized in that, A smart dinosaur toy used in a voice interaction system, the voice interaction system including a server and the smart dinosaur toy, the smart dinosaur toy being provided with at least one touch point; the method includes: Touch information of a target user on a target touch point is detected, wherein the at least one touch point includes the target touch point; Send the touch information to the server; If the server receives interactive voice information in response to the touch information, then the server performs voice interaction with the target user based on the interactive voice information, including: if a first interactive voice information is received, then the server performs voice interaction with the target user based on the first interactive voice information, wherein the first interactive voice information is the voice information corresponding to a first target interactive content determined by the server when the first interactive content is not completed; and if a second interactive voice information is received, then the server performs voice interaction with the target user based on the second interactive voice information, wherein the second interactive voice information is the voice information corresponding to a second target interactive content determined by the server when the first interactive content is completed, and the second target interactive content is associated with the first interactive content; the interactive voice information includes the first interactive voice information and the second interactive voice information, and the target interactive content includes the first target interactive content and the second target interactive content.

2. The method according to claim 1, characterized in that, After interacting with the target user via voice based on the interactive voice information, the method further includes: If a fifth interactive voice message is received, then a voice interaction is performed with the target user based on the fifth interactive voice message. The fifth interactive voice message is used to prompt the user to change the touch point when the target user does not like the target interactive content.

3. The method according to claim 1, characterized in that, The second interactive voice information includes the third interactive voice information and the fourth interactive voice information, and the second target interactive content includes the third target interactive content and the fourth target interactive content; If a second interactive voice message is received, then voice interaction is performed with the target user based on the second interactive voice message, including: If a third interactive voice message is received, then voice interaction is performed with the target user based on the third interactive voice message. The third interactive voice message is the voice message corresponding to the third target interactive content determined by the server when the first interactive content has been completed. The third target interactive content refers to the interactive content related to the second interactive content that the target touch point has not been interacted with and the third interactive content that the first touch point has not been interacted with when the current first number of touches of the smart dinosaur toy is greater than the first preset value. The first touch point is the touch point among a plurality of second touch points whose second number of touches is less than the first preset value. The second touch point is any touch point in the smart dinosaur toy other than the target touch point. If a fourth interactive voice message is received, then a voice interaction is performed with the target user based on the second interactive voice message. The fourth interactive voice message is the voice message corresponding to the fourth target interactive content determined by the server when the first interactive content has been completed. The fourth target interactive content refers to the interactive content determined according to the order of the depth of knowledge related to the smart dinosaur toy when the current first number of touches of the smart dinosaur toy is less than a first preset value.

4. The method according to claim 1, characterized in that, Before detecting touch information, the method further includes: Upon receiving an interaction request, the interaction request is sent to the server. The interaction request includes first voice information, which includes request voice information. The first voice information is provided by at least one first user, and the request voice information is provided by the target user. The at least one first user includes the target user. The first voice information is used to indicate the current number of users to the server. The current number of users is used to indicate the voice interaction strategy between the smart dinosaur toy and the target user. The voice interaction strategy is used to instruct the smart dinosaur toy to perform voice interaction with the target user. The request voice information is used to instruct the server to determine the target interaction content.

5. The method according to claim 4, characterized in that, The interaction request includes a first interaction request and a second interaction request; the first voice information includes a second voice information and a third voice information; the request voice information includes a first request voice information and a second request voice information; the second voice information includes the first request voice information; the third voice information includes the second request voice information; the target user includes a first target user and a second target user; and the voice interaction strategy includes a single-person voice interaction strategy and a multi-person voice interaction strategy. Upon receiving an interaction request, the interaction request is sent to the server, including: If a first interaction request is received, the first interaction request is sent to the server. The first request voice information is provided by the first target user. The target user refers to the first user whose frequency of outputting interaction requests is greater than a second preset value. The second voice information is used to instruct the server to perform voice interaction with the first target user when determining a single-person voice interaction strategy. If a second interaction request is obtained, the second interaction request is sent to the server. The third voice information is provided by multiple first users. The second request voice information refers to the request voice information with the highest accuracy probability corresponding to semantic recognition among the multiple request voice information in the third voice information. The second voice information is used to instruct the server to perform voice interaction with the second target user when determining the multi-person voice interaction strategy. The second target user is the first user corresponding to the second request voice information.

6. The method according to claim 5, characterized in that, After sending the second interaction request to the server, the method further includes: The system obtains target user response information, which is user response information obtained after a preset number of interactions with the target user. The target user response is used to instruct the server to perform semantic analysis based on the target user response information. The semantic analysis refers to determining whether a second voice message from the second user appears after the second target user responds, and if the second voice message appears, engaging in voice interaction with the second user, or if the second voice message does not appear, continuing to engage in voice interaction with the second target user. The target user's response information is sent to the server.

7. A smart toy, characterized in that, The method includes a processor, a memory, a communication interface, and one or more programs, said programs being stored in the memory and configured to be executed by the processor, said programs including instructions for performing the steps of the method as described in any one of claims 1-6.

8. An electronic device, characterized in that, The method includes a processor, a memory, a communication interface, and one or more programs, said programs being stored in the memory and configured to be executed by the processor, said programs including instructions for performing the steps of the method as described in any one of claims 1-6.

9. A computer-readable storage medium, characterized in that, A computer program for storing electronic data interchange is provided, wherein the computer program causes a computer to execute instructions for the steps of the method as described in any one of claims 1-6.