Interaction method and device, electronic equipment and storage medium
By obtaining audio information in the live broadcast room and generating target interaction information, the problem of low information interaction efficiency in live broadcast interaction is solved, real-time and efficient push of interactive information is achieved, saving resources and improving user experience.
Patent Information
- Application Number
- CN202311869520.1
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2023-12-29
- Publication Date
- 2025-07-01
AI Technical Summary
The existing live broadcast interaction scheme has low information interaction efficiency, resulting in poor comment experience for live broadcast viewers on the live broadcast platform and serious waste of resources.
Real-time interaction is achieved by obtaining audio information in the live broadcast room and when the target audio and interaction requests are detected, the target interaction information is generated and pushed, including text, video, images or special effects props.
It improves information interaction efficiency, saves server bandwidth and computing resources, and ensures real-time and accuracy of interaction.
Smart Images

Figure CN120238667A_ABST
Abstract
Description
Technical Field
[0001] Embodiments of the present disclosure relate to the field of live communication technologies, and in particular, to an interaction method, apparatus, electronic device, and storage medium. Background Art
[0002] With the continuous development of network technologies, various service scenarios emerge in an endless stream. In service scenarios, the scenario functions and interaction methods with objects are often enriched through various associated scenario services. For example, the interaction between a live object and live viewing objects can be realized through a live platform.
[0003] However, in existing live interaction solutions, the information interaction efficiency is low, which greatly reduces the comment experience of live viewing objects in the live platform. Summary of the Invention
[0004] The present disclosure provides an interaction method, apparatus, electronic device, and storage medium to achieve the effect of obtaining interaction information that meets the interaction requirements based on the audio information of the live object and the interaction request of the live viewing object during the live broadcast.
[0005] In a first aspect, an embodiment of the present disclosure provides an interaction method, which includes:
[0006] Obtaining audio information corresponding to a first object in a live room;
[0007] When the audio information corresponding to the first object includes target audio and a target request from a second object corresponding to the live room is received, obtaining target interaction information corresponding to the target audio.
[0008] In a second aspect, an embodiment of the present disclosure further provides an interaction apparatus, which includes:
[0009] An audio information acquisition module, configured to obtain audio information corresponding to a first object in a live room;
[0010] An interaction information acquisition module, configured to obtain target interaction information corresponding to the target audio when the audio information corresponding to the first object includes target audio and a target request from a second object corresponding to the live room is received.
[0011] In a third aspect, an embodiment of the present disclosure further provides an electronic device, where the electronic device includes:
[0012] One or more processors;
[0013] A storage device, configured to store one or more programs,
[0014] When the one or more programs are executed by the one or more processors, the one or more processors implement the interaction method according to any one of the embodiments of the present disclosure.
[0015] In a fourth aspect, an embodiment of the present disclosure further provides a storage medium including computer-executable instructions, which are used to execute the interaction method according to any one of the embodiments of the present disclosure when executed by a computer processor.
[0016] The technical solution of the embodiment of the present disclosure improves the information interaction efficiency while ensuring the real-time nature of the interaction. BRIEF DESCRIPTION OF THE DRAWINGS
[0017] In combination with the accompanying drawings and with reference to the following specific embodiments, the above and other features, advantages and aspects of the various embodiments of the present disclosure will become more obvious. Throughout the drawings, the same or similar reference numerals denote the same or similar elements. It should be understood that the drawings are schematic and the original and elements are not necessarily drawn to scale.
[0018] Figure 1 is a flowchart of an interaction method provided by an embodiment of the present disclosure;
[0019] Figure 2 is a flowchart of an interaction method provided by an embodiment of the present disclosure;
[0020] Figure 3 is a flowchart of an interaction method provided by an embodiment of the present disclosure;
[0021] Figure 4 is a flowchart of an interaction method provided by an embodiment of the present disclosure;
[0022] Figure 5 is a schematic structural diagram of an interaction device provided by an embodiment of the present disclosure;
[0023] Figure 6 is a schematic structural diagram of an electronic device provided by an embodiment of the present disclosure. DETAILED DESCRIPTION OF THE EMBODIMENTS
[0024] The embodiments of the present disclosure will be described in more detail below with reference to the accompanying drawings. Although some embodiments of the present disclosure are shown in the drawings, it should be understood that the present disclosure can be implemented in various forms and should not be construed as limited to the embodiments set forth herein. On the contrary, these embodiments are provided to more thoroughly and completely understand the present disclosure. It should be understood that the drawings and embodiments of the present disclosure are only for exemplary purposes and are not used to limit the protection scope of the present disclosure.
[0025] It should be understood that the various steps described in the method embodiments of the present disclosure may be executed in different orders and / or in parallel. In addition, the method embodiments may include additional steps and / or omit the steps shown. The scope of the present disclosure is not limited in this regard.
[0026] As used herein, the term "comprising" and its variations are open-ended, that is, "including but not limited to". The term "based on" means "at least partially based on". The term "one embodiment" means "at least one embodiment"; the term "another embodiment" means "at least one additional embodiment"; the term "some embodiments" means "at least some embodiments". The relevant definitions of other terms will be given in the following description.
[0027] It should be noted that the concepts such as "first", "second", etc. mentioned in the present disclosure are only used to distinguish different devices, modules or units, and are not used to limit the order of functions performed by these devices, modules or units or their interdependent relationships.
[0028] It should be noted that the modifications of "one" and "multiple" mentioned in the present disclosure are illustrative rather than restrictive. Those skilled in the art should understand that unless otherwise clearly stated in the context, it should be understood as "one or more".
[0029] The names of the messages or information exchanged between multiple devices in the embodiments of the present disclosure are only for illustrative purposes and are not used to limit the scope of these messages or information.
[0030] It can be understood that before using the technical solutions disclosed in the embodiments of the present disclosure, the types, usage scopes, usage scenarios, etc. of the personal information involved in the present disclosure should be informed to the user and the user's authorization should be obtained in an appropriate manner in accordance with relevant laws and regulations.
[0031] For example, when responding to receiving an active request from a user, a prompt message is sent to the user to clearly prompt the user that the operation requested by the user will require obtaining and using the user's personal information. Thus, the user can autonomously choose whether to provide personal information to software or hardware such as an electronic device, an application program, a server, or a storage medium that performs the operations of the technical solutions of the present disclosure according to the prompt message.
[0032] As an optional but non-limiting implementation manner, the manner of sending a prompt message to the user in response to receiving an active request from the user may be, for example, in the form of a pop-up window, and the prompt message may be presented in text in the pop-up window. In addition, the pop-up window may also carry a selection control for the user to choose "agree" or "disagree" to provide personal information to the electronic device.
[0033] It can be understood that the above-mentioned notification and the process of obtaining user authorization are only illustrative and do not limit the implementation manner of the present disclosure. Other manners that comply with relevant laws and regulations can also be applied to the implementation manner of the present disclosure.
[0034] It can be understood that the data involved in the present technical solution (including but not limited to the data itself, the acquisition or use of the data) should comply with the requirements of corresponding laws, regulations and related regulations.
[0035] Before introducing the present technical solution, it should be noted that the device for executing the interaction method provided by the embodiments of the present disclosure can be integrated in an application software that supports interaction functions, and the software can be installed in an electronic device. Optionally, the electronic device can be a mobile terminal or a PC terminal, etc. The application software can be a type of software for data interaction, and specific application software thereof will not be elaborated one by one here as long as it can realize data interaction. It can also be a specially developed application program and integrated in the software for realizing interaction, or integrated in the corresponding page, and the interaction can be realized through the page integrated in the PC terminal.
[0036] Figure 1 It is a schematic flowchart of an interaction method provided by an embodiment of the present disclosure. The embodiments of the present disclosure are applicable to any scenario that needs to obtain interaction information. Optionally, in a live broadcast scenario for obtaining interaction information corresponding to audio information, the method can be executed by an interaction device, and the device can be implemented in the form of software and / or hardware. Optionally, it can be implemented through an electronic device, and the electronic device can be a mobile terminal, a PC terminal or a server, etc.
[0037] As Figure 1 shown, the method includes:
[0038] S110. Obtain audio information corresponding to a first object in the live broadcast room.
[0039] Wherein, the first object in the live broadcast room can be any object associated with the live broadcast room. Optionally, the first object can be the host of the live broadcast room. The audio information can be understood as the multimedia data stream displayed in the live broadcast room. Exemplarily, if the first object corresponds to the host, the audio information can be the host audio.
[0040] In practical applications, during the live broadcast, in addition to being able to display the live broadcast screen corresponding to the live broadcast room, when receiving the audio information of the first object, the audio information corresponding to the first object in the live broadcast room can also be obtained. Further, interaction in the live broadcast room can be realized based on the obtained audio information.
[0041] S120. When the audio information corresponding to the first object includes the target audio and a target request from the second object corresponding to the live broadcast room is received, obtain the target interaction information corresponding to the target audio.
[0042] Among them, the target audio can be the key audio used to determine whether there is corresponding interaction information for the audio information. Optionally, the target audio can be a continuous audio information, or an audio information including at least one keyword, etc. The target audio includes at least one of a welcome audio, a thank-you audio, and a sharing audio. The second object can be understood as the object entering the live broadcast room of the first object. The second object can be the object whose terminal display interface shows the live broadcast room of the first object. Exemplarily, the second object can be the live broadcast room audience. The target request can be understood as an interaction request associated with the multimedia data stream. Generally, during the live broadcast room display process, it may include the live broadcast room owner side and the live broadcast room viewing side. Both sides can input requests for the page showing the live broadcast room. At least part of the requests input by at least one of these two sides for the live broadcast room display page can be determined as the target request. Optionally, the target request includes at least one of an enter live broadcast room request, a resource transfer request, and a resource sharing request, where the resource can be a virtual resource, and resource transfer can be understood as the resource being transferred from the resource sender to the resource receiver. It should be noted that the target audio can correspond to the target request. Optionally, if the target request is an enter live broadcast room request, the target audio can be a welcome audio; if the target request is a resource transfer request, the target audio can be a thank-you audio; if the target request is a resource sharing request, the target audio can be a sharing audio. The target interaction information can be understood as information that can implement an interaction function and has a preset representation form. The target interaction information can include any multimedia information that can be displayed on the display interface. Optionally, the interaction information can include at least one of audio, text, video, image, and animation special effects.
[0043] In practical applications, during the live broadcast, it is possible to receive the interaction request input by the second object corresponding to the live broadcast room for the live broadcast room and determine whether the received interaction request is a target request. Also, it is possible to determine whether the obtained audio information corresponding to the first object includes the target audio. Further, when it is determined that the audio information corresponding to the first object includes the target audio and the target request from the second object corresponding to the live broadcast room is received, the target interaction information corresponding to the target audio can be obtained. Thus, the interaction between the first object and the second object, or the interaction between the second object corresponding to the target request and other second objects can be realized based on the target interaction information.
[0044] The technical solution of the embodiment of the present invention obtains the audio information corresponding to the first object in the live broadcast room. Further, in the case where the audio information corresponding to the first object includes the target audio and the target request of the second object corresponding to the live broadcast room is received, the target interaction information corresponding to the target audio is obtained, which solves the problems existing in the related technology, such as the inability to achieve interaction backtracking, the easy miss of interaction information by the live broadcast viewing object, and the great waste of resources such as the terminal performance, server bandwidth, and computing power. It realizes the effect of obtaining the interaction information that meets the interaction requirements based on the audio information of the live broadcast object and the interaction request of the live broadcast viewing object during the live broadcast process, saves resources such as server bandwidth and computing power, and achieves the effect of improving the interaction information interaction efficiency while ensuring the real-time nature of the interaction.
[0045] Figure 2 FIG. 4 is a schematic flowchart of an interaction method provided by an embodiment of the present disclosure. On the basis of the above embodiment, the audio information including the target audio is sent to the first server, so that the first server converts the audio information including the target audio into a text to be processed. Further, the text to be processed sent by the first server is received, and the text to be processed and the operation information corresponding to the target request are sent to the second server, so that the second server determines the target interaction information based on the operation information and the text to be processed. The specific display manner thereof can refer to the description of this embodiment. Among them, the same or similar technical features as those in the foregoing embodiment will not be described in detail here.
[0046] As Figure 2 shown, the method of this embodiment may specifically include:
[0047] S210. Obtain the audio information corresponding to the first object in the live broadcast room.
[0048] S220. In the case where the audio information corresponding to the first object includes the target audio and the target request of the second object corresponding to the live broadcast room is received, send the audio information including the target audio to the first server, so that the first server converts the audio information including the target audio into a text to be processed.
[0049] Among them, the first server may be a server capable of performing a voice detection function. The text to be processed may be text information corresponding to the audio to be processed.
[0050] In practical applications, in the case where the audio information corresponding to the first object includes the target audio and the target request of the second object corresponding to the live broadcast room is received, the audio information including the target audio may be sent to the first server, so that the first server converts the audio information including the target audio into a text to be processed.
[0051] It should be noted that converting the audio information into the text to be processed may be to perform text detection processing on the audio information based on a preset algorithm and use the obtained text information as the text to be processed; alternatively, the audio information may also be processed through other audio-text conversion methods to obtain the text to be processed, and the embodiments of the present disclosure do not make specific limitations on this.
[0052] S230. Receive the text to be processed sent by the first server, and send the text to be processed and the operation information corresponding to the target request to the second server, so that the second server determines the target interaction information based on the operation information and the text to be processed.
[0053] Among them, the operation information may be an interaction operation generated during the live broadcast. The operation information can also be understood as operation information including the operation initiator object, the operation target object, the operation description, the operation type, etc., that is, an operation in which at least two parties influence each other. The operation information corresponding to the target request can be understood as information representing the interaction operation corresponding to the target request. The second server may be a service server.
[0054] In practical applications, when receiving the text to be processed, the operation information corresponding to the target request can be obtained. Optionally, when the target request is a resource transfer request, the operation information may be resource transfer operation information; when the target request is a request to enter the live broadcast room, the operation information may be operation information for entering the live broadcast room. Further, the received text to be processed and the operation information corresponding to the target request can be sent to the second server, so that the second server determines the target interaction information based on the text to be processed and the operation information. The interaction information may be comment information. Specifically, the content of the comment information can be determined based on the text to be processed and the operation information, and the comment information is displayed in the live broadcast room (live broadcast interface) of the first object.
[0055] It should be noted that during the display process of the live broadcast room of the first object, there will be multiple second objects associated with the live broadcast room, and each second object may initiate an interaction request for the live broadcast page. Correspondingly, during the process of obtaining the interaction request, multiple interaction requests corresponding to the second objects will be obtained. Therefore, in order to determine the second object corresponding to the target request based on the interaction information, when it is determined that the obtained interaction request includes the target request, the second object corresponding to the target request can be determined, and the object identifier of the second object can be obtained. Furthermore, the text to be processed and the object identifier of the second object corresponding to the operation information can be sent to the second server.
[0056] Optionally, sending the text to be processed and the operation information corresponding to the target request to the second server includes: sending the text to be processed and the object identifier of the second object corresponding to the operation information to the second server.
[0057] In practical applications, when receiving the text to be processed, the received text to be processed and the object identifier of the second object corresponding to the operation information can be sent to the second server.
[0058] It should be noted that before sending the object identifier of the second object to the second server, the object identifier of the second object corresponding to the operation information can also be determined.
[0059] In this embodiment, the determination process of the object identifier can include at least two cases. For different cases, different determination methods can be corresponding. The following can separately describe these two cases:
[0060] The first case: When receiving the target request of the second object corresponding to the live broadcast room, the object attribute information of the second object corresponding to the target request is used as the object identifier.
[0061] Among them, the object attribute information can be the object identifier, such as the object avatar or the object nickname. Optionally, it can be at least part of the characters in the object nickname.
[0062] The second case: When the target audio is included in the audio information, the second object and the object identifier of the second object that match the target audio are determined according to the target request.
[0063] In practical applications, after obtaining the audio information of the first object, it can be determined whether the obtained audio information includes the target audio. Further, when it is determined that the target audio is included in the audio information, the obtained interaction requests can be detected to determine the interaction request that matches the target audio among the interaction requests associated with multiple second objects obtained, and this interaction request is used as the target request to determine the second object corresponding to the target request. Furthermore, the object attribute information of the second object can be determined, and the object attribute information is used as the object identifier of the second object.
[0064] Further, when the second server receives the object identifier of the second object and the text to be processed, the second server can determine the target interaction information based on the object identifier and the text to be processed.
[0065] It should be noted that the target interaction information is a copywriting that can achieve the interaction effect between at least two parties and the corresponding at least two parties influence each other. And the target interaction information is generated based on the text to be processed corresponding to the first object and the object identifier of the second object. Therefore, the target interaction information can at least include the text to be processed corresponding to the first object and the object identifier of the second object.
[0066] It should be noted that the information type of the target interaction information may include at least one of text, video, image, and special effect props. Correspondingly, generating the target interaction information according to the object identifier and the text to be processed can include multiple implementation methods, and the following can respectively explain these implementation methods:
[0067] The first method: When the target interaction information includes text information, the text to be processed can be used as the interaction text, and the interaction text can be used as the target interaction information.
[0068] The third method: When the target interaction information includes image information, an image corresponding to the text to be processed after matching processing can be retrieved from the material library in the application software, and the image can be used as the interaction image; alternatively, the text to be processed can be processed based on a preset image generation algorithm to obtain a generated image, and the generated image can be used as the interaction image. Furthermore, the interaction image can be used as the target interaction information.
[0069] The fourth method: When the target interaction information includes video information, a video segment corresponding to the text to be processed can be determined based on the video information displayed in the live broadcast room, and the video segment can be used as the interaction video. Or, a video can be randomly retrieved from a pre-stored video library, and the video can be used as the interaction video. Furthermore, the interaction video can be used as the target interaction information.
[0070] The fifth method: When the target interaction information includes special effect props, the special effect prop corresponding to the text to be processed can be determined from multiple pre-set special effect props according to the keywords in the text to be processed, and the special effect prop can be used as the interaction special effect. Furthermore, the interaction special effect can be used as the target interaction information.
[0071] S240. Send the target interaction information to the third server so that the third server can push the target interaction information to at least one client.
[0072] Among them, the third server can be a live streaming server. The live streaming server can be understood as a server deployed at the interface of the Content Delivery Network (CDN). At least one client can be understood as a client associated with the live broadcast page. At least one client may include the client of the first object and the client of the second object.
[0073] In practical applications, after determining the target interaction information, the target interaction information can be sent to the third server. Furthermore, the target interaction information can be pushed to at least one client based on the third server, so as to display the interaction information based on the display interfaces corresponding to the respective clients.
[0074] It should be noted that in addition to sending the target interaction information to the third server, the audio information corresponding to the target interaction information can also be sent to the third server, so that the third server can push the target interaction information and the audio information to at least one client.
[0075] Optionally, sending the target interaction information to the third server so that the third server can push the target interaction information to at least one client includes: sending the target interaction information and the audio information corresponding to the target interaction information to the third server; based on the third server pushing the target interaction information and the audio information to the clients of the second object to display the target interaction information on the interaction public screen of the client.
[0076] Among them, the interaction public screen can be understood as the area in the client display interface for displaying interaction information.
[0077] In practical applications, when the target interaction information is received, the audio information corresponding to the target interaction information can be determined. Furthermore, the target interaction information and the audio information corresponding to the target interaction information are sent to the third server. Further, based on the third server, the target interaction information and the audio information can be pushed to the clients of all second objects associated with the first object. Furthermore, the target interaction information can be displayed on the interaction public screen of the client.
[0078] In practical applications, when the target interaction information is displayed on the interaction public screen, the content displayed in the target interaction information can also be triggered, and then, the triggered content can be displayed.
[0079] Based on this, on the basis of the above technical solutions, it further includes: when it is detected that the interaction content in the target interaction information is triggered, displaying the target interface corresponding to the interaction content.
[0080] In practical applications, the target interaction information can include the object identifier corresponding to the second object, so as to realize the interaction between the first object and the second object based on the target interaction information. And, in order to facilitate the first object and / or other second objects except the second object appearing in the target interaction information to understand the second object appearing in the copywriting, the object identifier corresponding to the second object in the target interaction information can be set to a triggerable state, and this information is used as the interaction content in the target interaction information. When a trigger operation on this interaction content by the object is received, the target interface corresponding to the second object is displayed.
[0081] Among them, the interactive content may be content representing object attribute information. Optionally, the interactive content may include an object identifier, for example, an object avatar or an object nickname. Optionally, the target interface may include the main interface or the session interface of the second object. The main interface can be understood as the interface where the object's home page is located.
[0082] In practical applications, after obtaining the target interactive information and before displaying the target interactive information on the display interface, the interactive content in the target interactive information can be set to a triggerable state in advance. Furthermore, the target interactive information can be displayed on the display interface.
[0083] Furthermore, when a trigger operation for the interactive content in the target interactive information is detected, the target interface corresponding to the interactive content can be displayed. Exemplarily, when a trigger operation for the object identifier in the target interactive information is detected, the target interface of the second object corresponding to the object identifier can be displayed. It should be noted that the display method of the target interface can be jump display, pop-up display, partial screen display, etc.
[0084] The technical solution of the embodiment of the present invention, by obtaining the audio information corresponding to the first object in the live broadcast room, further, in the case that the audio information corresponding to the first object includes the target audio and the target request of the second object corresponding to the live broadcast room is received, sending the audio information including the target audio to the first server, so that the first server converts the audio information including the target audio into a text to be processed, receiving the text to be processed sent by the first server, and sending the text to be processed and the operation information corresponding to the target request to the second server, so that the second server determines the target interactive information based on the operation information and the text to be processed, and sending the target interactive information to the third server, so that the third server pushes the target interactive information to at least one client, realizes the effect of generating interactive information that meets the interactive needs based on the audio information of the live broadcast object and the target request of the live broadcast viewing object during the live broadcast, saves server resources such as bandwidth and computing power, and achieves the effect of improving the accuracy of interactive information while ensuring the real-time nature of the interaction.
[0085] Figure 3 It is a schematic flowchart of an interaction method provided by an embodiment of the present disclosure. On the basis of the above embodiment, in the case of receiving the text to be processed sent by the first server, preprocessing the text to be processed based on the second server to obtain a text to be matched, and further, performing a matching process on the text to be matched and the object identifier based on the second server to obtain the target interactive information, and the specific display method thereof can refer to the description of this embodiment. Technical features that are the same as or similar to those in the foregoing embodiments will not be described in detail here.
[0086] As Figure 3 shown, the method of this embodiment may specifically include:
[0087] S310. Obtain the audio information corresponding to the first object in the live broadcast room.
[0088] S320. When the audio information corresponding to the first object includes the target audio and a target request from the second object corresponding to the live broadcast room is received, send the audio information including the target audio to the first server, so that the first server converts the audio information including the target audio into a text to be processed.
[0089] S330. Receive the text to be processed sent by the first server, and preprocess the text to be processed based on the second server to obtain a text to be matched.
[0090] Among them, preprocessing can be understood as processing the text to be processed based on a preset text processing method. The preset text processing method can include various methods. Optionally, it can include text representation form adjustment and / or text elimination, etc. The text to be matched can be understood as the text information obtained after preprocessing the text to be processed.
[0091] In practical applications, after obtaining the audio to be processed, the audio to be processed can be textually processed, and the obtained text information can be used as the text to be processed. Further, the text to be processed can be preprocessed to process the text to be processed into a text that meets the preset copywriting generation standard, and the processed text can be used as the text to be matched.
[0092] Optionally, preprocessing the text to be processed to obtain a text to be matched includes: representing the first type of text in the text to be processed in a first preset manner, and representing the second type of text in the text to be processed in a second preset manner to obtain an adjusted text; adjusting the third type of text in the adjusted text to obtain the text to be matched.
[0093] Among them, the first type of text can be any type of text information. Optionally, the first type of text can be letters. The first preset manner can be understood as a manner that is preset and used to limit the representation of the first type of text in the text. The first preset manner can be any representation manner. Optionally, it can be represented by lowercase letters. The second type of text can be any type of text information. Optionally, the second type of text can be numbers. The second preset manner can be understood as a manner that is preset and used to limit the representation of the second type of text in the text. The second preset manner can be any representation manner. Optionally, it can be represented by Arabic numerals. The third type of text can be any type of text information. Optionally, the third type of text can be stop words and / or punctuation marks. Stop words can be function words included in natural languages. These function words are very common. Compared with other words, function words have no actual meaning. That is to say, stop words can be words without actual meaning in natural languages. Optionally, stop words can include modal particles, adverbs, conjunctions, and connectives, etc. Punctuation marks are symbols used in writing to indicate pauses and tones. Optionally, punctuation marks include full stops (such as periods, question marks, and exclamation marks, etc.) and labels (such as quotation marks, parentheses, and ellipsis marks, etc.).
[0094] In practical applications, after obtaining the text to be processed, the first type of text in the text to be processed and the second type of text in the text to be processed can be obtained respectively. Further, the representation manner of the first type of text in the text to be processed can be adjusted to the first preset manner, and the representation manner of the second type of text in the text to be processed can be adjusted to the second preset manner. Then, the text obtained after adjustment can be used as the text to be adjusted. After that, the third type of text in the text to be adjusted can be detected, and the third type of text in the text to be adjusted can be adjusted. The adjusted text to be adjusted can be used as the text to be matched. The advantage of such a setting is that it improves the standardization degree of interactive information and enhances the conciseness of interactive information.
[0095] S340. Based on the second server, perform matching processing on the text to be matched and the object identifier to obtain the target interactive information.
[0096] In this embodiment, after obtaining the text to be matched, the text to be matched can be directly matched with the object identifier to obtain the target interactive information. The advantage of such a setting is that it improves the matching efficiency and matching accuracy between the text to be matched and the object identifier. At the same time, it enhances the conciseness of interactive information.
[0097] Optionally, based on the second server's matching process for the text to be matched and the object identifier, the target interaction information is obtained, including: determining the object identifier to be modified corresponding to the second object in the text to be matched based on the second server; obtaining the target object identifier according to the object identifier to be modified and the object identifier; and replacing the object identifier to be modified in the text to be matched with the target object identifier to obtain the target interaction information.
[0098] Among them, the object identifier to be modified can be understood as the object identifier corresponding to the second object included in the text to be matched.
[0099] It should be noted that before matching the object identifier to be modified and the object identifier, the object identifier can also be preprocessed. Specifically, the first type of text in the object identifier is represented in a first preset manner, and the second type of text in the object identifier is represented in a second preset manner to obtain the object identifier to be adjusted; the second type of text in the object identifier to be adjusted is adjusted to obtain the object identifier to be matched. Furthermore, the object identifier to be modified can be matched with the object identifier to be matched to obtain the target object identifier.
[0100] In practical applications, after obtaining the text to be matched, the text information that may conform to the object identifier of the second object can be determined based on the object identifier of the second object, and the determined text information is used as the object identifier to be modified. Further, the object identifier to be modified can be matched with the associated information to obtain the target object identifier. It should be noted that when performing the matching process, prefix matching and / or suffix matching methods can be used, and the embodiments of the present disclosure do not make specific limitations on this.
[0101] Optionally, the prefix character in the object identifier can be determined. Then, after obtaining the object identifier to be modified, the prefix character is matched with the object identifier to be modified. Further, in the case where the prefix character matches the object identifier to be modified, it can be determined that the object identifier to be modified and the object identifier match each other, and the target object identifier is obtained.
[0102] Optionally, the suffix character in the object identifier can be determined. Then, after obtaining the object identifier to be modified, the object identifier to be modified can be matched with the determined suffix character. Further, in the case where the suffix character matches the object identifier to be modified, it can be determined that the object identifier to be modified and the object identifier match each other, and the target object identifier is obtained.
[0103] In the actual application process, there may be a situation where multiple object identifiers to be modified match the object identifier at the same time. At this time, the number of matching characters corresponding to each object identifier to be modified can be determined respectively. Furthermore, the target matching character number with the largest value can be determined according to the number of matching characters of each, and the object identifier corresponding to the target matching character number can be used as the basis for generating the target interaction information. The advantage of this setting is that it realizes the effect of mutual matching between audio information and operation data. Furthermore, it enhances the relevance between the interaction information, the audio information, and the operation data, and improves the accuracy of the interaction information.
[0104] Furthermore, the object identifier to be modified in the text to be matched can be replaced based on the target object identifier, and the text to be matched after replacement is used as the target interaction information.
[0105] The technical solution of the embodiment of the present invention is to obtain the audio information of the first object in the live broadcast room. After that, when the audio information of the first object includes the target audio and the target request of the second object corresponding to the live broadcast room is received, the audio information including the target audio is sent to the first server, so that the first server converts the audio information including the target audio into the text to be processed. Furthermore, receive the text to be processed sent by the first server, preprocess the text to be processed based on the second server to obtain the text to be matched, and perform matching processing on the text to be matched and the object identifier based on the second server to obtain the target interaction information, realizing the effect of triggering audio keyword detection based on preset operation data, or triggering operation data detection based on preset keywords to complete information interaction, enhancing the intelligence of the information interaction function, and improving the matching degree between the interaction information and the operation data.
[0106] Figure 4 It is a schematic flowchart of an interaction method provided by an embodiment of the present disclosure. The embodiment of the present disclosure is an optional embodiment of the above-mentioned various disclosure embodiments. Taking the target request as the request to enter the live broadcast room, the first object as the object terminal to which the anchor belongs, and the second object as the object terminal to which the audience belongs as an example, the interaction method provided by the embodiment of the present disclosure will be described. As Figure 4 shown, the execution process of the method of the embodiment of the present disclosure can be as follows:
[0107] During the live broadcast, when the client to which the object (the host and / or the audience) belongs detects that the audience enters the live broadcast room, keyword detection is performed on the audio information of the host displayed in the live broadcast room, and the audio information of the host is obtained. After that, the object identifier of the new object in the live broadcast room of the host can be determined, and these object identifiers and audio information are sent to the server. Further, based on the voice detection function in the server, voice detection is performed on the audio information to obtain a voice detection result, and the voice detection result and the new object identifier are sent to the business server and the data server in the server. After that, based on the algorithms in the business server and the data server, the voice detection result and the new object identifier are matched to obtain a matching result. When the matching result is a successful match, a target interaction message is generated and sent to the client. Finally, the target interaction message is displayed based on the display interface of the client to which the host and / or the audience belongs.
[0108] The technical solution of the embodiment of the present invention, by obtaining the audio information corresponding to the first object in the live broadcast room, further, when the audio information corresponding to the first object includes the target audio and a target request of the second object corresponding to the live broadcast room is received, obtaining the target interaction information corresponding to the target audio, solves the problems existing in the related art such as the inability to achieve interaction backtracking, the live broadcast viewing object is likely to miss the interaction information, and a great waste of resources such as the terminal performance, the server bandwidth, and the computing power, etc., realizes the effect of obtaining the interaction information that meets the interaction requirements based on the audio information of the live broadcast object and the interaction request of the live broadcast viewing object during the live broadcast, saves resources such as the server bandwidth and the computing power, and, achieves the effect of improving the accuracy of the interaction information while ensuring the real-time nature of the interaction.
[0109] Figure 5 It is a schematic structural diagram of an interaction device provided by an embodiment of the present disclosure, as Figure 5 shown, the device includes: an audio information acquisition module 410 and an interaction information acquisition module 420.
[0110] Among them, the audio information acquisition module 410 is used to acquire the audio information corresponding to the first object in the live broadcast room; the interaction information acquisition module 420 is used to acquire the target interaction information corresponding to the target audio when the audio information corresponding to the first object includes the target audio and a target request of the second object corresponding to the live broadcast room is received.
[0111] On the basis of the above technical solutions, optionally, the device further includes: an audio conversion module and an interaction information determination module.
[0112] The audio conversion module is used to send the audio information including the target audio to the first server, so that the first server converts the audio information including the target audio into a text to be processed;
[0113] An interactive information determination module, configured to receive the text to be processed sent by the first server, and send the text to be processed and the operation information corresponding to the target request to the second server, so that the second server determines the target interactive information based on the operation information and the text to be processed.
[0114] Based on the above technical solutions, optionally, the device further includes: an interactive information pushing module.
[0115] The interactive information pushing module is configured to, after obtaining the target interactive information corresponding to the target audio, send the target interactive information to the third server, so that the third server pushes the target interactive information to at least one client.
[0116] Based on the above technical solutions, optionally, the interactive information determination module is specifically configured to send the text to be processed and the object identifier of the second object corresponding to the operation information to the second server.
[0117] Based on the above technical solutions, optionally, the device further includes: a text preprocessing module and an object identifier matching module.
[0118] The text preprocessing module is configured to preprocess the text to be processed based on the second server to obtain the text to be matched;
[0119] The object identifier matching module is configured to perform matching processing on the text to be matched and the object identifier based on the second server to obtain the target interactive information.
[0120] Based on the above technical solutions, optionally, the text preprocessing module includes: a text representation mode adjustment unit and a text adjustment unit.
[0121] The text representation mode adjustment unit is configured to represent the first type of text in the text to be processed in a first preset manner and represent the second type of text in the text to be processed in a second preset manner based on the second server to obtain the text to be adjusted; and
[0122] The text adjustment unit is configured to adjust the third type of text in the text to be adjusted to obtain the text to be matched.
[0123] Based on the above technical solutions, optionally, the object identifier matching module includes: a to-be-modified object identifier determination unit, a target object identifier determination unit, and an object identifier replacement unit.
[0124] A to-be-modified object identifier determination unit, configured to determine a to-be-modified object identifier corresponding to the second object in the to-be-matched text based on the second server;
[0125] A target object identifier determination unit, configured to obtain a target object identifier according to the to-be-modified object identifier and the object identifier;
[0126] An object identifier replacement unit, configured to replace the to-be-modified object identifier in the to-be-matched text based on the target object identifier to obtain the target interaction information.
[0127] Based on the above technical solutions, optionally, the interaction information push module includes: an audio information sending unit and an interaction information display unit.
[0128] The audio information sending unit is configured to send the target interaction information and audio information corresponding to the target interaction information to the third server;
[0129] The interaction information display unit is configured to push the target interaction information and the audio information to the client of the second object based on the third server to display the target interaction information on the interaction public screen of the client.
[0130] Based on the above technical solutions, optionally, the target request includes at least one of a request to enter a live broadcast room, a resource transfer request, and a resource sharing request, and the target audio includes at least one of a welcome audio, a thank-you audio, and a sharing audio.
[0131] Based on the above technical solutions, optionally, the device further includes: a target interface display module.
[0132] The target interface display module is configured to display a target interface corresponding to the interaction content when detecting a trigger of the interaction content in the target interaction information, where the interaction content includes an object identifier.
[0133] Based on the above technical solutions, optionally, the target interface includes a main interface or a conversation interface of the second object.
[0134] The technical solution of the embodiment of the present invention obtains the audio information corresponding to the first object in the live broadcast room. Further, when the audio information corresponding to the first object includes the target audio and the target request of the second object corresponding to the live broadcast room is received, the target interaction information corresponding to the target audio is obtained, which solves the problems existing in the related technologies, such as the inability to achieve interaction backtracking, the easy miss of interaction information by the live broadcast viewing object, and the great waste of resources such as the terminal performance, server bandwidth, and computing power. It realizes the effect of obtaining the interaction information that meets the interaction requirements based on the audio information of the live broadcast object and the interaction request of the live broadcast viewing object during the live broadcast, saves resources such as server bandwidth and computing power, and achieves the effect of improving the accuracy of the interaction information while ensuring the real-time nature of the interaction.
[0135] The interaction device provided by the embodiments of the present disclosure can execute the interaction method provided by any embodiment of the present disclosure, and has corresponding functional modules and beneficial effects for executing the method.
[0136] It should be noted that the various units and modules included in the above device are only divided according to the functional logic, but are not limited to the above division, as long as the corresponding functions can be realized; in addition, the specific names of the functional units are only for the convenience of mutual distinction and do not limit the protection scope of the embodiments of the present disclosure.
[0137] Figure 6 It is a schematic structural diagram of an electronic device provided by the embodiments of the present disclosure. Next, refer to Figure 6 which shows a schematic structural diagram of an electronic device (such as the terminal device or server in Figure 6 ) 500 suitable for implementing the embodiments of the present disclosure. The terminal device in the embodiments of the present disclosure may include, but is not limited to, mobile terminals such as mobile phones, laptop computers, digital broadcast receivers, PDAs (Personal Digital Assistants), PADs (Tablet Computers), PMPs (Portable Multimedia Players), in-vehicle terminals (such as in-vehicle navigation terminals), etc., and fixed terminals such as digital TVs, desktop computers, etc. Figure 6 The electronic device shown is only an example and should not bring any limitation to the functions and usage scope of the embodiments of the present disclosure.
[0138] Such as Figure 6As shown, the electronic device 500 may include a processing device (such as a central processing unit, a graphics processing unit, etc.) 501, which may perform various appropriate actions and processes according to a program stored in a read-only memory (ROM) 502 or a program loaded from a storage device 508 into a random access memory (RAM) 503. In the RAM 503, various programs and data required for the operation of the electronic device 500 are also stored. The processing device 501, the ROM 502, and the RAM 503 are connected to each other through a bus 504. An editing / output (I / O) interface 505 is also connected to the bus 504.
[0139] Generally, the following devices may be connected to the I / O interface 505: an input device 506 including, for example, a touch screen, a touchpad, a keyboard, a mouse, a camera, a microphone, an accelerometer, a gyroscope, etc.; an output device 507 including, for example, a liquid crystal display (LCD), a speaker, a vibrator, etc.; a storage device 508 including, for example, a magnetic tape, a hard disk, etc.; and a communication device 509. The communication device 509 may allow the electronic device 500 to communicate with other devices wirelessly or wiredly to exchange data. Although Figure 6 an electronic device 500 with various devices is shown, it should be understood that it is not required to implement or include all the shown devices. Instead, more or fewer devices may be implemented or included.
[0140] Specifically, according to an embodiment of the present disclosure, the process described above with reference to the flowchart may be implemented as a computer software program. For example, an embodiment of the present disclosure includes a computer program product, which includes a computer program carried on a non-transitory computer-readable medium, and the computer program includes program codes for executing the method shown in the flowchart. In such an embodiment, the computer program may be downloaded and installed from a network through the communication device 509, or installed from the storage device 508, or installed from the ROM 502. When the computer program is executed by the processing device 501, the above functions defined in the method of the embodiment of the present disclosure are executed.
[0141] The names of the messages or information interacted between multiple devices in the embodiments of the present disclosure are only for illustrative purposes and are not used to limit the scope of these messages or information.
[0142] The electronic device provided by the embodiment of the present disclosure and the interaction method provided by the above embodiment belong to the same inventive concept. Technical details not described in detail in this embodiment may be referred to in the above embodiment, and this embodiment has the same beneficial effects as the above embodiment.
[0143] The embodiment of the present disclosure provides a computer storage medium, on which a computer program is stored, and when the program is executed by a processor, the interaction method provided by the above embodiment is implemented.
[0144] It should be noted that the above-mentioned computer-readable medium in the present disclosure can be a computer-readable signal medium, a computer-readable storage medium, or any combination of the two. A computer-readable storage medium can be, for example, but not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination of the above. More specific examples of computer-readable storage media can include, but are not limited to: an electrical connection with one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the above. In the present disclosure, a computer-readable storage medium can be any tangible medium that contains or stores a program, which can be used by or in conjunction with an instruction execution system, apparatus, or device. In the present disclosure, a computer-readable signal medium can include a data signal propagated in a baseband or as part of a carrier wave, which carries computer-readable program code. Such a propagated data signal can take various forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination of the above. A computer-readable signal medium can also be any computer-readable medium other than a computer-readable storage medium, which can send, propagate, or transmit a program for use by or in conjunction with an instruction execution system, apparatus, or device. The program code contained on a computer-readable medium can be transmitted using any appropriate medium, including but not limited to: wires, optical cables, RF (radio frequency), etc., or any suitable combination of the above.
[0145] In some embodiments, the client and the server can communicate using any currently known or future-developed network protocol such as HTTP (HyperText Transfer Protocol), and can be interconnected with digital data communication in any form or medium (e.g., a communication network). Examples of communication networks include local area networks (“LANs”), wide area networks (“WANs”), the Internet (e.g., the Internet), and end-to-end networks (e.g., ad hoc end-to-end networks), as well as any currently known or future-developed networks.
[0146] The above-mentioned computer-readable medium can be included in the above-mentioned electronic device; or it can exist separately without being assembled into the electronic device.
[0147] The above computer-readable medium carries one or more programs, which, when executed by the electronic device, cause the electronic device to: obtain audio information corresponding to a first object in a live broadcast room; and obtain target interaction information corresponding to the target audio when the audio information corresponding to the first object includes target audio and a target request from a second object corresponding to the live broadcast room is received.
[0148] Computer program code for performing the operations of the present disclosure may be written in one or more programming languages or combinations thereof. The programming languages include, but are not limited to, object-oriented programming languages such as Java, Smalltalk, C++, and also include conventional procedural programming languages such as the "C" language or similar programming languages. The program code may execute entirely on the user's computer, partially on the user's computer, as a stand-alone software package, partially on the user's computer and partially on a remote computer, or entirely on the remote computer or server. In the case of a remote computer, the remote computer may be connected to the user's computer through any type of network, including a local area network (LAN) or a wide area network (WAN), or may be connected to an external computer (e.g., through the Internet using an Internet service provider).
[0149] The flowcharts and block diagrams in the accompanying drawings illustrate the possible architectures, functions, and operations of systems, methods, and computer program products according to various embodiments of the present disclosure. In this regard, each block in the flowchart or block diagram may represent a module, a program segment, or a part of code that contains one or more executable instructions for implementing a specified logical function. It should also be noted that, in some alternative implementations, the functions marked in the blocks may occur in a different order than that marked in the accompanying drawings. For example, two consecutive blocks shown may actually be executed substantially in parallel, and they may sometimes be executed in the reverse order, depending on the functions involved. It should also be noted that each block in the block diagram and / or flowchart, and the combinations of blocks in the block diagram and / or flowchart, may be implemented by a dedicated hardware-based system that performs the specified functions or operations, or may be implemented by a combination of dedicated hardware and computer instructions.
[0150] The above description is only a preferred embodiment of the present disclosure and an explanation of the applied technical principles. Those skilled in the art should understand that the scope of the disclosure involved in the present disclosure is not limited to the technical solutions formed by the specific combination of the above technical features, and should also cover other technical solutions formed by any combination of the above technical features or their equivalent features without departing from the above disclosure concept. For example, the technical solutions formed by mutually replacing the above features with the technical features (but not limited to) having similar functions disclosed in the present disclosure.
[0151] In addition, although the operations are depicted in a particular order, this should not be construed as requiring that the operations be performed in the particular order shown or in sequential order. In certain environments, multitasking and parallel processing may be advantageous. Similarly, although several specific implementation details are included in the above discussion, these should not be construed as limiting the scope of the present disclosure. Certain features described in the context of separate embodiments may also be implemented in combination in a single embodiment. Conversely, the various features described in the context of a single embodiment may also be implemented separately or in any suitable sub-combination in multiple embodiments.
[0152] Although the subject matter has been described in language specific to structural features and / or methodological act logic, it should be understood that the subject matter defined in the appended claims is not necessarily limited to the specific features or acts described above. Rather, the specific features and acts described above are merely example forms for implementing the claims.
Claims
1. An interaction method, characterized in that, Including: Obtaining audio information corresponding to a first object in a live broadcast room; When the audio information corresponding to the first object includes target audio and a target request from a second object corresponding to the live broadcast room is received, obtaining target interaction information corresponding to the target audio.
2. The method according to claim 1, wherein It further includes: Sending the audio information including the target audio to a first server, so that the first server converts the audio information including the target audio into a text to be processed; Receiving the text to be processed sent by the first server, and sending the text to be processed and operation information corresponding to the target request to a second server, so that the second server determines the target interaction information based on the operation information and the text to be processed.
3. The method according to claim 2, characterized in that, After obtaining the target interaction information corresponding to the target audio, it further includes: Sending the target interaction information to a third server, so that the third server pushes the target interaction information to at least one client.
4. The method according to claim 2, wherein The sending the text to be processed and the operation information corresponding to the target request to the second server includes: Sending the text to be processed and the object identifier of the second object corresponding to the operation information to the second server.
5. The method according to claim 2, wherein The method further includes: Preprocessing the text to be processed based on the second server to obtain a text to be matched; Performing a matching process on the text to be matched and the object identifier based on the second server to obtain target interaction information.
6. The method according to claim 5, characterized in that, The preprocessing the text to be processed based on the second server to obtain the text to be matched includes: Based on the second server, representing the first type of text in the text to be processed in a first preset manner, and representing the second type of text in the text to be processed in a second preset manner to obtain an adjusted text; and Adjusting the third type of text in the adjusted text to obtain the text to be matched.
7. The method according to claim 5, characterized in that, The performing a matching process on the text to be matched and the object identifier based on the second server to obtain target interaction information includes: Based on the second server, determining a to-be-modified object identifier corresponding to the second object in the text to be matched; Obtaining a target object identifier according to the to-be-modified object identifier and the object identifier; Replacing the to-be-modified object identifier in the text to be matched based on the target object identifier to obtain the target interaction information.
8. The method according to claim 4, characterized in that The sending the target interaction information to a third server, so that the third server pushes the target interaction information to at least one client includes: Sending the target interaction information and the audio information corresponding to the target interaction information to the third server; Based on the third server, pushing the target interaction information and the audio information to the client corresponding to the second object to display the target interaction information on the interactive public screen of the client.
9. The method according to claim 1, characterized in that The target request includes at least one of a request to enter the live broadcast room, a resource transfer request, and a resource sharing request, and the target audio includes at least one of a welcome audio, a thank-you audio, and a sharing audio.
10. The method according to claim 1, characterized in that The method further includes: When the interaction content in the target interaction information is detected, a target interface corresponding to the interaction content is displayed, where the interaction content includes an object identifier.
11. The method according to claim 10, wherein The target interface includes the main interface or the conversation interface of the second object.
12. An interactive device, characterized in that, It includes: An audio information acquisition module, configured to acquire audio information corresponding to a first object in a live broadcast room; An interaction information acquisition module, configured to acquire target interaction information corresponding to the target audio when the audio information corresponding to the first object includes the target audio and a target request of a second object corresponding to the live broadcast room is received.
13. An electronic device, characterized in that, The electronic device includes: One or more processors; A storage device, configured to store one or more programs, When the one or more programs are executed by the one or more processors, the one or more processors implement the interaction method according to any one of claims 1-11.
14. A storage medium containing computer-executable instructions, where the computer-executable instructions are used to execute the interaction method according to any one of claims 1-11 when executed by a computer processor.