Data Processing Method, Device and Server for Live Streaming Room
The cloud server handles user interaction requests, determines the target user language and processes interactive data, solving the problem of user interaction in different languages, improving the user experience, and providing a secure and private communication channel.
Patent Information
- Application Number
- CN202211706177.4
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2022-12-26
- Publication Date
- 2025-06-24
- Estimated Expiration
- 2042-12-26
AI Technical Summary
The existing live broadcast room data processing methods cannot effectively realize efficient communication and interaction between users in different languages, and cannot provide private and secure interaction channels, resulting in poor user interaction experience.
The cloud server receives the interaction request initiated by the user terminal, determines the language and terminal of the target user, applies preset processing rules to process the interactive data, and displays the processed interactive data through the live broadcast interface of the target user terminal.
It realizes efficient and convenient communication and interaction between users in different languages, improves the user's interactive experience, and ensures the security and privacy of communication by establishing a private data communication channel.
Smart Images

Figure CN116055756B_ABST
Abstract
Description
Technical Field
[0001] This application belongs to the technical field of cloud computing, and particularly relates to a method, device, and server for processing data in a live broadcast room. Background Art
[0002] With the rise and development of the video live broadcast industry, more and more activities such as cross-regional exhibitions are starting to be held in the form of online video live broadcasts.
[0003] Holding activities such as cross-regional exhibitions in the form of online video live broadcasts can provide many conveniences for users on the one hand. On the other hand, limited by the existing data processing methods in the live broadcast room, users in different regions cannot communicate and interact well with each other in the same live broadcast room due to the different languages they use; in addition, the existing live broadcast rooms cannot provide users with a relatively private and secure interaction channel for communicating and interacting involving key information, resulting in a relatively poor interaction experience for users.
[0004] In response to the above problems, no effective solution has been proposed yet. Summary of the Invention
[0005] This application provides a method, device, and server for processing data in a live broadcast room, which can enable viewer users to communicate and interact diversely with target user objects using different languages in the live broadcast room efficiently and conveniently, and improve the interaction experience of viewer users.
[0006] This application provides a method for processing data in a live broadcast room, which is applied to a cloud server and includes:
[0007] Receiving an interaction request initiated by a first user terminal; wherein, the interaction request carries at least interaction data and an object identifier of a target user object for which the interaction data is directed;
[0008] Determining a target language matching the target user and a target user terminal held by the target user object according to the object identifier of the target user object;
[0009] Determining a matching target processing rule from a preset set of processing rules according to the data type of the interaction data;
[0010] Processing the interaction data based on the target language according to the target processing rule to obtain processed interaction data;
[0011] Displaying the processed interaction data in the current live video displayed to the target user object through the live broadcast room interface of the target user terminal.
[0012] In one embodiment, the data type of the interaction data includes at least one of the following: text data, voice data, and emoticon images.
[0013] In one embodiment, when the data type of the interactive data includes voice data, based on the target language and according to the target processing rules, the interactive data is processed to obtain the processed interactive data, including:
[0014] Performing speech recognition on the voice data using a speech recognition model to obtain corresponding text data; and extracting the voice features of the first audience user from the voice data;
[0015] Processing the text data using a translation model matching the target language to obtain text data based on the target language;
[0016] Using a preset speech synthesis model to process the text data based on the target language according to the voice features of the first audience user to obtain corresponding synthesized voice data as the processed interactive data.
[0017] In one embodiment, when the data type of the interactive data includes expression images, based on the target language and according to the target processing rules, the interactive data is processed to obtain the processed interactive data, including:
[0018] Performing character detection on the expression image to determine whether there are meaningful text characters in the expression image;
[0019] When it is determined that there are meaningful text characters in the expression image, performing recognition processing on the expression image using an image recognition model to extract the text characters in the expression image as text data;
[0020] Processing the text data using a translation model matching the target language to obtain text data based on the target language;
[0021] Constructing annotation data for the expression image according to the text data based on the target language;
[0022] Combining the expression image and the annotation data to obtain an expression image carrying the annotation data as the processed interactive data.
[0023] In one embodiment, when the data type of the interactive data includes text data, based on the target language and according to the target processing rules, the interactive data is processed to obtain the processed interactive data, including:
[0024] Processing the text data using a translation model matching the target language to obtain text data based on the target language;
[0025] Constructing corresponding bullet screen data according to the text data in the target language as the processed interactive data.
[0026] In one embodiment, the method further includes:
[0027] Receiving a private message request initiated by a first user terminal; wherein, the private message request carries at least an object identifier of a target user object;
[0028] Responding to the private message request, and establishing a privacy data communication channel based on the live broadcast room interface between the first user terminal and the target user terminal according to relevant encryption communication protocols; wherein, the cloud server displays the processed interaction data in the current live video shown to the target user object through the live broadcast room interface of the target user terminal based on this privacy data communication channel.
[0029] In one embodiment, after establishing the privacy data communication channel based on the live broadcast room interface between the first user terminal and the target user terminal, the method further includes:
[0030] Encrypting the processed interaction data by using the public key data corresponding to the target user terminal to obtain ciphertext data of the interaction data;
[0031] Sending the ciphertext data of the interaction data to the target user terminal through the privacy data communication channel; wherein, the target user terminal decrypts the ciphertext data of the interaction data by using the private key data to obtain the processed interaction data; and displays the processed interaction data in the current live video shown to the target user object through the live broadcast room interface.
[0032] In one embodiment, before receiving the interaction request initiated by the first user terminal, the method further includes:
[0033] Receiving a connection request initiated by the first user terminal;
[0034] Responding to the connection request, and establishing a first data connection with the first user terminal;
[0035] Transmitting the live video stream data to the first user terminal through the first data connection; wherein, the first user terminal displays the current live video to the first user through the live broadcast room interface according to the live video stream data.
[0036] In one embodiment, while transmitting the live video stream data to the first user terminal through the first data connection, the method further includes:
[0037] Collecting characteristic parameters of the first user terminal;
[0038] Determining a first language that matches the first viewer user according to the characteristic parameters of the first user terminal;
[0039] Detect whether the live video stream data currently transmitted through the first data connection is live video stream data based on the first language;
[0040] In the case where it is determined that the live video stream data currently transmitted through the first data connection is not live video stream data based on the first language, determine the cloud CDN that caches the live video stream data based on the first language as the first target CDN;
[0041] Switch the first data connection to the first target CDN, so as to transmit the live video stream data based on the first language to the first user terminal through the first data connection.
[0042] In one embodiment, the characteristic parameters of the first user terminal include at least one of the following: the IP address of the first user terminal, the default language parameter of the browser of the first user terminal, and the cookie data of the first user terminal.
[0043] This application also provides a data processing method for a live broadcast room, which is applied to the first user terminal and includes:
[0044] Display the current live video based on the first language to the first viewer user through the live broadcast room interface;
[0045] Receive the interaction data input by the first user through the live broadcast room interface; and determine the target user object for which the interaction data is directed;
[0046] Generate a corresponding interaction request according to the interaction data; wherein, the interaction request also carries the object identifier of the target user object;
[0047] Send the interaction request to the cloud server; wherein, the cloud server determines the target language that matches the target user object according to the object identifier of the target user object; and processes the interaction data based on the target language to obtain the processed interaction data; the cloud server also displays the processed interaction data in the current live video displayed to the target user object through the live broadcast room interface of the target user terminal.
[0048] This application also provides a data processing device for a live broadcast room, which is applied to the cloud server and includes:
[0049] A receiving module, configured to receive an interaction request initiated by the first user terminal; wherein, the interaction request carries at least the interaction data and the object identifier of the target user object for which the interaction data is directed;
[0050] A first determination module, configured to determine the target language that matches the target user and the target user terminal held by the target user object according to the object identifier of the target user object;
[0051] A second determination module, configured to determine a matching target processing rule from a preset set of processing rules according to the data type of the interaction data;
[0052] A processing module, configured to process the interaction data based on the target language according to the target processing rule to obtain processed interaction data;
[0053] An outreach module, configured to display the processed interaction data in the current live video shown to the target user object through the live broadcast interface of the target user terminal.
[0054] This application also provides a server, including a processor and a memory for storing instructions executable by the processor. When the processor executes the instructions, the relevant steps of the data processing method for the live broadcast room are implemented.
[0055] This application also provides a computer-readable storage medium, on which computer instructions are stored. When the instructions are executed by a processor, the following steps are implemented: receiving an interaction request initiated by a first user terminal; wherein the interaction request carries at least interaction data and an object identifier of a target user object to which the interaction data is directed; determining a target language matching the target user and a target user terminal held by the target user object according to the object identifier of the target user object; determining a matching target processing rule from a preset set of processing rules according to the data type of the interaction data; processing the interaction data based on the target language according to the target processing rule to obtain processed interaction data; and displaying the processed interaction data in the current live video shown to the target user object through the live broadcast interface of the target user terminal.
[0056] This application also provides a computer program product, including a computer program. When the computer program is executed by a processor, the relevant steps of the data processing method for the live broadcast room are implemented.
[0057] Based on the data processing method, device, and server for a live streaming room provided in this application, after the cloud server of the cloud live streaming service platform receives an interaction request initiated by the first user terminal, it can first determine the target language that matches the target user according to the object identifier of the target user object, as well as the target user terminal held by the target user object; according to the data type of the interaction data, determine the matching target processing rule from the preset processing rule set; then process the interaction data based on the target language according to the target processing rule to obtain the processed interaction data; and display the processed interaction data in the current live video shown to the target user object through the live streaming room interface of the target user terminal. Thus, it can enable the viewer users to communicate and interact diversely with the target user objects using different languages efficiently and conveniently in the live streaming room, improving the interaction experience of the viewer users in the live streaming room. Further, the cloud server can also establish an exclusive private data communication channel between the viewer users and the target user objects in the live streaming room according to the specific needs of the viewer users, and then the viewer users can communicate and interact more securely and privately with the target user objects directly in the live streaming room interface, avoiding the leakage of private data during the communication and interaction process of the viewer users. BRIEF DESCRIPTION OF THE DRAWINGS
[0058] To illustrate the embodiments of this application more clearly, the following will briefly introduce the accompanying drawings required for the embodiments. The accompanying drawings in the following description are only some embodiments recorded in this application. For those of ordinary skill in the art, without creative efforts, other drawings can be obtained based on these drawings.
[0059] Figure 1 is a flowchart showing the data processing method for a live streaming room provided by the embodiments of this application;
[0060] Figure 2 is a schematic diagram of an embodiment applying the data processing method for a live streaming room provided by the embodiments of this application in a scenario example;
[0061] Figure 3 is a schematic diagram of an embodiment applying the data processing method for a live streaming room provided by the embodiments of this application in a scenario example;
[0062] Figure 4 is a schematic diagram of an embodiment applying the data processing method for a live streaming room provided by the embodiments of this application in a scenario example;
[0063] Figure 5 is a schematic diagram of an embodiment applying the data processing method for a live streaming room provided by the embodiments of this application in a scenario example;
[0064] Figure 6 It is a schematic flowchart of a data processing method for a live broadcast room provided by another embodiment of the present application;
[0065] Figure 7 It is a schematic diagram of the structural composition of a server provided by an embodiment of the present application;
[0066] Figure 8 It is a schematic diagram of the structural composition of a data processing device for a live broadcast room provided by an embodiment of the present application;
[0067] Figure 9 It is a schematic diagram of the structural composition of a data processing device for a live broadcast room provided by another embodiment of the present application. Detailed implementation manners
[0068] In order to enable those skilled in the art to better understand the technical solutions in the present application, the following will clearly and completely describe the technical solutions in the embodiments of the present application with reference to the accompanying drawings in the embodiments of the present application. Obviously, the described embodiments are only a part of the embodiments of the present application, rather than all the embodiments. Based on the embodiments in the present application, all other embodiments obtained by those of ordinary skill in the art without creative efforts shall fall within the protection scope of the present application.
[0069] It should be noted that all the information data related to users involved in this specification are obtained and used on the premise that the users are aware of and consent. And the acquisition, storage, use, processing, etc. of the above information data all comply with the relevant regulations of national laws and regulations.
[0070] Refer to Figure 1 As shown, the embodiments of the present application provide a data processing method for a live broadcast room, where the method is specifically applied to the cloud server side. Specifically, the method may include the following content:
[0071] S101: Receive an interaction request initiated by a first user terminal; wherein, the interaction request carries at least interaction data and the object identifier of the target user object to which the interaction data is directed;
[0072] S102: Determine the target language that matches the target user and the target user terminal held by the target user object according to the object identifier of the target user object;
[0073] S103: Determine a matching target processing rule from a preset set of processing rules according to the data type of the interaction data;
[0074] S104: Process the interaction data based on the target language according to the target processing rule to obtain the processed interaction data;
[0075] S105: Display the processed interaction data in the current live video shown to the target user object through the live broadcast room interface of the target user terminal.
[0076] Among them, the above-mentioned first user terminal may specifically be the user terminal held by the first audience user. The above-mentioned first audience user may specifically be understood as any audience user currently online in the live broadcast room.
[0077] The above-mentioned target user object may specifically include a user object that is in the same live broadcast room as the above-mentioned first audience user and with whom the first audience user currently wants to have a separate communication and interaction. Specifically, the above-mentioned target user object may include the host user and / or other audience users in the same live broadcast room as the first audience user. The above-mentioned target user object may include one or more user objects.
[0078] The above-mentioned live video may specifically be a cross-regional exhibition live video, a cross-regional academic conference live video, or a cross-regional commodity exchange live video, etc. Of course, the above-listed live videos are only illustrative. In specific implementation, according to the specific application scenario and processing requirements, the above-mentioned live video may also include other suitable types and content of live videos. This specification does not make any limitations in this regard.
[0079] Based on the above embodiments, the cloud server can first process the interaction data of the first audience user for the target user object in the same live broadcast room into processed interaction data based on the target language that matches the target user object, and then display the above-mentioned processed interaction data in the current live video shown to the target user object through the live broadcast room interface of the target user terminal, so that the audience user can efficiently and conveniently have diverse and language-barrier-free interaction and communication with the target user object such as the host user and / or other audience users in the same live broadcast room, effectively improving the interaction experience of the audience user in the live broadcast room.
[0080] In some embodiments, refer to Figure 2 As shown, the above-mentioned data processing method for the live broadcast room may specifically be applied to the cloud server side.
[0081] Among them, the above cloud server can specifically include a background server applied to one side of the cloud live broadcast service platform, which can realize functions such as data transmission and data processing. Specifically, the cloud server can be, for example, an electronic device with data operation, storage functions, and network interaction functions. Or, the cloud server can also be a software program running in the electronic device, providing support for data processing, storage, and network interaction. In this embodiment, the number of servers included in the cloud server is not specifically limited. The cloud server can specifically be a single server, or several servers, or a server cluster formed by several servers.
[0082] Specifically, the above cloud server can be connected to the live broadcast terminal and multiple user terminals in a wired or wireless manner.
[0083] Among them, the above live broadcast terminal and user terminals can specifically include a front end applied to the host user and audience users, which can realize functions such as data collection and data transmission. Specifically, the live broadcast terminal and user terminals can be, for example, electronic devices such as desktop computers, tablet computers, laptop computers, and smart phones. Or, the live broadcast terminal and user terminals can also be software applications that can run in the above electronic devices. For example, it can be a certain live broadcast APP running on a smart phone, etc.
[0084] Furthermore, multiple algorithm models can be deployed on the above cloud server. For example, a speech recognition model, a translation model, and an image recognition model; in addition, the above cloud server can also be connected to a database, a translation terminal, and multiple cloud CDNs. For example, cloud CDN1, cloud CDN2, cloud CDNn, etc.
[0085] Among them, the above cloud CDN (Content Delivery Network) can specifically refer to a content delivery network based on cloud technology. The above translation terminal can specifically be a manual translation terminal or an automatic translation terminal based on artificial intelligence.
[0086] During specific implementation, refer to Figure 2 As shown, the host user can use the held live broadcast terminal to conduct video live broadcasts such as exhibitions and product promotions. The live broadcast terminal can collect the live video stream data of the host user in real time and upload the live video stream data to the cloud server.
[0087] After receiving the live video stream data, the cloud server can, on the one hand, directly forward the above-mentioned live video stream data to the user terminal so that the viewer user can watch the live video in the language used by the host user through the live room interface of the user terminal; on the other hand, it can also send the above-mentioned live video stream data to the translation terminal to perform real-time translation on the above-mentioned live video stream data through the translation terminal, and then cache the translated live video stream data in different languages to the corresponding cloud CDN so that viewer users in different regions using different languages can pull and watch the live video in their own language relatively synchronously through the user terminal in the same live room.
[0088] Specifically, taking any first viewer user in the live room as an example. The current first viewer user watches the current live video in the first language through the live room interface displayed by the first user terminal held. Among them, the above-mentioned first language can be specifically understood as the language matching the first viewer user. For example, the mother tongue of the first viewer user, etc.
[0089] While watching the current live video through the live room interface displayed by the first user terminal, the first viewer user can also select the host user and / or other viewer users in the live room as the target user object; and specifically communicate and interact with the above-mentioned target user object in the live room interface.
[0090] Specifically, a user object list can also be displayed in the live room interface. Among them, the above-mentioned user object list can specifically contain the object identifiers of user objects such as the host user and online viewer users in the live room. The first viewer user can initiate a click operation in the user object list to specify the target user object to be specifically communicated and interacted with by selecting the object identifiers of one or more user objects.
[0091] After selecting the target user object, an interactive data input box can pop up on the live room interface. Correspondingly, the first viewer user can input specific interactive data through the interactive data input box.
[0092] The first user terminal can receive the above-mentioned interactive data, determine the target user object targeted by the interactive data, obtain the object identifier of the target user object; and then can generate a corresponding interactive request. Among them, the interactive request can carry at least the interactive data and the object identifier of the target user object.
[0093] Next, the first user terminal can send the above-mentioned interactive request to the cloud server. Correspondingly, the cloud server receives and obtains the above-mentioned first interactive request.
[0094] In some embodiments, the data types of the interactive data may specifically include at least one of the following: text data, voice data, expression images, etc. Of course, it should be noted that the above-listed data types of the interactive data are only illustrative. In specific implementation, according to the specific scenario and processing requirements, the above interactive data may also include other suitable types of interactive data. This specification does not make any limitations in this regard.
[0095] Based on the above embodiments, the viewer user can freely use one or more different data types of interactive data to communicate and interact with the target user object in the live broadcast room according to the specific situation, so as to meet the diverse interactive needs of the viewer user.
[0096] In some embodiments, during specific implementation, the cloud server can determine the user terminal held by the target user object as the target user terminal by querying the user database according to the identifier of the target user object.
[0097] Furthermore, the cloud server can also collect the characteristic parameters of the target user terminal according to the data connection established by the target user terminal for obtaining the live video data, and determine the language matching the target user as the target language according to the characteristic parameters of the target user terminal.
[0098] In addition, the cloud server can also query the database according to the identification information of the target user object, and determine the target language matching the target user object according to the user data of the target user object recorded in the database.
[0099] In some embodiments, before specific implementation, corresponding preset processing rules can be respectively configured for different data types of interactive data. Among them, each preset processing rule may include corresponding algorithm rules and algorithm models. Further, a corresponding preset processing rule set can be obtained by combining the above multiple preset processing rules. And, the matching relationship between the preset processing rule and the data type of the interactive data is also stored in the preset processing rule set.
[0100] In some embodiments, during specific implementation, a preset processing rule that matches can be determined from the preset processing rule set according to the data type of the interactive data as the target processing rule.
[0101] In some embodiments, referring to Figure 3 As shown, in the case where the data type of the interactive data includes voice data, when processing the interactive data based on the target language according to the target processing rule to obtain the processed interactive data, the specific implementation may include the following content:
[0102] S1: Use a speech recognition model to perform speech recognition on the speech data to obtain corresponding text data; and extract the speech features of the first audience user from the speech data.
[0103] S2: Use a translation model matching the target language to process the text data to obtain text data based on the target language.
[0104] S3: Use a preset speech synthesis model to process the text data based on the target language according to the speech features of the first audience user to obtain corresponding synthesized speech data as the processed interactive data.
[0105] Based on the above embodiments, the speech data input by the first audience user can be efficiently converted into synthesized speech data based on the target language and conforming to the speech characteristics of the first audience user when speaking at that time, which can more truly and comprehensively reflect the emotional information such as the tone and attitude of the first audience user when speaking at that time. Furthermore, the above synthesized speech data is used as the processed interactive data to reach the target user object, so that the target user object can not only conveniently and efficiently understand the relevant semantic content in the interactive data based on the target language used by itself, but also intuitively feel the real emotional information of the first audience user through the above synthesized speech data, thereby obtaining a relatively good communication and interaction effect.
[0106] Wherein, the speech features include at least one of the following: pitch, loudness, frequency, timbre, etc.
[0107] In specific implementation, when using a speech recognition model to perform speech recognition on speech data to obtain corresponding text data, a speech feature extraction model can also be used to process the speech data to extract the speech features of the first audience user when inputting the speech data. Then, a translation model matching the target language is determined from multiple translation models; and the translation model matching the target language is used to process the text data to obtain translated text data based on the target language. Further, a preset speech synthesis model can be used to perform speech synthesis on the text data based on the target language based on the speech features of the target user, so as to obtain processed interactive data that not only contains semantic content based on the target language but also can convey the real emotional information of the first audience user.
[0108] In some embodiments, refer to Figure 4 As shown, when the data type of the interactive data includes expression images, processing the interactive data based on the target language according to the target processing rules to obtain processed interactive data may specifically include the following contents:
[0109] S1: Perform character detection on the expression image to determine whether there are meaningful text characters in the expression image;
[0110] S2: In the case of determining that there are meaningful text characters in the expression image, use an image recognition model to perform recognition processing on the expression image to extract the text characters in the expression image as text data;
[0111] S3: Process the text data using a translation model matching the target language to obtain text data based on the target language;
[0112] S4: Construct annotation data for the expression image according to the text data based on the target language;
[0113] S5: Combine the expression image and the annotation data to obtain an expression image carrying the annotation data as the processed interaction data.
[0114] Based on the above embodiments, the expression image input by the first viewer user can be efficiently converted into processed interaction data carrying annotation data based on the target language, so that the target user object can conveniently and efficiently understand the true meaning expressed by the expression image sent by the first viewer user.
[0115] During specific implementation, in the case of determining that there are no meaningful text characters in the expression image, the expression image can be left unprocessed and directly determined as the processed interaction data.
[0116] When specifically determining whether there are meaningful text characters in the expression image, text character detection can be first performed on the expression image. In the case of determining that there are text characters in the expression image, then according to a meaningless character reference template matching the first language, it is judged whether there are meaningful text characters among the text characters in the expression image.
[0117] When specifically combining the expression image and the annotation data, a annotation box containing the annotation data can be generated; then the annotation box is spliced with the expression image, so that an expression image carrying the annotation data can be obtained.
[0118] In some embodiments, in the case where the data type of the interaction data includes text data, when processing the interaction data based on the target language according to the target processing rules to obtain the processed interaction data, during specific implementation, the following content may further be included:
[0119] S1: Process the text data using a translation model matching the target language to obtain text data based on the target language;
[0120] S2: Construct corresponding barrage data based on the text data in the target language as the processed interactive data.
[0121] Based on the above embodiments, the text data input by the first viewer user can be efficiently and accurately converted into text data based on the target language, so that the target user object can conveniently and quickly understand the semantic content expressed by the interactive data sent by the first viewer user.
[0122] In some embodiments, considering that in the actual communication and interaction process, the requirement for timeliness is often relatively high, but the requirement for the accuracy of semantic content is not very strict. For example, usually different user objects only need to know the general meaning of each other during communication and interaction. Further, considering that there may be many communication and interactions in the same live broadcast room at the same time. In this case, if a translation terminal with a higher accuracy is called to perform relevant translation processing on the text data, on the one hand, it will cause an excessive data processing volume of the translation terminal, which may even affect the normal processing of the live video stream data; on the other hand, it cannot better meet the timeliness of communication and interaction.
[0123] Based on the above considerations, in this specification, the text data is mainly processed by calling a relatively simplified translation model independent of the translation terminal to ensure that the key information that user objects care about has a higher accuracy when processing the translation of text data, without affecting the basic communication and interaction between user objects, effectively ensuring the timeliness of communication and interaction, reducing the data processing burden, and at the same time avoiding affecting the processing of live video stream data.
[0124] When specifically implemented, the above processing of text data using a translation model matching the target language may further include the following: determining the reception time of the interaction request; according to the initiation time, determining the theme information in the live video when the first viewer user initiates the interaction request through the first user terminal; and then using a translation model matching the target language, based on the theme information as a reference, to process the text data, so as to more efficiently translate the text data on the premise of ensuring that the key information related to the theme information has a higher accuracy, and quickly obtain text data based on the target language that meets the basic requirements of communication and interaction.
[0125] In some embodiments, referring to Figure 5 as shown, when the method is specifically implemented, it may further include the following:
[0126] S1: Receive a private message request initiated by the first user terminal; wherein, the private message request carries at least the object identifier of the target user object;
[0127] S2: In response to the private message request, establish a privacy data communication channel based on the live streaming room interface between the first user terminal and the target user terminal according to the relevant encrypted communication protocol; wherein, the cloud server displays the processed interaction data in the current live video presented to the target user object through the live streaming room interface of the target user terminal based on this privacy data communication channel. Among them, only the first viewer user and the target user object can perceive the processed interaction data.
[0128] Among them, the above privacy data communication channel is different from the ordinary live streaming room bullet screen public screen. Specifically, a privacy dialog box visible only to the first viewer user and the target user object can be additionally displayed in the live streaming room interfaces of the two user terminals that establish the privacy data communication channel. Among them, the above privacy dialog box is only used to display and input the interaction data for communication between the two user terminals with the established privacy data communication channel.
[0129] It should be noted that the processed interaction data displayed in the current live video presented to the target user object through the live streaming room interface based on the above privacy data communication channel can only be perceived by the first viewer user and the target user object, and other user objects in the live streaming room cannot perceive it.
[0130] Based on the above embodiments, in some live streaming scenarios such as business exhibitions, the first viewer user can also establish an independent and private privacy data communication channel with the target user object in the live streaming room. Then, based on this privacy data communication channel, high-security communication and interaction can be carried out without the perception of other user objects and without affecting the first viewer user and the target user object from watching the current live video, avoiding the leakage of privacy data involved in the communication process.
[0131] In some embodiments, after establishing the privacy data communication channel based on the live streaming room interface between the first user terminal and the target user terminal, when the method is specifically implemented, it may further include the following content:
[0132] S1: Encrypt the processed interaction data using the public key data corresponding to the target user terminal to obtain the ciphertext data of the interaction data;
[0133] S2: Send the ciphertext data of the interaction data to the target user terminal through the privacy data communication channel; wherein, the target user terminal decrypts the ciphertext data of the interaction data using the private key data to obtain the processed interaction data; and displays the processed interaction data in the current live video presented to the target user object through the live streaming room interface.
[0134] In specific implementation, after establishing a privacy data communication channel based on the live broadcast interface between the first user terminal and the target user terminal, the cloud server can also interact with the target user terminal according to relevant encryption communication protocols, and use the terminal identifier and random number generator of the target user terminal to generate public key data and private key data corresponding to the target user terminal; moreover, the cloud server stores the public key data therein, and the target user terminal stores the private key data therein. Similarly, the cloud server can also interact with the first user terminal according to relevant encryption communication protocols, and use the terminal identifier and random number generator of the first user terminal to generate public key data and private key data corresponding to the first user terminal; moreover, the cloud server stores the public key data therein, and the first user terminal stores the private key data therein.
[0135] Based on the above embodiments, it is possible to more effectively avoid the leakage of privacy data involved when the first user terminal and the target user terminal communicate and interact through the privacy data communication channel, and better protect the data security during the communication and interaction.
[0136] In some embodiments, before receiving the interaction request initiated by the first user terminal, when the method is specifically implemented, the following content may further be included:
[0137] S1: Receive a connection request initiated by the first user terminal;
[0138] S2: Respond to the connection request and establish a first data connection with the first user terminal;
[0139] S3: Transmit the live video stream data to the first user terminal through the first data connection; wherein, the first user terminal displays the current live video to the first user through the live broadcast interface according to the live video stream data.
[0140] Based on the above embodiments, the cloud server can quickly respond to the connection request initiated by the first user terminal. By establishing and using the first data connection, the first user terminal can efficiently and stably obtain and display the real-time live video.
[0141] In some embodiments, while transmitting the live video stream data to the first user terminal through the first data connection, when the method is specifically implemented, the following content may further be included:
[0142] S1: Collect the characteristic parameters of the first user terminal;
[0143] S2: Determine the first language that matches the first audience user according to the characteristic parameters of the first user terminal;
[0144] S3: Detect whether the live video stream data currently transmitted through the first data connection is live video stream data based on the first language;
[0145] S4: In the case where it is determined that the live video stream data currently transmitted through the first data connection is not the live video stream data based on the first language, determine the cloud CDN that caches the live video stream data based on the first language as the first target CDN;
[0146] S5: Switch the first data connection to the first target CDN to transmit the live video stream data based on the first language to the first user terminal through the first data connection.
[0147] Based on the above embodiments, the cloud server can automatically detect and identify the first language that matches the first viewer user, and then, in the case of actively determining that the live video currently provided to the first viewer user does not match the first user, automatically switch in a timely manner to the live video based on the first language that matches the first viewer user, so that the first viewer user can obtain a better interaction experience in the live broadcast room.
[0148] In some embodiments, the feature parameter package of the above first user terminal may specifically include at least one of the following: the IP address of the first user terminal, the default language parameter of the browser of the first user terminal, the cookie data of the first user terminal, etc.
[0149] Based on the above embodiments, the cloud server can accurately and automatically determine the language type that matches the first user by collecting and according to the feature parameters of the first user terminal.
[0150] Specifically, the cloud server can collect the feature parameters of the first user terminal through the first data connection.
[0151] In some embodiments, when the method is specifically implemented, the following content may further be included:
[0152] S1: Obtain the live video stream data collected in real time;
[0153] S2: Call the corresponding translation terminal to convert the live video stream data into video stream data based on multiple different languages; and cache the video stream data based on multiple different languages into the corresponding cloud CDNs respectively.
[0154] In some embodiments, after transmitting the live video stream data to the first user terminal through the first data connection, when the method is specifically implemented, the following content may further be included:
[0155] S1: Receive the language switching request initiated by the first user terminal;
[0156] S2: Determine the second language customized by the first viewer user according to the language switching request;
[0157] S3: Determine the cloud CDN that caches the live video stream data in the second language as the second target CDN;
[0158] S4: Switch the first data connection to the second target CDN to transmit the live video stream data in the second language to the first user terminal through the first data connection.
[0159] As can be seen from the above, based on the data processing method for the live broadcast room provided in the embodiments of the present application, after receiving the interaction request initiated by the first user terminal, the cloud server can first determine the target language that matches the target user according to the object identifier of the target user object, and the target user terminal held by the target user object; according to the data type of the interaction data, determine the matching target processing rule from the preset processing rule set; then process the interaction data based on the target language according to the target processing rule to obtain the processed interaction data; display the processed interaction data in the current live video displayed on the live broadcast room interface of the target user terminal to the target user object. Thus, it can enable the audience users to efficiently and conveniently conduct diverse communication and interaction with the target user objects using different languages in the live broadcast room, improve the interaction experience of the audience users in the live broadcast room, and meet the diverse interaction needs of the audience users. Further, the cloud server can also establish an exclusive private data communication channel between the audience users and the target user objects in the live broadcast room according to the needs of the audience users. Furthermore, the audience users can directly conduct relatively secure and private communication and interaction with the target user objects in the live broadcast room interface based on the above private data communication channel, avoiding the leakage of the private data involved in the communication and interaction of the audience users.
[0160] Refer to Figure 6 As shown, the embodiments of the present application also provide another data processing method for the live broadcast room, which is applied to the first user terminal. Among them, when the method is specifically implemented, it may include the following contents:
[0161] S601: Display the current live video in the first language to the first audience user through the live broadcast room interface;
[0162] S602: Receive the interaction data input by the first user through the live broadcast room interface; and determine the target user object targeted by the interaction data;
[0163] S603: Generate a corresponding interaction request according to the interaction data; wherein, the interaction request also carries the object identifier of the target user object;
[0164] S604: Send the interaction request to the cloud server; wherein, the cloud server determines a target language matching the target user object according to the object identifier of the target user object, processes the interaction data based on the target language to obtain processed interaction data; the cloud server also displays the processed interaction data in the current live video shown to the target user object through the live broadcast room interface of the target user terminal.
[0165] As can be seen from the above, the data processing method for the live broadcast room provided by the embodiments of the present application enables the viewer user to efficiently and conveniently conduct diverse communication and interaction with the target user object using different languages in the live broadcast room, improving the interaction experience of the viewer user. Further, a dedicated private data communication channel between the viewer user and the target user object can be established for the viewer user in the live broadcast room according to the needs of the viewer user. Then, the viewer user can directly conduct relatively secure and private communication and interaction with the target user object based on the above private data communication channel in the live broadcast room interface, avoiding the leakage of the privacy data of the viewer user.
[0166] The embodiments of the present application also provide a server, including a processor and a memory for storing processor-executable instructions. When specifically implemented, the processor can execute the following steps according to the instructions: receive an interaction request initiated by a first user terminal; wherein, the interaction request carries at least interaction data and the object identifier of the target user object to which the interaction data is directed; determine a target language matching the target user and the target user terminal held by the target user object according to the object identifier of the target user object; determine a matching target processing rule from a preset set of processing rules according to the data type of the interaction data; process the interaction data based on the target language according to the target processing rule to obtain processed interaction data; display the processed interaction data in the current live video shown to the target user object through the live broadcast room interface of the target user terminal.
[0167] To be able to complete the above instructions more accurately, refer to Figure 7 As shown, the embodiments of the present application also provide another specific server. Wherein, the server includes a network communication port 701, a processor 702, and a memory 703. The above structures are connected by internal cables so that each structure can perform specific data interactions.
[0168] Among them, the network communication port 701 is specifically used to receive an interaction request initiated by a first user terminal; wherein, the interaction request carries at least interaction data and the object identifier of the target user object to which the interaction data is directed.
[0169] The processor 702 can be specifically configured to determine a target language that matches the target user and the target user terminal held by the target user object according to the object identifier of the target user object; determine a matching target processing rule from a preset set of processing rules according to the data type of the interaction data; process the interaction data based on the target language according to the target processing rule to obtain processed interaction data; and display the processed interaction data in the current live video displayed to the target user object through the live broadcast interface of the target user terminal.
[0170] The memory 703 can be specifically configured to store corresponding instruction programs.
[0171] In this embodiment, the network communication port 701 can be bound to different communication protocols, so as to send or receive different data. For example, the network communication port can be a port responsible for web data communication, can also be a port responsible for FTP data communication, or can also be a port responsible for email data communication. In addition, the network communication port can also be a physical communication interface or communication chip. For example, it can be a wireless mobile network communication chip, such as GSM, CDMA, etc.; it can also be a Wifi chip; it can also be a Bluetooth chip.
[0172] In this embodiment, the processor 702 can be implemented in any suitable manner. For example, the processor can take the form of, for example, a microprocessor or a processor and a computer-readable medium storing computer-readable program code (such as software or firmware) executable by the (micro)processor, logic gates, switches, application specific integrated circuit (ASIC), programmable logic controller, and embedded microcontroller, etc. This specification does not make any limitations.
[0173] In this embodiment, the memory 703 can include multiple levels. In a digital system, as long as it can store binary data, it can be a memory; in an integrated circuit, a circuit with a storage function without a physical form is also called a memory, such as RAM, FIFO, etc.; in a system, a storage device with a physical form is also called a memory, such as a memory stick, TF card, etc.
[0174] An embodiment of the present application further provides a user terminal, including a processor and a memory for storing processor-executable instructions. When specifically implemented, the processor may execute the following steps according to the instructions: display a current live video in a first language to a first viewer user through a live room interface; receive interaction data input by the first user through the live room interface; and determine a target user object targeted by the interaction data; generate a corresponding interaction request according to the interaction data; wherein, the interaction request also carries an object identifier of the target user object; send the interaction request to a cloud server; wherein, the cloud server determines a target language matching the target user object according to the object identifier of the target user object; and processes the interaction data based on the target language to obtain processed interaction data; the cloud server also displays the processed interaction data in the current live video displayed to the target user object through the live room interface of the target user terminal.
[0175] An embodiment of the present application further provides a computer-readable storage medium based on the above data processing method for a live room. The computer-readable storage medium stores computer program instructions, and when the computer program instructions are executed, the following steps are implemented: receive an interaction request initiated by a first user terminal; wherein, the interaction request carries at least interaction data and an object identifier of a target user object targeted by the interaction data; determine a target language matching the target user and a target user terminal held by the target user object according to the object identifier of the target user object; determine a matching target processing rule from a preset set of processing rules according to the data type of the interaction data; process the interaction data based on the target language according to the target processing rule to obtain processed interaction data; display the processed interaction data in the current live video displayed to the target user object through the live room interface of the target user terminal.
[0176] In this embodiment, the above storage medium includes but is not limited to a random access memory (RAM), a read-only memory (ROM), a cache, a hard disk drive (HDD), or a memory card. The memory may be used to store computer program instructions. The network communication unit may be set according to standards specified by a communication protocol and is an interface for network connection communication.
[0177] In this embodiment, the functions and effects specifically implemented by the program instructions stored in the computer-readable storage medium may be explained in contrast with other embodiments and will not be elaborated here.
[0178] The embodiment of the present application also provides a computer program product, which includes a computer program. When the computer program is executed by a processor, the following steps are implemented: receiving an interaction request initiated by a first user terminal; wherein the interaction request carries at least interaction data and an object identifier of a target user object to which the interaction data is directed; determining a target language that matches the target user and a target user terminal held by the target user object according to the object identifier of the target user object; determining a matching target processing rule from a preset set of processing rules according to the data type of the interaction data; processing the interaction data based on the target language according to the target processing rule to obtain processed interaction data; and displaying the processed interaction data in the current live video displayed to the target user object through the live broadcast room interface of the target user terminal.
[0179] Referring to Figure 8 As shown, at the software level, the embodiment of the present application also provides a data processing device for a live broadcast room, which is applied to the cloud server side. The device may specifically include the following structural modules:
[0180] A receiving module 801, which may specifically be used to receive an interaction request initiated by a first user terminal; wherein the interaction request carries at least interaction data and an object identifier of a target user object to which the interaction data is directed;
[0181] A first determination module 802, which may specifically be used to determine a target language that matches the target user and a target user terminal held by the target user object according to the object identifier of the target user object;
[0182] A second determination module 803, which may specifically be used to determine a matching target processing rule from a preset set of processing rules according to the data type of the interaction data;
[0183] A processing module 804, which may specifically be used to process the interaction data based on the target language according to the target processing rule to obtain processed interaction data;
[0184] An access module 805, which may specifically be used to display the processed interaction data in the current live video displayed to the target user object through the live broadcast room interface of the target user terminal.
[0185] In some embodiments, the data type of the interaction data may specifically include at least one of the following: text data, voice data, expression images, etc.
[0186] In some embodiments, when the above-mentioned processing module 804 is specifically implemented, in the case where the data type of the interactive data includes voice data, based on the target language and according to the target processing rules, the interactive data is processed to obtain the processed interactive data as follows: Use a speech recognition model to perform speech recognition on the voice data to obtain corresponding text data; and extract the voice features of the first audience user from the voice data; Use a translation model matching the target language to process the text data to obtain text data based on the target language; Use a preset speech synthesis model to process the text data based on the target language according to the voice features of the first audience user to obtain corresponding synthesized voice data as the processed interactive data.
[0187] In some embodiments, when the above-mentioned processing module 804 is specifically implemented, in the case where the data type of the interactive data includes expression images, based on the target language and according to the target processing rules, the interactive data is processed to obtain the processed interactive data as follows: Perform character detection on the expression image to determine whether there are meaningful text characters in the expression image; In the case where it is determined that there are meaningful text characters in the expression image, use an image recognition model to perform recognition processing on the expression image to extract the text characters in the expression image as text data; Use a translation model matching the target language to process the text data to obtain text data based on the target language; According to the text data based on the target language, construct annotation data for the expression image; Combine the expression image and the annotation data to obtain an expression image carrying the annotation data as the processed interactive data.
[0188] In some embodiments, when the above-mentioned processing module 804 is specifically implemented, in the case where the data type of the interactive data includes text data, based on the target language and according to the target processing rules, the interactive data is processed to obtain the processed interactive data as follows: Use a translation model matching the target language to process the text data to obtain text data based on the target language; According to the text data in the target language, construct corresponding bullet screen data as the processed interactive data.
[0189] In some embodiments, when the device is specifically implemented, it can also be used to receive a private message request initiated by a first user terminal; wherein, the private message request carries at least the object identifier of the target user object; Respond to the private message request and establish a privacy data communication channel based on the live broadcast interface between the first user terminal and the target user terminal according to the relevant encryption communication protocol; wherein, the cloud server displays the processed interactive data in the current live video shown to the target user object through the live broadcast interface of the target user terminal based on this privacy data communication channel.
[0190] In some embodiments, after establishing a privacy data communication channel based on the live broadcast room interface between the first user terminal and the target user terminal, when the device is specifically implemented, it can also be used to encrypt the interaction data processed by the public key data corresponding to the target user terminal to obtain the ciphertext data of the interaction data; send the ciphertext data of the interaction data to the target user terminal through the privacy data communication channel; wherein, the target user terminal decrypts the ciphertext data of the interaction data by using the private key data to obtain the processed interaction data; and displays the processed interaction data in the current live video shown to the target user object through the live broadcast room interface.
[0191] In some embodiments, before receiving an interaction request initiated by the first user terminal, when the device is specifically implemented, it can also be used to receive a connection request initiated by the first user terminal; respond to the connection request to establish a first data connection with the first user terminal; transmit the live video stream data to the first user terminal through the first data connection; wherein, the first user terminal displays the current live video to the first user through the live broadcast room interface according to the live video stream data.
[0192] In some embodiments, while transmitting the live video stream data to the first user terminal through the first data connection, when the device is specifically implemented, it can also be used to collect the characteristic parameters of the first user terminal; determine the first language that matches the first audience user according to the characteristic parameters of the first user terminal; detect whether the live video stream data currently transmitted through the first data connection is the live video stream data based on the first language; in the case where it is determined that the live video stream data currently transmitted through the first data connection is not the live video stream data based on the first language, determine the cloud CDN that caches the live video stream data based on the first language as the first target CDN; switch the first data connection to the first target CDN to transmit the live video stream data based on the first language to the first user terminal through the first data connection.
[0193] In some embodiments, the characteristic parameters of the first user terminal may specifically include at least one of the following: the IP address of the first user terminal, the default language parameter of the browser of the first user terminal, the cookie data of the first user terminal, etc.
[0194] See Figure 9 As shown, the embodiment of the present application also provides another data processing method device for a live broadcast room, which is applied to the first user terminal side and may specifically include the following structural modules:
[0195] A display module 901, which can specifically be used to display the current live video based on the first language to the first audience user through the live broadcast room interface;
[0196] The receiving module 902 can be specifically configured to receive the interaction data input by the first user through the live broadcast interface; and determine the target user object targeted by the interaction data;
[0197] The generating module 903 can be specifically configured to generate a corresponding interaction request according to the interaction data; wherein, the interaction request also carries the object identifier of the target user object;
[0198] The sending module 904 can be specifically configured to send the interaction request to the cloud server; wherein, the cloud server determines the target language matching the target user object according to the object identifier of the target user object; and processes the interaction data based on the target language to obtain the processed interaction data; the cloud server also displays the processed interaction data in the current live video shown to the target user object through the live broadcast interface of the target user terminal.
[0199] It should be noted that the units, devices or modules etc. described in the above embodiments can be specifically implemented by computer chips or entities, or by products with certain functions. For the convenience of description, when describing the above devices, they are divided into various modules according to functions for separate description. Of course, when implementing this specification, the functions of each module can be implemented in the same or multiple software and / or hardware, or the modules implementing the same function can be realized by the combination of multiple sub-modules or sub-units, etc. The device embodiments described above are only illustrative. For example, the division of the units is only a logical function division, and there can be other division methods in actual implementation. For example, multiple units or components can be combined or integrated into another system, or some features can be ignored or not executed. Another point, the displayed or discussed coupling or direct coupling or communication connection to each other can be through some interfaces, and the indirect coupling or communication connection of the devices or units can be in electrical, mechanical or other forms.
[0200] As can be seen from the above, based on the data processing device for the live broadcast room provided by the embodiments of the present application, it can enable the viewer user to efficiently and conveniently conduct diverse communication and interaction with the target user object using different languages in the live broadcast room, improving the interaction experience of the viewer user. Further, a dedicated private data communication channel can be established between the viewer user and the target user object in the live broadcast room according to the needs of the viewer user. Then, the viewer user can directly conduct relatively safe and private communication and interaction with the target user object based on the above private data communication channel on the live broadcast interface, avoiding the leakage of the privacy data of the viewer user.
[0201] Although this specification provides method operation steps as described in the embodiments or flowcharts, more or fewer operation steps may be included based on conventional or non-creative means. The order of steps listed in the embodiments is only one way among the execution orders of numerous steps and does not represent the only execution order. When the actual device or client product is executed, it can be executed in the order of the method shown in the embodiments or the drawings or executed in parallel (for example, in a parallel processor or multi-threaded processing environment, or even in a distributed data processing environment). The terms "comprise", "include" or any other variant thereof are intended to cover non-exclusive inclusion, such that a process, method, product or device comprising a series of elements not only includes those elements but also includes other elements not expressly listed, or elements inherent to such process, method, product or device. Without further limitation, there is no exclusion of additional identical or equivalent elements in the process, method, product or device comprising the said elements. The terms such as first, second, etc. are used to denote names and do not denote any particular order.
[0202] As is also known to those skilled in the art, in addition to implementing the controller in the form of pure computer-readable program code, it is entirely possible to logically program the method steps so that the controller can be implemented in the form of logic gates, switches, application-specific integrated circuits, programmable logic controllers, embedded microcontrollers, etc. to achieve the same functions. Therefore, such a controller can be regarded as a hardware component, and the devices included therein for implementing various functions can also be regarded as the structures within the hardware component. Or even, the devices for implementing various functions can be regarded as either software modules for implementing the method or the structures within the hardware component.
[0203] This specification can be described in the general context of computer-executable instructions executed by a computer, such as program modules. Generally, program modules include routines, programs, objects, components, data structures, classes, etc. that perform specific tasks or implement specific abstract data types. This specification can also be practiced in a distributed computing environment where tasks are performed by remote processing devices connected through a communication network. In a distributed computing environment, program modules can be located in local and remote computer-readable storage media including storage devices.
[0204] From the descriptions of the above embodiments, those skilled in the art can clearly understand that this specification can be implemented by means of software plus a necessary general hardware platform. Based on such an understanding, the technical solution of this specification can essentially be embodied in the form of a software product, and this computer software product can be stored in a storage medium, such as ROM / RAM, magnetic disk, optical disk, etc., including several instructions for causing a computer device (which can be a personal computer, mobile terminal, server, or network device, etc.) to execute the methods described in each embodiment or some parts of the embodiments of this specification.
[0205] The embodiments in this specification are described in a progressive manner. For the same or similar parts between the embodiments, reference can be made to each other. The key point of each embodiment is to illustrate the differences from other embodiments. This specification can be used in many general or special computer system environments or configurations. For example: personal computers, server computers, handheld or portable devices, tablet devices, multi-processor systems, microprocessor-based systems, set-top boxes, programmable electronic devices, network PCs, minicomputers, mainframe computers, distributed computing environments including any of the above systems or devices, and so on.
[0206] Although this specification is depicted through embodiments, those of ordinary skill in the art know that this specification has many variations and changes without departing from the spirit of this specification. It is hoped that the appended claims will cover these variations and changes without departing from the spirit of this specification.
Claims
1. A method for processing data in a live broadcast room, characterized in that, Applied to a cloud server, including: Receiving an interaction request initiated by a first user terminal; wherein, the interaction request carries at least interaction data and an object identifier of a target user object to which the interaction data is directed; Determining a target language that matches the target user and a target user terminal held by the target user object according to the object identifier of the target user object; Determining a matching target processing rule from a preset set of processing rules according to the data type of the interaction data; Processing the interaction data based on the target language according to the target processing rule to obtain processed interaction data; Displaying the processed interaction data in the current live video shown to the target user object through the live broadcast room interface of the target user terminal; Before receiving the interaction request initiated by the first user terminal, the method further includes: receiving a connection request initiated by the first user terminal; responding to the connection request to establish a first data connection with the first user terminal; transmitting the live video stream data to the first user terminal through the first data connection; wherein, the first user terminal displays the current live video to the first user through the live broadcast room interface according to the live video stream data; While transmitting the live video stream data to the first user terminal through the first data connection, the method further includes: collecting characteristic parameters of the first user terminal; determining a first language that matches the first viewer user according to the characteristic parameters of the first user terminal; detecting whether the live video stream data currently transmitted through the first data connection is live video stream data based on the first language; in the case where it is determined that the live video stream data currently transmitted through the first data connection is not live video stream data based on the first language, determining a cloud CDN that caches live video stream data based on the first language as a first target CDN; switching the first data connection to the first target CDN to transmit the live video stream data based on the first language to the first user terminal through the first data connection.
2. The method according to claim 1, wherein The data type of the interaction data includes at least one of the following: text data, voice data, and emoticon images.
3. The method according to claim 2, wherein In the case where the data type of the interaction data includes voice data, processing the interaction data based on the target language according to the target processing rule to obtain processed interaction data includes: Performing speech recognition on the voice data using a speech recognition model to obtain corresponding text data; and extracting the speech characteristics of the first viewer user from the voice data; Processing the text data using a translation model that matches the target language to obtain text data based on the target language; Processing the text data based on the target language using a preset speech synthesis model according to the speech characteristics of the first viewer user to obtain corresponding synthesized voice data as the processed interaction data.
4. The method according to claim 2, characterized in that, In the case where the data type of the interaction data includes emoticon images, processing the interaction data based on the target language according to the target processing rule to obtain processed interaction data includes: Performing character detection on the emoticon image to determine whether there are meaningful text characters in the emoticon image; In the case where meaningful text characters are present in the expression image, use an image recognition model to perform recognition processing on the expression image to extract the text characters in the expression image as text data; Use a translation model matching the target language to process the text data to obtain text data based on the target language; Construct annotation data for the expression image according to the text data based on the target language; Combine the expression image and the annotation data to obtain an expression image carrying the annotation data as the processed interaction data.
5. The method according to claim 2, wherein In the case where the data type of the interaction data includes text data, based on the target language and according to the target processing rules, process the interaction data to obtain the processed interaction data, including: Use a translation model matching the target language to process the text data to obtain text data based on the target language; Construct corresponding barrage data according to the text data in the target language as the processed interaction data.
6. The method according to claim 1, characterized in that, The method further includes: Receive a private message request initiated by the first user terminal; wherein, the private message request carries at least the object identifier of the target user object; Respond to the private message request and establish a privacy data communication channel based on the live broadcast interface between the first user terminal and the target user terminal according to the relevant encryption communication protocol; wherein, the cloud server displays the processed interaction data in the current live video shown to the target user object through the live broadcast interface of the target user terminal based on this privacy data communication channel.
7. The method according to claim 6, characterized in that After establishing the privacy data communication channel based on the live broadcast interface between the first user terminal and the target user terminal, the method further includes: Use the public key data corresponding to the target user terminal to encrypt the processed interaction data to obtain the ciphertext data of the interaction data; Send the ciphertext data of the interaction data to the target user terminal through the privacy data communication channel; wherein, the target user terminal decrypts the ciphertext data of the interaction data using the private key data to obtain the processed interaction data; and displays the processed interaction data in the current live video shown to the target user object through the live broadcast interface.
8. The method according to claim 1, characterized in that The characteristic parameters of the first user terminal include at least one of the following: the IP address of the first user terminal, the default language parameter of the browser of the first user terminal, and the cookie data of the first user terminal.
9. A data processing method for a live broadcast room, characterized in that Applied to the first user terminal, it includes: Display the current live video based on the first language to the first viewer user through the live broadcast interface; Receive the interaction data input by the first user through the live broadcast interface; and determine the target user object targeted by the interaction data; Generate a corresponding interaction request according to the interaction data; wherein, the interaction request also carries the object identifier of the target user object; Send the interactive request to the cloud server; wherein, the cloud server determines a target language that matches the target user object according to the object identifier of the target user object; and processes the interactive data based on the target language to obtain the processed interactive data; the cloud server also displays the processed interactive data in the current live video shown to the target user object through the live broadcast room interface of the target user terminal; Wherein, the method further includes: the cloud server receives a connection request initiated by the first user terminal; in response to the connection request, establishes a first data connection with the first user terminal; transmits the live video stream data through the first data connection to the first user terminal; the first user terminal displays the current live video to the first user through the live broadcast room interface according to the live video stream data; and, the cloud server also collects the characteristic parameters of the first user terminal; determines a first language that matches the first viewer user according to the characteristic parameters of the first user terminal; detects whether the live video stream data currently transmitted through the first data connection is the live video stream data based on the first language; in the case where it is determined that the live video stream data currently transmitted through the first data connection is not the live video stream data based on the first language, determines a cloud CDN that caches the live video stream data based on the first language as the first target CDN; switches the first data connection to the first target CDN to transmit the live video stream data based on the first language through the first data connection to the first user terminal.
10. A data processing device for a live broadcast room, characterized in that, Applied to the cloud server, it includes: A receiving module, configured to receive an interactive request initiated by the first user terminal; wherein, the interactive request carries at least interactive data and the object identifier of the target user object to which the interactive data is directed; A first determination module, configured to determine a target language that matches the target user and the target user terminal held by the target user object according to the object identifier of the target user object; A second determination module, configured to determine a matching target processing rule from a preset set of processing rules according to the data type of the interactive data; A processing module, configured to process the interactive data based on the target language according to the target processing rule to obtain the processed interactive data; A reaching module, configured to display the processed interactive data in the current live video shown to the target user object through the live broadcast room interface of the target user terminal; Before receiving the interactive request initiated by the first user terminal, the device is further configured to: receive a connection request initiated by the first user terminal; in response to the connection request, establish a first data connection with the first user terminal; transmit the live video stream data through the first data connection to the first user terminal; wherein, the first user terminal displays the current live video to the first user through the live broadcast room interface according to the live video stream data; While transmitting the live video stream data to the first user terminal through the first data connection, the device is further configured to: collect the characteristic parameters of the first user terminal; determine a first language that matches the first viewer user according to the characteristic parameters of the first user terminal; detect whether the live video stream data currently transmitted through the first data connection is a live video stream data based on the first language; in the case where it is determined that the live video stream data currently transmitted through the first data connection is not a live video stream data based on the first language, determine a cloud CDN that caches the live video stream data based on the first language as the first target CDN; switch the first data connection to the first target CDN, so as to transmit the live video stream data based on the first language to the first user terminal through the first data connection.
11. A server, characterized in that, It includes a processor and a memory for storing processor-executable instructions, and when the processor executes the instructions, it implements the steps of the method according to any one of claims 1 to 8.
12. A computer-readable storage medium, characterized in that, Stored thereon are computer instructions, and when the instructions are executed by a processor, it implements the steps of the method according to any one of claims 1 to 8, or 9.
13. A computer program product, characterized in that, It contains a computer program, and when the computer program is executed by a processor, it implements the steps of the method according to any one of claims 1 to 8, or 9.
Citation Information
Patent Citations
Live broadcast processing method, equipment and device, and storage medium
CN108737845A
Live broadcast interaction method and device, storage medium and electronic equipment
CN112527168A
Information processing method and device, electronic equipment and storage medium
CN113179412A
Live broadcast method and apparatus, and electronic device
CN113301357A