Live streaming interaction method, apparatus, device and medium
The live streaming interaction method enables dynamic content adaptation by switching virtual object scenes based on viewer interactions, enhancing engagement and experience.
Patent Information
- Application Number
- JP2023534896
- Authority / Receiving Office
- JP · JP
- Patent Type
- Patents
- Current Assignee / Owner
- Priority Date
- 2020-12-11
- Filing Date
- 2021-11-09
- Publication Date
- 2025-08-07
- Estimated Expiration
- 2041-11-09
AI Technical Summary
Current live streaming technologies using virtual objects lack viewer interaction, resulting in passive viewing experiences due to pre-configured content.
A live streaming interaction method that allows viewers to influence the content by sending interaction information, triggering a switch from a first scene where the virtual object performs multimedia resources to a second scene where it responds to viewer interactions.
Enhances viewer engagement and interaction experience by allowing virtual objects to dynamically adapt their content based on viewer input, improving the variety and interest of the live streaming session.
Smart Images

Figure 0007720393000001 
Figure 0007720393000002 
Figure 0007720393000003
Abstract
Description
[Technical Field]
[0001] (CROSS-REFERENCE TO RELATED APPLICATIONS) This application claims priority from a Chinese patent application filed on December 11, 2020, bearing application number 202011463601.8 and entitled "Live Streaming Interaction Method, Apparatus, Device and Medium," the entire contents of which are incorporated herein by reference. The present disclosure relates to the field of live streaming technology, and in particular to live streaming interaction methods, apparatus, devices and media. [Background technology]
[0002] With the advancement of live streaming technology, watching live streaming has become an important entertainment activity in people's lives.
[0003] Currently, it is possible to use virtual objects instead of reality streamers for live streaming, but these virtual objects usually only perform live streaming according to pre-configured content, which means viewers can only passively watch and cannot decide what to watch, resulting in poor live streaming results. Summary of the Invention [Problem to be solved by the invention]
[0004] To solve or at least partially solve the above-mentioned technical problems, the present disclosure provides a live streaming interaction method, apparatus, device and medium. [Means for solving the problem]
[0005] An embodiment of the present disclosure provides a live streaming interaction method applied to a plurality of viewer terminals entering a live streaming room of a virtual object, the method comprising: Playing video content of the virtual object in a first live streaming scene on a live streaming interface and displaying interaction information from the plurality of viewer terminals; and in response to the interaction information satisfying a trigger condition, playing video content of the virtual object in a second live streaming scene on the live streaming interface, wherein the live streaming scene is used to indicate a type of live streaming content of the virtual object.
[0006] An embodiment of the present disclosure further provides a live streaming interaction method applied to a service end, the method comprising: receiving interaction information of a plurality of viewer terminals in a first live streaming scene, and determining whether a trigger condition for switching the live streaming scene is satisfied based on the interaction information; and transmitting second video data corresponding to a second live streaming scene to the plurality of viewer terminals when the trigger condition is met, wherein the live streaming scene is used to indicate a type of live streaming content of a virtual object in the live streaming room.
[0007] An embodiment of the present disclosure further provides a live streaming interaction device, the device being provided in a plurality of viewer terminals entering a live streaming room of a virtual object; a first live streaming module for playing video content of the virtual object in a first live streaming scene on a live streaming interface and displaying interaction information from the plurality of viewer terminals; and a second live streaming module for playing video content of the virtual object in a second live streaming scene on the live streaming interface in response to the interaction information satisfying a trigger condition, the live streaming scene being used to indicate a type of live streaming content of the virtual object.
[0008] An embodiment of the present disclosure further provides a live streaming interaction device, the device being provided at a service end, comprising: an information receiving module for receiving interaction information of a plurality of viewer terminals in a first live streaming scene, and determining whether a trigger condition for switching a live streaming scene is satisfied based on the interaction information; and a data transmitting module for transmitting second video data corresponding to a second live streaming scene to the plurality of viewer terminals when the trigger condition is met, wherein the live streaming scene is used to indicate a type of live streaming content of a virtual object in the live streaming room.
[0009] An embodiment of the present disclosure further provides an electronic device, the electronic device including a processor and a memory for storing executable instructions for the processor, the processor reading and executing the executable instructions from the memory to realize the live streaming interaction method provided by the embodiment of the present disclosure.
[0010] An embodiment of the present disclosure further provides a computer-readable storage medium, the storage medium storing a computer program, the computer program being used to perform the live streaming interaction method provided by the embodiment of the present disclosure.
[0011] The technical solution provided by the embodiments of the present disclosure has the following advantages over the prior art: The live streaming interaction solution provided by the embodiments of the present disclosure allows multiple viewer terminals entering a live streaming room of a virtual object to play video content of the virtual object in a first live streaming scene on a live streaming interface, and displays interaction information from the multiple viewer terminals; and, in response to the interaction information satisfying a trigger condition, can play video content of the virtual object in a second live streaming scene on the live streaming interface, where the live streaming scene is used to indicate the type of live streaming content of the virtual object. By adopting the above technical solution, the virtual object can switch from live streaming in the first live streaming scene to live streaming in the second live streaming scene based on the viewer's interaction information, thereby realizing interaction sessions of different live streaming scenes between the virtual object and the viewer, satisfying the various interaction needs of the viewer, improving the variety and interest of the live streaming of the virtual object, and further improving the viewer's interaction experience. [Brief explanation of the drawings]
[0012] These and other features, advantages, and aspects of each embodiment of the present disclosure will become more apparent by reference to the following specific embodiments in conjunction with the accompanying drawings, in which the same or similar reference numerals refer to the same or similar elements throughout the drawings, and in which the drawings are schematic and parts and elements are not necessarily drawn to scale. [Figure 1] 1 is a schematic flowchart of a live streaming interaction method according to an embodiment of the present disclosure. [Figure 2] FIG. 1 is a schematic diagram of live streaming interaction according to an embodiment of the present disclosure. [Figure 3]FIG. 10 is a schematic diagram of another live streaming interaction according to an embodiment of the present disclosure. [Figure 4] 1 is a schematic flowchart of another live streaming interaction method according to an embodiment of the present disclosure. [Figure 5] 1 is a schematic diagram of the configuration of a live streaming interaction device according to an embodiment of the present disclosure; [Figure 6] FIG. 1 is a schematic structural diagram of another live streaming interaction device according to an embodiment of the present disclosure. [Figure 7] 1 is a schematic diagram illustrating the configuration of an electronic device according to an embodiment of the present disclosure. DETAILED DESCRIPTION OF THE INVENTION
[0013] Hereinafter, the embodiments of the present disclosure will be described in more detail with reference to the drawings. Although some embodiments of the present disclosure are shown in the drawings, it should be understood that the present disclosure can be realized in various forms and should not be construed as being limited to the embodiments described herein, but rather these embodiments are provided to provide a deeper and more complete understanding of the present disclosure. It should also be understood that the drawings and embodiments of the present disclosure are used for illustrative purposes only and are not intended to limit the scope of protection of the present disclosure.
[0014] It should be understood that the steps described in the method embodiments of the present disclosure may be performed in different orders and / or in parallel. Also, method embodiments may include additional steps and / or omit the performance of steps that are illustrated. The scope of the present disclosure is not limited in this respect.
[0015] As used herein, the term "comprises" and variations thereof mean open-ended inclusion, i.e., "including but not limited to." The term "based on" means "based at least in part on." The term "in one embodiment" means "at least one embodiment," the term "another embodiment" means "at least one other embodiment," and the term "some embodiments" means "at least some embodiments." Relevant definitions of other terms are provided below.
[0016] It should be noted that the concepts of "first," "second," etc. referred to in this disclosure are used only to distinguish between different devices, modules, or units, and are not intended to limit the order or interdependence of functions performed by these devices, modules, or units.
[0017] It should be noted that the modifications "one" and "multiple" referred to in this disclosure are exemplary rather than limiting, and should be understood as "one or more" unless otherwise indicated herein, as would be understood by one of ordinary skill in the art.
[0018] The names of messages or information exchanged between devices in the embodiments of the present disclosure are not intended to limit the scope of these messages or information, but are for illustrative purposes only.
[0019] 1 is a schematic flowchart of a live streaming interaction method according to an embodiment of the present disclosure, which can be performed by a live streaming interaction device, which can be realized in software and / or hardware and generally can be integrated into an electronic device. As shown in FIG. 1, the method is applied to multiple viewer terminals entering a live streaming room of a virtual object, and includes steps 101 and 102. Step 101: Play video content of a virtual object in a first live streaming scene on a live streaming interface, and display interaction information from multiple viewer terminals.
[0020] The virtual object may be a three-dimensional model created in advance based on artificial intelligence (AI) technology. A controllable digital object may be installed on a computer, and a motion capture device and a facial capture device may be used to obtain the body movements and facial information of a real person to drive the virtual object. The specific types of virtual objects may include a variety of types, and different virtual objects may have different appearances. Specifically, the virtual objects may be virtual animals or virtual characters with different styles. In an embodiment of the present disclosure, by combining artificial intelligence technology with video live streaming technology, the virtual object can realize video live streaming in place of a real person.
[0021] The live streaming interface refers to a page for displaying a live streaming room of a virtual object, which may be a web page or a page in an application program client. The live streaming scene is a scene for indicating the type of live streaming content of a virtual object, and the live streaming scene of a virtual object may include various types, and the live streaming scene in the embodiment of the present disclosure may include a live streaming scene in which a virtual object performs multimedia resources and a live streaming scene in which a virtual object responds to interaction information, and the multimedia resources may include, but are not limited to, books to read, songs to sing, and painting themes.
[0022] In an embodiment of the present disclosure, the first live streaming scene is a live streaming scene in which a virtual object performs a multimedia resource, and the step of playing video content of the virtual object in the first live streaming scene in the live streaming interface can include the steps of: displaying multimedia resource information of a plurality of multimedia resources to be performed in a first area of the live streaming interface; and playing video content in which the virtual object performs a target multimedia resource, where the target multimedia resource is determined based on trigger information of a plurality of viewer terminals for the plurality of multimedia resources.
[0023] Since the multimedia resources can include books to be read, songs to be performed, and themes of paintings, etc., the multimedia resource information of the target to be performed can include books to be read, songs to be sung, and themes of paintings to be drawn, etc. The first area is an area provided in the live streaming interface for displaying the multimedia resource information of the target to be performed, and supports viewers' trigger operations on the multimedia resources. The trigger operations include one or more of clicking, double-tapping, swiping, and voice commands.
[0024] Furthermore, the terminal can receive multimedia resource information of the multiple multimedia resources to be performed sent from the service end, and display the multimedia resource information in a first area of the live streaming interface. Each terminal sends viewer trigger information for the multimedia resources to the service end, and the service end can determine a target multimedia resource from the multiple multimedia resources according to the trigger information, for example, determine the multimedia resource with the most trigger count as the target multimedia resource. The terminal can receive video data of the target multimedia resource sent from the service end, and play video content in which a virtual object performs the multimedia resource on the live streaming interface based on the video data.
[0025] In the above solution, the virtual object can perform a live streaming scene of the multimedia resource according to the viewer's selection, allowing the viewer to decide the content of viewing, improving the level of engagement and further improving the live streaming effect of the virtual object.
[0026] In an embodiment of the present disclosure, the step of playing video content of a virtual object in a first live-streaming scene on a live-streaming interface can include the steps of receiving first video data corresponding to the first live-streaming scene, the first video data including first scene data, first movement data, and first audio data, the first scene data being used to display a live-streaming room background screen in the first live-streaming scene, the first movement data being used to display facial movements and body movements of the virtual object in the first live-streaming scene, and matching the audio data with a target multimedia resource; and playing video content in which the virtual object performs the target multimedia resource in the first live-streaming scene on the live-streaming interface based on the first video data.
[0027] The first video data refers to data pre-arranged by the server for realizing the virtual object to perform live streaming in the first live-streaming scene, and the first video data may include first scene data, first movement data, and first audio data. The scene corresponding to the live-streaming room background screen may include a background scene and a screen field of view scene in the first live-streaming scene of the virtual object. The screen field of view may be the field of view when different lenses photograph the virtual object, and the display size and / or display direction of the scene image corresponding to the different screen field of view may be different. The first movement data may be used to generate facial movements and body movements of the virtual object in the first live-streaming scene. The audio data matches a target multimedia resource among multiple multimedia resources. For example, if the target multimedia resource is a song, the audio data is the audio of the song.
[0028] In an embodiment of the present disclosure, after detecting a viewer's trigger operation on the virtual object, the terminal can obtain first video data corresponding to a first live-streaming scene sent from the service end, generate corresponding video content through decoding the first video data, and play the video content in which the virtual object performs the target multimedia resource in the first live-streaming scene on a live-streaming interface. Furthermore, during the process of playing the video content in which the virtual object performs the target multimedia resource in the first live-streaming scene, the terminal can receive multiple pieces of interaction information from multiple live-streaming viewers and display the multiple pieces of interaction information on the live-streaming interface, where the specific display position can be set according to actual conditions. Optionally, during the process of playing the video content in which the virtual object performs the target multimedia resource in the first live-streaming scene, the live-streaming room background screen and the movement of the virtual object can be switched according to changes in the video content based on the first scene data and the first movement data.
[0029] 2 is a schematic diagram of a live streaming interaction according to an embodiment of the present disclosure. As shown in FIG. 2, a live streaming interface for a first live streaming scene of a virtual object 11 is displayed. The live streaming interface displays a live streaming screen of the virtual object 11 reading a book, with an e-reader placed in front of the virtual object 11 and the virtual object 11 reciting the book. The avatar of the virtual object 11, its name "A-chan", and a follow button 12 are further displayed in the upper left corner of the live streaming interface in FIG. 2.
[0030] Referring to Fig. 2, below the live streaming interface in Fig. 2, interaction information sent from different users watching the live streaming of the virtual object is further displayed, for example, "This story is great" sent from user A (viewer A) in the figure, "Hello" sent from user B (viewer B), and "I'm here for you" sent from user C (viewer C). At the bottom of the live streaming interface, an editing area 13 where the current user sends interaction information, and other function buttons, for example, a selection button 14, an interaction button 15, and an event and reward button 16 in the figure, are further displayed, and different function buttons have different functions.
[0031] Step 102: In response to the interaction information satisfying the trigger condition, play the video content of the virtual object in a second live streaming scene on the live streaming interface, where the live streaming scene is used to indicate the type of the live streaming content of the virtual object.
[0032] The trigger condition refers to a condition for determining whether to switch live streaming scenes based on viewer interaction information. In the embodiments of the present disclosure, the trigger condition can include at least one of the following: the number of interaction information reaches a predetermined threshold; the interaction information includes a first keyword; the number of second keywords in the interaction information reaches a keyword threshold; the duration of the first live streaming scene reaches a preset duration; and the first live streaming scene reaches a preset mark. The predetermined threshold, the first keyword, the second keyword, the keyword threshold, the preset duration, and the preset mark can all be set according to actual circumstances.
[0033] In an embodiment of the present disclosure, playing video content of the virtual object in the second live-streaming scene on the live-streaming interface includes playing video content of the virtual object responding to the interaction information on the live-streaming interface, where the second live-streaming scene is different from the first live-streaming scene and refers to a live-streaming scene in which the virtual object responds to the interaction information.
[0034] Specifically, the terminal receives answer audio data corresponding to one or more pieces of interaction information, and jointly generates answering video content based on the answer audio data and second scene data and second action data in the second live-streaming scene of the virtual object, and can play the video content in which the virtual object answers the interaction information on the live-streaming interface.
[0035] Optionally, the virtual object replies to target interaction information among the interaction information, and the live streaming interaction method may further include displaying the target interaction information and text information replying to the target interaction information in a second area of the live streaming interface.
[0036] The target interaction information refers to one or more pieces of information that require a response, which are determined by the service end based on a preset strategy from multiple pieces of interaction information sent by live streaming viewers. The preset strategy can be set according to actual circumstances, for example, to determine the target interaction information based on the points at which live streaming viewers send interaction information, or to search for target interaction information that matches preset keywords. The preset keywords may be previously extracted by mining according to hotspot information, or may be keywords related to the content of the live streaming. Alternatively, semantic recognition is performed on the interaction information to cluster interaction information with similar expression meanings to obtain a number of information sets, and the set with the most interaction information is the hottest topic for live streaming viewers to interact on, and the interaction information corresponding to this set is the target interaction information. The text information in response to the target interaction information refers to the response text information that matches the target interaction information, which is determined by the service end based on a corpus. The terminal may receive the text information in response to the target interaction information, and display the target interaction information and the text information in response to the target interaction information in a second area of the live streaming interface.
[0037] In the above solution, in the second live streaming scene, the terminal can play video content in which the virtual object responds to the interaction information on the live streaming interface, and display the current interaction information and corresponding response text information, so that the viewer can understand which viewer's interaction content the virtual object is responding to, and further improve the depth of interaction between the viewer and the virtual object, and enhance the interaction dialogue experience.
[0038] In an embodiment of the present disclosure, the step of playing video content of the virtual object in the second live-streaming scene on the live-streaming interface can include the steps of receiving second multimedia data corresponding to the second live-streaming scene, where the second multimedia data includes second scene data, second movement data, and second audio data, where the second scene data is used to show a live-streaming room background screen in the second live-streaming scene, the second movement data is used to show facial movements and body movements of the virtual object in the second live-streaming scene, and the second audio data is generated based on the target interaction information; and playing video content, in which the virtual object responds to the target interaction information in the second live-streaming scene, on the live-streaming interface based on the second multimedia data.
[0039] The second video data refers to data pre-arranged by the server to enable the virtual object to perform live streaming in the second live-streaming scene, and may include second scene data, second motion data, and second audio data. The meaning of each data is similar to that of the first video data, and will not be described in detail here. The difference is that the specific video data in the first live-streaming scene and the second live-streaming scene are different.
[0040] In an embodiment of the present disclosure, if the service end determines, based on the interaction information, that a trigger condition is met, it can send second video data corresponding to a second live-streaming scene to the terminal. After receiving the second video data, the terminal can generate corresponding video content through a decoding process on the second video data, and play the video content in which the virtual object responds to the target interaction information in the second live-streaming scene on the live-streaming interface. Furthermore, the terminal may display interaction information from multiple viewer terminals during the process of playing the video content in which the virtual object responds to the target interaction information in the second live-streaming scene. Optionally, during the process of playing the video content in which the virtual object responds to the target interaction information in the second live-streaming scene, the live-streaming room background screen and the behavior of the virtual object can be switched according to changes in the video content based on the second scene data and the second behavior data, but may be different from the live-streaming room background screen and the behavior of the virtual object in the first live-streaming scene.
[0041] Illustratively, Figure 3 is a schematic diagram of another live streaming interaction according to an embodiment of the present disclosure. As shown in Figure 3, the figure displays a live streaming screen in the process of a virtual object 11 responding to interaction information in a second live streaming scene. Compared with Figure 2, the electronic reader is no longer in front of the virtual object 11. Below the live streaming interface, interaction information sent by different users in the live streaming chat process is further displayed, such as "I miss you" sent by user A (viewer A), "Hello" sent by user B (viewer B), and "Let's chat" sent by user C (viewer C) in the figure.
[0042] The live streaming page in FIG. 3 further displays a second area 17. The second area 17 may display interaction information from a current viewer and text information responding to the virtual object's interaction information, allowing the viewer to understand which viewer's dialogue the virtual object is responding to. For example, the interaction information in the figure is "Let's chat," sent by viewer C, and the virtual object's response text is "It's getting late, let's do it tomorrow." The response text corresponds to the response audio data and matches the conversation content when the virtual object responded. Referring to FIGS. 2 and 3, the virtual object 11 behaves differently in FIGS. 2 and 3. In the first live streaming scene in FIG. 2, the virtual object 11 rests its chin on its left hand, while in the second live streaming scene in FIG. 3, the virtual object 11 raises its left hand and rests its chin on its right hand.
[0043] It should be noted that the first live-streaming scene is a live-streaming scene in which a virtual object plays a multimedia resource, and the second live-streaming scene is a live-streaming scene in which a virtual object responds to interaction information. The first and second live-streaming scenes can also be interchanged, that is, the first live-streaming scene can be a live-streaming scene in which a virtual object responds to interaction information, and the second live-streaming scene can be a live-streaming scene in which a virtual object plays a multimedia resource, but are not limited thereto. In addition, the first and second live-streaming scenes can be continuously alternated, so that the live-streaming scenes of the virtual object can be continuously switched.
[0044] In the embodiments of the present disclosure, live streaming of virtual objects in different live streaming scenes can be realized, the live streaming scenes can be switched according to the viewer's selection, and the live streaming room background screen and the behavior of virtual objects in different live streaming scenes can be made different, so as to meet the various interaction needs of viewers.
[0045] In a live streaming interaction solution provided by an embodiment of the present disclosure, a plurality of viewer terminals entering a live streaming room of a virtual object can play video content of the virtual object in a first live streaming scene on a live streaming interface, and display interaction information from the plurality of viewer terminals. In response to the interaction information satisfying a trigger condition, the live streaming interface can play video content of the virtual object in a second live streaming scene, where the live streaming scene is used to indicate the type of live streaming content of the virtual object. By adopting the above technical solution, the virtual object can realize live streaming switching from the first live streaming scene to live streaming in the second live streaming scene based on the viewer's interaction information, realizing interaction sessions of different live streaming scenes between the virtual object and the viewer to meet the various interaction needs of the viewer, improving the variety and interest of the live streaming of the virtual object, and further improving the viewer's interaction experience.
[0046] 4 is a schematic flowchart of another live streaming interaction method according to an embodiment of the present disclosure, which is based on the above embodiment and further optimizes the above live streaming interaction method. As shown in FIG. 4, the method is applied to the service end and includes steps 201-202. Step 201: receiving interaction information of a plurality of viewer terminals in a first live streaming scene, and determining whether a trigger condition for switching the live streaming scene is met based on the interaction information;
[0047] The live streaming scene is a scene for indicating the type of live streaming content of the virtual object, and the live streaming scene of the virtual object can include various types. In the embodiment of the present disclosure, the live streaming scene can include a live streaming scene in which the virtual object performs multimedia resources and a live streaming scene in which the virtual object responds to interaction information. The multimedia resources can include, but are not limited to, books to read, songs to sing, painting themes, etc. The interaction information refers to interaction text information sent by multiple viewers watching the live streaming in the first live streaming scene via their terminals.
[0048] Specifically, the service end can receive interaction information from multiple viewer terminals in the first live-streaming scene and determine whether a trigger condition for switching the live-streaming scene is met based on the interaction information and / or related information of the first live-streaming scene. In the embodiments of the present disclosure, the trigger condition can include at least one of: the number of pieces of interaction information reaches a predetermined threshold; the interaction information includes a first keyword; the number of second keywords in the interaction information reaches a keyword threshold; the duration of the first live-streaming scene reaches a preset duration; and the first live-streaming scene reaches a preset mark. The predetermined threshold, the first keyword, the second keyword, the keyword threshold, the preset duration, and the preset mark can all be set according to actual circumstances.
[0049] In an embodiment of the present disclosure, the first live-streaming scene is a live-streaming scene in which a virtual object performs a multimedia resource, and the live-streaming interaction method may further include: searching an audio database for first audio data matching the target multimedia resource; searching a virtual object movement database for first movement data for indicating facial movements and body movements of the virtual object in the first live-streaming scene corresponding to the target multimedia resource; determining, based on a scene identifier of the first live-streaming scene, first scene data for indicating a live-streaming room background screen in the first live-streaming scene; combining the first movement data, the first audio data, and the first scene data into first video data corresponding to the first live-streaming scene; and transmitting the first video data to multiple viewer terminals.
[0050] The audio database and the virtual object movement database may be pre-installed databases. The target multimedia resource is one of the multiple multimedia resources. The scene identifier refers to an identifier for distinguishing different live streaming scenes, and the service end may pre-install corresponding scene data for different live streaming scenes. The service end may search the audio database and the virtual object movement database to determine first audio data and first movement data matching the target multimedia resource, and further determine corresponding first scene data based on the scene identifier of the first live streaming scene. Then, the service end may combine the first movement data, first audio data, and first scene data to obtain first video data, and transmit the first video data to multiple viewer terminals.
[0051] After receiving the first video data, the viewer terminal can generate corresponding video content through decoding processing of the first video data, and play the video content in which the virtual object performs the target multimedia resource in the first live-streaming scene on the live-streaming interface. During the process of playing the video content in which the virtual object performs the target multimedia resource in the first live-streaming scene, the live-streaming room background screen and the movement of the virtual object can be switched according to changes in the video content based on the first scene data and the first movement data.
[0052] In an embodiment of the present disclosure, the live streaming interaction method may further include receiving trigger information for a plurality of multimedia resources displayed in a first live streaming scene from a plurality of viewer terminals, and determining a target multimedia resource from the plurality of multimedia resources based on the trigger information. The trigger information may be related information corresponding to a viewer's trigger operation on the multimedia resource, for example, the trigger information may include a trigger count, a trigger time, etc.
[0053] The viewer terminal can display multimedia resource information of multiple multimedia resources on a live streaming interface, receive trigger operations of the viewer on the multimedia resources, and send the trigger information of the multimedia resources to the service end. After receiving the trigger information, the service end can determine a target multimedia resource from the multiple multimedia resources, for example, determine the multimedia resource with the most trigger count as the target multimedia resource.
[0054] Step 202: if the trigger condition is met, send second video data corresponding to a second live-streaming scene to multiple viewer terminals, where the live-streaming scene is used to indicate the type of live-streaming content of the virtual object in the live-streaming room.
[0055] In an embodiment of the present disclosure, the trigger condition is determined by at least one of the following methods: when the number of similar interaction information among the interaction information reaches a predetermined threshold, the trigger condition is satisfied, and the similar interaction information is interaction information whose similarity is greater than a similarity threshold; when keywords in the interaction information are extracted and matched with first keywords and / or second keywords in a keyword database, the trigger condition is satisfied when the interaction information contains the first keyword and / or the number of second keywords in the interaction information reaches a keyword threshold; when the duration of the first live streaming scene reaches a preset duration; and when the first live streaming scene reaches a preset mark point.
[0056] Specifically, the service end performs semantic recognition on the interaction information and clusters interaction information whose similarity is greater than a similarity threshold, which may be referred to as similar interaction information. When the number of similar interaction information reaches a predetermined threshold, it may determine that the trigger condition for switching the live streaming scene is met. And / or, the service end may extract keywords from the interaction information based on a vocabulary and match the keywords with first keywords in a keyword database. If the matching is successful, it may determine that the first keyword is included in the interaction information and determine that the trigger condition is met. And / or, the service end may match keywords in the interaction information with second keywords. If the matching is successful, it may add 1 to the number of second keywords. If the number of second keywords reaches a keyword threshold, it may determine that the trigger condition is met. The first and second keywords may be keywords related to the second live streaming scene.
[0057] And / or the service end may obtain the duration of the first live-streaming scene and determine that the trigger condition is met when the duration reaches a preset duration. And / or the service end may determine that the trigger condition is met when the first live-streaming scene reaches a preset markpoint. The preset markpoint may be set in advance according to the multimedia resource in the first live-streaming scene. For example, if the multimedia resource is a book, the book may be semantically segmented to obtain multiple paragraphs for viewing, and a preset markpoint may be set at the end of each text paragraph. For example, if the multimedia resource is a song, the preset markpoint may be set based on the attribute features of the song.
[0058] In an embodiment of the present disclosure, the second video data is generated by the following steps: determining text information in a preset text library based on the target interaction information, the text information responding to the target interaction information; converting the text information into second audio data; searching a virtual object movement database for second movement data indicating facial movements and body movements of the virtual object in the first live-streaming scene corresponding to the target interaction information; determining second scene data for indicating a live-streaming room background screen in the second live-streaming scene based on a scene identifier of the second live-streaming scene; combining the second movement data, the second audio data, and the second scene data into second video data corresponding to the second live-streaming scene; and transmitting the second video data to multiple viewer terminals.
[0059] Optionally, searching for second motion data corresponding to the target interaction information in the virtual object motion database includes: recognizing emotion information fed back from the virtual object based on the target interaction information; and searching for second motion data corresponding to the emotion information in the virtual object motion database, wherein motion data corresponding to different emotion information, such as a clapping motion corresponding to a happy emotion and a banging motion corresponding to an angry emotion, are pre-stored in the virtual object motion database.
[0060] Since the second live-streaming scene is a live-streaming scene in which the virtual object responds to the interaction information, the second video data can be generated based on the target interaction information. Specifically, the service end determines text information matching the target interaction information in a preset text library through semantic recognition and analysis, and converts the text information into realistic voice data of the virtual object in real time using Text To Speech (TTS) technology to obtain second audio data. Then, the service end searches a virtual object action database to determine second action data corresponding to the emotion information indicated by the target interaction information, and determines second scene data based on the scene identifier of the second live-streaming scene. The service end combines the second audio data, the second action data, and the second scene data to obtain second video data, and transmits the second video data to multiple viewer terminals.
[0061] After receiving the second video data, the viewer terminal can generate corresponding video content through decoding process on the second video data, and play the video content in which the virtual object replies to the target interaction information in the second live-streaming scene on the live-streaming interface. Optionally, during the process of playing the video content in which the virtual object replies to the target interaction information in the second live-streaming scene, the live-streaming room background screen and the behavior of the virtual object can be switched according to changes in the video content based on the second scene data and the second behavior data, and may be different from the live-streaming room background screen and the behavior of the virtual object in the first live-streaming scene.
[0062] It can be understood that the above first live streaming scene is a live streaming scene in which a virtual object plays a multimedia resource, and the second live streaming scene is a live streaming scene in which a virtual object responds to interaction information. This is only an example, and the configuration of the first live streaming scene and the second live streaming scene can also be interchanged, and the first live streaming scene and the second live streaming scene can be constantly alternated, so that the live streaming scene of the virtual object is constantly switched.
[0063] In an embodiment of the present disclosure, the live streaming interaction method may further include transmitting the target interaction information and text information responding to the target interaction information to a plurality of viewer terminals.
[0064] The service end can determine target interaction information from multiple pieces of interaction information sent by live streaming viewers based on a preset scheme, and the preset scheme can be set according to actual circumstances, for example, to determine target interaction information based on the points of live streaming viewers who send interaction information, or to search for target interaction information that matches preset keywords, and the preset keywords may be those previously extracted by mining according to hotspot information, or may be keywords related to the content of the live streaming, or perform semantic recognition on the interaction information, cluster interaction information with similar expression meanings, and obtain a number of information sets, and the set with the most interaction information is the hottest topic for live streaming viewers to interact on, and the interaction information corresponding to this set is the target interaction information. Then, the service end can send the target interaction information and text information responding to the target interaction information to the viewer terminal, and the terminal can receive the text information responding to the target interaction information and display the target interaction information and the text information responding to the target interaction information in a second area of the live streaming interface.
[0065] In an embodiment of the present disclosure, a service end receives interaction information of a plurality of viewer terminals in a first live-streaming scene, and determines based on the interaction information whether a trigger condition for switching live-streaming scenes is met. If the trigger condition is met, the service end can transmit second video data corresponding to a second live-streaming scene to the plurality of viewer terminals, where the live-streaming scene is used to indicate the type of live-streaming content of the virtual object in the live-streaming room. By adopting the above technical solution, if the service end determines that the trigger condition for switching live-streaming scenes is met, the service end can transmit data of the second live-streaming scene to the viewer terminals, so that the viewer terminals can switch live-streaming scenes. The virtual object can switch from live-streaming in the first live-streaming scene to live-streaming in the second live-streaming scene based on the viewer interaction information, thereby realizing interaction sessions between the virtual object and the viewer in different live-streaming scenes to meet the various interaction needs of the viewer, improving the variety and interest of the live-streaming of the virtual object, and further improving the viewer's interaction experience.
[0066] 5 is a schematic diagram of a live streaming interaction device according to an embodiment of the present disclosure, which can be realized by software and / or hardware and generally can be integrated into an electronic device. As shown in FIG. 5, the device is installed in multiple viewer terminals that enter a live streaming room of a virtual object, a first live streaming module 301 for playing video content of the virtual object in a first live streaming scene on a live streaming interface and displaying interaction information from the plurality of viewer terminals; and a second live streaming module 302 for playing video content of the virtual object in a second live streaming scene on the live streaming interface in response to the interaction information satisfying a trigger condition, wherein the live streaming scene is used to indicate a type of live streaming content of the virtual object.
[0067] Optionally, said live streaming scenes include live streaming scenes in which said virtual objects act out multimedia resources, and live streaming scenes in which said virtual objects respond to interaction information.
[0068] Optionally, the first live-streaming scene is a live-streaming scene in which the virtual object plays a multimedia resource, and the first live-streaming module 301 specifically comprises: Displaying multimedia resource information of a plurality of multimedia resources to be performed in a first area of the live streaming interface; The virtual object is used to play video content that plays the target multimedia resource; The target multimedia resource is determined based on trigger information of the plurality of viewer terminals for the plurality of multimedia resources.
[0069] Optionally, the second live streaming module 302 specifically comprises: The live streaming interface is used to play video content in which the virtual object responds to the interaction information.
[0070] Optionally, the trigger condition includes at least one of: a number of the interaction information reaches a predetermined threshold; a first keyword is included in the interaction information; a number of second keywords in the interaction information reaches a keyword threshold; a duration of the first live streaming scene reaches a preset duration; and the first live streaming scene reaches a preset mark point.
[0071] Optionally, the virtual object responds with target interaction information from the interaction information, and the device: The live streaming interface further includes a response module for displaying the target interaction information and text information responding to the target interaction information in a second area of the live streaming interface.
[0072] Optionally, the first live streaming module 301 specifically comprises: receiving first video data corresponding to the first live streaming scene; used to play, on the live streaming interface, video content in which the virtual object plays the target multimedia resource in the first live streaming scene based on the first video data; The first video data includes first scene data, first action data, and first audio data, wherein the first scene data is used to show a live streaming room background screen in the first live streaming scene, the first action data is used to show facial expressions and body actions of the virtual object in the first live streaming scene, and the audio data matches the target multimedia resource.
[0073] Optionally, the second live streaming module is specifically: receiving second multimedia data corresponding to the second live streaming scene; playback, on the live streaming interface, video content in which the virtual object responds to the target interaction information in the second live streaming scene based on the second multimedia data; The second multimedia data includes second scene data, second action data, and second audio data, wherein the second scene data is used to display a live streaming room background screen in the second live streaming scene, the second action data is used to display facial expressions and body actions of the virtual object in the second live streaming scene, and the second audio data is generated based on the target interaction information.
[0074] The live streaming interaction device provided by the embodiments of the present disclosure can execute the live streaming interaction method provided by any embodiment of the present disclosure, and has functional modules and beneficial effects according to the execution method.
[0075] 6 is a schematic diagram of another live streaming interaction device according to an embodiment of the present disclosure, which can be realized by software and / or hardware, and can generally be integrated into an electronic device. As shown in FIG. 6, the device is installed at a service end, an information receiving module 401 for receiving interaction information of a plurality of viewer terminals in a first live streaming scene, and determining whether a trigger condition for switching a live streaming scene is met according to the interaction information; and a data transmitting module 402 for transmitting second video data corresponding to a second live streaming scene to the plurality of viewer terminals when the trigger condition is met, where the live streaming scene is used to indicate a type of live streaming content of a virtual object in the live streaming room.
[0076] Optionally, said live streaming scenes include live streaming scenes in which said virtual objects act out multimedia resources, and live streaming scenes in which said virtual objects respond to interaction information.
[0077] Optionally, the first live streaming scene is a live streaming scene in which the virtual object plays a multimedia resource, and the apparatus further comprises a data determination module, wherein the data determination module: Searching for first audio data matching the target multimedia resource in an audio database; searching for first movement data corresponding to the target multimedia resource in a virtual object movement database, the first movement data being for indicating facial movements and body movements of the virtual object in the first live streaming scene; Determine first scene data for showing a live streaming room background screen in the first live streaming scene according to a scene identifier of the first live streaming scene; combining the first motion data, the first audio data, and the first scene data into first video data corresponding to the first live streaming scene; It is used to transmit the first video data to the plurality of viewer terminals.
[0078] Optionally, the apparatus comprises: The method further includes a resource determination module for receiving trigger information for a plurality of multimedia resources to be displayed in the first live streaming scene from the plurality of viewer terminals, and determining the target multimedia resource from the plurality of multimedia resources based on the trigger information.
[0079] Optionally, said apparatus further comprises a second data module, said second data module comprising: Determine text information in a preset text library that responds to the target interaction information based on the target interaction information; converting the text information into second audio data; Searching a virtual object movement database for second movement data indicating facial movements and body movements of the virtual object in the first live streaming scene corresponding to the target interaction information; Determine second scene data for showing a live streaming room background screen in the second live streaming scene according to a scene identifier of the second live streaming scene; combining the second motion data, the second audio data, and the second scene data into second video data corresponding to the second live streaming scene; The second video data is used to transmit the second video data to the plurality of viewer terminals.
[0080] Optionally, said second data module comprises: Recognizing emotion information fed back from the virtual object based on the target interaction information; The second action data corresponding to the emotion information is used to search a virtual object action database.
[0081] Optionally, the apparatus comprises: The system further includes a response information transmitting module for transmitting the target interaction information and text information responding to the target interaction information to the plurality of viewer terminals.
[0082] Optionally, the apparatus further comprises a trigger condition module, the trigger condition module comprising: When the number of similar interaction information among the interaction information reaches a predetermined threshold, the trigger condition is satisfied, and the similar interaction information is interaction information whose similarity is greater than a similarity threshold; extracting keywords in the interaction information, matching the keywords with first keywords and / or second keywords in a keyword database, and satisfying a trigger condition if the interaction information includes the first keywords and / or if the number of the second keywords included in the interaction information reaches a keyword threshold; When the duration of the first live streaming scene reaches a preset duration, a trigger condition is met; When the first live streaming scene reaches a preset mark point, it is used to satisfy a trigger condition.
[0083] The live streaming interaction device provided by the embodiments of the present disclosure can execute the live streaming interaction method provided by any embodiment of the present disclosure, and has functional modules and beneficial effects according to the execution method.
[0084] FIG. 7 is a schematic diagram of an electronic device according to an embodiment of the present disclosure. Referring specifically to FIG. 7, a schematic diagram suitable for implementing an electronic device 500 according to an embodiment of the present disclosure is shown. The electronic device 500 according to an embodiment of the present disclosure may include, but is not limited to, mobile terminals such as mobile phones, laptops, digital broadcast receivers, personal digital assistants (PDAs), tablets, portable multimedia players (PMPs), and in-vehicle terminals (e.g., in-vehicle navigation terminals), as well as fixed terminals such as digital TVs and desktop computers. The electronic device shown in FIG. 7 is merely an example and should not impose any limitations on the functionality and scope of use of the embodiment of the present disclosure.
[0085] 7, the electronic device 500 may include a processing unit (e.g., a central processing unit, a graphics processor, etc.) 501 that may perform various appropriate operations and processes in accordance with a program stored in a read-only memory (ROM) 502 or loaded from a storage device 508 into a random access memory (RAM) 503. The RAM 503 may further store various programs and data necessary for the operation of the electronic device 500. The processing unit 501, the ROM 502, and the RAM 503 are connected to one another via a bus 504. An input / output (I / O) interface 505 is also connected to the bus 504.
[0086] Typically, input devices 506, including, for example, a touch screen, touch pad, keyboard, mouse, camera, microphone, accelerometer, gyroscope, etc.; output devices 507, including, for example, a liquid crystal display (LCD), speaker, vibrator, etc.; storage devices 508, including, for example, a magnetic tape, hard disk, etc.; and communication devices 509. The communication devices 509 enable the electronic device 500 to exchange data with other devices through wireless or wired communication. While FIG. 7 illustrates the electronic device 500 having various devices, it should be understood that it is not required to implement or include all of the devices shown. Instead, the electronic device 500 may implement or include more or fewer devices.
[0087] In particular, according to embodiments of the present disclosure, the processes described above with reference to the flowcharts may be implemented as a computer software program. For example, embodiments of the present disclosure include a computer program product including a computer program carried on a non-transitory computer-readable medium, the computer program including program code for performing the methods illustrated in the flowcharts. In such embodiments, the computer program may be downloaded and installed from a network via the communication device 509, or may be installed from the storage device 508 or from the ROM 502. When the computer program is executed by the processing device 501, it performs the above-described functions defined in the live streaming interaction method of the embodiments of the present disclosure.
[0088] It should be noted that the computer-readable medium described in this disclosure may be a computer-readable signal medium, a computer-readable storage medium, or any combination of the above. The computer-readable storage medium may be, for example, but is not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination thereof. More specific examples of the computer-readable storage medium may include, but are not limited to, an electrical connection having one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the above. In this disclosure, the computer-readable storage medium may be any tangible medium that contains or stores a program that can be used by or in combination with an instruction execution system, apparatus, or device. In this disclosure, the computer-readable signal medium may include a data signal propagated in baseband or as part of a carrier wave, with the computer-readable program code carried in the data signal. Such propagated data signals may take various forms, including, but not limited to, electromagnetic signals, optical signals, or any suitable combination of the above. A computer-readable signal medium may be any computer-readable medium other than a computer-readable storage medium, which can transmit, propagate, or transmit a program for use by or in connection with an instruction execution system, apparatus, or device. The program code contained in the computer-readable medium may be transmitted by any suitable medium, including, but not limited to, electrical wire, fiber optic cable, RF (radio frequency), etc., or any suitable combination of the above.
[0089] In some embodiments, clients and servers may communicate using any now known or later developed network protocol, such as HyperText Transfer Protocol (HTTP), and may be interconnected by any form or medium of digital data communication (e.g., a communications network). Examples of communications networks include a local network ("LAN"), a wide area network ("WAN"), the World Wide Web (e.g., the Internet), an end-to-end network (e.g., an ad-hoc end-to-end network), and any now known or later developed network.
[0090] The computer-readable medium may be included in the electronic device or may be separate from and not located on the electronic device.
[0091] When the computer-readable medium carries one or more programs, the one or more programs, when executed by the electronic device, cause the electronic device to perform the following steps: playing video content of the virtual object in a first live streaming scene on a live streaming interface, and displaying interaction information from the plurality of viewer terminals; and, in response to the interaction information satisfying a trigger condition, playing video content of the virtual object in a second live streaming scene on the live streaming interface, wherein the live streaming scene is used to indicate a type of live streaming content of the virtual object.
[0092] Alternatively, when the computer-readable medium carries one or more programs, and the one or more programs are executed by the electronic device, the electronic device is caused to perform the following steps: receiving interaction information of multiple viewer terminals in a first live-streaming scene, and determining, based on the interaction information, whether a trigger condition for switching a live-streaming scene is met; and if the trigger condition is met, sending second video data corresponding to a second live-streaming scene to the multiple viewer terminals, where the live-streaming scene is used to indicate a type of live-streaming content of a virtual object in the live-streaming room.
[0093] Computer program code for carrying out the operations of the present disclosure can be written in one or more programming languages or a combination thereof, including object-oriented programming languages such as Java, Smalltalk, C++, and further including, but not limited to, conventional procedural programming languages such as "C" or similar programming languages. The program code can run entirely on the user's computer, partially on the user's computer, as a separate software package, partially on the user's computer and partially on a remote computer, or entirely on a remote computer or server. When a remote computer is involved, the remote computer can be connected to the user's computer via any type of network, including a local area network (LAN) or a wide area network (WAN), or can be connected to an external computer (e.g., via the Internet using an Internet service provider).
[0094] The flowcharts and block diagrams in the figures illustrate the system architecture, functions, and operations that can be implemented according to the systems, methods, and computer program products of various embodiments of the present application. In this regard, each block in the flowcharts or block diagrams may represent a module, program segment, or portion of code, which includes one or more executable instructions for implementing a given logical function. Note that in some alternative implementations, the functions shown in the blocks may occur in an order different from that shown in the figures. For example, two blocks shown in succession may actually be executed essentially in parallel, or in some cases, in the reverse order, depending on the functionality involved. Furthermore, each block in the block diagrams and / or flowcharts, and combinations of blocks in the block diagrams and / or flowcharts, may be implemented in a dedicated hardware system or a combination of dedicated hardware and computer instructions to perform a given function or operation.
[0095] The units described in the embodiments of the present disclosure may be implemented in a software manner or a hardware manner, and the names of the units, if any, do not constitute limitations on the units themselves.
[0096] The functions described herein may be performed, at least in part, by one or more hardware logic components, for example, exemplary types of hardware logic components that may be utilized include, but are not limited to, field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), systems on chips (SOCs), complex programmable logic devices (CPLDs), etc.
[0097] In this disclosure, a machine-readable medium may be a tangible medium that contains or stores a program that may be used by or in combination with an instruction execution system, apparatus, or device. A machine-readable medium may be a machine-readable signal medium or a machine-readable storage medium. A machine-readable medium includes, but is not limited to, an electronic, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any suitable combination of the above. More specific examples of machine-readable storage media include an electrical connection by one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact magnetic disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the above.
[0098] According to one or more embodiments of the present disclosure, the present disclosure provides a live streaming interaction method applied to a plurality of viewer terminals entering a live streaming room of a virtual object, comprising: Playing video content of the virtual object in a first live streaming scene on a live streaming interface and displaying interaction information from the plurality of viewer terminals; and in response to the interaction information satisfying a trigger condition, playing video content of the virtual object in a second live streaming scene on the live streaming interface, wherein the live streaming scene is used to indicate a type of live streaming content of the virtual object.
[0099] According to one or more embodiments of the present disclosure, in the live streaming interaction method provided by the present disclosure, the live streaming scene includes a live streaming scene in which the virtual object performs a multimedia resource, and a live streaming scene in which the virtual object responds to interaction information.
[0100] According to one or more embodiments of the present disclosure, in the live-streaming interaction method provided by the present disclosure, the first live-streaming scene is a live-streaming scene in which the virtual object plays a multimedia resource, and the step of playing video content of the virtual object in the first live-streaming scene in a live-streaming interface includes: Displaying multimedia resource information of a plurality of multimedia resources to be performed in a first area of the live streaming interface; The method includes a step of playing video content in which the virtual object plays a target multimedia resource, the target multimedia resource being determined based on trigger information of the plurality of viewer terminals for the plurality of multimedia resources.
[0101] According to one or more embodiments of the present disclosure, in the live-streaming interaction method provided by the present disclosure, the step of playing video content of the virtual object in a second live-streaming scene on the live-streaming interface includes: Playing video content in the live streaming interface in which the virtual object responds about the interaction information.
[0102] According to one or more embodiments of the present disclosure, in the live streaming interaction method provided by the present disclosure, the trigger condition includes at least one of: the number of the interaction information reaches a predetermined threshold; the interaction information includes a first keyword; the number of second keywords in the interaction information reaches a keyword threshold; the time length of the first live streaming scene reaches a preset time length; and the first live streaming scene reaches a preset mark point.
[0103] According to one or more embodiments of the present disclosure, in a live streaming interaction method provided by the present disclosure, the virtual object responds to target interaction information among the interaction information, and the method includes: The method further includes displaying the target interaction information and text information responding to the target interaction information in a second area of the live streaming interface.
[0104] According to one or more embodiments of the present disclosure, in the live-streaming interaction method provided by the present disclosure, the step of playing video content of the virtual object in a first live-streaming scene in a live-streaming interface includes: receiving first video data corresponding to the first live-streaming scene, the first video data including first scene data, first movement data, and first audio data, the first scene data being used to show a live-streaming room background screen in the first live-streaming scene, the first movement data being used to show facial movements and body movements of the virtual object in the first live-streaming scene, and the audio data matching the target multimedia resource; and playing, on the live streaming interface, video content in which the virtual object plays the target multimedia resource in the first live streaming scene based on the first video data.
[0105] According to one or more embodiments of the present disclosure, in the live-streaming interaction method provided by the present disclosure, the step of playing video content of the virtual object in a second live-streaming scene on the live-streaming interface includes: receiving second multimedia data corresponding to the second live-streaming scene, the second multimedia data including second scene data, second action data, and second audio data, the second scene data being used to display a live-streaming room background screen in the second live-streaming scene, the second action data being used to display facial actions and body actions of the virtual object in the second live-streaming scene, and the second audio data being generated based on the target interaction information; and playing video content in the live streaming interface, in which the virtual object responds to the target interaction information in the second live streaming scene based on the second multimedia data.
[0106] According to one or more embodiments of the present disclosure, the present disclosure provides a live streaming interaction method applied to a service end, comprising: receiving interaction information of a plurality of viewer terminals in a first live streaming scene, and determining whether a trigger condition for switching the live streaming scene is satisfied based on the interaction information; and transmitting second video data corresponding to a second live streaming scene to the plurality of viewer terminals when the trigger condition is met, wherein the live streaming scene is used to indicate a type of live streaming content of a virtual object in the live streaming room.
[0107] According to one or more embodiments of the present disclosure, in the live streaming interaction method provided by the present disclosure, the live streaming scene includes a live streaming scene in which the virtual object performs a multimedia resource, and a live streaming scene in which the virtual object responds to interaction information.
[0108] According to one or more embodiments of the present disclosure, in the live-streaming interaction method provided by the present disclosure, the first live-streaming scene is a live-streaming scene in which the virtual object plays a multimedia resource; Searching for first audio data matching a target multimedia resource in an audio database, and searching for first movement data corresponding to the target multimedia resource in a virtual object movement database, the first movement data being for indicating facial movements and body movements of the virtual object in the first live streaming scene; determining first scene data for showing a live streaming room background screen in the first live streaming scene according to a scene identifier of the first live streaming scene; combining the first motion data, the first audio data, and the first scene data into first video data corresponding to the first live streaming scene; The method further includes transmitting the first video data to the plurality of viewer terminals.
[0109] According to one or more embodiments of the present disclosure, there is provided a live streaming interaction method, comprising: receiving trigger information for a plurality of multimedia resources to be displayed in the first live streaming scene from the plurality of viewer terminals; and determining the target multimedia resource from the plurality of multimedia resources based on the trigger information.
[0110] According to one or more embodiments of the present disclosure, in the live streaming interaction method provided by the present disclosure, the second video data is: determining, based on the target interaction information, text information in a preset text library that responds to the target interaction information; converting the text information into second audio data; searching a virtual object movement database for second movement data indicating facial movements and body movements of the virtual object in the first live streaming scene corresponding to the target interaction information; determining second scene data for showing a live streaming room background screen in the second live streaming scene according to a scene identifier of the second live streaming scene; combining the second motion data, the second audio data, and the second scene data into second video data corresponding to the second live streaming scene; and transmitting the second video data to the plurality of viewer terminals.
[0111] According to one or more embodiments of the present disclosure, in the live streaming interaction method provided by the present disclosure, the step of searching for second action data corresponding to the target interaction information in a virtual object action database includes: Recognizing emotion information fed back from the virtual object based on the target interaction information; and searching for second motion data corresponding to the emotion information in a virtual object motion database.
[0112] According to one or more embodiments of the present disclosure, there is provided a live streaming interaction method, the method comprising: The method further includes a step of transmitting the target interaction information and text information responding to the target interaction information to the plurality of viewer terminals.
[0113] According to one or more embodiments of the present disclosure, in the live streaming interaction method provided by the present disclosure, the trigger condition comprises: a method in which, when the number of similar interaction information among the interaction information reaches a predetermined threshold, a trigger condition is satisfied, and the similar interaction information is interaction information whose similarity is greater than a similarity threshold; extracting keywords from the interaction information, matching the keywords with first keywords and / or second keywords in a keyword database, and satisfying a trigger condition when the interaction information includes the first keywords and / or when the number of the second keywords in the interaction information reaches a keyword threshold; A method of satisfying a trigger condition when a duration of the first live streaming scene reaches a preset duration; whether the first live streaming scene reaches a preset mark point; and whether a trigger condition is met.
[0114] According to one or more embodiments of the present disclosure, the present disclosure provides a live streaming interaction device, comprising: a first live streaming module for playing video content of the virtual object in a first live streaming scene on a live streaming interface and displaying interaction information from the plurality of viewer terminals; and a second live streaming module for playing video content of the virtual object in a second live streaming scene on the live streaming interface in response to the interaction information satisfying a trigger condition, the live streaming scene being used to indicate a type of live streaming content of the virtual object.
[0115] According to one or more embodiments of the present disclosure, in the live streaming interaction device provided by the present disclosure, the live streaming scene includes a live streaming scene in which the virtual object performs a multimedia resource, and a live streaming scene in which the virtual object responds to interaction information.
[0116] According to one or more embodiments of the present disclosure, in the live-streaming interaction device provided by the present disclosure, the first live-streaming scene is a live-streaming scene in which the virtual object plays a multimedia resource, and the first live-streaming module specifically: Displaying multimedia resource information of a plurality of multimedia resources to be performed in a first area of the live streaming interface; The virtual object is used to play video content that plays the target multimedia resource; The target multimedia resource is determined based on trigger information of the plurality of viewer terminals for the plurality of multimedia resources.
[0117] According to one or more embodiments of the present disclosure, in the live-streaming interaction device provided by the present disclosure, the second live-streaming module specifically comprises: The live streaming interface is used to play video content in which the virtual object responds to the interaction information.
[0118] According to one or more embodiments of the present disclosure, in the live streaming interaction device provided by the present disclosure, the trigger condition includes at least one of: the number of the interaction information reaches a predetermined threshold; the interaction information includes a first keyword; the number of second keywords in the interaction information reaches a keyword threshold; the time length of the first live streaming scene reaches a preset time length; and the first live streaming scene reaches a preset mark point.
[0119] According to one or more embodiments of the present disclosure, in a live streaming interaction device provided by the present disclosure, the virtual object responds to target interaction information among the interaction information, and the device: The live streaming interface further includes a response module for displaying the target interaction information and text information responding to the target interaction information in a second area of the live streaming interface.
[0120] According to one or more embodiments of the present disclosure, in the live-streaming interaction device provided by the present disclosure, the first live-streaming module specifically comprises: receiving first video data corresponding to the first live streaming scene; used to play, on the live streaming interface, video content in which the virtual object plays the target multimedia resource in the first live streaming scene based on the first video data; The first video data includes first scene data, first action data, and first audio data, wherein the first scene data is used to show a live streaming room background screen in the first live streaming scene, the first action data is used to show facial expressions and body actions of the virtual object in the first live streaming scene, and the audio data matches the target multimedia resource.
[0121] According to one or more embodiments of the present disclosure, in the live-streaming interaction device provided by the present disclosure, the second live-streaming module specifically comprises: receiving second multimedia data corresponding to the second live streaming scene; playback, on the live streaming interface, video content in which the virtual object responds to the target interaction information in the second live streaming scene based on the second multimedia data; The second multimedia data includes second scene data, second action data, and second audio data, wherein the second scene data is used to display a live streaming room background screen in the second live streaming scene, the second action data is used to display facial expressions and body actions of the virtual object in the second live streaming scene, and the second audio data is generated based on the target interaction information.
[0122] According to one or more embodiments of the present disclosure, the present disclosure provides a live streaming interaction device, comprising: an information receiving module for receiving interaction information of a plurality of viewer terminals in a first live streaming scene, and determining whether a trigger condition for switching a live streaming scene is satisfied based on the interaction information; and a data transmitting module for transmitting second video data corresponding to a second live streaming scene to the plurality of viewer terminals when the trigger condition is met, wherein the live streaming scene is used to indicate a type of live streaming content of a virtual object in the live streaming room.
[0123] According to one or more embodiments of the present disclosure, in the live streaming interaction device provided by the present disclosure, the live streaming scene includes a live streaming scene in which the virtual object performs a multimedia resource, and a live streaming scene in which the virtual object responds to interaction information.
[0124] According to one or more embodiments of the present disclosure, in the live-streaming interaction device provided by the present disclosure, the first live-streaming scene is a live-streaming scene in which the virtual object plays a multimedia resource, and the device further includes a data determination module, wherein the data determination module: Searching for first audio data matching the target multimedia resource in an audio database; searching for first movement data corresponding to the target multimedia resource in a virtual object movement database, the first movement data being for indicating facial movements and body movements of the virtual object in the first live streaming scene; Determine first scene data for showing a live streaming room background screen in the first live streaming scene according to a scene identifier of the first live streaming scene; combining the first motion data, the first audio data, and the first scene data into first video data corresponding to the first live streaming scene; It is used to transmit the first video data to the plurality of viewer terminals.
[0125] According to one or more embodiments of the present disclosure, in the live streaming interaction device provided by the present disclosure, the device further includes a resource determination module, and the resource determination module: receiving trigger information for a plurality of multimedia resources to be displayed in the first live streaming scene from the plurality of viewer terminals; The trigger information is used to determine the target multimedia resource from the plurality of multimedia resources based on the trigger information.
[0126] According to one or more embodiments of the present disclosure, in the live streaming interaction device provided by the present disclosure, the device further includes a second data module, wherein the second data module: Determine text information in a preset text library that responds to the target interaction information based on the target interaction information; converting the text information into second audio data; Searching a virtual object movement database for second movement data indicating facial movements and body movements of the virtual object in the first live streaming scene corresponding to the target interaction information; Determine second scene data for showing a live streaming room background screen in the second live streaming scene according to a scene identifier of the second live streaming scene; combining the second motion data, the second audio data, and the second scene data into second video data corresponding to the second live streaming scene; The second video data is used to transmit the second video data to the plurality of viewer terminals.
[0127] According to one or more embodiments of the present disclosure, in the live streaming interaction device provided by the present disclosure, the second data module comprises: Recognizing emotion information fed back from the virtual object based on the target interaction information; The second action data corresponding to the emotion information is used to search a virtual object action database.
[0128] According to one or more embodiments of the present disclosure, there is provided a live streaming interaction device, the device comprising: The system further includes a response information transmitting module for transmitting the target interaction information and text information responding to the target interaction information to the plurality of viewer terminals.
[0129] According to one or more embodiments of the present disclosure, in the live streaming interaction device provided by the present disclosure, the device further includes a trigger condition module, and the trigger condition module: When the number of similar interaction information among the interaction information reaches a predetermined threshold, the trigger condition is satisfied, and the similar interaction information is interaction information whose similarity is greater than a similarity threshold; extracting keywords in the interaction information, matching the keywords with first keywords and / or second keywords in a keyword database, and satisfying a trigger condition when the interaction information includes the first keywords and / or when the number of the second keywords in the interaction information reaches a keyword threshold; When the duration of the first live streaming scene reaches a preset duration, a trigger condition is met; When the first live streaming scene reaches a preset mark point, it is used to satisfy a trigger condition.
[0130] According to one or more embodiments of the present disclosure, the present disclosure provides an electronic device, a processor; a memory for storing executable instructions for said processor; Including, The processor reads and executes the executable instructions from the memory to realize any one of the live streaming interaction methods provided by the present disclosure.
[0131] According to one or more embodiments of the present disclosure, the present disclosure provides a computer-readable storage medium having a computer program stored therein, the computer program being used to perform the live streaming interaction method described in any one of claims provided by the present disclosure.
[0132] The above description merely describes the preferred embodiments and applied technical principles of the present disclosure. As can be understood by those skilled in the art, the scope of the present disclosure is not limited to the technical solutions formed by the specific combinations of the above technical features, but also includes other technical solutions formed by any combination of the above technical features or their equivalent features without departing from the concept disclosed above, for example, by substituting the above features with technical features having similar functions disclosed in the present disclosure (but not limited to these).
[0133] Also, although operations are described employing a particular order, this should not be construed as requiring these operations to be performed in the particular order or sequence shown. In certain environments, multitasking and parallel processing may be advantageous. Similarly, although the above discussion includes some specific implementation details, these should not be construed as limitations on the scope of the disclosure. Some features that are described in the context of a single embodiment may also be implemented in combination in a single embodiment. Conversely, various features that are described in the context of a single embodiment may also be implemented in multiple embodiments separately or in any suitable subcombination.
[0134] Although the present subject matter has been described in language specific to structural features and / or logical operations of a method, it should be understood that the subject matter defined in the appended claims is not limited to the specific features or operations described above. Rather, the specific features and operations described above are merely example forms of implementing the claims.
Claims
1. A live streaming interaction method applied to a plurality of viewer terminals entering a live streaming room of a virtual object, comprising: Playing video content of the virtual object in a first live streaming scene on a live streaming interface and displaying interaction information from the plurality of viewer terminals; playing, in the live streaming interface, video content of the virtual object in a second live streaming scene in response to the interaction information satisfying a trigger condition, the trigger condition being a condition for determining whether to perform live streaming scene switching based on the interaction information, and the live streaming scene being used to indicate a type of live streaming content of the virtual object; the first live streaming scene is a live streaming scene in which the virtual object plays a multimedia resource; Playing video content of the virtual object in a first live streaming scene in a live streaming interface includes: Displaying multimedia resource information of a plurality of multimedia resources to be performed in a first area of the live streaming interface; playing video content in which the virtual object plays a target multimedia resource, the target multimedia resource being determined based on trigger information of the plurality of the viewer terminals for the plurality of multimedia resources; For video content of the virtual object in the second live streaming scene, determining, based on the target interaction information, text information in a preset text library that responds to the target interaction information; converting the text information into second audio data; searching for second movement data corresponding to the target interaction information in a virtual object movement database, where the second movement data is used to represent facial movements and body movements of the virtual object in the second live streaming scene; determining second scene data for showing a live streaming room background screen in the second live streaming scene according to a scene identifier of the second live streaming scene; transmitting the second motion data, the second audio data, and the second scene data to the plurality of viewer terminals to combine them into video content of the virtual object in the second live streaming scene; generated by the service end in a manner that includes A method characterized by:
2. The live streaming scene includes a live streaming scene in which the virtual object plays a multimedia resource, and a live streaming scene in which the virtual object responds to interaction information.
2. The method of claim 1 .
3. Playing video content of the virtual object in a second live streaming scene on the live streaming interface includes: playing video content in the live streaming interface, in which the virtual object responds about the interaction information; 2. The method of claim 1 .
4. the trigger condition includes at least one of: the number of pieces of interaction information reaches a predetermined threshold; the interaction information includes a first keyword; and the number of second keywords in the interaction information reaches a keyword threshold.
2. The method of claim 1 .
5. The virtual object responds to the target interaction information of the interaction information, and the method further comprises: and further comprising displaying the target interaction information and the text information responding to the target interaction information in a second area of the live streaming interface.
4. The method of claim 3.
6. The target interaction information is determined based on points of the live streaming viewers who send the interaction information, or the target interaction information is obtained by matching based on preset keywords; 6. The method of claim 5.
7. Playing video content of the virtual object in a first live streaming scene in a live streaming interface includes: receiving first video data corresponding to the first live-streaming scene, the first video data including first scene data, first action data, and first audio data, the first scene data being used to show a live-streaming room background screen in the first live-streaming scene, the first action data being used to show facial action and body action of the virtual object in the first live-streaming scene, and the audio data being matched with the target multimedia resource; and playing, on the live streaming interface, video content in which the virtual object plays the target multimedia resource in the first live streaming scene based on the first video data.
2. The method of claim 1 .
8. Playing video content of the virtual object in a second live streaming scene on the live streaming interface includes: receiving second multimedia data corresponding to the second live streaming scene, the second multimedia data including the second scene data, the second action data, and the second audio data; and playing, on the live streaming interface, video content in which the virtual object responds to the target interaction information in the second live streaming scene based on the second multimedia data.
6. The method of claim 5.
9. The scene corresponding to the live streaming room background screen includes a background scene and a screen field of view scene in the live streaming scene of the virtual object, and the screen field of view scene is a field of view when different lenses capture the virtual object; 9. The method according to claim 7 or 8.
10. A live streaming interaction method applied to a service end, comprising: receiving interaction information of a plurality of viewer terminals in a first live streaming scene, and determining whether a trigger condition for switching the live streaming scene is satisfied based on the interaction information; sending second video data corresponding to a second live-streaming scene to the plurality of viewer terminals when the trigger condition is satisfied, wherein the trigger condition is a condition for determining whether to switch the live-streaming scene based on the interaction information, and the live-streaming scene is used to indicate a type of live-streaming content of a virtual object in a live-streaming room; The method comprises: receiving trigger information for a plurality of multimedia resources to be displayed in a first live streaming scene from the plurality of viewer terminals; Further comprising: determining a target multimedia resource from the plurality of multimedia resources based on the trigger information, wherein the target multimedia resource is a video content in which the virtual object plays the multimedia resource to the plurality of viewer terminals; The second video data is determining, based on the target interaction information, text information in a preset text library that responds to the target interaction information; converting the text information into second audio data; searching for second movement data corresponding to the target interaction information in a virtual object movement database, where the second movement data is used to represent facial movements and body movements of the virtual object in the second live streaming scene; determining second scene data for showing a live streaming room background screen in the second live streaming scene according to a scene identifier of the second live streaming scene; combining the second motion data, the second audio data, and the second scene data into second video data corresponding to the second live streaming scene; transmitting the second video data to the plurality of viewer terminals; A method characterized by:
11. The live streaming scene includes a live streaming scene in which the virtual object plays a multimedia resource, and a live streaming scene in which the virtual object responds to interaction information.
11. The method of claim 10.
12. the first live streaming scene is a live streaming scene in which the virtual object plays a multimedia resource; searching for first audio data matching a target multimedia resource in an audio database, and searching for first movement data corresponding to the target multimedia resource in a virtual object movement database, where the first movement data is used to represent facial movements and body movements of the virtual object in the first live streaming scene; determining first scene data for showing a live streaming room background screen in the first live streaming scene according to a scene identifier of the first live streaming scene; combining the first motion data, the first audio data, and the first scene data into first video data corresponding to the first live streaming scene; transmitting the first video data to the plurality of viewer terminals; The method of claim 11 further comprising:
13. The step of searching for second action data corresponding to the target interaction information in a virtual object action database includes: Recognizing emotion information fed back from the virtual object based on the target interaction information; searching a virtual object action database for second action data corresponding to the emotion information; 11. The method of claim 10, comprising:
14. transmitting the target interaction information and text information responding to the target interaction information to the plurality of viewer terminals; The method of claim 10 further comprising:
15. The target interaction information is determined based on points of the live streaming viewers who send the interaction information, or the target interaction information is obtained by matching based on preset keywords; 15. The method according to any one of claims 10 to 14.
16. The trigger condition is: a method in which, when the number of similar interaction information among the interaction information reaches a predetermined threshold, a trigger condition is satisfied, and the similar interaction information is interaction information whose similarity is greater than a similarity threshold; extracting keywords in the interaction information, matching the keywords with first keywords and / or second keywords in a keyword database, and satisfying a trigger condition when the interaction information includes the first keywords and / or when the number of the second keywords reaches a keyword threshold; 11. The method of claim 10.
17. A live streaming interaction device provided in a plurality of viewer terminals entering a live streaming room of a virtual object, a first live streaming module for playing video content of the virtual object in a first live streaming scene and displaying interaction information from the plurality of viewer terminals in a live streaming interface; a second live streaming module for playing video content of the virtual object in a second live streaming scene on the live streaming interface in response to the interaction information satisfying a trigger condition, the trigger condition being a condition for determining whether to perform live streaming scene switching based on the interaction information, and the live streaming scene being used to indicate a type of live streaming content of the virtual object; Including, The first live streaming scene is a live streaming scene in which the virtual object performs a multimedia resource, and the first live streaming module further displays multimedia resource information of a plurality of target multimedia resources in a first area of the live streaming interface, and is used to play video content in which the virtual object performs a target multimedia resource, and the target multimedia resource is determined according to trigger information of a plurality of the viewer terminals for the plurality of multimedia resources; For video content of the virtual object in the second live streaming scene, determining, based on the target interaction information, text information in a preset text library that responds to the target interaction information; converting the text information into second audio data; searching for second movement data corresponding to the target interaction information in a virtual object movement database, where the second movement data is used to represent facial movements and body movements of the virtual object in the second live streaming scene; determining second scene data for showing a live streaming room background screen in the second live streaming scene according to a scene identifier of the second live streaming scene; transmitting the second motion data, the second audio data, and the second scene data to the plurality of viewer terminals so that a second live streaming module combines them into video content of the virtual object in the second live streaming scene; generated by the service end in a manner that includes A live streaming interaction device comprising:
18. A live streaming interaction device provided at a service end, comprising: an information receiving module for receiving interaction information of a plurality of viewer terminals in a first live-streaming scene, and determining, according to the interaction information, whether a trigger condition for switching the live-streaming scene is met; a data transmitting module for transmitting second video data corresponding to a second live-streaming scene to the plurality of viewer terminals when the trigger condition is satisfied, wherein the trigger condition is a condition for determining whether to perform live-streaming scene switching based on the interaction information, and the live-streaming scene is used to indicate a type of live-streaming content of a virtual object in a live-streaming room; a resource determination module for receiving trigger information for a plurality of multimedia resources displayed in a first live streaming scene from the plurality of viewer terminals, and determining a target multimedia resource from the plurality of multimedia resources according to the trigger information, wherein the target multimedia resource is a video content in which the virtual object plays the multimedia resource to the plurality of viewer terminals; Including, The live streaming interaction device further comprises: determining, based on the target interaction information, text information in a preset text library that responds to the target interaction information; converting the text information into second audio data; searching for second movement data corresponding to the target interaction information in a virtual object movement database, where the second movement data is used to represent facial movements and body movements of the virtual object in the second live streaming scene; determining second scene data for showing a live streaming room background screen in the second live streaming scene according to a scene identifier of the second live streaming scene; combining the second motion data, the second audio data, and the second scene data into second video data corresponding to the second live streaming scene; transmitting the second video data to the plurality of viewer terminals; and configured to generate the second video data by executing A live streaming interaction device comprising:
19. 1. An electronic device comprising: a processor; a memory for storing executable instructions for said processor; Including, The processor reads and executes the executable instructions from the memory to perform the live streaming interaction method of any one of claims 1 to 16. An electronic device characterized by:
20. A computer-readable storage medium having a computer program stored thereon, the computer program being used to perform the live streaming interaction method according to any one of claims 1 to 16. A computer-readable storage medium comprising:
Citation Information
Patent Citations
Interface display method and device, terminal and storage medium
CN110286976A
Live broadcast interaction method and device, electronic equipment and storage medium
CN110519611A