Content display method, device, medium and program product

The performance of the anchor is evaluated through AI algorithms and triggered the display of custom content, which solves the problem of single human-computer interaction on the live broadcast platform and achieves rich live broadcast experience and interactive effects.

CN120302097APending Publication Date: 2025-07-11GUANGZHOU KUGOU COMP TECH CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Application Number
CN202510278617.8
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-03-10
Publication Date
2025-07-11

AI Technical Summary

Technical Problem

The human-computer interaction method of the existing live broadcast platform is relatively single, and the anchor may not receive audience feedback for a long time, resulting in poor live broadcast experience and limiting interactivity and fun.

Method used

The live broadcast performance of the anchor is evaluated in real time through AI algorithms, and the pre-configured custom content is triggered to display pre-configured custom content in the live broadcast room based on the scoring results, meeting the user's personalized setting needs and enriching the diversity of live broadcast images and interactive methods.

Benefits of technology

It improves human-computer interaction efficiency and user experience, encourages anchors to improve their performance, enhances live broadcast effects, and increases the interactive fun of the live broadcast room.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120302097A_ABST
    Figure CN120302097A_ABST
Patent Text Reader

Abstract

The invention discloses a content display method and device, a medium and a program product, and relates to the technical field of computers, and the method comprises the following steps: receiving a content configuration operation for a first live broadcast room, the content configuration operation being used for configuring a first customized content displayed in the first live broadcast room according to a preset triggering requirement; displaying live broadcast content of the first live broadcast room, wherein the live broadcast content comprises content represented by the first main body according to a preset content template; and under the condition that the matching relationship between the live broadcast content and the content template meets a preset triggering requirement, displaying the first user-defined content in the first live broadcast room. The personalized setting requirement of the user on the display content of the live broadcast room can be met, the diversity of live broadcast pictures is enriched, and the man-machine interaction mode is increased.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] Embodiments of the present application relate to the field of computer technology, and in particular, to a content display method, device, medium, and program product. Background Art

[0002] Live singing is a popular form of entertainment interaction. The host plays music and sings in the live room, and the audience in the live room can listen to the songs in real time and interact with the host.

[0003] In related technologies, most live platforms provide a gift-giving function for live broadcasts. Users can give virtual gifts (virtual resources) to the host according to the host's singing performance, and special effect pictures corresponding to the virtual gifts will be displayed in the live room, enhancing the interaction between the audience and the host.

[0004] However, the above interaction method relies on the audience to actively reward, and the human-computer interaction method is relatively single. The host may not receive feedback for a long time, resulting in a poor live broadcast experience and limiting the fun and interactivity of the live broadcast. Summary of the Invention

[0005] Embodiments of the present application provide a content display method, device, medium, and program product, which can enrich the human-computer interaction method. The technical solutions are as follows:

[0006] On the one hand, a content display method is provided. The method includes:

[0007] Receiving a content configuration operation for a first live room, where the content configuration operation is used to configure first custom content to be displayed in the first live room according to a preset trigger requirement;

[0008] Displaying the live broadcast content of the first live room, where the live broadcast content includes the content presented by a first subject according to a preset content template;

[0009] When the matching relationship between the live broadcast content and the content template meets the preset trigger requirement, displaying the first custom content in the first live room.

[0010] On the other hand, a content display device is provided. The device includes:

[0011] A receiving module, configured to receive a content configuration operation for a first live room, where the content configuration operation is used to configure first custom content to be displayed in the first live room according to a preset trigger requirement;

[0012] A content display module, configured to display the live broadcast content of the first live room, where the live broadcast content includes the content presented by a first subject according to a preset content template;

[0013] The content display module is further configured to display the first custom content in the first live broadcast room when the matching relationship between the live content and the content template meets the preset trigger requirement.

[0014] In an alternative embodiment, the content display module is further configured to display the first custom content in the first live broadcast room when the matching degree between the live content and the content template reaches a preset matching degree threshold.

[0015] In an alternative embodiment, the receiving module is further configured to receive at least two content configuration operations for the first live broadcast room, where the at least two content configuration operations are respectively used to configure the first custom content displayed according to different trigger requirements, and the first custom content corresponding to different trigger requirements is different;

[0016] The content display module is further configured to display the first custom content corresponding to the matching degree in the first live broadcast room based on the matching degree between the live content and the content template, where the first custom content corresponding to different preset matching degree intervals is different.

[0017] In an alternative embodiment, the content display module is further configured to display the first custom content in the first live broadcast room with a prominent performance intensity corresponding to the matching degree based on the matching degree between the live content and the content template, where there is a positive correlation between the matching degree and the prominent performance intensity.

[0018] In an alternative embodiment, the receiving module is further configured to receive a co-hosting request, where the co-hosting request is a request from a viewer account to co-host with the first entity in the first live broadcast room; in response to the co-hosting request, receive an operation to obtain at least one candidate custom content;

[0019] The apparatus further includes:

[0020] A sending module, configured to send the at least one candidate custom content to the viewer account, where the candidate custom content is used to provide the second custom content displayed for the viewer account when the preset trigger requirement is met during the co-hosting process.

[0021] In an alternative embodiment, the live content includes the singing content of the first entity, and the content template includes a song accompaniment;

[0022] The content display module is further configured to obtain a singing score for the matching between the singing content and the song accompaniment; and display the first custom content in the first live broadcast room when the singing score reaches a score threshold.

[0023] In an optional embodiment, the content display module is further configured to obtain beat point data of the song accompaniment, where the beat point data includes timestamps of song beat points of the song accompaniment in the audio of the song accompaniment; analyze the beat point data to determine a target beat point of the song accompaniment, and the timestamp of the target beat point is used to divide a singing scoring segment, and the target beat point meets at least one of the following conditions: the difference between the beat intensity of the target beat point and the adjacent beat point reaches a preset intensity threshold, the sound frequency at the target beat point reaches a maximum value, the target beat point corresponds to the start timestamp of the chorus in the song accompaniment; divide the song accompaniment into at least one singing scoring segment based on the timestamp of the target beat point, and the i-th singing scoring segment corresponds to the i-th time period in the song accompaniment, where i is a positive integer; for the i-th singing scoring segment, obtain the i-th singing segment corresponding to the i-th singing scoring segment from the singing content; calculate a similarity score between the i-th singing scoring segment and the i-th singing segment as the singing score for the match between the singing content and the song accompaniment during the i-th time period.

[0024] In an optional embodiment, the receiving module is further configured to, in response to receiving a beat point setting operation, set a specified beat point in the beat point data as the target beat point.

[0025] In an optional embodiment, when the first custom content is static content, the content display module is further configured to generate a dynamic performance animation corresponding to the first custom content; determine a first performance form of the dynamic performance animation based on the beat point data, where the first performance form matches the song beat points; display the dynamic performance animation in the first live room in the first performance form.

[0026] In an optional embodiment, the live content includes action content of a first subject, and the content template includes an action video;

[0027] The content display module is further configured to obtain an action score for the match between the action content and the action video; and display the first custom content in the first live room when the action score reaches a score threshold.

[0028] In an alternative embodiment, the content display module is further configured to collect action images of the action content based on a first frequency, and video images of the action video based on the first frequency, to obtain at least two pairs of action scoring images, where the m-th pair of action scoring images includes the m-th action image and the m-th video image collected at the m-th moment, and m is a positive integer; calculate a similarity score between the m-th action image and the m-th video image to obtain an action score of the m-th pair of action scoring images; and obtain an action score for matching between the action content and the action video based on the sum of the action scores of the at least two pairs of action scoring images respectively.

[0029] In an alternative embodiment, the receiving module is further configured to receive a display mode configuration operation for the first custom content, where the display mode configuration operation is used to indicate a manner of displaying the first custom content in combination with the live broadcast screen of the first live broadcast room.

[0030] The content display module is further configured to display the first custom content that matches the current live broadcast screen of the first live broadcast room in the first live broadcast room.

[0031] On the other hand, a computer device is provided, where the computer device includes a processor and a memory, and at least one instruction, at least one program, a code set, or an instruction set is stored in the memory, and the at least one instruction, the at least one program, the code set, or the instruction set is loaded and executed by the processor to implement the content display method as described in any one of the above embodiments of this application.

[0032] On the other hand, a computer-readable storage medium is provided, where at least one instruction, at least one program, a code set, or an instruction set is stored in the storage medium, and the at least one instruction, the at least one program, the code set, or the instruction set is loaded and executed by a processor to implement the content display method as described in any one of the above embodiments of this application.

[0033] On the other hand, a computer program product or a computer program is provided, where the computer program product or the computer program includes computer instructions, and the computer instructions are stored in a computer-readable storage medium. A processor of a computer device reads the computer instructions from the computer-readable storage medium, and the processor executes the computer instructions, so that the computer device executes the content display method as described in any one of the above embodiments.

[0034] The beneficial effects brought by the technical solutions provided in the embodiments of this application at least include:

[0035] By means of custom content with pre-configured triggering conditions, and triggering the display of the custom content according to the live broadcast performance of the first subject during the live broadcast, the personalized setting requirements of users for the content displayed in the live broadcast room can be met, the diversity of the live broadcast screen can be enriched, and the human-computer interaction method can be increased. When the performance of the subject meets the triggering conditions, instant feedback is generated, the human-computer interaction efficiency and the user experience during the live broadcast are improved, the enthusiasm of the subject to improve the live broadcast performance is encouraged to trigger the display of the custom content, and the live broadcast effect is enhanced. Brief Description of the Drawings

[0036] In order to more clearly illustrate the technical solutions in the embodiments of the present application, the drawings required for the description of the embodiments will be briefly introduced below. Obviously, the drawings in the following description are only some embodiments of the present application. For those of ordinary skill in the art, without creative efforts, other drawings can also be obtained based on these drawings.

[0037] Figure 1 is a schematic diagram of a content display system provided by an exemplary embodiment of the present application;

[0038] Figure 2 is a flowchart of a content display method provided by an exemplary embodiment of the present application;

[0039] Figure 3 is a flowchart of a content display method provided by another exemplary embodiment of the present application;

[0040] Figure 4 is a flowchart of a content display method provided by still another exemplary embodiment of the present application;

[0041] Figure 5 is a block diagram of the structure of a content display device provided by an exemplary embodiment of the present application;

[0042] Figure 6 is a block diagram of the structure of a content display device provided by another exemplary embodiment of the present application;

[0043] Figure 7 is a block diagram of the structure of a computer device provided by an exemplary embodiment of the present application. Detailed Embodiments

[0044] To make the objectives, technical solutions, and advantages of the present application clearer, the embodiments of the present application will be further described in detail below in conjunction with the drawings.

[0045] The terms used in this application are for the purpose of describing specific embodiments only and are not intended to limit this application. The singular forms "a", "the", and "said" used in this application and the appended claims are also intended to include the plural forms unless the context clearly indicates otherwise. It should also be understood that the term "and / or" as used herein refers to and encompasses any and all possible combinations of one or more of the associated listed items.

[0046] It should be noted that the information and data involved in this application are all information and data authorized by users or fully authorized by all parties, and the collection, use, and processing of relevant data need to comply with the relevant laws, regulations, and standards of relevant countries and regions.

[0047] First, a brief introduction to the nouns involved in the embodiments of this application:

[0048] Artificial Intelligence (AI): It is a new technical science that studies, develops theories, methods, technologies, and application systems for simulating, extending, and expanding human intelligence. It aims to enable machines to perform tasks that usually require human intelligence, such as learning, reasoning, problem-solving, understanding language, recognizing images, planning decisions, etc. AI enables systems to have the capabilities of self-adaptation, self-learning, and self-optimization by leveraging computer technology and data to achieve intelligent behaviors and interactions.

[0049] Among them, AI algorithms are a series of computational methods and rules used in the field of artificial intelligence to enable computer systems to achieve intelligent behaviors and complete specific tasks. For example, using AI algorithms to score singing.

[0050] Song beat points: They are the basic time units in music, used to measure rhythm and tempo. Beat points are the basis for forming music rhythm and usually repeat at uniform time intervals. Among them, the downbeat refers to the first beat of each measure in music, usually a strong beat, marking the beginning of a new measure. The chorus part usually starts with a downbeat.

[0051] Live singing is a popular online interactive method. Through a live streaming application, the host can open a live room to display the live content, and viewers can enter the live room to interact with the host. During the live broadcast, the host can play the accompaniment of a specified song and sing through the audio and video playback function. While listening to the host singing, viewers can also use the interaction mechanisms such as bullet screens and comments provided by the live streaming application to have real-time interactive communication with the host.

[0052] In the related art, the live broadcast application also provides a resource gifting function for the audience users, which allows the audience to give virtual resources / virtual gifts to the anchor in the operation interface based on their subjective evaluation of the anchor's singing performance, such as virtual props obtained by exchanging virtual resources. When the audience completes the gifting operation, the system will present the special effects animation or visual effects corresponding to the virtual gift on the display interface of the live broadcast room according to the preset rules, so as to enhance the interactive experience between the user and the anchor and enrich the visual presentation of the live broadcast.

[0053] However, the above-mentioned interactive mode mainly relies on the subjective behavior of the audience to initiate interaction, resulting in a relatively simple human-computer interaction method. In actual application scenarios, the anchor may face a long period of lack of audience feedback. For example, during the continuous live broadcast, although the anchor sang many songs, the live broadcast room did not appear the special effects screens generated by the audience's gift of resources, which reduced the anchor's live broadcast enthusiasm, affected his live broadcast experience, and also limited the fun and diversity of live broadcast interaction.

[0054] The present application provides a content display method, whereby the anchor can upload custom content in advance as content material to be triggered and displayed in the live broadcast room. In the live broadcast scenario, the AI ​​algorithm performs real-time scoring and periodic cumulative scoring on the anchor's live broadcast performance to obtain the scoring results. The custom content is automatically displayed in the live broadcast room based on whether the scoring results meet the preset trigger conditions, thereby increasing the diversity of human-computer interaction methods.

[0055] Next, the content display system involved in the embodiment of the present application is described. For illustration, please refer to Figure 1 The system involves a first terminal 110, a second terminal 120 and a server 130, wherein a communication connection exists among the first terminal 110, the second terminal 120 and the server 130.

[0056] Optionally, the first terminal 110 and the second terminal 120 run a live broadcast application, which can provide a live broadcast function. The first terminal 110 has a host account logged in, and the second terminal 120 has a viewer account logged in. After the host account opens a live broadcast room, the viewer account can enter the live broadcast room to watch the live content.

[0057] The server 130 is used to provide technical support for the implementation of various functions of the live broadcast application, for example, for live broadcast streaming data transmission, the server 130 is used to store data related to the live broadcast application, perform calculations on a large amount of data generated during the live broadcast process, and receive the live broadcast video stream and audio stream pushed by the first terminal 110, and stably transmit them to the second terminal 120. Or, live broadcast performance evaluation, etc., evaluate the live broadcast performance of the anchor and generate a performance score.

[0058] Exemplarily, the first terminal 110 receives and stores the first custom content uploaded by the host account. The types of the first custom content include, but are not limited to, at least one of text content, static / dynamic image content, video content, animation content, and special effect content.

[0059] In some embodiments, after receiving the first custom content, the first terminal 110 processes the first custom content to generate a content form suitable for display in the live broadcast room. For example, special effect content is generated based on the first custom content, with a dynamic change effect.

[0060] The host opens the live broadcast room through the first terminal 110, selects the song accompaniment and sings through the singing scoring function provided by the live broadcast application program. The first terminal 110 obtains the beat point data of the song accompaniment, and the beat point data includes each song beat point in the song accompaniment, including at least two target beat points. The target beat point is the beat point used to determine the moment to display the first custom content.

[0061] During the host's singing process, the server 130 will use the AI algorithm to score the host's singing level in real time, generate and display the singing score for each line of lyrics. When the song accompaniment plays to the target beat point, it is determined whether to display the first custom content based on the size relationship between the sum of the singing scores before the target beat point and the preset score threshold.

[0062] When the sum of the singing scores reaches the preset score threshold, the first custom content is automatically displayed; when the sum of the singing scores does not reach the preset score threshold, the first custom content is not displayed.

[0063] Since the number of target beat points is at least two, the above process is triggered each time a target beat point is reached.

[0064] In some embodiments, the manner of triggering the display of the first custom content based on the singing score can also be as follows.

[0065] For each target beat point, the singing scores of the host from the start of singing to the current target beat point are accumulated to obtain a phased singing score. The manner of displaying the first custom content at this target node is determined based on the score interval to which the phased singing score belongs, and different score intervals correspond to different display effects.

[0066] For example, there are 3 target beat points A, B, and C. The start timestamp of the host's singing is O, and the end timestamp is E. The entire singing process is divided into 4 stages: [O, A], [A, B], [B, C], [C, E], and the singing scores S1, S2, S3, S4 corresponding to the 4 stages are calculated respectively (S1, S2, S3, S4 are all real numbers).

[0067] Among them, the first stage singing score corresponding to the target beat point A is S1, the second stage singing score corresponding to the target beat point B is S1 + S2, and the third stage singing score corresponding to the target beat point C is S1 + S2 + S3.

[0068] In some embodiments, when the host finishes singing, the singing scores of the whole song can be accumulated to obtain the singing score of the whole song, and the singing score of the whole song S1 + S2 + S3 + S4 is obtained. Based on the score range to which the singing score of the whole song belongs, the first custom content is displayed in a corresponding form at the termination timestamp E.

[0069] For example, when the score range is [0, 50], the first custom content is displayed in the live broadcast room in the first size; when the score range is [50, 70], the first custom content is displayed in the live broadcast room in the second size; when the score range is [70, 90], the first custom content is displayed in the live broadcast room in the third size; when the score range is [90, 100], the first custom content is displayed in the live broadcast room in the fourth size.

[0070] Among them, there is a positive correlation between the upper limit of the score range and the size.

[0071] In some embodiments, if the sum of the singing scores between the current target beat point and the previous target beat point does not reach the preset score threshold, then regardless of the score range to which the stage singing score belongs, the first custom content is not displayed.

[0072] On the live broadcast interfaces of the first terminal 110 and the second terminal 120, there are live broadcast pictures of the host singing and the first custom content displayed when the trigger condition is met.

[0073] In some embodiments, when the host account is live broadcasting and the host does not appear in the live broadcast room, the default display content set in advance by the host account is displayed in the live broadcast picture.

[0074] In some embodiments, the process of scoring the host's singing process through the AI algorithm and triggering the display of the first custom content executed by the above server 130 can also be executed by the first terminal 110, and this embodiment does not limit this.

[0075] The above first terminal 110 and second terminal 120 can be terminal devices in various forms such as mobile phones, tablet computers, desktop computers, portable laptops, smart TVs, vehicle-mounted terminals, and smart home devices, and this application embodiment does not limit this.

[0076] It should be noted that the above server 130 can be an independent physical server, or a server cluster or distributed system composed of multiple physical servers, or a cloud server that provides basic cloud computing services such as cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communications, middleware services, domain name services, security services, content delivery networks (CDNs), and big data and artificial intelligence platforms.

[0077] In some embodiments, the above server 130 can also be implemented as a node in a blockchain system.

[0078] Combined with the above noun introduction and application scenarios, the content display method provided in this application will be described. This method can be executed by a server or a terminal, or jointly executed by a server and a terminal. In the embodiments of this application, an example in which this method is executed by a terminal will be described. As Figure 2 shown, Figure 2 is a flowchart of a content display method provided by an exemplary embodiment of this application. This method includes the following steps.

[0079] Step 210, receiving a content configuration operation for the first live broadcast room.

[0080] Among them, the content configuration operation is used to configure the first custom content to be displayed in the first live broadcast room according to a preset trigger requirement.

[0081] A live broadcast application program is installed in the terminal. The host can start the first live broadcast room by logging in to the host account in the live broadcast application program and upload the first custom content through the content configuration operation. The audience can enter the first live broadcast room by logging in to the audience account in the live broadcast application program to watch the live broadcast content.

[0082] Optionally, the first custom content is the content uploaded by the host account, and the types of the first custom content include but are not limited to at least one of the following.

[0083] 1. Text content;

[0084] 1.1 Host announcements: For example, announcements about the preview of the next live broadcast time, the preview of the next live broadcast content, the program / activity arrangement of this live broadcast, etc.;

[0085] 1.2 Interactive topics: For example, topic content used to guide the audience to participate in interactive discussions;

[0086] 1.3 Thank-you messages: For example, words used by the host to express gratitude for the support and interaction of the audience;

[0087] 2. Image content;

[0088] 2.1 Static images: for example, photos / pictures uploaded locally by the host related to the live streaming theme;

[0089] 2.2 Dynamic images: for example, special effects / gifs / gif emojis with interactive effects;

[0090] 3. Video content;

[0091] 3.1 Background videos: for example, dynamic videos used as the background of the live streaming room, such as video materials of natural scenery, city night views, etc.;

[0092] 3.2 Animated videos: for example, animated special effect videos;

[0093] 4. Audio content;

[0094] 4.1 Background music: for example, adding background music to the live streaming room, or switching the original music in the live streaming room to the current background music;

[0095] 4.2 Sound effects: for example, sound effects related to the host's expressions;

[0096] 4.3 Voice content: for example, voice audio pre-recorded by the host;

[0097] 5. Interactive control content;

[0098] 5.1 Voting control: for example, a control used to initiate a voting activity. The host can set various voting questions and voting options, and the audience can trigger different options through the terminal screen to vote;

[0099] 5.2 Lottery control: for example, a control that the host can set for a lottery activity, and display content including lottery rules, number of lottery draws, winning list display, etc. The audience can trigger the control through the terminal screen to participate in the lottery;

[0100] 5.3 Links: for example, product links, activity links, social media homepage links recommended by the host, etc. The audience can click on the link to jump to the specified page or view the specified information.

[0101] Among them, the quantity and type of the first custom content can be arbitrary. When there are multiple first custom contents, the triggering requirements corresponding to each first custom content can be the same or different.

[0102] Optionally, receive at least two content configuration operations for the first live streaming room. The at least two content configuration operations are respectively used to configure the first custom content displayed according to different triggering requirements, where the first custom contents corresponding to different triggering requirements are different.

[0103] Exemplarily, at least two content configuration operations are configured with t first custom contents, where the tᵢ-th first custom content corresponds to the tᵢ-th trigger requirement, t is a positive integer, and tᵢ is a positive integer not exceeding t.

[0104] For example, t = 2. The 1st first custom content is a static image, and the 1st trigger requirement is that a specified song starts playing in the first live room, which is used to switch the background image of the first live room to display this static image; the 2nd first custom content is special effect content, and the 2nd trigger requirement is that the behavior performance of the subject appearing in the first live room meets the specified behavior standard (e.g., the volume of the subject reaches the preset decibel), which is used to display the special effect content in the first live room in a highlighted and flashing form.

[0105] Step 220: Display the live content of the first live room.

[0106] Among them, the live content includes the content presented by the first subject according to the preset content template.

[0107] The preset content template is a reference benchmark for guiding the first subject to perform and act in the first live room. The content expressed in the preset content template and the content expressed by the first subject in the live content belong to the same type.

[0108] Optionally, the content types of the preset content template and the live content and corresponding examples are as follows:

[0109] (1) Audible content: The live content is singing content, and the preset content template is the song accompaniment (including the original singer's vocals); or, the live content is dubbing content, and the preset content template is the original sound of the radio drama (including the original dubbing vocals); or, the live content is the performance content of the first subject playing a specified piece on an instrument; the preset template content is the audio content of playing the same piece on the same instrument;

[0110] (2) Visual content: The live content is action content, and the preset content template is an action reference video (including the picture of the subject performing the specified action), for example, a dance video, a broadcast exercise video, a martial arts teaching video, etc.

[0111] Step 230: When the matching relationship between the live content and the content template meets the preset trigger requirement, display the first custom content in the first live room.

[0112] Optionally, when the matching degree between the live content and the content template reaches the preset matching degree threshold, display the first custom content in the first live room.

[0113] Exemplarily, based on the matching degree between the live content and the content template, display the first custom content in the first live room with a highlighting performance intensity corresponding to the matching degree, where there is a positive correlation between the matching degree and the highlighting performance intensity.

[0114] Divide different highlighting performance intensities based on the matching degree interval to which the matching degree belongs. The size of the upper limit value of the matching degree interval is positively correlated with the highlighting performance intensity.

[0115] For example, the live content is singing content, and the content template is a song accompaniment (including the original singer's vocals). Use the content template as a benchmark to analyze the matching degree between it and the live content. The first custom content is preset multimedia content, with a quantity of 1, and the highlighting performance intensity refers to the display brightness when displaying the first custom content.

[0116] There are a total of 3 matching degree intervals, which are as follows: (0, a), (a, b), (b, c), where a < b < c, and a, b, and c are all preset positive numbers.

[0117] The matching degree interval (0, a) corresponds to the first display brightness, the matching degree interval (a, b) corresponds to the second display brightness, and the matching degree interval (b, c) corresponds to the third display brightness. The display brightness increases as the upper limit of the matching degree interval increases.

[0118] In some embodiments, at least two first custom contents can be set, and one of the first custom contents is selected for display according to the matching degree between the live content and the content template, and the display of other first custom contents is cancelled.

[0119] Exemplarily, based on the matching degree between the live content and the content template, display the first custom content corresponding to the matching degree in the first live room, where the first custom content corresponding to the matching degree in different preset matching degree intervals is different.

[0120] Different preset matching degree intervals correspond to different first custom contents, and the highlighting performance intensities of different custom contents are the same.

[0121] Among them, the size of the upper limit value of the matching degree interval is positively correlated with the highlighting performance intensity. For different types of custom contents, their highlighting performance intensities are different intensities divided based on the common characteristic attributes of each type of custom content as the performance basis.

[0122] Exemplarily, the prominent performance types and different intensity examples are as follows: (1) The prominent performance type refers to the size of the first custom content in the live broadcast screen, and the prominent performance intensity refers to the size of the first custom content; (2) The prominent performance type refers to the display brightness of the first custom content, and the prominent performance intensity refers to the brightness of the first custom content; (3) The prominent performance type refers to adding a flashing effect to the first custom content, and the prominent performance intensity refers to the flashing frequency of the flashing effect.

[0123] For example, the number of the first custom content is 3, which are different preset multimedia contents: Custom Content 1, Custom Content 2, and Custom Content 3.

[0124] There are 3 matching degree intervals, which are as follows: (0, a), (a, b), (b, c), where a < b < c, and a, b, and c are all preset positive numbers.

[0125] The matching degree interval (0, a) corresponds to Custom Content 1, the matching degree interval (a, b) corresponds to Custom Content 2, and the matching degree interval (b, c) corresponds to Custom Content 3. The first custom content corresponding to different matching degree intervals can be specified by the user or randomly assigned.

[0126] Exemplarily, taking the first subject as the live broadcaster and the scenario where the first subject sings in the first live broadcast room as an example for illustration.

[0127] Optionally, the live broadcast content includes the singing content of the first subject, and the content template includes the song accompaniment.

[0128] Obtain the singing score that matches between the singing content and the song accompaniment.

[0129] Exemplarily, perform feature extraction on the singing content through an AI algorithm, convert it into a first feature representation, perform feature extraction on the song accompaniment in the same way, and convert it into a second feature representation. The feature representation dimensions of the first feature representation and the second feature representation are the same.

[0130] Calculate the cosine similarity between the first feature representation and the second feature representation, and use the cosine similarity as the singing score.

[0131] When the singing score reaches the score threshold, display the first custom content in the first live broadcast room.

[0132] In order to increase the number of times the first custom content is triggered to be displayed in the first live broadcast room, the singing score will be calculated periodically during the singing process of the first subject, and it is determined whether to trigger the display of the first custom content based on the singing score of each stage, which can enrich the content richness in the live broadcast room and improve the live broadcast effect.

[0133] Exemplarily, obtain the beat point data of the song accompaniment, where the beat point data contains the timestamps of the song beat points of the song accompaniment in the audio of the song accompaniment.

[0134] Analyze the beat point data to determine the target beat points of the song accompaniment, and the timestamps of the target beat points are used to divide the singing scoring segments.

[0135] Among them, the target beat points meet at least one of the following conditions: the difference between the beat intensity of the target beat point and the adjacent beat point reaches a preset intensity threshold, the sound frequency at the target beat point reaches a maximum value, and the target beat point corresponds to the start timestamp of the chorus in the song accompaniment.

[0136] Exemplarily, the total duration of the song accompaniment is 3 minutes (00:00 - 03:00), and there are 3 target beat points in total. Their start timestamps are as follows: the timestamp of target beat point 1 in the song accompaniment is 00:30, the timestamp of target beat point 2 in the song accompaniment is 01:30, and the timestamp of target beat point 3 in the song accompaniment is 02:30.

[0137] Divide the song accompaniment into at least one singing scoring segment based on the timestamps of the target beat points. The i-th singing scoring segment corresponds to the i-th time period in the song accompaniment, where i is a positive integer.

[0138] k target beat points divide the song accompaniment into k singing scoring segments, where k is a positive integer.

[0139] Exemplarily, the methods of dividing the singing scoring segments include the following several types:

[0140] (1) When the timestamp of the target beat point does not coincide with the start timestamp of the song accompaniment, combine the start timestamp of the song accompaniment and the timestamps corresponding to each target beat point, and divide the time period between two adjacent timestamps into a singing scoring segment.

[0141] For example, the above 3 target beat points divide the song accompaniment into the following singing scoring segments:

[0142] 【00:00 - 00:30】, 【00:30 - 01:30】, 【01:30 - 02:30】. For the segment between 【02:30 - 03:00】, its singing score can be obtained, but only the first custom content is triggered and displayed at the target beat point.

[0143] (2) Divide each time period between the timestamp of each target beat point and the start timestamp of the song accompaniment into a singing scoring segment.

[0144] For example, the above 3 target beat points divide the song accompaniment into the following singing scoring segments:

[0145] For the segments / complete differences within 【00:00~03:00】, such as 【00:00~00:30】, 【00:00~01:30】, and 【00:00~02:30】, the singing score can be obtained, but the first custom content is only displayed at the target beat points.

[0146] For the i-th singing score segment, the i-th singing segment corresponding to the i-th singing score segment is obtained from the singing content.

[0147] Align the start and end timestamps of the i-th singing score segment with those of the i-th singing segment.

[0148] Calculate the similarity score between the i-th singing score segment and the i-th singing segment as the singing score for the match between the singing content and the song accompaniment within the i-th time period.

[0149] In some embodiments, the target beat point can also be specified by the user, that is, the moment to trigger the display of the first custom content is user-defined.

[0150] Optionally, in response to receiving a beat point setting operation, the specified beat point in the beat point data is set as the target beat point.

[0151] Exemplarily, upon receiving a beat point setting operation, the live application interface displays the beat point data of the song accompaniment, which includes the timestamp of each song beat point in the song accompaniment and the corresponding relationship between each song beat point and the lyrics.

[0152] Upon receiving a selection operation for at least two beat points in the beat point data, the selected specified beat points are set as the target beat points.

[0153] The display method of the first custom content includes, but is not limited to: (1) The application automatically processes / edits the first custom content, determines the display position of the first custom content, and obtains the processed first custom content. When the performance of the first subject meets the trigger requirements, the first custom content is displayed at the specified position on the live broadcast screen; (2) The user who uploads the first custom content defines its display position, whether to process / edit it, and the method during processing / editing.

[0154] The following is an example of the application automatically processing the first custom content.

[0155] When the first custom content is static content, a dynamic performance animation corresponding to the first custom content is generated. Based on the beat point data, the first performance form of the dynamic performance animation is determined, where the first performance form matches the song beat points.

[0156] Display a dynamic performance animation in the first live broadcast room in a first presentation form.

[0157] For example, the first custom content is a static image that contains target content (such as an item / animal / plant / person / virtual object). After identifying the target content, perform segmentation / matting processing on the static image, crop the target image along the shape of the target content, and add dynamic special effects (such as zooming / fading / transforming from 2D to 3D) to the target image to obtain a dynamic performance animation.

[0158] Obtain multiple beat points adjacent to the target beat point based on the beat point data, and determine that the first presentation form is to display the dynamic performance animation with corresponding brightness according to the rhythm intensity of the target beat point and the adjacent beat points. The rhythm intensity is positively correlated with the display brightness of the dynamic performance animation.

[0159] In summary, the content display method provided in this application can meet the personalized setting requirements of users for the display content in the live broadcast room, enrich the diversity of the live broadcast screen, and increase the human-computer interaction methods by triggering the display of custom content according to the live broadcast performance of the first subject during the live broadcast through custom content pre-configured with trigger conditions. When the performance of the subject meets the trigger condition, an immediate feedback is generated, which improves the human-computer interaction efficiency and the user experience during the live broadcast, encourages the subject to enhance the enthusiasm for the live broadcast performance to trigger the display of custom content, and enhances the live broadcast effect.

[0160] In addition to the singing scenario listed in the above first embodiment, there is also an action scenario corresponding in the first live broadcast room. For example, the live broadcast content includes the action content of the first subject, and the content template includes an action video. Figure 3 It is a flowchart of the content display method provided by another exemplary embodiment of this application. The above step 230 can also be implemented as the following steps 310 to 320.

[0161] Step 310, obtain an action score for the match between the action content and the action video.

[0162] Collect action images of the action content based on the first frequency, and collect video images of the action video based on the first frequency to obtain at least two pairs of action score images.

[0163] Among them, the mth pair of action score images contains the mth action image and the mth video image collected at the mth moment, where m is a positive integer.

[0164] Exemplarily, when the live broadcast content is action content, the action video contains the action pictures of the reference subject. The action pictures based on the action video are displayed by the first subject in the live broadcast screen, and the playing action video is synchronously displayed.

[0165] Take screenshots of the live broadcast screen in real time to obtain action score image pairs. The captured images include the action images of the first subject performing an action and the video images of the reference subject performing an action.

[0166] Take screenshots of the live broadcast screen at the first frequency to obtain at least two action score image pairs.

[0167] Calculate the similarity score between the m-th action image and the m-th video image to obtain the action score of the m-th action score image pair.

[0168] Exemplarily, input the m-th action score image pair into a pre-trained model and output the action score result.

[0169] Exemplarily, the model analyzes the m-th action image to obtain the first key point data of the first subject corresponding to the m-th action image. The first key point data includes at least two limb key points of the first subject; the model analyzes the m-th video image to obtain the second key point data of the reference subject corresponding to the m-th video image. The second key point data includes at least two limb key points of the reference subject.

[0170] Extract features from the first key point data and the second key point data respectively, and map them into a three-dimensional space to obtain a first vector corresponding to the first key point data and a second vector corresponding to the second key point data.

[0171] Calculate the cosine similarity between the first vector and the second vector as the action score of the m-th action score image pair.

[0172] Repeat the above steps, and obtain the action score for the matching between the action content and the action video based on the sum of the action scores of at least two action score image pairs respectively.

[0173] For example, the entire action content and the action video are divided into q action score image pairs. The action scores corresponding to the q action score image pairs are s1, s2... sq respectively, where s1, s2... sq are all real numbers. Calculate the mean of s1, s2... sq to obtain the complete action score between the entire action content and the action video, which is used to reflect the matching degree between the action content performed by the first subject and the action video.

[0174] In some embodiments, the action video can also be divided into at least two action score intervals, calculate the action scores of each action score interval, and display the first custom content in the first live broadcast room when the action score of each interval reaches a preset score threshold.

[0175] Among them, the trigger display moments of the first custom content include the following: (1) the moments specified by the user; (2) obtaining the audio part of the action video, analyzing the beat point data of the audio part, and determining the moment corresponding to the target beat point as the trigger display moment of the first custom content. The determination method of the target beat point refers to step 230 and will not be elaborated here; (3) analyzing the action video, obtaining at least one target action frame in the action video, and determining the moment corresponding to the target action frame as the trigger display moment of the first custom content. Among them, the target action frame refers to the change in the action of the reference subject in the action video that conforms to the preset change requirements compared with its previous frame.

[0176] Exemplarily, the division method can be automatically divided by the live application program (for example, evenly divided / randomly divided based on the duration) or user-defined.

[0177] For each action score interval, based on the start and end timestamps of the action score interval and the timestamps of the action score image pairs collected, the average value of the action scores corresponding to the action score image pairs belonging to the same action score interval is taken to obtain the interval action scores corresponding to each action score interval.

[0178] Step 320, when the action score reaches the score threshold, display the first custom content in the first live room.

[0179] Optionally, when there is one first custom content set, based on the matching degree / action score between the action content and the action video, display the first custom content in the first live room with the highlighting performance intensity corresponding to the matching degree, where there is a positive correlation between the matching degree and the highlighting performance intensity.

[0180] Optionally, when there are multiple first custom contents set, based on the score interval to which the matching degree / action score between the action content and the action video belongs, display the first custom content corresponding to the matching degree in the first live room, where the first custom contents corresponding to the matching degrees in different preset score intervals are different.

[0181] Exemplarily, by obtaining the audio part of the action video, analyzing the beat point data of the audio part, and determining the moment corresponding to the target beat point as the trigger display moment of the first custom content, the number of target beat points is at least two, and the number of the first custom contents is 1.

[0182] Exemplarily, the total duration of the action video is 3 minutes (00:00 - 03:00), and there are 4 target beat points in its audio part. Their start timestamps are as follows: The timestamp of target beat point 1 in the action video is 00:30, the timestamp of target beat point 2 in the action video is 01:20, the timestamp of target beat point 3 in the action video is 01:50, and the timestamp of target beat point 4 in the action video is 03:00.

[0183] Combining the start timestamp of the action video and the timestamps corresponding to each target beat point, the time period between two adjacent timestamps is divided into an action scoring interval.

[0184] For example, among the above 4 target beat points, if target beat point 4 coincides with the end timestamp of the action video, then based on the above 4 target beat points, the action video is divided into the following action scoring intervals:

[0185] 【00:00 - 00:30】, 【00:30 - 01:20】, 【01:20 - 01:50】, 【01:50 - 03:00】.

[0186] For the interval 【00:00 - 00:30】, the corresponding scoring threshold is 80. It contains 30 action scoring image pairs, and the average action score is 90, reaching the scoring threshold. When the action video plays to the 00:30 moment, the first custom content is displayed in the first live room.

[0187] For the interval 【00:30 - 01:20】, the corresponding scoring threshold is 80. It contains 50 action scoring image pairs, and the average action score is 85, reaching the scoring threshold. When the action video plays to the 01:20 moment, the first custom content is displayed in the first live room.

[0188] For the interval 【01:20 - 01:50】, the corresponding scoring threshold is 80. It contains 30 action scoring image pairs, and the average action score is 75, not reaching the scoring threshold. When the action video plays to the 01:50 moment, the first custom content is not displayed either.

[0189] For the interval 【01:50 - 03:00】, the corresponding scoring threshold is 70. It contains 70 action scoring image pairs, and the average action score is 75, reaching the scoring threshold. When the action video plays to the 03:00 moment, the first custom content is displayed in the first live room.

[0190] That is to say, in the above example, the first custom content is displayed 3 times in the first live room.

[0191] In summary, the content display method provided by this application can meet the user's personalized setting requirements for the content displayed in the live broadcast room, enrich the diversity of the live broadcast screen, and increase the human-computer interaction methods by triggering the display of custom content according to the live performance of the first subject during the live broadcast with the custom content pre-configured with trigger conditions. When the performance of the subject meets the trigger conditions, immediate feedback is generated, improving the human-computer interaction efficiency and the user experience during the live broadcast, encouraging the subject to enhance the enthusiasm for live performance to trigger the display of custom content, and enhancing the live broadcast effect.

[0192] Figure 4 FIG. 4 is a flowchart of a content display method provided by another exemplary embodiment of this application, and this method includes the following steps.

[0193] Step 410, receive a content configuration operation for the first live broadcast room.

[0194] The content configuration operation is used to configure the first custom content to be displayed in the first live broadcast room according to a preset trigger requirement.

[0195] A live broadcast application program is installed in the terminal. The host can start the first live broadcast room by logging in to the host account in the live broadcast application program and upload the first custom content through the content configuration operation. The audience can enter the first live broadcast room to watch the live broadcast content by logging in to the audience account in the live broadcast application program.

[0196] Optionally, the first custom content is the content uploaded by the host account, and the content type and quantity of the first custom content can be arbitrary.

[0197] Step 420, display the live broadcast content of the first live broadcast room.

[0198] Among them, the live broadcast content includes the content presented by the first subject according to a preset content template.

[0199] The preset content template is a reference benchmark for guiding the first subject to perform and conduct activities in the first live broadcast room, and the content expressed in the preset content template and the content expressed by the first subject in the live broadcast content belong to the same type.

[0200] Step 430, receive a co-hosting request.

[0201] Among them, the co-hosting request is a request from the audience account to co-host with the first subject in the first live broadcast room.

[0202] After the audience account enters the first live broadcast room, it can initiate a co-hosting request to the host account based on the interaction function provided by the live broadcast application program to establish a real-time voice or video connection to achieve co-live interaction in the first live broadcast room.

[0203] Exemplarily, the audience account corresponds to a second subject, where the second subject is different from the first subject.

[0204] When the co-hosting request fails, the main body performing in the first live room remains the first main body; when the co-hosting request is approved, the main bodies performing in the first live room include the first main body and the second main body.

[0205] For example, if the co-hosting request is a voice connection request, the second main body will not be displayed in the video of the first live room, but the audio sent by the second main body through the co-hosting request will be played in real time, so that other viewers, the first main body, and the second main body in the first live room can all perceive the voice performance of the second main body.

[0206] If the co-hosting request is a video connection request, the second main body may be displayed in the video of the first live room.

[0207] When the second main body turns on the camera through the terminal logged in with the viewer account and the camera captures the second main body itself, the second main body will be correspondingly displayed in the video of the first live room.

[0208] Step 440: In response to the co-hosting request, receive the operation of obtaining at least one candidate custom content.

[0209] Among them, the candidate custom content is the custom content provided by the host account, and the types of the candidate custom content include but are not limited to at least one of the following.

[0210] 1. Text content; for example, host announcements, interactive topics, gratitude messages;

[0211] 2. Image content; for example, static images, dynamic images;

[0212] 3. Video content; for example, background videos, animated videos;

[0213] 4. Audio content; for example, background music, sound effects, voice content;

[0214] 5. Interactive control content; for example, voting controls, lottery controls, links.

[0215] Step 450: Send at least one candidate custom content to the viewer account.

[0216] Among them, the candidate custom content is used to provide the second custom content displayed for the viewer account when the preset trigger requirements are met during the co-hosting process.

[0217] Exemplarily, after the viewer account and the host account successfully co-host, the video of the first live room includes two display areas: the first display area is used to display the live video corresponding to the host account, and the second display area is used to display the co-hosting video corresponding to the viewer account.

[0218] Step 460, in the case of meeting the preset trigger requirements during the co-hosting process, display the second custom content for the viewer account.

[0219] Among them, the first custom content and the second custom content are independent of each other when displayed, that is, the display situation of the first custom content depends on the live performance of the first entity, and the display situation of the second custom content depends on the live performance of the second entity.

[0220] Optionally, in the case where the matching relationship between the live content and the content template meets the preset trigger requirements, display the first custom content in the first live room.

[0221] Among them, the preset trigger requirement refers to the requirement for triggering the display of the second custom content. When the performance of the second entity corresponding to the viewer account during the co-hosting process meets the preset trigger requirements, the corresponding second custom content will be displayed for the viewer account.

[0222] The second custom content is the content selected by the viewer account from the candidate custom content after receiving the candidate custom content. In some embodiments, the viewer account can not only select the type of the second custom content, but also set the preset trigger requirements for the display of the second custom content, and this setting process takes effect after being confirmed by the host account.

[0223] Among them, the live content includes the first live content and the second live content, and they are presented in the live video screen at the same time. Among them, the first live content is the content presented by the first entity according to the preset content template, and the second live content is the content presented by the second entity according to the preset content template.

[0224] During the co-hosting process, in response to the matching relationship between the first live content and the content template meeting the preset trigger requirements corresponding to the first custom content, display the first custom content in the first live room; and, in response to the matching relationship between the second live content and the content template meeting the preset trigger requirements corresponding to the second custom content, display the second custom content for the viewer account.

[0225] Exemplarily, the live content of the first live room is singing content, the preset content template is the song accompaniment, and the first entity and the second entity sing different segments of the same song together in the first live room.

[0226] Obtain the beat point data of the song accompaniment and analyze it to obtain the target beat points in the beat point data. Based on the target beat points, divide the song accompaniment into at least two first scoring segments corresponding to the first entity and at least two second scoring segments corresponding to the second entity.

[0227] Exemplarily, for the x-th first scoring segment, obtain the x-th first singing segment corresponding to the x-th first scoring segment from the first live content, where x is a positive integer.

[0228] Align the start and end timestamps of the x-th first scoring segment with those of the x-th first singing segment.

[0229] Calculate the similarity score between the x-th first scoring segment and the x-th first singing segment as the singing score for their matching.

[0230] When the singing score reaches the preset scoring threshold, display the first custom content at the target beat point corresponding to the x-th first scoring segment.

[0231] Exemplarily, for the y-th second scoring segment, obtain the y-th second singing segment corresponding to the y-th second scoring segment from the second live content, where y is a positive integer.

[0232] Align the start and end timestamps of the y-th second scoring segment with those of the y-th second singing segment.

[0233] Calculate the similarity score between the y-th second scoring segment and the y-th second singing segment as the singing score for their matching.

[0234] When the singing score reaches the preset scoring threshold, display the second custom content at the target beat point corresponding to the y-th second scoring segment.

[0235] In summary, the content display method provided by this application, through the custom content with pre-configured trigger conditions, can meet the user's personalized setting requirements for the display content in the live broadcast by triggering the display of custom content according to the live performance of the first subject during the live broadcast, enrich the diversity of the live broadcast screen, and increase the human-computer interaction methods. When the performance of the subject meets the trigger conditions, instant feedback is generated, improving the human-computer interaction efficiency and the user experience during the live broadcast, encouraging the subject to enhance the enthusiasm for the live performance to trigger the display of custom content, and enhancing the live broadcast effect.

[0236] Figure 5 is the structural block diagram of a content display device provided by an exemplary embodiment of this application, as Figure 5 shown, the device includes the following parts.

[0237] A receiving module 510, configured to receive a content configuration operation for the first live broadcast room, where the content configuration operation is used to configure the first custom content to be displayed in the first live broadcast room according to a preset trigger requirement;

[0238] A content display module 520 is configured to display the live content of the first live broadcast room, where the live content includes the content presented by the first subject according to a preset content template;

[0239] The content display module 520 is further configured to display the first custom content in the first live broadcast room when the matching relationship between the live content and the content template meets the preset trigger requirement.

[0240] In an optional embodiment, the content display module 520 is further configured to display the first custom content in the first live broadcast room when the matching degree between the live content and the content template reaches a preset matching degree threshold.

[0241] In an optional embodiment, the receiving module 510 is further configured to receive at least two content configuration operations for the first live broadcast room, where the at least two content configuration operations are respectively used to configure the first custom content displayed according to different trigger requirements, and the first custom content corresponding to different trigger requirements is different;

[0242] The content display module 520 is further configured to display the first custom content corresponding to the matching degree in the first live broadcast room based on the matching degree between the live content and the content template, where the first custom content corresponding to different preset matching degree intervals is different.

[0243] In an optional embodiment, the content display module 520 is further configured to display the first custom content in the first live broadcast room with a highlighting performance intensity corresponding to the matching degree based on the matching degree between the live content and the content template, where there is a positive correlation between the matching degree and the highlighting performance intensity.

[0244] In an optional embodiment, the receiving module 510 is further configured to receive a co-hosting request, where the co-hosting request is a request from a viewer account to co-host with the first subject in the first live broadcast room; in response to the co-hosting request, receive an operation to obtain at least one candidate custom content;

[0245] The apparatus further includes:

[0246] A sending module 530 is configured to send the at least one candidate custom content to the viewer account, where the candidate custom content is used to provide the second custom content displayed for the viewer account when the preset trigger requirement is met during the co-hosting process.

[0247] In an optional embodiment, the live content includes the singing content of the first subject, and the content template includes a song accompaniment;

[0248] The content display module 520 is further configured to obtain a singing score for the matching between the singing content and the song accompaniment; and when the singing score reaches a score threshold, display the first custom content in the first live room.

[0249] In an optional embodiment, the content display module 520 is further configured to obtain beat point data of the song accompaniment, where the beat point data includes timestamps of song beat points of the song accompaniment in the audio of the song accompaniment; analyze the beat point data to determine target beat points of the song accompaniment, and the timestamps of the target beat points are used to divide singing score segments, and the target beat points meet at least one of the following conditions: the difference between the beat intensity of the target beat point and that of an adjacent beat point reaches a preset intensity threshold, the sound frequency at the target beat point reaches a maximum value, the target beat point corresponds to the start timestamp of the chorus in the song accompaniment; divide the song accompaniment into at least one singing score segment based on the timestamps of the target beat points, where the i-th singing score segment corresponds to the i-th time period in the song accompaniment, and i is a positive integer; for the i-th singing score segment, obtain the i-th singing segment corresponding to the i-th singing score segment from the singing content; calculate a similarity score between the i-th singing score segment and the i-th singing segment as the singing score for the matching between the singing content and the song accompaniment in the i-th time period.

[0250] In an optional embodiment, the receiving module 510 is further configured to, in response to receiving a beat point setting operation, set a specified beat point in the beat point data as the target beat point.

[0251] In an optional embodiment, when the first custom content is static content, the content display module 520 is further configured to generate a dynamic performance animation corresponding to the first custom content; determine a first performance form of the dynamic performance animation based on the beat point data, where the first performance form matches the song beat points; and display the dynamic performance animation in the first live room in the first performance form.

[0252] In an optional embodiment, the live content includes action content of a first subject, and the content template includes an action video.

[0253] The content display module 520 is further configured to obtain an action score for the matching between the action content and the action video; and when the action score reaches a score threshold, display the first custom content in the first live room.

[0254] In an alternative embodiment, the content display module 520 is further configured to collect action images of the action content based on a first frequency, and video images of the action video based on the first frequency, to obtain at least two pairs of action scoring images, where the m-th pair of action scoring images includes the m-th action image and the m-th video image collected at the m-th moment, and m is a positive integer; calculate a similarity score between the m-th action image and the m-th video image to obtain an action score of the m-th pair of action scoring images; and obtain an action score for the match between the action content and the action video based on the sum of the action scores of the at least two pairs of action scoring images respectively.

[0255] In an alternative embodiment, the receiving module 510 is further configured to receive a display mode configuration operation for the first custom content, where the display mode configuration operation is used to indicate a manner of displaying the first custom content in combination with the live broadcast screen of the first live broadcast room.

[0256] The content display module 520 is further configured to display the first custom content that matches the current live broadcast screen of the first live broadcast room in the first live broadcast room.

[0257] In summary, the content display device provided in this application can meet the user's personalized setting requirements for the display content in the live broadcast room, enrich the diversity of the live broadcast screen, and increase the human-computer interaction method by triggering the display of custom content according to the live broadcast performance of the first subject during the live broadcast through the custom content with pre-configured trigger conditions. When the performance of the subject meets the trigger conditions, instant feedback is generated, improving the human-computer interaction efficiency and the user experience during the live broadcast, encouraging the subject to enhance the enthusiasm for the live broadcast performance to trigger the display of custom content, and enhancing the live broadcast effect.

[0258] It should be noted that: for the content display device provided in the above embodiment, only the above-mentioned division of each functional module is used for illustration. In practical applications, the above functions can be allocated to different functional modules according to needs, that is, the internal structure of the device is divided into different functional modules to complete all or part of the functions described above. In addition, the content display device provided in the above embodiment and the content display method embodiment belong to the same concept, and the specific implementation process can be found in the method embodiment, which will not be elaborated here.

[0259] Figure 7The structural block diagram of a computer device 700 provided by an exemplary embodiment of the present application is shown. The computer device 700 may be: a smart phone, a tablet computer, a Moving Picture Experts Group Audio Layer III player (MP3), a Moving Picture Experts Group Audio Layer IV (MP4) player, a laptop computer, or a desktop computer. The computer device 700 may also be referred to by other names such as a user device, a portable terminal, a laptop terminal, a desktop terminal, etc.

[0260] Generally, the computer device 700 includes: a processor 701 and a memory 702.

[0261] The processor 701 may include one or more processing cores, such as a 4-core processor, an 8-core processor, etc. The processor 701 may be implemented in at least one hardware form of Digital Signal Processing (DSP), Field-Programmable Gate Array (FPGA), or Programmable Logic Array (PLA). The processor 701 may also include a main processor and a coprocessor. The main processor is a processor for processing data in the wake state, also known as the Central Processing Unit (CPU); the coprocessor is a low-power processor for processing data in the standby state. In some embodiments, the processor 701 may be integrated with a Graphics Processing Unit (GPU), and the GPU is responsible for rendering and drawing the content to be displayed on the display screen. In some embodiments, the processor 701 may also include an Artificial Intelligence (AI) processor, and the AI processor is used to process computational operations related to machine learning.

[0262] The memory 702 may include one or more computer-readable storage media, and the computer-readable storage media may be non-transitory. The memory 702 may also include high-speed random access memory and non-volatile memory, such as one or more disk storage devices, flash storage devices. In some embodiments, the non-transitory computer-readable storage media in the memory 702 is used to store at least one instruction, and the at least one instruction is used to be executed by the processor 701 to implement the content display method provided by the method embodiment of the present application.

[0263] In some embodiments, the computer device 700 further includes some other components 703, and the types and quantities of the other components 703 can be selected based on the functional requirements of the computer device 700. Those skilled in the art can understand that Figure 7 the structure shown in

[0264] does not constitute a limitation on the computer device 700, and it may include more or fewer components than shown in the figure, or combine certain components, or adopt different component arrangements. Optionally, the computer-readable storage medium may include: Read Only Memory (ROM), Random Access Memory (RAM), Solid State Drives (SSD), or optical discs, etc. Among them, the random access memory may include Resistance Random Access Memory (ReRAM) and Dynamic Random Access Memory (DRAM). The serial numbers of the embodiments of the present application are only for description and do not represent the advantages and disadvantages of the embodiments.

[0265] The embodiments of the present application further provide a computer device, which includes a processor and a memory. At least one instruction, at least one program, a code set, or an instruction set is stored in the memory, and the at least one instruction, the at least one program, the code set, or the instruction set is loaded and executed by the processor to implement the content display method as described in any one of the above embodiments of the present application.

[0266] The embodiments of the present application further provide a computer-readable storage medium, in which at least one instruction, at least one program, a code set, or an instruction set is stored, and the at least one instruction, the at least one program, the code set, or the instruction set is loaded and executed by a processor to implement the content display method as described in any one of the above embodiments of the present application.

[0267] The embodiments of the present application further provide a computer program product or a computer program. The computer program product or the computer program includes computer instructions, and the computer instructions are stored in a computer-readable storage medium. The processor of the computer device reads the computer instructions from the computer-readable storage medium, and the processor executes the computer instructions, so that the computer device executes the content display method as described in any one of the above embodiments.

[0268] Those of ordinary skill in the art can understand that all or part of the steps to implement the above embodiments can be completed by hardware, or can be completed by instructing relevant hardware through a program. The program can be stored in a computer-readable storage medium. The above-mentioned storage medium can be a read-only memory, a disk, an optical disc, etc.

[0269] The above are only alternative embodiments of the present application and are not intended to limit the present application. Any modifications, equivalent replacements, improvements, etc. made within the spirit and principle of the present application shall be included within the protection scope of the present application.

Claims

1. A content display method, characterized in that, The method includes: Receiving a content configuration operation for a first live streaming room, where the content configuration operation is used to configure first custom content to be displayed in the first live streaming room according to a preset trigger requirement; Displaying the live streaming content of the first live streaming room, where the live streaming content includes content presented by a first entity according to a preset content template; When the matching relationship between the live streaming content and the content template meets the preset trigger requirement, displaying the first custom content in the first live streaming room.

2. The method according to claim 1, wherein The step of, when the matching relationship between the live streaming content and the content template meets the preset trigger requirement, displaying the first custom content in the first live streaming room includes: When the matching degree between the live streaming content and the content template reaches a preset matching degree threshold, displaying the first custom content in the first live streaming room.

3. The method according to claim 2, characterized in that, The step of receiving a content configuration operation for a first live streaming room includes: Receiving at least two content configuration operations for the first live streaming room, where the at least two content configuration operations are respectively used to configure first custom content to be displayed according to different trigger requirements, and where the first custom content corresponding to different trigger requirements is different; The step of, when the matching degree between the live streaming content and the content template reaches a preset matching degree threshold, displaying the first custom content in the first live streaming room includes: Based on the matching degree between the live streaming content and the content template, displaying in the first live streaming room the first custom content corresponding to the matching degree, where the first custom content corresponding to different preset matching degree intervals is different.

4. The method according to claim 2, wherein The step of, when the matching degree between the live streaming content and the content template reaches a preset matching degree threshold, displaying the first custom content in the first live streaming room includes: Based on the matching degree between the live streaming content and the content template, displaying the first custom content in the first live streaming room with a highlighting performance intensity corresponding to the matching degree, where there is a positive correlation between the matching degree and the highlighting performance intensity.

5. The method according to any one of claims 1 to 4, characterized in that, The method further includes: Receiving a co-hosting request, where the co-hosting request is a request from a viewer account to co-host with the first entity in the first live streaming room; In response to the co-hosting request, receiving an operation to obtain at least one candidate custom content; Sending the at least one candidate custom content to the viewer account, where the candidate custom content is used to provide, to the viewer account, second custom content to be displayed for the viewer account when the preset trigger requirement is met during the co-hosting process.

6. The method according to any one of claims 1 to 4, characterized in that The live streaming content includes singing content of a first entity, and the content template includes song accompaniment; The step of, when the matching relationship between the live streaming content and the content template meets the preset trigger requirement, displaying the first custom content in the first live streaming room includes: Obtaining a singing score for the match between the singing content and the song accompaniment; When the singing score reaches a score threshold, displaying the first custom content in the first live streaming room.

7. The method according to claim 6, characterized in that, Obtaining the singing score for the matching between the singing content and the song accompaniment includes: Obtaining the beat point data of the song accompaniment, where the beat point data includes the timestamps of the song beat points of the song accompaniment in the audio of the song accompaniment; Analyzing the beat point data to determine the target beat points of the song accompaniment, and the timestamps of the target beat points are used to divide the singing score segments. The target beat points meet at least one of the following conditions: the difference in beat intensity between the target beat point and the adjacent beat points reaches a preset intensity threshold, the sound frequency at the target beat point reaches a maximum value, and the target beat point corresponds to the start timestamp of the chorus in the song accompaniment; Dividing the song accompaniment into at least one singing score segment based on the timestamps of the target beat points. The i-th singing score segment corresponds to the i-th time period in the song accompaniment, where i is a positive integer; For the i-th singing score segment, obtaining the i-th singing segment corresponding to the i-th singing score segment from the singing content; Calculating the similarity score between the i-th singing score segment and the i-th singing segment as the singing score for the matching between the singing content and the song accompaniment within the i-th time period.

8. The method according to claim 7, wherein The method further includes: In response to receiving a beat point setting operation, setting the specified beat point in the beat point data as the target beat point.

9. The method according to claim 7, wherein The displaying the first custom content in the first live room includes: When the first custom content is static content, generating a dynamic performance animation corresponding to the first custom content; Determining the first performance form of the dynamic performance animation based on the beat point data, where the first performance form matches the song beat points; Displaying the dynamic performance animation in the first live room in the first performance form.

10. The method according to any one of claims 1 to 4, characterized in that, The live content includes the action content of the first subject, and the content template includes an action video; When the matching relationship between the live content and the content template meets the preset trigger requirement, the displaying the first custom content in the first live room includes: Obtaining the action score for the matching between the action content and the action video; When the action score reaches the score threshold, displaying the first custom content in the first live room.

11. The method according to claim 10, wherein The obtaining the action score for the matching between the action content and the action video includes: Collecting action images of the action content based on a first frequency, and collecting video images of the action video based on the first frequency to obtain at least two pairs of action score images, where the m-th pair of action score images includes the m-th action image collected at the m-th moment and the m-th video image, and m is a positive integer; Calculating the similarity score between the m-th action image and the m-th video image to obtain the action score of the m-th pair of action score images; Obtaining the action score for the matching between the action content and the action video based on the sum of the action scores of the at least two pairs of action score images respectively.

12. The method according to any one of claims 1 to 4, characterized in that, After receiving the content configuration operation for the first live broadcast room, it further includes: Receiving a display mode configuration operation for the first custom content, where the display mode configuration operation is used to indicate the way of displaying the first custom content in combination with the live broadcast screen of the first live broadcast room; The displaying the first custom content in the first live broadcast room includes: Displaying the first custom content that matches the current live broadcast screen of the first live broadcast room in the first live broadcast room.

13. A computer device, characterized in that, The computer device includes a processor and a memory, and at least one program is stored in the memory, and the at least one program is loaded and executed by the processor to implement the content display method according to any one of claims 1 to 12.

14. A computer-readable storage medium, characterized in that, At least one program is stored in the storage medium, and the at least one program is loaded and executed by a processor to implement the content display method according to any one of claims 1 to 12.

15. A computer program product, characterized in that, It includes a computer program, and when the computer program is executed by a processor, it implements the content display method according to any one of claims 1 to 12.