Comment information display method and device, computer device, and storage medium

By displaying comment information on the video clip playback controls on the video playback interface, and by comparing and identifying the comment content to trigger the playback of the target video clip, the problem of limited comment information dimensions is solved, and the human-computer interaction rate is improved.

CN115734016BActive Publication Date: 2025-12-12TENCENT TECHNOLOGY (SHENZHEN) CO LTD
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
CN202111016208.9
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2021-08-31
Publication Date
2025-12-12
Estimated Expiration
2041-08-31

AI Technical Summary

Technical Problem

The limited information dimensions of comments on the video playback interface result in a low human-computer interaction rate.

Method used

The video playback interface displays comment information for the video clip playback control. By recognizing the comment content and comparing it with the tag information of the candidate video clips, the matching target video clip is obtained, and its playback is triggered through the control.

Benefits of technology

It expands the information dimensions of comment information, improves the interactive experience of interactive objects, and enhances the human-computer interaction rate.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN115734016B_ABST
    Figure CN115734016B_ABST
Patent Text Reader

Abstract

The application discloses a display method and device of comment information, a computer device and a storage medium, and belongs to the multimedia technical field. The method comprises the following steps: displaying a video playing interface; acquiring a video segment matching result of comment content in the video playing interface; in response to the fact that the video segment matching result comprises a target video segment matched with the comment content, displaying target comment information comprising a video segment playing control on the video playing interface, and the video segment playing control is used for triggering the playing of the target video segment. In this way, the displayed comment information can provide information in the video segment dimension, the dimension of the information provided by the comment information is expanded, the interactive experience of the interactive object is improved, and the human-computer interaction rate is improved.
Need to check novelty before this filing date? Find Prior Art

Description

TECHNICAL FIELD

[0001] Embodiments of the present application relate to the technical field of multimedia, and in particular, relate to a display method and device of comment information, a computer device, and a storage medium. BACKGROUND

[0002] With the development of multimedia technology, more and more application programs and web pages support playing videos, and can display comment information published by an interactive object in a video playing interface to bring the interactive object a real-time interactive experience.

[0003] In the related art, the comment information displayed in the video playing interface is information including at least one of text or picture, the dimension of the information provided by the comment information is relatively limited, and the interactive experience of the interactive object is poor, resulting in a low human-computer interaction rate. SUMMARY

[0004] Embodiments of the present application provide a display method and device of comment information, a computer device, and a storage medium, which can be used to expand the dimension of information provided by the comment information and improve the human-computer interaction rate. The technical solution is as follows:

[0005] In one aspect, a display method of comment information is provided, and the method includes:

[0006] displaying a video playing interface;

[0007] obtaining a video segment matching result of comment content in the video playing interface;

[0008] in response to the video segment matching result including a target video segment matched with the comment content, displaying target comment information including a video segment playing control on the video playing interface, the video segment playing control being used to trigger playing of the target video segment.

[0009] In one possible implementation, displaying target comment information including a video segment playing control on the video playing interface includes:

[0010] controlling the target comment information including the video segment playing control to move from a reference position of the video playing interface into the video playing interface, so that the target comment information including the video segment playing control is displayed on the video playing interface.

[0011] In another aspect, a display device of comment information is provided, and the device includes:

[0012] a first display module configured to display a video playing interface;

[0013] The first obtaining module is configured to obtain a video clip matching result of the comment content in the video playing interface.

[0014] The second display module is configured to, in response to the video clip matching result including a target video clip matched with the comment content, display target comment information including a video clip playing control on the video playing interface, where the video clip playing control is used to trigger playing of the target video clip.

[0015] In a possible implementation, the video playing interface plays a first video, and the first obtaining module is configured to identify text information corresponding to the comment content; compare the text information with label information of a candidate video clip, obtain a video clip matching result of the comment content based on a comparison result, and the candidate video clip is a video clip associated with the first video.

[0016] In a possible implementation, the video playing interface plays a first video, and the first obtaining module is configured to identify text information corresponding to the comment content; send the text information to a server, where the server is configured to compare the text information with label information of a candidate video clip, obtain a video clip matching result of the comment content based on a comparison result, and return the video clip matching result, and the candidate video clip is a video clip associated with the first video; and receive the video clip matching result returned by the server.

[0017] In a possible implementation, the first obtaining module is further configured to, in response to the comparison result indicating that the text information matches the label information of the first candidate video clip successfully, obtain a target video clip matched with the comment content based on the first candidate video clip, and take a result including the target video clip as the video clip matching result.

[0018] In a possible implementation, the second display module is further configured to, in response to the video clip matching result not including a target video clip matched with the comment content, display prompt information on the video playing interface, where the prompt information is used to prompt that the comment content has not been successfully matched to a video clip.

[0019] In a possible implementation, the video playing interface displays a comment information publishing control, and the second display module is further configured to, in response to a triggering operation of the comment information publishing control and in response to the video clip matching result not including a target video clip matched with the comment content, display the prompt information on the video playing interface.

[0020] In a possible implementation, the apparatus further includes:

[0021] The playing module is configured to play the target video clip on the video playing interface in response to a triggering operation of the video clip playing control.

[0022] In a possible implementation, the video playing interface plays a first video, and the playing module is configured to play the target video clip on the video playing interface and pause playing of the first video.

[0023] In a possible implementation, the playing module is further configured to, in response to a resuming playing operation of the first video, cancel playing of the target video clip and resume playing of the first video.

[0024] In a possible implementation, the apparatus further includes:

[0025] The second obtaining module is configured to obtain a resuming playing operation of the first video in response to a triggering operation of a target region on the video playing interface, the target region being any region on the video playing interface except a region playing the target video clip.

[0026] In a possible implementation, the region playing the target video clip supports at least one of size adjustment and position transformation.

[0027] In a possible implementation, the video playing interface displays a content editing box, and the comment content is content displayed in the content editing box.

[0028] In a possible implementation, the video playing interface displays a video clip matching control and a comment information publishing control, and the first obtaining module is configured to obtain a video clip matching result of the comment content in the video playing interface in response to a triggering operation of the video clip matching control.

[0029] The second display module is configured to display target comment information including a video clip playing control on the video playing interface in response to a triggering operation of the comment information publishing control and in response to the video clip matching result including a target video clip matching the comment content.

[0030] In a possible implementation, the apparatus further includes:

[0031] The state transformation module is configured to transform a state of the video clip matching control from an unactivated state to an activated state in response to a triggering operation of the video clip matching control.

[0032] In a possible implementation, the state transformation module is further configured to transform the state of the video clip matching control from the activated state to the unactivated state in response to a triggering operation of the comment information publishing control.

[0033] In a possible implementation, the state transforming module is further configured to, in response to the video segment matching result not including the target video segment matched with the comment content, restore the state of the video segment matching control from the enabled state to the disabled state.

[0034] In a possible implementation, the second display module is configured to control the target comment information including the video segment play control to move from the reference position of the video playing interface into the video playing interface, so that the target comment information including the video segment play control is displayed on the video playing interface.

[0035] In another aspect, a computer device is provided, which includes a processor and a memory, and the memory stores at least one computer program, which is loaded and executed by the processor, so that the computer device implements the comment information display method described above.

[0036] In another aspect, a computer readable storage medium is also provided, which stores at least one computer program, which is loaded and executed by a processor, so that a computer implements the comment information display method described above.

[0037] In another aspect, a computer program product is also provided, which includes a computer program or computer instructions, which is loaded and executed by a processor, so that a computer executes the comment information display method described above.

[0038] The technical solutions provided by the embodiments of the present application at least bring the following beneficial effects:

[0039] The technical solutions provided by the embodiments of the present application can display a comment information including a video segment play control on a video playing interface, and the video segment play control can trigger the playing of a target video segment matched with the comment content. That is, the comment information displayed by the embodiments of the present application can provide information in the dimension of video segments, expand the dimension of information provided by the comment information, and the interactive object can watch the target video segment by triggering the video segment play control in the comment information, which is beneficial to improving the interactive experience of the interactive object and further improving the human-computer interaction rate. BRIEF DESCRIPTION OF DRAWINGS

[0040] In order to more clearly illustrate the technical solutions in the embodiments of the present application, the drawings needed to be used in the embodiments description will be briefly introduced. Obviously, the drawings in the following description are only some embodiments of the present application, and other drawings can be obtained by those skilled in the art without any creative effort based on these drawings.

[0041] Figure 1 is a schematic diagram of an implementation environment of a comment information display method provided by an embodiment of the present application;

[0042] Figure 2 is a flowchart of a comment information display method provided by an embodiment of the present application;

[0043] Figure 3 is a schematic diagram of a video playing interface provided by an embodiment of the present application;

[0044] Figure 4 is a schematic diagram of a video playing interface provided by an embodiment of the present application;

[0045] Figure 5 is a schematic diagram of a video playing interface provided by an embodiment of the present application;

[0046] Figure 6 is a schematic diagram of a video playing interface provided by an embodiment of the present application;

[0047] Figure 7 is a schematic diagram of a video playing interface provided by an embodiment of the present application;

[0048] Figure 8 is a schematic diagram of a video playing interface provided by an embodiment of the present application;

[0049] Figure 9 is a schematic diagram of a comment information display process provided by an embodiment of the present application;

[0050] Figure 10 is a schematic diagram of a comment information display device provided by an embodiment of the present application;

[0051] Figure 11 is a schematic diagram of a comment information display device provided by an embodiment of the present application;

[0052] Figure 12 is a schematic diagram of a terminal structure provided by an embodiment of the present application. DETAILED DESCRIPTION

[0053] In order to make the purpose, technical solutions and advantages of the present application more clear, the embodiments of the present application will be further described in detail with reference to the drawings.

[0054] It should be noted that the terms "first", "second", and the like in the present application are used to distinguish similar objects, and do not necessarily indicate a specific order or a chronological sequence. It should be understood that the data thus used can be interchanged under appropriate circumstances, so that the embodiments of the present application described herein can be implemented in an order other than that illustrated or described herein. The embodiments described in the following exemplary embodiments do not represent all embodiments consistent with the present application. Rather, they are merely examples of apparatuses and methods consistent with some aspects of the present application as detailed in the appended claims.

[0055] Figure 1 is a schematic diagram of an implementation environment of a comment information display method provided by an embodiment of the present application, which includes a terminal 11 and a server 12.

[0056] The terminal 11 is installed and runs an application or a webpage supporting video playing and display of comment information, and when the application or the webpage needs to display comment information on a video playing interface, the method provided by the present application can be applied to display. Exemplarily, the application or the webpage is a video application or webpage, a live application or webpage, a browser application or webpage, etc. Exemplarily, the terminal 11 is a device used by any interactive object, and the account of the any interactive object is logged in.

[0057] The server 12 is a background server of the application or the webpage installed by the terminal supporting video playing and display of comment information, and is used to provide background service for the application or the webpage.

[0058] Optionally, the terminal 11 can be any electronic product that can interact with a user through one or more of a keyboard, a touchpad, a touch screen, a remote controller, voice interaction, a handwriting device, etc., such as a PC (Personal Computer), a mobile phone, a smart phone, a PDA (Personal Digital Assistant), a wearable device, a PPC (Pocket PC), a tablet computer, a smart car machine, a smart television, a smart speaker, or a vehicle terminal, etc. The server 12 can be a server, a server cluster composed of multiple servers, or a cloud computing service center. The terminal 11 and the server 12 establish a communication connection through a wired or wireless network.

[0059] Those skilled in the art should understand that the above-mentioned terminal 11 and server 12 are only examples, and other existing or future terminal or server, such as those applicable to the present application, should also be included in the protection scope of the present application, and are hereby included by reference.

[0060] Based on the above Figure 1In the illustrated implementation environment, the embodiment of the present application provides a display method of comment information. The display method is taken as an example to be executed by the terminal 11. As shown in Figure 2 As shown, the display method of comment information provided by the embodiment of the present application includes the following steps 201 to 203.

[0061] In step 201, a video playing interface is displayed.

[0062] The video playing interface is an interface for playing a video currently watched by an interactive object. The video currently watched by the interactive object is referred to as a first video in the embodiment of the present application. That is, the first video is played on the video playing interface. The playing manner of the first video on the video playing interface is not limited in the embodiment of the present application. For example, the first video is played in a full screen mode or a window mode, which is related to the playing setting of an application or a webpage providing the video playing interface.

[0063] In some embodiments, the video playing interface is an interface in an application supporting video playing and display of comment information, or the video playing interface is an interface in a webpage supporting video playing and display of comment information. The embodiment of the present application is not limited in this regard. For example, the first video can be a video that has been recorded and completed, or can be a live video. The embodiment of the present application is not limited in this regard.

[0064] In a possible implementation manner, the terminal displays the video playing interface in the following manner: the terminal acquires the first video in response to a playing instruction of the first video, and displays the video playing interface and plays the first video on the video playing interface. For example, the terminal can further acquire historical comment information corresponding to the first video, and then display the historical comment information corresponding to the current playing progress on the video playing interface based on the playing progress of the first video. The historical comment information is comment information published by at least one interactive object when watching the first video. Of course, in some embodiments, the terminal can not display the historical comment information corresponding to the first video when playing the first video.

[0065] In an example embodiment, the first video is a video stored on a server, and the terminal acquires the first video and the historical comment information corresponding to the first video by interacting with the server. In an example embodiment, the first video is a video stored locally on the terminal, and the terminal extracts the first video from the local storage and acquires the historical comment information corresponding to the first video by interacting with the server.

[0066] In a possible implementation manner, at least one of a content editing box, a video segment matching control and a comment information publishing control is displayed on the video playing interface. The content editing box is used to display comment content. In an example embodiment, the content editing box is provided with a voice input interface, and an interactive object can input comment voice through triggering the voice input interface provided by the content editing box, and then the terminal displays comment content obtained by performing voice recognition on the comment voice in the content editing box. In an example embodiment, an interactive object can directly edit comment content in the form of text, pictures and the like in the content editing box. The video segment matching control is used to trigger obtaining of a video segment matching result, and the comment information publishing control is used to trigger publishing of comment information.

[0067] The display positions and forms of the content editing box, the video segment matching control and the comment information publishing control can be set according to experience, or can be flexibly adjusted according to application scenarios, and embodiments of the present application do not limit this. For example, the content editing box, the video segment matching control and the comment information publishing control are all displayed in a lower right corner region of the video playing interface, the form of the content editing box is an editable input box, the form of the video segment matching control is a triggerable icon or a triggerable button, and the form of the comment information publishing control is a triggerable icon or a triggerable button.

[0068] For example, first information can be displayed on the video segment matching control, to prompt an interactive object that the control is used to trigger obtaining of a video segment matching result. Second information can be displayed on the comment information publishing control, to prompt that the control is used to trigger publishing of comment information. The first information and the second information are set according to experience, or are flexibly adjusted according to application scenarios, and embodiments of the present application do not limit this.

[0069] For example, a video playing interface is as shown in FIG. 3. Figure 3 For example, a video playing interface is as shown in FIG. 3. Figure 3 A content editing box 301, a video segment matching control 302 and a comment information publishing control 303 are displayed in the video playing interface as shown in FIG. 3. First information "AI" is displayed on the video segment matching control 302, and second information "publish" is displayed on the comment information publishing control 303. For example, Figure 3 For example, a video playing interface is as shown in FIG. 3.

[0070] The plurality of historical comment information is displayed on the video picture of the first video. For example, the plurality of historical comment information refers to comment information matched with the time point corresponding to the current video picture of the first video. The plurality of historical comment information can be comment information published by the interactive object currently watching the first video, or comment information published by other interactive objects during watching the first video, and the embodiments of the present application do not limit the same. The plurality of historical comment information can move on the video playing interface along with the interactive object watching the first video. For example, the plurality of historical comment information can display the avatar of the interactive object publishing the historical comment information, and the interactive object can know the interactive object publishing the comment information through the avatar on the historical comment information. The interactive object can interact with the publisher of the historical comment information by triggering the historical comment information.

[0071] For example, in the video playing interface shown in Figure 3 The video playing interface shown in the figure also displays the playing progress bar of the first video, the playing control control of the first video, the playing time information, the playing resolution, the playing speed, the historical playing statistical information, the sharing control, the comment information switch control, and the like. The playing control control of the first video includes, but is not limited to, the pause / play control, the next episode control, the loop playing control, the volume control control, the setting control, the playing picture proportion selection control, the full screen playing control, and the like. The comment information switch control is used to select to open or close the comment information display function by the interactive object, such as Figure 3 As shown in the figure, the comment information switch control displays the word "pop" on it.

[0072] For example, the comment information can also be called the bullet screen, which is a kind of comment subtitle and can be superimposed on the video picture. When watching the video, the interactive object can express his / her watching feeling by publishing the bullet screen, and can also see the bullet screen published by other interactive objects at the current video playing time. The interactive objects can communicate and interact with each other through the bullet screen.

[0073] In step 202, the video segment matching result of the comment content in the video playing interface is obtained.

[0074] The comment content in the video playing interface is comment content provided by an interactive object (referred to as a target interactive object) currently watching the first video. Illustratively, the comment content in the video playing interface refers to comment content obtained by performing voice recognition on comment voice input by the target interactive object. Illustratively, the comment content in the video playing interface refers to content displayed in a content editing box on the video playing interface. In an example embodiment, the content displayed in the content editing box can refer to content in the form of text, pictures, etc. edited directly by the target interactive object in the content editing box, or can refer to content obtained by performing voice recognition on target comment voice input by the target interactive object through a voice input interface provided by the content editing box.

[0075] If the target interactive object wants to comment on the video content of the first video at a certain time point during the process of watching the first video, the target interactive object can input comment voice or edit comment content in the content editing box. Illustratively, the content that the target interactive object can edit in the content editing box includes but is not limited to text, pictures, etc. The text refers to characters input or uploaded by the target interactive object, and the pictures can refer to pictures selected by the target interactive object from alternative pictures (such as expression pictures, cartoon pictures) provided by an application or a webpage playing the first video, or can refer to pictures selected by the target interactive object from the local terminal and uploaded, etc. For example, as shown in FIG. 1, the target interactive object directly leaks the plot at a tense plot stage, a key part of emotional entanglement between the male and female leads, and a segment of a real hammer, so that the comment content edited by the target interactive object in the content editing box is the text “16th set 5 minutes and 10 seconds, real hammer”. Figure 3

[0076] After obtaining the comment content in the video playing interface, a video segment matching result of the comment content in the video playing interface is further obtained. In one possible implementation, after obtaining the comment content in the video playing interface, an operation of obtaining the video segment matching result of the comment content in the video playing interface is directly performed. In another possible implementation, a video segment matching control is displayed on the video playing interface, and after obtaining the comment content in the video playing interface, an operation of obtaining the video segment matching result of the comment content in the video playing interface is further performed in response to a triggering operation of the video segment matching control.

[0077] The video segment matching control is used to trigger the obtaining of the video segment matching result. The triggering operation of the video segment matching control refers to a human-computer interaction operation triggered on the video segment matching control. The type of the triggering operation of the video segment matching control is not limited in the embodiments of the present application. Illustratively, the triggering operation of the video segment matching control includes but is not limited to at least one of a single-click operation, a double-click operation, a hovering touch operation, a pressure touch operation, and a sliding operation. ​

[0078] In a possible implementation, after obtaining the triggering operation of the video segment matching control, the state of the video segment matching control can be changed from the unactivated state to the activated state, so as to inform the target interactive object that the video segment matching control has been successfully triggered by using the activated state. That is, the state of the video segment matching control is changed from the unactivated state to the activated state in response to the triggering operation of the video segment matching control. The unactivated state and the activated state are two different control display states, and embodiments of the present application do not limit the unactivated state and the activated state. Exemplarily, the video segment matching control in the unactivated state is as shown in 302 in FIG. 3A, and the video segment matching control in the activated state is as shown in 401 in FIG. 4A. Figure 3 Figure 4

[0079] The video segment matching result of the comment content is used to indicate whether the comment content is successfully matched to a video segment. Embodiments of the present application refer to a video segment to which the comment content is successfully matched as a target video segment. The target video segment refers to a video segment matched with the comment content, and can present the meaning of the comment content in the form of a video. The time length of the target video segment is set according to experience or is flexibly adjusted according to an application scenario, and embodiments of the present application do not limit this. Exemplarily, the time length of the target video segment is 60 seconds.

[0080] In exemplary embodiments, the video segment matching result of the comment content has two possibilities, one of which includes the target video segment, and the other of which does not include the target video segment. If the video segment matching result of the comment content includes the target video segment, it means that the comment content is successfully matched to a video segment. If the video segment matching result of the comment content does not include the target video segment, it means that the comment content is not successfully matched to a video segment.

[0081] In exemplary embodiments, the video segment matching result that does not include the target video segment includes prompt information, which is used to prompt that the comment content is not successfully matched to a video segment. Embodiments of the present application do not limit the form of the prompt information. Exemplarily, the prompt information can include only information indicating the matching failure, or can include both information indicating the matching failure and information indicating the reason for the matching failure. Exemplarily, the reason for the matching failure includes but is not limited to an incorrect comment content description, a non-existent video segment due to deletion indicated by the comment content, network failure, and the like.

[0082] ​​In a possible implementation, the manner in which the terminal obtains the video segment matching result of the comment content in the video playing interface can refer to that the terminal locally obtains the video segment matching result of the comment content in the video playing interface, or that the terminal obtains the video segment matching result of the comment content in the video playing interface by interacting with the server, which is not limited in the embodiments of the application.

[0083] In a possible implementation, the process in which the terminal locally obtains the video segment matching result of the comment content in the video playing interface includes the following steps A and step B.

[0084] Step A: Identify the text information corresponding to the comment content.

[0085] After obtaining the comment content in the video playing interface, the terminal identifies the text information corresponding to the comment content, which can represent the meaning of the comment content to a certain extent.

[0086] For example, for a case in which the comment content only includes text content, the process of identifying the text information corresponding to the comment content is a semantic recognition process. Semantic recognition is a natural language processing technology that enables a computer to understand and analyze text content and automatically process the text content. Through semantic recognition, the meaning of the text content can be understood. For example, for a case in which the comment content only includes a picture, the process of identifying the text information corresponding to the comment content is a picture recognition process. Picture recognition can identify relevant information of the picture, for example, if the picture is a face image of an actor, the picture recognition can identify the face in the picture as the face of the actor. For example, for a case in which the comment content includes text content and a picture, the process of identifying the text information corresponding to the comment content is a comprehensive process of semantic recognition and picture recognition. The manner of semantic recognition and picture recognition is not limited in the embodiments of the application, for example, the semantic recognition and picture recognition can be implemented by using a neural network model.

[0087] Step B: Compare the text information with label information of a candidate video segment, and obtain the video segment matching result of the comment content based on the comparison result. The candidate video segment is a video segment associated with the first video.

[0088] After the text information corresponding to the comment content is identified, the text information is compared with label information of a candidate video segment, and then a video segment matching result of the comment content is obtained based on a comparison result. The candidate video segment is a video segment associated with the first video. For example, the candidate video segment is a segment cut from at least one of the first video and an extended video corresponding to the first video. The extended video corresponding to the first video is a video associated with the first video and used to extend the video content of the first video. For example, the extended video corresponding to the first video includes, but is not limited to, a blooper video corresponding to the first video, a paid video corresponding to the first video, and the like.

[0089] The application does not limit the way of cutting the candidate video segment. For example, one frame of video is cut as a candidate video segment, or a video segment with a reference time length is cut as a candidate video segment. The reference time length is set according to experience or flexibly adjusted according to an application scenario. For example, the reference time length is 10 seconds. For example, which video segments are cut from the first video and the extended video corresponding to the first video as candidate video segments is set according to experience or flexibly adjusted according to actual video content. The application does not limit this.

[0090] The label information of the candidate video segment is used to describe the content of the candidate video segment. The label information of the candidate video segment is composed of one or more labels. Different labels are used to describe different aspects of the content of the candidate video segment. For example, the label information of the candidate video segment is used to describe one or more aspects of the content of the plot details attribute, the episode title, the time point, the person, the thing, the object, the scene, the event, and the development stage of the candidate video segment. For example, the number of candidate video segments is multiple. Each candidate video segment has its own label information. The label information of different candidate video segments can be the same or different. The application does not limit this.

[0091] For example, before the terminal compares the text information with the label information of the candidate video segment, the terminal needs to obtain the label information of the candidate video segment. For example, the candidate video segment is stored locally in the terminal, and then the terminal obtains the label information of the candidate video segment locally.

[0092] Exemplarily, the server pre-interpretively transcodes and extracts the tag information of the candidate video segment associated with the first video and stores, such as stores in a cloud storage space. That is, the candidate video segment is stored in the server, and the terminal obtains the tag information of the candidate video segment by interacting with the server. Exemplarily, the terminal sends a request to the server to obtain the tag information of the candidate video segment corresponding to the first video, and the request carries the video identifier of the first video. After receiving the request sent by the terminal, the server extracts the tag information of the candidate video segment according to the video identifier of the first video, and then returns the tag information of the candidate video to the terminal.

[0093] After obtaining the text information and the tag information of the candidate video segment, the text information is compared with the tag information of the candidate video segment. In a possible implementation, the way of comparing the text information with the tag information of the candidate video segment is: obtaining the similarity between the text information and the tag information of the candidate video segment, and judging whether the text information matches the tag information of the candidate video segment successfully according to the similarity. The way of obtaining the similarity between the text information and the tag information of the candidate video segment is not limited in the embodiments of the present application. Exemplarily, the ratio of the number of tags in the text information that hit the tag information of a certain candidate video segment to the total number of tags in the tag information of the candidate video segment is taken as the similarity between the text information and the tag information of the candidate video segment.

[0094] Exemplarily, the feature extraction model is called to perform feature extraction on the text information and the tag information of the candidate video segment respectively, to obtain the feature of the text information and the feature of the tag information of the candidate video segment; and the similarity between the feature of the text information and the feature of the tag information of the candidate video segment is taken as the similarity between the text information and the tag information of the candidate video segment. Exemplarily, the feature of the text information and the feature of the tag information of the candidate video segment are both vectors, and the similarity between the feature of the text information and the feature of the tag information of the candidate video segment is obtained by calculating the similarity between the two vectors. For example, the similarity between the two vectors refers to the cosine similarity, the Jaccard similarity, etc.

[0095] After comparing the text information with the label information of the candidate video segments, a comparison result is obtained. There are two possibilities for the comparison result: the comparison result indicates that the text information matches the label information of the first candidate video segment successfully; the comparison result indicates that the text information does not match the label information of each candidate video segment. The manner in which the embodiments of the present application determine whether the text information matches the label information of a certain candidate video segment is not limited. For example, if the similarity between the text information and the label information of a certain candidate video segment is not less than a similarity threshold, it is determined that the text information matches the label information of the candidate video segment; if the similarity between the text information and the label information of a certain candidate video segment is less than the similarity threshold, it is determined that the text information does not match the label information of the candidate video segment. The similarity threshold is set according to experience or is flexibly adjusted according to an actual application scenario, and the embodiments of the present application are not limited in this regard.

[0096] If the text information matches the label information of one or more candidate video segments, the one or more candidate video segments are referred to as the first candidate video segment, and at this time, the comparison result indicating that the text information matches the label information of the first candidate video segment is obtained. If the text information does not match the label information of each candidate video segment, the comparison result indicating that the text information does not match the label information of each candidate video segment is obtained.

[0097] The embodiments of the present application do not limit the form of the comparison result. For example, the comparison result indicating that the text information matches the label information of the first candidate video segment is a result including the identifier of the first candidate video segment; the comparison result indicating that the text information does not match the label information of each candidate video segment is a result including empty information or a result including a prompt information that the matching fails.

[0098] After obtaining the comparison result, the video segment matching result of the comment content is obtained based on the comparison result. In a possible implementation manner, the manner in which the video segment matching result of the comment content is obtained based on the comparison result is as follows: in response to the comparison result indicating that the text information matches the label information of the first candidate video segment successfully, the target video segment matching the comment content is obtained based on the first candidate video segment, and a result including the target video segment is taken as the video segment matching result.

[0099] In a possible implementation manner, the manner in which the video segment matching result of the comment content is obtained based on the comparison result is as follows: in response to the comparison result indicating that the text information does not match the label information of each candidate video segment, a result not including any video segment is taken as the video segment matching result. For example, the result not including any video segment may or may not include a prompt information, and the embodiments of the present application are not limited in this regard.

[0100] In a possible implementation, the number of the first candidate video clips can be one or more, and the embodiments of the present application do not limit this. In a possible implementation, for the case where the number of the first candidate video clips is one, the manner of obtaining the target video clip matching the comment content based on the first candidate video clip is that the first candidate video clip is taken as the target video clip matching the comment content.

[0101] In another possible implementation, for the case where the number of the first candidate video clips is one, the manner of obtaining the target video clip matching the comment content based on the first candidate video clip is that the position of the first candidate video clip in the source video is determined, and a video clip of a first time length before and after the center position of the position of the first candidate video clip in the source video is clipped to form the target video clip. The first time length is half of a specified time length, and the specified time length is a specified time length that the target video clip should have. The specified time length is set according to experience or is flexibly adjusted according to an application scenario, for example, the specified time length is 60 seconds. Of course, the process of obtaining the target video clip based on the first candidate video clip can also be implemented in other manners, and the embodiments of the present application do not repeat the description.

[0102] In a possible implementation, for the case where the number of the first candidate video clips is multiple, the manner of obtaining the target video clip matching the comment content based on the first candidate video clip is that one candidate video clip is selected from the multiple first candidate video clips as a target candidate video clip, and the target video clip matching the comment content is obtained based on the target candidate video clip. The embodiments of the present application do not limit the manner of selecting one candidate video clip from the multiple first candidate video clips as the target candidate video clip, and for example, one candidate video clip is randomly selected from the multiple first candidate video clips as the target candidate video clip. For example, one candidate video clip with the highest similarity between the label information and the text information in the multiple first candidate video clips is taken as the target candidate video clip. After the target candidate video clip is determined, the target video clip matching the comment content is obtained based on the target candidate video clip. The manner of obtaining the target video clip matching the comment content based on the target candidate video clip can refer to the manner of obtaining the target video clip matching the comment content based on the first candidate video clip for the case where the number of the first candidate video clip is one, and details are not repeated here.

[0103] The above describes a process of the terminal locally obtaining the video segment matching result of the comment content. Exemplarily, the terminal can also obtain the video segment matching result of the comment content by interacting with the server. In a possible implementation manner, the manner in which the terminal obtains the video segment matching result of the comment content by interacting with the server includes but is not limited to the following two manners.

[0104] Manner one: The terminal identifies the text information corresponding to the comment content; the terminal sends the text information to the server, the server is configured to compare the text information with the label information of the candidate video segment, obtain the video segment matching result of the comment content based on the comparison result, and return the video segment matching result; and the terminal receives the video segment matching result returned by the server.

[0105] In this manner, the terminal needs to perform the operation of identifying the text information corresponding to the comment content, and the operation of obtaining the video segment matching result based on the text information is performed by the server. The manner in which the server performs the operation of obtaining the video segment matching result based on the text information is described with reference to the related process in the case of the terminal locally obtaining the video segment matching result, which will not be described herein again.

[0106] Manner two: The terminal sends the comment content to the server, the server is configured to identify the text information corresponding to the comment content, compare the text information with the label information of the candidate video segment, obtain the video segment matching result of the comment content based on the comparison result, and return the video segment matching result; and the terminal receives the video segment matching result returned by the server.

[0107] In this manner two, the entire process of obtaining the video segment matching result based on the comment content is performed by the server. The implementation manner is described with reference to the process of the terminal locally obtaining the video segment matching result, which will not be described herein again.

[0108] As can be known from the content introduced in this step 202, after obtaining the comment content in the video playing interface, the terminal obtains the video segment matching result of the comment content in the video playing interface. The video segment matching result can include the target video segment matched with the comment content, or can not include the target video segment matched with the comment content.

[0109] In step 203, in response to the video segment matching result including the target video segment matched with the comment content, the target comment information including the video segment playing control is displayed on the video playing interface, and the video segment playing control is configured to trigger the playing of the target video segment.

[0110] After the video clip matching result is obtained, it is determined whether the video clip matching result includes a target video clip matching the comment content. If the video clip matching result includes the target video clip matching the comment content, the target comment information including the video clip play control is displayed on the video playing interface. That is, in response to the video clip matching result including the target video clip matching the comment content, the target comment information including the video clip play control is displayed on the video playing interface.

[0111] In a possible implementation, after it is determined that the video clip matching result includes the target video clip matching the comment content, in response to the video clip matching result including the target video clip matching the comment content, the operation of displaying the target comment information including the video clip play control on the video playing interface is directly performed. In another possible implementation, the comment information publishing control is displayed on the video playing interface. In response to the triggering operation of the comment information publishing control, and in response to the video clip matching result including the target video clip matching the comment content, the operation of displaying the target comment information including the video clip play control on the video playing interface is further performed.

[0112] The comment information publishing control is used to trigger the publishing of the comment information. When the triggering operation of the comment information publishing control is obtained, it indicates that the target interactive object wants to publish the comment information. The triggering operation of the comment information publishing control refers to the human-computer interaction operation triggered on the comment information publishing control. The type of the triggering operation of the comment information publishing control is not limited in the embodiments of the present application. For example, the triggering operation of the comment information publishing control includes at least one of a single-click operation, a double-click operation, a hovering touch operation, a pressure touch operation, and a sliding operation, but is not limited to the above.

[0113] The video clip play control is used to trigger the playing of the target video clip. That is, by triggering the video clip play control included in the target comment information, the playing of the target video clip can be realized. The form of the video clip play control is not limited in the embodiments of the present application, and can be set according to experience or flexibly adjusted according to application scenarios.

[0114] In the exemplary embodiments, the target comment information includes the comment content in addition to the video clip play control. The comment content includes at least one of text and a picture. Based on this mode, the target comment information is a comment information in the form of comment content+video clip play control, which can provide multi-dimensional information such as text, picture, and video clip, so that the target interactive object and other interactive objects watching the first video can see the comment content and watch the target video clip matching the comment content by triggering the video clip play control in the target comment information, thereby improving the interactive experience of the interactive objects. For example, the target comment information displayed on the video playing interface is as shown in FIG. 2.Figure 5 as shown in 501 in FIG. 5. Figure 5 The target comment information as shown in 501 in FIG. 5 includes comment content 5011 and a video segment playback control 5012.

[0115] In an example embodiment, the target comment information can further include associated information of the comment content, which is information related to the comment content. For example, the associated information includes at least one of an account of an interactive object editing the comment content, an avatar of the interactive object editing the comment content, and a nickname of the interactive object editing the comment content.

[0116] In a possible implementation, the implementation of displaying the target comment information including the video segment playback control on the video playback interface is that: the target comment information including the video segment playback control is controlled to move from a reference position of the video playback interface into the video playback interface, so that the target comment information including the video segment playback control is displayed on the video playback interface. The reference position is set according to experience or is flexibly adjusted according to application scenarios, and the example embodiments of the present application do not limit this. For example, the reference position refers to the right side of the video playback interface; or the reference position refers to the left side of the video playback interface; or the reference position refers to the upper side of the video playback interface, and the like.

[0117] For example, the way of controlling the target comment information to move from the reference position of the video playback interface into the video playback interface is that: the target comment information is controlled to move from the reference position of the video playback interface into the video playback interface at a first reference speed. The first reference speed is set according to experience or is flexibly adjusted according to application scenarios, and the example embodiments of the present application do not limit this. For example, the first reference speed is a relatively slow speed, so as to ensure that the target comment information can move into the video playback interface relatively gently.

[0118] The above implementation of displaying the target comment information on the video playback interface is only an example, and the example embodiments of the present application are not limited thereto. For example, the target comment information is directly displayed on the video playback interface; or the target comment information is displayed on the video playback interface through a display special effect.

[0119] In a possible implementation, after the target comment information including the video segment playback control is displayed on the video playback interface, the implementation further includes: controlling the target comment information to move on the video playback interface. The example embodiments of the present application do not limit the way of controlling the target comment information to move on the video playback interface. For example, the target comment information is controlled to move on the video playback interface from right to left at a second reference speed. The second reference speed is set according to experience or is flexibly adjusted according to application scenarios, and the example embodiments of the present application do not limit this.

[0120] In the process of controlling the movement of the target comment information on the video playing interface, in response to a selection operation of the target comment information, the movement of the target comment information is paused, and the target comment information is displayed by using a first display mode, which is different from the display mode of the unselected comment information. Exemplarily, the selection operation of the target comment information is an operation of placing a mouse in a display area of the target comment information. Exemplarily, the selection operation of the target comment information is an operation of touching the display area of the target comment information by using a finger.

[0121] The first display mode is not limited in the embodiments of the present application. Exemplarily, for the case that the display background color of the unselected comment information is the first color, the first display mode is a display mode of adjusting the display background color to the second color. For example, the first color is white, and the second color is gray. For example, the target comment information displayed by using the gray background color is as shown in 601 in FIG. 6. Figure 6

[0122] In the exemplary embodiments, for the case that the display mode of the unselected comment information is to not display the interactive control by default, the first display mode can display the interactive control in the surrounding area of the comment information. The interactive control includes, but is not limited to, the like and the disagree control. The surrounding area can refer to the lower area, the right area, etc. In the exemplary embodiments, the first display mode can also be to adjust the background color and display the interactive control.

[0123] In the exemplary embodiments, in response to a cancel selection operation of the target comment information, the movement of the target comment information is continued until the target comment information moves out of the video playing interface.

[0124] In a possible implementation, after the target comment information including the video segment playing control is displayed on the video playing interface, the method further includes: in response to a trigger operation of the video segment playing control, playing the target video segment on the video playing interface. Since the video segment playing control is used to trigger the playing of the target video segment, after the trigger operation of the video segment playing control is acquired, the target video segment is played on the video playing interface. Exemplarily, the process of playing the target video segment is the process of displaying the video picture of the target video segment in the area where the target video segment is played.

[0125] The area on the video playing interface where the target video segment is played is not limited in the embodiments of the present application. Exemplarily, the target video segment is played in the upper right corner area on the video playing interface; or the target video segment is played in the nearby area of the target comment information displayed on the video playing interface. For example, as shown in Figure 6

[0126] Exemplarily, as shown in FIG. 6, the target video segment is played in the nearby area 602 of the target comment information 601.​​Figure 6 As shown in the video playing interface 600, in the region 602 for playing the target video segment, in addition to the video picture of the target video segment, there are also displayed a close control 6021, a full screen playing control 6022 and a progress bar 6023. The close control 6021 is used to cancel playing the target video segment, the full screen playing control 6022 is used to play the target video segment in full screen on the video playing interface, and the progress bar 6023 is used to present the progress of the target video segment that has been played.

[0127] In a possible implementation, the region for playing the target video segment supports at least one of size adjustment and position transformation. For example, the size adjustment of the region for playing the target video segment is achieved by pulling one of the four corners of the region, and the size adjustment includes zooming in and zooming out. For example, the position transformation of the region for playing the target video segment is achieved by dragging the region.

[0128] In a possible implementation, during the process of playing the target video segment on the video playing interface, the first video can be played simultaneously or paused, and the present application does not limit this. For the case of playing the first video and the target video segment simultaneously, the audio of the first video and the audio of the target video segment can be mixed and played; only the audio of the first video can be played, and the audio of the target video segment is not played; only the audio of the target video segment can be played, and the audio of the first video is not played.

[0129] The present application takes the case of pausing the first video while playing the target video segment on the video playing interface as an example. For example, the video playing interface when playing the target video segment is as shown in FIG. 6B. Figure 6 As shown in the video playing interface 600, in the region 602 for playing the target video segment, in addition to the video picture of the target video segment, there are also displayed a close control 6021, a full screen playing control 6022 and a progress bar 6023. The close control 6021 is used to cancel playing the target video segment, the full screen playing control 6022 is used to play the target video segment in full screen on the video playing interface, and the progress bar 6023 is used to present the progress of the target video segment that has been played. Figure 6 As shown in the video playing interface 600, in the region 602 for playing the target video segment, in addition to the video picture of the target video segment, there are also displayed a close control 6021, a full screen playing control 6022 and a progress bar 6023. The close control 6021 is used to cancel playing the target video segment, the full screen playing control 6022 is used to play the target video segment in full screen on the video playing interface, and the progress bar 6023 is used to present the progress of the target video segment that has been played.

[0130] In an example embodiment, for the case of playing the target video segment on the video playing interface and pausing the first video, it further includes: in response to a resuming playing operation of the first video, canceling playing the target video segment and resuming playing the first video. If the resuming playing operation of the first video is acquired, it indicates that the target interactive object does not need to watch the target video segment any more and needs to continue watching the first video. At this time, the playing of the target video segment is canceled and the playing of the first video is resumed. For example, the video playing interface after canceling the playing of the target video segment and resuming the playing of the first video is as shown in FIG. 6A. Figure 7 As shown in the video playing interface 600, in the region 602 for playing the target video segment, in addition to the video picture of the target video segment, there are also displayed a close control 6021, a full screen playing control 6022 and a progress bar 6023. The close control 6021 is used to cancel playing the target video segment, the full screen playing control 6022 is used to play the target video segment in full screen on the video playing interface, and the progress bar 6023 is used to present the progress of the target video segment that has been played.

[0131] In a possible implementation, the manner of obtaining the resuming play operation of the first video is: in response to a triggering operation of a target region in the video play interface, the resuming play operation of the first video is obtained, the target region being any region in the video play interface except the region playing the target video segment. That is, if a triggering operation of any region in the video play interface except the region playing the target video segment is detected, the resuming play operation of the first video is obtained.

[0132] In another possible implementation, the manner of obtaining the resuming play operation of the first video is: in response to a triggering operation of a close control in the region playing the target video segment, the resuming play operation of the first video is obtained.

[0133] In a possible implementation, the video segment matching result can also not include the target video segment matched with the comment content. In this case, in response to the video segment matching result not including the target video segment matched with the comment content, prompt information is displayed on the video play interface, the prompt information being used to prompt that the comment content is not successfully matched to a video segment. Exemplarily, the prompt information can be included in the video segment matching result, and the prompt information is obtained at the same time of obtaining the video segment matching result. Exemplarily, the prompt information is generated locally on the terminal.

[0134] In an exemplary embodiment, after it is determined that the video segment matching result does not include the target video segment matched with the comment content, in response to the video segment matching result not including the target video segment matched with the comment content, the operation of displaying prompt information on the video play interface is directly performed. In an exemplary embodiment, for the case that the comment information publishing control is displayed on the video play interface, in response to a triggering operation of the comment information publishing control, and in response to the video segment matching result not including the target video segment matched with the comment content, the operation of displaying prompt information on the video play interface is further performed.

[0135] Exemplarily, while the prompt information is displayed, the comment information including only the comment content can also be displayed on the video play interface.

[0136] Exemplarily, the video play interface after the prompt information is displayed is as shown in FIG. 8I, in which the prompt information 801 of “video segment matching fails” is displayed. Figure 8 Figure 8 In an exemplary embodiment, for the case that the prompt information is displayed, the display color of the comment content edited in the content editing box and the video segment matching control is adjusted to a lighter color, as shown in FIG. 8J. Figure 8 ​The display color of the video segment matching control is restored to normal after the prompt information display ends. Exemplarily, the display duration of the prompt information is set according to experience or is flexibly adjusted according to application scenarios, for example, the display duration of the prompt information is 2 seconds.

[0137] In a possible implementation, for the case that the state of the video segment matching control is changed from the un-enabled state to the enabled state in response to the triggering operation of the video segment matching control, the state of the video segment matching control is further restored from the enabled state to the un-enabled state after the state of the video segment matching control is changed from the un-enabled state to the enabled state. Restoring the state of the video segment matching control to the un-enabled state can lay a foundation for the interactive object to publish the comment information including the video segment playing control next time. For example, the video segment matching control restored to the un-enabled state is as shown in 502 in FIG. 5. Figure 5 The timing of restoring the state of the video segment matching control to the un-enabled state is not limited in the embodiments of the present application, and can be flexibly set according to actual application scenarios.

[0138] In an exemplary embodiment, for the case that the comment information publishing control is displayed on the video playing interface, the state of the video segment matching control is restored from the enabled state to the un-enabled state after the triggering operation of the comment information publishing control is acquired. That is, the state of the video segment matching control is restored from the enabled state to the un-enabled state in response to the triggering operation of the comment information publishing control. In this case, whether the video segment matching result includes the target video segment matched with the comment content or not, the operation of restoring the state of the video segment matching control from the enabled state to the un-enabled state is triggered according to the triggering operation of the comment information publishing control.

[0139] In an exemplary embodiment, the state of the video segment matching control is restored from the enabled state to the un-enabled state in a case that it is determined that the video segment matching result does not include the target video segment matched with the comment content. That is, the state of the video segment matching control is restored from the enabled state to the un-enabled state in response to the video segment matching result not including the target video segment matched with the comment content.

[0140] In an exemplary embodiment, in a case that it is determined that the video segment matching result includes the target video segment matched with the comment content, the video segment matching control can be first kept in the enabled state, and then the state of the video segment matching control is restored from the enabled state to the un-enabled state when the triggering operation of the comment information publishing control is acquired.

[0141] Of course, in the example embodiment, in the case that the state of the video segment matching control is changed from the un-enabled state to the enabled state in response to the triggering operation of the video segment matching control, the video segment matching control can also be kept in the enabled state regardless of whether the triggering operation of the comment information publishing control is acquired and regardless of whether the video segment matching result includes the target video segment matching the comment content, until the triggering operation of the video segment matching control is acquired again, and the state of the video segment matching control is changed from the enabled state to the un-enabled state.

[0142] In the example embodiment, in response to the triggering operation of the comment information publishing control, the comment content in the video playing interface is cleared to facilitate the interactive object to provide the comment content that the interactive object wants to publish next time.

[0143] It should be noted that the terminal publishing the target comment information is taken as an example for description in the embodiments of the present application. After the terminal publishing the target comment information successfully publishes the target comment information, the server will distribute the target comment information to other terminals watching the first video, so that the other terminals can display the target comment information on the video playing interface when the video is played to the publishing time of the target comment information. The other terminals display the target comment information and the operation process after the target comment information is displayed are the same as the terminal publishing the target comment information, and the embodiments of the present application will not be described again.

[0144] In the example embodiment, the display process of the comment information is as shown in Figure 9 The target interactive object edits the comment content in the content editing box, and the terminal identifies the text information corresponding to the comment content in real time. The target interactive object triggers the video segment matching control, and the terminal sends the text information to the server. The server performs video segment matching calculation, and clips the target video segment matching the comment content. The server returns the target video segment to the terminal. The target interactive object triggers the comment information publishing control, the terminal generates the target comment information including the comment information and the video segment playing control, and the terminal displays the target comment information on the video playing interface. The first interactive object (which can be the target interactive object or other interactive object) watches the target video segment by triggering the video segment playing control in the target comment information. The first interactive object can drag, zoom in, zoom out or close the area playing the target video segment in the process of watching the target video segment.

[0145] The embodiment of the present application provides a display mode of target comment information capable of providing information of a video segment dimension. The target video segment capable of being played based on the target comment information is obtained through rapid matching and clipping of a video segment based on identification of comment content edited on an interactive object (for example, some interactive objects like to see spoilers, when watching a video for the second time or the third time, directly revealing the result in the stage of intense or tragic plot, or adding comment content such as a blooper), and the target comment information including a video segment play control used for triggering playing of the target video segment matched with the comment content is generated and displayed in real time. The display mode of the comment information can expand the content for the video spoiler type or blooper and special paid type, diversify the content form of the comment information, and enable the interactive objects who like to see spoilers or bloopers or even more stories behind the scenes to receive more dimensional visual interpretation in the process of watching the video, thereby increasing the watching experience of the interactive objects and increasing the discussion space for the content. Meanwhile, the display mode can also bring more commercial opportunities for payment, such as exclusive payment for members or even advance on-demand payment. The capability is more suitable for the interactive object group who like to watch suspense or detective drama categories and like to be decrypted, and can greatly improve the watching experience of the interactive objects and enable the interactive objects to watch the decryption in real time without any pressure.

[0146] In addition, the target comment information including the comment content and the video segment play control can reduce video jumping, the current interface can be used to view more video segments, and the video segments have a strong correlation with the currently played video, so that the use time of the interactive object on the application or the webpage is increased, and commercial revenue is improved.

[0147] The technical scheme provided by the embodiment of the present application can display comment information including a video segment play control on a video playing interface, and the video segment play control can trigger playing of a target video segment matched with comment content. That is, the comment information displayed by the embodiment of the present application can provide information of a video segment dimension, the dimension of the information provided by the comment information is expanded, the interactive object can watch the target video segment by triggering the video segment play control in the comment information, the interactive experience of the interactive object is improved, and the human-computer interaction rate is improved.

[0148] Referring to Figure 10 The embodiment of the present application provides a display device of comment information, and the device comprises:

[0149] The first display module 1001 is configured to display a video playing interface.

[0150] The first acquisition module 1002 is configured to acquire a video segment matching result of comment content in the video playing interface.

[0151] The second display module 1003 is configured to, in response to the video segment matching result including the target video segment matched with the comment content, display, on the video playing interface, the target comment information including a video segment playing control, and the video segment playing control is configured to trigger playing of the target video segment.

[0152] In a possible implementation, the first video is played on the video playing interface, and the first obtaining module 1002 is configured to identify text information corresponding to the comment content; compare the text information with label information of a candidate video segment, and obtain a video segment matching result of the comment content based on a comparison result, the candidate video segment being a video segment associated with the first video.

[0153] In a possible implementation, the first video is played on the video playing interface, and the first obtaining module 1002 is configured to identify text information corresponding to the comment content; send the text information to a server, the server being configured to compare the text information with label information of a candidate video segment, obtain a video segment matching result of the comment content based on a comparison result, and return the video segment matching result, the candidate video segment being a video segment associated with the first video; and receive the video segment matching result returned by the server.

[0154] In a possible implementation, the first obtaining module 1002 is further configured to, in response to the comparison result indicating that the text information matches the label information of the first candidate video segment successfully, obtain, based on the first candidate video segment, a target video segment matched with the comment content, and take the target video segment as the video segment matching result.

[0155] In a possible implementation, the second display module 1003 is further configured to, in response to the video segment matching result not including the target video segment matched with the comment content, display, on the video playing interface, prompt information, the prompt information being configured to prompt that the comment content is not successfully matched to a video segment.

[0156] In a possible implementation, the video playing interface displays a comment information publishing control, and the second display module 1003 is further configured to, in response to a trigger operation of the comment information publishing control and in response to the video segment matching result not including the target video segment matched with the comment content, display, on the video playing interface, the prompt information.

[0157] In a possible implementation, referring to Figure 11 The apparatus further includes:

[0158] The playing module 1004 is configured to, in response to a trigger operation of the video segment playing control, play the target video segment on the video playing interface.

[0159] In a possible implementation, the video playing interface plays the first video, and the playing module 1004 is configured to play the target video clip on the video playing interface and pause playing of the first video.

[0160] In a possible implementation, the playing module 1004 is further configured to, in response to a resuming playing operation of the first video, cancel playing of the target video clip and resume playing of the first video.

[0161] In a possible implementation, referring to Figure 11 The apparatus further includes:

[0162] The second obtaining module 1005 is configured to, in response to a triggering operation of a target region in the video playing interface, obtain a resuming playing operation of the first video, the target region being any region in the video playing interface except a region playing the target video clip.

[0163] In a possible implementation, the region playing the target video clip supports at least one of size adjustment and position transformation.

[0164] In a possible implementation, the video playing interface displays a content editing box, and the comment content is content displayed in the content editing box.

[0165] In a possible implementation, the video playing interface displays a video clip matching control and a comment information publishing control, and the first obtaining module 1002 is configured to, in response to a triggering operation of the video clip matching control, obtain a video clip matching result of the comment content in the video playing interface.

[0166] The second display module 1003 is configured to, in response to a triggering operation of the comment information publishing control and in response to the video clip matching result including a target video clip matching the comment content, display, on the video playing interface, target comment information including a video clip playing control.

[0167] In a possible implementation, referring to Figure 11 The apparatus further includes:

[0168] The state transformation module 1006 is configured to, in response to a triggering operation of the video clip matching control, transform a state of the video clip matching control from an unenabled state to an enabled state.

[0169] In a possible implementation, the state transformation module 1006 is further configured to, in response to a triggering operation of the comment information publishing control, transform the state of the video clip matching control from the enabled state to the unenabled state.

[0170] In a possible implementation, the state transformation module 1006 is further configured to, in response to the video segment matching result not including the target video segment matched with the comment content, restore the state of the video segment matching control from the enabled state to the disabled state.

[0171] In a possible implementation, the second display module 1003 is configured to control the target comment information including the video segment playing control to move from the reference position of the video playing interface into the video playing interface, so as to display the target comment information including the video segment playing control on the video playing interface.

[0172] The technical scheme provided by the embodiments of the present application can display a comment information including a video segment playing control on a video playing interface, and the video segment playing control can trigger playing of a target video segment matched with the comment content. That is, the comment information displayed by the embodiments of the present application can provide information in the dimension of a video segment, thereby expanding the dimension of information provided by the comment information, and the interactive object can watch the target video segment by triggering the video segment playing control in the comment information, which is beneficial to improving the interactive experience of the interactive object and further improving the human-computer interaction rate.

[0173] It should be noted that the apparatus provided by the above embodiments is only used as an example to divide the above functional modules, and in actual application, the above functions can be completed by different functional modules according to needs, that is, the internal structure of the device is divided into different functional modules to complete all or part of the above described functions. In addition, the apparatus and method embodiments provided by the above embodiments belong to the same concept, and the specific implementation process is described in the method embodiments, which will not be repeated here.

[0174] Figure 12 Fig. 1 is a structural schematic diagram of a terminal provided by an embodiment of the present application. Exemplarily, the terminal can be a smart phone, a tablet computer, an MP3 (Moving Picture Experts Group Audio Layer III) player, an MP4 (Moving Picture Experts Group Audio Layer IV) player, a notebook computer, a desktop computer or a vehicle-mounted terminal. The terminal can also be referred to as a user equipment, a portable terminal, a laptop terminal, a desktop terminal or other names.

[0175] Generally, the terminal includes a processor 1201 and a memory 1202.

[0176] The processor 1201 can include one or more processing cores, such as a 4-core processor, an 8-core processor, and the like. The processor 1201 can be implemented in the form of at least one of a DSP (Digital Signal Processing), an FPGA (Field-Programmable Gate Array), a PLA (Programmable Logic Array). The processor 1201 can also include a main processor and a coprocessor. The main processor is a processor for processing data in an awake state, also known as a CPU (Central Processing Unit). The coprocessor is a low-power processor for processing data in a standby state. In some embodiments, the processor 1201 can be integrated with a GPU (Graphics Processing Unit) for rendering and drawing content required to be displayed by the display screen. In some embodiments, the processor 1201 can further include an AI (Artificial Intelligence) processor for processing machine learning-related computing operations.

[0177] The memory 1202 can include one or more computer-readable storage media that can be non-transitory. The memory 1202 can also include a high-speed random access memory, and a nonvolatile memory such as one or more disk storage devices, flash storage devices. In some embodiments, the non-transitory computer-readable storage medium in the memory 1202 is used to store at least one instruction for being executed by the processor 1201 to enable the terminal to implement the comment information display method provided by the method embodiments in the present application.

[0178] In some embodiments, the terminal can also optionally include a peripheral device interface 1203 and at least one peripheral device. The processor 1201, the memory 1202, and the peripheral device interface 1203 can be connected through a bus or a signal line. Each peripheral device can be connected to the peripheral device interface 1203 through a bus, a signal line, or a circuit board. Specifically, the peripheral device includes at least one of a radio frequency circuit 1204, a display screen 1205, a camera assembly 1206, an audio circuit 1207, and a power supply 1209.

[0179] The peripheral interface 1203 can be used to connect at least one I / O (Input / Output) related peripheral device to the processor 1201 and the memory 1202. In some embodiments, the processor 1201, the memory 1202 and the peripheral interface 1203 are integrated on the same chip or circuit board; in some other embodiments, any one or two of the processor 1201, the memory 1202 and the peripheral interface 1203 can be implemented on a separate chip or circuit board, and the present embodiments are not limited in this regard.

[0180] The radio frequency circuit 1204 is configured to receive and send RF (Radio Frequency) signals, also known as electromagnetic signals. The radio frequency circuit 1204 communicates with communication networks and other communication devices through electromagnetic signals. The radio frequency circuit 1204 converts electrical signals to electromagnetic signals for transmission, or converts electromagnetic signals received into electrical signals. Optionally, the radio frequency circuit 1204 includes an antenna system, an RF transceiver, one or more amplifiers, a tuner, an oscillator, a digital signal processor, a codec chipset, a subscriber identity module card, and the like. The radio frequency circuit 1204 can communicate with other terminals through at least one wireless communication protocol. The wireless communication protocol includes, but is not limited to, a metropolitan area network, various generations of mobile communication networks (2G, 3G, 4G and 5G), a wireless local area network and / or a WiFi (Wireless Fidelity) network. In some embodiments, the radio frequency circuit 1204 can also include NFC (Near Field Communication) related circuitry, and the present application is not limited in this regard.

[0181] The display screen 1205 is configured to display a UI (User Interface). The UI can include graphics, text, icons, video, and any combination thereof. When the display screen 1205 is a touch display screen, the display screen 1205 is further configured to capture touch signals on or above the surface of the display screen 1205. The touch signals can be input to the processor 1201 as control signals for processing. In this case, the display screen 1205 can also be configured to provide virtual buttons and / or virtual keyboard, also known as soft buttons and / or soft keyboard. In some embodiments, the display screen 1205 can be one, disposed on the front panel of the terminal; in other embodiments, the display screen 1205 can be at least two, respectively disposed on different surfaces of the terminal or in a folding design; in other embodiments, the display screen 1205 can be a flexible display screen, disposed on a curved surface or a folding surface of the terminal. Even, the display screen 1205 can also be disposed in an irregular shape, i.e. a special-shaped screen. The display screen 1205 can be made of LCD (Liquid Crystal Display), OLED (Organic Light-Emitting Diode), etc.

[0182] The camera assembly 1206 is configured to capture images or videos. Optionally, the camera assembly 1206 includes a front camera and a rear camera. Typically, the front camera is disposed on the front panel of the terminal, and the rear camera is disposed on the back of the terminal. In some embodiments, the rear camera is at least two, which are any one of a main camera, a depth-of-field camera, a wide-angle camera, and a telephoto camera, to realize the background blur function by fusing the main camera and the depth-of-field camera, the panoramic shooting and VR (Virtual Reality) shooting function by fusing the main camera and the wide-angle camera, or other fusion shooting functions. In some embodiments, the camera assembly 1206 can further include a flash. The flash can be a single-color temperature flash or a dual-color temperature flash. The dual-color temperature flash refers to the combination of a warm light flash and a cold light flash, which can be used for light compensation under different color temperatures.

[0183] The audio circuit 1207 can include a microphone and a speaker. The microphone is used to collect sound waves of the user and the environment, and convert the sound waves into an electrical signal input to the processor 1201 for processing, or input to the radio frequency circuit 1204 to realize voice communication. For the purpose of stereo sound collection or noise reduction, the microphone can be multiple, respectively arranged at different parts of the terminal. The microphone can also be an array microphone or an omnidirectional collection type microphone. The speaker is used to convert the electrical signal from the processor 1201 or the radio frequency circuit 1204 into sound waves. The speaker can be a traditional diaphragm speaker, or a piezoelectric ceramic speaker. When the speaker is a piezoelectric ceramic speaker, not only can the electrical signal be converted into a sound wave that humans can hear, but also can be converted into a sound wave that humans cannot hear for ranging purposes. In some embodiments, the audio circuit 1207 can also include a headphone jack.

[0184] The power supply 1209 is used to supply power to each component in the terminal. The power supply 1209 can be alternating current, direct current, disposable battery or rechargeable battery. When the power supply 1209 includes a rechargeable battery, the rechargeable battery can support wired charging or wireless charging. The rechargeable battery can also be used to support fast charging technology.

[0185] In some embodiments, the terminal also includes one or more sensors 1210. The one or more sensors 1210 include but are not limited to: an acceleration sensor 1211, a gyroscope sensor 1212, a pressure sensor 1213, an optical sensor 1215, and a proximity sensor 1216.

[0186] The acceleration sensor 1211 can detect the acceleration in three coordinate axes of the coordinate system established by the terminal. For example, the acceleration sensor 1211 can be used to detect the components of the gravitational acceleration in three coordinate axes. The processor 1201 can control the display screen 1205 to display the user interface in a landscape view or a portrait view according to the gravitational acceleration signal collected by the acceleration sensor 1211. The acceleration sensor 1211 can also be used for game or user motion data collection.

[0187] The gyroscope sensor 1212 can detect the body direction and rotation angle of the terminal, and the gyroscope sensor 1212 can collect 3D actions of the user on the terminal in cooperation with the acceleration sensor 1211. The processor 1201 can realize the following functions according to the data collected by the gyroscope sensor 1212: motion sensing (such as changing the UI according to the user's tilt operation), image stabilization when shooting, game control, and inertial navigation.

[0188] The pressure sensor 1213 can be disposed at a side frame of the terminal and / or under the display screen 1205. When the pressure sensor 1213 is disposed at the side frame of the terminal, a holding signal of the terminal by a user can be detected, and left-hand or right-hand recognition or a shortcut operation can be performed by the processor 1201 according to the holding signal collected by the pressure sensor 1213. When the pressure sensor 1213 is disposed under the display screen 1205, a pressure operation of the display screen 1205 by a user can be performed by the processor 1201 to control an operable control on a UI interface. The operable control includes at least one of a button control, a scroll bar control, an icon control, and a menu control.

[0189] The optical sensor 1215 is configured to collect ambient light intensity. In an embodiment, the processor 1201 can control display brightness of the display screen 1205 according to the ambient light intensity collected by the optical sensor 1215. Specifically, when the ambient light intensity is high, the display brightness of the display screen 1205 is increased, and when the ambient light intensity is low, the display brightness of the display screen 1205 is decreased. In another embodiment, the processor 1201 can also dynamically adjust a shooting parameter of the camera assembly 1206 according to the ambient light intensity collected by the optical sensor 1215.

[0190] The proximity sensor 1216, also referred to as a distance sensor, is usually disposed at a front panel of the terminal. The proximity sensor 1216 is configured to collect a distance between a user and a front face of the terminal. In an embodiment, when the proximity sensor 1216 detects that the distance between the user and the front face of the terminal gradually decreases, the display screen 1205 is switched from a bright screen state to a dim screen state by the processor 1201, and when the proximity sensor 1216 detects that the distance between the user and the front face of the terminal gradually increases, the display screen 1205 is switched from the dim screen state to the bright screen state by the processor 1201.

[0191] Those skilled in the art can understand that the structures shown in the above embodiments do not constitute a limitation on the terminal, and the terminal can include more or fewer components than those shown in the figures, or combine certain components, or adopt a different arrangement of components. Figure 12 Those skilled in the art can understand that the structures shown in the above embodiments do not constitute a limitation on the terminal, and the terminal can include more or fewer components than those shown in the figures, or combine certain components, or adopt a different arrangement of components.

[0192] In an exemplary embodiment, a computer device is also provided, which includes a processor and a memory having at least one computer program stored therein. The at least one computer program is loaded and executed by one or more processors to enable the computer device to implement any of the above-mentioned display methods of comment information.

[0193] In an exemplary embodiment, a computer readable storage medium is also provided, which has at least one computer program stored therein. The at least one computer program is loaded and executed by a processor of a computer device to enable the computer to implement any of the above-mentioned display methods of comment information.

[0194] In a possible implementation manner, the computer readable storage medium can be a read-only memory (ROM), a random access memory (RAM), a compact disc read-only memory (CD-ROM), a magnetic tape, a floppy disk, an optical data storage device, or the like.

[0195] In the example embodiment, a computer program product is also provided, which includes a computer program or computer instructions loaded and executed by a processor to enable a computer to implement any of the above display methods of comment information.

[0196] It should be understood that "multiple" mentioned herein refers to two or more. The "and / or" describes the association relationship of the associated objects, which means that there can be three relationships, for example, A and / or B can represent the three cases of A existing alone, A and B existing together, and B existing alone. The character " / " generally represents that the front and rear associated objects are in an "or" relationship.

[0197] The above only describes example embodiments of the present application and is not intended to limit the present application. Any modification, equivalent replacement, improvement, etc. made within the spirit and principle of the present application shall be included in the protection scope of the present application.

Claims

1. A method for displaying comment information, characterized in that, The method includes: Displays the video playback interface; Obtain the video segment matching results of the comment content in the video playback interface; In response to the video segment matching result including a target video segment matching the comment content, target comment information including a video segment playback control is displayed on the video playback interface. The video segment playback control is used to trigger the playback of the target video segment. The target video segment is obtained based on a first candidate video segment, which is a video segment whose tag information matches the text information corresponding to the comment content. The tag information of the candidate video segment is used to describe the content of the candidate video segment. The candidate video segment is a segment extracted from an extended video corresponding to a first video, where the first video is the video playback interface. The video played in the video, the extended video is a video associated with the first video used to expand the video content of the first video, the extended video includes at least one of the behind-the-scenes video corresponding to the first video or the paid video corresponding to the first video; the method of obtaining the target video segment includes: determining the position of the first candidate video segment in the extended video, taking the center position of the first candidate video segment in the extended video as the center, cutting video segments of a first duration before and after, the cut video segments constitute the target video segment, the first duration is half of a specified duration, the specified duration is the duration of the target video segment; In response to the triggering operation of the video clip playback control, the target video clip is played in the video playback interface and the first video is paused. The video playback interface displays the video frame of the target video clip and the paused video frame of the first video. In response to a trigger operation on a target area in the video playback interface, the playback of the target video segment is canceled, and the playback of the first video is resumed. The target area is any area in the video playback interface other than the area where the target video segment is played.

2. The method according to claim 1, characterized in that, The step of obtaining the video segment matching result of the comment content in the video playback interface includes: Identify the text information corresponding to the comment content; The text information is compared with the tag information of the candidate video segments, and the video segment matching result of the comment content is obtained based on the comparison result.

3. The method according to claim 1, characterized in that, The step of obtaining the video segment matching result of the comment content in the video playback interface includes: Identify the text information corresponding to the comment content; The text information is sent to the server, which compares the text information with the tag information of the candidate video segments, obtains the video segment matching result of the comment content based on the comparison result, and returns the video segment matching result. Receive the video segment matching result returned by the server.

4. The method according to claim 2 or 3, characterized in that, The process of obtaining video segment matching results for the comment content based on the comparison results includes: In response to the comparison result indicating that the text information and the tag information of the first candidate video segment are successfully matched, a target video segment matching the comment content is obtained based on the first candidate video segment, and the result including the target video segment is used as the video segment matching result.

5. The method according to any one of claims 1-3, characterized in that, The method further includes: If the video segment matching result does not include a target video segment that matches the comment content, a prompt message is displayed on the video playback interface to indicate that the comment content has not successfully matched a video segment.

6. The method according to claim 5, characterized in that, The video playback interface displays a comment posting control. In response to the video segment matching result not including the target video segment matching the comment content, a prompt message is displayed on the video playback interface, including: In response to the triggering operation of the comment posting control, and in response to the video segment matching result not including the target video segment that matches the comment content, the prompt information is displayed on the video playback interface.

7. The method according to claim 1, characterized in that, The area where the target video segment is played supports at least one of the following functions: resizing and repositioning.

8. The method according to any one of claims 1-3 and 6-7, characterized in that, The video playback interface displays a content editing box, and the comment content is the content displayed in the content editing box.

9. The method according to any one of claims 1-3 and 6-7, characterized in that, The video playback interface displays video clip matching controls and comment posting controls. Obtaining the video clip matching results for the comment content in the video playback interface includes: In response to the triggering operation of the video segment matching control, the video segment matching result of the comment content in the video playback interface is obtained; The response to the video segment matching result including the target video segment matching the comment content, is to display the target comment information, including video segment playback controls, on the video playback interface, including: In response to the triggering operation of the comment information posting control, and in response to the video segment matching result including a target video segment that matches the comment content, the target comment information, including the video segment playback control, is displayed on the video playback interface.

10. The method according to claim 9, characterized in that, The method further includes: In response to the triggering operation of the video clip matching control, the state of the video clip matching control is changed from the disabled state to the enabled state.

11. The method according to claim 10, characterized in that, After changing the state of the video segment matching control from disabled to enabled, the method further includes: In response to the triggering operation of the comment posting control, the state of the video clip matching control is restored from the enabled state to the disabled state.

12. The method according to claim 10, characterized in that, After changing the state of the video segment matching control from disabled to enabled, the method further includes: In response to the video segment matching result not including the target video segment that matches the comment content, the state of the video segment matching control is restored from the enabled state to the disabled state.

13. A device for displaying comment information, characterized in that, The device includes: The first display module is used to display the video playback interface; The first acquisition module is used to acquire the video segment matching result of the comment content in the video playback interface; The second display module is configured to, in response to the video segment matching result including a target video segment matching the comment content, display target comment information including a video segment playback control on the video playback interface, wherein the video segment playback control is used to trigger the playback of the target video segment; wherein the target video segment is obtained based on a first candidate video segment, the first candidate video segment being a video segment whose tag information in the candidate video segment successfully matches the text information corresponding to the comment content; the tag information of the candidate video segment is used to describe the content of the candidate video segment, the candidate video segment is a segment extracted from an extended video corresponding to a first video, and the first video is the video. The video played in the playback interface, the extended video is a video associated with the first video used to expand the video content of the first video, the extended video includes at least one of the behind-the-scenes video corresponding to the first video or the paid video corresponding to the first video; the method of obtaining the target video segment includes: determining the position of the first candidate video segment in the extended video, taking the center position of the first candidate video segment in the extended video as the center, cutting video segments of a first duration before and after, the cut video segments constitute the target video segment, the first duration is half of a specified duration, the specified duration is the duration of the target video segment; The playback module is configured to, in response to a trigger operation of the video clip playback control, play the target video clip and pause the first video in the video playback interface, wherein the video playback interface displays the video frame of the target video clip and the paused video frame of the first video; and, in response to a trigger operation of a target area in the video playback interface, cancel the playback of the target video clip and resume the playback of the first video, wherein the target area is any area in the video playback interface other than the area where the target video clip is played.

14. The apparatus according to claim 13, characterized in that, The first acquisition module is used to identify the text information corresponding to the comment content; compare the text information with the tag information of the candidate video segment, and obtain the video segment matching result of the comment content based on the comparison result.

15. The apparatus according to claim 13, characterized in that, The first acquisition module is used to identify the text information corresponding to the comment content; send the text information to the server, the server is used to compare the text information with the tag information of the candidate video segment, obtain the video segment matching result of the comment content based on the comparison result, and return the video segment matching result; and receive the video segment matching result returned by the server.

16. The apparatus according to claim 14 or 15, characterized in that, The first acquisition module is configured to, in response to the comparison result indicating that the text information and the tag information of the first candidate video segment are successfully matched, acquire a target video segment that matches the comment content based on the first candidate video segment, and use the result including the target video segment as the video segment matching result.

17. The apparatus according to any one of claims 13-15, characterized in that, The second display module is further configured to, in response to the video segment matching result not including a target video segment that matches the comment content, display a prompt message on the video playback interface, the prompt message being used to indicate that the comment content has not successfully matched a video segment.

18. The apparatus according to claim 17, characterized in that, The video playback interface displays a comment posting control. The second display module is further configured to respond to the triggering operation of the comment posting control and to display the prompt information on the video playback interface if the video segment matching result does not include a target video segment that matches the comment content.

19. The apparatus according to claim 13, characterized in that, The area where the target video segment is played supports at least one of the following functions: resizing and repositioning.

20. The apparatus according to any one of claims 13-15 and 18-19, characterized in that, The video playback interface displays a content editing box, and the comment content is the content displayed in the content editing box.

21. The apparatus according to any one of claims 13-15 and 18-19, characterized in that, The video playback interface displays a video clip matching control and a comment information posting control. The first acquisition module is used to respond to the triggering operation of the video clip matching control and acquire the video clip matching result of the comment content in the video playback interface. The second display module is configured to display target comment information, including video clip playback controls, on the video playback interface in response to the triggering operation of the comment information posting control and in response to the video clip matching result including a target video clip matching the comment content.

22. The apparatus according to claim 21, characterized in that, The device further includes: The state transition module is used to change the state of the video segment matching control from an inactive state to an active state in response to the trigger operation of the video segment matching control.

23. The apparatus according to claim 22, characterized in that, The state transition module is also used to respond to the triggering operation of the comment information posting control by restoring the state of the video clip matching control from the enabled state to the disabled state.

24. The apparatus according to claim 22, characterized in that, The state transition module is further configured to, in response to the video segment matching result not including a target video segment that matches the comment content, restore the state of the video segment matching control from the enabled state to the disabled state.

25. A computer device, characterized in that, The computer device includes a processor and a memory, the memory storing at least one computer program, the at least one computer program being loaded and executed by the processor to enable the computer device to implement the method for displaying comment information as described in any one of claims 1 to 12.

26. A computer-readable storage medium, characterized in that, The computer-readable storage medium stores at least one computer program, which is loaded and executed by a processor to enable the computer to implement the method for displaying comment information as described in any one of claims 1 to 12.

27. A computer program product, characterized in that, The computer program product includes a computer program or computer instructions, which are loaded and executed by a processor to enable the computer to implement the method for displaying comment information as described in any one of claims 1 to 12.

Citation Information

Patent Citations

  • Multimedia data playing method and device and storage medium

    CN110572716A

  • Video playing method and device, commenting method and device, equipment and storage medium

    CN113194349A