Video playing method and device, electronic equipment, storage medium and product

By displaying controls associated with the target scene in the comment section, the target video is generated and displayed, solving the problem that users have difficulty efficiently browsing the visual content of specific scenes in the story. This achieves efficient scene-level related recommendations and improves the user experience.

CN120980290APending Publication Date: 2025-11-18DOUYIN VISION CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Application Number
CN202511232117.7
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-08-29
Publication Date
2025-11-18

AI Technical Summary

Technical Problem

After browsing specific scenes in a story, users find it difficult to efficiently find and browse visual content related to those scenes, and existing content platforms cannot achieve scene-level related recommendations.

Method used

By displaying controls associated with the target scene in the comment section, the target video is displayed in response to user triggers. The target video is generated based on comments and various works, and video content matching the target scene is generated using large language models and material recognition technology.

Benefits of technology

This allows users to efficiently jump to visual content related to their target scenario while browsing comments, improving browsing efficiency and immersion.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120980290A_ABST
    Figure CN120980290A_ABST
Patent Text Reader

Abstract

The invention relates to a video playing method and device, electronic equipment, a storage medium and a product, and relates to the field of computers. The video playing method comprises the following steps: displaying a comment of a first work corresponding to a story, wherein the comment relates to a target scene in the story; and in response to the comment or a first control associated with the comment being triggered, displaying a playing interface of a target video corresponding to the target scene, the target video being generated according to content associated with the target scene in one or more works corresponding to the story.
Need to check novelty before this filing date? Find Prior Art

Description

TECHNICAL FIELD

[0001] The present disclosure relates to the field of computer, and in particular, to a video playing method and device, electronic device, storage medium and product. BACKGROUND

[0002] With the development of mobile internet technology, users can browse various works, such as network novels, network videos, etc. through personal computers, mobile phones, tablets and other terminals. Or, for the content on the non-network platform, such as offline published paper books, offline released movies or TV series broadcasted by TV stations, users can also publish comments on these works on the network platform. SUMMARY

[0003] According to some embodiments of the present disclosure, a video playing method is provided, including: displaying a comment of a first work corresponding to a story, the comment relating to a target scene in the story; in response to the comment or a first control associated with the comment being triggered, displaying a playing interface of a target video corresponding to the target scene, wherein the target video is generated according to content associated with the target scene in one or more works corresponding to the story.

[0004] According to some other embodiments of the present disclosure, a video playing device is provided, including: a first display module configured to display a comment of a first work corresponding to a story, the comment relating to a target scene in the story; a second display module configured to, in response to the comment or a first control associated with the comment being triggered, display a playing interface of a target video corresponding to the target scene, wherein the target video is generated according to content associated with the target scene in one or more works corresponding to the story.

[0005] According to some embodiments of the present disclosure, an electronic device is provided, including: a memory; and a processor coupled to the memory, the processor being configured to execute a video playing method of any of the embodiments described in the present disclosure based on instructions stored in the memory.

[0006] According to some embodiments of the present disclosure, a computer readable storage medium is provided, having stored thereon a computer program, which, when executed by a processor, performs a video playing method of any of the embodiments described in the present disclosure.

[0007] According to some embodiments of the present disclosure, a computer program product is provided, which, when running on a computer, causes the computer to implement a video playing method of any of the embodiments described in the present disclosure.

[0008] Other features, aspects, and advantages of the present disclosure will become apparent from the following detailed description of the exemplary embodiments with reference to the following drawings. BRIEF DESCRIPTION OF DRAWINGS

[0009] Embodiments of this disclosure are described below with reference to the accompanying drawings. It should be understood that the drawings described below are merely illustrative of some embodiments of this disclosure and are not intended to limit the scope of this disclosure. In the drawings:

[0010] Figure 1 A flowchart illustrating a video playback method according to some embodiments of the present disclosure is shown.

[0011] Figure 2 A flowchart illustrating a method for generating a target video according to some embodiments of the present disclosure is shown.

[0012] Figure 3 A schematic diagram of a video playback interface according to some embodiments of the present disclosure is shown.

[0013] Figure 4 A schematic diagram of a comment interface according to some embodiments of the present disclosure is shown.

[0014] Figure 5 A schematic diagram of a novel reading interface according to some embodiments of the present disclosure is shown.

[0015] Figure 6 A schematic diagram of the structure of a video playback device according to some embodiments of the present disclosure is shown.

[0016] Figure 7 A block diagram of an electronic device according to some embodiments of the present disclosure is shown.

[0017] Figure 8 Block diagrams of electronic devices according to other embodiments of the present disclosure are shown. Detailed Implementation

[0018] The technical solutions of the embodiments of this disclosure will be clearly and completely described below with reference to the accompanying drawings. It should be understood that this disclosure can be implemented in various forms and should not be construed as limited to the embodiments set forth herein.

[0019] It should be understood that the various steps described in the method embodiments of this disclosure may be performed in different orders and / or in parallel. Furthermore, method embodiments may include additional steps and / or omit the steps shown. The scope of this disclosure is not limited in this respect. Unless otherwise specifically stated, the relative arrangement of components and steps set forth in these embodiments should be interpreted as merely exemplary and does not limit the scope of this disclosure.

[0020] As used in this disclosure, the term "comprising" and its variations are open-ended terms that include at least the following elements / features but do not exclude other elements / features, i.e., "including but not limited to". The term "based on" means "at least partially based on".

[0021] It should be noted that the concepts of "first," "second," etc., used in this disclosure are used only to distinguish different devices, modules, or units, and are not intended to define the order of functions performed by these devices, modules, or units or their interdependencies. Unless otherwise specified, the concepts of "first," "second," etc., are not intended to imply that the objects described herein must be in a given temporal, spatial, rank, or any other given order.

[0022] It should be noted that the terms "a" and "a plurality of" used in this disclosure are illustrative rather than restrictive, and those skilled in the art should understand that, unless otherwise expressly indicated in the context, they should be understood as "one or more".

[0023] The names of messages or information exchanged between multiple devices in the embodiments of this disclosure are for illustrative purposes only and are not intended to limit the scope of such messages or information.

[0024] The user information (including but not limited to user device information, user personal information, etc.) and data (including but not limited to data used for analysis, data stored, data displayed, etc.) involved in this disclosure are all information and data authorized by the user or fully authorized by all parties. Furthermore, the collection, use and processing of the relevant data shall comply with the relevant laws, regulations and standards of the relevant countries and regions, and corresponding operation portals shall be provided for users to choose to authorize or refuse.

[0025] The embodiments of this disclosure are described in detail below with reference to the accompanying drawings; however, this disclosure is not limited to these specific embodiments. These specific embodiments can be combined with each other, and the same or similar concepts or processes may not be described again in some embodiments. Furthermore, in one or more embodiments, specific features, structures, or characteristics can be combined in any suitable manner that will be apparent to those skilled in the art from this disclosure.

[0026] A story may contain one or more compelling scenes. These scenes may be key moments in the story, or scenes that become popular and generate considerable discussion after the work is disseminated. Users may post their thoughts or comments about these scenes in the comment sections of these works. Additionally, users may create their own versions of these scenes.

[0027] For example, some stories may have multiple genres. Some stories might first appear as novels and then be adapted into film or television. Later, user-created video trailers might also be published on content platforms. However, after watching one version of the story, if a user wants to watch other types of works, they need to search for them manually or browse other genre pages. This results in high search costs and low search efficiency for users.

[0028] Some users, after browsing a work or a scene within it, may become so engrossed that they wish to explore other works depicting that scene. For example, after reading a key chapter of a novel, a user might want to view a visual representation of that chapter; after watching a crucial scene in a film, they might want to see more detailed clips; and after watching a scene in an anime, a user might imagine what it would be like if that scene were performed by a live-action actor.

[0029] Some content platforms support displaying related works. For example, after finishing an ebook, other recommended ebooks may be displayed, possibly including those related to the already read one; after watching a movie, similar movies may be recommended. However, the recommended works are usually works with different stories, and scene-level related content recommendations are not supported.

[0030] Analysis revealed that after watching a specific scene in a story, if that scene is particularly memorable or engaging, users will leave comments and discuss it in the work's comment section. This disclosure provides a video playback method that links comments discussing a specific scene to a target video corresponding to that scene. This allows users browsing comments to efficiently trigger playback of the target video.

[0031] Figure 1 A flowchart illustrating a video playback method according to some embodiments of the present disclosure is shown. Figure 1 As shown, the video playback method of this embodiment includes steps S11 to S12. This method can be executed on a user's client or electronic device. The user is, for example, a first user.

[0032] In step S11, a commentary on the first work corresponding to the story is displayed, which relates to the target scene in the story.

[0033] The first work describes the story in a specific genre, such as a novel, video, comic, music, etc. The story can be carried solely by the first work, or it can be carried by other works, such as a second work or more. The type of the second work can be the same as or different from the first work. Even if the first and second works are of the same type, their content may not be entirely identical. For example, if both the first and second works are videos, the two works can be performed by different actors, or one can be live-action and the other animated, etc.

[0034] Comments on the first work can be displayed on the page corresponding to the first work. That is, this page is a comment page for the first work, or the comments on the first work can be displayed on a collection page of comments for multiple works. This disclosure does not impose any restrictions on this.

[0035] The target scenario comes from one or more scenes in the story. The target scenario can be pre-specified or dynamically determined. Whether a comment relates to the target scenario can be determined through keyword matching or semantic analysis based on machine models.

[0036] In step S12, in response to the comment or the first control associated with the comment being triggered, a playback interface for the target video corresponding to the target scene is displayed, wherein the target video is generated based on content associated with the target scene from one or more works corresponding to the story.

[0037] Comments can include at least one of text, images, audio, and video. Comments can be posted by a first user or by users other than the first user. That is, comments in the embodiments of this disclosure can be posted by any user.

[0038] A comment relating to the target scenario can itself act as a triggerable control, or the portion of a comment that relates to the target scenario (such as keywords corresponding to the target scenario) can also act as a triggerable control. In these cases, for text-based comments, the comment itself, or the triggerable content within the comment, can be displayed differently from comments that cannot be triggered.

[0039] The primary control associated with a comment can be displayed in the vicinity of the comment. For example, each comment can correspond to a display element, which includes the comment text and the primary control corresponding to that comment text.

[0040] The comment, or the first control associated with it, can be linked to a link to the target video. Therefore, in response to a trigger on the comment or the first control, the playback interface of the target video can be displayed.

[0041] The target video can be generated based on a first work, a second work different from the first work, or both. The generation of the target video can include one or more of the following: text, images, video, and sound, and the information used to generate the target video must include at least content related to the target scene.

[0042] Through the above embodiments, when comments on a work involve a target scene, those comments can be linked to a target video describing the same scene. This allows users to easily browse visual content of that scene while viewing discussions about it. This enables users to efficiently consume visual content about the same scene within the story, improving browsing efficiency and enhancing the immersive experience.

[0043] As mentioned earlier, the target scenario can be pre-specified or dynamically determined. In some embodiments, scenarios frequently mentioned in the comments of the first work can be identified as target scenarios. For example, the comments of the first work can be clustered to obtain multiple categories; a target category can be determined based on the number of comments in each category; and the theme of the target category can be determined based on the comments in the target category, serving as the target scenario for the story. Comments can be converted into vectors, and clustering can be performed based on these vectors. The target category can be a category with a number of comments exceeding a specified threshold, or a category whose number of comments ranks higher than a specified rank. After determining the target category, the theme of the category can be extracted based on the content of the comments in that category, for example, using a topic analysis model or a large language model. Thus, popular scenarios can be identified as target scenarios based on the level of user discussion surrounding the first work.

[0044] The target video is pre-generated before displaying triggerable comments or the first control. The target video can be generated and published by any user, by the creator of the first work, or automatically generated on the application's server side. The target video can be manually edited and generated by the user, or automatically generated using a model. An exemplary embodiment of the target video generation method is described below.

[0045] Figure 2 A flowchart illustrating a method for generating a target video according to some embodiments of the present disclosure is shown. Figure 2 As shown, the method for generating the target video in this embodiment includes steps S21 to S23.

[0046] In step S21, material description information is determined based on the description of the target scene of the story in the comments.

[0047] Since comments may contain content unrelated to the target scene, they can be filtered to retain only descriptions of the target scene as descriptive material. For example, a comment might read, "This episode was so exciting, I couldn't keep up! I especially loved watching A cry!" Here, "This episode was so exciting, I couldn't keep up! I especially loved it" expresses the user's emotion, while "A cries" describes the target scene. A refers to a character in the story or the actor who plays that character in the first production.

[0048] The material description information describes the material used to generate the target video, such as a description of the content involved in the material. For example, the material description information could come from comments, such as "A is crying," etc. Furthermore, the material description information can be determined based on the content involved in the target scene in at least one of the first work associated with the story, or other works besides the first work, or the context of the involved content. For example, if in the first or second work of the story, the scene where A is crying is in front of a convenience store in heavy rain, then the material description information could also include "convenience store," "heavy rain," etc. The material description information can be determined by searching for the target scene in the works and based on the search results.

[0049] The determination of material description information can be achieved in other ways. For example, the description of the target scene can be input into a large model such as a Large Language Model (LLM), and a first processing instruction can be input into the large model. The first processing instruction is used to instruct the large model to output material description information for generating the video of the target scene based on the input description of the target scene. Alternatively, other methods can be used to determine the material description information. For example, the material description information can be pre-defined for each target scene. Another example is that multiple material types can be pre-specified, and the content description for each type of material can be determined based on the description information of the target scene.

[0050] In step S22, materials are obtained from one or more works corresponding to the story based on the material description information.

[0051] That is, when acquiring materials, one can acquire materials only from the first work, or only from other works different from the first work, or acquire materials from both the first work and other works.

[0052] The acquired materials are matched with the material description information. In some embodiments, if the material description information includes a material type, materials belonging to that material type are acquired.

[0053] For example, determining the material description information based on the description of the target scene in the comments includes: determining one or more material types and the corresponding material description information for each material type; and obtaining materials from one or more works corresponding to the story based on the material description information includes: obtaining materials from one or more works corresponding to the story based on one or more material types and the corresponding material description information for each material type. Based on the description of the target scene, the type of the target scene can be determined. The type of the target scene includes, for example, visual, auditory, etc. For example, if the description of the target scene describes character appearance, character interaction, scenery, or architectural appearance, then the target scene belongs to the visual type; if the description of the target scene describes sound or character dialogue, then the target scene belongs to the auditory type; if the description of the target scene describes multiple aspects of the environment, then the target scene may belong to both visual and auditory types. Then, the material type matching the type of the target scene can be determined. For example, visual types correspond to images, videos, etc., and auditory types correspond to timbre, music, speech, etc. The obtained materials belong to the material type matching the material description information. In addition, text-based materials can be obtained for any type of target scene.

[0054] In some embodiments, obtaining materials from one or more works corresponding to a story includes: for each material type, determining works that include that material type from the one or more works corresponding to the story; and obtaining the material from the determined works based on material description information corresponding to that material type. For example, for image types, covers and screenshots can be obtained from video works, or illustrations can be obtained from novels; for another example, timbre and voice can be obtained from audiobooks and films, and music can be obtained from films, music albums, etc.; and for yet another example, text can be obtained from novels. Thus, the required materials can be obtained quickly and accurately.

[0055] In some embodiments, for each material type and the material description information corresponding to that material type, in response to the absence of a work including the material type among one or more works corresponding to the story, or the absence of content matching the material description information among the works including the material type, materials matching the material description information are determined from works excluding the material type. Thus, if no material type matching the target scene can be directly obtained, other material types can be used. For example, if the target scene is a visual type, but the story only has text-based works and does not include illustrations, then images, videos, or other materials cannot be directly obtained. In this case, text content related to the target scene can be obtained from the works, and the visual content of the target video can be generated using this text. This allows the methods of the embodiments of this disclosure to be applicable to more application scenarios.

[0056] In step S23, the target video is generated based on the acquired materials.

[0057] The acquired footage can be processed using a pre-set video generation template to generate the target video. Alternatively, this step can be accomplished using a model for video generation. For example, the acquired footage and a generation instruction for the target video can be input into the model, which instructs the model to generate a video based on the input footage. The target video is then output by the model.

[0058] Through the above embodiments, materials can be obtained based on the description of the target scene, and a target video can be generated based on the materials. This enables the generated target video to match the target scene and achieves an automated video generation process, improving the efficiency of video generation.

[0059] Below, we will describe methods for obtaining materials based on several types of works, using examples as examples.

[0060] In some embodiments, one or more works corresponding to the story include film and television works. Obtaining material from one or more works corresponding to the story, based on material description information, includes: performing speech recognition on the audio of the film and television work, and performing image recognition on one or more frames of the film and television work; determining the content of one or more segments of the film and television work based on the results of speech recognition and image recognition; and obtaining at least one of video, image, and music from the segments corresponding to the material description information as the material.

[0061] Image recognition and speech recognition can both be achieved using image recognition models. Image recognition determines the content of different segments of a work from a visual perspective, while speech recognition determines the content from an auditory perspective. Therefore, by combining both visual and auditory perspectives, the content to be expressed in each segment of a video can be determined more accurately, allowing for the identification of segments that best match the descriptive information. Once the matching segments are identified, the remaining material can be extracted more accurately and quickly.

[0062] In some embodiments, one or more works corresponding to the story include audiobooks. Obtaining materials from one or more works corresponding to the story, based on material description information, includes: determining the text corresponding to the audiobook; determining multiple audio segments of the audiobook and the character corresponding to each audio segment based on the text; and determining the audio segment of the character corresponding to the material description information from the multiple audio segments as timbre material or background sound material. Some audiobooks are performed by characters with different timbres. For example, the narrator, character A, and character B may all use different timbres. By parsing the text, the content corresponding to each character in the audiobook can be determined, thereby allowing for more accurate extraction of the audio or timbre corresponding to the character as material.

[0063] In some embodiments, obtaining materials from one or more works corresponding to the story, based on the material description information, includes at least one of the following: obtaining text materials from a novel corresponding to the story; obtaining music materials from music corresponding to the story; and obtaining image materials from user-generated content or AI-generated content corresponding to the story. When obtaining text materials from a novel, content matching the description information of the target scene can be searched directly within the novel. When obtaining music materials from music corresponding to the story, the search can be conducted by the name of the music or by the category of the music. For example, if the target scene describes a sad atmosphere, music with a sad melody can be searched. The obtained image materials can also be content from user-created or AI-generated works, for example, content from character illustrations created based on novel text.

[0064] The generation of a target video can be user-triggered. For example, while browsing works of the story (e.g., the first or second work), a user may trigger the generation of a target video if they become interested in the currently viewed content. In some embodiments, the target video is generated by a viewer of the second work of the story, and this target video is associated with the target scene in the second work, which may be the same as or different from the first work. Comments in the first work that relate to the target scene are identified; these comments are then linked to the target video, or a first control is added to the comment to link to the target video. Thus, after a user triggers the generation of the target video, comments relating to the target scene can serve as the entry point for the generated target video. This allows the user or other users to quickly jump to the target video while browsing comments. This improves user creation efficiency and provides a distribution channel for the generated target video, thereby increasing the number of views.

[0065] Figure 3 A schematic diagram of a video playback interface according to some embodiments of the present disclosure is shown. Figure 3 As shown, the first installment of a story is played on video playback interface 3, and user comments 31 to 33 are displayed in the form of bullet comments on playback interface 3. Comment 33, "I love watching the scene where the male and female protagonists meet," relates to the target scene, so the bullet comment containing comment 33 is converted into a triggerable control and associated with a link to the target video. In response to the user triggering control 33, for example, video interface 34 can be displayed, which is used to play the target video describing the meeting of the male and female protagonists. Video interface 34 can be located in a completely independent new page from interface 3, or it can be located within interface 3, for example, floating above the content shown on interface 3, etc., and this disclosure does not impose any restrictions on this.

[0066] Figure 4 A schematic diagram of a comment interface according to some embodiments of this disclosure is shown. Figure 4 As shown, a comment area 41 is displayed floating above the playback interface or information display interface 4 of the work. This area includes multiple comments, such as comment 411 posted by user 2 and comment 413 posted by user 3. Comment 411 relates to the target scene, so a control 412 is displayed nearby. This control is used to trigger the display of the target video related to comment 411. The display format of the target video is not limited and can be referenced. Figure 3 The relevant explanations are in the text.

[0067] In addition to adding an entry point to the target video in the comments, this entry point can also be added to the novel itself, allowing users to quickly browse visual content corresponding to their current reading content while browsing text-based works. In some embodiments, in response to the generation of the target video, text corresponding to the target scene is determined in the novel corresponding to the story; the text corresponding to the target scene is linked to the target video. For example, the text of the target scene can be converted into text carrying a link, or a second control carrying a link can be added to the adjacent area of ​​the text. Thus, when readers of the novel reach the corresponding content, they can browse the corresponding target video by triggering the text or triggering the second control.

[0068] Figure 5 A schematic diagram of a novel reading interface according to some embodiments of the present disclosure is shown. Figure 5 As shown, in the reading interface 5, three paragraphs 51 to 53 are exemplaryly displayed, where paragraph 52 is related to the target scene. A control 54 can then be displayed after paragraph 52 (or at any other adjacent location). Control 54 is used to trigger the display of the target video corresponding to the target scene involved in paragraph 52. The display format of the target video is not limited and can be referenced... Figure 3 The relevant explanations are in the text.

[0069] The methods of the embodiments of this disclosure have been described above. The apparatus for implementing the methods of the above embodiments is described below.

[0070] Figure 6 A schematic diagram of a video playback device according to some embodiments of the present disclosure is shown. For example... Figure 6 As shown, the video playback device 6 of this embodiment includes: a first display module 61 configured to display comments on a first work corresponding to a story, the comments relating to a target scene in the story; and a second display module 62 configured to display a playback interface for a target video corresponding to the target scene in response to the comments or the triggering of a first control associated with the comments, wherein the target video is generated based on content associated with the target scene from one or more works corresponding to the story.

[0071] In some embodiments, the video playback device 6 further includes: a generation module configured to determine material description information based on the description of the target scene of the story in the comments; to obtain material from one or more works corresponding to the story based on the material description information; and to generate the target video based on the obtained material.

[0072] In some embodiments, the generation module is configured to: determine one or more material types and material description information corresponding to each material type based on the description of the target scene of the story in the comments; and obtain materials from one or more works corresponding to the story based on the one or more material types and the material description information corresponding to each material type.

[0073] In some embodiments, the generation module is configured to: for each material type, determine works that include the material type from one or more works corresponding to the story; and obtain the material from the determined works based on material description information corresponding to the material type.

[0074] In some embodiments, the generation module is configured to: for each material type and the material description information corresponding to the material type, in response to the absence of a work including the material type in one or more works corresponding to the story, or the absence of content matching the material description information in the works including the material type, determine the material matching the material description information from the works excluding the material type.

[0075] In some embodiments, one or more works corresponding to the story include film and television works, and the generation module is configured to: perform speech recognition on the sound of the film and television works, and perform image recognition on one or more frames of the film and television works; determine the content of one or more segments of the film and television works based on the results of the speech recognition and the results of the image recognition; and obtain at least one of video, image, and music from the segments corresponding to the material description information as the material.

[0076] In some embodiments, one or more works corresponding to the story include audiobooks, and the generation module is configured to: determine text corresponding to the audiobook; determine multiple audio segments of the audiobook and a character corresponding to each audio segment based on the text; and determine, from the multiple audio segments, an audio segment of the character corresponding to the material description information as timbre material or background sound material.

[0077] In some embodiments, the generation module is configured to: obtain text materials from a novel corresponding to the story; obtain music materials from music corresponding to the story; and obtain image materials from user-generated content or AI-generated content corresponding to the story.

[0078] In some embodiments, the video playback device 6 further includes: a determining module configured to cluster comments on the first work to obtain multiple categories; determine a target category based on the number of comments in each category; and determine the theme of the target category as the target scene of the story based on the comments in the target category.

[0079] In some embodiments, the video playback device 6 further includes: a first linking module configured to trigger the generation of the target video by a viewer of a second work of the story, the target video being associated with the target scene in the second work, the second work being the same as or different from the first work; identifying comments in the comments of the first work that relate to the target scene; linking the comments to the target video, or adding the first control for linking to the target video to the comments.

[0080] In some embodiments, the video playback device 6 further includes: a second linking module configured to, in response to the generation of the target video, determine text corresponding to the target scene in a novel corresponding to the story; and link the text corresponding to the target scene to the target video.

[0081] Figure 7 A block diagram of an electronic device according to some embodiments of the present disclosure is shown.

[0082] Memory 71 is used to store one or more computer-readable instructions. Memory 71 may include any combination of various forms of computer-readable storage media, such as volatile memory and / or non-volatile memory, including but not limited to random access memory (RAM), dynamic random access memory (DRAM), static random access memory (SRAM), read-only memory (ROM), and flash memory. Memory 71 may, for example, store operating systems, application programs, bootloaders, databases, and other programs, as well as various application programs and various data.

[0083] The processor 72 is configured to execute computer-readable instructions to implement the method described in any of the foregoing embodiments. Specific implementations of each step of the method can be found in the above embodiments; repeated details will not be elaborated upon here.

[0084] The processor 72 can be configured to perform the steps of the foregoing embodiments. The processor 72 can be embodied in various processing devices, such as a central processing unit (CPU), a network processor (NP), etc.; it can also be a digital signal processor (DSP), an application-specific integrated circuit (ASIC), a field-programmable gate array (FPGA), or other programmable logic devices, discrete gate or transistor logic devices, or discrete hardware components. The central processing unit (CPU) can be an x86 or ARM architecture, etc.

[0085] The processor 72 and the memory 71 can communicate with each other directly or indirectly. For example, the processor 72 and the memory 71 can communicate via a network. The network can include a wireless network, a wired network, and / or any combination of wireless and wired networks. The processor 72 and the memory 71 can also communicate with each other via a system bus, which is not limited in this disclosure.

[0086] It should be noted that Figure 7 The components of the electronic device 7 shown are merely exemplary and not limiting. The electronic device 7 may have other components depending on the specific application requirements. The processor 72 can control other components in the electronic device 7 to perform desired functions.

[0087] Electronic device 7 can be implemented by software, firmware and / or hardware, and can be integrated into a device with the relevant application installed.

[0088] Figure 8 Block diagrams of electronic devices according to other embodiments of the present disclosure are shown.

[0089] Figure 8 The electronic device 8 shown can be a computer system with a dedicated hardware structure, which can perform corresponding functions when the relevant application is installed.

[0090] Electronic devices include, but are not limited to, mobile terminals such as smartphones, laptops, personal digital assistants (PDAs), tablet computers (PCs), PMPs (portable multimedia players), in-vehicle terminals (such as in-vehicle navigation terminals), wearable devices, and fixed terminals such as digital televisions and desktop computers.

[0091] like Figure 8 As shown, the Central Processing Unit (CPU) 81 performs various processes based on a program stored in the Read-Only Memory (ROM) 82 or a program loaded from the storage section 88 into the Random Access Memory (RAM) 83. The RAM 83 stores data required as needed when the CPU 81 performs various processes, etc. The CPU is merely exemplary; it could also be other types of processors, such as the various processors described above. The ROM 82, RAM 83, and storage section 88 can be various forms of computer-readable storage media. It should be noted that although... Figure 8 The diagram shows ROM 82, RAM 83 and storage section 88, but one or more of them may be combined or located in the same or different memory or storage modules.

[0092] CPU 81, ROM 82 and RAM 83 are interconnected via bus 84. Input / output interface 85 is also connected to bus 84.

[0093] The following components are connected to the input / output interface 85: input section 86, such as a touchscreen, touchpad, keyboard, mouse, image sensor, microphone, accelerometer, gyroscope, etc.; output section 87, including displays such as cathode ray tube (CRT), liquid crystal display (LCD), speakers, vibrators, etc.; storage section 88, including hard disk, magnetic tape, etc.; and communication section 89, including network interface cards such as LAN cards, modems, etc. The communication section 89 allows communication processing to be performed via a network such as the Internet. It is easy to understand that, although... Figure 8 The portion of the electronic device 8 shown communicates via bus 84, but it may also communicate via a network or other means, wherein the network may include a wireless network, a wired network, and / or any combination of wireless and wired networks.

[0094] As needed, drive 810 is also connected to input / output interface 85. Removable media 811, such as disks, optical disks, magneto-optical disks, semiconductor memories, etc., are installed on drive 810 as needed, so that computer programs read from them can be installed into storage section 88 as needed.

[0095] When the above series of processes are implemented through software, the program constituting the software can be installed from a network such as the Internet or a storage medium such as a removable medium 811.

[0096] According to embodiments of this disclosure, the processes described above with reference to the flowcharts can be implemented as computer software programs. For example, some embodiments of this disclosure include a computer program product that, when run on a computer, causes the computer to perform the methods described in any of the foregoing embodiments. The computer program product includes computer instructions carried on a computer-readable medium, containing program code for performing the methods shown in the flowcharts. In such embodiments, the computer instructions can be downloaded and installed from a network via communication section 89, or installed from storage section 88, or installed from ROM 82. When the computer program is executed by CPU 81, the methods of embodiments of this disclosure are performed.

[0097] It should be noted that, in the context of this disclosure, a computer-readable medium can be a tangible medium that may contain or store programs for use by or in conjunction with an instruction execution system, apparatus, or device.

[0098] A computer-readable medium may be a computer-readable storage medium, a computer-readable signal medium, or any combination thereof.

[0099] Computer-readable storage media include, but are not limited to, electrical, magnetic, optical, electromagnetic, infrared, or semiconductor systems, apparatuses, or devices, or any combination thereof. More specific examples of computer-readable storage media may include, but are not limited to, electrical connections having one or more wires, portable computer disks, hard disks, random access memory (RAM), read-only memory (ROM), erasable programmable read-only memory (EPROM or flash memory), optical fiber, portable compact disk read-only memory (CD-ROM), optical storage devices, magnetic storage devices, or any suitable combination thereof. In this disclosure, a computer-readable storage medium can be any tangible medium containing or storing a program that can be used by or in conjunction with an instruction execution system, apparatus, or device. Computer instructions are stored on the computer-readable storage medium that, when executed by a processor, implement the methods described in any of the foregoing embodiments.

[0100] Computer-readable signal media may include data signals propagated in baseband or as part of a carrier wave, carrying computer-readable program code. Such propagated data signals may take various forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination thereof. Computer-readable signal media may also be any computer-readable medium other than computer-readable storage media, capable of sending, propagating, or transmitting programs for use by or in connection with an instruction execution system, apparatus, or device. The program code contained on the computer-readable medium may be transmitted using any suitable medium, including but not limited to: wires, optical fibers, RF (radio frequency), etc., or any suitable combination thereof.

[0101] The aforementioned computer-readable medium may be included in the aforementioned electronic device; or it may exist independently and not assembled into the electronic device.

[0102] In some embodiments, a computer program is also provided, comprising: instructions that, when executed by a processor, cause the processor to perform the methods described in any of the foregoing embodiments. For example, the instructions may be embodied in computer program code.

[0103] In embodiments of this disclosure, computer program code for performing the operations of this disclosure can be written in one or more programming languages ​​or a combination thereof. These programming languages ​​include, but are not limited to, object-oriented programming languages ​​such as Java, Smalltalk, and C++, as well as conventional procedural programming languages ​​such as the "C" language or similar programming languages. The program code can be executed entirely on the user's computer, partially on the user's computer, as a standalone software package, partially on the user's computer and partially on a remote computer, or entirely on a remote computer or server. In cases involving remote computers, the remote computer can be connected to the user's computer via any type of network (including a local area network (LAN) or a wide area network (WAN)), or it can be connected to an external computer (e.g., via the Internet using an Internet service provider).

[0104] The flowcharts and block diagrams in the accompanying drawings illustrate the architecture, functionality, and operation of possible implementations of systems, methods, and computer program products according to various embodiments of this disclosure. In this regard, each block in a flowchart or block diagram may represent a module, segment, or portion of code containing one or more executable instructions for implementing a specified logical function. It should also be noted that in some alternative implementations, the functions indicated in the blocks may occur in a different order than those indicated in the drawings. For example, two consecutively indicated blocks may actually be executed substantially in parallel, and they may sometimes be executed in reverse order, depending on the functions involved. It should also be noted that each block in the block diagrams and / or flowcharts, and combinations of blocks in the block diagrams and / or flowcharts, can be implemented using a dedicated hardware-based system that performs the specified function or operation, or using a combination of dedicated hardware and computer instructions.

[0105] The functions described above can be performed, at least in part, by one or more hardware logic components. For example, without limitation, exemplary hardware logic components that can be used include: field-programmable gate arrays (FPGAs), application-specific integrated circuits (ASICs), application-specific standard products (ASSPs), system-on-a-chip (SoCs), complex programmable logic devices (CPLDs), and so on.

[0106] While specific embodiments of this disclosure have been described in detail by way of example, those skilled in the art should understand that the examples are for illustrative purposes only and not intended to limit the scope of this disclosure. Those skilled in the art should understand that modifications can be made to the above embodiments without departing from the scope and spirit of this disclosure. The scope of this disclosure is defined by the appended claims.

Claims

1. A video playback method, comprising: Displays comments on the first work corresponding to the story, with the comments relating to the target scene in the story; In response to the comment, or the first control associated with the comment being triggered, a playback interface for the target video corresponding to the target scene is displayed, wherein the target video is generated based on content associated with the target scene from one or more works corresponding to the story.

2. The video playback method according to claim 1 further includes: Based on the description of the target scene of the story in the comments, determine the material description information; Based on the material description information, obtain materials from one or more works corresponding to the story; The target video is generated based on the acquired materials.

3. The video playback method according to claim 2, wherein: The step of determining the material description information based on the description of the target scene of the story in the comments includes: determining one or more material types and material description information corresponding to each material type based on the description of the target scene of the story in the comments; The step of obtaining materials from one or more works corresponding to the story based on the material description information includes: obtaining materials from one or more works corresponding to the story based on the one or more material types and the material description information corresponding to each material type.

4. The video playback method according to claim 3, wherein, The step of obtaining materials from one or more works corresponding to the story based on the one or more material types and the material description information corresponding to each material type includes: For each material type, identify works that include that material type from one or more works corresponding to the story; The material is obtained from the determined work based on the material description information corresponding to the material type.

5. The video playback method according to claim 3, wherein, The step of obtaining materials from one or more works corresponding to the story based on the one or more material types and the material description information corresponding to each material type includes: For each material type and the material description information corresponding to the material type, in response to the fact that there is no work including the material type in one or more works corresponding to the story, or that there is no content matching the material description information in the works including the material type, the material matching the material description information is determined from the works excluding the material type.

6. The video playback method according to any one of claims 2 to 5, wherein, The one or more works corresponding to the story include film and television works, and the step of obtaining materials from the one or more works corresponding to the story based on the material description information includes: The system performs speech recognition on the audio of the film and television works, and performs image recognition on one or more frames of the film and television works. Based on the results of the speech recognition and the image recognition, the content of one or more segments of the film or television work is determined; From the segments corresponding to the material description information, at least one of video, image, and music is obtained as the material.

7. The video playback method according to any one of claims 2 to 5, wherein, One or more works corresponding to the story include audiobooks, and obtaining materials from one or more works corresponding to the story based on the material description information includes: Determine the text corresponding to the audiobook; Based on the text, determine multiple audio segments of the audiobook and the character corresponding to each audio segment; From the plurality of audio segments, determine the audio segment of the character corresponding to the material description information, and use it as timbre material or background sound material.

8. The video playback method according to any one of claims 2 to 5, wherein, The step of obtaining materials from one or more works corresponding to the story based on the material description information includes at least one of the following: Obtain textual material from the novel corresponding to the story; Obtain musical materials from the music corresponding to the story. Image materials are obtained from user-generated or AI-generated content corresponding to the story.

9. The video playback method according to claim 1, further comprising: The comments on the first work are clustered to obtain multiple categories; Determine the target category based on the number of comments in each category; Based on the comments in the target category, determine the theme of the target category as the target scene of the story.

10. The video playback method according to claim 1, further comprising: The target video is generated by the viewer of the second work of the story. The target video is associated with the target scene in the second work. The second work may be the same as or different from the first work. Identify comments related to the target scene in the comments of the first work; Link the comment to the target video, or add the first control to the comment for linking to the target video.

11. The video playback method according to any one of claims 2 to 5, further comprising: In response to the generation of the target video, text corresponding to the target scene is determined in the novel corresponding to the story; Link the text corresponding to the target scene to the target video.

12. A video playback device, comprising: The first display module is configured to display comments on a first work corresponding to the story, the comments relating to a target scene in the story; The second display module is configured to display a playback interface of a target video corresponding to the target scene in response to the comment or the triggering of a first control associated with the comment, wherein the target video is generated based on content associated with the target scene from one or more works corresponding to the story.

13. An electronic device, comprising: Memory; as well as A processor coupled to the memory, the processor being configured to execute the video playback method as described in any one of claims 1 to 11 based on instructions stored in the memory.

14. A computer-readable storage medium having a computer program stored thereon, which, when executed by a processor, implements the video playback method according to any one of claims 1 to 11.

15. A computer program product, when run on a computer, causes the computer to implement the video playback method according to any one of claims 1 to 11.