Content generation method and apparatus, device, and storage medium

By analyzing published video content to generate video scripts and generating video content based on events in user interactive content, the problem of low content generation efficiency in the prior art is solved, and efficient and accurate video content generation is achieved.

WO2025118806A1PCT designated stage expired Publication Date: 2025-06-12BEIJING ZITIAO NETWORK TECH CO LTD
View PDF 5 Cites 0 Cited by

Patent Information

Application Number
PCT/CN2024/123070
Authority / Receiving Office
WO · WO
Patent Type
Applications
Current Assignee / Owner
Priority Date
2023-12-04
Filing Date
2024-09-30
Publication Date
2025-06-12

AI Technical Summary

Technical Problem

The prior art is inefficient in recording and sharing user interactive content, and it is difficult to quickly generate high-quality video content.

Method used

By analyzing the published video content, a video script is generated, which at least indicates the event type that matches the first time period of the video to be generated, and a matching event is determined from the user interaction content, and corresponding video content is generated based on these interaction events.

Benefits of technology

It improves the efficiency and accuracy of content generation, can conveniently and accurately record user interaction events, and generate high-quality video content.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN2024123070_12062025_PF_FP_ABST
    Figure CN2024123070_12062025_PF_FP_ABST
Patent Text Reader

Abstract

Embodiments of the present disclosure relate to a content generation method and apparatus, a device, and a storage medium. The method provided herein comprises: determining a video script to be applied, wherein said video script is generated for the analysis of a set of published video content, and said video script at least indicates an event type matching a first time period of a video to be generated; from interactive content associated with a user, determining an interaction event matching the event type; and on the basis of the interaction event, generating video content corresponding to said video script. By means of said method, the embodiments of the present disclosure can use a video script to collect an interaction event from interactive content of a user and generate video content, and can conveniently and accurately record the interaction event of the user, improving the content generation efficiency.
Need to check novelty before this filing date? Find Prior Art

Description

Content generation method, device, equipment and storage medium

[0001] This application claims priority to the Chinese invention patent application entitled “Method, apparatus, device and storage medium for content generation” filed on December 4, 2023, with application number 202311651164.6, the entire contents of which are incorporated by reference into this application. Technical Field

[0002] Example embodiments of the present disclosure generally relate to the field of computers, and more particularly, to methods, devices, apparatuses, and computer-readable storage media for content generation. Background Art

[0003] With the development of computer technology and Internet technology, people can use the Internet and streaming technology to interact online. For example, users can use the Internet to interact with other users in a virtual environment to meet their needs for social interaction, entertainment, content sharing, etc.

[0004] Typically, to enable wider interaction and content sharing, users can record their interactions. For example, they can film or record the interaction to share the content in the form of images or videos. Therefore, how to better capture content and improve its quality is currently a pressing concern.

[0005] Summary of the Invention

[0006] In a first aspect of the present disclosure, a method for content generation is provided. The method includes: determining a video script to be applied, the video script being generated based on an analysis of a set of published video content, the video script at least indicating an event type that matches a first time period of the video to be generated; determining an interactive event matching the event type from interactive content associated with a user; and generating video content corresponding to the video script based on the interactive event.

[0007] In a second aspect of the present disclosure, a content generation apparatus is provided. The apparatus includes: a script determination module configured to determine a video script to be applied, the video script being generated based on an analysis of a set of published video content, the video script at least indicating an event type that matches a first time period of the video to be generated; an event determination module configured to determine an interactive event matching the event type from interactive content associated with a user; and a content generation module configured to generate video content corresponding to the video script based on the interactive event.

[0008] In a third aspect of the present disclosure, an electronic device is provided. The device includes at least one processing unit; and at least one memory coupled to the at least one processing unit and storing instructions for execution by the at least one processing unit. When executed by the at least one processing unit, the instructions cause the device to perform the method of the first aspect.

[0009] In a fourth aspect of the present disclosure, a computer-readable storage medium is provided, wherein a computer program is stored on the computer-readable storage medium, and the computer program can be executed by a processor to implement the method of the first aspect.

[0010] It should be understood that the content described in this summary section is not intended to limit the key features or important features of the embodiments of the present disclosure, nor is it intended to limit the scope of the present disclosure. Other features of the present disclosure will become easily understood through the following description. BRIEF DESCRIPTION OF THE DRAWINGS

[0011] The above and other features, advantages and aspects of the embodiments of the present disclosure will become more apparent with reference to the following detailed description in conjunction with the accompanying drawings. In the accompanying drawings, the same or similar reference numerals represent the same or similar elements, wherein:

[0012] FIG1 shows a schematic diagram of an example environment in which embodiments according to the present disclosure may be implemented;

[0013] FIG2 shows a flowchart of a process for example script generation according to some embodiments of the present disclosure;

[0014] FIG3 illustrates an example of a video script according to some embodiments of the present disclosure;

[0015] 4A and 4B respectively show flowcharts of a video content generation process according to some embodiments of the present disclosure;

[0016] FIG5 is a flowchart illustrating an example process of content generation according to some embodiments of the present disclosure;

[0017] FIG6 is a schematic structural block diagram of an apparatus for content generation according to some embodiments of the present disclosure; and

[0018] FIG7 illustrates a block diagram of an electronic device capable of implementing various embodiments of the present disclosure. DETAILED DESCRIPTION

[0019] The following describes embodiments of the present disclosure in more detail with reference to the accompanying drawings. Although certain embodiments of the present disclosure are shown in the accompanying drawings, it should be understood that the present disclosure can be implemented in various forms and should not be construed as limited to the embodiments described herein. Rather, these embodiments are provided to provide a more thorough and complete understanding of the present disclosure. It should be understood that the drawings and embodiments of the present disclosure are for illustrative purposes only and are not intended to limit the scope of protection of the present disclosure.

[0020] It should be noted that the titles of any section / subsection provided herein are not limiting. Various embodiments are described throughout this document, and any type of embodiment may be included under any section / subsection. Furthermore, the embodiments described in any section / subsection may be combined in any manner with any other embodiments described in the same section / subsection and / or in different sections / subsections.

[0021] In the description of the embodiments of the present disclosure, the term "including" and similar terms should be understood as open inclusion, that is, "including but not limited to". The term "based on" should be understood as "based at least in part on". The term "one embodiment" or "the embodiment" should be understood as "at least one embodiment". The term "some embodiments" should be understood as "at least some embodiments". Other explicit and implicit definitions may be included below. The terms "first", "second", etc. may refer to different or the same objects. Other explicit and implicit definitions may be included below.

[0022] The embodiments of the present disclosure may involve user data, data acquisition and / or use, etc. These aspects shall comply with the corresponding laws, regulations and relevant provisions. In the embodiments of the present disclosure, all data collection, acquisition, processing, processing, forwarding, use, etc. are carried out on the premise that the user is aware of and confirms them. Accordingly, when implementing the various embodiments of the present disclosure, the types, scope of use, and usage scenarios of the data or information that may be involved should be informed to the user and the user's authorization should be obtained in an appropriate manner in accordance with the relevant laws and regulations. The specific notification and / or authorization method may vary according to the actual situation and application scenario, and the scope of the present disclosure is not limited in this respect.

[0023] If this specification and the solutions in the examples involve the processing of personal information, such processing will be done only with a legitimate basis (such as with the consent of the subject of personal information or as necessary for the performance of a contract) and only within the prescribed or agreed scope. A user's refusal to process personal information other than that required for basic functions will not affect the user's use of basic functions.

[0024] As briefly mentioned above, in order to achieve a wider range of interaction and content sharing, users can record the interaction process. Furthermore, in order to enhance the value of content, records are usually generated based on some specific interaction events in the user interaction content.

[0025] In some solutions, complete interactive content is provided to users through recording and playback methods, such as image capture and video capture. Users can then record the video content of a specific interactive event, which is then edited by a video editor to form a complete video content. Accordingly, users can use the video content to share their interactive process (or more specifically, their specific interactive event).

[0026] However, in these solutions, the editor is required to perform editing based on the complete interactive content recording, which requires multiple screening and editing to generate the final video content, and the efficiency of video content generation is low.

[0027] Embodiments of the present disclosure provide a content generation solution. According to the solution, a video script to be applied can be determined, where the video script is generated based on an analysis of a set of published video content and at least indicates an event type that matches a first time period of the video to be generated; interactive events matching the event type are determined from interactive content associated with a user; and video content corresponding to the video script is generated based on the interactive events.

[0028] In this way, the embodiments of the present disclosure can utilize video scripts to capture interactive events in user interactive content and generate video content, which can conveniently and accurately record user interactive events and improve content generation efficiency.

[0029] Various example implementations of this solution are described in detail below in conjunction with the accompanying drawings.

[0030] Sample Environment

[0031] FIG1 shows a schematic diagram of an example environment 100 in which embodiments of the present disclosure can be implemented. As shown in FIG1 , the example environment 100 may include an electronic device 110 .

[0032] In this example environment 100, an application 120 for providing content video generation functionality may be running in the electronic device 110. The application 120 may be used to generate video content 140 based on acquired video material. For example, the application 120 may be used by an editor associated with the application 120 based on video material provided by the user 130 (e.g., interactive content of the user 130 in a virtual scene) to generate video content 140 according to the instructions of the user 130. In some embodiments, the application 120 may also serve as an application that provides interaction between users. For example, the application 120 may provide a virtual scene for the user 130 so that the user 130 can interact with other users in the virtual scene. Accordingly, the application 120 may obtain the above-mentioned video material based on the interactive content of the user 130 and other users, and complete the generation of the video content. Thus, examples of the application 120 may include, but are not limited to, online gaming applications.

[0033] Furthermore, the application 120 can generate video content 140 based on the acquired video material according to the instructions of the user 130. For example, according to the instructions of the user 130, the interactive content associated with the user 130 is uploaded and organized by the editor (for example, the service provider providing the application 120). Furthermore, the electronic device 110 can determine multiple video segments (for example, according to multiple time intervals indicated by the editor) from the interactive content of the user 130 with other users, and obtain the video content 140 based on these video segments (for example, splicing). In some embodiments, the editor can be the maintainer of the application 120 (for example, when the application 120 is provided by an interactive platform, the editor can be the video editing service provider indicated by the interactive platform). The editor can edit the video material uploaded by the user 130 to generate the corresponding video content 140 for use by the user 130. For example, the user 130 can share the video 140 with other users to achieve social purposes.

[0034] Typically, the user 130 can interact with the application 120 via the electronic device 110 and / or its attached devices, for example, interacting in a virtual scene provided by the application 120 and instructing the application 120 to generate the video content 140 .

[0035] Generally, the electronic device 110 may be an independent physical server, or a server cluster or distributed system composed of multiple physical servers. It may also be a cloud server that provides basic cloud computing services such as cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communications, middleware services, domain name services, security services, content distribution networks, and big data and artificial intelligence platforms. For example, computing systems / servers such as mainframes, edge computing nodes, computing devices in a cloud environment, etc.

[0036] In some embodiments, if the computing power of the electronic device 110 meets the requirements, the electronic device 110 can also be any type of mobile terminal, fixed terminal or portable terminal, including a mobile phone, a desktop computer, a laptop computer, a notebook computer, a netbook computer, a tablet computer, a media computer, a multimedia tablet, a handheld computer, a portable game terminal, a VR / AR device, a personal communication system (PCS) device, a personal navigation device, a personal digital assistant (PDA), an audio / video player, a digital camera / camcorder, a positioning device, a television receiver, a radio broadcast receiver, an e-book device, a gaming device or any combination of the foregoing, including accessories and peripherals of these devices or any combination thereof. In some embodiments, the electronic device 110 can also support any type of interface for the user (such as a "wearable" circuit, etc.).

[0037] In addition, if the electronic device 110 is a terminal device used by the editor, the electronic device 110 can also communicate with other electronic devices such as a server to, for example, obtain online support for the content of the application 120 provided by the server, and provide the generated video content 140 to the server for sharing, etc.

[0038] In this case, a communication connection can be established between the electronic device 110 embodied as a terminal device and the server. The communication connection can be established in a wired manner or a wireless manner. The communication connection may include, but is not limited to, a Bluetooth connection, a mobile network connection, a Universal Serial Bus (USB) connection, a Wireless Fidelity (WiFi) connection, etc., and the embodiments of the present disclosure are not limited in this respect.

[0039] It should be understood that the structure and function of the various elements in the environment 100 are described for illustrative purposes only and do not imply any limitation on the scope of the present disclosure.

[0040] Some example embodiments of the present disclosure will be described below with continued reference to the accompanying drawings.

[0041] As described above, embodiments of the present disclosure can utilize video scripts to generate video content. For example, specific interactive events within interactive content can be marked and captured based on the video scripts, and videos can be composed and generated based on the identified interactive events. For ease of understanding, the method for generating video scripts will be explained first, followed by the content generation process implemented using video scripts.

[0042] Sample script generation

[0043] The following is a flowchart of a process 200 for generating an example script according to some embodiments of the present disclosure, described in conjunction with Figure 2. For ease of understanding, the process 200 will also be described in conjunction with the environment 100 shown in Figure 1. For example, the process 200 may be performed by an electronic device 110, exemplified as a server.

[0044] In block 210 , the electronic device 110 obtains the published video content through an analysis module.

[0045] In an embodiment of the present disclosure, the analysis module can be used to obtain published video content, or in other words, to collect published video content. For example, other users generate and record video content for the interaction process with user 130, and the video content records part of the interaction process between other users and user 130. In some embodiments, the analysis module can, for example, periodically capture video content published by, for example, user 130, other users, and editors on the video platform. In some embodiments, the analysis module can be configured in the electronic device 110, or it can be configured independently of the electronic device 110. In the case where the analysis module is configured independently of the electronic device 110, the electronic device 110 can obtain the published video content included in the analysis module by communicating with the analysis module. In some embodiments, the published video content can be, for example, high-profile content on a video website, game scripts, game strategies, etc.

[0046] At block 220 , the electronic device 110 determines a set of target events from the published video content.

[0047] In an embodiment of the present disclosure, electronic device 110 can identify a set of target events from the published video content. Typically, such target events are representative events, such as user 130 completing a specific interactive event within the interactive content. Interactive events can be specific events within a scenario. For example, in a game scenario, interactive events can trigger certain pre-set reward content. For example, an event where user 130 defeats or assists in defeating another user. For example, an event where user 130 defeats multiple other users consecutively. Typically, target events are core components of the interactive content or are related to the core gameplay of the interactive content. For example, in a competitive game, such a target event could be a user defeating or assisting in defeating another user. Furthermore, such target events typically provide positive incentives for user 130's interactions. For example, user 130 may acquire a pre-set special item (or a relatively rare item), win a game or competition against other users, or unleash a pre-set specific skill. In some embodiments, these target events directly associated with actions taken by user 130 can also be referred to as the user's "highlight events" or "highlight moments" (e.g., user 130's victory in a competition). In some embodiments, when a target event is triggered, the electronic device 110 may provide some content such as identification information, special sound effects, etc. to prompt the user 130 and motivate him.

[0048] In some embodiments, the video script may also include other content added to the video by the user 130 or the editor, such as inserted audio effects (e.g., background sound effects of the "original sound" in different virtual scenes), video effects (e.g., additional video effects, image styles, etc.), transition effects (e.g., switching effects of screen content), etc. For example, other content may include "A music effect" inserted at the XX playback time of the video. The electronic device 110 may parse the published video content to determine a set of target events included in the published video content (e.g., all target events included in the published video content). In some embodiments, the electronic device 110 may determine a set of target events based on at least one of the screen information, audio information, and text information of the published content. Specifically, the electronic device 110 may identify whether the published video content includes the target event by, for example, parsing the content included in the video screen information (e.g., whether the reward icon, prompt information, etc. for the target event are included). The electronic device 110 may also determine whether the target event exists by parsing the audio information. For another example, the presence of a target event can be determined by detecting whether there is a special sound effect corresponding to the target event (for example, an incentive sound effect, such as an announcement of the specific content of the target event, such as the user 130 winning a confrontation with other users). The electronic device 110 can also determine whether there is a target event by parsing text information. For example, by parsing whether there is a text prompt associated with the target event in the video content (for example, text prompt content associated with the video content, text prompts for interactive content, script information associated with the game, etc.). Thus, the video content can be parsed through at least one of the picture information, audio information, and text information to mine the target event included in the published video.

[0049] In some embodiments, the electronic device 110 may further determine a group of target events from the published video content based on popularity information of the published video content, wherein the popularity of the video content portion corresponding to the group of target events is greater than a threshold.

[0050] Specifically, the electronic device 110 can determine the popularity information of the published content in each time period based on the interaction of the published video content (e.g., browsing distribution, like distribution, sharing distribution, etc.). Furthermore, the electronic device 110 can extract the video content portion with a popularity greater than a threshold from the published video content, thereby determining a corresponding set of target events.

[0051] In block 230 , the electronic device 110 generates a video script corresponding to the published video content based on a set of target events.

[0052] In some embodiments, the electronic device 110 may construct a video script corresponding to the published video content based on time information and type information of a set of target events.

[0053] In an embodiment of the present disclosure, the electronic device 110 may record the type information and time information of each target event. For example, the time information indicates the time period of the corresponding event in the published video content, or the insertion time and end time of the corresponding target event in the published video content. The type information indicates the event type of the corresponding event. In some embodiments, the type information may be divided based on the content of the target event, for example, it may be that the user 130 achieves a type A target event, the user 130 achieves a type B target event, and so on.

[0054] Furthermore, electronic device 110 can construct a video script corresponding to the published video content based on the target event time information and type information included in the published video content. Taking a target event as an example, electronic device 110 can construct a video script corresponding to the published video content based on the target event time information and type information. For example, the video script indicates that the user achieves target event type A during the X1-X2 time period, and the user achieves target event type B during the X3-X4 time period.

[0055] In an embodiment of the present disclosure, the video script at least indicates the event type that matches the time period (for ease of description, it is described as the first time period) included in the video to be generated. For example, the event type that matches the first time period X1-X2 (for example, the user achieves type A target event) is indicated to associate the target event.

[0056] In some embodiments, the video script may further indicate a time period corresponding to other content in the published video content (eg, at least one of the audio effects, video effects, and transition effects described above).

[0057] In some embodiments, the video script may further indicate an audio effect corresponding to a time period of the video to be generated (for ease of description, described as a second time period). That is, the video script may indicate that an audio effect is to be added to the second time period, so that the video script can be used to insert an audio effect different from the "original sound."

[0058] In some embodiments, the video script may further indicate a video effect corresponding to a time period of the video to be generated (described as a third time period for ease of description). That is, the video script may indicate that a video effect is to be added to the third time period so that the video script can be used to insert an audio effect that is different from the additional video effect.

[0059] In some embodiments, the video script may further indicate a transition effect corresponding to the time period of the video to be generated (for ease of description, this will be described as the fourth time period). That is, the video script may indicate that a transition effect is added to the fourth time period to achieve the connection between different screen contents. This allows the video script to be used to complete the editing of multiple video contents.

[0060] It should be understood that, depending on the scenario, the first time period, the second time period, the third time period, and the fourth time period may be the same or different. For example, the first time period may indicate the time periods X1-X2 and X2-X3, the second time period may indicate the time periods X2-X3, and the third time period may indicate X1-X2 and X2-X3.

[0061] Thus, the electronic device 110 can obtain the corresponding video script after understanding its content structure based on the analysis of the published video content. Subsequently, the video script can be used to generate video content with a content structure similar to that of the published video content. For example, there can be multiple "slots" in the video script (for example, each operation can correspond to a target event or other content), so that complete video content 140 can be produced by inserting content into the "slots". In some embodiments, it is also possible not to independently provide "slots" for other content, but to allow editors to add other content for target events through methods such as prompt information.

[0062] In some embodiments, to facilitate subsequent use of the video script, the electronic device 110 may also add prompt information associated with the video script. For example, the prompt information may include the virtual interactive scene to which the video script refers, the interactive objects in the virtual interactive scene to which it applies, and so on. Thus, the electronic device 110 may subsequently select a more appropriate video script based on the actual interaction situation.

[0063] It should be understood that the electronic device 110 may process a plurality of published video contents in advance to obtain a corresponding plurality of video scripts (or a group of video scripts) for use.

[0064] In some embodiments, the electronic device 110 may generate a video script using a target model, for example. Such a target model may be implemented based on an appropriate machine learning model, examples of which may include but are not limited to a language model.

[0065] Specifically, the electronic device 110 may generate input information to the target model based on the extracted set of target events. For example, the electronic device 110 may generate guidance information (also referred to as guidance words) to the target model based on the time information of the set of target events in the video and the corresponding event description information.

[0066] In yet other embodiments, the electronic device 110 may further provide descriptive information related to virtual objects in the virtual scene. Such virtual objects may include, for example, manipulable virtual characters in the virtual scene. For example, the electronic device 110 may generate guidance information to a target model based on the set of target events and the descriptive information about the virtual objects.

[0067] In some embodiments, such description information may include, for example, scene description information about the virtual scene. For example, the scene description information may include information such as the world view setting of the virtual scene.

[0068] In some embodiments, such description information may also include role description information about a virtual object (eg, a virtual character). For example, the role description information may include role setting information about a specific virtual character.

[0069] Furthermore, the electronic device 110 may generate a video script corresponding to the published video content based on the output information of the target model.

[0070] For example, the electronic device 110 may directly use the output information of the target model as the video script. Alternatively, the electronic device 110 may also edit or modify the output information to determine the final video script.

[0071] In some embodiments, the generated video script may also be requested by, for example, an editor, so that the editor may adjust the video script according to needs to improve the quality of use of the video script.

[0072] It should be understood that, in the video script, the target event included in the same time point (or time period) can be one or more. In some embodiments, in order to improve the presentation effect, some constraints can also be configured accordingly, for example, the target event of type A and the target event of type B will not be presented simultaneously, to avoid, for example, image overlap, audio overlap, etc., which have a negative impact on the presentation effect.

[0073] Therefore, by learning and understanding the published video content, the above information can be fitted into a storyline (for example, a "highlight" storyline), so that the video content made based on the video script fits the highlight storyline. And because the storyline is generated based on the published video content, the popularity of the published video can be used to adjust the focus of the video content and improve the quality of the video content, making the entire highlight storyline more suitable for the current user environment. In addition, this method can also build the storyline by yourself and publish it online to observe the data indicators of the submissions, and can automatically iterate the storyline periodically and continuously.

[0074] Sample content generation

[0075] The following will discuss in detail the process by which electronic device 110 generates video content 140 using a video script. This will be described in conjunction with Figures 3, 4A, and 4B. Figure 3 shows an example of a video script 300 according to some embodiments of the present disclosure. Figures 4A and 4B respectively show flowcharts of video content generation processes 400A and 400B according to some embodiments of the present disclosure. For ease of understanding, this description will also be made in conjunction with environment 100 shown in Figure 1. For example, both process 400A and process 400B can be performed by electronic device 110, which is exemplified as a server.

[0076] In an embodiment of the present disclosure, the electronic device 110 may determine a video script to be applied. For example, the video script may be determined from a set of video scripts generated by analyzing a set of published video content as described above. The video script at least indicates an event type that matches the first time period of the video to be generated. In some embodiments, a set of preset scripts may be obtained by utilizing a set of published video content, such as described above. Furthermore, the electronic device 110 may receive a selection of a video script to be applied from the set of preset scripts. For example, after the electronic device 110 determines that video content 140 needs to be generated (e.g., user 130 sends a generation request to the electronic device 110, or an editor instructs the electronic device 110 to generate video content 140), the electronic device 110 may determine the video script to be applied. In some embodiments, the electronic device 110 may determine the video script to be applied based on the content of the video content 140 that the user 130 desires to generate. For example, if the desired content is to record the "highlight moments" of character B in virtual scene A, the electronic device 110 may determine the video script to be applied based on keywords such as "character highlight moments," "character highlight moments in virtual scene A," and "character B highlight moments." For example, based on the above description, the video script to be applied is selected through the matching results of the keywords and the prompt information.

[0077] In some embodiments, the electronic device 110 may also provide the user 130 with a video script to be applied in advance based on the user 130's request (for example, a set of video scripts may be provided based on the user 130's request, and the video script to be applied may be determined based on the user 130's selection). Thus, the user 130 may interact based on the content indicated by the video script, thereby providing video materials more efficiently and accurately.

[0078] In some embodiments, the electronic device 110 may also update the existing video script based on a preset operation of the user 130. Specifically, the electronic device 110 may provide the user 130 with an initial video script, such as a default generated video script.

[0079] Furthermore, the electronic device 110 may generate guidance information for updating the initial video script based on a preset operation of the user 130. For example, the user may input a text content to describe a specific requirement for updating the initial video script.

[0080] Furthermore, the electronic device 110 can, for example, provide the guidance information and the initial video script to the model, and the model can generate a video script to be applied based on the guidance information and the initial video script. Based on this approach, the embodiments of the present disclosure can support users to further customize or optimize the video script, so that the generated content better meets the user's expectations.

[0081] In some embodiments, after the user 130 completes the interaction, the electronic device 110 may also determine the corresponding video script to be applied based on the analysis results of the collected interaction content of the user 130 (for example, analyzing the interaction events included therein, for example, the interaction event can be determined based on the above-mentioned target event, or in other words, the interaction event corresponds to the above-mentioned target event).

[0082] Typically, to enrich video content, the published video content and the generated video content 140 may include content related to multiple users (for example, the video content 140 may include both interactive content of the user 130 and interactive content of other users).

[0083] In some embodiments, the electronic device 110 can determine the video script to be applied based on the virtual object corresponding to the user 130 in the interactive content. Specifically, after determining the interactive content of the user 130 (for example, using character B to interact with other users' characters in virtual scene A), the electronic device 110 can determine the video script to be applied based on the virtual object corresponding to the user 130 (that is, the character B controlled by the user 130). In this way, the video script can be selected more targeted. Furthermore, the video content 140 generated by the electronic device 110 using the video script can be associated with the virtual character (for example, character B). As described above, the electronic device 110 can generate the video content 140 of the character B by determining the video script that was generated for the interactive content of the character B. For example, in a specific scenario, the video script can be, for example, the video content published by other users for controlling the interactive content of "character B". In this way, the video content 140 can focus on the main interactive content of the user 130.

[0084] With reference to FIG3 , for example, in FIG3 , the video script may include recommendation information 310 to recommend a virtual object (e.g., character B) applicable to the video script 300. In some embodiments, the video script 300 may further include recommendation information 320 and recommendation information 330, so as to utilize the recommendation information 320 to recommend other music that can be used as audio effects, and utilize the recommendation information 330 to recommend video parameters (e.g., the aspect ratio is 16:9). In some embodiments, the visual style of the “slot” in the video script 300 may be labeled based on the specific content of the “event type” (e.g., the special effects, clips, stickers recommended in the video script 300) for easier understanding. For example, the presentation style of slot 340 may be “Skill A”.

[0085] In the video script 300, at least the event type that matches the first time period of the video to be generated is indicated. In some embodiments, the video script 300 may also indicate other contents described above (e.g., audio effects, video effects, transition effects, etc.). For example, the event types included in the 3s-10s time period are "releasing a preset specific skill" (e.g., "skill A") and other content "video effects" (e.g., "video effects A"). For another example, the 10s-30s time period includes "interactive event A", "interactive event B", and other content "video effects" (e.g., "video effects B") and "audio effects".

[0086] Furthermore, electronic device 110 determines interactive events that match an event type from the interactive content associated with user 130. The event type of the interactive event can correspond to the event type of the target event described above, thereby making the interactive event correspond to the target event. In embodiments of the present disclosure, electronic device 110 can determine interactive content associated with user 130. In some embodiments, the interactive content can include interactive content of user 130 in a virtual scene, such as game content. For example, in an example where user 130 utilizes character B in virtual scene A, the interactive goal can be for the points earned by the user's character in virtual scene A to reach a target value. For example, user 130 can control character B to compete against character C controlled by another user. After a game begins, if the points earned by any character reach the target value, the game is considered complete. Accordingly, the game process between the users can be referred to as game content. Accordingly, electronic device 110 can determine interactive events that match the aforementioned event type based on the interactive content associated with a user (e.g., user 130). In some embodiments, the interactive event may be one of the target events used when constructing the video script. For example, the interactive event may be "releasing a preset specific skill." Another example is defeating another user, defeating another user continuously (e.g., two or three times in a row, or defeating two or three different users in a row), etc.

[0087] Accordingly, in some embodiments, when generating a video script, the electronic device 110 may also mark the target events and event types associated with the "user interactive behavior" to facilitate more efficient determination of the "interactive event". In some embodiments, the "interactive event" may be, for example, the "highlight event" or "highlight moment" mentioned above. As a result, the final video content 140 includes, for example, highlight content associated with the game content. In other words, the final video content obtained may be the "highlight video" of the user 130. As a result, the user 130 can use the "highlight video" to share the game situation and enhance the social experience of the user 130.

[0088] Furthermore, the electronic device 110 generates video content corresponding to the video script based on the collected interactive events. In an embodiment of the present disclosure, the electronic device 110 can, based on the event type of the collected interactive events, correspondingly determine interactive events that match the event type from the interactive content associated with the user (e.g., user 130), and use the video script (e.g., inserting the content into a "slot" in the video script) to form the complete video content 140.

[0089] In some embodiments, the electronic device 110 may generate and present the video content 140 to the editor. For example, after generating the video content 140, the electronic device 110 may provide the video content 140 to the editor for editing by the editor. Furthermore, based on the editor's confirmation or editing operation on the video content 140, the electronic device 110 may provide the confirmed video content 140 or the edited video content 140 to the user 130. This allows the editor to perform more refined processing on the video content to improve its quality.

[0090] In some embodiments, the electronic device 110 may have multiple alternative interactive events corresponding to the same slot in the video script (for example, character B uses skill A multiple times). In this case, the electronic device 110 may determine the final interactive event to be used by, for example, providing it to the editor for selection.

[0091] In some embodiments, the electronic device 110 can generate video content using different video scripts. For example, the first video content is generated using the first video script, and the second video content is generated using the second video script. In this case, the electronic device 110 can also present the second video content corresponding to the second video script to the editor, and determine whether to use the first video content or the second video content as the target video content (for example, video content 140) based on the editor's choice of the first video content or the second video content. After determining the target video content, the electronic device 110 can provide the target video content to the user 130. As a result, multiple video scripts can be used at the same time to generate video content for selection, and the user 130 can be satisfied as much as possible by widening the range of selection.

[0092] For example, please refer to Figures 4A and 4B. First, in process 400A, after a user initiates an interaction, electronic device 110 may generate video content 140 that depicts highlight events during a game of user 130. In block 410, after user 130 begins interacting, electronic device 110 may capture user 130's "highlight events" during the game. For example, after user 130 generates an interactive event in which they defeat players consecutively, electronic device 110 may capture this "highlight event." In block 420, after user 130 completes their game, highlight editing may begin. For example, after capturing the "highlight event," an editor may perform video editing. In block 430, electronic device 110 may select a video script based on user 130's "highlight event" (e.g., based on the character featured in the "highlight event"). Furthermore, in block 450, electronic device 110 may utilize the video script to generate video content (e.g., video content 140). In some embodiments, before box 450, box 440 may also be included, and the electronic device 110 may provide the video script to be applied determined by executing box 430 to the editor, so that the editor can adjust the video script to be applied according to actual needs (for example, adjust the insertion position, duration, etc. of the "highlight event").

[0093] In process 400B, the electronic device 110 may first execute box 460 to obtain the video script to be applied. For example, a video script to be applied may be selected based on the instructions of the editor, historical video content popularity information, and the like. Further, in box 470, the electronic device 110 may instruct the user 130 to interact according to the content included in the video script (for example, indicating the "highlight events" that the user 130 can make). Further, in box 480, after the user 130 completes the game, the highlight editing may also be entered. Further, in box 490, the electronic device 110 may use the video script to generate video content (for example, video content 140).

[0094] It should be understood that in process 400B, the editor may also be allowed to adjust the video script (for example, after executing box 480 and before executing box 490, the editor is allowed to adjust the video script) to obtain better quality video content 140.

[0095] In this way, the degree of linkage with the application 120 can be improved, so that the ability to bury points and edit can be coordinated during the interaction process of the user 130. Taking the application 120 that provides virtual scene interaction as an example, the application 120 can link the provided content with video production to enrich the interactive experience of the user 130. In addition, this method can also improve the output efficiency of the editor, so that after providing multiple story lines, the editor's focus will shift from "creative conception" to "creative review" to improve the output of the story line. This can also make it possible to adjust the focus of the work on the interactive platform side to monitoring quality and algorithm tuning.

[0096] In this way, the embodiments of the present disclosure can utilize video scripts to capture interactive events in user interactive content and generate video content, which can conveniently and accurately record user interactive events and improve content generation efficiency.

[0097] Example Process

[0098] 5 shows a flow chart of an example process 500 for content generation according to some embodiments of the present disclosure. The process 500 may be implemented at the electronic device 110. The process 500 is described below with reference to FIG1.

[0099] 5 , at block 510 , the electronic device 110 determines a video script to be applied. In an embodiment of the present disclosure, the video script is generated based on an analysis of a set of published video contents, and the video script at least indicates an event type that matches a first time period of the video to be generated.

[0100] In block 520 , the electronic device 110 determines an interaction event matching the event type from the interaction content associated with the user.

[0101] In block 530 , the electronic device 110 generates video content corresponding to the video script based on the interactive event.

[0102] In some embodiments, determining the video script to be applied includes receiving a selection of the video script to be applied from a set of preset scripts.

[0103] In some embodiments, determining the video script to be applied includes: determining the video script to be applied based on a virtual object corresponding to the user in the interactive content.

[0104] In some embodiments, the published video content is associated with a virtual object.

[0105] In some embodiments, the process 500 further includes: presenting the video content to the editor; and providing the user with confirmed video content or edited video content based on the editor's confirmation operation or editing operation on the video content.

[0106] In some embodiments, the video content is first video content corresponding to the first video script, and process 500 also includes: presenting second video content corresponding to the second video script to the editor; receiving the editor's selection of the first video content or the second video content; and providing the user with target video content generated based on the selected video content.

[0107] In some embodiments, the video script further indicates at least one of the following: an audio effect corresponding to a second time period of the video to be generated; a video effect corresponding to a third time period of the video to be generated; and a transition effect corresponding to a fourth time period of the video to be generated.

[0108] In some embodiments, the interactive content includes game content, and the video content includes highlight content associated with the game content.

[0109] In some embodiments, a video script is generated based on the following process: an analysis module obtains published video content; a set of target events is determined from the published video content; and based on the time information and type information of the set of target events, a video script corresponding to the published video content is constructed, wherein the time information indicates the time period of the corresponding event in the published video content, and the type information indicates the event type of the corresponding event.

[0110] In some embodiments, determining a set of target events from the published video content includes determining a set of target events based on at least one of picture information, audio information, and text information of the published video content.

[0111] In some embodiments, determining a group of target events from the published video content includes: determining a group of target events from the published video content based on popularity information of the published video content, wherein the popularity of a video content portion corresponding to the group of target events is greater than a threshold.

[0112] In some embodiments, generating a video script corresponding to a published video content based on a set of target events includes: constructing a video script based on time information and type information of a set of target events, wherein the time information indicates the time period of the corresponding event in the published video content, and the type information indicates the event type of the corresponding event.

[0113] In some embodiments, generating a video script corresponding to the published video content based on a set of target events includes: generating input information to a target model based on the set of target events; and generating a video script corresponding to the published video content based on output information of the target model.

[0114] In some embodiments, generating input information to the target model based on a set of target events includes generating input information to the target model based on the set of target events and descriptive information associated with the target virtual object.

[0115] In some embodiments, the description information includes at least one of: scene description information of a virtual scene associated with the target virtual object; and character description information about the target virtual object.

[0116] In some embodiments, determining the video script to be applied includes: obtaining an initial video script; generating guidance information for updating the initial video script based on a preset operation of a user; and obtaining the video script to be applied generated based on the guidance information.

[0117] Example devices and equipment

[0118] Embodiments of the present disclosure also provide corresponding apparatuses for implementing the above-described methods or processes. FIG6 shows a schematic structural block diagram of an apparatus 600 for generating content according to certain embodiments of the present disclosure. Apparatus 600 may be implemented as or included in electronic device 110. Each module / component in apparatus 600 may be implemented by hardware, software, firmware, or any combination thereof.

[0119] Apparatus 600 includes a script determination module 610 configured to determine a video script to be applied. The video script is generated based on an analysis of a set of published video content, and the video script at least indicates an event type that matches a first time period of the video to be generated. Apparatus 600 also includes an event determination module 620 configured to determine interactive events matching the event type from interactive content associated with the user. Apparatus 600 also includes a content generation module 630 configured to generate video content corresponding to the video script based on the interactive events.

[0120] In some embodiments, determining the video script to be applied includes receiving a selection of the video script to be applied from a set of preset scripts.

[0121] In some embodiments, determining the video script to be applied includes: determining the video script to be applied based on a virtual object corresponding to the user in the interactive content.

[0122] In some embodiments, the published video content is associated with a virtual object.

[0123] In some embodiments, the apparatus 600 further includes: a first providing module configured to present the video content to the editor; and provide the user with confirmed video content or edited video content based on the editor's confirmation operation or editing operation on the video content.

[0124] In some embodiments, the video content is first video content corresponding to the first video script, and the device 600 also includes: a second providing module, configured to present second video content corresponding to the second video script to the editor; receive the editor's selection of the first video content or the second video content; and provide the user with target video content generated based on the selected video content.

[0125] In some embodiments, the video script further indicates at least one of the following: an audio effect corresponding to a second time period of the video to be generated; a video effect corresponding to a third time period of the video to be generated; and a transition effect corresponding to a fourth time period of the video to be generated.

[0126] In some embodiments, the interactive content includes game content, and the video content includes highlight content associated with the game content.

[0127] In some embodiments, a video script is generated based on the following process: an analysis module obtains published video content; a set of target events is determined from the published video content; and based on the time information and type information of the set of target events, a video script corresponding to the published video content is constructed, wherein the time information indicates the time period of the corresponding event in the published video content, and the type information indicates the event type of the corresponding event.

[0128] In some embodiments, determining a set of target events from the published video content includes determining a set of target events based on at least one of picture information, audio information, and text information of the published video content.

[0129] In some embodiments, determining a group of target events from the published video content includes: determining a group of target events from the published video content based on popularity information of the published video content, wherein the popularity of a video content portion corresponding to the group of target events is greater than a threshold.

[0130] In some embodiments, generating a video script corresponding to a published video content based on a set of target events includes: constructing a video script based on time information and type information of a set of target events, wherein the time information indicates the time period of the corresponding event in the published video content, and the type information indicates the event type of the corresponding event.

[0131] In some embodiments, generating a video script corresponding to the published video content based on a set of target events includes: generating input information to a target model based on the set of target events; and generating a video script corresponding to the published video content based on output information of the target model.

[0132] In some embodiments, generating input information to the target model based on a set of target events includes generating input information to the target model based on the set of target events and descriptive information associated with the target virtual object.

[0133] In some embodiments, the description information includes at least one of: scene description information of a virtual scene associated with the target virtual object; and character description information about the target virtual object.

[0134] In some embodiments, the determination module 610 is further configured to: obtain an initial video script; generate guidance information for updating the initial video script based on a preset operation of the user; and obtain a video script to be applied generated based on the guidance information.

[0135] FIG7 shows a block diagram of an electronic device 700 in which one or more embodiments of the present disclosure may be implemented. It should be understood that the electronic device 700 shown in FIG7 is merely exemplary and should not be construed as limiting the functionality and scope of the embodiments described herein. The electronic device 700 shown in FIG7 can be used to implement the electronic device 110 of FIG1 .

[0136] As shown in FIG7 , electronic device 700 is a general-purpose electronic device. Components of electronic device 700 may include, but are not limited to, one or more processors or processing units 710, memory 720, storage device 730, one or more communication units 740, one or more input devices 750, and one or more output devices 760. Processing unit 710 may be a real or virtual processor and is capable of performing various processes according to programs stored in memory 720. In a multi-processor system, multiple processing units execute computer-executable instructions in parallel to enhance the parallel processing capabilities of electronic device 700.

[0137] The electronic device 700 typically includes a plurality of computer storage media. Such media can be any accessible media that can be obtained by the electronic device 700, including but not limited to volatile and non-volatile media, removable and non-removable media. The memory 720 can be a volatile memory (e.g., a register, a cache, a random access memory (RAM)), a non-volatile memory (e.g., a read-only memory (ROM), an electrically erasable programmable read-only memory (EEPROM), flash memory), or some combination thereof. The storage device 730 can be a removable or non-removable medium and can include a machine-readable medium, such as a flash drive, a disk, or any other medium that can be used to store information and / or data (e.g., training data for training) and can be accessed within the electronic device 700.

[0138] The electronic device 700 may further include additional removable / non-removable, volatile / non-volatile storage media. Although not shown in FIG. 7 , a disk drive for reading from or writing to a removable, non-volatile disk (e.g., a “floppy disk”) and an optical drive for reading from or writing to a removable, non-volatile optical disk may be provided. In these cases, each drive may be connected to a bus (not shown) by one or more data media interfaces. The memory 720 may include a computer program product 727 having one or more program modules configured to perform various methods or actions of various embodiments of the present disclosure.

[0139] The communication unit 740 enables communication with other electronic devices via a communication medium. Additionally, the functions of the components of the electronic device 700 can be implemented as a single computing cluster or multiple computing machines that can communicate via a communication connection. Thus, the electronic device 700 can operate in a networked environment using a logical connection with one or more other servers, a network personal computer (PC), or another network node.

[0140] Input device 750 may be one or more input devices, such as a mouse, keyboard, or trackball. Output device 760 may be one or more output devices, such as a display, a speaker, or a printer. Electronic device 700 may also communicate with one or more external devices (not shown) via communication unit 740 as needed, such as storage devices, display devices, or the like, with one or more devices that allow a user to interact with electronic device 700, or with any device that allows electronic device 700 to communicate with one or more other electronic devices (e.g., a network card, a modem, etc.). Such communication may be performed via an input / output (I / O) interface (not shown).

[0141] According to an exemplary implementation of the present disclosure, a computer-readable storage medium is provided, on which computer-executable instructions are stored, wherein the computer-executable instructions are executed by a processor to implement the method described above. According to an exemplary implementation of the present disclosure, a computer program product is also provided, which is tangibly stored on a non-transitory computer-readable medium and includes computer-executable instructions, and the computer-executable instructions are executed by a processor to implement the method described above.

[0142] Various aspects of the present disclosure are described herein with reference to flowcharts and / or block diagrams of methods, apparatuses, devices, and computer program products implemented according to the present disclosure. It should be understood that each block of the flowcharts and / or block diagrams, and combinations of blocks in the flowcharts and / or block diagrams, can be implemented by computer-readable program instructions.

[0143] These computer-readable program instructions can be provided to a processing unit of a general-purpose computer, a special-purpose computer, or other programmable data processing device, thereby producing a machine, such that when these instructions are executed by the processing unit of the computer or other programmable data processing device, a device is generated that implements the functions / actions specified in one or more blocks in the flowchart and / or block diagram. These computer-readable program instructions can also be stored in a computer-readable storage medium, where these instructions cause the computer, programmable data processing device, and / or other device to operate in a specific manner. Thus, the computer-readable medium storing the instructions comprises an article of manufacture that includes instructions for implementing various aspects of the functions / actions specified in one or more blocks in the flowchart and / or block diagram.

[0144] Computer-readable program instructions can be loaded onto a computer, other programmable data processing apparatus, or other device so that a series of operational steps are performed on the computer, other programmable data processing apparatus, or other device to produce a computer-implemented process, thereby causing the instructions executed on the computer, other programmable data processing apparatus, or other device to implement the functions / actions specified in one or more boxes in the flowchart and / or block diagram.

[0145] The flow charts and block diagrams in the accompanying drawings show the possible architecture, functions and operations of the systems, methods and computer program products according to multiple implementations of the present disclosure. In this regard, each box in the flow chart or block diagram can represent a part for a module, program segment or instruction, and a part for a module, program segment or instruction comprises one or more executable instructions for realizing the logical function of the specification. In some alternative implementations, the functions marked in the box can also occur in a sequence different from that marked in the accompanying drawings. For example, two continuous boxes can actually be executed substantially in parallel, and they can sometimes be executed in the opposite order, depending on the functions involved. It should also be noted that each box in the block diagram and / or flow chart, and the combination of the boxes in the block diagram and / or flow chart can be realized by a special hardware-based system that performs the function or action of the specification, or can be realized by a combination of special hardware and computer instructions.

[0146] While various implementations of the present disclosure have been described above, the foregoing description is intended to be illustrative, not exhaustive, and not limited to the disclosed implementations. Many modifications and variations will be apparent to those skilled in the art without departing from the scope and spirit of the described implementations. The terminology used herein is selected to best explain the principles of the implementations, their practical applications, or improvements to existing technologies, or to enable others skilled in the art to understand the various implementations disclosed herein.

Claims

1. A method for generating content, comprising: Determine a video script to be applied, the video script being generated based on an analysis of a set of published video contents, the video script at least indicating an event type matching a first time period of the video to be generated; Determining, from the interactive content associated with the user, an interactive event matching the event type; as well as Based on the interactive event, video content corresponding to the video script is generated.

2. The method according to claim 1, wherein determining the video script to be applied comprises: A selection of the video script to be applied is received from a group of preset scripts.

3. The method according to claim 1, wherein determining the video script to be applied comprises: The video script to be applied is determined based on the virtual object corresponding to the user in the interactive content. The method according to claim 3 , wherein the published video content is associated with the virtual object.

5. The method according to claim 1, further comprising: Presenting the video content to the editor; as well as Based on the confirmation operation or editing operation of the editor on the video content, the confirmed video content or the edited video content is provided to the user.

6. The method according to claim 5, wherein the video content is first video content corresponding to a first video script, the method further comprising: Presenting second video content corresponding to the second video script to the editor; receiving a selection by the editor regarding the first video content or the second video content; as well as Target video content generated based on the selected video content is provided to the user.

7. The method of claim 1, wherein the video script further indicates at least one of the following: an audio effect corresponding to the second time period of the video to be generated; a video effect corresponding to a third time period of the video to be generated; A transition effect corresponding to a fourth time period of the video to be generated.

8. The method according to claim 1, wherein the interactive content includes game content, and the video content includes highlight content associated with the game content.

9. The method of claim 1, wherein the video script is generated based on the following process: The analysis module obtains the published video content; determining a set of target events from the published video content; as well as Based on the set of target events, a video script corresponding to the published video content is generated.

10. The method of claim 9, wherein determining a set of target events from the published video content comprises: Determine based on at least one of the picture information, audio information and text information of the published video content The set of target events.

11. The method of claim 9, wherein determining a set of target events from the published video content comprises: Based on the popularity information of the published video content, the group of target events is determined from the published video content, wherein the popularity of the video content portion corresponding to the group of target events is greater than a threshold.

12. The method according to claim 9, wherein generating a video script corresponding to the published video content based on the set of target events comprises: The video script is constructed based on the time information and type information of the set of target events, wherein the time information indicates a time period of the corresponding event in the published video content, and the type information indicates an event type of the corresponding event.

13. The method according to claim 9, wherein generating a video script corresponding to the published video content based on the set of target events comprises: generating input information to a target model based on the set of target events; as well as Based on the output information of the target model, the video script corresponding to the published video content is generated.

14. The method of claim 13, wherein generating input information to a target model based on the set of target events comprises: The input information to the target model is generated based on the set of target events and descriptive information associated with a target virtual object.

15. The method according to claim 14, wherein the description information includes an indication of at least one of the following: scene description information of a virtual scene associated with the target virtual object; Character description information about the target virtual object.

16. The method of claim 1, wherein determining the video script to be applied comprises: Get the initial video script; Based on the preset operation of the user, generating guidance information for updating the initial video script; as well as Obtain the video script to be applied that is generated based on the guide information.

17. A content generation device, comprising: a script determination module configured to determine a video script to be applied, the video script being generated based on an analysis of a set of published video contents, the video script at least indicating an event type matching a first time period of the video to be generated; An event determination module, configured to determine an interactive event matching the event type from interactive content associated with the user; as well as The content generation module is configured to generate video content corresponding to the video script based on the interactive event.

18. An electronic device, comprising: at least one processing unit; as well as At least one memory, the at least one memory being coupled to the at least one processing unit and storing instructions for execution by the at least one processing unit, the instructions causing the electronic device to perform the method according to any one of claims 1 to 16 when executed by the at least one processing unit.

19. A computer-readable storage medium having a computer program stored thereon, wherein the computer program can be executed by a processor to implement the method according to any one of claims 1 to 16.

Citation Information

Patent Citations

  • Video editing method and device, computer equipment and storage medium

    CN115120968A

  • Method and device for generating interactive video, equipment and storage medium

    CN115988266A

  • Video generation method and device, electronic equipment and storage medium

    CN116647714A

  • Content generation method and device, equipment and storage medium

    CN117641072A

  • Automatic video montage generation

    US20220108727A1