Video editing method and apparatus, electronic device, and computer-readable storage medium

By obtaining video material copy, determining the editing template and automatically intercepting and splicing material clips, the problem of cumbersome operations during video editing is solved, and efficient video editing is achieved.

WO2025139743A1PCT designated stage expired Publication Date: 2025-07-03BEIJING ZITIAO NETWORK TECH CO LTD
View PDF 5 Cites 0 Cited by

Patent Information

Application Number
PCT/CN2024/137928
Authority / Receiving Office
WO · WO
Patent Type
Applications
Current Assignee / Owner
Priority Date
2023-12-29
Filing Date
2024-12-09
Publication Date
2025-07-03

AI Technical Summary

Technical Problem

During the video editing process, users need to select and match materials from a huge material library, which makes the operation cumbersome and time-consuming, making it difficult to complete video editing efficiently.

Method used

By obtaining the copy of the video material, determining the target video editing template, intercepting the material clip and filling it into the template slot, generating the second video material that presents the editing effect, and splicing it into the original material, realizing automatic video editing.

Benefits of technology

It simplifies user operation process, reduces the time-consuming video editing, improves processing efficiency, and avoids manual editing and splicing operations.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN2024137928_03072025_PF_FP_ABST
    Figure CN2024137928_03072025_PF_FP_ABST
Patent Text Reader

Abstract

The present disclosure relates to a video editing method and apparatus, an electronic device, and a computer-readable storage medium, which can improve the processing efficiency of video editing for video materials to be processed. The method comprises: acquiring a first video material and target video copy corresponding to the first video material; determining a target video editing template on the basis of the target video copy, wherein the target video editing template is used for indicating a filling slot for at least one video clip and an editing effect applied to the at least one video clip; capturing at least one material segment from the first video material, wherein the at least one material segment corresponds to the filling slot for the at least one video clip and is used for filling the target video editing template to form the at least one video clip; generating a second video material on the basis of the at least one material segment and the target video editing template, wherein the second video material is used for presenting the editing effect applied to the at least one material segment; and stitching the second video material before the first video material to generate a target video.
Need to check novelty before this filing date? Find Prior Art

Description

Video editing method, device, electronic device and computer-readable storage medium

[0001] This application claims priority to the Chinese invention patent application entitled “Video editing method, device, electronic device and computer-readable storage medium” filed on December 29, 2023, with application number 202311865439.6. The entire contents of that application are incorporated by reference into this application. Technical Field

[0002] The present disclosure relates to the field of video processing technology, and in particular to a video editing method, device, electronic device, and computer-readable storage medium. Background Art

[0003] Video editing is a crucial step in the user's video creation process. By combining text, music, sound effects, stickers, special effects, and other materials, the visual and auditory effects of the video can be enriched, making the message and emotion more prominent. However, video editing is often time-consuming and labor-intensive. Users need to select and match appropriate materials from a vast library, and then design, add, and adjust the video editing style, making the process quite tedious. Summary of the Invention

[0004] In order to solve the above technical problems or at least partially solve the above technical problems, the present disclosure provides a video editing method, device, electronic device, storage medium and program product.

[0005] According to a first aspect of an embodiment of the present disclosure, a video editing method is provided, which includes: obtaining a first video material and a target video text corresponding to the first video material; determining a target video editing template based on the target video text, wherein the target video editing template is used to indicate a filling slot of at least one video clip and an editing effect applied to the at least one video clip; intercepting at least one material clip from the first video material; the at least one material clip, corresponding to the filling slot of the at least one video clip, is used to fill the target video editing template to form the at least one video clip; generating a second video material based on the at least one material clip and the target video editing template; wherein the second video material is used to present the editing effect applied to the at least one material clip; and generating a target video by splicing the second video material in front of the first video material.

[0006] In some embodiments of the present disclosure, obtaining the first video material and the target video copy corresponding to the first video material includes: obtaining the target video copy of the video material to be processed; based on the target video copy, dividing the video material to be processed into multiple video materials, each of the multiple video materials is a first video material; or obtaining the first video material and the video copy corresponding to the first video material includes: obtaining the to-be-processed video material copy of the video material to be processed; based on the to-be-processed video material copy, dividing the video material to be processed into multiple video materials, each of the multiple video materials is a first video material; obtaining the target video copy of the first video material; generating the target video by splicing the second video material in front of the first video material includes: generating an edited video corresponding to the first video material by splicing the second video material in front of the first video material to generate the edited video corresponding to each video material; splicing the edited videos corresponding to each video material to obtain the target video.

[0007] In some embodiments of the present disclosure, the target video editing template is also used to indicate a video title to fill a slot; the method also includes: generating a video title of the first video material based on a video text of the first video material, and the video title is used to fill the video title filling slot; generating a second video material based on the at least one material clip and the target video editing template, including: generating a second video material based on the at least one material clip, the video title and the target video editing template; wherein the second video material is used to present the editing effect applied to the at least one material clip and the video title.

[0008] In some embodiments of the present disclosure, determining the target video editing template based on the target video copy includes: determining the category label of the category to which the target video copy belongs; determining the target video editing template corresponding to the category label from multiple video editing templates, different category labels corresponding to different video editing templates.

[0009] In some embodiments of the present disclosure, the filling slot of each video clip is also used to indicate the duration of the corresponding video clip; intercepting at least one material clip from the first video material includes: slicing the first video material to obtain at least one slice; performing highlight clip extraction processing on the at least one slice to obtain at least one highlight clip, and the at least one slice corresponds to the at least one highlight clip one-to-one; determining N first material clips from the at least one highlight clip, the duration of each first material clip is the duration of a video clip, each first material clip corresponds to a highlight clip, and N is a natural number; when N is less than M, determining MN second material clips from the first video material, the duration of each second material clip is the duration of a video clip other than the video clip corresponding to the N first material clips, M is the number of the at least one material clip, and M is a positive integer; determining the N first material clips and the MN second material clips as the at least one material clip; when N is equal to M, determining the N first material clips as the at least one material clip.

[0010] In some embodiments of the present disclosure, when N is less than M, MN second material segments are determined from the first video material, including: when the video duration of the remaining video other than the N first material segments in the first video material is greater than the target duration, determining the MN second material segments from the remaining video, the target duration being the sum of the durations of the video segments in the at least one video segment other than the video segments corresponding to the N first material segments; when the video duration of the remaining video is less than or equal to the target duration, determining K second material segments from the remaining video, K being a natural number less than or equal to MN; and determining MNK second material segments from the first video material.

[0011] In some embodiments of the present disclosure, the filling slot of each video clip is also used to indicate the arrangement order of the corresponding video clips. The second video material is generated based on the at least one material clip and the target video editing template, including: based on the target video editing template, the at least one material clip is synthesized according to the arrangement order of the corresponding video clips to generate the second video material.

[0012] According to a second aspect of an embodiment of the present disclosure, a video editing device is provided, which includes: an acquisition module for acquiring a first video material and a target video text corresponding to the first video material; a determination module for determining a target video editing template based on the target video text, wherein the target video editing template is used to indicate a filling slot of at least one video clip and an editing effect applied to the at least one video clip; a capture module for capturing at least one material clip from the first video material; the at least one material clip corresponds to the filling slot of the at least one video clip and is used to fill the target video editing template to form the at least one video clip; a generation module for generating a second video material based on the at least one material clip and the target video editing template; wherein the second video material is used to present the editing effect applied to the at least one material clip; and a target video is generated by splicing the second video material in front of the first video material.

[0013] In some embodiments of the present disclosure, the acquisition module is specifically used to obtain the target video copy of the video material to be processed; based on the target video copy, the video material to be processed is divided into multiple video materials, and each video material in the multiple video materials is a first video material; or, the acquisition of the first video material and the video copy corresponding to the first video material includes: obtaining the to-be-processed video material copy of the video material to be processed; based on the to-be-processed video material copy, the video material to be processed is divided into multiple video materials, and each video material in the multiple video materials is a first video material; obtaining the target video copy of the first video material; the generation module is specifically used to generate an edited video corresponding to the first video material by splicing the second video material in front of the first video material, so as to generate an edited video corresponding to each video material; and splicing the edited videos corresponding to each video material to obtain the target video.

[0014] In some embodiments of the present disclosure, the target video editing template is also used to indicate a video title to fill a slot; the generation module is also used to generate a video title for the first video material based on a video text of the first video material, and the video title is used to fill the video title filling slot; the generation module is specifically used to generate a second video material based on the at least one material clip, the video title and the target video editing template; wherein the second video material is used to present the editing effect applied to the at least one material clip and the video title.

[0015] In some embodiments of the present disclosure, the determination module is specifically used to determine the category label of the category to which the target video copy belongs; and determine the target video editing template corresponding to the category label from multiple video editing templates, where different category labels correspond to different video editing templates.

[0016] In some embodiments of the present disclosure, the filling slot of each video clip is also used to indicate the duration of the corresponding video clip; the interception module is specifically used to slice the first video material to obtain at least one slice; perform highlight clip extraction processing on the at least one slice to obtain at least one highlight clip, and the at least one slice corresponds to the at least one highlight clip one-to-one; from the at least one highlight clip, N first material clips are determined, the duration of each first material clip is the duration of a video clip, each first material clip corresponds to a highlight clip, and N is a natural number; when N is less than M, MN second material clips are determined from the first video material, the duration of each second material clip is the duration of a video clip other than the video clip corresponding to the N first material clips, M is the number of the at least one material clip, and M is a positive integer; the N first material clips and the MN second material clips are determined as the at least one material clip; when N is equal to M, the N first material clips are determined as the at least one material clip.

[0017] In some embodiments of the present disclosure, the interception module is specifically used to determine the MN second material segments from the remaining video in the first video material, when the video duration of the remaining video excluding the N first material segments in the first video material is greater than the target duration, and the target duration is the sum of the durations of the video segments in the at least one video segment excluding the video segments corresponding to the N first material segments; determine K second material segments from the remaining video, when the video duration of the remaining video is less than or equal to the target duration, and K is a natural number less than or equal to MN; and determine MNK second material segments from the first video material.

[0018] In some embodiments of the present disclosure, the filling slot of each video clip is also used to indicate the arrangement order of the corresponding video clips. The generation module is specifically used to synthesize the at least one material clip according to the arrangement order of the corresponding video clips based on the target video editing template to generate a second video material.

[0019] According to a third aspect of an embodiment of the present disclosure, an electronic device is provided, comprising a processor, a memory, and a computer program stored in the memory and executable on the processor, wherein the computer program, when executed by the processor, implements the video editing method as described in the first aspect.

[0020] According to a fourth aspect of the embodiments of the present disclosure, a computer-readable storage medium is provided, on which a computer program is stored. When the computer program is executed by a processor, the video editing method as described in the first aspect is implemented.

[0021] According to a fifth aspect of the embodiments of the present disclosure, a computer program product is provided, wherein the computer program product includes a computer program. When the computer program product runs on a processor, the processor executes the computer program to implement the video editing method as described in the first aspect.

[0022] In a sixth aspect of an embodiment of the present disclosure, a chip is provided, which includes a processor and a communication interface, wherein the communication interface is coupled to the processor, and the processor is used to run program instructions to implement the video editing method as described in the first aspect. BRIEF DESCRIPTION OF THE DRAWINGS

[0023] The accompanying drawings, which are incorporated in and constitute a part of this specification, illustrate embodiments consistent with the present disclosure and, together with the description, serve to explain the principles of the present disclosure.

[0024] In order to more clearly illustrate the embodiments of the present disclosure or the technical solutions in the prior art, the following briefly introduces the drawings required for use in the embodiments or the description of the prior art. Obviously, for ordinary technicians in this field, other drawings can be obtained based on these drawings without any creative work.

[0025] FIG1 is a flow chart of a video editing method provided by an embodiment of the present disclosure;

[0026] FIG2 is a schematic diagram of one interface of the video editing method provided by an embodiment of the present disclosure;

[0027] FIG3 is a second schematic diagram of an interface of a video editing method provided by an embodiment of the present disclosure;

[0028] FIG4 is a third schematic diagram of an interface of a video editing method provided by an embodiment of the present disclosure;

[0029] FIG5 is a fourth schematic diagram of an interface of a video editing method provided by an embodiment of the present disclosure;

[0030] FIG6 is a structural block diagram of a video editing device provided by an embodiment of the present disclosure;

[0031] FIG7 is a structural block diagram of an electronic device provided by an embodiment of the present disclosure. DETAILED DESCRIPTION

[0032] In order to more clearly understand the above-mentioned objectives, features and advantages of the present disclosure, the scheme of the present disclosure will be further described below. It should be noted that the embodiments of the present disclosure and the features therein can be combined with each other in the absence of conflict.

[0033] The technical solution provided by the embodiments of the present disclosure has the following advantages over the prior art: In the embodiments of the present disclosure, a first video material and a target video text corresponding to the first video material are obtained; a target video editing template is determined based on the target video text, and the target video editing template is used to indicate the filling slot of at least one video clip and the editing effect applied to the at least one video clip; at least one material clip is intercepted from the first video material; the at least one material clip, corresponding to the filling slot of the at least one video clip, is used to fill the target video editing template to form the at least one video clip; a second video material is generated based on the at least one material clip and the target video editing template; wherein the second video material is used to present the editing effect applied to the at least one material clip; and a target video is generated by splicing the second video material in front of the first video material. In this way, the present solution designs a video editing template, which is used to indicate the filling slots of at least one video clip and the editing effects applied to the at least one video clip, and then obtains the target video text corresponding to the first video material; determines the target video editing template based on the target video text, and then generates the second video material corresponding to the first video material according to the target video editing template, and generates the target video by splicing the second video material in front of the first video material, thereby realizing automatic video editing processing of the first video material, without the need for the user to manually clip the first video material, nor the need for the user to manually splice the clipped video segments, nor the need for the user to manually add editing effects to the clipped video segments to generate the second video material, nor the need for the user to manually splice the second video material in front of the first video material to generate the target video. Therefore, the disclosed solution simplifies the user's operation process for video editing processing of the first video material, reduces the time consumption of video editing processing, and improves the processing efficiency of video editing processing.

[0034] In the following description, many specific details are set forth to facilitate a full understanding of the present disclosure, but the present disclosure may also be implemented in other ways different from those described herein; it is obvious that the embodiments in the specification are only part of the embodiments of the present disclosure, rather than all of the embodiments.

[0035] The terms "first", "second", etc. in the specification and claims of the present disclosure are used to distinguish similar objects, and are not used to describe a specific order or sequence. It should be understood that the data used in this way can be interchangeable under appropriate circumstances, so that the embodiments of the present disclosure can be implemented in an order other than those illustrated or described herein, and the objects distinguished by "first", "second", etc. are generally of the same type, and the number of objects is not limited. For example, the first object can be one or more. In addition, "and / or" in the specification and claims represents at least one of the connected objects, and the character " / " generally indicates that the objects related to each other are in an "or" relationship.

[0036] The electronic devices in the embodiments of the present disclosure may be mobile electronic devices or non-mobile electronic devices. Mobile electronic devices may include mobile phones, tablet computers, laptop computers, PDAs, in-vehicle electronic devices, wearable devices, ultra-mobile personal computers (UMPCs), netbooks, or personal digital assistants (PDAs); non-mobile electronic devices may include personal computers (PCs), televisions (TVs), ATMs, or self-service kiosks; and the embodiments of the present disclosure do not specifically limit these.

[0037] The execution subject of the video editing method provided by the embodiment of the present disclosure can be the above-mentioned electronic device (including mobile electronic devices and non-mobile electronic devices), or it can be a functional module and / or functional entity in the electronic device that can implement the video editing method. The specific execution subject can be determined according to actual usage requirements and is not limited by the embodiment of the present disclosure.

[0038] The video editing method provided by the embodiment of the present disclosure is described in detail below through specific embodiments and application scenarios in conjunction with the accompanying drawings.

[0039] As shown in FIG1 , an embodiment of the present disclosure provides a video editing method, which may include the following steps 101 to 105 .

[0040] 101. Obtain a first video material and a target video text corresponding to the first video material.

[0041] The target video copy is the text content corresponding to the audio information in the first video material. The target video copy can be existing subtitle text in the first video material, or it can be text content obtained by recognizing the audio information of the first video material. The specific target video copy can be determined based on actual conditions and is not limited here.

[0042] 102. Determine a target video editing template based on the target video copy.

[0043] The target video editing template is used to indicate a filling slot of at least one video segment and an editing effect applied to the at least one video segment.

[0044] Wherein, a target video editing template is determined from multiple video editing templates based on the target video copy.

[0045] 103. Capture at least one material segment from the first video material.

[0046] The at least one material clip corresponds to a filling slot of the at least one video clip and is used to fill the target video editing template to form the at least one video clip.

[0047] It can be understood that because the target video editing template is used to indicate the filling slot of at least one video clip, it is necessary to extract at least one material clip from the first video material, and the at least one material clip corresponds one-to-one to the filling slot of at least one video clip.

[0048] 104. Generate a second video material according to the at least one material clip and the target video editing template.

[0049] The second video material is used to present the editing effect applied to the at least one material segment.

[0050] It can be understood that at least one material clip is filled into the filling slot of at least one video clip to generate a second material video, or, according to the target video editing template, at least one material clip is spliced ​​and an editing effect is applied to at least one material clip to obtain a second video material.

[0051] 105. Generate a target video by splicing the second video material in front of the first video material.

[0052] In the embodiment of the present disclosure, a video editing template is designed, which is used to indicate the filling slots of at least one video clip and the editing effects applied to the at least one video clip, and then a target video text corresponding to the first video material is obtained; a target video editing template is determined based on the target video text, and then a second video material corresponding to the first video material is generated according to the target video editing template, and the target video is generated by splicing the second video material in front of the first video material, thereby realizing automatic video editing of the first video material. There is no need for the user to manually clip the first video material, nor is there any need for the user to manually splice the clipped video segments, nor is there any need for the user to manually add editing effects to the clipped video segments to generate the second video material. At the same time, there is no need for the user to manually splice the second video material in front of the first video material to generate the target video. Therefore, the disclosed solution simplifies the user's operation process for video editing of the first video material, reduces the time consumption of video editing, and improves the processing efficiency of video editing.

[0053] In some embodiments of the present disclosure, steps 101 to 105 can be used to generate a header for the video material to be processed (hereinafter referred to as header video editing processing), wherein the first video material is the video material to be processed, the second video material is the header of the first video material, and the target video is the first video material including the header.

[0054] In some embodiments of the present disclosure, steps 101 to 105 can be used to generate chapters of the video material to be processed (hereinafter referred to as chapter video editing processing), wherein the first video material is one of the multiple chapters included in the video material to be processed, the second video material is a chapter introduction of the first video material, and the target video is the video material to be processed including multiple chapters. In this case, in the above step 101, it is necessary to first divide the video material to be processed into chapters to obtain multiple chapters. In the above step 105, it is necessary to splice the multiple chapters with the added chapter introductions to obtain the target video.

[0055] In some embodiments of the present disclosure, the above step 101 can be specifically implemented through the following steps 101a to 101b, or the above step 101 can be specifically implemented through the following steps 101c to 101e; the above step 105 can be specifically implemented through the following steps 105a to 105b.

[0056] 101a. Obtain the target video text of the video material to be processed.

[0057] 101b. Based on the target video text, divide the to-be-processed video material into a plurality of video materials, each of the plurality of video materials being a first video material.

[0058] It can be understood that the target video copy is the video copy of the video material to be processed. Then, combined with step 102, for chapter video editing processing, the video editing template of the chapter is determined based on the video copy of the video material to be processed, that is, the video editing templates corresponding to multiple video materials are the same.

[0059] 101c. Obtain the video material text to be processed of the video material to be processed.

[0060] 101d. Based on the text of the video material to be processed, divide the video material to be processed into multiple video materials.

[0061] Each of the multiple video materials is a first video material.

[0062] 101e. Obtain the target video text of the first video material.

[0063] It can be understood that the target video copy is the video copy of the first video material, and then combined with step 102, for the chapter video editing processing, the video editing template of the chapter is determined based on the video copy of the first video material, that is, the video editing templates corresponding to multiple video materials are determined according to the video copy of each video material. Therefore, the video editing templates corresponding to multiple video materials may be the same or different, and can be determined specifically according to actual conditions, which is not limited here.

[0064] Each of the multiple video materials is a chapter of the video material to be processed. Based on the voice information of the target video text, the video material to be processed can be divided into chapters to obtain multiple video materials, each of which is a part of the video material to be processed that describes different things or events.

[0065] The plurality of video materials are each a first video material, and the above steps 102 to 104 are performed on each of the plurality of video materials to obtain a second video material corresponding to each video material.

[0066] 105a. Generate an edited video corresponding to the first video material by splicing the second video material in front of the first video material, so as to generate an edited video corresponding to each video material.

[0067] The edited video corresponding to the first video material is the video material after adding chapter introductions to the first video material.

[0068] 105b. Splice the edited videos corresponding to each video material to obtain the target video.

[0069] In the embodiment of the present disclosure, the target video is obtained by splicing the edited videos corresponding to the plurality of video materials, that is, the video material after adding chapters to the video material to be processed is obtained.

[0070] In the embodiment of the present disclosure, in the process of adding chapters to the video material to be processed, the same video editing template can be determined for multiple chapters based on the video copy of the video material to be processed, or video editing templates can be determined for different chapters based on the video copy of the chapters. The specific details can be determined based on actual conditions and are not limited here.

[0071] In some embodiments of the present disclosure, the target video editing template is further used to indicate the video title to fill the slot; the video editing method provided by the embodiment of the present disclosure may further include the following step 106, and the above step 104 may be specifically implemented by the following step 104a.

[0072] 106. Generate a video title for the first video material according to the video text of the first video material, where the video title is used to fill the video title filling slot.

[0073] The video title is used to summarize the main information expressed by the first video material.

[0074] 104a: Generate a second video material according to the at least one material clip, the video title, and the target video editing template.

[0075] The second video material is used to present the editing effect applied to the at least one material segment and the video title.

[0076] In the disclosed embodiment, the video editing template includes a video title filling slot, which can make the generated second video material include the information mainly expressed by the first video material, making it easier for users to understand what the first video material is trying to express.

[0077] In some embodiments of the present disclosure, the above step 102 can be specifically implemented through the following steps 102a and 102b.

[0078] 102a. Determine the category label of the category to which the target video copy belongs.

[0079] Among them, category tags may include video blogs, technology, fashion, film and television, animation, news, etc. If the video copy has no obvious category characteristics, the video copy will be labeled with a general tag.

[0080] It can be understood that the category to which the video copy belongs is the category to which the video material belongs.

[0081] In some embodiments of the present disclosure, step 102a can also be replaced by determining a category label for the category to which the first video material belongs. This can be accomplished by extracting a preset number of video frames from the first video material and then performing content recognition on the preset number of video frames to determine the category to which the first video material belongs, thereby obtaining a category label for the category to which the first video material belongs.

[0082] 102b. Determine a target video editing template corresponding to the category label from a plurality of video editing templates, where different category labels correspond to different video editing templates.

[0083] Among them, different category labels correspond to different video editing templates, and the target video editing template includes a video title filling slot and at least one video clip filling slot. The video title filling slot is used to indicate the video title of the first video material, and each video clip filling slot is used to indicate a video clip of the first video material.

[0084] It can be understood that multiple correspondences are stored in the electronic device, each correspondence is used to indicate a category label and at least one video editing template corresponding to the category label; each correspondence can also be used to indicate a video editing template and a category label corresponding to the video editing template.

[0085] In some embodiments of the present disclosure, a category label may correspond to at least one video editing template, and a video editing template may also correspond to one or more category labels.

[0086] The multiple video editing templates in the above step 102b correspond to different category labels respectively. If there is only one video editing template corresponding to the category label among the multiple video editing templates, then the one video editing template is the target video editing template. If there are multiple editing templates corresponding to the category label among the multiple video editing templates, then you can arbitrarily select one video editing template from the multiple editing templates as the target video editing template; you can also display multiple template identifiers corresponding to the multiple editing templates to the user (which can also include an introduction to the differences between different template identifiers), and then determine the video editing template corresponding to the template identifier selected by the user from the multiple template identifiers as the target video editing template; the specific determination can be based on actual conditions and is not limited here.

[0087] In the disclosed embodiment, a video editing template is determined based on the category label of the category to which the video copy belongs, which can increase the diversity of video editing and improve user experience.

[0088] It can be understood that when the video editing process is a title video editing process, the target video editing template is a title video editing template, the title video editing template includes a title slot and a filling slot for at least one video clip, the video title of the first video material is the title title, the above step 104 is specifically based on the title video editing template, the title is filled into the title slot, and at least one material clip is filled into the filling slot of the at least one video clip to generate the target title, the above step 105 is specifically to synthesize the target title and the first video material into the target video.

[0089] It can be understood that when the video editing process is a chapter video editing process, the target video editing template is a chapter video editing template. The chapter video editing template includes a chapter title slot and a filling slot for at least one video clip. The video title of the first video material is the chapter title. The above step 104 is specifically based on the chapter video editing template, filling the chapter title into the chapter title slot, and filling at least one material clip into the filling slot of the at least one video clip to generate a target chapter introduction. The above step 105 is specifically splicing the chapter introduction before the first video material to obtain the edited video corresponding to the first video material (a chapter including a chapter introduction), and then splicing the edited videos corresponding to each video material to obtain the target video.

[0090] In the disclosed embodiment, a target video editing template that matches the category label is determined by using the category label of the video text of the video material to be processed, and then at least one material segment required by the target video editing template is extracted from the first video material, and the video title of the first video material required by the target video editing template is obtained based on the video text of the first video material, and then the video title and at least one material segment of the first video material are filled into the target video editing template to generate a second video material corresponding to the first video material, so as to synthesize the target video based on the second video material and the first video material. In this way, by setting different video editing templates for different categories of video materials to be processed, and then performing video editing on the video materials to be processed according to the video editing template that matches the category to which the video material to be processed belongs, automatic video editing of the video materials to be processed is achieved, which can improve the video editing efficiency of the video materials to be processed.

[0091] In some embodiments of the present disclosure, the filling slot of each video segment is also used to indicate the duration of the corresponding video segment; the above-mentioned step 103 can be specifically carried out through the following steps 103a to 103f.

[0092] 103a. Slice the first video material to obtain at least one slice.

[0093] In some embodiments of the present disclosure, the first video material may be sliced ​​based on the continuity of the video image. For example, the first video material may be sliced ​​using the locations where scene changes or shot changes are identified in the first video material as slicing points to obtain at least one slice.

[0094] In some embodiments of the present disclosure, the first video material may also be sliced ​​according to a preset duration (which may be determined based on actual conditions and is not limited here).

[0095] 103b. Perform highlight segment extraction processing on the at least one slice to obtain at least one highlight segment.

[0096] The at least one slice corresponds to the at least one highlight segment in a one-to-one manner.

[0097] In the disclosed embodiment, the specific process of extracting highlight segments is not limited and can be determined based on actual conditions.

[0098] Exemplarily, highlight segments may be extracted from slices using a related highlight segment extraction method, or using a highlight segment extraction model.

[0099] In some embodiments of the present application, extracting highlight segments for a slice (hereinafter referred to as a target slice) may be as follows: scoring each frame in the target slice, determining the frame with the highest score as the highlight frame, and then extending based on this highlight frame (extracting a continuous preset number of video frames forward and / or backward) to obtain the highlight segment corresponding to the target slice.

[0100] Among them, the scoring will be based on the composition, color, aesthetics and other dimensions of the video frame.

[0101] For example, if a 1-second highlight clip is needed, based on the highlight frame with the highest score, 0.5-second video frames can be taken forward and backward to obtain a 1-second highlight clip.

[0102] In some embodiments of the present disclosure, a highlight clip is a video clip that includes a highlight event. A highlight event is a special plot such as a person, thing, action, story and / or sound effect in a video that attracts the audience. Different highlight events can be defined in different types of videos. For example, in a football game video, goals, tackles and saves are defined as highlight events; in a battle-type operation game video, the release actions of various skills are defined as highlight events; compared with other relatively bland parts of the video, highlight events can arouse people's interest and attention.

[0103] In some embodiments of the present disclosure, a highlight segment refers to a segment of video material to be processed that includes a highlight event and has relatively bright colors and high picture quality.

[0104] Exemplarily, extracting a highlight segment for a slice (hereinafter referred to as a target slice) can be as follows: according to the category label of the category to which the first video material belongs, determining multiple highlight events that match the category label, and determining the segment of the highlight event among the multiple highlight events identified in the target slice as the highlight segment.

[0105] 103c. Determine N first material segments from the at least one highlight segment.

[0106] The duration of each first material segment is the duration of a video segment, each first material segment corresponds to a highlight segment, and N is a natural number;

[0107] Each first material segment is a highlight segment of the at least one highlight segment, which is obtained by processing according to the duration of a video segment, and N is a natural number.

[0108] It can be understood that the N first material clips are recorded as N of the at least one material clip.

[0109] In some embodiments of the present disclosure, each highlight segment in at least one highlight segment is scored. When determining N first material segments from the at least one highlight segment, the first material segment is preferentially selected from the highlight segments with the highest scores and those whose durations match the durations of the corresponding video segments. Highlight segments whose durations are shorter than the durations of any video segments are discarded, and from among the highlight segments whose durations are equal to or greater than the durations of the corresponding video segments, the highest-scoring highlight segments that match the durations of the corresponding video segments are selected as the first material segments.

[0110] In Example 1, the process for obtaining a first clip is to select, from among at least one highlight clip, a highlight clip with a duration greater than or equal to the corresponding duration and the highest score, excluding highlight clips already obtained from the first clip, to obtain the first clip. For example, the at least one highlight clip includes three highlight clips: 10s highlight clip 1, 15s highlight clip 2, 8s highlight clip 3, and 6s highlight clip 4. The scores of these four highlight clips are, from highest to lowest, highlight clip 3, highlight clip 2, highlight clip 4, and highlight clip 1. At least one video clip (hereinafter referred to as M video clips) includes clip a with a duration of 10s, clip b with a duration of 6s, and clip c with a duration of 8s. Clip a is first extracted from highlight clip 2, clip b is then extracted from highlight clip 3, and finally clip c is extracted from highlight clip 1. The extracted clips a, b, and c constitute the three first clips.

[0111] 103d. When N is less than M, determine MN second material segments from the first video material.

[0112] The duration of each second material segment is the duration of a video segment other than the video segments corresponding to the N first material segments, M is the number of the at least one material segment, and M is a positive integer.

[0113] Each second material segment is the first video material, and is obtained by processing according to the duration of a video segment among the at least one video segment, excluding the duration of the video segment corresponding to the N first material segments.

[0114] It can be understood that the MN second material clips are the remaining MN material clips in the at least one material clip (excluding the N first material clips).

[0115] 103e. Determine the N first material clips and the MN second material clips as the at least one material clip.

[0116] The N first material segments extracted above are the optimal material segments selected from the first video material. When N is less than M, it is necessary to select suboptimal material segments from the first video material again.

[0117] Example 2, for the above steps 103d and 103e, continuing from the above example 1, M is 5, and at least one video clip (M video clips) also includes a material clip d with a duration of 9 seconds, and a material clip e with a duration of 12 seconds. Then it is necessary to determine two second material clips from the first video material as material clip d and material clip e, one of which has a duration of 9 seconds and the other has a duration of 12 seconds. Highlight clip 4 (duration of 6 seconds) does not meet the conditions for extracting the second material clip, so it is necessary to determine two second material clips from the first video material. Then the three first material clips and two second material clips extracted above are the five material clips required by the target video editing template.

[0118] 103f. When N is equal to M, determine the N first material clips as the at least one material clip.

[0119] Example 3, continuing from the above Example 1, M is 3, and the three first material clips extracted above are the three material clips required by the target video editing template.

[0120] In the embodiment of the present disclosure, through the above steps 103a to 103f, based on the length of each video clip, at least one material clip (at least one video clip) that meets the requirements of the target video editing template can be obtained, and then a second video material can be generated based on the target video editing template and at least one material clip, thereby obtaining a target video that meets the requirements, which can improve the processing efficiency of video editing.

[0121] In some embodiments of the present disclosure, priority is given to ensuring that each second material segment does not duplicate any first material segment, and that each second material segment does not duplicate each other. If this cannot be guaranteed, efforts are made to ensure that each second material segment is not completely identical to any first material segment, and that each second material segment is not completely identical to each other. For example, step 103d can be implemented through steps 103d1 to 103d3.

[0122] 103d1. When the video duration of the remaining videos in the first video material except the N first material segments is greater than the target duration, determine the MN second material segments from the remaining videos.

[0123] The target duration is the sum of the durations of the video segments in the at least one video segment excluding the video segments corresponding to the N first material segments. In other words, the target duration is the sum of the durations of the M video segments in the M material segments excluding the video segments corresponding to the N first material segments.

[0124] 103d2. When the video duration of the remaining video is less than or equal to the target duration, determine K second material segments from the remaining video.

[0125] Wherein, K is a natural number less than or equal to MN. K can be equal to 0 or a positive integer less than or equal to MN.

[0126] 103d3. Determine MNK second material segments from the first video material.

[0127] In the embodiment of the present disclosure, the logic for determining MN second material segments is to preferentially select suboptimal video segments from the remaining videos of the first video material excluding the N first material segments as the second material segments. When the duration of the remaining video is greater than the target duration, MN second material segments are selected from the remaining video, and the obtained at least one video segment (at least one material segment) does not repeat with each other. When the duration of the remaining video is less than or equal to the target duration, K second material segments are preferentially determined from the remaining video, and there is no duplication between the K second material segments, and the K second material segments do not repeat with the above-mentioned N first material segments, and then MNK second material segments are determined from the first video material, and each of the MNK second material segments may overlap with one or more material segments among the N first material segments and the K second material segments.

[0128] For example, assuming that the length of the first video material is 20 seconds, the first video material needs to be edited for the title video, and the matched title video editing template requires 4 (M) video clips, each of which is 6 seconds long. The first video material is sliced ​​to obtain 2 slices, and 2 (N) first material clips (6s*2=12s) are obtained based on the 2 slices. Then, the remaining video of the first video material except the 2 first material clips is extracted to obtain 1 (K) second material clip (6s). Because the remaining video of the first video material except the 2 first material clips and 1 second material clip that have been obtained is only 2 seconds, it is not enough to extract another second material clip. Therefore, it is necessary to extract 1 (MNK) second material clip from the first video material, which has repeated content with at least one of the 2 first material clips and 1 second material clip extracted previously, so that the title video editing template requires 4 video clips.

[0129] In the embodiment of the present disclosure, by extracting MN second material segments according to the method provided in the above steps 103d1 to 103d3, second material segments that better meet the needs can be obtained. Finally, the second video material synthesized based on at least one video segment (i.e., N first material segments and MN second material segments) has the least number of repeated frames, which can improve the user experience.

[0130] In some embodiments of the present disclosure, the filling slot of each video segment is also used to indicate the arrangement order of the corresponding video segments. The above step 104 can be specifically implemented through the following step 104b.

[0131] 104b. Based on the target video editing template, synthesize the at least one material segment according to the arrangement order of the corresponding video segments to generate a second video material.

[0132] In the disclosed embodiment, if the fill slots of at least one video clip in the target video editing template are ordered, the arrangement order is determined based on the order of the fill slots of the at least one video clip, and the order of extracting the at least one material clip is also determined based on the order of the fill slots of the at least one video clip. By filling the fill slots of the at least one video clip with the at least one material clip according to the arrangement order of the corresponding material clips, a second video material that better meets the requirements can be obtained.

[0133] In some embodiments of the present disclosure, if there is no order in the filling slots of each video clip, the at least one clip and the video title of the first video material can be synthesized into the second video material according to the timestamp of the first frame of each clip in the at least one extracted clip.

[0134] In some embodiments of the present disclosure, the above step 101 may further include the following step 101f, and after the above step 105, the following step 107 may further be included.

[0135] 101f. In response to a triggering operation on a video editing control in a video editing panel, obtaining a first video material and a target video text corresponding to the first video material to enter a video editing process;

[0136] 107. Display the target video.

[0137] In the disclosed embodiment, the target video is displayed to facilitate users to view the video editing effect.

[0138] In some embodiments of the present disclosure, before the above 101f, the video editing method provided by the embodiment of the present application may further include the following step 108.

[0139] 108. In response to a triggering operation on a video processing control displayed on the video editing interface, display a video editing panel.

[0140] Among them, the video editing panel displays video editing controls, which include header editing controls or chapter editing controls. The header editing controls are used to indicate header video editing processing, and the chapter editing controls are used to indicate chapter video editing processing.

[0141] In some embodiments of the present disclosure, the video editing panel may display only the title editing control, or only the chapter editing control, or both the title editing control and the chapter editing control at the same time. Other editing controls may also be displayed in the video editing panel. The specific details may be determined based on actual conditions and are not limited here.

[0142] In some embodiments of the present disclosure, one or more video editing controls may be displayed in a video editing panel. The video editing panel may display one or more title editing controls, or one or more chapter editing controls. The specific details may be determined based on actual circumstances and are not limited here. Different title editing controls correspond to different logic for editing the title video. Different chapter editing controls correspond to different logic for editing the chapter video.

[0143] In some embodiments of the present disclosure, a video editing control (a header editing control and / or a chapter editing control) may be a control with a switch. When the video editing function is used for the first time, the switch of the video editing control (hereinafter referred to as the video editing switch) is turned off by default. When the video editing function is used subsequently, it is displayed according to the recorded state of the video editing switch. That is, if the last recorded state of the video editing switch is on, it is turned on by default this time. If the last recorded state of the video editing switch is off, it is turned off by default this time.

[0144] In some embodiments of the present disclosure, the video editing controls can also be set according to whether the video editing interface is displayed in landscape or portrait mode. For example, when the video editing interface is displayed in portrait mode, the video editing controls are grayed out (i.e., video editing cannot be performed on the video material to be processed based on the video editing controls). At this time, when you click the switch, a pop-up window will pop up prompting "Currently only supports landscape video editing."

[0145] It can be understood that if the video editing control is a header editing control, then in response to the triggering operation of the header editing control, based on the header video editing processing indicated by the header editing control, the video material to be processed is edited to generate a target video, which includes a synthesized header mixed-cut clip.

[0146] It can be understood that if the video editing control is a chapter editing control, then in response to the triggering operation of the chapter editing control, based on the chapter video editing processing indicated by the chapter editing control, video editing processing is performed on the video material to be processed to generate a target video. The target video includes multiple chapters, and each chapter includes a synthetic chapter mixed clip.

[0147] In this way, this solution can trigger automatic video editing of the video material to be processed by displaying the video processing controls on the video editing interface, and then triggering the display of the video editing panel through the video processing controls, and by triggering the video editing controls displayed in the video editing panel. There is no need for users to select and match suitable materials in the huge material library, and make style adjustments, etc. to realize video editing of the video material to be processed. It simplifies the user's operation process for video editing of the video material to be processed (header video editing or chapter video editing), reduces the time consumption of video editing of the video material to be processed, and improves the processing efficiency of video editing of the video material to be processed.

[0148] In some embodiments of the present disclosure, a commercial material filtering option is also displayed in the video editing panel, which is not selected by default (i.e., the commercial material filtering function is turned off by default). After the commercial material filtering function is turned on, the video editor only returns commercially available materials.

[0149] In some embodiments of the present disclosure, when the video editing panel is displayed for the first time, the commercial material screening option is unselected by default. When the video editing panel is displayed subsequently, it is displayed according to the status of the last recorded commercial material screening option. That is, if the status of the last recorded commercial material screening option is unselected, it is unselected by default this time; if the status of the last recorded commercial material screening option is selected, it is selected by default this time.

[0150] In some embodiments of the present disclosure, the video editing panel may also display text editing controls (for video editing of text in subtitles, picture-in-picture, etc.) and sound effect video editing controls (when the subtitles include keywords that match sound effects, corresponding sound effects are added when the keywords are displayed). The specific details can be determined based on actual conditions and are not limited here.

[0151] In some embodiments of the present disclosure, when a text editing control is displayed in the video editing panel, a Delete Existing Subtitles option is also displayed for the text editing control in the video editing panel (selected by default, i.e., the Delete Existing Subtitles function is enabled by default). The Delete Existing Subtitles option in a selected state is used to indicate that the existing subtitles of the video material to be processed are deleted before text editing is performed on the video material to be processed, and the Delete Existing Subtitles option in an unselected state is used to indicate that the existing subtitles of the video material to be processed are used for video editing when text editing is performed on the video material to be processed.

[0152] In some embodiments of the present disclosure, the video editing panel also displays a function introduction icon corresponding to each video editing control. The function introduction icon is used to trigger the display of a function introduction to the video editing mode indicated by the corresponding video editing control. The function introduction includes a text-based introduction and / or an animated graphic-based introduction. In this way, if the user does not understand the video editing mode corresponding to each video editing control, the function introduction to the video editing mode indicated by the video editing control can be triggered by triggering the function introduction icon corresponding to the video editing control, thereby facilitating user understanding and use and improving user experience.

[0153] In some embodiments of the present disclosure, a function introduction interface corresponding to the function introduction icon is displayed in a floating manner in the video editing panel. The function introduction interface can be a pop-up window display or the following function introduction floating layer, which is not limited here.

[0154] In some embodiments of the present disclosure, the above step 108 can be specifically implemented through the following steps 108a and 108b.

[0155] 108a. In response to a triggering operation on a video processing control displayed on the video editing interface, a function introduction floating layer is displayed.

[0156] Among them, the function introduction floating layer is used to display the entry control of the video editing panel and the function introduction of the video editing process indicated by the video editing control. The function introduction includes a text introduction and / or an animated image introduction.

[0157] In some embodiments of the present disclosure, when the video editing panel includes multiple video editing controls, the function introduction corresponding to each video editing control in the function introduction floating layer can be displayed separately (the function introduction corresponding to different video editing controls can be switched through a switching operation), and the function introduction corresponding to each video editing control in the function introduction floating layer can also be displayed at the same time. The specific details can be determined based on actual conditions and are not limited here.

[0158] Among them, the specific function introduction can be set according to the corresponding video editing control and is not limited here.

[0159] Exemplarily, the function introduction floating layer is a half-screen function introduction floating layer. Specifically, an animated image introduction corresponding to a video editing control is displayed above the function introduction floating layer, and a text introduction corresponding to the video editing control is displayed below. Users can switch the animated image introductions corresponding to different video editing controls by sliding left and right. After the animated image introduction is switched, the text introduction below is synchronously updated.

[0160] 108b. In response to a triggering operation on the entry control, display the video editing panel.

[0161] Exemplarily, the entry control may be a “try it” control.

[0162] In some embodiments of the present disclosure, it can be arranged that when a new user (after downloading and installing) uses the video editing function for the first time, the above-mentioned step 108 includes the above-mentioned step 108a and step 108b. When an old user uses the video editing function, the video editing panel is directly displayed by triggering the video processing control, and the above-mentioned function introduction floating layer is not displayed.

[0163] In the embodiment of the present disclosure, before displaying the video editing panel, displaying a function introduction floating layer helps the user understand and compare the video editing processing corresponding to each video editing control, and then facilitates the subsequent selection of which video editing control to use for video editing processing of the video material to be processed, which can improve the user experience.

[0164] In some embodiments of the present disclosure, the above step 101f can be specifically carried out through the following steps 201 to 205.

[0165] 201. In response to a triggering operation on the video editing control, determine whether a video editing function has been authorized.

[0166] If it is determined that the video editing function is not authorized, the following steps 202 to 205 are executed; if it is determined that the video editing function is authorized, the following step 204 is executed.

[0167] 202. If the video editing function is not authorized, an authorization pop-up window is displayed, in which an authorization option and a non-authorization option are displayed.

[0168] The authorization option is used to indicate that the video editing function is authorized, and the non-authorization option is used to indicate that the video editing function is not authorized.

[0169] Exemplarily, the authorization pop-up window may include text-based introduction information about the video editing function, as well as an "Allow" option and a "Do Not Allow" option, wherein the "Allow" option corresponds to the authorization option, and the "Do Not Allow" option corresponds to the non-authorization option.

[0170] 203. In response to a triggering operation on the authorization option, determine to authorize the video editing function.

[0171] In response to the triggering operation of the authorization option, the video editing function is authorized for the electronic device, and the following step 204 is executed.

[0172] 204. If the video editing function is authorized, perform video editing on the video material to be processed based on the video editing control to generate the target video.

[0173] 205. In response to the triggering operation of the unauthorized option, prohibiting video editing processing of the video material to be processed.

[0174] It can be understood that in response to the triggering operation of the unauthorized option, the electronic device may not perform video editing on the video material to be processed, and may also display a prompt message, which is used to prompt that the video editing function cannot be used without authorization.

[0175] It is understood that when the video editing control is clicked, if it is detected that the electronic device is not authorized for the video editing function, an authorization pop-up window will pop up. Select the authorization option to start video editing of the video material to be processed. Select the unauthorized option to close the authorization pop-up window and not start (prohibit) video editing of the video material to be processed.

[0176] In some embodiments of the present disclosure, if the authorization option is selected this time, the authorization pop-up window will not pop up the next time the video editing function is used; if the non-authorization option is selected this time, the authorization pop-up window will still pop up the next time the video editing function is used, and the user can only use the video editing function after selecting the authorization option.

[0177] In the embodiment of the present disclosure, by setting the video editing function to be allowed when authorized, and prohibited when not authorized, the user's safety can be guaranteed, user misoperation can be avoided, and the user experience can be improved.

[0178] In some embodiments of the present disclosure, it is also possible to determine whether the video editing function has been authorized in response to a triggering operation of a video processing control. If the video editing function has not been authorized, an authorization pop-up window is displayed, in which an authorization option and a non-authorization option are displayed. If the video editing function has been authorized, a video editing panel is displayed. The specific details can be determined based on actual conditions and are not limited here.

[0179] In some embodiments of the present disclosure, the video editing method provided by the embodiments of the present disclosure may further include at least one of the following steps 109 to 111.

[0180] 109. In response to the triggering operation of the cancel control, the video editing process is stopped and the performed video editing results are cleared.

[0181] It can be understood that, in the process of performing video editing on the video material to be processed, in response to the triggering operation of the cancel control, the process of performing video editing on the video material to be processed is stopped, and the performed video editing results are cleared.

[0182] The cancel control is used to end the video editing process.

[0183] It is understood that the video editing method provided by the embodiment of the present disclosure supports terminating the video editing process during the video editing process. In this way, when the user suddenly does not want to continue the video editing process, the video editing process can be terminated by triggering the cancel control, which can improve the user experience.

[0184] 110. In response to a triggering operation on a target control, the video editing process is switched to background execution, and the progress of the video editing process is displayed at a target location.

[0185] It can be understood that in the process of video editing the video material to be processed, in response to the triggering operation of the target control, the process of video editing the video material to be processed will be switched to background execution, and the process progress of video editing the video material to be processed will be displayed at a preset position.

[0186] The process progress may be a progress percentage, or may be process description information of the current processing stage in the video editing process.

[0187] The target control is a control for switching the video editing process to background execution. For example, the target control can be a control for switching to the tool homepage of an application with video editing functions, a control for switching to the homepage of an electronic device, or other types of controls. The specific target control can be determined based on actual conditions and is not limited here.

[0188] It is understood that the video editing method provided by the embodiment of the present disclosure supports switching the video editing process to the background during the video editing process. For example, during the video editing process, by triggering the operation of exiting to the tool homepage, the tool homepage is displayed, and the video editing process is kept running in the background.

[0189] Among them, the preset position can be the draft cover of the tool homepage, or it can be a hanging window control displayed in an application with video editing function, or it can be a hanging window control displayed on an electronic device, or it can be other locations or areas. The specific location can be determined according to actual conditions and is not limited here.

[0190] Exemplarily, in response to a trigger operation of exiting to the tool homepage, the video editing process is kept running in the background, and the progress is displayed on the draft cover of the tool homepage.

[0191] The disclosed embodiment supports switching the video editing process to background execution, so that the user can also perform other operations during the video editing process of the video material to be processed, which can improve the user experience.

[0192] 111. During the process of video editing the video material to be processed, a loading animation is displayed, where the loading animation includes process description information of the video editing process for the video material to be processed.

[0193] In some embodiments of the present disclosure, the loading animation may be scrolling the process description information once every preset time period (such as 5 seconds) (ie, switching to display the process description information).

[0194] In some embodiments of the present disclosure, the loading animation may also display process description information in real time according to the actual video editing process, that is, the process description information of the current stage of the video editing process is displayed.

[0195] Exemplarily, the process description information can be "Acquiring video text...", "Determining the category label of the category to which the video belongs...", "Matching the video editing template corresponding to the category label...", "Extracting highlight clips...", etc. The specific process description information can be determined based on actual conditions and is not limited here.

[0196] In the disclosed embodiment, during the video editing process, displaying a loading animation can help users understand the general process of the video editing process and improve the user experience.

[0197] In some embodiments of the present disclosure, the video editing method provided by the embodiments of the present disclosure may further include at least one of the following steps 112 and 113.

[0198] 112. Display target prompt information.

[0199] Among them, when the video editing processing of the video material to be processed is successful, the target prompt information is used to prompt that the video editing processing of the video material to be processed is successful; when the video editing processing of the video material to be processed fails, the target prompt information is used to prompt that the video editing processing of the video material to be processed fails, and / or the reason for the failure of video editing processing of the video material to be processed.

[0200] In the embodiment of the present disclosure, after the video material to be processed is successfully edited (the material is added successfully), a prompt message "Video editing successful" or "Video editing processing successful" may be popped up.

[0201] In the embodiment of the present disclosure, after the video editing process of the video material to be processed fails (the adding of the material fails), a prompt message of "Video editing failed" or "Video editing process failed" may be popped up.

[0202] In some embodiments of the present disclosure, in the event that video editing fails, the reason for the failure may also be prompted. For example, if the second video material (title or chapter clip) cannot be returned due to the lack of audio content, a pop-up window may prompt "No speech content was recognized, and the title and chapter cannot be added for the time being"; if the recognition fails due to a network problem, a pop-up window may prompt "Network error, please try again."

[0203] In the disclosed embodiment, by displaying target prompt information, the user can be promptly prompted with the result of the video editing process, which facilitates the user to perform subsequent operations and improves the user experience.

[0204] 113. Display target status information.

[0205] Among them, when the video editing processing of the video material to be processed is successful, the target status information is used to indicate that the video editing processing of the video material to be processed is successful; when the video editing processing of the video material to be processed fails, the target status information is used to indicate that the video editing processing of the video material to be processed fails.

[0206] The disclosed embodiment supports displaying target status information corresponding to the success or failure of video editing processing at a preset position (refer to the above description of the preset position, which will not be repeated here). In this way, the user can be prompted whether the video editing processing of the video material to be processed is successful or failed. Furthermore, when the user cannot pay attention to the video editing process of the video material to be processed in real time, the user can understand the result of the video editing processing of the video material to be processed based on the target status information, which can improve the user experience.

[0207] In the disclosed embodiments, target state information may be displayed at a preset location for a certain period of time. For example, the target state information may be displayed at the preset location for the lifecycle of the corresponding application and may no longer be displayed after the application is restarted. The target state information may also be undisplayed after entering the timeline editing panel.

[0208] In the embodiment of the present disclosure, the target status information indicating the failure of video editing processing can also be cancelled after the user triggers the process of re-editing the video material to be processed through a trigger operation (such as clicking a retry control, clicking a redo control, or clicking any video editing processing control).

[0209] In some embodiments of the present disclosure, after the above step 113, the video editing method provided by the embodiments of the present disclosure may further include any one of the following steps 114 to 117.

[0210] 114. When it is detected that the application corresponding to the video editing function is closed, clear the target status information.

[0211] 115. When detecting that the timeline editing panel of the video editing interface is displayed, clear the target state information and position the timeline preview axis to the first position on the timeline where video editing has been performed.

[0212] Among them, positioning the timeline preview axis to the first position on the timeline where video editing has been performed can facilitate the user to determine at which position the video editing process was started, thereby improving the user experience.

[0213] 116. When it is detected that video editing is to be performed again on the video material to be processed, the target state information and the target video are cleared.

[0214] Among them, the video editing process is performed again on the video material to be processed. This can be done by using the video editor again after the video editor successfully adds the material; or by clicking the redo control, which is not limited here.

[0215] Clearing the target status information and the target video automatically clears the results of the previous video editing process, effectively overwriting the old results with the new results. This prevents the results of the two video editing processes from being mixed together, resulting in the final video editing result containing different editing styles and degrading the user experience.

[0216] 117. When it is detected that the video editing process performed on the video material to be processed is canceled, the target video is cleared.

[0217] The embodiment of the present disclosure supports the operation of undoing the video editing process that was previously performed.

[0218] It can be understood that adding materials using video editing is a whole operation step. When the user clicks the undo control, all video editing materials generated by the previous video editing process must be completely removed.

[0219] In some embodiments of the present disclosure, a free icon can be displayed at the video processing control according to actual needs, and the function can be used normally after clicking it.

[0220] For example, as shown in FIG2 , the area indicated by the mark “21” is the area for displaying the video material to be processed, the area indicated by the mark “22” is the timeline editing panel, and the area indicated by the mark “23” is the toolbar area for editing processing, wherein the toolbar area includes a “Video Processing” control. Clicking the “Video Processing” control shown in FIG2 displays a function introduction floating layer as indicated by the mark “31” in FIG3 , wherein the function introduction floating layer includes an introduction area in the form of an animated image for the video editing control 1 as indicated by the mark “311”, an introduction area in the form of a text for the video editing control 1 as indicated by the mark “312”, and a “Try It Out” control as indicated by the mark “313”. Click the "Try It" control shown in Figure 3 to display the video editing panel indicated by the mark "41" in Figure 4. The video editing panel includes the title editing control 1 (the switch of the current title editing control 1 is on, that is, the title editing control 1 is selected), the function introduction icon 1 corresponding to the title editing control 1, the chapter editing control 2 (the switch of the current chapter editing control 2 is off, that is, the chapter editing control 2 is not selected), the function introduction icon 2 corresponding to the chapter editing control 2, and the "Start" control. Click the "Start" control shown in Figure 4 to display the video editing loading animation area indicated by the mark "51" in Figure 5. At the same time, a pop-up window is displayed to prompt "Video editing still needs to run for a while, do you need to return to the tool homepage and run it in the background?" The user can choose "Run in the background" or "Wait" according to the prompt information.

[0221] Figure 6 is a structural block diagram of a video editing device shown in an embodiment of the present disclosure. As shown in Figure 6, it includes: an acquisition module 601, used to acquire a first video material and a target video copy corresponding to the first video material; a determination module 602, used to determine a target video editing template based on the target video copy, and the target video editing template is used to indicate the filling slot of at least one video clip and the editing effect applied to the at least one video clip; an interception module 603, used to intercept at least one material clip in the first video material; the at least one material clip, corresponding to the filling slot of the at least one video clip, is used to fill the target video editing template to form the at least one video clip; a generation module 604, used to generate a second video material based on the at least one material clip and the target video editing template; wherein the second video material is used to present the editing effect applied to the at least one material clip; and the target video is generated by splicing the second video material in front of the first video material.

[0222] In some embodiments of the present disclosure, the acquisition module 601 is specifically used to obtain the target video copy of the video material to be processed; based on the target video copy, the video material to be processed is divided into multiple video materials, each of the multiple video materials is a first video material; or, the acquisition of the first video material and the video copy corresponding to the first video material includes: obtaining the video material copy of the video material to be processed; based on the video material copy, the video material to be processed is divided into multiple video materials, each of the multiple video materials is a first video material; obtaining the target video copy of the first video material; the generation module 604 is specifically used to generate an edited video corresponding to the first video material by splicing the second video material in front of the first video material, so as to generate an edited video corresponding to each video material; and splicing the edited videos corresponding to each video material to obtain the target video.

[0223] In some embodiments of the present disclosure, the target video editing template is also used to indicate a video title to fill a slot; the generation module 604 is also used to generate a video title for the first video material based on the video text of the first video material, and the video title is used to fill the video title filling slot; the generation module 604 is specifically used to generate a second video material based on the at least one material clip, the video title and the target video editing template; wherein the second video material is used to present the editing effect applied to the at least one material clip and the video title.

[0224] In some embodiments of the present disclosure, the determination module 602 is specifically used to determine the category label of the category to which the target video copy belongs; and determine the target video editing template corresponding to the category label from multiple video editing templates, where different category labels correspond to different video editing templates.

[0225] In some embodiments of the present disclosure, the filling slot of each video clip is also used to indicate the duration of the corresponding video clip; the interception module 603 is specifically used to slice the first video material to obtain at least one slice; perform highlight clip extraction processing on the at least one slice to obtain at least one highlight clip, and the at least one slice corresponds to the at least one highlight clip one-to-one; from the at least one highlight clip, determine N first material clips, the duration of each first material clip is the duration of a video clip, each first material clip corresponds to a highlight clip, and N is a natural number; when N is less than M, determine MN second material clips from the first video material, the duration of each second material clip is the duration of a video clip other than the video clip corresponding to the N first material clips, M is the number of the at least one material clip, and M is a positive integer; the N first material clips and the MN second material clips are determined as the at least one material clip; when N is equal to M, determine the N first material clips as the at least one material clip.

[0226] In some embodiments of the present disclosure, the interception module 603 is specifically used to determine the MN second material segments from the remaining video in the first video material, when the video duration of the remaining video other than the N first material segments in the first video material is greater than the target duration, and the target duration is the sum of the durations of the video segments in the at least one video segment other than the video segments corresponding to the N first material segments; determine K second material segments from the remaining video, when the video duration of the remaining video is less than or equal to the target duration, and K is a natural number less than or equal to MN; and determine MNK second material segments from the first video material.

[0227] In some embodiments of the present disclosure, the filling slot of each video clip is also used to indicate the arrangement order of the corresponding video clips. The generation module 604 is specifically used to synthesize the at least one material clip according to the arrangement order of the corresponding video clips based on the target video editing template to generate a second video material.

[0228] In the embodiments of the present disclosure, each module can implement the video editing method provided by the above method embodiments and can achieve the same technical effect. To avoid repetition, it will not be described here.

[0229] FIG7 is a schematic structural diagram of an electronic device provided in an embodiment of the present disclosure, which is used to exemplify an electronic device that implements any video editing method in an embodiment of the present disclosure and should not be understood as a specific limitation on the embodiment of the present disclosure.

[0230] As shown in Figure 7, electronic device 700 may include a processor (e.g., a central processing unit, a graphics processing unit, etc.) 701, which can perform various appropriate actions and processes according to a program stored in a read-only memory (ROM) 702 or a program loaded from a storage device 708 into a random access memory (RAM) 703. Various programs and data required for the operation of electronic device 700 are also stored in RAM 703. Processor 701, ROM 702, and RAM 703 are connected to each other via a bus 704. An input / output (I / O) interface 705 is also connected to bus 704.

[0231] Typically, the following devices may be connected to the I / O interface 705: an input device 706 including, for example, a touch screen, a touchpad, a keyboard, a mouse, a camera, a microphone, an accelerometer, a gyroscope, etc.; an output device 707 including, for example, a liquid crystal display (LCD), a speaker, a vibrator, etc.; a storage device 708 including, for example, a magnetic tape, a hard disk, etc.; and a communication device 709. The communication device 709 may allow the electronic device 700 to communicate with other devices wirelessly or by wire to exchange data. Although the electronic device 700 is shown as having various devices, it should be understood that it is not required to implement or have all of the devices shown. More or fewer devices may be implemented or have alternatively.

[0232] In particular, according to an embodiment of the present disclosure, the process described above with reference to the flowchart can be implemented as a computer software program. For example, an embodiment of the present disclosure includes a computer program product, which includes a computer program carried on a non-transitory computer-readable medium, and the computer program includes a program code for executing the method shown in the flowchart. In such an embodiment, the computer program can be downloaded and installed from the network through the communication device 709, or installed from the storage device 708, or installed from the ROM 702. When the computer program is executed by the processor 701, the functions defined in any video editing method provided by the embodiment of the present disclosure can be executed.

[0233] It should be noted that the computer-readable medium mentioned above in the present disclosure may be a computer-readable signal medium or a computer-readable storage medium, or any combination of the two. Computer-readable storage media may be, for example, but not limited to, electrical, magnetic, optical, electromagnetic, infrared, or semiconductor systems, devices, or components, or any combination of the above. More specific examples of computer-readable storage media may include, but are not limited to: an electrical connection with one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the above. In the present disclosure, a computer-readable storage medium may be any tangible medium that contains or stores a program that can be used by or in conjunction with an instruction execution system, device, or component. In the present disclosure, a computer-readable signal medium may include a data signal propagated in baseband or as part of a carrier wave, which carries computer-readable program code. Such a propagated data signal may take a variety of forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination of the above. A computer-readable signal medium may also be any computer-readable medium other than a computer-readable storage medium that can transmit, propagate, or transport a program for use by or in conjunction with an instruction execution system, apparatus, or device. The program code contained on the computer-readable medium may be transmitted using any suitable medium, including but not limited to wires, optical cables, RF (radio frequency), etc., or any suitable combination thereof.

[0234] In some embodiments, the client and server can communicate using any currently known or future developed network protocol, such as HTTP (HyperText Transfer Protocol), and can be interconnected with any form or medium of digital data communication (e.g., a communication network). Examples of communication networks include a local area network ("LAN"), a wide area network ("WAN"), an internet (e.g., the Internet), and a peer-to-peer network (e.g., an ad hoc peer-to-peer network), as well as any currently known or future developed network.

[0235] The computer-readable medium may be included in the electronic device, or may exist independently without being incorporated into the electronic device.

[0236] The computer-readable medium carries one or more programs. When the one or more programs are executed by the electronic device, the electronic device executes the video editing method.

[0237] In embodiments of the present disclosure, computer program code for performing the operations of the present disclosure may be written in one or more programming languages ​​or a combination thereof, including but not limited to object-oriented programming languages ​​such as Java, Smalltalk, C++, and conventional procedural programming languages ​​such as "C" or similar programming languages. The program code may be executed entirely on the computer, partially on the computer, as a separate software package, partially on the computer and partially on a remote computer, or entirely on the remote computer or server. In cases involving a remote computer, the remote computer may be connected to the computer via any type of network, including a local area network (LAN) or a wide area network (WAN), or may be connected to an external computer (e.g., via the Internet using an Internet service provider).

[0238] The flowcharts and block diagrams in the accompanying drawings illustrate the possible implementation architecture, functions and operations of the systems, methods and computer program products according to various embodiments of the present disclosure. In this regard, each box in the flowchart or block diagram can represent a module, program segment, or a part of code, and the module, program segment, or a part of code contains one or more executable instructions for realizing the specified logical function. It should also be noted that in some alternative implementations, the functions marked in the box can also occur in a different order than that marked in the accompanying drawings. For example, two boxes represented in succession can actually be executed substantially in parallel, and they can sometimes be executed in the opposite order, depending on the functions involved. It should also be noted that each box in the block diagram and / or flowchart, and the combination of the boxes in the block diagram and / or flowchart, can be implemented with a dedicated hardware-based system that performs the specified function or operation, or can be implemented with a combination of dedicated hardware and computer instructions.

[0239] The units involved in the embodiments described in this disclosure may be implemented in software or hardware, wherein the name of a unit does not necessarily limit the unit itself.

[0240] The functions described above herein may be performed, at least in part, by one or more hardware logic components. For example, and without limitation, exemplary types of hardware logic components that may be used include: field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), systems on chip (SOCs), complex programmable logic devices (CPLDs), and the like.

[0241] In the context of the present disclosure, a computer-readable medium can be a tangible medium that can contain or store a program for use by an instruction execution system, device or equipment or used in combination with an instruction execution system, device or equipment. A computer-readable medium can be a computer-readable signal medium or a computer-readable storage medium. A computer-readable medium can include, but is not limited to, an electronic, magnetic, optical, electromagnetic, infrared, or semiconductor system, device or equipment, or any suitable combination of the foregoing. A more specific example of a computer-readable storage medium can include an electrical connection based on one or more lines, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the foregoing.

[0242] The above description is merely a preferred embodiment of the present disclosure and an illustration of the technical principles employed. Those skilled in the art should understand that the scope of disclosure involved in the present disclosure is not limited to the technical solutions formed by the specific combination of the above-mentioned technical features, but also includes other technical solutions formed by any combination of the above-mentioned technical features or their equivalents without departing from the above-mentioned disclosed concepts. For example, a technical solution formed by replacing the above-mentioned features with (but not limited to) technical features with similar functions disclosed in this disclosure.

[0243] In addition, although each operation is described in a specific order, this should not be understood as requiring these operations to be performed in the specific order shown or in a sequential order. Under certain circumstances, multitasking and parallel processing may be advantageous. Similarly, although some specific implementation details have been included in the above discussion, these should not be interpreted as limiting the scope of the present disclosure. Some features described in the context of a separate embodiment can also be implemented in a single embodiment in combination. On the contrary, the various features described in the context of a single embodiment can also be implemented in multiple embodiments individually or in any suitable sub-combination mode.

[0244] Although the subject matter has been described in language specific to structural features and / or methodological logical acts, it should be understood that the subject matter defined in the appended claims is not necessarily limited to the specific features or acts described above. Rather, the specific features and acts described above are merely example forms of implementing the claims.

Claims

1. A video editing method, the method comprising: Obtaining a first video material and a target video copy corresponding to the first video material; Determining a target video editing template based on the target video copy, the target video editing template being used to indicate filling slots for at least one video segment and editing effects applied to the at least one video segment; Intercepting at least one material segment from the first video material; the at least one material segment corresponding to the filling slots of the at least one video segment and being used to fill the target video editing template to form the at least one video segment; Generating a second video material according to the at least one material segment and the target video editing template; wherein the second video material is used to present the editing effects applied to the at least one material segment; Generating a target video by splicing the second video material in front of the first video material.

2. The method according to claim 1, wherein, The obtaining the first video material and the target video copy corresponding to the first video material includes: Obtaining the target video copy of the video material to be processed; Dividing the video material to be processed into a plurality of video materials based on the target video copy, each of the plurality of video materials being a respective one of the first video materials; Or, The obtaining the first video material and the video copy corresponding to the first video material includes: Obtaining the video copy of the video material to be processed of the video material to be processed; Dividing the video material to be processed into a plurality of video materials based on the video copy of the video material to be processed, each of the plurality of video materials being a respective one of the first video materials; Obtaining the target video copy of the first video material; The generating a target video by splicing the second video material in front of the first video material includes: Generating an edited video corresponding to the first video material by splicing the second video material in front of the first video material, so as to generate an edited video corresponding to each video material; Splicing the edited videos corresponding to each video material to obtain the target video.

3. The method according to claim 1, wherein The target video editing template is further used to indicate a video title filling slot; The method further includes: Generating a video title of the first video material according to the video copy of the first video material, the video title being used to fill the video title filling slot; The generating a second video material according to the at least one material segment and the target video editing template includes: Generating a second video material according to the at least one material segment, the video title and the target video editing template; wherein the second video material is used to present the editing effects applied to the at least one material segment and the video title.

4. The method according to claim 1, wherein, The determining a target video editing template based on the target video copy includes: Determining a category label of the category to which the target video copy belongs; Determining a target video editing template corresponding to the category label from a plurality of video editing templates, different category labels corresponding to different video editing templates.

5. The method according to claim 1, wherein, The filling slot of each video segment is further used to indicate the duration of the corresponding video segment; Intercepting at least one material segment from the first video material includes: Performing slicing processing on the first video material to obtain at least one slice; Performing highlight segment extraction processing on each of the at least one slice to obtain at least one highlight segment, and the at least one slice corresponds to the at least one highlight segment one by one; Determining N first material segments from the at least one highlight segment, the duration of each first material segment is the duration of a video segment, each of the first material segments corresponds to a highlight segment, and N is a natural number; When N is less than M, determining M - N second material segments from the first video material, the duration of each second material segment is the duration of a video segment other than the video segments corresponding to the N first material segments, M is the number of the at least one material segment, and M is a positive integer; Determining the N first material segments and the M - N second material segments as the at least one material segment; When N is equal to M, determining the N first material segments as the at least one material segment.

6. The method according to claim 5, wherein When N is less than M, determining M - N second material segments from the first video material includes: When the video duration of the remaining video in the first video material other than the N first material segments is greater than the target duration, determining the M - N second material segments from the remaining video, and the target duration is the sum of the durations of the video segments other than the video segments corresponding to the N first material segments among the at least one video segment; When the video duration of the remaining video is less than or equal to the target duration, determining K second material segments from the remaining video, and K is a natural number less than or equal to M - N; Determining M - N - K second material segments from the first video material.

7. The method according to any one of claims 1 to 6, wherein, Each filling slot of the video segment is further used to indicate the arrangement order of the corresponding video segment. Generating a second video material according to the at least one material segment and the target video editing template includes: Based on the target video editing template, synthesizing the at least one material segment in the arrangement order of the corresponding video segments respectively to generate the second video material.

8. A video editing device, including: An acquisition module, configured to acquire a first video material and a target video copy corresponding to the first video material; A determination module, configured to determine a target video editing template based on the target video copy, and the target video editing template is used to indicate the filling slots of at least one video segment and the editing effects applied to the at least one video segment; An interception module, configured to intercept at least one material segment from the first video material; the at least one material segment corresponds to the filling slots of the at least one video segment and is used to fill the target video editing template to form the at least one video segment; A generation module, configured to generate a second video material according to the at least one material segment and the target video editing template; wherein, the second video material is used to present the editing effects applied to the at least one material segment; Generate a target video by splicing the second video material in front of the first video material.

9. An electronic device, comprising: A memory and a processor, the memory is used to store a computer program; The processor is used to execute the video editing method according to any one of claims 1 to 7 when calling the computer program.

10. A computer-readable storage medium, on which a computer program is stored, and when the computer program is executed by a processor, the video editing method according to any one of claims 1 to 7 is implemented.

Citation Information

Patent Citations

  • Video generation method and device, electronic equipment and storage medium

    CN116266856A

  • Video content generation method and device, electronic equipment and storage medium

    CN116471450A

  • Video processing method and device, electronic equipment and computer readable storage medium

    CN116471451A

  • Special effect video determination method and device, electronic equipment and storage medium

    CN116847147A

  • Video creation, editing, and sharing for social media

    US20170294212A1