Video editing method and apparatus, and electronic device and storage medium

Through the structure and editing information of the video editing template, users can customize the import of materials and generate video editing results, solving the problem of high material requirements for existing video templates and realizing the flexible use of video templates in short videos, medium videos and long videos.

WO2025139877A1PCT designated stage expired Publication Date: 2025-07-03BEIJING ZITIAO NETWORK TECH CO LTD
View PDF 6 Cites 0 Cited by

Patent Information

Application Number
PCT/CN2024/139575
Authority / Receiving Office
WO · WO
Patent Type
Applications
Current Assignee / Owner
Priority Date
2023-12-28
Filing Date
2024-12-16
Publication Date
2025-07-03

AI Technical Summary

Technical Problem

Existing video templates have high requirements for materials and many restrictions, which lead to inflexible use, which is especially difficult to apply in video production other than short videos.

Method used

Provides a video editing method, displays video paragraphs and editing effects through video editing templates, allows users to customize the import of materials, and generates video editing results based on templates and materials, reducing the requirements for the number and duration of materials, and is suitable for short videos, medium videos and long videos.

Benefits of technology

It improves the flexibility of using video templates, making video templates suitable not only for short videos, but also for medium videos, reducing the requirements for materials and enhancing users' creative freedom.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN2024139575_03072025_PF_FP_ABST
    Figure CN2024139575_03072025_PF_FP_ABST
Patent Text Reader

Abstract

Provided in the embodiments of the present disclosure are a video editing method and apparatus, and an electronic device and a storage medium. The video editing method comprises: in response to a trigger operation on a video editing template, presenting a video editing interface; in response to a material import operation in respect of a target video segment, presenting identifiers of at least one image material in a segment editing area of the target video segment, wherein at least one video segment comprises the target video segment, and the at least one image material is material which is imported into the target video segment on the basis of the material import operation; and in response to a trigger operation for editing processing, generating a video editing result on the basis of the video editing template and the at least one image material.
Need to check novelty before this filing date? Find Prior Art

Description

Video editing method and device, electronic device and storage medium

[0001] This application claims priority to Chinese Patent Application No. 202311843312.4 filed on December 28, 2023, and the contents of the above-mentioned Chinese patent application disclosure are hereby incorporated by reference in their entirety as a part of this application. Technical Field

[0002] Embodiments of the present disclosure relate to a video editing method and apparatus, an electronic device, and a storage medium. Background Art

[0003] Video content creation is ubiquitous in daily life. Users use it to document their lives, showcase their personalities, and contribute value. There are two common approaches to video production: recording and editing videos themselves, or creating videos using video templates. Because creating videos using video templates is convenient, quick, and offers rich effects, it has become the primary method for sharing video content. Summary of the Invention

[0004] At least one embodiment of the present disclosure provides a video editing method, comprising: in response to a triggering operation on a video editing template, displaying a video editing interface, the video editing template including structural information and editing information, the structural information being used to indicate at least one video segment, the editing information being used to indicate at least one editing effect applied in the at least one video segment, the video editing interface including segment editing areas corresponding to the at least one video segment respectively; in response to a material importing operation for a target video segment, displaying an identifier of at least one image material in the segment editing area of ​​the target video segment, the at least one video segment including the target video segment, the at least one image material being a material imported into the target video segment based on the material importing operation; and in response to a triggering operation for editing processing, generating a video editing result based on the video editing template and the at least one image material, the portion of the video editing result corresponding to the target video segment being an editing result obtained based on a target editing effect and at least one image material, the target editing effect matching an editing effect of the at least one editing effect located within the target video segment.

[0005] At least one embodiment of the present disclosure further provides a video editing device, comprising: a first display unit, configured to display a video editing interface in response to a triggering operation on a video editing template, wherein the video editing template includes structural information and editing information, the structural information is used to indicate at least one video segment, and the editing information is used to indicate at least one editing effect applied in the at least one video segment, and the video editing interface includes segment editing areas corresponding to the at least one video segment; a second display unit, configured to display an identifier of at least one image material in the segment editing area of ​​the target video segment in response to a material importing operation on a target video segment, the at least one video segment includes the target video segment, and the at least one image material is a material imported into the target video segment based on the material importing operation; and a result generating unit, configured to generate a video editing result based on the video editing template and the at least one image material in response to a triggering operation on an editing process, wherein a portion of the video editing result corresponding to the target video segment is an editing result obtained based on a target editing effect and the at least one image material, and the target editing effect matches an editing effect of the at least one editing effect located within the target video segment.

[0006] At least one embodiment of the present disclosure also provides an electronic device, comprising: a processor; a memory, comprising one or more computer program modules; wherein the one or more computer program modules are stored in the memory and configured to be executed by the processor, and the one or more computer program modules include instructions for implementing the video editing method described in any embodiment of the present disclosure.

[0007] At least one embodiment of the present disclosure further provides a storage medium for storing non-transitory computer-readable instructions. When the non-transitory computer-readable instructions are executed by a computer, the video editing method described in any embodiment of the present disclosure can be implemented. BRIEF DESCRIPTION OF THE DRAWINGS

[0008] The above and other features, advantages, and aspects of the various embodiments of the present disclosure will become more apparent with reference to the following detailed description in conjunction with the accompanying drawings. Throughout the drawings, like reference numerals represent like elements. It should be understood that the drawings are schematic and that the originals and elements are not necessarily drawn to scale.

[0009] FIG1A shows a schematic flow chart of a video editing method provided by at least one embodiment of the present disclosure;

[0010] FIG1B shows a schematic diagram of a video editing method provided by at least one embodiment of the present disclosure;

[0011] FIG1C shows a schematic diagram of structural information provided by at least one embodiment of the present disclosure;

[0012] FIG1D shows a schematic diagram of a video editing interface provided by at least one embodiment of the present disclosure;

[0013] FIG2 shows a schematic diagram of adjusting an editing effect in a video editing template to obtain an editing effect in a video editing result, provided by some embodiments of the present disclosure;

[0014] FIG3 shows a flowchart of the method of step S30 in FIG1A provided by at least one embodiment of the present disclosure;

[0015] FIG4 shows a schematic diagram of a preview page provided by at least some embodiments of the present disclosure;

[0016] 5A to 5F illustrate schematic diagrams of modifying subtitles provided by at least some embodiments of the present disclosure;

[0017] FIG6 is a schematic block diagram of a video editing device provided by some embodiments of the present disclosure;

[0018] FIG7 is a schematic block diagram of an electronic device provided by some embodiments of the present disclosure;

[0019] FIG8 is a schematic block diagram of another electronic device provided by some embodiments of the present disclosure; and

[0020] FIG9 is a schematic diagram of a storage medium provided in some embodiments of the present disclosure. DETAILED DESCRIPTION

[0021] The following describes embodiments of the present disclosure in more detail with reference to the accompanying drawings. Although certain embodiments of the present disclosure are shown in the accompanying drawings, it should be understood that the present disclosure can be implemented in various forms and should not be construed as limited to the embodiments described herein. Rather, these embodiments are provided to provide a more thorough and complete understanding of the present disclosure. It should be understood that the drawings and embodiments of the present disclosure are for illustrative purposes only and are not intended to limit the scope of protection of the present disclosure.

[0022] It should be understood that the various steps described in the method embodiments of the present disclosure may be performed in different orders and / or in parallel. In addition, the method embodiments may include additional steps and / or omit the steps shown. The scope of the present disclosure is not limited in this respect.

[0023] As used herein, the term "including" and its variations are open-ended, i.e., "including but not limited to." The term "based on" means "based, at least in part, on." The term "one embodiment" means "at least one embodiment," the term "another embodiment" means "at least one additional embodiment," and the term "some embodiments" means "at least some embodiments." Other terms are defined in the following description.

[0024] It should be noted that the concepts of "first" and "second" mentioned in this disclosure are only used to distinguish different devices, modules or units, and are not used to limit the order or interdependence of the functions performed by these devices, modules or units.

[0025] It should be noted that the modifications of "one" and "plurality" mentioned in this disclosure are illustrative rather than restrictive, and those skilled in the art will understand that unless the context clearly indicates otherwise, they should be understood as "one or more." "Plurality" should be understood as two or more.

[0026] The names of the messages or information exchanged between multiple devices in the embodiments of the present disclosure are only used for illustrative purposes and are not used to limit the scope of these messages or information.

[0027] As video templates become more diverse and diverse, the demand and scenarios for creating and sharing videos based on them will continue to grow. However, most current video templates are designed for short videos. These templates have fixed slots, requiring users to import the same number of assets, clips, and duration as the template. Consequently, short-video templates place high demands on the assets they import and impose many restrictions, making them inflexible.

[0028] In response to the above problems, at least one embodiment of the present disclosure provides a video editing method and apparatus, an electronic device, and a storage medium, which can reduce the requirements and restrictions of video templates on user-imported materials and improve the flexibility of using video templates.

[0029] At least one embodiment of the present disclosure provides a method for generating videos using video templates. This method not only reduces the requirements for source material, making the use of video templates more flexible, but also applies to both short and medium-length videos. Short videos generally refer to user-generated videos with a duration of less than one minute, i.e., user-generated content (UGC). Long videos generally refer to professionally produced videos with a duration of more than 30 minutes and higher content quality, i.e., professionally generated content (PGC). Medium-length videos are somewhere between short and long videos. Although user-generated content is user-generated, the user's professional level is higher than that of short videos, i.e., professionally generated content (PUGC). Medium-length videos are also generally between short and long videos. From the user's perspective, short videos do not require deliberate time allocation and deliver fragmented content. They can capture the key content of the video in a shorter time, but they may not necessarily be retained. Medium-length and long videos, on the other hand, require finding a suitable time and place to watch, requiring more effort, and the video content can be retained in memory for a longer period of time.

[0030] It should be noted that, in the embodiments of the present disclosure, the meanings of video template and template video are the same, both referring to templates used as reference when making videos, and the template is presented in the form of a video.

[0031] At least one embodiment of the present disclosure provides a video editing method, a video editing device, an electronic device, and a computer-readable storage medium. The video editing method includes: in response to a trigger operation on a video editing template, displaying a video editing interface, the video editing template including structural information and editing information, the structural information being used to indicate at least one video segment, the editing information being used to indicate at least one editing effect applied in the at least one video segment, the video editing interface including segment editing areas corresponding to the at least one video segment; in response to a material import operation for a target video segment, displaying an identifier of at least one image material in the segment editing area of ​​the target video segment, the at least one video segment including the target video segment, the at least one image material being a material imported into the target video segment based on the material import operation; and in response to a trigger operation for editing processing, generating a video editing result based on the video editing template and the at least one image material, the portion of the video editing result corresponding to the target video segment being an editing result in which the target editing effect is applied to the at least one image material, the target editing effect matching the editing effect of the at least one editing effect located within the target video segment. This video editing method does not limit the at least one imported image material to match the video editing template, but applies the editing effect of the video editing template to the at least one imported image material. There are no requirements for the quantity and time length of the at least one imported image material. This not only reduces the requirements for image materials, making the use of video editing templates more flexible, but also this method is applicable to medium videos in addition to short videos.

[0032] Hereinafter, embodiments of the present disclosure will be described in detail with reference to the accompanying drawings.

[0033] Figure 1A shows a flow chart of a video editing method provided by at least one embodiment of the present disclosure; Figure 1B shows a schematic diagram of a video editing method provided by at least one embodiment of the present disclosure; Figure 1C shows a schematic diagram of structural information provided by at least one embodiment of the present disclosure; Figure 1D shows a schematic diagram of a video editing interface provided by at least one embodiment of the present disclosure.

[0034] As shown in FIG. 1A , in at least one embodiment, the method includes the following operations.

[0035] Step S10: In response to a triggering operation on a video editing template, a video editing interface is displayed. The video editing template includes structural information and editing information. The structural information is used to indicate at least one video segment. The editing information is used to indicate at least one editing effect applied to at least one video segment. The video editing interface includes segment editing areas corresponding to the at least one video segment.

[0036] Step S20: In response to the material import operation for the target video segment, the identifier of at least one image material is displayed in the segment editing area of ​​the target video segment, at least one video segment includes the target video segment, and at least one image material is the material imported into the target video segment based on the material import operation.

[0037] Step S30: In response to the triggering operation of the editing process, a video editing result is generated according to the video editing template and at least one image material, wherein the portion of the video editing result corresponding to the target video segment is the editing result in which the target editing effect is applied to at least one image material, and the target editing effect matches the editing effect of at least one editing effect located within the target video segment.

[0038] For example, in step S10, the triggering operation for the video editing template is, for example, a click operation on a use icon on the main page of the video editing template. As shown in FIG1B , the main page of the video editing template is page 101, which includes a use icon 102. When the client receives a click operation on the use icon 102, the video editing interface 103 shown in FIG1D is displayed.

[0039] For example, the video editing method is applied to a video application, and multiple video editing template covers are displayed on the template recommendation page of the video application. If the user selects the cover of a video editing template, a display video of the video editing template is played for the user's reference. When the display video ends, the user enters the main page of the video editing template. The user can operate on the main page (for example, page 101) to use the video editing template or not. If the user performs a trigger operation on the main page, it means that the user has selected to use the video editing template, and then enters the video editing interface 103 of the video editing template to perform video editing in the video editing interface 103.

[0040] In some embodiments of the present disclosure, a video editing template includes structural information and editing information. The structural information indicates at least one video segment, and the editing information indicates at least one editing effect applied to the video segment. The editing information may include, for example, at least one of subtitles, dubbing, music, text, filters, and transition animations.

[0041] As shown in FIG1C , the video editing template includes structural information 100. The structural information 100 is set by, for example, the producer of the video editing template. For example, the publishing link and the vimo platform support adding the video editing template structural information. For example, the structural information 100 shown in FIG1 is added via the vimo platform.

[0042] As shown in FIG1C , the structural information 100 includes three parts: overall description, intelligent capability, and template structure decomposition.

[0043] The template structure consists of the first section 1 and allows for the addition of more sections. Section 1 represents a video segment, and the operator can fill in information such as the title and description of the video segment. For example, by clicking the Add More Sections control 2, you can add a second section, a third section, and so on, each representing a video segment.

[0044] The overall description part is controlled by the operator, for example, to fill in the description of the video editing template, such as content, style, etc. The overall description part does not limit the maximum number of characters entered.

[0045] Intelligent capabilities, such as intelligent subtitle and dubbing, can be selected by operators. Operators can choose whether to select intelligent capabilities and which one or more intelligent capabilities to select. If an intelligent capability is selected, the video editing template will have that intelligent capability.

[0046] The structure of each section of the structural information is similar. For example, Section 1 is described below. Section 1 includes a title, description, start and end times, and start and end segments. The title, for example, provides a concise description of the template image material for Section 1 in the video editing template. The template image material, for example, is the material used in the video editing template. The description, for example, provides a general description of the template image material for Section 1.

[0047] In some embodiments of the present disclosure, the structural information indicates each of at least one video segment by a time interval or a sequence number of consecutive video segments. For example, in the structural information of FIG1C , each video segment is identified by a start and end time, or by a start and end segment.

[0048] For example, the template image material is at least one continuous video segment, the start and end time is selected from the at least one continuous video segment from the Kth second to the Lth second (time interval) as the material of the first part 1, and the start and end segments are, for example, selected from the at least one continuous video segment from the jth segment to the ith segment as the material of the first part 1. j and i are examples of serial numbers of the video segments.

[0049] It's important to note that the structure information's division of start and end times and segments is intended to facilitate understanding of video editing templates and facilitate editing for users. User-imported footage doesn't necessarily follow this structure information. For example, users can import footage of any format and then perform processing such as segmentation and highlight slicing based on the structure information.

[0050] In the above embodiment, the video editing template includes the template image material. In some other embodiments of the present disclosure, the video editing template may also not include the template image material.

[0051] In some embodiments of the present disclosure, the video editing template of the medium video can directly reuse the production and publishing process of the short video template. After exporting the project file through the editing tool, select "Publish Template" to obtain the video editing template.

[0052] After the medium video template is released, it supports storage and management in the vimo background, and reuses the short video template logic. For example, by manually marking it on vimo, it is determined whether the template is of the "medium video template" type. Vimo supports batch addition of identifiers of the "medium video template" type. For the medium video template type, it is necessary to add an attribute field of the "medium video template" type to the metadata, and distinguish the video type through the type attribute field identifier (for example, identified as video_type_id). For those marked as medium video template types, the medium video template logic is performed on the entire template link.

[0053] For editing information, it can be similar to a short video template. Producers can add text information, dubbing information, music information, sticker information, filter information, transition animation information and other editing information according to their own ideas and needs. These editing information indicate the editing effects, and the editing effects include at least one of subtitles, dubbing, music, text, filters, and transition animations.

[0054] As shown in FIG1D , the video editing interface 103 includes at least one paragraph editing area corresponding to each video segment. Each paragraph editing area serves as an editing area for a theme, and displays video clips of different themes. For example, if the first part 1 in FIG1C indicates a paragraph editing area 11 corresponding to the video segment, and the content entered in the title of the first part 1 is "finished dish", then the title of the paragraph editing area 11 corresponding to the video segment indicated by the first part 1 is "finished dish". For example, if the structural information also includes a second part, and the content entered in the title of the second part is "food preparation process", then the video segment indicated by the second part can correspond to the paragraph editing area 12, and the title of the paragraph editing area 12 is "food preparation process".

[0055] Each video segment is regarded as a theme, and each segment editing area is regarded as an editing area of ​​a theme. In the following description, the video segment is referred to as a theme, and the segment editing area is referred to as an editing area of ​​a theme.

[0056] The video editing interface 103 can be divided into multiple parts, each part is used to display different content. For example, in the example of Figure 1D, the video editing interface 103 includes a navigation bar title 113, in which the title is "Select Material", and the video editing interface 103 also includes M topics arranged in sequence. The editing area, M is a positive integer. For example, the first topic is "Finished Dish", and the second topic is "Food Making Process". The user can import the corresponding creative video clips or pictures for each topic in the editing area; or choose not to import creative video clips or pictures in at least one of the M topics. Below the editing area of ​​each topic can be included a related description 163 of the video segment corresponding to the topic. If there is no related description of the topic in the video editing template, for example, if no corresponding characters are entered in the description part of Figure 1C, the related description will not be displayed.

[0057] It should be noted that in embodiments of the present disclosure, a video editing template can be generated based on a video editing draft. The video editing draft contains assets (such as video, audio, and images) and editing information. The video editing template needs to indicate which assets in the video editing draft need to be replaced. Medium video templates differ from short video templates in two key aspects. First, medium video templates have a paragraph structure, where a video segment corresponds to a timeline interval and includes one or more video clips. Short video templates, on the other hand, lack a paragraph structure. A short video template is simply a video editing draft with some video clips designated as assets that need to be replaced. Second, in terms of template usage, medium video templates apply the video editing effects within each paragraph separately. The application logic within a paragraph is based on the editing operations indicated by the editing information, which are then applied to the assets imported into the paragraph after changes. (Because the number of assets is not bound to the clips in the template, editing operations cannot be directly reused on assets.) In contrast, the application logic of short video templates simply imports the video clips in the template into the assets, and the editing operations are applied directly to the assets, without the need for segmentation or changing the editing operations.

[0058] Regarding step S20, as shown in FIG1D , the editing area for each theme includes an add entry 123. For example, in response to a click on add entry 123, a list of albums is pulled up, allowing the user to select image materials from the list to import into the theme. Specifically, the material import operation includes, for example, clicking on add entry 123 and selecting image materials from the list of albums to import into the theme.

[0059] In some embodiments of the present disclosure, the adding entry 123 does not have a slot concept and does not limit the number of segments and the length of the segments.

[0060] For example, after a material import operation is performed on a target video segment (also referred to as a target theme), an identifier of at least one image material is displayed in a segment editing area of ​​the target theme.

[0061] In some embodiments of the present disclosure, for example, if a video editing template includes M themes, a target theme can be selected from these M themes for importing video material. For example, if a user sequentially selects N themes for importing video material, each of these N themes will serve as the target video segment. M and N are both positive integers. In the example of FIG1D , in response to a user clicking on the add entry 123 in the "Finished Dish" section, an album list 133 is displayed, allowing the user to select at least one video material related to the finished dish from the album list 133. The identifier of each of the at least one video material is then displayed in the segment editing area 11 for the target theme "Finished Dish." The identifier of each of the at least one video material can be, for example, a frame extracted from each video material, such as the first frame of each video material being extracted as the identifier of the video material. For example, in the example of FIG1D , if a user imports video material 1 and video material 8 from the album list 133, and the first frames of video material 1 and video material 8 display the number 1 and the number 8, respectively, the images displaying the number 1 and the number 8 are displayed in the segment editing area 11.

[0062] The user can click on the add entry 123 in the "Food Preparation Process" to pull up the album list 133 again, allowing the user to select at least one video material related to the food preparation process from the album list 133. In the embodiment of the present disclosure, the user can select materials for each theme or only for some of the M themes.

[0063] Compared with short videos, medium videos have obvious editing structure routines. One or more image materials imported by the user are grouped and sorted according to their structural information. In this embodiment, the video editing interface of the video editing template includes M themes arranged in sequence, so that one or more image materials are grouped according to the theme, so that the video structure is rationalized. In the embodiment of the present disclosure, there is basically no limit on the amount and duration of user-imported materials, and batch selection and deletion of currently selected materials are supported. However, in order to avoid extreme situations, in some embodiments of the present disclosure, the amount and duration of imported materials can be subject to maximum restrictions. For example, the duration of a single material (i.e., a video clip) can only be up to 30 minutes. If it exceeds, the prompt "Only videos within 30 minutes can be imported" will be displayed. The upper limit of the number of imported materials for each theme is 100 clips. If it exceeds, the prompt "Only 100 clips can be imported at most" will be displayed. For example, the total duration of at least one image material is limited by the upper limit of the total duration of the user material in Figure 1C.

[0064] In some embodiments of the present disclosure, the album list may include prompts. For example, when no video clips are selected, the default prompt is "Select the video clips to import." When video clips are selected, the prompt changes to "%d clips selected" based on the number of clips. As shown in Figure 1D , if video clips 1 and 8 are selected, the prompt becomes "2 clips selected." As shown in Figure 1D , the album list may also include a confirm add button 143. This confirm add button 143 can be lit or grayed out. After importing a clip, the confirm add button 143 is lit. The logic is that the button is lit when a video clip has been imported and grayed out otherwise. Clicking the add button places the imported video clips into the corresponding frames in their order. As shown in Figure 1D , clicking the add button 143 places the labels for video clips 1 and 8 in the section editing area 11 with the theme "Finished Dish." If the current theme frame is already populated with video clips, the album list will display the added video clips by default.

[0065] In some embodiments of the present disclosure, the addition entry 123 moves backward following the position of the filled material, and the image material added later is automatically sorted behind the existing image material. If it exceeds one line, it will automatically wrap. The added image material will display the length of its original image material, and will be intelligently intercepted and split when entering the next step of synthesis.

[0066] In some embodiments of the present disclosure, image material deletion operations are supported. In response to a deletion operation on an image material, the image material is deleted, and the image material of the subject is automatically postponed to the front row. For example, the deletion operation can be performed by selecting the delete button in the album list page and the video editing interface 103. In some embodiments of the present disclosure, users can also sort and edit the imported image materials. For example, the positions of two image materials can be swapped, or editing operations such as beautification and cropping can be performed on a certain image material.

[0067] In the embodiments of the present disclosure, the image material may be a video clip or a picture.

[0068] In step S30, the triggering operation for the editing process is, for example, selecting an export control or a preview control after completing editing in the video editing interface. For example, selecting the export control can generate an exported composite video or a generated file based on the video editing template and at least one image source. For example, selecting the preview control can generate a preview object on the preview page based on the video editing template and at least one image source. The preview object can be a video, an image, or the like.

[0069] For example, a video editing result includes multiple video sections, each corresponding to a target video segment. Specifically, each video section represents the editing result of applying a target editing effect to at least one image source within the target video segment. The target editing effect is derived from at least one editing effect located within the target video segment.

[0070] For example, the first image material is imported into the theme "Finished Dish", and for the target video segment of the theme "Finished Dish", the target editing effect is applied to the first image material, and the target editing effect is an editing effect that matches the editing effect located in the target video segment in the video editing template.

[0071] For example, the target editing effect is the same as the editing effect in the target video segment in the video editing template, or the target editing effect is obtained by changing the editing effect in the target video segment in the video editing template.

[0072] In some embodiments of the present disclosure, step S30 includes: in response to a triggering operation of editing processing, adjusting the editing effect according to a matching strategy to obtain a target editing effect; applying the target editing effect to at least a portion of at least one image material according to a video editing template to obtain a video portion of a target video segment; and generating a video editing result based on the video portion corresponding to each target video segment in at least one video segment.

[0073] In some embodiments of the present disclosure, the matching strategy includes: when the editing effect includes a first copy, the target editing effect includes a second copy, the copy format of the second copy is the same as the copy format of the first copy, and the copy content of the second copy matches at least one image material.

[0074] In some embodiments of the present disclosure, copywriting can refer to all texts other than subtitles, such as decorative texts such as titles. For example, each theme adds a copywriting content, and the copywriting content takes effect based on the theme.

[0075] FIG2 shows a schematic diagram of adjusting the editing effects in a video editing template to obtain the editing effects in a video editing result, provided by some embodiments of the present disclosure.

[0076] As shown in Figure 2(a), in the video editing template, the copy for Topic 1 is Text 1, the copy for Topic 2 is Text 2, and the copy for Topic 3 is Text 3. That is, Text 1 applies only to Topic 1, File 2 applies only to Topic 2, and File 3 applies only to Topic 3. For example, Topic 1 is the "Finished Dish" theme in Figure 1D, and Topic 2 is the "Food Preparation Process" theme in Figure 1D. Topic 3 is, for example, another added theme. Text 1, Text 2, and Text 3 are examples of the first copy.

[0077] The text format includes, for example, the font, color, and style of the text.

[0078] For example, if text 1 is in regular font, the font of the text in Topic 1 in the video editing result will also be regular font; if text 2 is in gradient color, the font of the text in Topic 2 in the video editing result will also be the same gradient color; if text 3 is in italic font, the font of the text in Topic 3 in the video editing result will also be italic font.

[0079] For example, step S30 also includes obtaining the content of the second copy. For example, in the example of Figure 1C , the video editing interface 103 may also include an information prompt section 104, where the user can enter prompt information to explain and describe the topic. In this embodiment, obtaining the content of the second copy includes: analyzing the prompt information and at least one video material to obtain an analysis result; and generating the content of the second copy based on the analysis result.

[0080] For example, a neural network used for natural language processing can be used to analyze the prompt information to obtain analysis results, thereby more accurately understanding the user's needs. For example, a sequence-to-sequence neural network model, a long short-term memory network, etc. can be used to understand the prompt information and obtain analysis results. For example, image processing technology can be used to perform feature recognition on image materials to obtain analysis results for at least one image material. The analysis results of the prompt information and the analysis results of the at least one image material are combined to generate the content of the second copy that matches the prompt information and the at least one image material.

[0081] For example, if the prompt message is "sweet and sour" and at least one image material is a picture of a beverage, the second copy content can be "sour and sweet, refreshing".

[0082] In some embodiments of the present disclosure, for example, the influence of the analysis result on the content of the second copy is greater than the influence of at least one image material on the content of the second copy. For example, when the neural network generates the content of the second copy, the weight of the analysis result is greater than the weight of the at least one image material. This can better meet the needs of users and the generated copy content is more in line with user expectations.

[0083] In the above embodiment, a copy adapted to the image material is intelligently generated during the generation of the video editing result. For example, the copy content is intelligently written based on the text structure of the video editing template, the image material, and the prompt information. For example, the intelligently generated copy only changes the copy content in the video editing template, does not change the copy format in the video editing template, and is consistent with the copy format in the video editing template. In some embodiments of the present disclosure, the number of words in the text of the intelligently generated copy is close to the number of words in the video editing template, for example, the number of words between the two is within 5 words. For example, the copy of the video editing template is "Home Record", the influencing material is food, and the prompt information is food production, then the content of the corresponding copy can be "Making Food Record". In other embodiments of the present disclosure, the copy in the video editing template can also be directly presented during the generation of the video editing result, or the user can personalize it on the preview page. Please refer to the description below for the preview page.

[0084] The format of the first copy is applied to the content of the second copy, thereby obtaining the second copy included in the target editing effect. Thereafter, the second copy is applied to at least one image material corresponding to the target video segment (for example, image material 1 and image material 8 imported in theme 1).

[0085] In some embodiments of the present disclosure, the matching strategy also includes: in response to the first copy being applied to the entire paragraph of the target video segment, the second copy is applied to the entire part of the video portion; in response to the first copy being applied to a partial time period in the target video segment, the second copy is applied to a partial time period in the video portion.

[0086] In some embodiments of the present disclosure, it is determined whether the text of a certain theme is displayed for the entire theme or only for a certain duration. When the first text in the video editing template covers the entire target theme, the second text is also effective for the entire target theme. When the first text in the video editing template only covers a certain duration of the target theme, the second text is effective for the duration of the target theme consistent with that duration.

[0087] As shown in Figures 2(a) and (b), in the video editing template, text 2 (an example of the first text) is displayed from the 0th second to the 60th second of Topic 2, so in the video editing result, text 2 (an example of the second text) is also only displayed from the 0th second to the 60th second of Topic 2 of the video editing result. In this example, the first text and the second text have the same content and format. In the video editing template, text 3 is displayed for the entire duration of Topic 3, so in the video editing result, text 3 is also displayed for the entire duration of Topic 3. Figures 2(c) and (d) are also similar and will not be repeated here.

[0088] When the user modifies the display duration of the text and then modifies the video material subsequently, it shall be displayed according to the duration range modified by the user; if the duration of the video material becomes shorter, the duration shall be correspondingly shortened.

[0089] In some other embodiments of the present disclosure, the text can be displayed according to the effective duration range of the text in the video editing template. As shown in Fig. 2(e), for example, in the video editing template, text 3 is displayed throughout the duration of theme 3 (for example, a total of 2 minutes), but the duration of theme 3 in the video editing result is a total of 3 minutes, then the text is only displayed within the first 2 minutes of theme 3 in the video editing result. In the video editing template, text 1 is displayed throughout the duration of theme 1 (for example, a total of 1 minute), but the duration of theme 1 in the video editing result is a total of 2 minutes, then the text is only displayed within the first 1 minute of theme 3 in the video editing result.

[0090] In some embodiments of the present disclosure, the text in the target video segment can directly adopt the text in the video editing template, which can improve the speed of synthesizing the video editing result. In some embodiments of the present disclosure, the content of the text can be intelligently rewritten to make the text more compatible with the imported materials.

[0091] In some embodiments of the present disclosure, for example, image recognition is performed on the second material to identify the scene and color tone of the second material, and then the text content of the video is generated based on the understanding of the second material. The text format adopted in the video editing template is directly applied to this text content to obtain the second text.

[0092] In some embodiments of the present disclosure, the matching strategy includes: when the editing effect includes a transition animation, the target editing effect includes a transition animation.

[0093] If the video editing template includes M themes and a material import operation is performed on N themes, and when N < M, for the first theme without imported video material, the transition animation between the first theme and the next theme after the first theme in the video editing template is deleted. That is, if a certain theme in the video editing template does not have imported video material, the transition animation immediately following that theme is also deleted. In this embodiment, only the transition animation corresponding to the target video segment in the video editing template is applied to the target video segment with imported video material.

[0094] As shown in Fig. 2(a), there is a transition animation (abbreviated as "transition") 1 between theme 1 and theme 2, and a transition 2 between theme 2 and theme 3. The transition animation is used to make the video segments of adjacent themes transition naturally, making the switching of the content in the video more natural and smooth.

[0095] As shown in Figure 2(b), if the initial material is not imported into Topic 1, then Transition 1 after Topic 1 is also deleted, and the video editing templates including Transition 2 between Topic 2 and Topic 3 are also applied between Topic 2 and Topic 3 in the video editing results.

[0096] In some embodiments of the present disclosure, as shown in Figure 1D, the structural information further includes an information prompt section 104, which is used to obtain prompt information. Step S30 includes: receiving prompt information input in the information prompt section; and performing semantic analysis on the prompt information to obtain an analysis result.

[0097] For example, a neural network for natural language processing can be used to analyze the prompt information to obtain analysis results, thereby more accurately understanding the user's needs. For example, a sequence-to-sequence neural network model or a long short-term memory network can be used to understand the prompt information.

[0098] FIG3 shows a flowchart of the method of step S30 in FIG1A provided by at least one embodiment of the present disclosure.

[0099] As shown in FIG3 , the method includes steps S301 to S304 .

[0100] Step S301: In response to a triggering operation of an editing process, the prompt information and at least one image material are analyzed to obtain an analysis result.

[0101] Step S302: In response to the editing effect including subtitles and / or dubbing, the subtitles and / or dubbing content in the target editing effect is obtained according to a matching strategy of matching the subtitles and / or dubbing content in the target editing effect with the analysis result.

[0102] Step S303: according to the matching strategy that the format of the subtitles in the target editing effect is the same as the format of the subtitles in the editing effect, the format of the subtitles in the target editing effect is obtained.

[0103] Step S304: according to the matching strategy that the sound parameters of the dubbing in the target editing effect are the same as the sound parameters of the dubbing in the editing effect, the sound parameters of the dubbing in the target editing effect are obtained.

[0104] In some embodiments of the present disclosure, if the video editing template includes tags for intelligently adding subtitles and / or intelligently adding dubbing, subtitles and / or dubbing are automatically added to the video editing result.

[0105] The method can adapt content with subtitles and / or dubbing to prompt information, thereby improving user experience.

[0106] For step S301, the analysis of the prompt information and at least one image material to obtain the analysis result is similar to the above embodiment and will not be repeated here. For example, based on the image material imported by the user and the prompt information input, the content and appeal are understood, and appropriate subtitles and / or dubbing content are generated. If it is recognized that the image material imported by the user includes sound, then the original sound in the image material and the image content of the image material and the prompt information can be combined to generate appropriate subtitles and / or dubbing content. Alternatively, regardless of whether the image material itself includes sound, its intelligent dubbing is defaulted.

[0107] For example, if it is determined through image recognition technology that the image material is an image about making delicious food, and the prompt information is "weekend", then it is determined that the subtitles in the preview video can include "Happy Weekend".

[0108] For steps S302 to S304 , the matching strategy includes a matching strategy for subtitle content and / or dubbing content, a matching strategy for subtitle format, and a matching strategy for sound parameters.

[0109] For subtitle and / or dubbing content, the matching strategy can be to match the subtitle and / or dubbing content in the target editing effect with the analysis results. Matching subtitle and / or dubbing content with analysis results is similar to matching text content with analysis results described above and will not be repeated here.

[0110] For subtitle formats, the matching strategy can be to ensure that the subtitle format in the target editing effect is the same as the subtitle format in the editing effect. For example, the subtitle format in the target editing effect can be the same as the subtitle format in the editing effect in the video editing template. If the template video does not contain subtitle format information, a default font style, such as "white background with black borders," can be used. Subtitle formats include, for example, the subtitle text font, text color, and text style.

[0111] The sound parameters of dubbing may include, for example, the timbre, frequency, pitch, etc. of the dubbing. For the sound parameters, the matching strategy is, for example, that the sound parameters of the dubbing in the target editing effect are the same as the sound parameters of the dubbing in the editing effect. For example, the timbre of the dubbing in the target editing effect is the same as the timbre, frequency and pitch of the dubbing in the editing effect in the video editing template. That is, the subtitle format and the sound parameters of the sound can directly apply the format and sound parameters in the editing effect in the video editing template. When there is no timbre in the editing effect in the video editing template (that is, the video editing template is the original sound of the video), the default timbre is used, and the default timbre can be pre-set. That is, when the image material of the video editing template has an original sound and the video editing template retains the original sound, when the video editing template is used to synthesize the video editing result, the original sound in the video editing template will not be retained by default.

[0112] For example, when the video editing template already contains audio content such as text reading and recording, automatic muting is performed.

[0113] In some embodiments of the present disclosure, for example, subtitles and / or dubbing may be effective for the entire video editing result.

[0114] For example, all imported video clips automatically have subtitles added to the final video, displaying them according to the duration of the final video. If the duration of a single user clip is less than a preset threshold, such as 3 seconds, no subtitles or dubbing will be added to that clip. As shown in Figures 2(a) to 2(e), the subtitle duration is the same as the final video duration.

[0115] In some embodiments of the present disclosure, subtitles and dubbing can be in a one-to-one correspondence. The underlying content of the dubbing is the subtitle content. If the subtitle content is modified, and the modification is allowed to be applied to the dubbing, the dubbing reading will automatically change. In some embodiments of the present disclosure, playing the dubbing and displaying the subtitles are decoupled. For example, Table 1 shows an example of decoupling dubbing and subtitles, provided in at least some embodiments of the present disclosure.

[0116] Table 1

[0117] As shown in Scenario 1, Scenario 2, Scenario 3 and Scenario 4 in Table 1, dubbing and subtitles are decoupled. If the video editing template includes the "Smart Add Dubbing" tag and the "Smart Add Subtitles" tag, for example, the structural information of Figure 1C, the synthesized video editing result is also expected to have dubbing and subtitles. If the video editing template does not include the "Smart Add Dubbing" tag and the "Smart Add Subtitles" tag, the synthesized video editing result is also expected to not include dubbing and subtitles. If the video editing template includes the "Smart Add Dubbing" tag but does not include the "Smart Add Subtitles" tag, the synthesized video editing result is expected to include dubbing but not subtitles. If the video editing template does not include the "Smart Add Dubbing" tag but includes the "Smart Add Subtitles" tag, the synthesized video editing result is expected to not include dubbing but include subtitles.

[0118] In some embodiments of the present disclosure, when the target editing effect only includes subtitles, the dubbing mark in the preview video corresponding to the video editing result is displayed as an off state. When the target editing effect only includes dubbing, the subtitle mark in the preview video corresponding to the video editing result is displayed as an off state. For example, in scene 2, the dubbing mark is in an on state and the subtitle mark is in an off state; in scene 3, the dubbing mark is in an off state and the subtitle mark is in an on state. Those skilled in the art can set arbitrary icons to represent dubbing and subtitles respectively. The dubbing icon includes two different display states, respectively representing an off state and an on state, and the subtitle icon also includes two different display states, respectively representing an off state and an on state.

[0119] In some embodiments of the present disclosure, step S30 includes, for example, splicing the video parts corresponding to each target video segment to obtain a video editing result.

[0120] Table 2 shows an example of a material import method provided by some embodiments of the present disclosure.

[0121] As shown in Table 2, the template video includes Topic 1, Topic 2, and Topic 3 arranged in sequence. The image material imported into Topic 1 includes Segment A, Segment B, and Segment C, and the three segments are arranged in the order of Segment A, Segment B, and Segment C. The image material of Topic 2 includes Segment D, and the image material of Topic 3 includes Segment E, Segment F, and Segment G, and the three segments are arranged in the order of Segment E, Segment F, and Segment G. The video portion of Topic 1, the video portion of Topic 2, and the video portion of Topic 3 are spliced ​​together in the order of the themes to obtain the video editing result. As described above, the video portion of each theme is obtained based on the image material and the target editing effect. For example, the target editing effect is applied to the image material. The video editing result is obtained by, for example, applying the first target editing effect to Segments A, B, and C, applying the second target editing effect to Segment D, and applying the third target editing effect to Segments E, F, and G, and then combining them in order.

[0122] If the video material of subject 1 includes segment A, segment B, and segment C, and the order of the three segments is segment A, segment B, and segment C, and no material is imported into subject 2 and subject 3, the multiple segments are spliced ​​together in the order of segment A, segment B, and segment C, and the first target editing effect is applied to the spliced ​​video of segment A, segment B, and segment C to obtain a video editing result.

[0123] If the video materials of Theme 1 include Clip A, Clip B, and Clip C, and the sorting of the three clips is Clip A, Clip B, and Clip C, Theme 2 is not imported, and the video materials of Theme 3 include Clip E, Clip F, and Clip G, and the sorting of the three clips is Clip E, Clip F, and Clip G. The multiple clips are spliced according to the theme order to obtain a spliced video combined in the order of Clip A, Clip B, Clip C, Clip E, Clip F, and Clip G, and a first target editing effect is applied to the parts of Clip A, Clip B, and Clip C in the spliced video, and a third target editing effect is applied to the parts of Clip F and Clip G in the spliced video to obtain a video editing result.

[0124] If no materials are imported in Theme 1, the video materials of Theme 2 include Clip D, and the video materials of Theme 3 include Clip E, Clip F, and Clip G, and the sorting of the three clips is Clip E, Clip F, and Clip G. The multiple clips are spliced according to the theme order to obtain a spliced video combined in the order of Clip D, Clip E, Clip F, and Clip G, and a second target editing effect is applied to the part of Clip D in the spliced video, and a third target editing effect is applied to the parts of Clip F and Clip G in the spliced video to obtain a video editing result.

[0125] In some embodiments of the present disclosure, when N < M, for the first theme without imported optimized materials, video clips are selected from the second theme with imported video materials as the target clips of the first theme; or the first theme is skipped during the process of splicing multiple video parts in the order of M themes, and the video parts of the themes after the first theme are spliced.

[0126] For example, when the second theme (e.g., the "Food Making Process" theme in FIG. 1D) has no materials, it is automatically postponed to the third theme for splicing with the first theme, and the materials of the third theme are directly arranged behind the materials of the first theme.

[0127] For example, if the video editing template is set to "Default Filling", even if no materials are imported in this theme, a segment (e.g., randomly selected) will be selected from the imported materials for filling to ensure the integrity of the overall preview effect.

[0128] In some embodiments of the present disclosure, at least one editing effect includes not only the above-mentioned paragraph editing effects for each theme, such as text, subtitles, and dubbing, but also a global editing effect. The global editing effect includes, for example, at least one of music, filters, and subtitles. In this embodiment, step S30 includes: in response to the triggering operation of the editing process, adjusting the time length of the global editing effect to match the time length of the target image material used in the video editing result, and obtaining the adjusted global editing effect. On this basis, for example, based on the adjusted global editing effect and the video portion corresponding to each target video paragraph, a video editing result is generated.

[0129] The target image material is, for example, the image material included in the video editing result. In some embodiments of the present disclosure, not all of the at least one image material imported by the user need be used to generate the video editing result. For example, at least one image material can be used to perform highlight processing to obtain a highlight segment to generate the video editing result. In this embodiment, the duration of the image material used in the target video segment is less than the duration of the at least one image material imported into the target video segment.

[0130] In some embodiments of the present disclosure, for example, the target image material included in the video editing result is obtained by splicing highlight clips from one or more target video segments. If the global editing effect includes music, the duration of the music playback is adjusted to match the duration of the target image material. For example, if the duration of the music playback in the video editing template is a first duration and the duration of the target image material is a second duration, the duration of the music playback is adjusted from the first duration to the second duration.

[0131] For example, in response to the second duration being greater than the first duration, the playback time of the music is extended to the second duration, and the time when the music starts playing in the video editing result is the same as the time when the music starts playing in the video editing template; or in response to the second duration being less than the first duration, the playback time of the music is shortened to the second duration, and the music gradually weakens to disappear in the video editing result.

[0132] In an embodiment of the present disclosure, the music is global to the video and there is only one. For example, a video editing template includes a piece of music, which serves as the background music of the video editing template. The background music in the video editing template can be directly applied to the generated video editing result, but the time length of the background music is appropriately adjusted.

[0133] As shown in FIG2(a), video editing template 201 includes Theme 1, Theme 2, and Theme 3. At least one template image material for Theme 1 includes Segments 1 and 2, at least one template image material for Theme 2 includes Segments 3, 4, 5, 6, and 7, and at least one template image material for Theme 3 includes Segments 8, 9, 10, and 11. The music in video editing template 201 runs through the entire template video. For example, the entire video editing template uses the music as background music, and the original duration of the music (i.e., the first duration) is consistent with the duration of the video in video editing template 201.

[0134] As shown in Figure 2(b), if the user only imports video materials into Theme 2 and Theme 3 in the video editing interface, but does not import video materials into Theme 1, so that the playback duration of the target video materials in the video editing result is less than the combined duration of the multiple template video materials in video editing template 201 (hereinafter referred to as "video duration"), the original duration of the music is directly shortened to the playback duration of the target video materials used in Theme 2 and Theme 3. In addition, the background music in the video editing result gradually weakens until it disappears, that is, the background music is faded out.

[0135] As shown in Figure 2(c), although the user has imported video materials into Theme 1, Theme 2, and Theme 3 in the video editing interface, if the playback duration of the target video material in the adopted video material is shorter than the video duration of video editing template 201, the original duration of the music is directly shortened to match the playback duration of the target video material. In addition, the background music in the video editing result gradually weakens until it disappears, that is, the background music is faded out.

[0136] As shown in Figure 2(d), the user imported image materials into Theme 1, Theme 2 and Theme 3 respectively in the video editing interface, but the playback duration of the target image material in the image materials used is longer than the video time length of the video editing template 201, so the original duration of the music is extended to be consistent with the playback duration of the target image material.

[0137] In some embodiments of the present disclosure, when the playback duration of the target image material is longer than the playback duration of the template image material in the video editing template, the end time of the music in the video editing template is extended backward, and the start time of the music in the video editing result is aligned with the video editing template.

[0138] In some embodiments of the present disclosure, if the total duration of the music is long, the remaining music will be played. For example, if the total duration of the music is 3 minutes, but the video editing template only uses the duration range of 0-1 minutes (i.e., the original duration is 1 minute), if the target video material exceeds 1 minute, the remaining content after 1 minute of the music will be played.

[0139] If the total duration of the music is shorter than the playback duration of the target video, the music can be extended intelligently. This can be done by, for example, using a neural network to automatically generate a music segment following the current music, or by searching a music library for a melody or music piece that matches the current music.

[0140] In some embodiments of the present disclosure, if there is a situation where intelligent extension is difficult to achieve, the playback time is extended by automatically looping the current music.

[0141] In some embodiments of the present disclosure, global editing effects include filters. For example, in response to the target video material's playback duration being longer than the video's duration, the filter's duration in the video editing template is extended; or in response to the target video material's playback duration being shorter than the video's duration, the filter's duration in the video editing template is shortened.

[0142] As shown in (b) and (c) of Figure 2, the first filter's effective duration is, for example, the original duration. When the target video material's playback duration is shorter than the original duration, the filter's effective duration is shortened to the target video material's playback duration. As shown in (d) of Figure 2, when the target video material's playback duration is longer than the original duration, the filter's effective duration is extended to the target video material's playback duration. The filter configuration is adjusted in a similar manner to the music configuration.

[0143] In some embodiments of the present disclosure, subtitles can also be used as a global editing effect. The processing logic of subtitles as a global editing effect is similar to that of the aforementioned filters and music, and will not be repeated here.

[0144] In some of the above embodiments, step S30 generates a video editing result based on a matching strategy. In some other embodiments of the present disclosure, step S30 may include: in response to a triggering operation of editing processing, using a neural network to process at least one image material and editing effect to obtain a video editing result.

[0145] For example, in response to a triggering operation of editing processing, the image material and the editing effect are analyzed through a neural network, and an editing effect adapted to the image material is applied to the image material to obtain a video editing result.

[0146] The neural network may include, for example, a neural network for image processing, such as a convolutional neural network, a deep learning neural network, etc.

[0147] In some embodiments of the present disclosure, step S30 includes: in response to a triggering operation for editing, pre-processing at least one image material to obtain a target image material, and then applying a corresponding editing effect to the target image material to obtain a video editing result. Pre-processing may include, for example, highlight recognition and material segmentation.

[0148] For example, step S30 includes: extracting a target segment from at least one image material in response to a triggering operation of the editing process; and generating a video editing result according to the video editing template and the target segment.

[0149] Extracting the target segment from the at least one image material may, for example, involve performing highlight recognition on the at least one image material to obtain the target segment at the highlight moment of the at least one image material. The target segment at the highlight moment may, for example, be an optimal segment obtained by capturing the optimal portion of at least one of the at least one image material.

[0150] In this embodiment, at least one image material can be automatically identified to obtain a preferred segment of at least one image material to generate a video editing result.

[0151] In some embodiments of the present disclosure, step S30 includes: in response to a triggering operation of editing processing, if there is a first image material in at least one image material, dividing the first image material into multiple sub-segments, wherein the first image material is a material with a duration greater than a preset duration; and generating a video editing result based on the video editing template and the multiple sub-segments.

[0152] In each theme, the user can import at least one image material as shown in FIG1D . In at least one image material, there may be a first image material whose time length is greater than the preset time length. In the embodiment of the present disclosure, the first image material whose time length is greater than the preset time length can be segmented and processed, and the first image material is segmented into multiple sub-segments. These sub-segments are still used as materials for the theme corresponding to the first image material and are not used across themes to avoid user surprises. Dividing the first image material into multiple sub-segments is conducive to identifying and utilizing the sub-segments, at least partially avoiding the problem of long processing time caused by too long image materials, and is conducive to the flexible use of sub-segments.

[0153] In some embodiments of the present disclosure, for example, within a video editing interface, the video material and prompt information imported by the user are analyzed and identified in advance. Once an editing trigger is received, the video material is directly extracted (i.e., cropped) and segmented based on the analysis results, thereby reducing the waiting time for generating the video editing results. For example, analysis can begin as soon as the user imports one or more materials, without waiting for the user to trigger the editing process before analysis begins.

[0154] In some embodiments of the present disclosure, when a user starts to select materials for another theme, it is considered that the previous theme has been processed, and compatibility with user modifications needs to be considered.

[0155] In some embodiments of the present disclosure, the method further includes playing a preview video on a preview page based on the video editing results. Playing the preview video allows the user to view a draft video generated based on the image material, facilitating timely modification by the user when necessary. In some embodiments of the present disclosure, if the analysis result indicates that the prompt information includes a preset playback duration, the playback duration of at least part of the material is consistent with the preset playback duration.

[0156] For example, if the prompt information input by the user is "Learn 3 delicious dishes in 20 minutes", then the semantic analysis of the prompt information shows that the playback time of the video the user wants to produce is 20 minutes, and the length of the target image material extracted from the image material and used for the video editing result is also 20 minutes.

[0157] FIG4 shows a schematic diagram of a preview page provided by at least some embodiments of the present disclosure.

[0158] As shown in FIG. 4 , the preview page 401 includes a video playback window 411 , a material display window 421 , and an editing entry window 431 .

[0159] The video playback window 411 is used to display or play a preview video. The material display window 421 includes multiple material sub-windows for displaying at least a portion of each of the multiple image materials in the preview video. At least a portion of each of the multiple image materials is, for example, a target segment in each of the multiple image materials. For example, the multiple material sub-windows include a material sub-window 4211 for displaying Topic 1 and a material sub-window 4212 for displaying Topic 2. That is, in some embodiments of the present disclosure, each theme corresponds to a material sub-window for displaying the material used in the theme, and each material sub-window includes multiple material window units, each material window unit being used to display a material in the theme. As shown in Figure 4, four material window units are displayed in Topic 1, namely, material window unit A, material window unit B, material window unit C and material window unit D, and the four material window units respectively display different materials in Topic 1.

[0160] In some embodiments of the present disclosure, the material display window 421 further includes a material adding sub-window 4213 for adding materials. In this embodiment, the video editing method further includes: in response to a triggering operation on the material adding sub-window when a material window unit in the multiple material sub-windows 4213 is selected, displaying multiple materials for selection; and adding a target material after the material window unit according to the adding selection operation on the multiple materials for selection.

[0161] As shown in FIG4 , if a click is performed on the material add sub-window 4213 while material window unit B is selected, an album list is displayed to display multiple video materials for selection. If the user selects a target video material from the multiple video materials for selection, the target video material is added after material window unit B. For example, a material window unit E is added between material window unit B and material window unit C to display the target video material. For more information about the album list, please refer to the description above.

[0162] The editing entry window 431 is used to display the entry for editing the material. For example, the editing entry window 431 includes at least one of a subtitle entry 4311 for editing subtitles, a music entry 4312 for editing music, a filter entry 4313 for editing filters, and a text entry 4314 for editing text. The editing entry window 431 may also include entrances to other functions, which those skilled in the art can set in the editing entry window according to editing requirements. In some embodiments of the present disclosure, when clicking on the subtitle entry 4311 to enter the subtitle editing page, a dubbing editing entry for editing dubbing may be included in the subtitle editing page. The dubbing editing entry may be resident on the subtitle editing page, and the dubbing editing entry may be displayed regardless of whether dubbing is performed. If there is no dubbing in the video editing template, after clicking to enter the dubbing editing entry, dubbing is disabled by default.

[0163] In some embodiments of the present disclosure, for subtitle entry 4311, if the video editing template includes a tag for intelligently adding subtitles or a tag for intelligently adding dubbing, the conditions for intelligently adding subtitles to the preview video are met, and subtitle entry 4311 is displayed in editing entry window 431. If the template video does not include a tag for intelligently adding subtitles or a tag for intelligently adding dubbing, subtitle entry 4311 is not displayed in editing entry window 431.

[0164] In some embodiments of the present disclosure, some video editing templates may have only dubbing or only subtitles. When only dubbing is used, the subtitle entry 4311 is normally displayed in the editing entry window 431, but the "subtitle" display on the subtitle editing page is turned off by default, and the subtitles only serve to read the dubbing content. In some template videos, only subtitles may be used. The subtitle entry 4311 is normally displayed in the editing entry window 431, but the "dubbing" display on the subtitle editing page is turned off by default, and only the subtitles are displayed without reading or dubbing.

[0165] On the subtitle editing page, users can modify the subtitle content, set the subtitle duration, read the text aloud, and delete it.

[0166] 5A to 5F illustrate schematic diagrams of modifying subtitles provided by at least some embodiments of the present disclosure.

[0167] For example, as shown in Figure 5A, in the subtitle level 1 page (i.e., subtitle editing page) 501, the subtitles "abcd", "efg" and "higk" are displayed. If the subtitle "efg" is selected in the subtitle editing page 501, and the subtitle "efg" is also displayed on the screen at this time, the subtitle level 2 page 502 shown in Figure 5B is entered, or the subtitle content is directly clicked to enter the subtitle editing page 502. Click the "Edit" entry 512 in the subtitle editing page 502 to enter the subtitle content modification. In some embodiments of the present disclosure, the selected state of the subtitle follows the sliding operation to select different subtitles, and the selected subtitles are automatically displayed on the screen.

[0168] As shown in Figures 5C to 5D, when the subtitle content is modified, the input box 504 and the keyboard 503 are automatically pulled up to modify the subtitle content using the input box 504 and the keyboard 503. The "Delete" button on the upper screen is supported; or the subtitles are deleted in the deletion method of the subtitle secondary page. For example, as shown in Figure 5B, the delete button 522 is located at the last one of the subtitle secondary page. The subtitle position can be dragged, and after dragging, it is applied globally by default, and all subtitles are changed to the new position. For example, after the modification of the subtitle content is completed, it can be applied to the preview video by clicking the effective icon 505. In some embodiments of the present disclosure, after clicking the effective icon 505, a pop-up window 506 as shown in Figure 5D may appear, and the pop-up window 506 is used to ask the user whether to apply the subtitle to dubbing. If the user chooses to apply to dubbing, the dubbing will be modified to be consistent with the subtitle accordingly.

[0169] As shown in FIG. 5B and FIG. 5D , the subtitle secondary page 502 may further include a subtitle duration setting button 532 .

[0170] For example, after a user selects a corresponding subtitle and clicks subtitle duration setting button 532, a segment range 542 corresponding to the subtitle is displayed at the bottom center of the subtitle secondary page, supporting scalable selected duration ranges. Other subtitles 552 can be displayed above the segment range 542 corresponding to the subtitle. Selecting other subtitles 552 switches to displaying other subtitle duration ranges.

[0171] In some embodiments of the present disclosure, a "Change Batch" button 562 may be permanently displayed on the first-level subtitle page 501 and the second-level subtitle page 502. In response to a click on the "Change Batch" button 562, new subtitle content is regenerated. When sliding subtitles on the first-level subtitle page 501 or the second-level subtitle page 502, the "Change Batch" button 562 may disappear. When the sliding stops, the "Change Batch" button 562 continues to appear.

[0172] In some embodiments of the present disclosure, the first-level subtitle page 501 and the second-level subtitle page 502 may further include a first view icon for viewing the previous subtitle and a second view icon for viewing the next subtitle. Clicking the first view icon returns to the previous version, and clicking the second view icon continues generating new subtitles.

[0173] In some embodiments of the present disclosure, the subtitles level 1 page 501 and the subtitles level 2 page 502 may further include a subtitles switch entry 511. When subtitles are applied, the icon of the subtitles switch entry 511 is in a selected state, and the subtitles are displayed in the preview video. After clicking to close, the subtitles switch entry 511 becomes unselected, and the subtitles are hidden and not displayed in the preview video, and are only used for reading. When switching between the display and hidden states, a prompt "Subtitles are displayed" is displayed when clicking to expand, and a prompt "Subtitles are hidden" is displayed when clicking to hide.

[0174] As described above, as shown in FIG5F , the subtitle editing page 502 may include a dubbing editing entry 562 for editing the dubbing. Selecting dubbing editing entry 562 opens the dubbing editing page 507, where the timbre can be modified. The currently selected timbre is selected by default in the dubbing editing page 507. For example, in FIG5F , timbre 1 is the currently selected timbre. Clicking a different timbre will immediately preview its effect and read the current subtitle content aloud. Clicking the "Disable" button in the dubbing editing page disables the dubbing.

[0175] The dubbing volume and the original video sound volume can be set in the dubbing editing page. When there is dubbing, the original video sound defaults to 0.

[0176] In some embodiments of the present disclosure, you can choose whether to apply the set dubbing to the global context in the dubbing editing page. If the dubbing is applied to the global context, the new dubbing effect will take effect on all subtitles.

[0177] In some embodiments of the present disclosure, for example, if the user clicks the effective icon, the new dubbing modification is saved and the dubbing takes effect. If the effective icon is not clicked or the exit icon is clicked, the dubbing modification is not saved and the user returns to the previous page of the dubbing editing page.

[0178] In some embodiments of the present disclosure, clicking on music entry 4312 enters the music editing page. On the music editing page, music recommendations can be made based on similar music. For example, music with a similarity greater than a predetermined value to the video editing template can be recommended. Alternatively, music recommendations can be made based on an understanding of the second material and the prompt information.

[0179] In the music editing page, users can select music and set the volume of the music. In some embodiments of the present disclosure, the music can also be processed at the right time.

[0180] In some embodiments of the present disclosure, the editing entry window also includes a delete icon for selecting and deleting multiple preferred materials in the preview video. For example, the editing entry window 431 on the preview page includes a "Delete" entry located at the end of the editing entry window 431. Select a single material and click the "Delete" icon on the preview page to delete the selected material. After the selected material is deleted, the other materials in the theme containing the selected material are automatically postponed, and materials in other themes will not skip the theme and be postponed.

[0181] In some embodiments of the present disclosure, the editing entry window 431 and the material display window 421 support left and right sliding to display more editing entries and materials. For example, sliding the editing entry window 431 to the left displays a delete icon.

[0182] In some embodiments of the present disclosure, the editing entry window also includes a reorganization entry, and the method also includes: in response to the reorganization entry being selected, displaying a sorting panel, the sorting panel including the second material and prompt information; receiving an organization operation on the second material in the sorting panel, and reorganizing the second material.

[0183] For example, the preview page's edit entry window also includes a reorganization entry. After clicking the reorganization entry, a sorting panel is displayed. For example, the sorting panel has the same logic as the original structure framework, except that the sorting panel includes the previously imported second material and prompt information, which is otherwise the same as the structure framework, allowing for modification of prompt information and importing of second materials. In response to entering the sorting panel through the reorganization entry, the previously filled prompt information is displayed in the structure framework, and each second material is displayed according to each theme, and the copy is also displayed according to the original copy. Organization operations include, for example, modifying, adding, deleting, and dragging and sorting imported materials (with cross-theme sorting) as well as modifying prompt information. Modifications, additions, deletions, and dragging and sorting (with cross-theme sorting) of prompt information and imported materials in the structure framework are received, and after receiving the validation instruction, the above modifications are saved. When the prompt information and material filling content change, the algorithm material synthesis and all intelligent effects of the algorithm are regenerated (including: material splicing selection, overall packaging synthesis, intelligent copy rewriting, intelligent subtitle / dubbing generation, and intelligent music recommendation). After the validation is completed, the page returns to the previous level. If the user modifies the structure framework and receives an exit command, all modifications will be invalid.

[0184] In some embodiments of the present disclosure, in addition to reorganizing the entry, the order of the materials can also be reordered through special operations (for example, long pressing a certain clip material). For example, special operations can support cross-group ordering.

[0185] In some embodiments of the present disclosure, the material can also be intercepted. For example, the interception range of the material is positioned as the actual selected range by default, and the interception range can be adjusted instead of a fixed duration.

[0186] In some embodiments of the present disclosure, various operations such as adding and deleting materials can be performed on the preview page, or the video editing interface can be redisplayed through the reorganization entry in the preview page, in which operations such as adding, deleting, and sorting image materials can be performed, and the capture time range of a single image material can be adjusted. In the video editing interface redisplayed through the reorganization entry in the preview page, prompt information can also be modified.

[0187] When the video material changes (for example, adding, deleting, or reordering footage), the newly imported video material is analyzed to identify the target segments and segment the footage. The original footage (previously imported footage) retains its original processing effects and does not need to be reprocessed. If the original footage is reordered after importing the new footage, the new video parts are re-stitched. Based on the new video material, it is repackaged, including the corresponding filters, music, text, and transition effects. Please refer to the description above.

[0188] In some embodiments of the present disclosure, for newly imported image materials, subtitles and dubbing need to be regenerated. For original materials, the original processing effects are retained and there is no need to regenerate subtitles and dubbing. For deleted materials, the corresponding subtitles / dubbing are also deleted. For new image material sorting, subtitles and dubbing are regenerated according to the new sorting.

[0189] In some embodiments of the present disclosure, newly imported video footage does not require re-generating copy. The copy for existing footage retains its original processing effects and does not require re-processing. If the deletion of footage results in a theme completely lacking footage, the copy for that theme is also deleted. For newly ordered video footage, the copy placement is also re-processed based on the new order.

[0190] In some embodiments of the present disclosure, when only the clipping duration of the video material changes, the clip is selected based on the user's clipping range, no new highlight recognition is performed, and the video is reassembled based on the new clipping duration. Based on the latest clip duration, the video material is repackaged according to the music, filter, and text processing methods described above. For example, the music, filter, and text in the regenerated video editing results only change the playback duration and effective duration; the content itself remains unchanged. When only the clipping duration of the video material changes, new subtitles and dubbing are generated based on the new video material duration. If the duration of a single user clip is less than 3 seconds, no subtitles or dubbing are assigned to that clip. In this case, the existing subtitles and dubbing can be discarded because they are too short to be displayed and played. When only the clipping duration of the video material changes, there is no need to regenerate the text. For text that is effective for the global theme, the text display duration is based on the length of the video material. For text that is effective based on the theme portion, if the video material is longer, the text display is based on the original duration; if the video material is shorter, the text display is based on the shortened duration.

[0191] In some embodiments of the present disclosure, if the theme and style of the prompt information changes, all subtitles and / or dubbing will be refreshed.

[0192] In some embodiments of the present disclosure, when subtitles and dubbing change, if the subtitles are longer than before the modification, they will continue to be displayed according to the duration range of the subtitles before the modification, and the subtitles will automatically wrap, with the line wrap logic aligned with the current subtitle wrap. If the subtitles are shorter than before the modification, they will continue to be displayed according to the original duration range, with as many subtitles as possible displayed. If the subtitle duration range spans multiple segments, resulting in multiple subtitles overlapping, they can be displayed according to the re-effect.

[0193] When subtitles are too long, causing the corresponding dubbing to also be too long, the dubbing will be read at the length of the lengthened reading, supporting the full playback of the imported video material. When multiple dubbing audio clips overlap in a clip, all audio clips will be played together. When subtitles are shortened, causing the corresponding dubbing to also be shortened, the dubbing will be read at the length of the shortened reading.

[0194] In some embodiments of the present disclosure, there is a conflict between the algorithm re-synthesizing the video editing results and the user editing. For example, in the process of synthesizing the video editing results, the user performs editing operations on the material, dubbing, etc. When the material modified by the user again coincides with the material processed by the current algorithm, the material effect after the user's active modification is displayed first. For example, if the user has manually intercepted the duration range of a certain segment of material, it is selected according to its duration, rather than selecting the target segment as described above. If in the process of the algorithm re-synthesizing the video editing results, in response to the selection operation of the subtitle editing entrance or the dubbing editing entrance, a notification message is displayed on the preview page to inform the user that subtitles and dubbing are being generated. Wait for the video editing results to be re-synthesized, and the preview page will display normally and support operations. If in the process of the algorithm re-synthesizing the video editing results, the subtitles on the screen are selected, normal editing and modification are performed, and the intelligent subtitle progress is loaded normally. When the user confirms to save the changes, the subtitle content modified by the user in the segment is displayed first, and the intelligent generation progress of the subtitle segment is stopped. If the user does not save the changes, the intelligent subtitles are loaded and displayed normally. When the rewritten copy has not yet been generated, the original template copy is displayed. If resynthesis is triggered after secondary editing, the current copy is displayed when the new rewritten copy has not yet been generated. Users can edit and modify the content normally, and the rewritten copy is loaded normally, effectively displaying the user-modified text in that segment, and the intelligent rewriting progress of that segment is stopped.

[0195] In some embodiments of the present disclosure, a video editing template can be converted from a project file. For example, after exporting the project file through the editing tool, select "Publish Template". By adding the structural routine information of the video template in the publishing link and the vimo platform, and through the template editing characteristics, it is possible to identify at which nodes to split each group of shots and the corresponding themes of each group of shots. After the medium video template is published, it can be stored and managed in the vimo background. For example, by adding a medium video template identifier on vimo, it is determined whether the template is a medium video template type through manual identification by the operator. For example, the "medium video template" type is added to the material category, and a new attribute field of the "medium video template" type is added to the metadata.

[0196] For inputting template structure information, for example, on the Vimo platform, a "Modify Structure Information" entry is added, sequentially through the Template Management portal, Template Material Management portal, and Operation Details portal. The Modify Structure Information page is divided into three main sections: Overall Description, Intelligent Capabilities, and Template Structure Analysis. The Overall Description section provides an overall description of the video editing template, with no maximum character limit. When displayed on the application page, if the design range is exceeded, a fading mask will be displayed.

[0197] The Intelligent Capabilities section allows operators to select intelligent features, including intelligent subtitle and dubbing. The Template Structure Decomposition section automatically breaks down the template into multiple sections based on the template structure. Each section contains the following content: title, description, start and end time, start and end segments, maximum total length of user material, default padding, and default placeholder. The title has no character limit, and a fading overlay appears when the display exceeds the design range on the application page. The description also has no character limit, and a fading overlay appears when the display exceeds the design range on the application page. The start and end time sections describe the start and end seconds of the video clip. The start and end segments describe the start and end segments. The maximum total length of user material is used to account for relatively long imported materials. The algorithm can use this limit to minimize the length of the final composited material, achieving optimal results. Default padding selects a segment from the imported materials to fill in if no material is imported into the group, ensuring the integrity of the overall preview. The default placeholder means that no assets have been imported into this group. When you enter the preview page, the assets in this group are displayed as blank placeholders. After confirming the template structure, a secondary confirmation pop-up window will pop up. Clicking on the secondary confirmation pop-up window will officially take effect and launch the video editing template.

[0198] In some embodiments of the present disclosure, as long as a material is entered, the "Preview" button will light up, and you can enter the preview page. When the material is empty, the "Preview" button will be grayed out, and you cannot enter the preview page. The "Preview" button can be fixed at the bottom of the page and will not follow the scrolling. When other content exceeds one screen, you need to scroll down to view more.

[0199] Figure 6 is a schematic block diagram of a video editing device provided in some embodiments of the present disclosure. As shown in Figure 6, the video editing device 600 includes a first display unit 110, a second display unit 120, and a result generation unit 130. For example, the video editing device 600 can be used in a user terminal or any device or system that requires previewing design materials, and the embodiments of the present disclosure are not limited thereto.

[0200] The first display unit 110 is configured to display a video editing interface in response to a triggering operation on a video editing template, wherein the video editing template includes structural information and editing information, the structural information indicating at least one video segment, the editing information indicating at least one editing effect applied to the at least one video segment, and the video editing interface including segment editing areas corresponding to the at least one video segment. For example, the first display unit 110 may execute step S10 of the video editing method shown in FIG1A .

[0201] The second display unit 120 is configured to display an identifier of at least one image material in the paragraph editing area of ​​the target video paragraph in response to a material import operation for the target video paragraph, wherein the at least one video paragraph includes the target video paragraph, and the at least one image material is a material imported into the target video paragraph based on the material import operation. For example, the second display unit 120 can execute step S20 of the video editing method shown in Figure 1A.

[0202] The result generation unit 130 is configured to generate a video editing result based on the video editing template and the at least one video material in response to a triggering operation of the editing process, wherein the portion of the video editing result corresponding to the target video segment is an editing result in which a target editing effect is applied to the at least one video material, and the target editing effect matches an editing effect of the at least one editing effect located within the target video segment. For example, the page display unit 130 may execute step S30 of the video editing method shown in FIG1A .

[0203] The video editing device not only reduces the requirements for materials and makes the use of video templates more flexible, but also the method is applicable to medium videos in addition to short videos.

[0204] For example, the first display unit 110, the second display unit 120, and the result generation unit 130 can be hardware, software, firmware, or any feasible combination thereof. For example, the first display unit 110, the second display unit 120, and the result generation unit 130 can be dedicated or general-purpose circuits, chips, or devices, or can be a combination of a processor and memory. The embodiments of the present disclosure do not limit the specific implementation of the first display unit 110, the second display unit 120, and the result generation unit 130.

[0205] It should be noted that in the embodiments of the present disclosure, the various units of the video editing device 600 correspond to the various steps of the aforementioned video editing method. For the specific functions of the video editing device 600, reference can be made to the description of the video editing method above and will not be repeated here. The components and structure of the video editing device 600 shown in FIG6 are merely exemplary and non-restrictive. The video editing device 600 may further include other components and structures as needed.

[0206] FIG7 is a schematic block diagram of an electronic device provided in some embodiments of the present disclosure. As shown in FIG7 , the electronic device 200 includes a processor 210 and a memory 220. The memory 220 is used to store non-transitory computer-readable instructions (e.g., one or more computer program modules). The processor 210 is used to execute non-transitory computer-readable instructions, which, when executed by the processor 210, can execute one or more steps in the video editing method described above. The memory 220 and the processor 210 can be interconnected via a bus system and / or other forms of connection mechanisms (not shown).

[0207] For example, the processor 210 may be a central processing unit (CPU), a digital signal processor (DSP), or other processing units with data processing capabilities and / or program execution capabilities, such as a field programmable gate array (FPGA). For example, the central processing unit (CPU) may be an X86 or ARM architecture. The processor 210 may be a general-purpose processor or a dedicated processor, and may control other components in the electronic device 200 to perform desired functions.

[0208] For example, the memory 220 may include any combination of one or more computer program products, which may include various forms of computer-readable storage media, such as volatile memory and / or non-volatile memory. Volatile memory may include, for example, random access memory (RAM) and / or cache memory. Non-volatile memory may include, for example, read-only memory (ROM), a hard disk, an erasable programmable read-only memory (EPROM), a portable compact disk read-only memory (CD-ROM), a USB memory, a flash memory, etc. One or more computer program modules may be stored on the computer-readable storage medium, and the processor 210 may execute one or more computer program modules to implement various functions of the electronic device 200. Various applications and various data, as well as various data used and / or generated by the applications, may also be stored in the computer-readable storage medium.

[0209] It should be noted that, in the embodiment of the present disclosure, the specific functions and technical effects of the electronic device 200 can be referred to the above description of the video editing method, which will not be repeated here.

[0210] Figure 8 is a schematic block diagram of another electronic device provided in some embodiments of the present disclosure. This electronic device 300 is, for example, suitable for implementing the video editing method provided in embodiments of the present disclosure. The electronic device 300 may be a user terminal, etc. It should be noted that the electronic device 300 shown in Figure 8 is merely an example and does not impose any limitations on the functionality and scope of use of the embodiments of the present disclosure.

[0211] As shown in FIG8 , the electronic device 300 may include a processing device (e.g., a central processing unit, a graphics processing unit, etc.) 310, which can perform various appropriate actions and processes according to a program stored in a read-only memory (ROM) 320 or a program loaded from a storage device 380 into a random access memory (RAM) 330. Various programs and data required for the operation of the electronic device 300 are also stored in the RAM 330. The processing device 310, the ROM 320, and the RAM 330 are connected to each other via a bus 340. An input / output (I / O) interface 350 is also connected to the bus 340.

[0212] Typically, the following devices may be connected to the I / O interface 350: an input device 360 ​​including, for example, a touch screen, a touchpad, a keyboard, a mouse, a camera, a microphone, an accelerometer, a gyroscope, etc.; an output device 370 including, for example, a liquid crystal display (LCD), a speaker, a vibrator, etc.; a storage device 380 including, for example, a magnetic tape, a hard disk, etc.; and a communication device 390. The communication device 390 may allow the electronic device 300 to communicate with other electronic devices wirelessly or by wire to exchange data. Although FIG9 shows the electronic device 300 with various devices, it should be understood that it is not required to implement or have all of the devices shown, and the electronic device 300 may alternatively implement or have more or fewer devices.

[0213] For example, according to an embodiment of the present disclosure, the video editing method shown in Figure 1A can be implemented as a computer software program. For example, an embodiment of the present disclosure includes a computer program product, which includes a computer program carried on a non-transitory computer-readable medium, and the computer program includes program code for executing the above-mentioned video editing method. In such an embodiment, the computer program can be downloaded and installed from the network through the communication device 390, or installed from the storage device 380, or installed from the ROM 320. When the computer program is executed by the processing device 310, the functions defined in the video editing method provided by the embodiment of the present disclosure can be executed.

[0214] At least one embodiment of the present disclosure further provides a storage medium for storing non-transitory computer-readable instructions. When executed by a computer, these non-transitory computer-readable instructions can implement the video editing method described in any embodiment of the present disclosure. Utilizing this storage medium not only reduces the requirements for source material, making the use of video templates more flexible, but also allows the method to be applied not only to short videos but also to medium-length videos.

[0215] FIG9 is a schematic diagram of a storage medium provided by some embodiments of the present disclosure. As shown in FIG9 , a storage medium 400 is used to store non-transitory computer-readable instructions 410. For example, when non-transitory computer-readable instructions 410 are executed by a computer, one or more steps of the video editing method described above may be performed.

[0216] For example, the storage medium 400 can be applied to the electronic device 200 described above. For example, the storage medium 400 can be the memory 220 in the electronic device 200 shown in FIG7 . For example, the relevant description of the storage medium 400 can refer to the corresponding description of the memory 220 in the electronic device 200 shown in FIG7 , and will not be repeated here.

[0217] The video editing method, video editing device, electronic device, and storage medium provided by the embodiments of the present disclosure are described above in conjunction with Figures 1A to 9. The video editing method provided by the embodiments of the present disclosure not only reduces the requirements for the material, making the use of video templates more flexible, but also is applicable to medium-length videos in addition to short videos.

[0218] It should be noted that the storage medium (computer-readable medium) mentioned above in the present disclosure may be a computer-readable signal medium or a non-transitory computer-readable storage medium or any combination of the two. Non-transitory computer-readable storage media may be, for example, but not limited to, electrical, magnetic, optical, electromagnetic, infrared, or semiconductor systems, devices or devices, or any combination of the above. More specific examples of non-transitory computer-readable storage media may include, but are not limited to: an electrical connection with one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the above. In the present disclosure, a non-transitory computer-readable storage medium may be any tangible medium containing or storing a program that can be used by or in conjunction with an instruction execution system, device or device. In the present disclosure, a computer-readable signal medium may include a data signal propagated in baseband or as part of a carrier wave, which carries a computer-readable program code. Such propagated data signals may take a variety of forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination thereof. A computer-readable signal medium may also be any computer-readable medium other than a non-transitory computer-readable storage medium that can send, propagate, or transmit a program for use by or in conjunction with an instruction execution system, apparatus, or device. The program code contained on the computer-readable medium may be transmitted using any suitable medium, including but not limited to wires, optical cables, RF (radio frequency), etc., or any suitable combination thereof.

[0219] In some embodiments, the client and server can communicate using any currently known or future developed network protocol, such as the Hyper Text Transfer Protocol (HTTP), and can be interconnected with any form or medium of digital data communication (e.g., a communication network). Examples of communication networks include local area networks (LANs), wide area networks (WANs), internetworks (e.g., the Internet), and peer-to-peer networks (e.g., ad hoc peer-to-peer networks), as well as any currently known or future developed networks.

[0220] The computer-readable medium may be included in the electronic device, or may exist independently without being incorporated into the electronic device.

[0221] The computer-readable medium carries one or more programs, and when the one or more programs are executed by the electronic device, the electronic device performs one or more steps in the video editing method described above. The computer program code for performing the operations of the present disclosure can be written in one or more programming languages ​​or a combination thereof, and the programming languages ​​include but are not limited to object-oriented programming languages, such as Java, Smalltalk, C++, and also conventional procedural programming languages, such as "C" language or similar programming languages. The program code can be executed entirely on the user's computer, partially on the user's computer, as an independent software package, partially on the user's computer and partially on a remote computer, or entirely on a remote computer or server. In the case of a remote computer, the remote computer can be connected to the user's computer through any type of network, such as a local area network (LAN) or a wide area network (WAN), or can be connected to an external computer (for example, using an Internet service provider to connect via the Internet).

[0222] The flowcharts and block diagrams in the accompanying drawings illustrate the possible implementation architecture, functions and operations of the systems, methods and computer program products according to various embodiments of the present disclosure. In this regard, each box in the flowchart or block diagram can represent a module, program segment, or a part of code, and the module, program segment, or a part of code contains one or more executable instructions for realizing the specified logical function. It should also be noted that in some alternative implementations, the functions marked in the box can also occur in a different order than that marked in the accompanying drawings. For example, two boxes represented in succession can actually be executed substantially in parallel, and they can sometimes be executed in the opposite order, depending on the functions involved. It should also be noted that each box in the block diagram and / or flowchart, and the combination of the boxes in the block diagram and / or flowchart, can be implemented using a dedicated hardware-based system that performs the specified function or operation, or can be implemented using a combination of dedicated hardware and computer instructions.

[0223] The units involved in the embodiments described in this disclosure may be implemented in software or hardware, wherein the name of a unit does not necessarily limit the unit itself.

[0224] The functions described above herein may be performed, at least in part, by one or more hardware logic components. For example, and without limitation, exemplary types of hardware logic components that may be used include: field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), systems on chip (SOCs), complex programmable logic devices (CPLDs), and the like.

[0225] In the present disclosure, a machine-readable medium may be a tangible medium that may contain or store a program for use by or in conjunction with an instruction execution system, device, or apparatus. A machine-readable medium may be a machine-readable signal medium or a machine-readable storage medium. A machine-readable medium may include, but is not limited to, an electronic, magnetic, optical, electromagnetic, infrared, or semiconductor system, device, or apparatus, or any suitable combination of the foregoing. More specific examples of machine-readable storage media may include an electrical connection based on one or more lines, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the foregoing.

[0226] The above description is only a partial embodiment of the present disclosure and an illustration of the technical principles used. Those skilled in the art should understand that the scope of disclosure involved in the present disclosure is not limited to the technical solutions formed by the specific combination of the above-mentioned technical features, but should also cover other technical solutions formed by any combination of the above-mentioned technical features or their equivalents without departing from the above-mentioned disclosed concepts. For example, the above-mentioned features can be replaced with (but not limited to) technical features with similar functions disclosed in this disclosure.

[0227] In addition, although each operation is described in a specific order, this should not be understood as requiring these operations to be performed in the specific order shown or in a sequential order. Under certain circumstances, multitasking and parallel processing may be advantageous. Similarly, although some specific implementation details have been included in the above discussion, these should not be interpreted as limiting the scope of the present disclosure. Some features described in the context of a separate embodiment can also be implemented in a single embodiment in combination. On the contrary, the various features described in the context of a single embodiment can also be implemented in multiple embodiments individually or in any suitable sub-combination mode.

[0228] Although the subject matter has been described in language specific to structural features and / or methodological logical acts, it should be understood that the subject matter defined in the appended claims is not necessarily limited to the specific features or acts described above. Rather, the specific features and acts described above are merely example forms of implementing the claims.

Claims

1. A video editing method, comprising: In response to a triggering operation on a video editing template, presenting a video editing interface, wherein the video editing template includes structure information and editing information, the structure information is used to indicate at least one video segment, the editing information is used to indicate at least one editing effect applied in the at least one video segment, and the video editing interface includes paragraph editing areas respectively corresponding to the at least one video segment; In response to a material import operation for a target video segment, presenting identifiers of at least one video material in the paragraph editing area of the target video segment, the at least one video segment includes the target video segment, and the at least one video material is the material imported into the target video segment based on the material import operation; and In response to a triggering operation for editing processing, generating a video editing result according to the video editing template and the at least one video material, wherein a part of the video editing result corresponding to the target video segment is an editing result obtained based on a target editing effect and the at least one video material, and the target editing effect matches the editing effect located in the target video segment among the at least one editing effect.

2. The method according to claim 1, wherein The generating a video editing result according to the video editing template and the at least one video material in response to a triggering operation for editing processing includes: In response to the triggering operation for editing processing, adjusting the editing effect according to a matching strategy to obtain the target editing effect; According to the video editing template, applying the target editing effect to at least a part of the at least one video material to obtain a video part of the target video segment; and Generating the video editing result based on the video part corresponding to each target video segment in the at least one video segment.

3. The method according to claim 2, wherein The matching strategy includes: When the editing effect includes a first copywriting, the target editing effect includes a second copywriting, wherein the copywriting format of the second copywriting is the same as that of the first copywriting, and the copywriting content of the second copywriting matches the at least one video material.

4. The method according to claim 3, wherein, The matching strategy further includes: In response to the first copywriting being applied to the entire paragraph of the target video segment, the second copywriting is applied to the entire part of the video part; In response to the first copywriting being applied to a partial time period in the target video segment, the second copywriting is applied to a partial time period in the video part.

5. The method according to claim 3, wherein The generating a video editing result according to the video editing template and the at least one video material in response to a triggering operation for editing processing further includes: Obtaining the content of the second copywriting, wherein the structure information further includes an information prompt paragraph for obtaining prompt information, and obtaining the content of the second copywriting includes: Analyzing the prompt information and the at least one video material to obtain an analysis result; and Generating the content of the second copywriting according to the analysis result.

6. The method according to claim 5, wherein, The degree of influence of the analysis result on the content of the second copywriting is greater than the degree of influence of the at least one video material on the content of the second copywriting.

7. The method according to claim 2, wherein The matching strategy includes: When the editing effect includes a transition animation, the target editing effect includes the transition animation.

8. The method according to claim 2, wherein, The structure information further includes an information prompt paragraph for obtaining prompt information. The adjusting the editing effect according to the matching strategy to obtain the target editing effect in response to the trigger operation of the editing process includes: Analyzing the prompt information and the at least one video material in response to the trigger operation of the editing process to obtain an analysis result; In response to the editing effect including subtitles and / or dubbing, obtaining the content of the subtitles and / or dubbing in the target editing effect according to the matching strategy of matching the content of the subtitles and / or dubbing in the target editing effect with the analysis result; Obtaining the format of the subtitles in the target editing effect according to the matching strategy that the format of the subtitles in the target editing effect is the same as the format of the subtitles in the editing effect; and Obtaining the sound parameters of the dubbing in the target editing effect according to the matching strategy that the sound parameters of the dubbing in the target editing effect are the same as the sound parameters of the dubbing in the editing effect.

9. The method according to claim 2, wherein, The generating the video editing result based on each video part corresponding to each target video paragraph in the at least one video paragraph includes: Splicing each video part corresponding to each target video paragraph to obtain the video editing result.

10. The method according to claim 1, wherein, The generating the video editing result according to the video editing template and the at least one video material in response to the trigger operation of the editing process includes: Processing the at least one video material and the editing effect by using a neural network in response to the trigger operation of the editing process to obtain the video editing result.

11. The method according to claim 2, wherein, The at least one editing effect further includes a global editing effect applied to the at least one video paragraph. The generating the video editing result according to the video editing template and the at least one video material in response to the trigger operation of the editing process further includes: Adjusting the time length of the global editing effect to match the time length of the target video material used in the video editing result in response to the trigger operation of the editing process to obtain an adjusted global editing effect; The generating the video editing result based on each video part corresponding to each target video paragraph in the at least one video paragraph includes: Generating the video editing result based on the adjusted global editing effect and each video part corresponding to each target video paragraph.

12. The method according to claim 10 or 11, wherein The global editing effect includes at least one of music, filters, and subtitles.

13. The method according to claim 1, wherein The structure information indicates each of the at least one video paragraph by a time interval or the serial number of consecutive video segments.

14. The method according to claim 1, wherein The generating the video editing result according to the video editing template and the at least one video material in response to the trigger operation of the editing process includes: In response to the triggering operation of the editing process, extract a target segment from the at least one video material; and Generate the video editing result according to the video editing template and the target segment.

15. The method according to claim 1, wherein The generating the video editing result in response to the triggering operation of the editing process according to the video editing template and the at least one video material includes: In response to the triggering operation of the editing process, if there is a first video material among the at least one video material, divide the first video material into multiple sub-segments, where the first video material is a material with a duration longer than a preset duration; and Generate the video editing result based on the video editing template and the multiple sub-segments.

16. The method according to any one of claims 1-15, further comprising: Play a preview video on a preview page according to the video editing result.

17. The method according to claim 16, wherein When the editing effect only includes subtitles, the dubbing identifier on the preview page is displayed in a closed state, and when the editing effect only includes dubbing, the subtitle identifier on the preview page is displayed in the closed state.

18. The method according to claim 16 or 17, wherein The preview page includes a video playback window, a material display window, and an editing entry window, The video playback window is used to play the preview video, the material display window includes multiple material sub-windows for displaying at least a part of each of the multiple video materials in the preview video, and the editing entry window is used to display an entry for editing the video material.

19. The method according to claim 18, wherein, The material display window further includes a material addition sub-window, The method further comprises: In response to a triggering operation on the material addition sub-window when a material window unit in the multiple material sub-windows is selected, display multiple selectable materials; Add a target material after the material window unit according to an addition selection operation on the multiple selectable materials.

20. The method according to claim 18, wherein, The editing entry window includes at least one of a subtitle entry for editing the subtitle, a filter entry for editing the filter, a copywriting entry for editing the copywriting, and a music entry for editing the music, The editing entry window further includes: a delete icon for selecting and deleting multiple preferred materials in the preview video.

21. The method according to claim 20, wherein, The editing entry window further includes a reorganization entry, and the method further comprises: In response to the reorganization entry being selected, display a sorting panel, where the sorting panel includes the at least one video material; and Receive an organization operation on the at least one video material in the sorting panel, and reorganize the at least one video material.

22. A video editing device, comprising: A first display unit configured to display a video editing interface in response to a trigger operation on a video editing template, wherein the video editing template includes structure information and editing information, the structure information is used to indicate at least one video segment, and the editing information is used to indicate at least one editing effect applied to the at least one video segment, and the video editing interface includes paragraph editing areas corresponding to the at least one video segment respectively; A second display unit configured to display identifiers of at least one video material in the paragraph editing area of the target video segment in response to a material import operation for the target video segment, the at least one video segment includes the target video segment, and the at least one video material is the material imported into the target video segment based on the material import operation; and A result generation unit configured to generate a video editing result according to the video editing template and the at least one video material in response to a trigger operation for editing processing, wherein the part of the video editing result corresponding to the target video segment is an editing result obtained based on a target editing effect and the at least one video material, and the target editing effect matches the editing effect located in the target video segment among the at least one editing effect.

23. An electronic device, comprising: A processor; A memory including one or more computer program instructions; Wherein, the one or more computer program instructions are stored in the memory and, when executed by the processor, implement the video editing method according to any one of claims 1-21.

24. A computer-readable storage medium that non-temporarily stores computer-readable instructions, wherein, When the computer-readable instructions are executed by the processor, the video editing method according to any one of claims 1-21 is implemented.

Citation Information

Patent Citations

  • Short video generation method and device

    CN111866587A

  • Video clip synthesis method and electronic equipment

    CN113891113A

  • Video editing method, video editing device and storage medium

    CN114697700A

  • Video processing method and device, equipment and storage medium

    CN115250335A

  • Video editing method and device

    JP2007318450A