Video editing method and device, electronic equipment and storage medium
The video paragraphs and editing effects are displayed through the video editing template, allowing users to display video materials in the paragraph editing area and generate video editing results. This solves the problem of high material requirements for existing video templates, and realizes the flexible use of video templates and the effect applicable to short and medium videos.
Patent Information
- Application Number
- CN202311843312.4
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2023-12-28
- Publication Date
- 2025-07-01
AI Technical Summary
Existing video templates have high requirements for materials, many restrictions, and are not flexible in use. Especially for short videos, users need to import the number and duration of materials consistent with the template, which cannot meet the needs of medium videos.
Provides a video editing method, displaying video paragraphs and editing effects through video editing templates, allowing users to display video materials in the paragraph editing area, and generate video editing results based on the templates and materials, reducing the number and duration requirements for materials, and is suitable for short and medium videos.
It realizes the flexible use of video templates, reduces the limitations on materials, makes the video editing method suitable for short and medium videos, and improves the user's editing freedom and efficiency.
Smart Images

Figure CN120238696A_ABST
Abstract
Description
Technical Field
[0001] Embodiments of the present disclosure relate to a video editing method and apparatus, an electronic device, and a storage medium. Background Art
[0002] Video content production can be seen everywhere in daily life. Users record their lives, showcase their personalities, and output values by producing video content. There are usually two ways to produce videos. One is to record videos by oneself and edit them to perfection. The other is to produce videos through video templates. Since producing videos through video templates is convenient, fast, and has rich effects, it has become the main way to share video content. Summary of the Invention
[0003] Most current video templates have fixed slots. Users need to import the same number of materials as the video template, and the number and duration of video clips also need to be consistent with the video template. Therefore, video templates have high requirements and many restrictions on the materials imported by users, and the use of video templates is not flexible. To address the above problems, at least one embodiment of the present disclosure provides a video editing method and apparatus, an electronic device, and a storage medium, which can reduce the requirements and restrictions of video templates on the materials imported by users and improve the flexibility of using video templates.
[0004] At least one embodiment of the present disclosure provides a video editing method, including: in response to a trigger operation on a video editing template, displaying a video editing interface, where the video editing template includes structure information and editing information, the structure information is used to indicate at least one video segment, the editing information is used to indicate at least one editing effect applied in the at least one video segment, and the video editing interface includes paragraph editing areas respectively corresponding to the at least one video segment; in response to a material import operation for a target video segment, displaying identifiers of at least one video material in the paragraph editing area of the target video segment, where the at least one video segment includes the target video segment, and the at least one video material is the material imported into the target video segment based on the material import operation; and in response to a trigger operation for editing processing, generating a video editing result according to the video editing template and the at least one video material, where a part of the video editing result corresponding to the target video segment is an editing result obtained based on a target editing effect and at least one video material, and the target editing effect matches the editing effect located in the target video segment among the at least one editing effect.
[0005] At least one embodiment of the present disclosure further provides a video editing device, including: a first display unit configured to display a video editing interface in response to a trigger operation on a video editing template, where the video editing template includes structure information and editing information, the structure information is used to indicate at least one video segment, the editing information is used to indicate at least one editing effect applied to the at least one video segment, and the video editing interface includes paragraph editing areas corresponding to the at least one video segment respectively; a second display unit configured to display identifiers of at least one video material in the paragraph editing area of the target video segment in response to a material import operation for the target video segment, the at least one video segment includes the target video segment, and the at least one video material is the material imported into the target video segment based on the material import operation; and a result generation unit configured to generate a video editing result according to the video editing template and the at least one video material in response to a trigger operation for editing processing, where a part of the video editing result corresponding to the target video segment is an editing result obtained based on a target editing effect and the at least one video material, and the target editing effect matches the editing effect located in the target video segment among the at least one editing effect.
[0006] At least one embodiment of the present disclosure further provides an electronic device, including: a processor; a memory including one or more computer program modules; where the one or more computer program modules are stored in the memory and configured to be executed by the processor, and the one or more computer program modules include instructions for implementing the video editing method according to any embodiment of the present disclosure.
[0007] At least one embodiment of the present disclosure further provides a storage medium for storing non-temporary computer-readable instructions, which can implement the video editing method according to any embodiment of the present disclosure when executed by a computer. Description of the Drawings
[0008] Combined with the drawings and referring to the following specific embodiments, the above and other features, advantages and aspects of the embodiments of the present disclosure will become more obvious. Throughout the drawings, the same reference numerals represent the same elements. It should be understood that the drawings are schematic and the original elements and elements are not necessarily drawn to scale.
[0009] Figure 1A A flowchart showing a video editing method provided by at least one embodiment of the present disclosure;
[0010] Figure 1B A schematic diagram showing a video editing method provided by at least one embodiment of the present disclosure;
[0011] Figure 1C A schematic diagram showing a kind of structural information provided by at least one embodiment of the present disclosure;
[0012] Figure 1D A schematic diagram showing a kind of video editing interface provided by at least one embodiment of the present disclosure;
[0013] Figure 2 A schematic diagram showing a kind of editing effect obtained by adjusting the editing effect in a video editing template to the editing effect in a video editing result provided by some embodiments of the present disclosure;
[0014] Figure 3 Showing what is provided by at least one embodiment of the present disclosure Figure 1A The method flow chart of step S30 in;
[0015] Figure 4 A schematic diagram showing a kind of preview page provided by at least some embodiments of the present disclosure;
[0016] Figures 5A to 5F A schematic diagram showing a kind of subtitle modification provided by at least some embodiments of the present disclosure;
[0017] Figure 6 A schematic block diagram of a video editing device provided by some embodiments of the present disclosure;
[0018] Figure 7 A schematic block diagram of an electronic device provided by some embodiments of the present disclosure;
[0019] Figure 8 A schematic block diagram of another electronic device provided by some embodiments of the present disclosure; and
[0020] Figure 9 A schematic diagram of a storage medium provided by some embodiments of the present disclosure. Detailed implementation manners
[0021] Embodiments of the present disclosure will be described in more detail below with reference to the accompanying drawings. Although some embodiments of the present disclosure are shown in the drawings, it should be understood that the present disclosure can be implemented in various forms and should not be construed as limited to the embodiments set forth herein. On the contrary, these embodiments are provided to more thoroughly and completely understand the present disclosure. It should be understood that the drawings and embodiments of the present disclosure are only for exemplary purposes and are not used to limit the protection scope of the present disclosure.
[0022] It should be understood that the various steps recorded in the method embodiments of the present disclosure can be executed in different orders and / or executed in parallel. In addition, the method embodiments may include additional steps and / or omit the steps shown. The scope of the present disclosure is not limited in this regard.
[0023] As used herein, the term "including" and its variations are open-ended, i.e., "including but not limited to". The term "based on" means "at least partially based on". The term "one embodiment" means "at least one embodiment"; the term "another embodiment" means "at least one additional embodiment"; the term "some embodiments" means "at least some embodiments". Relevant definitions of other terms will be given in the following description.
[0024] It should be noted that the concepts such as "first", "second", etc. mentioned in this disclosure are only used to distinguish different devices, modules or units, and are not used to limit the order or interdependence of the functions performed by these devices, modules or units.
[0025] It should be noted that the modifications of "one" and "a plurality of" mentioned in this disclosure are illustrative rather than restrictive. Those skilled in the art should understand that, unless clearly stated otherwise in the context, it should be understood as "one or more". "A plurality of" should be understood as two or more.
[0026] The names of the messages or information exchanged between multiple devices in the embodiments of this disclosure are only for illustrative purposes and are not used to limit the scope of these messages or information.
[0027] With the enrichment of video template effects and the increase in types, the requirements and scenarios for video production and sharing based on video templates will increase day by day. However, most of the current video templates are for short videos. Video templates for short videos have fixed slots, and users need to import the same number of materials as the video template, and the number and duration of video clips also need to be consistent with the video template. Therefore, video templates for short videos have high requirements and many restrictions on the materials imported by users, and the use of video templates is not flexible.
[0028] At least one embodiment of the present disclosure provides a method for generating a video using a video template. This method not only reduces the requirements for materials, making the use of video templates more flexible, but also applies to medium-length videos in addition to short videos. Short videos generally refer to user-generated videos with a duration of less than 1 minute, that is, User-generated Content (UGC). Long videos generally refer to videos with a duration of more than 30 minutes produced by professional structures, with higher content quality, that is, Professionally-generated Content (PGC). Medium-length videos are between short videos and long videos. Although they are user-generated original content, the professional level of users is a bit higher than that of short videos, that is, Professionally-User-generated (PUGC), and the duration of medium-length videos is usually also between short videos and long videos. From the user's perspective, short videos do not require setting aside specific time, convey fragmented content, and can obtain the key content of the video in a short time, but it may not be retained. Medium-length videos and long videos require finding a suitable time and place to watch, and more effort needs to be invested, and the video content can be retained in memory for a longer time.
[0029] It should be noted that in the embodiments of the present disclosure, the video template and the template video have the same meaning, both referring to the template used when making a video, and this template is presented in the form of a video.
[0030] At least one embodiment of the present disclosure provides a video editing method, a video editing device, an electronic device, and a computer-readable storage medium. The video editing method includes: in response to a trigger operation on a video editing template, presenting a video editing interface, where the video editing template includes structure information and editing information, the structure information is used to indicate at least one video segment, the editing information is used to indicate at least one editing effect applied to the at least one video segment, and the video editing interface includes paragraph editing regions respectively corresponding to the at least one video segment; in response to a material import operation for a target video segment, presenting identifiers of at least one video material in the paragraph editing region of the target video segment, the at least one video segment includes the target video segment, and the at least one video material is the material imported into the target video segment based on the material import operation; and in response to a trigger operation for editing processing, generating a video editing result according to the video editing template and the at least one video material, where the part of the video editing result corresponding to the target video segment is the editing result with the target editing effect applied to the at least one video material, and the target editing effect matches the editing effect located in the target video segment among the at least one editing effect. This video editing method does not limit the matching of the at least one imported video material with the video editing template, but applies the editing effect of the video editing template to the at least one imported video material, and has no requirements for the quantity and time length of the at least one imported video material. It not only reduces the requirements for video materials and makes the use of video editing templates more flexible, but also this method is applicable to medium-length videos in addition to short videos.
[0031] Next, embodiments of the present disclosure will be described in detail with reference to the accompanying drawings.
[0032] Figure 1A The flowchart shows a video editing method provided by at least one embodiment of the present disclosure; Figure 1B The diagram shows a video editing method provided by at least one embodiment of the present disclosure; Figure 1C The diagram shows a structure information provided by at least one embodiment of the present disclosure; Figure 1D The diagram shows a video editing interface provided by at least one embodiment of the present disclosure.
[0033] As Figure 1A shown, in at least one embodiment, the method includes the following operations.
[0034] Step S10: In response to a trigger operation on a video editing template, presenting a video editing interface, where the video editing template includes structure information and editing information, the structure information is used to indicate at least one video segment, the editing information is used to indicate at least one editing effect applied to the at least one video segment, and the video editing interface includes paragraph editing regions respectively corresponding to the at least one video segment.
[0035] Step S20: In response to a material import operation for a target video segment, display identifiers of at least one video material in the segment editing area of the target video segment. The at least one video segment includes the target video segment, and the at least one video material is the material imported into the target video segment based on the material import operation.
[0036] Step S30: In response to a trigger operation for editing processing, generate a video editing result according to a video editing template and at least one video material. Among them, the part of the video editing result corresponding to the target video segment is the editing result with the target editing effect applied to at least one video material, and the target editing effect matches the editing effect located in the target video segment among at least one editing effect.
[0037] For example, in Step S10, the trigger operation for the video editing template is, for example, a click operation on the usage icon on the main page of the video editing template. As Figure 1B shown, the main page of the video editing template is Page 101, and this Page 101 includes a usage icon 102. If the client receives a click operation on the usage icon 102, display Figure 1D the video editing interface 103 shown.
[0038] For example, this video editing method is applied to a video application. Covers of multiple video editing templates are displayed on the template recommendation page of the video application. If the user performs a selection operation on the cover of a certain video editing template, play the display video of this video editing template for the user to refer to. When the display video ends, enter the main page of this video editing template. The user can perform operations on the main page (for example, Page 101) to use this video editing template or not use this video editing template. If the user performs a trigger operation on the main page, it means that the user selects to use this video editing template, and then enters the video editing interface 103 of this video editing template to perform video editing in the video editing interface 103.
[0039] In some embodiments of the present disclosure, the video editing template includes structure information and editing information. The structure information is used to indicate at least one video segment, and the editing information is used to indicate at least one editing effect applied in an indicated video segment. The editing information can, for example, include at least one of subtitles, dubbing, background music, copywriting, filters, and transition animations.
[0040] As Figure 1C shown, the video editing template includes structure information 100. This structure information 100 is, for example, set by the producer of the video editing template. For example, the publishing link and the vimo platform support adding video editing template structure information. For example, add the structure information 100 shown in Figure 1 through the vimo platform.
[0041] As Figure 1C shown, the structure information 100 includes three major parts: an overall description, intelligent capabilities, and template structure breakdown.
[0042] The template structure breakdown includes a first part 1 and allows for the addition of more parts. The first part 1 represents a video segment, and information such as the title and description of this video segment can be filled in by the operator. For example, by clicking the control 2 for adding more parts, a second part, a third part, etc. can be added, and each part represents a video segment.
[0043] The overall description part is, for example, controlled by the operator himself / herself to fill in the description of, for example, the content and style of the video editing template. The overall description part does not limit the maximum number of input characters, for example.
[0044] The intelligent capabilities, for example, support the operator to check. For example, they include intelligent subtitle addition, intelligent dubbing addition, etc. The operator can choose whether to check the intelligent capabilities and which one or more intelligent capabilities to check. If a certain intelligent capability is checked during operation, the video editing template has this intelligent capability.
[0045] The structure of each part in the structure information is similar. For example, the first part 1 is used as an example for illustration. For example, the first part 1 includes a title, a description, start and end times, and start and end segments. The title is, for example, a refined description of the template video material in the first part 1 of the video editing template. The template video material is, for example, the material used in the video editing template. The description is, for example, a general description of the template video material in the first part 1.
[0046] In some embodiments of the present disclosure, the structure information indicates each of at least one video segment through a time interval or the serial number of consecutive video segments. For example, in Figure 1C the structure information, each video segment is determined by the start and end times or by the start and end segments.
[0047] For example, the template video material is at least one consecutive video segment, the start and end times are, for example, selecting from the at least one consecutive video segment the time interval from the Kth second to the Lth second as the material for the first part 1, and the start and end segments are, for example, selecting from the at least one consecutive video segment the jth segment to the ith segment as the material for the first part 1. j and i are examples of the serial numbers of the video segments.
[0048] It should be noted that the division of the start and end times and the start and end segments in the structure information is for the understanding of the video editing template to facilitate creating a finished video for the user, while the material imported by the user does not necessarily follow the division of the structure information. For example, the user can import materials in any form, and then perform processing such as splitting and high - light slicing on the materials imported by the user according to the structure information.
[0049] In the above embodiments, the video editing template includes template video materials. In some other embodiments of the present disclosure, the video editing template may not include template video materials.
[0050] In some embodiments of the present disclosure, the production and release processes of the short video template can be directly reused for the video editing template of medium videos. After exporting through the editing tool project file, the video editing template can be obtained by selecting "Publish Template".
[0051] After the medium video template is published, it supports storage and management in the vimo background and reuses the short video template logic. For example, it is completed by manually marking during the operation on vimo to determine whether the template is of the "medium video template" type. Vimo supports batch addition of the identifier of whether it is of the "medium video template" type. For the medium video template type, an attribute field of the "medium video template" type needs to be newly added to the metadata, and the video types are distinguished through the type attribute field identifier (for example, identified as video_type_id). Those marked as the medium video template type perform the medium video template logic throughout the overall template link.
[0052] For editing information, for example, similar to the short video template, producers can add editing information such as text information, dubbing information, music information, sticker information, filter information, and transition animation information according to their own ideas and needs. These editing information indicate the editing effects, and the editing effects include at least one of subtitles, dubbing, background music, text, filters, and transition animations.
[0053] As Figure 1D shown, the video editing interface 103 includes paragraph editing areas corresponding to at least one video paragraph respectively. Each paragraph editing area serves as an editing area for a theme and displays video clips of different themes respectively. For example, Figure 1C the paragraph editing area 11 corresponding to the video paragraph indicated by the first part 1 in
[0054] If the content entered in the title of the first part 1 is "Finished Dishes", then the title of the paragraph editing area 11 corresponding to the video paragraph indicated by the first part 1 is "Finished Dishes". For example, the structure information further includes a second part. If the content entered in the title of the second part is "Food Making Process", then the video paragraph indicated by the second part can correspond to the paragraph editing area 12, and the title of this paragraph editing area 12 is "Food Making Process".
[0055] Each video paragraph serves as a theme, and each paragraph editing area serves as an editing area for a theme. In the following description, the video paragraph will be referred to as a theme, and the paragraph editing area will be referred to as the editing area of the theme.
[0055] The video editing interface 103 can be divided into multiple parts, and each part is used to display different contents. For example, in Figure 1DIn the example, the video editing interface 103 includes a navigation bar title 113, and in this example, the title is "Select Materials". The video editing interface 103 also includes editing areas for each of the M themes arranged in sequence, where M is a positive integer. For example, the first theme is "Finished Dishes" and the second theme is "Food Making Process". The user can import corresponding creative video clips or pictures for each theme in the editing area of each theme; or the user can choose not to import creative video clips or pictures in at least one of the M themes. Below the editing area of each theme, there may be relevant descriptions 163 of the video paragraph corresponding to this theme. If there is no relevant description of this theme in the video editing template, for example, in Figure 1C the description part, no corresponding characters are input, then the relevant description is not displayed.
[0056] It should be noted that in the embodiments of the present disclosure, the video editing template can be generated based on a video editing draft, and the video editing draft contains materials (such as videos, audios, pictures) and editing information. The video editing template needs to indicate which materials in the video editing draft need to be replaced. There are mainly two differences between the medium video template and the short video template. On the one hand, the medium video template has a paragraph structure. The video paragraph corresponds to a timeline interval, and the video paragraph includes one or more video clips, while the short video template has no structural paragraphs. The short video template is just a video editing draft in which some video clips are designated as materials that need to be replaced and filled; on the other hand, from the perspective of template usage, for the medium video template, each paragraph is respectively applied with the video editing effects within that paragraph. The application logic within the paragraph is that the editing operations indicated by the editing information within that paragraph are applied to the materials imported within the paragraph after being changed (because the number of materials is not bound to the clips in the template and the editing operations cannot be directly reused on the materials), while the application logic of the short video template is to import the video clips in the template into the materials, and the editing operations directly act on the materials, without the need to divide paragraphs or change the editing operations.
[0057] For step S20, as Figure 1D shown, in the editing area of each theme, there is an addition entry 123. For example, in response to a click operation on the addition entry 123, an album list is pulled up for the user to select and import the image materials for this theme from the album list. That is, the material import operation is, for example, a click operation on the addition entry 123 and the selection and import of the image materials for this theme from the album list.
[0058] In some embodiments of the present disclosure, the addition entry 123 has no slot concept and does not limit the number of clips and the clip duration.
[0059] For example, after performing a material import operation on a target video segment (also known as a target theme), the identifiers of at least one video material are displayed in the segment editing area of the target theme.
[0060] In some embodiments of the present disclosure, for example, if the video editing template includes a total of M themes, the target theme can be selected from the M themes for the material import operation. For example, if the user sequentially selects N themes for the material import operation, the N themes are all used as the target video segments. Both M and N are positive integers. Information prompt paragraph Information prompt paragraph In Figure 1D In the example of, for example, in response to the user clicking on the addition entry 123 in "Finished Dishes", the photo album list 133 is pulled up, allowing the user to select at least one video material related to the finished dishes from the photo album list 133. After that, the identifier of each of the at least one video material is displayed in the segment editing area 11 of the target theme "Finished Dishes". The identifier of each of the at least one video material is, for example, a frame image extracted from each video material. For example, the first frame image of each video material is extracted as the identifier of the video material. For example, in Figure 1D In the example of, the user imports video material 1 and video material 8 from the photo album list 133, and the first frame images of video material 1 and video material 8 are images showing the number 1 and the number 8 respectively. Then, images showing the number 1 and the number 8 are displayed in the segment editing area 11.
[0061] The user can click on the addition entry 123 in "Food Making Process" to pull up the photo album list 133 again, allowing the user to select at least one video material related to the food making process from the photo album list 133 again. In the embodiments of the present disclosure, the user can perform a material selection operation for each theme, or only for some of the M themes.
[0062] Medium videos have obvious editing structure routines compared to short videos. According to its structural information, one or more video materials imported by the user are grouped and sorted. In this embodiment, the video editing interface of the video editing template includes M themes arranged in sequence, so as to group one or more video materials according to the theme, making the video structure reasonable. In the embodiments of the present disclosure, there are basically no restrictions on the amount and duration of the imported materials by the user, and batch selection and deletion of the currently selected materials are supported. However, to avoid extreme situations, in some embodiments of the present disclosure, the maximum limits can be set for the amount and duration of the imported materials. For example, the duration of a single material (i.e., a video clip) can be at most 30 minutes, and if it exceeds, a prompt "The maximum imported video can only be within 30 minutes" will be given. The upper limit for the number of segments of materials imported for each theme is 100, and if it exceeds, a prompt "The maximum number of segments that can be imported is 100" will be given. For example, through Figure 1CThe upper limit of the total duration of user materials in [[]] limits the total duration of at least one video material.
[0063] In some embodiments of the present disclosure, the album list may include prompt texts. For example, when no video material is selected, the default prompt text is "Select the video materials to be imported"; when video materials have been selected, according to the number of materials, the prompt text is changed to " %d segments have been selected". As Figure 1D in, if video material 1 and video material 8 are selected, the prompt text is "2 segments have been selected". As Figure 1D in, the album list may further include a confirmation addition button 143. The confirmation addition button 143 has two states: lit and grayed out. After importing a segment, the state of the confirmation addition button 143 is the lit state, that is, the lighting logic is that the button is lit when a video material is imported, otherwise it is grayed out; after clicking the addition button, the video materials imported by the user are placed in the corresponding frame in order. As Figure 1D shown, after clicking the addition button 143, the identifiers of video material 1 and video material 8 are placed in the paragraph editing area 11 with the theme of "finished dish". If the current theme frame has been filled with video materials, enter the album list, and the added video materials are displayed by default.
[0064] In some embodiments of the present disclosure, the addition entry 123 moves backward following the position of the filled materials. The later-added video materials are automatically sorted behind the existing video materials. If there are more than one line, they will automatically wrap. The added video materials will display their original video material durations, and will be intelligently intercepted and segmented when entering the next synthesis step.
[0065] In some embodiments of the present disclosure, the deletion operation of video materials is supported. In response to the deletion operation of video materials, the video materials are deleted, and the video materials of this theme are automatically moved forward. For example, the deletion button can be selected on the album list page and the video editing interface 103 to perform the deletion operation. In some embodiments of the present disclosure, the user can also sort and edit the imported video materials. For example, swap the positions of two video materials, or perform editing operations such as beauty and cropping on a certain video material.
[0066] In the embodiments of the present disclosure, the video materials can be video segments or pictures.
[0067] For step S30, the triggering operation for the editing process is, for example, the selection operation of the export control or the preview control after the editing is completed through the video editing interface. For example, if the export control is selected, a synthesized exported video or an exported generated file can be generated according to the video editing template and at least one video material. For example, if the preview control is selected, a preview object for the preview page is generated according to the video editing template and at least one video material. The preview object can be a video or a picture, etc.
[0068] The video editing result includes, for example, multiple video parts, and each video part corresponds to a target video segment, that is, each video part is the editing result of the target editing effect being applied to at least one video material in the target video segment. The target editing effect is obtained according to the editing effects located within the target video segment among at least one editing effect.
[0069] For example, when importing the first video material into the theme "finished dish", for the target video segment of the theme "finished dish", the target editing effect is applied to the first video material, and the target editing effect is an editing effect that matches the editing effect located within the target video segment in the video editing template.
[0070] For example, the target editing effect is the same as the editing effect located within the target video segment in the video editing template, or the target editing effect is obtained by varying the editing effect located within the target video segment in the video editing template.
[0071] In some embodiments of the present disclosure, step S30 includes: in response to the triggering operation of the editing process, adjusting the editing effect according to the matching strategy to obtain the target editing effect; according to the video editing template, applying the target editing effect to at least a part of at least one video material to obtain the video part of the target video segment; and generating a video editing result based on the video parts corresponding to each target video segment in at least one video segment.
[0072] In some embodiments of the present disclosure, the matching strategy includes: when the editing effect includes a first piece of text, the target editing effect includes a second piece of text, the text format of the second piece of text is the same as the text format of the first piece of text, and the text content of the second piece of text matches at least one video material.
[0073] In some embodiments of the present disclosure, the text can refer to all texts other than subtitles, such as title-like decorative texts. For example, a piece of text content is added to each theme, and the text content takes effect based on the theme.
[0074] Figure 2 Shows a schematic diagram of adjusting the editing effect in the video editing template to obtain the editing effect in the video editing result provided by some embodiments of the present disclosure.
[0075] As shown in Figure 2 (a), in the video editing template, the copywriting for theme 1 is text 1, the copywriting for theme 2 is text 2, and the copywriting for theme 3 is text 3. That is, text 1 corresponds to theme 1 and only takes effect for theme 1, text 2 corresponds to theme 2 and only takes effect for theme 2, and text 3 corresponds to theme 3 and only takes effect for theme 3. Theme 1 is, for example Figure 1D the "finished dish" theme in Figure 1D and theme 2 is, for example
[0076] the "food making process" theme in. Theme 3 is, for example, other added themes. Text 1, text 2, and text 3 are examples of the first copywriting.
[0077] For example, if text 1 is in regular script, then the font of the copywriting in theme 1 in the video editing result is also regular script; if text 2 is in a gradient color, then the copywriting in theme 2 in the video editing result is also the same gradient color; if text 3 is in italic font, then the copywriting in theme 3 in the video editing result is also italic font.
[0078] For example, step S30 further includes obtaining the content of the second copywriting. For example, in Figure 1C the example of, the video editing interface 103 may further include an information prompt paragraph 104, and the user can input prompt information in the information prompt paragraph 104 to explain and describe the theme. In this embodiment, obtaining the content of the second copywriting includes: analyzing the prompt information and at least one video material to obtain an analysis result; and generating the content of the second copywriting according to the analysis result.
[0079] For example, a neural network for natural language processing can be used to analyze the prompt information to obtain an analysis result, so as to more accurately understand the user's requirements. For example, a sequence-to-sequence neural network model, a long short-term memory network, etc. are used to understand the prompt information to obtain an analysis result. For example, image processing technology is used to perform feature recognition on the video material to obtain an analysis result of at least one video material. Combining the analysis result of the prompt information and the analysis result of at least one video material, the content of the second copywriting that matches the prompt information and at least one video material is generated.
[0080] For example, if the prompt information is "sour and sweet" and at least one video material is a picture of a beverage, then the content of the second copywriting can be "sour and sweet, refreshing".
[0081] In some embodiments of the present disclosure, for example, the degree of influence of the analysis result on the content of the second copywriting is greater than the degree of influence of at least one video material on the content of the second copywriting. For example, when the second copywriting content is generated by a neural network, the weight of the analysis result is greater than the weight of at least one video material, which can better meet the user's requirements and make the generated copywriting content more in line with the user's expectations.
[0082] In the above embodiments, during the generation of the video editing result, a copywriting adapted to the video material is intelligently generated. For example, the copywriting content is intelligently written based on the text structure of the video editing template, the video material, and the prompt information. For example, the intelligently generated copywriting only modifies the copywriting content in the video editing template and does not modify the copywriting format in the video editing template, which is consistent with the copywriting format in the video editing template. In some embodiments of the present disclosure, the number of words in the intelligently generated copywriting is close to the number of words in the video editing template. For example, the difference in the number of words between the two is within 5 words. For example, the copywriting in the video editing template is "Home Record", the influencing material is food, and the prompt information is food production, then the content of the corresponding copywriting can be "Record of Making Food". In some other embodiments of the present disclosure, the copywriting in the video editing template can also be directly presented during the generation of the video editing result, or the user can perform personalized writing on the preview page. For the preview page, please refer to the following description.
[0083] Apply the format of the first copywriting to the content of the second copywriting to obtain the second copywriting included in the target editing effect. Then, apply this second copywriting to at least one video material corresponding to the target video segment (for example, video material 1 and video material 8 imported in theme 1).
[0084] In some embodiments of the present disclosure, the matching strategy further includes: in response to the first copywriting being applied to the entire paragraph of the target video segment, the second copywriting is applied to the entire part of the video part; in response to the first copywriting being applied to a partial time period in the target video segment, the second copywriting is applied to a partial time period in the video part.
[0085] In some embodiments of the present disclosure, it is determined whether the copywriting of a certain theme is displayed directly for the entire theme or only for a certain duration. When the first copywriting in the video editing template covers the entire duration of a certain theme, the second copywriting also takes effect for the entire target theme. When the first copywriting in the video editing template only covers a certain duration of a certain theme, the effective duration range of the second file in the target theme is the same as that duration.
[0086] Such as Figure 2As shown in (a) and (b), in the video editing template, text 2 (an example of the first copywriting) is displayed from the 0th second to the 60th second of theme 2. Then, in the video editing result, text 2 (an example of the second copywriting) is also only displayed from the 0th second to the 60th second of theme 2 in the video editing result. In this example, the content and format of the first copywriting and the second copywriting are the same. In the video editing template, text 3 is displayed throughout the duration of theme 3. Then, in the video editing result, text 3 is also, for example, displayed throughout the duration of theme 3. Figure 2 (c) and (d) are similar and will not be elaborated.
[0087] When the user modifies the display duration of the copywriting and then modifies the video material, it will be displayed according to the modified duration range; if the duration of the video material becomes shorter, the duration will be correspondingly shortened.
[0088] In some other embodiments of the present disclosure, the copywriting can be displayed according to the effective duration range of the copywriting in the video editing template. As Figure 2 As shown in (e), for example, in the video editing template, text 3 is displayed throughout the duration of theme 3 (for example, a total of 2 minutes), but the duration of theme 3 in the video editing result is a total of 3 minutes. Then, the copywriting is only displayed within the first 2 minutes of theme 3 in the video editing result. In the video editing template, text 1 is displayed throughout the duration of theme 1 (for example, a total of 1 minute), but the duration of theme 1 in the video editing result is a total of 2 minutes. Then, the copywriting is only displayed within the first 1 minute of theme 3 in the video editing result.
[0089] In some embodiments of the present disclosure, the copywriting in the target video segment can directly adopt the copywriting in the video editing template, which can improve the speed of synthesizing the video editing result. In some embodiments of the present disclosure, the content of the copywriting can be intelligently rewritten to make the copywriting more compatible with the imported materials.
[0090] In some embodiments of the present disclosure, for example, image recognition is performed on the second material to identify the scene and color tone of the second material, and then the copywriting content is generated based on the understanding of the second material. The copywriting format used in the video editing template is directly applied to this copywriting content to obtain the second copywriting.
[0091] In some embodiments of the present disclosure, the matching strategy includes: when the editing effect includes a transition animation, the target editing effect includes a transition animation.
[0092] When the video editing template includes M themes and material import operations are performed on N themes, in the case of N < M, for the first theme without imported video material, the transition animation between the first theme and the next theme after the first theme in the video editing template is deleted. That is, if a certain theme in the video editing template does not have imported video material, the transition animation immediately following that theme is also deleted. In this embodiment, only the transition animation corresponding to the target video segment in the video editing template is applied to the target video segment with imported video material.
[0093] As Figure 2 (a) shows, there is a transition animation (abbreviation "transition") 1 between theme 1 and theme 2, and there is a transition 2 between theme 2 and theme 3. The transition animation is used to make the video segments of adjacent themes transition naturally, making the switching of content in the video more natural and smooth.
[0094] As Figure 2 (b) shows, if theme 1 does not have imported initial material, then the transition 1 after theme 1 is also deleted, and the transition 2 included between theme 2 and theme 3 in the video editing template is also applied between theme 2 and theme 3 in the video editing result.
[0095] In some embodiments of the present disclosure, as Figure 1D shown, the structure information further includes an information prompt paragraph 104, and the information prompt paragraph 104 is used to obtain prompt information. Step S30 includes: receiving the prompt information input in the information prompt paragraph; and performing semantic analysis on the prompt information to obtain an analysis result.
[0096] For example, a neural network for natural language processing can be used to analyze the prompt information to obtain an analysis result, so as to more accurately understand the user's requirements. For example, a sequence-to-sequence neural network model, a long short-term memory network, etc. are used to understand the prompt information.
[0097] Figure 3 shows a Figure 1A method flowchart of step S30 provided by at least one embodiment of the present disclosure.
[0098] As Figure 3 shown, the method includes steps S301 to S304.
[0099] Step S301: In response to a trigger operation for editing processing, analyze the prompt information and at least one video material to obtain an analysis result.
[0100] Step S302: In response to the editing effect including subtitles and / or dubbing, according to the matching strategy that the content of the subtitles and / or dubbing in the target editing effect matches the analysis result, obtain the content of the subtitles and / or dubbing in the target editing effect.
[0101] Step S303: Obtain the format of the subtitles in the target editing effect according to the matching strategy that the format of the subtitles in the target editing effect is the same as the format of the subtitles in the editing effect.
[0102] Step S304: Obtain the voice parameters of the dubbing in the target editing effect according to the matching strategy that the voice parameters of the dubbing in the target editing effect are the same as the voice parameters of the dubbing in the editing effect.
[0103] In some embodiments of the present disclosure, if the video editing template includes tags for intelligent subtitle addition and / or intelligent dubbing addition, subtitles and / or dubbing are automatically added to the video editing result.
[0104] This method can make the content with subtitles and / or dubbing adapt to the prompt information, thereby improving the user experience.
[0105] Regarding step S301, the analysis of the prompt information and at least one video material to obtain the analysis result is similar to the foregoing embodiments and will not be elaborated here. For example, according to the video material imported by the user and the input prompt information, content and requirements are understood to generate appropriate subtitles and / or dubbing content. If it is recognized that the video material imported by the user includes sound, then the original sound in the video material, the image content of the video material, and the prompt information can be combined to generate appropriate subtitles and / or dubbing content. Or, regardless of whether the video material itself includes sound, it is defaulted to be intelligently dubbed.
[0106] For example, by using image recognition technology to determine that the video material is an image about making food, and the prompt information is "weekend", then it is determined that the subtitles in the preview video may include "Happy Weekend".
[0107] For steps S302 to S304, the matching strategy includes the matching strategy of the content of the subtitles and / or the content of the dubbing, the matching strategy of the subtitle format, and the matching strategy of the voice parameters.
[0108] For the content of the subtitles and / or the content of the dubbing, the matching strategy may be that the content of the subtitles and / or the dubbing in the target editing effect matches the analysis result. The matching of the content of the subtitles / dubbing and the analysis result is similar to the matching of the copywriting content and the analysis result described above and will not be elaborated here.
[0109] For the subtitle format, the matching strategy may be that the format of the subtitles in the target editing effect is the same as the format of the subtitles in the editing effect. For example, the format of the subtitles in the target editing effect may be the same as the format of the subtitles in the editing effect in the video editing template. When there is no subtitle format information in the template video, a default font style such as "white background with black border" can be adopted. The subtitle format includes, for example, the text font, text color, text style, etc. of the subtitles.
[0110] The voice parameters of the dubbing can include, for example, the timbre, frequency, pitch, etc. of the dubbing. For the voice parameters, the matching strategy is, for example, that the voice parameters of the dubbing in the target editing effect are the same as those of the dubbing in the editing effect. For example, the timbre of the dubbing in the target editing effect is the same as the timbre, frequency, and pitch, etc. of the dubbing in the editing effect of the video editing template. That is, the subtitle format and the voice parameters of the voice can directly apply the format and voice parameters in the editing effect of the video editing template. When there is no timbre in the editing effect of the video editing template (i.e., the video editing template is the original sound of the video), the default timbre is adopted, and the default timbre can be preset. That is, when the video editing template has original sound for the video material and the video editing template retains the original sound, when using the video editing template to synthesize the video editing result, the original sound in the video editing template is not retained by default.
[0111] For example, when there is already audio content such as text reading and recording in the video editing template, automatic muting processing is performed.
[0112] In some embodiments of the present disclosure, for example, the subtitle and / or the dubbing can take effect on the entire video editing result.
[0113] For example, for the final video finished product of all the video materials imported by the user, subtitles are automatically added, that is, the subtitles are displayed according to the duration of the final video finished product. When the duration of a single segment of the user is less than the preset duration threshold, for example, 3s, subtitles and / or dubbing are not provided for this segment. As Figure 2 (a) to 2(e) show, the subtitle duration is the same as the final video duration.
[0114] In some embodiments of the present disclosure, the subtitle and the dubbing can be in one-to-one correspondence. The content of the dubbing is based on the subtitle content at the bottom layer. If the subtitle content is modified, it is allowed to apply the modified content to the dubbing at the same time, and the reading content of the dubbing also changes automatically. In some embodiments of the present disclosure, playing the dubbing and displaying the subtitle are decoupled. For example, Table 1 shows an example of the decoupling of the dubbing and the subtitle provided by at least some embodiments of the present disclosure.
[0115] Table 1
[0116]
[0117] As shown in Scenario 1, Scenario 2, Scenario 3, and Scenario 4 in Table 1, the dubbing and the subtitle are decoupled. If the video editing template includes a "smart add dubbing" label and a "smart add subtitle" label, for example, Figure 1CRegarding the structural information, the synthesized video editing result is expected to have voiceovers and subtitles. If the video editing template does not include the "Intelligent Addition of Voiceovers" tag and the "Intelligent Addition of Subtitles" tag, the synthesized video editing result is expected not to include voiceovers and subtitles. If the video editing template includes the "Intelligent Addition of Voiceovers" tag but does not include the "Intelligent Addition of Subtitles" tag, the synthesized video editing result is expected to include voiceovers but not subtitles. If the video editing template does not include the "Intelligent Addition of Voiceovers" tag but includes the "Intelligent Addition of Subtitles" tag, the synthesized video editing result is expected not to include voiceovers but to include subtitles.
[0118] In some embodiments of the present disclosure, when the target editing effect only includes subtitles, the voiceover identifier in the preview video corresponding to the video editing result is displayed as the off state, and when the target editing effect only includes voiceovers, the subtitle identifier in the preview video corresponding to the video editing result is displayed as the off state. For example, in Scenario 2, the voiceover identifier is in the on state and the subtitle identifier is in the off state; in Scenario 3, the voiceover identifier is in the off state and the subtitle identifier is in the on state. Those skilled in the art can set any icons to represent voiceovers and subtitles respectively. The icon for voiceovers includes two different display states representing the off state and the on state respectively, and the subtitle icon also includes two different display states representing the off state and the on state respectively.
[0119] In some embodiments of the present disclosure, step S30, for example, includes splicing the video parts corresponding to each target video segment to obtain a video editing result.
[0120] Table II shows an example of a material import method provided by some embodiments of the present disclosure.
[0121] Subject 1 Subject 2 Subject 3 Corresponding sorting and splicing result A / B / C D E / F / G A / B / C / D / E / F / G A / B / C None None A / B / C A / B / C None E / F / G A / B / C / E / F / G None D E / F / G D / E / F / G
[0122] As shown in Table II, the template video includes Theme 1, Theme 2, and Theme 3 arranged in sequence. The video materials imported in Theme 1 include Clip A, Clip B, and Clip C, and the sorting of the three clips is Clip A, Clip B, and Clip C. The video materials of Theme 2 include Clip D, and the video materials of Theme 3 include Clip E, Clip F, and Clip G, and the sorting of the three clips is Clip E, Clip F, and Clip G. The video parts of Theme 1, Theme 2, and Theme 3 are spliced in the theme order to obtain a video editing result. As described above, the video part of each theme is obtained based on the video materials and the target editing effect. For example, applying the target editing effect to the video materials. This video editing result is, for example, obtained by applying the first target editing effect to Clip A, Clip B, and Clip C, applying the second target editing effect to Clip D, applying the third target editing effect to Clip E, Clip F, and Clip G, and then combining them in order.
[0123] If the video materials of Theme 1 include Clip A, Clip B, and Clip C, and the sorting order of the three clips is Clip A, Clip B, and Clip C, no materials are imported for Theme 2, and no materials are imported for Theme 3 either. The multiple clips are spliced according to the theme order to obtain a spliced video combined in the order of Clip A, Clip B, and Clip C, and a first target editing effect is applied to the video spliced with Clip A, Clip B, and Clip C to obtain a video editing result.
[0124] If the video materials of Theme 1 include Clip A, Clip B, and Clip C, and the sorting order of the three clips is Clip A, Clip B, and Clip C, Theme 2 is not imported, and the video materials of Theme 3 include Clip E, Clip F, and Clip G, and the sorting order of the three clips is Clip E, Clip F, and Clip G. The multiple clips are spliced according to the theme order to obtain a spliced video combined in the order of Clip A, Clip B, Clip C, Clip E, Clip F, and Clip G, and a first target editing effect is applied to the parts of Clip A, Clip B, and Clip C in the spliced video, and a third target editing effect is applied to the parts of Clip F and Clip G in the spliced video to obtain a video editing result.
[0125] If no materials are imported for Theme 1, the video materials of Theme 2 include Clip D, and the video materials of Theme 3 include Clip E, Clip F, and Clip G, and the sorting order of the three clips is Clip E, Clip F, and Clip G. The multiple clips are spliced according to the theme order to obtain a spliced video combined in the order of Clip D, Clip E, Clip F, and Clip G, and a second target editing effect is applied to the part of Clip D in the spliced video, and a third target editing effect is applied to the parts of Clip F and Clip G in the spliced video to obtain a video editing result.
[0126] In some embodiments of the present disclosure, when N < M, for the first theme without imported optimized materials, video clips are selected from the second theme with imported video materials as the target clips of the first theme; or the first theme is skipped during the process of splicing multiple video parts in the order of M themes, and the video parts of the themes after the first theme are spliced.
[0127] For example, when the second theme (for example, Figure 1D the "food production process" theme in) has no materials, it will automatically be postponed to the third theme for splicing with the first theme, and the materials of the third theme are directly arranged behind the materials of the first theme.
[0128] For example, if the video editing template is set to "default filling", even if no materials are imported for this theme, a segment (for example, randomly selected) will be selected from the imported materials for filling to ensure the integrity of the overall preview effect.
[0129] In some embodiments of the present disclosure, at least one editing effect includes, in addition to the above-mentioned paragraph editing effects for each theme such as copywriting, subtitles, and dubbing, a global editing effect. The global editing effect includes, for example, at least one of music, filters, and subtitles. In this embodiment, step S30 includes: in response to a trigger operation for editing processing, adjusting the time length of the global editing effect to match the time length of the target video material used in the video editing result, to obtain an adjusted global editing effect. On this basis, for example, based on the adjusted global editing effect and the video part corresponding to each target video paragraph, a video editing result is generated.
[0130] The target video material is, for example, the video material included in the video editing result. In some embodiments of the present disclosure, not all of the at least one video material imported by the user need to be used to generate the video editing result. For example, at least one video material can be used for highlight processing to obtain a highlight segment to generate the video editing result. In this embodiment, the time length of the video material used in the target video paragraph is less than the time length of the at least one video material imported for the target video paragraph.
[0131] In some embodiments of the present disclosure, for example, the target video material included in the video editing result is spliced from the highlight segments of one or more target video paragraphs. If the global editing effect includes music, adjust the playing time length of the music to match the time length of the target video material. For example, if the playing time length of the music in the video editing template is the first duration and the time length of the target video material is the second duration, then adjust the playing time length of the music from the first duration to the second duration.
[0132] For example, in response to the second duration being greater than the first duration, extend the playing time length of the music to the second duration, and the start time of the music playing in the video editing result is the same as the start time of the music playing in the video editing template; or in response to the second duration being less than the first duration, shorten the playing time length of the music to the second duration, and the music gradually fades out in the video editing result.
[0133] In the embodiments of the present disclosure, there is only one piece of music for the entire video. For example, the video editing template includes one piece of music, which serves as the background music of the video editing template. The background music in the video editing template can be directly applied in the generated video editing result, but the time length of the background music is appropriately adjusted.
[0134] Such as Figure 2As shown in (a), the video editing template 201 includes Theme 1, Theme 2, and Theme 3. At least one template video material of Theme 1 includes Clip 1 and Clip 2, at least one template video material of Theme 2 includes Clip 3, Clip 4, Clip 5, Clip 6, and Clip 7, and at least one template video material of Theme 3 includes Clip 8, Clip 9, Clip 10, and Clip 11. The music in the video editing template 201 runs through the entire template video. For example, the entire video editing template uses this music as the background music, and the original duration of the music (i.e., the first duration) is the same as the video time length of the video editing template 201.
[0135] As Figure 2 As shown in (b), if the user only imports video materials into Theme 2 and Theme 3 in the video editing interface and does not import video materials into Theme 1, resulting in the playing duration of the target video materials in the video editing result being less than the time length of the combined multiple template video materials of the video editing template 201 (hereinafter simply referred to as "video time length"), then directly shorten the original duration of the music to the playing duration of the target video materials adopted by Theme 2 and Theme 3. Moreover, the background music in the video editing result gradually fades out until it disappears, that is, fade out the background music.
[0136] As Figure 2 As shown in (c), although the user imports video materials into Theme 1, Theme 2, and Theme 3 respectively in the video editing interface, but the playing duration of the target video materials in the adopted video materials is less than the video time length of the video editing template 201, then also directly shorten the original duration of the music to be the same as the playing duration of the target video materials. Moreover, the background music in the video editing result gradually fades out until it disappears, that is, fade out the background music.
[0137] As Figure 2 As shown in (d), the user imports video materials into Theme 1, Theme 2, and Theme 3 respectively in the video editing interface, but the playing duration of the target video materials in the adopted video materials is greater than the video time length of the video editing template 201, then extend the original duration of the music to be the same as the playing duration of the target video materials.
[0138] In some embodiments of the present disclosure, when the playing duration of the target video materials is longer than the playing duration of the template video materials in the video editing template, extend the end time of the music in the video editing template backward, and align the start time of the music in the video editing result with the video editing template.
[0139] In some embodiments of the present disclosure, if the total duration of the music is long, the remaining music is played continuously. For example, if the total duration of the music is 3 minutes, but the video editing template only intercepts the duration range of 0-1 minute for use (i.e., the original duration is 1 minute), at this time, if the target video material exceeds 1 minute, the remaining content after 1 minute of music can be played continuously.
[0140] If the total duration of the music is shorter than the playing duration of the target video material, the music can be extended by means of intelligent music extension. The way of intelligent music extension is, for example, to automatically generate a music segment connected to the current music by applying a neural network, or to find a melody or music adapted to the current music from the music library.
[0141] In some embodiments of the present disclosure, if it is difficult to achieve intelligent extension, the current music is played in a loop automatically to extend the playing duration.
[0142] In some embodiments of the present disclosure, the global editing effect includes filters. For example, in response to the playing duration of the target video material being greater than the video time length, the filter effect time length in the video editing template is extended; or in response to the playing duration of the target video material being less than the video time length, the filter effect time length in the video editing template is shortened.
[0143] As Figure 2 shown in (b) and (c) of Figure 2 the first filter effect time length is, for example, the original duration. When the playing duration of the target video material is less than the original duration, the filter effect time length is shortened to the playing time length of the target video material. As
[0144] shown in (d) of Figure 2 when the playing duration of the target video material is greater than the original duration, the filter effect time length is extended to the playing time length of the target video material. The adjustment method of the filter configuration is similar to the adjustment method of the music configuration.
[0144] In some embodiments of the present disclosure, subtitles can also be used as a global editing effect. The processing logic of subtitles as a global editing effect is similar to that of the aforementioned filters and music, and will not be elaborated here.
[0145] In some of the above embodiments, step S30 generates a video editing result according to a matching strategy. In other embodiments of the present disclosure, step S30 may include: in response to a trigger operation of the editing process, using a neural network to process at least one video material and an editing effect to obtain a video editing result.
[0146] For example, in response to a trigger operation of the editing process, the neural network analyzes the video material and the editing effect to obtain an editing effect adapted to the video material and applies it to the video material to obtain a video editing result.
[0147] A neural network may include, for example, a neural network for image processing, such as a convolutional neural network, a deep learning neural network, etc.
[0148] In some embodiments of the present disclosure, step S30 includes: in response to a trigger operation of an editing process, first preprocessing at least one video material to obtain a target video material, and then applying a corresponding editing effect to the target video material to obtain a video editing result. The preprocessing may include, for example, highlight recognition, material segmentation, etc.
[0149] For example, step S30 includes: in response to a trigger operation of an editing process, extracting a target segment from at least one video material; and generating a video editing result according to a video editing template and the target segment.
[0150] Extracting a target segment from at least one video material may be, for example, performing highlight recognition on at least one video material to obtain a target segment of the highlight moment of at least one video material. The target segment of the highlight moment is, for example, the optimal material obtained by intercepting the optimal part of at least one of at least one video material.
[0151] In this embodiment, it is possible to automatically identify at least one video material to obtain a preferred segment of at least one video material and generate a video editing result.
[0152] In some embodiments of the present disclosure, step S30 includes: in response to a trigger operation of an editing process, if there is a first video material in at least one video material, splitting the first video material into multiple sub - segments, where the first video material is a material with a duration greater than a preset duration; and generating a video editing result based on a video editing template and the multiple sub - segments.
[0153] In each theme, the user can import at least one video material as shown in Figure 1D There may be a first video material with a time length greater than the preset duration in at least one video material. In the embodiments of the present disclosure, for the first video material with a time length greater than the preset duration, splitting processing can be performed to split the first video material into multiple sub - segments. These sub - segments are still used as the materials corresponding to the theme of the first video material and are not used across themes to avoid unexpected situations for the user. Splitting the first video material into multiple sub - segments is beneficial for identifying and utilizing the sub - segments, at least partially avoiding the problem of too long processing time caused by too long video materials, and is also beneficial for the flexible use of the sub - segments.
[0154] In some embodiments of the present disclosure, for example, in a video editing interface page, the imported video materials and prompt information of the user are analyzed and recognized in advance. After obtaining the trigger operation for editing and processing, the video materials are directly extracted (i.e., intercepted) and segmented according to the analysis results, thereby reducing the waiting time for generating the video editing result. For example, as long as the user imports one or more materials, the materials are analyzed, without waiting for the user to execute the trigger operation for editing and processing to start the analysis.
[0155] In some embodiments of the present disclosure, when the user starts to select materials of another theme, it is considered that the previous theme has been processed, and the situation of the user's modification needs to be considered for compatibility.
[0156] In some embodiments of the present disclosure, the method further includes playing a preview video on the preview page according to the video editing result. By playing the preview video, the user can watch the video draft generated from the video materials, which is convenient for the user to modify in time when the user needs to modify. In some embodiments of the present disclosure, when the analysis result indicates that the prompt information includes a preset playing duration, the playing duration of at least some materials is the same as the preset playing duration.
[0157] For example, if the prompt information input by the user is "Learn 3 kinds of delicious foods in 20 minutes", then through semantic analysis of the prompt information, it is obtained that the playing duration of the video that the user hopes to produce is 20 minutes. Then, the duration of the target video materials extracted from the video materials and applied to the video editing result is also 20 minutes.
[0158] Figure 4 The figure shows a schematic diagram of a preview page provided by at least some embodiments of the present disclosure.
[0159] As Figure 4 shown, the preview page 401 includes a video playback window 411, a material display window 421, and an editing entry window 431.
[0160] The video playback window 411 is used to display or play the preview video. The material display window 421 includes a plurality of material sub-windows for displaying at least part of each of the plurality of video materials in the preview video. At least part of each of the plurality of video materials is, for example, the target segment of each of the plurality of video materials. For example, the plurality of material sub-windows include a material sub-window 4211 for displaying the materials of theme 1 and a material sub-window 4212 for displaying the materials of theme 2. That is, in some embodiments of the present disclosure, each theme corresponds to a material sub-window for displaying the materials adopted by the theme. Each material sub-window includes a plurality of material window units, and each material window unit is used to display one material in the theme. As Figure 4As shown, there are 4 material window units shown in Theme 1, namely material window unit A, material window unit B, material window unit C, and material window unit D. The 4 material window units respectively display different materials in this Theme 1.
[0161] In some embodiments of the present disclosure, the material display window 421 further includes a material addition sub-window 4213 for adding materials. In this embodiment, the video editing method further includes: in response to a trigger operation on the material addition sub-window when a material window unit in the plurality of material sub-windows 4213 is selected, displaying a plurality of selectable materials; and adding a target material after the material window unit according to an addition selection operation on the plurality of selectable materials.
[0162] As Figure 4 shown, if the material addition sub-window 4213 is clicked when the material window unit B is selected, an album list is pulled up to display a plurality of selectable video materials. If the user performs an addition selection operation on the target material among the plurality of selectable video materials, the target material is added after the material window unit B. For example, a material window unit E is added between the material window unit B and the material window unit C to display the target material. For the album list, please refer to the above description.
[0163] The editing entry window 431 is used to display the entry for editing materials. For example, the editing entry window 431 includes at least one of a subtitle entry 4311 for editing subtitles, a music entry 4312 for editing music, a filter entry 4313 for editing filters, and a copywriting entry 4314 for editing copywriting. The editing entry window 431 may further include entries for other functions, and those skilled in the art can set them in the editing entry window according to editing needs. In some embodiments of the present disclosure, when the subtitle entry 4311 is clicked to enter the subtitle editing page, a voice-over editing entry for editing voice-over may be included in this subtitle editing page. The voice-over editing entry may be resident in the subtitle editing page, and the voice-over editing entry can be displayed whether voice-over is performed or not. If there is no voice-over in the video editing template, after clicking to enter the voice-over editing entry, the voice-over is defaulted to the disabled state.
[0164] In some embodiments of the present disclosure, for the subtitle entry 4311, if the video editing template includes a label for intelligent subtitle addition or a label for intelligent voice-over addition, the condition for intelligent subtitle addition to the preview video is met, and then the subtitle entry 4311 is displayed in the editing entry window 431. If the template video does not include a label for intelligent subtitle addition and a label for intelligent voice-over addition, the subtitle entry 4311 is not displayed in the editing entry window 431.
[0165] In some embodiments of the present disclosure, some video editing templates may have only voiceovers or only subtitles. When there is only a voiceover, the subtitle entry 4311 is normally displayed in the editing entry window 431, but in the subtitle editing page, "Subtitle" is displayed as default in the closed state, and the subtitle only serves as the content for voiceover reading. In some template videos, there may be only subtitles. The subtitle entry 4311 is normally displayed in the editing entry window 431, but in the subtitle editing page, "Voiceover" is displayed as default in the closed state, and only the subtitles are displayed without voiceover reading.
[0166] In the subtitle editing page, the user can modify the content of the subtitle, set the subtitle duration, text reading, and deletion, etc.
[0167] Figures 5A to 5F FIG. shows a schematic diagram of modifying subtitles provided by at least some embodiments of the present disclosure.
[0168] For example, as Figure 5A shown, in the subtitle primary page (i.e., the subtitle editing page) 501, subtitles "abcd", "efg", and "higk" are displayed. If the subtitle "efg" is selected in the subtitle editing page 501 and this subtitle "efg" is also displayed on the screen at this time, then it enters Figure 5B the subtitle secondary page 502 shown, or directly click on the subtitle content to enter the subtitle editing page 502. Click on the "Edit" entry 512 in the subtitle editing page 502 to enter the subtitle content modification. In some embodiments of the present disclosure, the selected state of the subtitle follows the sliding operation to select different subtitles, and the selected subtitle is automatically displayed on the screen.
[0169] As Figures 5C to 5D shown, when modifying the subtitle content, the input box 504 and the keyboard 503 are automatically pulled up to modify the subtitle content using the input box 504 and the keyboard 503. Support the "Delete" button for input on the screen; or delete the subtitle in the way of the subtitle secondary page. For example, as Figure 5B shown, the delete button 522 is located at the last of the subtitle secondary page. The subtitle position can be dragged, and after dragging, it is applied globally by default, and all subtitles are changed to the new position. For example, after modifying the subtitle content, click on the effect icon 505 to apply it in the preview video. In some embodiments of the present disclosure, after clicking on the effect icon 505, Figure 5D the pop-up window 506 shown may appear. This pop-up window 506 is used to ask the user whether to apply this subtitle to the voiceover. If the user selects to apply it to the voiceover, then correspondingly, the voiceover is modified to be consistent with the subtitle.
[0170] As Figure 5B and 5D shown, the subtitle secondary page 502 may also include a subtitle duration setting button 532.
[0171] For example, when the user selects the corresponding subtitle and clicks the subtitle duration setting button 532, the segment range 542 corresponding to the subtitle is displayed at the lower part of the secondary subtitle page, and stretching selection of the duration range is supported. Above the displayed segment range 542 corresponding to the subtitle, other subtitles 552 can be displayed. Selecting other subtitles 552 can switch to the display of other subtitle duration ranges.
[0172] In some embodiments of the present disclosure, a "change to another batch" button 562 can be permanently resident on the primary subtitle page 501 and the secondary subtitle page 502. In response to a click operation on the "change to another batch" button 562, new subtitle content is regenerated. When the subtitle is scrolled on the primary subtitle page 501 or the secondary subtitle page 502, the "change to another batch" button 562 can disappear, and when the scrolling stops, the "change to another batch" button 562 appears again.
[0173] In some embodiments of the present disclosure, the primary subtitle page 501 and the secondary subtitle page 502 can further include a first viewing icon for viewing the previous subtitle and a second viewing icon for viewing the next subtitle. Clicking the first viewing icon returns to the previous version, and clicking the second viewing icon continues to generate new subtitles.
[0174] In some embodiments of the present disclosure, the primary subtitle page 501 and the secondary subtitle page 502 can further include a subtitle switch entry 511. When there is a subtitle for an application, the icon of the subtitle switch entry 511 is presented in a selected state, and at this time the subtitle is displayed in the preview video. After clicking to close, the subtitle switch entry 511 becomes unselected, and at this time the subtitle is hidden and not displayed in the preview video, and is only for reading aloud. When switching between the display and hidden states, a prompt "subtitle has been displayed" is given when clicking to expand, and a prompt "subtitle has been hidden" is given when clicking to hide.
[0175] As described above, as Figure 5F shown, in the subtitle editing page 502, there can be a dubbing editing entry 562 for editing the dubbing. Enter the dubbing editing page 507 by selecting the dubbing editing entry 562, and the timbre can be modified in the dubbing editing page 507. In the dubbing editing page 507, the currently applied timbre is selected by default. For example Figure 5F in which timbre 1 is the currently applied timbre and is in a selected state. Clicking to switch to other timbres immediately previews the effect and reads aloud the current subtitle content. If the "disable" button is clicked in the dubbing editing page, the dubbing is not read aloud.
[0176] In the dubbing editing page, setting the dubbing volume and the video original sound volume is supported. When there is dubbing, the video original sound is defaulted to 0.
[0177] In some embodiments of the present disclosure, it is possible to select whether to apply the set dubbing globally in the dubbing editing page. If the dubbing is applied globally, all subtitles will take effect with the new dubbing effect.
[0178] In some embodiments of the present disclosure, for example, if the user clicks the effect icon, the new dubbing modification is saved and the dubbing takes effect. If the user does not click the effect icon or clicks the exit icon, the dubbing modification is not saved and the user returns to the previous page of the dubbing editing page.
[0179] In some embodiments of the present disclosure, clicking on the music entry 4312 enters the music editing page. In the music editing page, music can be recommended through similar music. For example, music with a similarity greater than a predetermined value in the video editing template is recommended. Or, music is recommended based on the understanding of the second material and the prompt information.
[0180] In the music editing page, the user can independently select music and can set the volume of the music. In some embodiments of the present disclosure, it is also possible to perform beat processing on the music.
[0181] In some embodiments of the present disclosure, the editing entry window further includes: a delete icon for selecting and deleting multiple preferred materials in the preview video. For example, in the editing entry window 431 on the preview page, there is a "Delete" entry, which is located at the last position of the editing entry window 431. When a single material is selected and the "Delete" icon is clicked on the preview page, the selected material is deleted. After the selected material is deleted, the other materials in the theme where the selected material is located automatically move forward, and the materials in other themes do not move across themes.
[0182] In some embodiments of the present disclosure, the editing entry window 431 and the material display window 421 support left and right sliding to display more editing entries and materials. For example, sliding the editing entry window 431 to the left reveals the delete icon.
[0183] In some embodiments of the present disclosure, the editing entry window further includes a reorganization entry. The method further includes: in response to the reorganization entry being selected, displaying a sorting panel, where the sorting panel includes the second material and the prompt information; receiving the organization operation performed on the second material in the sorting panel, and reorganizing the second material.
[0184] For example, the editing entry window on the preview page further includes a reorganization entry. After clicking the reorganization entry, a sorting panel is displayed. For example, the sorting panel has the same logic as the original structural framework. Except that the sorting panel includes the second material imported previously and prompt information, the rest is the same as the above structural framework, so as to modify the prompt information and import the second material. In response to entering the sorting panel through the reorganization entry, the previously filled prompt information is displayed in the structural framework, and each second material is displayed according to each theme, and the copywriting is also displayed according to the original copywriting. The organization operations include, for example, modification, addition, deletion, and drag-and-drop sorting (able to sort across themes) of imported materials, as well as modification of prompt information, etc. Receive the modification, addition, deletion, and drag-and-drop sorting (able to sort across themes) of the prompt information and imported materials in the structural framework, and save the above modifications after receiving the effective instruction. When there are changes in the prompt information and the content filled with materials, all the synthesis of algorithm materials and all intelligent effects of the algorithm need to be regenerated once (including: material splicing selection, overall packaging synthesis, intelligent copywriting rewriting, intelligent subtitle / dubbing generation, intelligent music recommendation), and after taking effect, return to the previous page. If the user receives an exit instruction after modifying the structural framework, all modifications will not take effect.
[0185] In some embodiments of the present disclosure, in addition to through the reorganization entry, the sorting of materials can also be performed by a special operation (for example, long pressing on a certain clip material). For example, the special operation can support cross-group sorting.
[0186] In some embodiments of the present disclosure, materials can also be intercepted. For example, the default positioning of the interception range of the materials is the actually selected range, and adjustment of the interception range is supported, rather than a fixed duration.
[0187] In some embodiments of the present disclosure, various operations such as adding and deleting materials can be performed on the preview page, or the video editing interface can be redisplayed through the reorganization entry on the preview page. In the video editing interface, operations such as adding, deleting, and sorting video materials can be performed, and the operation of adjusting the interception duration range of a single video material can also be performed. In the video editing interface redisplayed through the reorganization entry on the preview page, the prompt information can also be modified.
[0188] When the video materials change (for example, adding, deleting, and sorting materials), the target segments are analyzed and identified for the newly imported video materials, and the materials are segmented; for the original materials (previously imported materials), the original processing effects are continued to be retained without reprocessing. If, on the basis of importing new video materials, the original materials are also reordered, then the new video part is re-spliced. Based on the new video materials, repackaging is performed, including corresponding filters, music, text, and transition effects. Please refer to the above description.
[0189] In some embodiments of the present disclosure, for newly imported video materials, subtitles and dubbing need to be regenerated. For original materials, the original processing effects are continued to be retained without regenerating subtitles and dubbing. For deleted materials, the corresponding subtitles / dubbing are also deleted accordingly. For the sorting of new video materials, subtitles and dubbing are regenerated according to the new sorting.
[0190] In some embodiments of the present disclosure, there is no need to regenerate the copywriting for newly imported video materials. For the copywriting of original materials, the original processing effects are continued to be retained without reprocessing. For deleted materials, if a certain theme has no materials at all, the corresponding copywriting for that theme is also deleted. For the sorting of new video materials, the display position of the copywriting is also reprocessed according to the new sorting.
[0191] In some embodiments of the present disclosure, when only the intercepted duration of the material changes, it is selected according to the user's intercepted range, no new highlight recognition is performed, and the material is re-spliced according to the newly intercepted material. Based on the duration of the latest material, the material is repackaged according to the processing methods of music, filters, and copywriting described above. For example, the music, filters, and copywriting in the regenerated video editing results only change the playing duration and the effective duration, without changing the content itself. When only the intercepted duration of the video material changes, new subtitles and dubbing are generated according to the duration of the new video material. When the duration of a single segment of the user is less than 3s, subtitles and dubbing are not provided for this segment. At this time, the existing subtitles and dubbing may not be retained because it is too short to be displayed and played. When only the intercepted duration of the video material changes, there is no need to regenerate the copywriting. For the text that takes effect globally for the theme, the length of the text display duration is the same as the length of the material. For the copywriting that takes effect only for part of the theme, if the video material becomes longer, the copywriting is still displayed according to the original duration; if the material becomes shorter, the text is displayed according to the duration after the material becomes shorter.
[0192] In some embodiments of the present disclosure, if the theme and style of the prompt information change, all subtitles and / or dubbing will be refreshed.
[0193] In some embodiments of the present disclosure, when the subtitles and dubbing change, if the subtitles are longer than the subtitles before modification, they continue to be displayed within the duration range of the subtitles before modification, and the subtitles automatically wrap, and the wrapping logic aligns with the current subtitle wrapping. If the subtitles are shorter than the subtitles before modification, they continue to be displayed within the original duration range of the subtitles, and as much of the subtitles as possible are displayed. If the subtitle duration range spans multiple segments, resulting in multiple subtitles overlapping, they can be displayed according to the new effect.
[0194] When the subtitles are too long, resulting in a correspondingly long voiceover, the voiceover is read normally according to the lengthened length, and full playback of the imported video material is supported. When there are multiple overlapping voiceover audio segments in a clip, the multiple audio segments are played together. When the subtitles become shorter, resulting in a correspondingly short voiceover, the voiceover is read normally according to the shortened length.
[0195] In some embodiments of the present disclosure, there is a conflict between the process of algorithmically re-synthesizing the video editing result and the user's editing. For example, during the process of synthesizing the video editing result, the user performs editing operations on materials, voiceovers, etc. When the material modified by the user again overlaps with the material being processed by the current algorithm, the effect of the material actively modified by the user is preferentially displayed. For example, if the user has manually intercepted the duration range of a certain material, it is selected according to its duration, rather than according to the target segment selected as described above. If, during the process of algorithmically re-synthesizing the video editing result, in response to a selection operation on the subtitle editing entry or the voiceover editing entry, a notification message is displayed on the preview page to inform the user that subtitles and voiceovers are being generated. Wait for the video editing result to be re-synthesized, and the preview page is normally displayed and supports operations. If, during the process of algorithmically re-synthesizing the video editing result, the on-screen subtitle is selected, normal editing and modification are performed, and the intelligent subtitle progress is normally loaded. When the user confirms to save the modification, at this time, the subtitle content modified by the user for this segment is preferentially displayed, and the intelligent generation progress of this segment of subtitle stops. When the user does not save the modification, the intelligent subtitle is normally loaded and displayed. When the rewritten copy has not been generated, the text content is displayed as the original template copy at this time; if it is a re-synthesis triggered after secondary editing, when the new rewritten copy has not been generated, the current copy is displayed. The user can normally edit and modify the content, the rewritten copy is normally loaded, and the text modified by the user for this segment is effectively displayed, and the intelligent rewriting progress of this segment of copy stops.
[0196] In some embodiments of the present disclosure, the video editing template can be obtained by converting the project file. For example, after exporting the project file of the editing tool, select "Publish Template". By adding the structural routine information of the video template in the publishing link and the vimo platform, and through the characteristics of template editing, identify at what nodes to split each group of shots and the corresponding themes of each group of shots. After the medium video template is published, it can be stored and managed in the vimo background. For example, by adding a medium video template identifier on vimo, it is judged whether the template is of the medium video template type through manual identification by the operation. For example, the material category is added with the type of "medium video template", and an attribute field of the type of "medium video template" is newly added to the metadata.
[0197] For the input of template structure information, for example, on the vimo platform, the "Modify Structure Information" entry is added successively through the template management entry, template material management entry, and operation details entry. In the page for modifying structure information, it is mainly divided into three major parts: overall description, intelligent capabilities, and template structure breakdown. The overall description part is used to provide an overall description of the video editing template, without limiting the maximum number of input characters. When displayed on the application page, if it exceeds the design range, a fading mask layer will be displayed.
[0198] The intelligent capabilities part supports the operation to check intelligent capabilities, including intelligent subtitle addition and intelligent dubbing addition. The template structure breakdown part is self - disassembled into multiple parts according to the template structure. The content filled in each part includes: title, description, start and end time, start and end segments, upper limit of the total duration of user materials, default filling, and default placeholder. The title does not limit the maximum number of input characters. When displayed on the application page and exceeding the design range, a fading mask layer will be displayed. The description also does not limit the maximum number of input characters. When displayed on the application page and exceeding the design range, a fading mask layer will be displayed. The start and end time is used to describe from which second the video segment starts to which second it ends. The start and end segments are used to describe from which segment to which segment it starts and ends. The upper limit of the total duration of user materials is used to explain that when importing relatively long materials, the algorithm can refer to the upper limit to prevent the finally intercepted and synthesized materials from being too long, so as to achieve the optimal effect. The default filling means that when no materials are imported in this group of parts, a segment is selected from the imported materials for filling to ensure the integrity of the overall preview effect. The default placeholder means that when no materials are imported in this group of parts and entering the preview page, the material display of this group is blank placeholder. After confirmation in the template structure breakdown part, a secondary confirmation pop - up window will be displayed. After clicking the secondary pop - up window, the video editing template will be officially launched and go live.
[0199] In some embodiments of the present disclosure, as long as one material is filled in, the "Preview" button will be lit and the preview page can be entered. When the material is empty, the "Preview" button will be dimmed and the preview page cannot be entered. The "Preview" button can be fixed at the bottom of the page without following the scroll. When other content exceeds one screen, it is necessary to scroll down to view more.
[0200] Figure 6 It is a schematic block diagram of a video editing device provided for some embodiments of the present disclosure. As Figure 6 shown, the video editing device 600 includes a first display unit 110, a second display unit 120, and a result generation unit 130. For example, the video editing device 600 can be applied to a user terminal, or can be applied to any device or system that needs to implement the preview of design materials. The embodiments of the present disclosure do not limit this.
[0201] The first display unit 110 is configured to display a video editing interface in response to a trigger operation on a video editing template. The video editing template includes structure information and editing information. The structure information is used to indicate at least one video segment, and the editing information is used to indicate at least one editing effect applied to the at least one video segment. The video editing interface includes paragraph editing areas corresponding to the at least one video segment respectively. For example, the first display unit 110 may execute step S10 of the video editing method as shown in Figure 1A shown in
[0202] The second display unit 120 is configured to display the identifiers of at least one video material in the paragraph editing area of the target video segment in response to a material import operation for the target video segment. The at least one video segment includes the target video segment, and the at least one video material is the material imported into the target video segment based on the material import operation. For example, the second display unit 120 may execute step S20 of the video editing method as shown in Figure 1A shown in
[0203] The result generation unit 130 is configured to generate a video editing result according to the video editing template and the at least one video material in response to a trigger operation for an editing process. The video editing result corresponding to the part of the target video segment is the editing result in which the target editing effect is applied to the at least one video material, and the target editing effect matches the editing effect located in the target video segment among the at least one editing effect. For example, the page display unit 130 may execute step S30 of the video editing method as shown in Figure 1A shown in
[0204] This video editing device not only reduces the requirements for materials, making the use of video templates more flexible, but also this method is applicable not only to short videos but also to medium-length videos.
[0205] For example, the first display unit 110, the second display unit 120, and the result generation unit 130 may be hardware, software, firmware, and any feasible combination thereof. For example, the first display unit 110, the second display unit 120, and the result generation unit 130 may be dedicated or general-purpose circuits, chips, or devices, etc., or may also be a combination of a processor and a memory. Regarding the specific implementation forms of the first display unit 110, the second display unit 120, and the result generation unit 130, the embodiments of the present disclosure do not limit this.
[0206] It should be noted that in the embodiments of the present disclosure, each unit of the video editing device 600 corresponds to each step of the foregoing video editing method. For the specific functions of the video editing device 600, reference may be made to the relevant descriptions of the video editing method in the foregoing text, which will not be elaborated herein. Figure 6 The components and structures of the video editing device 600 shown are exemplary and not restrictive. According to needs, the video editing device 600 may further include other components and structures.
[0207] Figure 7 It is a schematic block diagram of an electronic device provided in some embodiments of the present disclosure. As Figure 7 shown, the electronic device 200 includes a processor 210 and a memory 220. The memory 220 is used to store non-temporary computer-readable instructions (such as one or more computer program modules). The processor 210 is used to run the non-temporary computer-readable instructions, and when the non-temporary computer-readable instructions are run by the processor 210, one or more steps in the foregoing video editing method can be executed. The memory 220 and the processor 210 may be interconnected through a bus system and / or other forms of connection mechanisms (not shown).
[0208] For example, the processor 210 may be a central processing unit (CPU), a digital signal processor (DSP), or other forms of processing units with data processing capabilities and / or program execution capabilities, such as a field programmable gate array (FPGA), etc.; for example, the central processing unit (CPU) may be of the X86 or ARM architecture, etc. The processor 210 may be a general-purpose processor or a dedicated processor, and may control other components in the electronic device 200 to perform desired functions.
[0209] For example, the memory 220 may include any combination of one or more computer program products. The computer program products may include various forms of computer-readable storage media, such as volatile memory and / or non-volatile memory. Volatile memory may include, for example, random access memory (RAM) and / or cache memory, etc. Non-volatile memory may include, for example, read-only memory (ROM), hard disk, erasable programmable read-only memory (EPROM), portable compact disc read-only memory (CD-ROM), USB memory, flash memory, etc. One or more computer program modules may be stored on the computer-readable storage media, and the processor 210 may run one or more computer program modules to implement various functions of the electronic device 200. Various application programs and various data, as well as various data used and / or generated by the application programs, etc. may also be stored in the computer-readable storage media.
[0210] It should be noted that in the embodiments of the present disclosure, for the specific functions and technical effects of the electronic device 200, reference may be made to the description of the video editing method in the foregoing text, and details are not described herein again.
[0211] Figure 8 FIG. is a schematic block diagram of another electronic device provided in some embodiments of the present disclosure. The electronic device 300 is, for example, suitable for implementing the video editing method provided in the embodiments of the present disclosure. The electronic device 300 may be a user terminal or the like. It should be noted that Figure 8 The illustrated electronic device 300 is merely an example and will not impose any limitations on the functions and usage scope of the embodiments of the present disclosure.
[0212] As Figure 8 shown, the electronic device 300 may include a processing device (such as a central processing unit, a graphics processing unit, etc.) 310, which may perform various appropriate actions and processes according to a program stored in a read-only memory (ROM) 320 or a program loaded from a storage device 380 into a random access memory (RAM) 330. In the RAM 330, various programs and data required for the operation of the electronic device 300 are also stored. The processing device 310, the ROM 320, and the RAM 330 are connected to each other through a bus 340. An input / output (I / O) interface 350 is also connected to the bus 340.
[0213] Generally, the following devices may be connected to the I / O interface 350: an input device 360 including, for example, a touch screen, a touchpad, a keyboard, a mouse, a camera, a microphone, an accelerometer, a gyroscope, etc.; an output device 370 including, for example, a liquid crystal display (LCD), a speaker, a vibrator, etc.; a storage device 380 including, for example, a magnetic tape, a hard disk, etc.; and a communication device 390. The communication device 390 may allow the electronic device 300 to communicate with other electronic devices wirelessly or wiredly to exchange data. Although Figure 9 the illustrated electronic device 300 has various devices, it should be understood that it is not required to implement or include all the illustrated devices, and the electronic device 300 may alternatively implement or include more or fewer devices.
[0214] For example, according to the embodiments of the present disclosure, Figure 1AThe video editing method shown can be implemented as a computer software program. For example, an embodiment of the present disclosure includes a computer program product, which includes a computer program carried on a non-transitory computer-readable medium, and the computer program includes program code for executing the above video editing method. In such an embodiment, the computer program can be downloaded and installed from a network through a communication device 390, or installed from a storage device 380, or installed from a ROM 320. When the computer program is executed by a processing device 310, the functions defined in the video editing method provided by the embodiments of the present disclosure can be executed.
[0215] At least one embodiment of the present disclosure further provides a storage medium for storing non-temporary computer-readable instructions, which can implement the video editing method described in any embodiment of the present disclosure when executed by a computer. Using this storage medium not only reduces the requirements for materials, making the use of video templates more flexible, but also this method is applicable to medium-length videos in addition to short videos.
[0216] Figure 9 Schematic diagram of a storage medium provided by some embodiments of the present disclosure. As Figure 9 shown, the storage medium 400 is used to store non-temporary computer-readable instructions 410. For example, when the non-temporary computer-readable instructions 410 are executed by a computer, one or more steps in the video editing method described above can be executed.
[0217] For example, the storage medium 400 can be applied to the above-mentioned electronic device 200. For example, the storage medium 400 can be Figure 7 the memory 220 in the shown electronic device 200. For example, the relevant description of the storage medium 400 can refer to Figure 7 the corresponding description of the memory 220 in the shown electronic device 200, which will not be elaborated here.
[0218] In the above, in combination with Figures 1A to 9 the video editing method, video editing device, electronic device and storage medium provided by the embodiments of the present disclosure are described. The video editing method provided by the embodiments of the present disclosure not only reduces the requirements for materials, making the use of video templates more flexible, but also this method is applicable to medium-length videos in addition to short videos.
[0219] It should be noted that the above-mentioned storage medium (computer-readable medium) in the present disclosure can be a computer-readable signal medium, a non-transitory computer-readable storage medium, or any combination of the two. The non-transitory computer-readable storage medium can be, for example, but not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination of the above. More specific examples of the non-transitory computer-readable storage medium can include, but are not limited to: an electrical connection with one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the above. In the present disclosure, the non-transitory computer-readable storage medium can be any tangible medium that contains or stores a program, and this program can be used by or in combination with an instruction execution system, apparatus, or device. In the present disclosure, the computer-readable signal medium can include a data signal propagated in a baseband or as part of a carrier wave, in which the computer-readable program code is carried. Such a propagated data signal can take various forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination of the above. The computer-readable signal medium can also be any computer-readable medium other than the non-transitory computer-readable storage medium, and this computer-readable signal medium can send, propagate, or transmit a program for use by or in combination with an instruction execution system, apparatus, or device. The program code contained on the computer-readable medium can be transmitted by any suitable medium, including but not limited to: wires, optical cables, RF (radio frequency), etc., or any suitable combination of the above.
[0220] In some embodiments, the client and the server can communicate using any currently known or future-developed network protocol such as the Hyper Text Transfer Protocol (HTTP), and can be interconnected with digital data communication in any form or medium (for example, a communication network). Examples of the communication network include a local area network (LAN), a wide area network (WAN), the Internet (for example, the Internet), and a peer-to-peer network (for example, an ad hoc peer-to-peer network), as well as any currently known or future-developed network.
[0221] The above-mentioned computer-readable medium can be included in the above-mentioned electronic device; it can also exist separately without being assembled into the electronic device.
[0222] The above computer-readable medium carries one or more programs which, when executed by the electronic device, cause the electronic device to perform one or more steps of the video editing method described above. Computer program code for carrying out operations of the present disclosure may be written in one or more programming languages or combinations thereof. The programming languages include, but are not limited to, object-oriented programming languages such as Java, Smalltalk, C++, and also conventional procedural programming languages such as the "C" language or similar programming languages. The program code may execute entirely on the user's computer, partly on the user's computer, as a stand-alone software package, partly on the user's computer and partly on a remote computer, or entirely on the remote computer or server. In the case of a remote computer, the remote computer may be connected to the user's computer through any type of network, such as a local area network (LAN) or a wide area network (WAN), or alternatively, may be connected to an external computer (e.g., through the Internet using an Internet service provider).
[0223] The flowcharts and block diagrams in the accompanying drawings illustrate the possible architectures, functions, and operations of systems, methods, and computer program products according to various embodiments of the present disclosure. In this regard, each block in the flowchart or block diagram may represent a module, a segment of a program, or a part of code, which contains one or more executable instructions for implementing the specified logical function. It should also be noted that in some alternative implementations, the functions noted in the blocks may occur in a different order than noted in the drawings. For example, two consecutive blocks shown may actually be executed substantially in parallel, or they may sometimes be executed in the reverse order, depending on the functions involved. It should also be noted that each block in the block diagrams and / or flowcharts, and combinations of blocks in the block diagrams and / or flowcharts, may be implemented by a dedicated hardware-based system for performing the specified functions or operations, or by a combination of dedicated hardware and computer instructions.
[0224] The units described in the embodiments of the present disclosure may be implemented in software or in hardware. In some cases, the name of the unit does not constitute a limitation on the unit itself.
[0225] The functions described above herein may be performed, at least in part, by one or more hardware logic components. By way of example, and without limitation, exemplary types of hardware logic components that may be used include: field programmable gate arrays (FPGA), application specific integrated circuits (ASIC), application specific standard products (ASSP), system on a chip (SOC), complex programmable logic devices (CPLD), etc.
[0226] In the present disclosure, a machine-readable medium may be a tangible medium that can contain or store a program for use by or in connection with an instruction execution system, apparatus, or device. The machine-readable medium may be a machine-readable signal medium or a machine-readable storage medium. The machine-readable medium may include, but is not limited to, electronic, magnetic, optical, electromagnetic, infrared, or semiconductor systems, apparatus, or devices, or any suitable combination of the foregoing. More specific examples of the machine-readable storage medium would include an electrical connection based on one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disc read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the foregoing.
[0227] The above description is only partial embodiments of the present disclosure and an illustration of the applied technical principles. Those skilled in the art should understand that the scope of the disclosure involved in the present disclosure is not limited to the technical solutions formed by the specific combination of the above technical features, and should also cover other technical solutions formed by any combination of the above technical features or their equivalent features without departing from the above disclosure concept. For example, the technical solutions formed by mutually replacing the above features with the (but not limited to) technical features having similar functions disclosed in the present disclosure.
[0228] In addition, although the operations are depicted in a particular order, this should not be construed as requiring that the operations be performed in the particular order shown or in sequential order. In certain environments, multitasking and parallel processing may be advantageous. Similarly, although several specific implementation details are included in the above discussion, these should not be construed as limiting the scope of the present disclosure. Certain features described in the context of separate embodiments may also be implemented in combination in a single embodiment. Conversely, the various features described in the context of a single embodiment may also be implemented separately or in any suitable sub-combination in multiple embodiments.
[0229] Although the subject matter has been described in language specific to structural features and / or methodological logical acts, it should be understood that the subject matter defined in the appended claims is not necessarily limited to the specific features or acts described above. Rather, the specific features and acts described above are merely example forms for implementing the claims.
Claims
1. A video editing method, comprising: In response to a trigger operation on a video editing template, presenting a video editing interface, wherein the video editing template includes structure information and editing information, the structure information is used to indicate at least one video segment, the editing information is used to indicate at least one editing effect applied to the at least one video segment, and the video editing interface includes paragraph editing areas respectively corresponding to the at least one video segment; In response to a material import operation for a target video segment, presenting identifiers of at least one video material in the paragraph editing area of the target video segment, the at least one video segment includes the target video segment, and the at least one video material is the material imported into the target video segment based on the material import operation; and In response to a trigger operation for editing processing, generating a video editing result according to the video editing template and the at least one video material, wherein a part of the video editing result corresponding to the target video segment is an editing result obtained based on a target editing effect and the at least one video material, and the target editing effect matches the editing effect located within the target video segment among the at least one editing effect.
2. The method according to claim 1, wherein, In response to a trigger operation for editing processing, generating a video editing result according to the video editing template and the at least one video material, including: In response to a trigger operation for editing processing, adjusting the editing effect according to a matching strategy to obtain the target editing effect; According to the video editing template, applying the target editing effect to at least a part of the at least one video material to obtain a video part of the target video segment; and Generating the video editing result based on the video part corresponding to each target video segment in the at least one video segment.
3. The method according to claim 2, wherein, The matching strategy includes: When the editing effect includes the first copywriting, the target editing effect includes the second copywriting, wherein the copywriting format of the second copywriting is the same as that of the first copywriting, and the copywriting content of the second copywriting matches the at least one video material.
4. The method according to claim 3, wherein The matching strategy further includes: In response to the first copywriting being applied to the entire paragraph of the target video segment, the second copywriting is applied to the entire part of the video part; In response to the first copywriting being applied to a partial time period in the target video segment, the second copywriting is applied to a partial time period in the video part.
5. The method according to claim 3, wherein, In response to a trigger operation for editing processing, generating a video editing result according to the video editing template and the at least one video material, further includes: Obtaining the content of the second copywriting, wherein the structure information further includes an information prompt paragraph for obtaining prompt information, and obtaining the content of the second copywriting includes: Analyzing the prompt information and the at least one video material to obtain an analysis result; and Generating the content of the second copywriting according to the analysis result.
6. The method according to claim 5, wherein The influence degree of the analysis result on the content of the second copywriting is greater than the influence degree of the at least one video material on the content of the second copywriting.
7. The method according to claim 2, wherein The matching strategy includes: In the case where the editing effect includes the transition animation, the target editing effect includes the transition animation.
8. The method according to claim 2, wherein The structure information further includes an information prompt paragraph for obtaining prompt information. In response to a trigger operation for editing processing, adjusting the editing effect according to the matching strategy to obtain the target editing effect includes: In response to a trigger operation for editing processing, analyzing the prompt information and the at least one video clip to obtain an analysis result; In response to the editing effect including subtitles and / or dubbing, obtaining the content of the subtitles and / or dubbing in the target editing effect according to the matching strategy of matching the content of the subtitles and / or dubbing in the target editing effect with the analysis result; Obtaining the format of the subtitles in the target editing effect according to the matching strategy that the format of the subtitles in the target editing effect is the same as the format of the subtitles in the editing effect; and Obtaining the voice parameters of the dubbing in the target editing effect according to the matching strategy that the voice parameters of the dubbing in the target editing effect are the same as the voice parameters of the dubbing in the editing effect.
9. The method according to claim 2, wherein, Generating the video editing result based on the video part corresponding to each target video segment in the at least one video segment includes: Splicing the video parts corresponding to each target video segment to obtain the video editing result.
10. The method according to claim 1, wherein, In response to a trigger operation for editing processing, generating a video editing result according to the video editing template and the at least one video clip includes: In response to a trigger operation for editing processing, using a neural network to process the at least one video clip and the editing effect to obtain the video editing result.
11. The method according to claim 2, wherein, The at least one editing result further includes a global editing effect applied to the at least one video segment. In response to the trigger operation of the editing processing, generating a video editing result according to the video editing template and the at least one video clip further includes: In response to the trigger operation of the editing processing, adjusting the time length of the global editing effect to match the time length of the target video clip used in the video editing result to obtain an adjusted global editing effect; Generating the video editing result based on the video part corresponding to each target video segment in the at least one video segment includes: Generating the video editing result based on the adjusted global editing effect and the video part corresponding to each target video segment.
12. The method according to claim 10, wherein The global editing effect includes at least one of music, filters, and subtitles.
13. The method according to claim 1, wherein The structure information indicates each of the at least one video segment by a time interval or the serial number of continuous video clips.
14. The method according to claim 1, wherein, In response to the trigger operation of the editing processing, generating a video editing result according to the video editing template and the at least one video clip includes: In response to the trigger operation of the editing processing, extracting a target segment from the at least one video clip; and Generating the video editing result according to the video editing template and the target segment.
15. The method according to claim 1, wherein, In response to the triggering operation of the editing process, a video editing result is generated according to the video editing template and the at least one video material, including: In response to the triggering operation of the editing process, if there is a first video material among the at least one video material, the first video material is divided into multiple sub - segments, where the first video material is a material with a duration longer than a preset duration; and Based on the video editing template and the multiple sub - segments, the video editing result is generated.
16. The method according to claim 1, further comprising: According to the video editing result, play the preview video on the preview page.
17. The method according to claim 16, wherein In the case where the editing effect only includes subtitles, the dubbing identifier on the preview page is displayed in a closed state, and in the case where the editing effect only includes dubbing, the subtitle identifier on the preview page is displayed in a closed state.
18. The method according to claim 16, wherein The preview page includes a video playback window, a material display window, and an editing entry window, The video playback window is used to play the preview video, the material display window includes multiple material sub - windows for displaying at least a part of each of the multiple video materials in the preview video, and the editing entry window is used to display the entry for editing the video materials.
19. The method according to claim 18, wherein, The material display window further includes a material addition sub - window, The method further comprises: In response to a triggering operation on the material addition sub - window when a material window unit in the multiple material sub - windows is selected, display multiple selectable materials; According to the addition and selection operation on the multiple selectable materials, add the target material after the material window unit.
20. The method according to claim 18, wherein, The editing entry window includes at least one of a subtitle entry for editing subtitles, a filter entry for editing filters, a copywriting entry for editing copywriting, and a music entry for editing music, The editing entry window further includes: a delete icon for selecting and deleting multiple preferred materials in the preview video.
21. The method according to claim 20, wherein, The editing entry window further includes a reorganization entry, and the method further comprises: In response to the selection of the reorganization entry, display a sorting panel, where the sorting panel includes the at least one video material; and Receive an organization operation on the at least one video material in the sorting panel, and reorganize the at least one video material.
22. A video editing device, comprising: A first display unit, configured to display a video editing interface in response to a triggering operation on a video editing template, where the video editing template includes structure information and editing information, the structure information is used to indicate at least one video paragraph, the editing information is used to indicate at least one editing effect applied in the at least one video paragraph, and the video editing interface includes paragraph editing areas corresponding to the at least one video paragraph respectively; A second display unit configured to display, in a paragraph editing area of the target video paragraph, identifiers of at least one video material in response to a material import operation for the target video paragraph, where the at least one video paragraph includes the target video paragraph, and the at least one video material is material imported into the target video paragraph based on the material import operation; and A result generation unit configured to generate a video editing result according to the video editing template and the at least one video material in response to a trigger operation for an editing process, where a part of the video editing result corresponding to the target video paragraph is an editing result obtained based on a target editing effect and the at least one video material, and the target editing effect matches the editing effect located within the target video paragraph among the at least one editing effect.
23. An electronic device, comprising: A processor; A memory including one or more computer program instructions; Wherein, the one or more computer program instructions are stored in the memory and, when executed by the processor, implement the video editing method according to any one of claims 1-21.
24. A computer-readable storage medium that non-temporarily stores computer-readable instructions, wherein, When the computer-readable instructions are executed by the processor, the video editing method according to any one of claims 1-21 is implemented.