Video editing method and apparatus, and electronic device and computer-readable storage medium
By identifying keywords in video copy and applying corresponding text effect templates, text materials are generated, the video editing process is simplified, editing efficiency is improved, and the problem of cumbersome material selection in the existing technology is solved.
Patent Information
- Application Number
- PCT/CN2024/137917
- Authority / Receiving Office
- WO · WO
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2023-12-29
- Filing Date
- 2024-12-09
- Publication Date
- 2025-07-03
AI Technical Summary
During the video editing process, users need to select and match materials from a huge library of materials, which makes the editing process time-consuming and cumbersome.
By identifying keywords in video copy, determining their category labels and matching text effect templates, generating text materials and updating initial video editing data, simplifying the editing process.
Improve the efficiency of video editing, reduce the operation steps of users in material selection and style design, and reduce editing time.
Smart Images

Figure CN2024137917_03072025_PF_FP_ABST
Abstract
Description
Video editing method, device, electronic device and computer-readable storage medium
[0001] This application claims priority to the Chinese invention patent application entitled “Video editing method, device, electronic device and computer-readable storage medium” filed on December 29, 2023, with application number 202311865095.9. The entire contents of that application are incorporated by reference into this application. Technical Field
[0002] The present disclosure relates to the field of video processing technology, and in particular to a video editing method, device, electronic device, and computer-readable storage medium. Background Art
[0003] Video editing is a crucial step in the user's video creation process. By combining text, music, sound effects, stickers, special effects, and other materials, you can enrich the visual and auditory effects of your video, making the message and emotion more prominent. However, video editing is often time-consuming and labor-intensive. Users must select and combine appropriate materials from a vast library, and then design, add, and adjust the video's style, a tedious process. Summary of the Invention
[0004] In order to solve the above technical problems or at least partially solve the above technical problems, the present disclosure provides a video editing method, device, electronic device, storage medium and program product.
[0005] According to a first aspect of an embodiment of the present disclosure, a video editing method is provided, comprising: obtaining a video text based on initial video editing data; determining a target text effect template based on the video text, the target text effect template being used to indicate an editing operation of at least one text effect; generating a text material corresponding to the video text by applying the target text effect template to a target keyword in the video text; wherein the text material is used to present the target keyword as the at least one text effect; and updating the initial video editing data using the text material to obtain target video editing data; wherein the target video editing data is used to present the text material.
[0006] In some embodiments of the present disclosure, determining the target text effect template based on the video copy includes: identifying the target keyword in the video copy; determining the category label of the category to which the target keyword belongs; determining the text effect editing type that matches the category label; and determining the target text effect template from multiple text effect templates included in the text effect editing type.
[0007] In some embodiments of the present disclosure, the target text effect template is determined from the multiple text effect templates included in the text effect editing type, including: when the text effect editing type is a first type, determining the target text effect template from the multiple text effect templates included in the first type; when the text effect editing type is a second type, determining the category identifier of the category to which the video copy belongs; and determining the target text effect template that matches the category identifier from the multiple text effect templates included in the second type.
[0008] In some embodiments of the present disclosure, when the text effect editing type is the second type, after determining the category identifier of the category of the video copy, the method also includes: when there is no text effect template matching the category identifier among the multiple text effect templates included in the second type, determining the first text effect template among the multiple text effect templates included in the second type as the target text effect template.
[0009] In some embodiments of the present disclosure, after determining the category label of the category to which the target keyword belongs, the method further includes: determining that there is a sound effect template matching the category label; and determining the sound effect template as the target text effect template.
[0010] In some embodiments of the present disclosure, obtaining video copy based on initial video editing data includes: displaying multiple copy effect modes, the multiple copy effect modes including a first copy effect mode, the first copy effect mode is used to indicate the generation of subtitles for keywords of the video copy; in response to a selection operation of the first copy effect mode, obtaining video copy based on the initial video editing data.
[0011] In some embodiments of the present disclosure, the multiple copy effect modes also include a second copy effect mode, and the second copy effect mode is used to indicate the generation of full-text subtitles, and the display form of the keywords in the full-text subtitles is different from the display form of the text other than the keywords; after displaying the multiple copy effect modes, the method also includes: in response to the selection operation of the second copy effect mode, obtaining video copy according to the initial video editing data; generating video subtitles from the video copy; displaying the keywords in the video subtitles in a first display form, and displaying the text other than the keywords in the video subtitles in a second display form, the first display form being different from the second display form; and synthesizing the video subtitles and the initial video editing data into first video editing data.
[0012] According to a second aspect of an embodiment of the present disclosure, a video editing device is provided, comprising: an acquisition module for acquiring a video text based on initial video editing data; a determination module for determining a target text effect template based on the video text, wherein the target text effect template is used to indicate an editing operation of at least one text effect; a generation module for generating a text material corresponding to the video text by applying the target text effect template to a target keyword in the video text; wherein the text material is used to present the target keyword as the at least one text effect; and an update module for updating the initial video editing data using the text material to obtain target video editing data; wherein the target video editing data is used to present the text material.
[0013] In some embodiments of the present disclosure, the determination module is specifically used to identify the target keyword in the video copy; determine the category label of the category to which the target keyword belongs; determine the text effect editing type that matches the category label; and determine the target text effect template from multiple text effect templates included in the text effect editing type.
[0014] In some embodiments of the present disclosure, the determination module is specifically used to determine the target text effect template from the multiple text effect templates included in the first type when the text effect editing type is the first type; determine the category identifier of the category to which the video copy belongs when the text effect editing type is the second type; and determine the target text effect template that matches the category identifier from the multiple text effect templates included in the second type.
[0015] In some embodiments of the present disclosure, the determination module is also used to, when the text effect editing type is the second type, determine the category identifier of the category of the video copy, and when there is no text effect template matching the category identifier among the multiple text effect templates included in the second type, determine the first text effect template among the multiple text effect templates included in the second type as the target text effect template.
[0016] In some embodiments of the present disclosure, the determination module is further configured to, after determining the category label of the category to which the target keyword belongs, determine whether there is a sound effect template matching the category label; and determine the sound effect template as the target text effect template.
[0017] In some embodiments of the present disclosure, the acquisition module is specifically used to display multiple copy effect modes, and the multiple copy effect modes include a first copy effect mode, and the first copy effect mode is used to indicate the generation of subtitles for keywords of the video copy; in response to the selection operation of the first copy effect mode, the video copy is acquired according to the initial video editing data.
[0018] In some embodiments of the present disclosure, the multiple copy effect modes also include a second copy effect mode, and the second copy effect mode is used to indicate the generation of full-text subtitles, and the display form of the keywords in the full-text subtitles is different from the display form of the text other than the keywords; the acquisition module is also used to, after displaying the multiple copy effect modes, respond to the selection operation of the second copy effect mode, obtain video copy according to the initial video editing data; generate video subtitles from the video copy; display the keywords in the video subtitles in a first display form, and display the text other than the keywords in the video subtitles in a second display form, the first display form being different from the second display form; and synthesize the video subtitles and the initial video editing data into first video editing data.
[0019] According to a third aspect of an embodiment of the present disclosure, an electronic device is provided, comprising a processor, a memory, and a computer program stored in the memory and executable on the processor, wherein the computer program, when executed by the processor, implements the video editing method as described in the first aspect.
[0020] According to a fourth aspect of the embodiments of the present disclosure, a computer-readable storage medium is provided, on which a computer program is stored. When the computer program is executed by a processor, the video editing method as described in the first aspect is implemented.
[0021] According to a fifth aspect of the embodiments of the present disclosure, a computer program product is provided, wherein the computer program product includes a computer program. When the computer program product runs on a processor, the processor executes the computer program to implement the video editing method as described in the first aspect.
[0022] In a sixth aspect of an embodiment of the present disclosure, a chip is provided, which includes a processor and a communication interface, wherein the communication interface is coupled to the processor, and the processor is used to run program instructions to implement the video editing method as described in the first aspect. BRIEF DESCRIPTION OF THE DRAWINGS
[0023] The accompanying drawings, which are incorporated in and constitute a part of this specification, illustrate embodiments consistent with the present disclosure and, together with the description, serve to explain the principles of the present disclosure.
[0024] In order to more clearly illustrate the embodiments of the present disclosure or the technical solutions in the prior art, the following briefly introduces the drawings required for use in the embodiments or the description of the prior art. Obviously, for ordinary technicians in this field, other drawings can be obtained based on these drawings without any creative work.
[0025] FIG1 is a flow chart of a video editing method provided by an embodiment of the present disclosure;
[0026] FIG2 is a schematic diagram of one interface of the video editing method provided by an embodiment of the present disclosure;
[0027] FIG3 is a second schematic diagram of an interface of a video editing method provided by an embodiment of the present disclosure;
[0028] FIG4 is a third schematic diagram of an interface of a video editing method provided by an embodiment of the present disclosure;
[0029] FIG5 is a fourth schematic diagram of an interface of a video editing method provided by an embodiment of the present disclosure;
[0030] FIG6 is a structural block diagram of a video editing device provided by an embodiment of the present disclosure;
[0031] FIG7 is a structural block diagram of an electronic device provided by an embodiment of the present disclosure. DETAILED DESCRIPTION
[0032] In order to more clearly understand the above-mentioned objectives, features and advantages of the present disclosure, the scheme of the present disclosure will be further described below. It should be noted that the embodiments of the present disclosure and the features therein can be combined with each other in the absence of conflict.
[0033] In the following description, many specific details are set forth to facilitate a full understanding of the present disclosure, but the present disclosure may also be implemented in other ways different from those described herein; it is obvious that the embodiments in the specification are only part of the embodiments of the present disclosure, rather than all of the embodiments.
[0034] The terms "first", "second", etc. in the specification and claims of the present disclosure are used to distinguish similar objects, and are not used to describe a specific order or sequence. It should be understood that the data used in this way can be interchangeable under appropriate circumstances, so that the embodiments of the present disclosure can be implemented in an order other than those illustrated or described herein, and the objects distinguished by "first", "second", etc. are generally of the same type, and the number of objects is not limited. For example, the first object can be one or more. In addition, "and / or" in the specification and claims represents at least one of the connected objects, and the character " / " generally indicates that the objects related to each other are in an "or" relationship.
[0035] The electronic devices in the embodiments of the present disclosure may be mobile electronic devices or non-mobile electronic devices. Mobile electronic devices may include mobile phones, tablet computers, laptop computers, PDAs, in-vehicle electronic devices, wearable devices, ultra-mobile personal computers (UMPCs), netbooks, or personal digital assistants (PDAs); non-mobile electronic devices may include personal computers (PCs), televisions (TVs), ATMs, or self-service kiosks; and the embodiments of the present disclosure do not specifically limit these.
[0036] The execution subject of the video editing method provided by the embodiment of the present disclosure can be the above-mentioned electronic device (including mobile electronic devices and non-mobile electronic devices), or it can be a functional module and / or functional entity in the electronic device that can implement the video editing method. The specific execution subject can be determined according to actual usage requirements and is not limited by the embodiment of the present disclosure.
[0037] The video editing method provided by the embodiment of the present disclosure is described in detail below through specific embodiments and application scenarios in conjunction with the accompanying drawings.
[0038] As shown in FIG1 , an embodiment of the present disclosure provides a video editing method, which may include the following steps 101 to 104 .
[0039] 101. Obtain video text based on initial video editing data.
[0040] The initial video editing data can be a video, in which case subtitles are identified from the video to obtain the video text. Alternatively, the initial video editing data can be a video draft containing text snippets. The text snippets are text assets generated based on the video text (text information) and text effects (think of it as a video of text images). The video text can then be identified from the text snippets in the video draft. The specific method can be determined based on actual circumstances and is not limited here.
[0041] 102. Determine a target text effect template based on the video copy.
[0042] The target text effect template is used to indicate an editing operation of at least one text effect.
[0043] Different video copy can correspond to different text effect templates. Each text effect template is used to indicate at least one text effect editing operation. Text effect editing operations can include editing operations for the text display format, and can also include editing operations for the sound effects corresponding to the text. The specific operation can be determined based on actual conditions and is not limited here.
[0044] In some embodiments of the present disclosure, the above step 102 can be specifically implemented through the following steps 102a to 102d.
[0045] 102a. Identify the target keyword in the video copy.
[0046] In the embodiment of the present disclosure, different categories of keywords are set, and it is possible to traverse the video text to see whether one or more keywords of a certain category exist. In the embodiment of the present disclosure, keywords in the video text can also be identified by a keyword recognition model.
[0047] 102b. Determine the category label of the category to which the target keyword belongs.
[0048] In some embodiments, the categories to which keywords belong can be divided into numbers, key nouns, and adjectives according to the first-level classification. According to the second-level classification, the categories to which keywords belong can include numbers such as time, money, quantity, fraction, proportion, etc.; key nouns can include product names, material names, quantifiers, brand names, store names, abbreviations, personal names, identities, proper nouns, place names, etc.; adjectives can include positive evaluations, negative evaluations, positive emotions, negative emotions, etc. The categories to which keywords belong can also be divided in other ways, or can also include other categories. The specific method can be determined according to actual conditions and is not limited here.
[0049] 102c. Determine the text effect editing type that matches the category label.
[0050] In the disclosed embodiment, there are multiple text effect editing types. Different category tags can match the same text effect editing type or different text effect editing types. The specific type can be determined based on actual conditions and is not limited here.
[0051] 102d. Determine the target text effect template from multiple text effect templates included in the text effect editing type.
[0052] In the embodiment of the present disclosure, different text effect editing types include different text effect templates.
[0053] In the disclosed embodiments, different keywords may belong to different category tags. Different category tags may match different text effect editing types. Different text effect editing types may include different text effect templates. Therefore, different text effect templates may be matched to different keywords, thereby increasing the diversity of matched text effect templates, increasing the diversity of text effect editing, and thus improving text editing effects.
[0054] In some embodiments of the present disclosure, text effect editing types may include a first type and a second type, and may also include other types, or may be divided according to other rules. The specific types may be determined based on actual conditions and are not limited here.
[0055] In some embodiments, category tags matching the first type may include time, place, subscription / follow, like, comment, view description, coupon, sponsor, positive evaluation, negative evaluation, positive emotion, negative emotion, etc.
[0056] In some embodiments of the present disclosure, the multiple text effect templates included in the first type may no longer be divided into types. The multiple text effect templates included in the first type may also be divided into text effect templates of types such as time, place, subscription / follow, like, comment, view description, coupon, sponsor, positive evaluation, negative evaluation, positive emotion, and negative emotion; the specific type can be determined based on actual conditions and is not limited here.
[0057] In some embodiments, category tags matching the second type may include video blog, technology, fashion, film and television, animation, news, etc.
[0058] In some embodiments of the present disclosure, the multiple text effect templates included in the second type may no longer be divided into types, and the multiple text effect templates included in the second type may also be divided into text effect templates of video blogs, technology, fashion, film and television, animation, news and other types; the specific types can be determined based on actual conditions and are not limited here.
[0059] In some embodiments of the present disclosure, the above step 102d can be specifically implemented through the following steps 102d1 to 102d3.
[0060] 102d1. When the text effect editing type is the first type, determine the target text effect template from multiple text effect templates included in the first type.
[0061] In some embodiments, any one of the plurality of text effect templates included in the first type may be determined as the target text effect template. Alternatively, a plurality of first text effect editing identifiers may be displayed to the user, and the first text effect template indicated by the first text effect editing identifier selected by the user may be determined as the target text effect template.
[0062] For example, the category label of the target keyword is subscription. If the subscription matches the first type, a text effect template is selected from the multiple text effect templates included in the first type as the text effect template corresponding to the subscription.
[0063] 102d2. When the text effect editing type is the second type, determine the category identifier of the category to which the video copy belongs.
[0064] 102d3. Determine the target text effect template that matches the category identifier from the multiple text effect templates included in the second type.
[0065] In some embodiments, the plurality of text effect templates included in the second type are a plurality of second text effect templates. The plurality of second text effect templates include at least one second text effect template that matches the category identifier. Any one of the at least one second text effect template can be determined as the target text effect template, or at least one second text editing identifier can be displayed to the user, and the second text effect template indicated by the second text editing identifier selected by the user can be determined as the target text effect template.
[0066] For example, if the category label of the category to which the target keyword belongs is a key noun or number (except time), etc., and the matched text effect editing type is the second type, then according to the category identifier of the category to which the initial video editing data belongs, a text effect template suitable for the category identifier can be randomly selected as the target text effect template, and the entire initial video editing data can uniformly use the same text effect template for editing operations.
[0067] In some embodiments of the present disclosure, after the above 102d2, the video editing method provided by the embodiments of the present disclosure may further include the following 102d4.
[0068] 102d4. If there is no text effect template matching the category identifier among the multiple text effect templates included in the second type, determine the first text effect template among the multiple text effect templates included in the second type as the target text effect template.
[0069] The first text effect template is a universal text effect template among the text effect templates included in the second type, that is, a text effect template applicable to any category to which the video copy belongs. Therefore, when no text effect template corresponding to the category identifier is matched, a random text effect template is selected from the universal text effect templates as the target text effect template.
[0070] In the disclosed embodiment, the target keyword may be subjected to text effect editing processing corresponding to the target text effect template, thereby achieving text effect editing processing on the initial video editing data to obtain the target video editing data.
[0071] In some embodiments of the present disclosure, after the above 102b, the video editing method provided by the embodiment of the present disclosure may further include the following 102e and step 102f.
[0072] 102e. Determine whether there is a sound effect template matching the category label.
[0073] 102f. Determine the sound effect template as the target text effect template.
[0074] In some embodiments, when the target keyword's category tag is time, money, or score, there is a sound effect template that matches the category tag. For example, if the category tag is time, the corresponding sound effect template may include a countdown sound effect or an alarm clock sound effect. If the category tag is money, the corresponding sound effect template may include a cash register sound effect. If the category tag is score, the corresponding sound effect template may include a scoring sound effect or a cheering sound effect.
[0075] In the embodiment of the present disclosure, there is no restriction on which category tags have matching sound effect templates, nor is there a restriction on the sound effects of the sound effect templates, which can be determined based on actual conditions.
[0076] It is understood that for some keywords that may have sound effect characteristics, text effect editing and sound effect editing can be performed simultaneously. For example, for a keyword in the time category, text effect editing can be performed based on the first type of text effect template corresponding to the time category, and sound effect editing can also be performed based on the sound effect template of the countdown sound effect or alarm sound effect corresponding to the time category.
[0077] In the disclosed embodiment, text effect editing and sound effect editing are performed simultaneously on certain categories of keywords, which can effectively increase the diversity of video editing.
[0078] In some embodiments of the present disclosure, the same video editing process can be performed on each keyword identified from the initial video editing data, or different video editing processes can be performed according to the category to which the identified keyword belongs. The specific process can be determined based on actual conditions and is not limited here.
[0079] 103. Generate text material corresponding to the video copy by applying the target text effect template to the target keyword in the video copy.
[0080] In some embodiments, the text material is used to present the target keyword as the at least one text effect.
[0081] It can be understood that the target keyword in the video copy is edited by performing at least one text effect indicated by the target text effect template to generate text material.
[0082] 104. Update the initial video editing data using the text material to obtain target video editing data. The target video editing data is used to present the text material.
[0083] The present disclosure determines a target text effect template based on the video copy, and applies the target text effect template to the target keywords in the video copy to generate text materials corresponding to the video copy. This enables text effect editing of the keywords in the video copy of the initial video editing data based on the target text effect template, thereby generating target video editing data including text materials corresponding to the keywords. In this way, video editing can be quickly achieved without the user having to manually obtain the video copy of the initial video editing data, and achieving editing operations on the keywords in the video copy to generate target video editing data. This simplifies the user's operation process for video editing processing of the initial video editing data, reduces the time consumption of video editing processing, and improves the processing efficiency of video editing.
[0084] In some embodiments of the present disclosure, the above step 101 can be specifically implemented through the following steps 101a and 101b.
[0085] 101a: Display multiple text effect modes, including a first text effect mode, which is used to instruct the generation of subtitles for keywords in the video text.
[0086] The first text effect mode is specifically the generation of target video editing data including text materials from step 102 to step 104 above.
[0087] 101b. In response to a selection operation on a first copy effect mode, obtaining a video copy according to the initial video editing data.
[0088] In the disclosed embodiment, a variety of text effect modes are provided, and the user can select the required text effect mode according to the needs, perform corresponding video editing on the initial video editing data, and obtain target video editing data that meets the requirements.
[0089] In some embodiments of the present disclosure, the multiple text effect modes also include a second text effect mode. The second text effect mode is used to indicate the generation of full-text subtitles, where the display format of keywords in the full-text subtitles is different from the display format of text other than the keywords. After step 101a above, the video editing method provided in the embodiments of the present disclosure may also include steps 105 to 107 described below.
[0090] 105. In response to a selection operation of a second copy effect mode, obtain a video copy according to the initial video editing data.
[0091] 106. Generate video subtitles based on the video text.
[0092] Keywords in the video subtitles are displayed in a first display form, and text other than the keywords in the video subtitles is displayed in a second display form, where the first display form is different from the second display form.
[0093] In the embodiments of the present disclosure, the specific display forms of the first display form and the second display form are not limited and can be determined according to actual conditions.
[0094] In some embodiments of the present disclosure, the first display form and the second display form may differ in at least one of the following display parameters: font type, font style, font color, border, shadow, background, glow, bubble, stroke, animation, font size, and font rotation angle.
[0095] In some embodiments of the present disclosure, the first display form and the second display form may be preset or determined separately according to different initial video editing data. The specific display forms may be determined according to actual conditions and are not limited here.
[0096] 107. Combine the video subtitles and the initial video editing data into first video editing data.
[0097] In the disclosed embodiment, selecting the second text effect mode triggers video editing of the initial video editing data, thereby generating first video editing data including video subtitles displayed in different display formats for keywords and non-keywords. This simplifies the video editing process and improves video editing efficiency.
[0098] In some embodiments of the present disclosure, for the second text effect mode, if the keyword is a keyword with sound effect characteristics, the keyword can also be edited with corresponding sound effects.
[0099] In some embodiments of the present disclosure, the above step 101a can be specifically implemented through the following step 101a1.
[0100] 101a1. In response to a triggering operation on a video processing control displayed on a video editing interface, a video editing panel is displayed. The video editing panel displays at least one text effect mode, each text effect mode being used to indicate an editing mode of a video text.
[0101] In some embodiments of the present disclosure, at least one text effect mode may include the first text effect mode and the second text effect mode. The first text effect mode corresponds to no full-text subtitles, where subtitles are generated only for keywords in the video text; the second text effect mode corresponds to full-text subtitles, where keywords and non-keywords (text other than keywords) in the full-text subtitles are displayed in different formats.
[0102] In some embodiments of the present disclosure, when a user first enters the video editing panel, the default selected text effect mode is determined based on the display ratio (horizontal or vertical display format) of the initial video editing data. For example, the first text effect mode is selected by default for horizontal display, and the second text effect mode is selected by default for vertical display. Upon subsequent re-entry of the video editing panel, the text effect mode previously selected by the user is determined as the default selected text effect mode.
[0103] The present disclosure displays a video processing control on a video editing interface, triggers the display of a video editing panel through the video processing control, and triggers a target text effect mode in at least one text effect mode displayed in the video editing panel to trigger the corresponding text effect editing processing of the initial video editing data. This eliminates the need for the user to select and match appropriate materials from a large material library and perform style adjustments to achieve text effect editing processing of the initial video editing data. In this way, the user's operation process for performing text effect editing processing on the initial video editing data is simplified, the time consumption for video editing of the initial video editing data is reduced, and the processing efficiency of video editing of the initial video editing data is improved.
[0104] In some embodiments of the present disclosure, when the keywords in the video text of the initial video editing data include one or more keywords with sound effect characteristics, the keywords can be edited for text effects while also editing for sound effects. Thus, when a keyword is encountered during playback of the target video editing data (or the first video editing data), the sound effect corresponding to the keyword is output.
[0105] In the disclosed embodiment, the sound effect editing mode (corresponding to the sound effect editing function) can be set separately in the video editing panel, or the sound effect editing function can be attached to the text effect mode. The specific details can be determined based on actual conditions and are not limited here.
[0106] In some embodiments of the present disclosure, a sound effect editing mode is also displayed in the video editing panel, or a sound effect editing option (selected by default, i.e., the sound effect editing function is enabled by default) and a delete existing subtitles option (selected by default, i.e., the delete existing subtitles function is enabled by default) are also displayed in the video editing panel. Among them, the sound effect editing option in a selected state is used to indicate that sound effect editing is performed while text effect editing is performed. The sound effect editing option that is not in a selected state is used to indicate that sound effect editing is not performed when text effect editing is performed. The delete existing subtitles option in a selected state is used to indicate that the existing subtitles of the initial video editing data are deleted before text effect editing is performed on the initial video editing data. The delete existing subtitles option in an unselected state is used to indicate that when text effect editing is performed on the initial video editing data, the existing subtitles of the initial video editing data are used to perform text effect editing processing.
[0107] In the disclosed embodiment, whether to perform sound effect editing on the initial video editing data can be determined based on the sound effect editing mode or the sound effect editing option. Furthermore, whether to use the existing subtitles in the initial video editing data for text effect editing can be determined based on the option to delete existing subtitles. This provides users with a variety of optional functions that can be set according to actual usage needs, thereby improving the user experience.
[0108] If the user has already added subtitles before using the video editor, the Clear Existing Subtitles option will be displayed when the second text effect mode is selected. After checking Clear Existing Subtitles, the existing subtitles will be cleared when the video editing process begins. If the user has not added subtitles before using the video editor, or if the user selects the first text effect mode, the Clear Existing Subtitles option will not be displayed or will be grayed out and unavailable for selection. The specific method can be determined based on actual circumstances and is not limited here.
[0109] In some embodiments of the present disclosure, a commercial material filtering option is also displayed in the video editing panel, which is not selected by default (i.e., the commercial material filtering function is turned off by default). After the commercial material filtering function is turned on, the video editor only returns commercially available materials.
[0110] In some embodiments of the present disclosure, when the video editing panel is displayed for the first time, the commercial material screening option is unselected by default. When the video editing panel is displayed subsequently, it is displayed according to the status of the last recorded commercial material screening option. That is, if the status of the last recorded commercial material screening option is unselected, it is unselected by default this time; if the status of the last recorded commercial material screening option is selected, it is selected by default this time.
[0111] In some embodiments of the present disclosure, the video editing panel also displays a function introduction icon corresponding to each text effect mode. The function introduction icon is used to trigger the display of a function introduction to the text effect editing function indicated by the corresponding text effect mode. The function introduction includes an introduction in text form and / or an introduction in animated form. In this way, if the user does not understand the text effect editing function corresponding to each text effect mode, the function introduction to the text effect editing function indicated by the text effect mode can be triggered by triggering the function introduction icon corresponding to the text effect mode, thereby facilitating user understanding and use and improving user experience.
[0112] In some embodiments of the present disclosure, a function introduction interface corresponding to the function introduction icon is displayed in a floating manner in the video editing panel. The function introduction interface can be a pop-up window display or the following function introduction floating layer, which is not limited here.
[0113] In some embodiments of the present disclosure, the above step 101a1 can be specifically implemented through the following steps 101a1a and 101a1b.
[0114] 101a1a. In response to a triggering operation on a video processing control displayed on the video editing interface, a function introduction floating layer is displayed. The function introduction floating layer is used to display the entry control of the video editing panel and the function introduction of the text effect editing function indicated by each text effect mode. The function introduction includes a text introduction and / or an animated graphic introduction.
[0115] In some embodiments of the present disclosure, the function introduction corresponding to each text effect mode can be displayed separately in the function introduction floating layer (the function introduction corresponding to different text effect modes can be switched by switching operations). The function introduction corresponding to each text effect mode in the function introduction floating layer can also be displayed simultaneously. The specific method can be determined according to actual conditions and is not limited here.
[0116] The specific function introduction can be set according to the corresponding copy effect mode, and is not limited here.
[0117] For example, the function introduction overlay is a half-screen function introduction overlay. A GIF-based introduction corresponding to a text effect mode is displayed above the overlay, and a text-based introduction corresponding to that text effect mode is displayed below. Users can swipe left or right to switch between GIF-based introductions corresponding to different text effect modes. After switching between GIF-based introductions, the text-based introduction below is updated synchronously.
[0118] 101a1b. In response to a triggering operation on the entry control, the video editing panel is displayed.
[0119] Exemplarily, the entry control may be a “try it” control.
[0120] In some embodiments of the present disclosure, it can be arranged that when a new user (after downloading and installing) uses the video editing function for the first time, the above-mentioned step 101a1 includes the above-mentioned steps 101a1a and 101a1b. When an old user uses the video editing function, the video editing panel is directly displayed by triggering the video processing control, and the above-mentioned function introduction floating layer is not displayed.
[0121] In the embodiment of the present disclosure, before displaying the video editing panel, displaying a function introduction floating layer helps users understand and compare the text effect editing functions corresponding to each copy effect mode, and then facilitates the subsequent selection of which copy effect mode to use for text effect editing processing of the initial video editing data, which can improve the user experience.
[0122] In some embodiments of the present disclosure, the above step 101b can be specifically carried out through the following steps 101b1 to 101b5.
[0123] 101b1. In response to a triggering operation on the first text effect mode, determine whether the video editing function has been authorized.
[0124] If it is determined that the video editing function is not authorized, the following steps 101b2 to 101b5 are executed; if it is determined that the video editing function is authorized, the following step 101b4 is executed.
[0125] 101b2. If the video editing function is not authorized, an authorization pop-up window is displayed, in which an authorization option and a non-authorization option are displayed.
[0126] The authorization option is used to indicate that the video editing function is authorized, and the non-authorization option is used to indicate that the video editing function is not authorized.
[0127] Exemplarily, the authorization pop-up window may include text-based introduction information about the video editing function, as well as an "Allow" option and a "Do Not Allow" option, wherein the "Allow" option corresponds to the authorization option, and the "Do Not Allow" option corresponds to the non-authorization option.
[0128] 101b3. In response to the triggering operation on the authorization option, determine to authorize the video editing function.
[0129] In response to the triggering operation of the authorization option, the electronic device is authorized to perform the following step 101b4 for the video editing function.
[0130] 101b4. When the video editing function is authorized, obtain video text according to the initial video editing data based on the first text effect mode.
[0131] 101b5. In response to the triggering operation of the unauthorized option, prohibiting the text effect editing processing of the initial video editing data.
[0132] It can be understood that in response to the triggering operation of the unauthorized option, the electronic device may not perform text effect editing processing on the initial video editing data, and may also display a prompt message, which is used to prompt that the video editing function cannot be used without authorization.
[0133] It is understood that when clicking on the first text effect mode, if it is detected that the electronic device has not authorized the video editing function, an authorization pop-up window will pop up. Select the authorization option to start text effect editing processing on the initial video editing data. Select the unauthorized option to close the authorization pop-up window and not start (prohibit) text effect editing processing on the initial video editing data.
[0134] In some embodiments of the present disclosure, if the authorization option is selected this time, the next time the video editing function is used, the authorization pop-up window will not be displayed; if the authorization option is not selected this time, the authorization pop-up window will still be displayed the next time the video editing function is used. The video editing function can only be used after the user selects the authorization option.
[0135] It should be noted that when the user selects the second text effect mode, the relevant authorization judgment process in the above steps 101b1 to 101b5 will also be executed, which will not be repeated here.
[0136] In the embodiment of the present disclosure, by setting the video editing function to be allowed when authorized, and prohibited when not authorized, the user's safety can be guaranteed, user misoperation can be avoided, and the user experience can be improved.
[0137] In some embodiments of the present disclosure, it is also possible to determine whether the video editing function has been authorized in response to a triggering operation of a video processing control. If the video editing function is not authorized, an authorization pop-up window is displayed, which displays an authorization option and a non-authorization option. If the video editing function is authorized, a video editing panel is displayed. The specific determination can be based on actual circumstances and is not limited here.
[0138] In some embodiments of the present disclosure, the video editing method provided by the embodiments of the present disclosure may further include at least one of the following steps 108 to 110.
[0139] 108. During the process of performing text effect editing processing on the initial video editing data, in response to a triggering operation on a cancel control, the process of performing text effect editing processing on the initial video editing data is stopped, and the performed text effect editing processing results are cleared.
[0140] The flow of text effect editing processing includes the relevant processes in steps 101 to 107 described above.
[0141] The cancel control is a control used to end the text effect editing process.
[0142] It is understood that the video editing method provided by the embodiments of the present disclosure supports terminating the text effect editing process during the text effect editing process. In this way, if the user suddenly does not want to perform the text effect editing process, they can trigger the cancel control to end the text effect editing process, which can improve the user experience.
[0143] 109. During the process of performing text effect editing processing on the initial video editing data, in response to the triggering operation of the target control, the process of performing text effect editing processing on the initial video editing data is switched to background execution, and the process progress of performing text effect editing processing on the initial video editing data is displayed at a preset position.
[0144] In some embodiments, the process progress may be a progress percentage, or may be process description information of the current processing stage in the text effect editing process.
[0145] The target control is a control used to switch the text effect editing process to background execution. For example, the target control can be a control that switches to the tool homepage of an application with video editing capabilities. The target control can also be a control that switches to the homepage of an electronic device, or other types of controls. The specific target control can be determined based on actual circumstances and is not limited here.
[0146] It is understood that the video editing method provided in the embodiments of the present disclosure supports switching the text effect editing process to background operation during the text effect editing process. For example, during the text effect editing process, by triggering an exit to the tool homepage, the tool homepage is displayed, and the text effect editing process is kept running in the background.
[0147] In some embodiments, the preset location can be the draft cover of the tool homepage, or a hanging window control displayed in an application with video editing functions, or a hanging window control displayed on an electronic device, or other locations or areas. The specific location can be determined based on actual conditions and is not limited here.
[0148] Exemplarily, in response to a trigger operation of exiting to the tool homepage, the text effect editing processing flow is kept running in the background, and the progress is displayed on the draft cover of the tool homepage.
[0149] The disclosed embodiment supports switching the process of text effect editing processing to background execution, so that during the process of text effect editing processing on the initial video editing data, the user can also perform other operations, which can improve the user experience.
[0150] 110. During the process of performing text effect editing processing on the initial video editing data, a loading animation is displayed. The loading animation includes flow description information of performing text effect editing processing on the initial video editing data.
[0151] In some embodiments of the present disclosure, the loading animation may be scrolling the process description information once every preset time period (such as 5 seconds) (ie, switching to display the process description information).
[0152] In some embodiments of the present disclosure, the loading animation may also display the process description information in real time according to the actual text effect editing process flow, that is, the process description information of the stage at which the text effect editing process flow is currently located is displayed.
[0153] Exemplarily, the process description information can be "identifying keywords...", "analyzing the picture scene...", "searching for good-looking text styles...", "matching appropriate sound effects...", etc. The specific process description information can be determined according to actual conditions and is not limited here.
[0154] In the disclosed embodiment, during the text effect editing process, displaying a loading animation can help users understand the general process of the text effect editing process and improve the user experience.
[0155] In some embodiments of the present disclosure, the video editing method provided by the embodiments of the present disclosure may further include at least one of the following steps 111 and 112.
[0156] 111. Display target prompt information.
[0157] If the text effect editing process for the initial video editing data is successful, the target prompt information is used to prompt that the text effect editing process for the initial video editing data is successful. If the text effect editing process for the initial video editing data fails, the target prompt information is used to prompt that the text effect editing process for the initial video editing data fails and / or the reason for the failure of the text effect editing process for the initial video editing data.
[0158] In the disclosed embodiment, after the text effect editing processing is successfully performed on the initial video editing data (the material is successfully added), a prompt message "Video editing successful" or "Text effect editing successful" may be displayed.
[0159] In the disclosed embodiment, after the text effect editing processing of the initial video editing data fails (the addition of materials fails), a prompt message "Video editing failed" or "Text effect editing processing failed" may be popped up.
[0160] In some embodiments of the present disclosure, if video editing fails, the reason for the failure can be displayed. For example, if no material (subtitles / text templates) is returned for the text portion due to the lack of audio content, a pop-up window may appear indicating "No speech content was recognized, and text effects cannot be added for editing." If recognition fails due to network issues, a pop-up window may appear indicating "Network error, please try again."
[0161] In the disclosed embodiment, by displaying target prompt information, the user can be promptly prompted with the result of the video editing process, which facilitates the user to perform subsequent operations and improves the user experience.
[0162] 112. Display target status information.
[0163] In some embodiments, if the text effect editing process is successfully performed on the initial video editing data, the target state information is used to indicate that the text effect editing process is successfully performed on the initial video editing data. If the text effect editing process is unsuccessful, the target state information is used to indicate that the text effect editing process is unsuccessful.
[0164] The disclosed embodiment supports displaying target status information corresponding to the success or failure of text effect editing processing at a preset position (refer to the above description of the preset position, which will not be repeated here). In this way, the user can be prompted whether the text effect editing processing of the initial video editing data is successful or failed. In this way, when the user cannot pay attention to the text effect editing processing flow of the initial video editing data in real time, the result of the text effect editing processing of the initial video editing data can be understood based on the target status information, which can improve the user experience.
[0165] In the disclosed embodiments, target state information may be displayed at a preset location for a certain period of time. For example, the target state information may be displayed at the preset location for the lifecycle of the corresponding application and may no longer be displayed after the application is restarted. The target state information may also be undisplayed after entering the timeline editing panel.
[0166] In an embodiment of the present disclosure, the target status information indicating that the text effect editing process has failed can also be canceled after the user triggers the process of re-editing the text effect editing process on the initial video editing data through a trigger operation (such as clicking a retry control, clicking a redo control, or clicking any text effect editing control).
[0167] In some embodiments of the present disclosure, after the above step 112, the video editing method provided by the embodiments of the present disclosure may further include any one of the following steps 113 to 116.
[0168] 113. When it is detected that the application corresponding to the video editing is closed, clear the target status information.
[0169] 114. When detecting that the timeline editing panel of the video editing interface is displayed, clear the target state information and position the timeline preview axis to the first position on the timeline where text effect editing has been performed.
[0170] In some embodiments, positioning the timeline preview axis to the first position on the timeline where text effect editing has been performed can facilitate the user in determining at which position the text effect editing process was started, thereby improving the user experience.
[0171] 115. When it is detected that text effect editing is performed again on the initial video editing data, the target state information and the target video editing data are cleared.
[0172] In some embodiments, the text effect editing process is performed again on the initial video editing data. This can be done by using the video editor again after the video editor successfully adds the material; or by clicking the redo control, which is not limited here.
[0173] Clearing the target state information and target video editing data automatically clears the results of the previous video editing process, which is equivalent to overwriting the old results with the new results of the text effect editing process. This can prevent the results of the two video editing processes from being mixed together, resulting in the final video editing result containing different video editing styles and degrading the user experience.
[0174] 116. When it is detected that the text effect editing process performed on the initial video editing data is undone, the target video editing data is cleared.
[0175] The embodiment of the present disclosure supports the operation of undoing the text effect editing process that was previously performed.
[0176] It is understood that adding material using video editing is a whole operation step. When the user clicks the undo control, all video editing materials generated by the previous video editing process must be completely removed.
[0177] In some embodiments of the present disclosure, a free icon can be displayed at the video processing control according to actual needs, and the function can be used normally after clicking it.
[0178] Exemplarily, as shown in FIG2 , the area indicated by mark “21” is the area for displaying the initial video editing data, the area indicated by mark “22” is the timeline editing panel, and the area indicated by mark “23” is the toolbar area for editing processing, wherein the toolbar area includes a “video processing” control. Click the “Video Processing” control shown in FIG2 to display the function introduction floating layer as indicated by mark “31” in FIG3 . The function introduction floating layer includes an introduction area in the form of an animated image for copy effect mode 1 as indicated by mark “311”, an introduction area in the form of a text for copy effect mode 1 as indicated by mark “312”, and a “Try it” control as indicated by mark “313”. Click the “Try it” control shown in FIG3 to display the video editing panel as indicated by mark “41” in FIG4 . The video editing panel includes text effect mode 1 (currently selected), function introduction icon 1 corresponding to text effect mode 1, text effect mode 2 (currently unselected), function introduction icon 2 corresponding to text effect mode 2, Delete existing subtitles option (Delete existing subtitles option is currently selected), and a "Start" control. Clicking the "Start" control shown in Figure 4 displays the video editing loading animation area indicated by the mark "51" in Figure 5, and a pop-up window displays a prompt message that prompts "Video editing still needs to run for a while, do you need to return to the tool homepage and run it in the background?" The user can choose "Run in the background" or "Wait" based on the prompt message.
[0179] Figure 6 is a structural block diagram of a video editing device shown in an embodiment of the present disclosure. As shown in Figure 6, it includes: an acquisition module 601, used to obtain a video copy based on initial video editing data; a determination module 602, used to determine a target text effect template based on the video copy, and the target text effect template is used to indicate an editing operation of at least one text effect; a generation module 603, used to generate a text material corresponding to the video copy by applying the target text effect template to a target keyword in the video copy; wherein the text material is used to present the target keyword as the at least one text effect; an update module 604, used to update the initial video editing data through the text material to obtain target video editing data; wherein the target video editing data is used to present the text material.
[0180] In some embodiments of the present disclosure, the determination module 602 is specifically used to identify the target keyword in the video copy; determine the category label of the category to which the target keyword belongs; determine the text effect editing type that matches the category label; and determine the target text effect template from multiple text effect templates included in the text effect editing type.
[0181] In some embodiments of the present disclosure, the determination module 602 is specifically used to determine the target text effect template from the multiple text effect templates included in the first type when the text effect editing type is the first type; determine the category identifier of the category to which the video copy belongs when the text effect editing type is the second type; and determine the target text effect template that matches the category identifier from the multiple text effect templates included in the second type.
[0182] In some embodiments of the present disclosure, the determination module 602 is also used to, when the text effect editing type is the second type, determine the category identifier of the category of the video copy, and when there is no text effect template matching the category identifier among the multiple text effect templates included in the second type, determine the first text effect template among the multiple text effect templates included in the second type as the target text effect template.
[0183] In some embodiments of the present disclosure, the determination module 602 is further configured to, after determining the category label of the category to which the target keyword belongs, determine whether there is a sound effect template matching the category label; and determine the sound effect template as the target text effect template.
[0184] In some embodiments of the present disclosure, the acquisition module 601 is specifically used to display multiple copy effect modes, and the multiple copy effect modes include a first copy effect mode, and the first copy effect mode is used to indicate the generation of subtitles for keywords of the video copy; in response to the selection operation of the first copy effect mode, the video copy is obtained according to the initial video editing data.
[0185] In some embodiments of the present disclosure, the multiple copy effect modes also include a second copy effect mode, and the second copy effect mode is used to indicate the generation of full-text subtitles, and the display form of the keywords in the full-text subtitles is different from the display form of the text other than the keywords; the acquisition module 601 is also used to, after displaying the multiple copy effect modes, respond to the selection operation of the second copy effect mode, obtain video copy according to the initial video editing data; generate video subtitles from the video copy; display the keywords in the video subtitles in a first display form, and display the text other than the keywords in the video subtitles in a second display form, the first display form being different from the second display form; and synthesize the video subtitles and the initial video editing data into first video editing data.
[0186] In the embodiments of the present disclosure, each module can implement the video editing method provided by the above method embodiments and can achieve the same technical effect. To avoid repetition, it will not be described here.
[0187] FIG7 is a schematic structural diagram of an electronic device provided in an embodiment of the present disclosure, which is used to exemplify an electronic device that implements any video editing method in an embodiment of the present disclosure and should not be understood as a specific limitation on the embodiment of the present disclosure.
[0188] As shown in Figure 7, electronic device 700 may include a processor (e.g., a central processing unit, a graphics processing unit, etc.) 701, which can perform various appropriate actions and processes according to a program stored in a read-only memory (ROM) 702 or a program loaded from a storage device 708 into a random access memory (RAM) 703. Various programs and data required for the operation of electronic device 700 are also stored in RAM 703. Processor 701, ROM 702, and RAM 703 are connected to each other via a bus 704. An input / output (I / O) interface 705 is also connected to bus 704.
[0189] Typically, the following devices may be connected to the I / O interface 705: an input device 706 including, for example, a touch screen, a touchpad, a keyboard, a mouse, a camera, a microphone, an accelerometer, a gyroscope, etc.; an output device 707 including, for example, a liquid crystal display (LCD), a speaker, a vibrator, etc.; a storage device 708 including, for example, a magnetic tape, a hard disk, etc.; and a communication device 709. The communication device 709 may allow the electronic device 700 to communicate with other devices wirelessly or by wire to exchange data. Although the electronic device 700 is shown as having various devices, it should be understood that it is not required to implement or have all of the devices shown. More or fewer devices may be implemented or have alternatively.
[0190] In particular, according to an embodiment of the present disclosure, the process described above with reference to the flowchart can be implemented as a computer software program. For example, an embodiment of the present disclosure includes a computer program product, which includes a computer program carried on a non-transitory computer-readable medium, and the computer program includes a program code for executing the method shown in the flowchart. In such an embodiment, the computer program can be downloaded and installed from the network through the communication device 709, or installed from the storage device 708, or installed from the ROM 702. When the computer program is executed by the processor 701, the functions defined in any video editing method provided by the embodiment of the present disclosure can be executed.
[0191] It should be noted that the computer-readable medium mentioned above in the present disclosure may be a computer-readable signal medium or a computer-readable storage medium, or any combination of the two. Computer-readable storage media may be, for example, but not limited to, electrical, magnetic, optical, electromagnetic, infrared, or semiconductor systems, devices, or components, or any combination of the above. More specific examples of computer-readable storage media may include, but are not limited to: an electrical connection with one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the above. In the present disclosure, a computer-readable storage medium may be any tangible medium that contains or stores a program that can be used by or in conjunction with an instruction execution system, device, or component. In the present disclosure, a computer-readable signal medium may include a data signal propagated in baseband or as part of a carrier wave, which carries computer-readable program code. Such a propagated data signal may take a variety of forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination of the above. A computer-readable signal medium may also be any computer-readable medium other than a computer-readable storage medium that can transmit, propagate, or transport a program for use by or in conjunction with an instruction execution system, apparatus, or device. The program code contained on the computer-readable medium may be transmitted using any suitable medium, including but not limited to wires, optical cables, RF (radio frequency), etc., or any suitable combination thereof.
[0192] In some embodiments, the client and server can communicate using any currently known or future developed network protocol, such as HTTP (HyperText Transfer Protocol), and can be interconnected with any form or medium of digital data communication (e.g., a communication network). Examples of communication networks include a local area network ("LAN"), a wide area network ("WAN"), an internet (e.g., the Internet), and a peer-to-peer network (e.g., an ad hoc peer-to-peer network), as well as any currently known or future developed network.
[0193] The computer-readable medium may be included in the electronic device, or may exist independently without being incorporated into the electronic device.
[0194] The above-mentioned computer-readable medium carries one or more programs. When the above-mentioned one or more programs are executed by the electronic device, the electronic device is enabled to: obtain a video copy based on the initial video editing data; determine a target text effect template based on the video copy, and the target text effect template is used to indicate the editing operation of at least one text effect; generate text material corresponding to the video copy by applying the target text effect template to the target keyword in the video copy; wherein, the text material is used to present the target keyword as the at least one text effect; update the initial video editing data through the text material to obtain target video editing data; wherein, the target video editing data is used to present the text material.
[0195] In embodiments of the present disclosure, computer program code for performing the operations of the present disclosure may be written in one or more programming languages or a combination thereof, including but not limited to object-oriented programming languages such as Java, Smalltalk, C++, and conventional procedural programming languages such as "C" or similar programming languages. The program code may be executed entirely on the computer, partially on the computer, as a separate software package, partially on the computer and partially on a remote computer, or entirely on the remote computer or server. In cases involving a remote computer, the remote computer may be connected to the computer via any type of network, including a local area network (LAN) or a wide area network (WAN), or may be connected to an external computer (e.g., via the Internet using an Internet service provider).
[0196] The flowcharts and block diagrams in the accompanying drawings illustrate the possible implementation architecture, functions and operations of the systems, methods and computer program products according to various embodiments of the present disclosure. In this regard, each box in the flowchart or block diagram can represent a module, program segment, or a part of code, and the module, program segment, or a part of code contains one or more executable instructions for realizing the specified logical function. It should also be noted that in some alternative implementations, the functions marked in the box can also occur in a different order than that marked in the accompanying drawings. For example, two boxes represented in succession can actually be executed substantially in parallel, and they can sometimes be executed in the opposite order, depending on the functions involved. It should also be noted that each box in the block diagram and / or flowchart, and the combination of the boxes in the block diagram and / or flowchart, can be implemented with a dedicated hardware-based system that performs the specified function or operation, or can be implemented with a combination of dedicated hardware and computer instructions.
[0197] The units involved in the embodiments described in this disclosure may be implemented in software or hardware, wherein the name of a unit does not necessarily limit the unit itself.
[0198] The functions described above herein may be performed, at least in part, by one or more hardware logic components. For example, and without limitation, exemplary types of hardware logic components that may be used include: field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), systems on chip (SOCs), complex programmable logic devices (CPLDs), and the like.
[0199] In the context of the present disclosure, a computer-readable medium can be a tangible medium that can contain or store a program for use by an instruction execution system, device or equipment or used in combination with an instruction execution system, device or equipment. A computer-readable medium can be a computer-readable signal medium or a computer-readable storage medium. A computer-readable medium can include, but is not limited to, an electronic, magnetic, optical, electromagnetic, infrared, or semiconductor system, device or equipment, or any suitable combination of the foregoing. A more specific example of a computer-readable storage medium can include an electrical connection based on one or more lines, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the foregoing.
[0200] The above description is merely a preferred embodiment of the present disclosure and an illustration of the technical principles employed. Those skilled in the art should understand that the scope of disclosure involved in the present disclosure is not limited to the technical solutions formed by the specific combination of the above-mentioned technical features, but also includes other technical solutions formed by any combination of the above-mentioned technical features or their equivalents without departing from the above-mentioned disclosed concepts. For example, a technical solution formed by replacing the above-mentioned features with (but not limited to) technical features with similar functions disclosed in this disclosure.
[0201] In addition, although each operation is described in a specific order, this should not be understood as requiring these operations to be performed in the specific order shown or in a sequential order. Under certain circumstances, multitasking and parallel processing may be advantageous. Similarly, although some specific implementation details have been included in the above discussion, these should not be interpreted as limiting the scope of the present disclosure. Some features described in the context of a separate embodiment can also be implemented in a single embodiment in combination. On the contrary, the various features described in the context of a single embodiment can also be implemented in multiple embodiments individually or in any suitable sub-combination mode.
[0202] Although the subject matter has been described in language specific to structural features and / or methodological logical acts, it should be understood that the subject matter defined in the appended claims is not necessarily limited to the specific features or acts described above. Rather, the specific features and acts described above are merely example forms of implementing the claims.
Claims
1. A video editing method, comprising: Obtaining a video copy based on initial video editing data; Determining a target text effect template according to the video copy, where the target text effect template is used to indicate editing operations of at least one text effect; Generating text material corresponding to the video copy by applying the target text effect template to target keywords in the video copy; wherein the text material is used to present the target keywords as the at least one text effect; and Updating the initial video editing data with the text material to obtain target video editing data; wherein the target video editing data is used to present the text material.
2. The method according to claim 1, wherein determining the target text effect template according to the video copy comprises: Identifying the target keywords in the video copy; Determining a category label of the category to which the target keywords belong; Determining a text effect editing type matching the category label; And Determining the target text effect template from multiple text effect templates included in the text effect editing type.
3. The method according to claim 2, wherein determining the target text effect template from multiple text effect templates included in the text effect editing type comprises: When the text effect editing type is the first type, determining the target text effect template from multiple text effect templates included in the first type; When the text effect editing type is the second type, determining a category identifier of the category to which the video copy belongs; and Determining the target text effect template matching the category identifier from multiple text effect templates included in the second type.
4. The method according to claim 3, wherein after determining the category identifier of the category to which the video copy belongs when the text effect editing type is the second type, and the method further comprises: When there is no text effect template matching the category identifier among multiple text effect templates included in the second type, determining the first text effect template among multiple text effect templates included in the second type as the target text effect template.
5. The method according to claim 2, wherein after determining the category label of the category to which the target keywords belong, and the method further comprises: Determining that there is a sound effect template matching the category label; And Determining the sound effect template as the target text effect template.
6. The method according to any one of claims 1 to 5, wherein obtaining the video copy based on the initial video editing data comprises: Displaying multiple copy effect modes, the multiple copy effect modes including a first copy effect mode, and the first copy effect mode is used to indicate generating subtitles for keywords of the video copy; And In response to a selection operation on the first copy effect mode, obtaining the video copy according to the initial video editing data.
7. The method according to claim 6, wherein the plurality of copywriting effect modes further includes a second copywriting effect mode, the second copywriting effect mode is used to indicate the generation of full subtitles, and the display form of the keywords in the full subtitles is different from the display form of the text other than the keywords; After displaying the multiple copy effect modes, and the method further comprises: In response to a selection operation on the second copywriting effect mode, obtain a video copywriting according to the initial video editing data; Generate video subtitles from the video copywriting, where keywords in the video subtitles are displayed in a first display form, and text other than keywords in the video subtitles is displayed in a second display form, and the first display form is different from the second display form; and Synthesize the video subtitles and the initial video editing data into first video editing data.
8. A video editing device, comprising: An acquisition module for obtaining a video copywriting according to initial video editing data; A determination module for determining a target text effect template according to the video copywriting, where the target text effect template is used to indicate editing operations for at least one text effect; A generation module for generating text materials corresponding to the video copywriting by applying the target text effect template to target keywords in the video copywriting; where the text materials are used to present the target keywords as the at least one text effect; and An update module for updating the initial video editing data through the text materials to obtain target video editing data; where the target video editing data is used to present the text materials.
9. An electronic device, comprising: A memory and a processor, where the memory is used to store a computer program; The processor is used to execute the video editing method according to any one of claims 1 to 7 when calling the computer program.
10. A computer-readable storage medium, on which a computer program is stored, and the computer program, when executed by a processor, implements the video editing method according to any one of claims 1 to 7.
11. A computer program product, comprising computer-executable instructions, where the computer-executable instructions, when executed by a processor, implement the video editing method according to any one of claims 1-7.
Citation Information
Patent Citations
Method and device for displaying presentation information
CN107786887A
Subtitle generation method and device and electronic equipment
CN110798636A
Subtitle file processing method and device
CN113068077A
Video editing method and device, computer equipment and storage medium
CN114449310A
Copywriting generation method and apparatus, model training method and apparatus, and device and storage medium
WO2023221934A1