Video editing method and apparatus, and electronic device and storage medium
By automatically generating subtitles using text templates in video editing methods, the problem of low efficiency in manually adding text in existing technologies is solved, thus achieving efficient subtitle production.
Patent Information
- Application Number
- PCT/CN2025/111501
- Authority / Receiving Office
- WO · WO
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2024-07-31
- Filing Date
- 2025-07-30
- Publication Date
- 2026-02-05
AI Technical Summary
Current methods for adding text to videos require manual positioning and adjustment, resulting in a large workload and low efficiency.
This paper provides a video editing method that displays an editing interface for text templates, obtains initial text and text editing information, generates text templates, and uses these templates to replace the initial text to form text fragments in the video editing draft, thereby achieving automated subtitle production.
It improves the efficiency of subtitle production, reduces the workload of subtitle production, and simplifies the process of adding text to videos.
Smart Images

Figure CN2025111501_05022026_PF_FP_ABST
Abstract
Description
Video editing methods and apparatus, electronic devices and storage media
[0001] This application claims priority to Chinese Patent Application No. 202411045028.7, filed on July 31, 2024, the disclosure of which is incorporated herein by reference in its entirety. Technical Field
[0002] Embodiments of this disclosure relate to a video editing method and apparatus, an electronic device, and a storage medium. Background Technology
[0003] Video content creation is ubiquitous in daily life; it's used to record life, express individuality, and convey value. With the development of internet technology, video creators can add text to their videos (e.g., as subtitles), conveying information to viewers, such as displaying today's topic. Summary of the Invention
[0004] Currently, adding text typically requires manual operation, which is not only labor-intensive but also inefficient. To address these issues, at least one embodiment of this disclosure provides a video editing method, apparatus, electronic device, and storage medium that can improve the efficiency of creating text in videos and reduce the workload.
[0005] At least one embodiment of this disclosure provides a video editing method, comprising: displaying an editing interface for a text template, wherein the editing interface includes an editing area and a preview area, the editing area being used to edit the text template, and the preview area being used to preview the editing effect of the text template; in response to an editing operation on the editing interface, acquiring initial text and text editing information, the text editing information being used to indicate a text effect applied to the initial text, the text effect including text visual effects and / or text sound effects; generating the text template based on the initial text and the text editing information, the text template being used to replace the initial text with target text to form a text fragment in a video editing draft, the text fragment being generated based on the target text and the text editing information, and the target text being presented with the text effect in the editing result corresponding to the video editing draft.
[0006] At least one embodiment of this disclosure also provides a video editing apparatus, comprising: a display unit configured to display an editing interface for a text template, the editing interface including an editing area and a preview area, the editing area being used to edit the text template, and the preview area being used to preview the editing effect of the text template; an acquisition unit configured to acquire initial text and text editing information in response to an editing operation on the editing interface, the text editing information being used to indicate a text effect applied to the initial text, the text effect including a text visual effect and / or a text sound effect; and a generation unit configured to generate the text template based on the initial text and the text editing information, the text template being used to replace the initial text with target text to form a text fragment in a video editing draft, the text fragment being generated based on the target text and the text editing information, and the target text being presented with the text effect in the editing result corresponding to the video editing draft.
[0007] At least one embodiment of this disclosure also provides an electronic device, including: a processor; a memory including one or more computer program instructions; wherein the one or more computer program instructions are stored in the memory and, when executed by the processor, implement the video editing method described in any embodiment of this disclosure.
[0008] At least one embodiment of this disclosure also provides a computer-readable storage medium that non-temporarily stores computer-readable instructions, wherein the video editing method described in any embodiment of this disclosure is implemented when the computer-readable instructions are executed by a processor. Attached Figure Description
[0009] The above and other features, advantages, and aspects of the embodiments of this disclosure will become more apparent from the accompanying drawings and the following detailed description. Throughout the drawings, the same reference numerals denote the same elements. It should be understood that the drawings are schematic, and the originals and elements are not necessarily drawn to scale.
[0010] Figure 1A shows a flowchart illustrating a video editing method provided in at least one embodiment of this disclosure;
[0011] Figure 1B shows a schematic diagram of the user interface 100 of an application for performing a video editing method according to at least one embodiment of the present disclosure;
[0012] Figure 2A shows a schematic diagram of an editing interface provided by at least one embodiment of the present disclosure;
[0013] Figure 2B illustrates a schematic diagram of obtaining initial text according to at least one embodiment of this disclosure;
[0014] Figure 2C shows a schematic diagram of a word segmentation and typesetting loading pop-up window provided in at least one embodiment of this disclosure;
[0015] Figure 2D shows a schematic diagram of word segmentation and typesetting provided by at least one embodiment of the present disclosure;
[0016] Figure 3A illustrates a schematic diagram of an initial text addition animation provided by at least one embodiment of the present disclosure;
[0017] Figure 3B shows a schematic diagram of a page after adding progressive animations according to at least one embodiment of the present disclosure;
[0018] Figure 3C shows a schematic diagram of a page after setting fill attributes according to at least one embodiment of the present disclosure;
[0019] Figure 3D shows a schematic diagram of a page after adding word-by-word animation according to at least one embodiment of the present disclosure;
[0020] Figure 3E illustrates a schematic diagram of adding a scaling attribute according to at least one embodiment of the present disclosure;
[0021] Figure 3F shows a schematic diagram of adding keyframes according to at least one embodiment of the present disclosure;
[0022] Figures 4A to 4C illustrate a schematic diagram of setting a keyframe curve according to at least one embodiment of the present disclosure;
[0023] Figure 5 is a schematic block diagram of a video editing device provided in some embodiments of this disclosure;
[0024] Figure 6 is a schematic block diagram of an electronic device provided in some embodiments of this disclosure;
[0025] Figure 7 is a schematic block diagram of another electronic device provided in some embodiments of this disclosure; and
[0026] Figure 8 is a schematic diagram of a storage medium provided in some embodiments of this disclosure. Detailed Implementation
[0027] Embodiments of this disclosure will now be described in more detail with reference to the accompanying drawings. While some embodiments of this disclosure are shown in the drawings, it should be understood that this disclosure can be implemented in various forms and should not be construed as limited to the embodiments set forth herein. Rather, these embodiments are provided to provide a more thorough and complete understanding of this disclosure. It should be understood that the accompanying drawings and embodiments of this disclosure are for illustrative purposes only and are not intended to limit the scope of protection of this disclosure.
[0028] It should be understood that the steps described in the method embodiments of this disclosure may be performed in different orders and / or in parallel. Furthermore, the method embodiments may include additional steps and / or omit the steps shown. The scope of this disclosure is not limited in this respect.
[0029] The term "comprising" and its variations as used herein are open-ended inclusions, meaning "including but not limited to". The term "based on" means "at least partially based on". The term "one embodiment" means "at least one embodiment"; the term "another embodiment" means "at least one additional embodiment"; the term "some embodiments" means "at least some embodiments". Definitions of other terms will be given in the description below.
[0030] It should be noted that the concepts of "first" and "second" mentioned in this disclosure are used only to distinguish different devices, modules or units, and are not used to limit the order of functions performed by these devices, modules or units or their interdependencies.
[0031] It should be noted that the terms "one" and "more" used in this disclosure are illustrative rather than restrictive, and those skilled in the art should understand that, unless explicitly stated otherwise in the context, they should be understood as "one or more". "More" should be understood as two or more.
[0032] The names of messages or information exchanged between multiple devices in the embodiments of this disclosure are for illustrative purposes only and are not intended to limit the scope of such messages or information.
[0033] Currently, adding text to videos typically involves manually locating the video frames where the text needs to be added, manually entering the text in each frame, and adjusting the text style in each frame, such as adjusting the text size and color. Therefore, the current method of adding text to videos requires manual operation, which is not only labor-intensive but also inefficient.
[0034] At least one embodiment of this disclosure provides a video editing method that can generate text templates, so that the video producer only needs to call the text template to generate the text to be used in the video editing draft (e.g., as subtitles in the video editing draft), thereby improving the efficiency of subtitle production and reducing the workload of subtitle production.
[0035] Figure 1A shows a flowchart illustrating a video editing method provided in at least one embodiment of this disclosure.
[0036] As shown in Figure 1A, in at least one embodiment, the method includes the following operations.
[0037] Step S10: Display the text template editing interface, which includes an editing area and a preview area. The editing area is used to edit the text template, and the preview area is used to preview the editing effect of the text template.
[0038] Step S20: In response to the editing operation on the editing interface, obtain the initial text and text editing information. The text editing information is used to indicate the text effects applied to the initial text. The text effects include text visual effects and / or text sound effects.
[0039] Step S30: Generate a text template based on the initial text and text editing information. The text template is used to replace the initial text with the target text to form text fragments in the video editing draft. The text fragments are generated based on the target text and text editing information. In the editing results corresponding to the video editing draft, the target text is presented as text effects.
[0040] The embodiments of this disclosure can be applied to a client, such as an application that performs the video editing methods provided in the embodiments of this disclosure.
[0041] The following example of step S10 is illustrated with reference to Figure 1B.
[0042] Figure 1B shows a schematic diagram of an application interface 100 for performing a video editing method according to at least one embodiment of the present disclosure.
[0043] As shown in Figure 1B, the user interface 100 includes a "New Subtitle Template" option 101 and a "New Text Template" option 102. The user interface 100 can, for example, serve as the application's homepage.
[0044] For example, selecting the "Create New Subtitle Template" option 101 in the operation interface 100 displays the text template editing interface. Text templates created through the "Create New Subtitle Template" option 101 are automatically used as subtitles in the video editing draft.
[0045] For example, selecting the "New Text Template" option 102 in the operation interface 100 displays the text template editing interface. Text templates created through the "New Text Template" option 102 are automatically used as text templates in the video editing draft, excluding subtitles, such as the narration text template in the video editing draft.
[0046] As described above, in some embodiments of this disclosure, the text template can refer to a subtitle template for a video editing draft, or a narration template for a video editing draft. In some embodiments of this disclosure, different types of text templates in the video editing draft can be generated through different entry points.
[0047] In some embodiments of this disclosure, user permissions can be set, and one or more of the "Create New Caption Template" option 101 and "Create New Text Template" option 102 can be displayed according to the user's permissions. The user, for example, refers to the creator of the text template.
[0048] As shown in Figure 1B, in addition to the "New Subtitle Template" option 101 and the "New Text Template" option 102, the operation interface 100 can also include a draft, in which a draft previously created using the application can be displayed.
[0049] The user interface 100 also includes displaying some basic information, such as username, homepage, material tools, cloud space, etc.
[0050] Figure 2A shows a schematic diagram of an editing interface provided by at least one embodiment of the present disclosure. This editing interface 200 is, for example, a page accessed by a user clicking the "New Subtitle Template" option 101. The editing interface includes an editing area 20 and a preview area 202. The editing area 20 is used to edit the text template, and the preview area is used to preview the editing effect of the text template. The editing area 20 includes, for example, a track area 20-1 and an editing area 20-2.
[0051] For example, editing area 20-2 in editing area 20 includes style controls, and the "Style" button 221 is an example of a style control. In this editing interface 200, selecting the "Style" button 221 will display the style panel 201 for setting the style of the text.
[0052] For step S20, for example, initial text and text editing information are obtained by performing editing operations in the style panel 201. For example, the editing operations performed on the style panel 201 may include entering text in the text input sub-control to obtain initial text and performing layout operations on the layout sub-control to obtain text editing information.
[0053] For example, the style panel 201 includes a text input box 211 (an example of a text input sub-control). For example, a user can enter initial text in the text input box 211.
[0054] For example, the style panel 201 includes controls for word count per line, line count per page, font, and font size. Text editing information is obtained by configuring these controls. This text editing information indicates the text effects applied to the initial text, including visual and / or audio effects. For instance, setting the font to KaiTi and the color to red will display the text in the subtitle as red KaiTi. In some embodiments of this disclosure, the text editing information may also include configuration of text audio (e.g., background sound, timbre of the read-out text), which can be applied to the initial text to obtain audio effects.
[0055] In step S30, the text template can be published on the platform for users to choose from. For example, after a user selects the text template, they can replace the initial text with the target text to create text snippets in the video editing draft. The text editing information in the text template is applied to the target text, so that the text effect of the target text is similar to that of the initial text.
[0056] For example, if a user double-clicks in the preview area 202 (attempting to enter characters), a first prompt message will pop up. This first prompt message is used to remind the user to enter the target text to replace the initial text in the text input box 201. The first prompt message is, for example, "Please enter the target text in the text input box". At the same time, the text input box 211 can be highlighted to serve as a prompt.
[0057] In some embodiments of this disclosure, manual line breaks are restricted in the text input box 211. If a manual line break operation is received in the text input box, a second prompt message pops up to remind the user to perform the line break operation in the text layout. This second prompt message may be, for example, "Please perform the line break operation in the text layout." Please refer to the description below for information on text layout.
[0058] Figure 2B illustrates a schematic diagram of obtaining initial text according to at least one embodiment of the present disclosure.
[0059] For example, Figure 2B is the page after initial text is entered into text input box 211 in Figure 2A. For instance, if "Flexible editing,magical AI tools,team collaboration,and stock assets.Make video creation like never" is entered into text input box 211 as the initial text, the characters displayed in preview area 202 will also be modified accordingly.
[0060] In some embodiments of this disclosure, the editing area 20 presents an editing track for a text template. For example, as shown in Figures 2A and 2B, the editing area 20 includes a track region 20-1 for presenting the editing track of the text template. The editing track includes at least a text track for placing at least one text track segment formed based on the initial text.
[0061] For example, video editing methods also include displaying the initial text in the editing track while the text input sub-control receives the initial text.
[0062] For example, the initial text is automatically displayed in the text track of the subtitle track area 203. Since the text has not been formatted at this time, the initial text is displayed as a single line in the preview area 202. Due to the width limitation of the preview area 202, only a portion of the characters in the initial text is displayed in the preview area 202, such as "collaboration, and stock assets".
[0063] As shown in Figure 2B, when the text content in the text input box 211 changes, the "Update Subtitle" button 2311 can be highlighted.
[0064] In some embodiments of this disclosure, in the text template, the initial text is displayed in an initial video, the initial video includes at least one frame, the initial text includes multiple characters, and the characters located in the same frame of the initial video are used as a text track segment. The time position of the at least one text track segment in the editing track corresponds to the time position of at least one frame of the initial video in the text template.
[0065] For example, the initial video might be a video composed of consecutive image frames, each image frame representing a single frame. Multiple characters are broken down into at least one text track segment, with each text track segment corresponding one-to-one with at least one frame. The temporal position of the at least one text track segment within the text track corresponds to the temporal position of its corresponding frame within the initial video. That is, the temporal position of a character displayed within the edit track corresponds to the temporal position of the frame containing that character within the initial video.
[0066] In some embodiments of this disclosure, adjacent text track segments are separated by a separator mark, which is located closer to the later text track segment in time. For a description of the text tracks, please refer to Figure 2D below.
[0067] In some embodiments of this disclosure, a single subtitle is added by default, and the single subtitle is automatically displayed on the subtitle track, and the number of text segments on the subtitle track is limited. When there is more than one text segment on the subtitle track, a pop-up window prompts "The subtitle template needs to be created from a single text segment"; when the number of text segments on the subtitle track is 0, a pop-up window prompts "The subtitle template needs to be created from a single text segment".
[0068] In some embodiments of this disclosure, for example, the default duration of the edit track is K seconds, meaning the default playback duration of the text is K seconds, where K is greater than 0. In some embodiments of this disclosure, the edit track supports operations such as dragging, shortening, and extending within the track duration.
[0069] In some embodiments of this disclosure, the subtitle template project page 200 includes a resource window 204. The resource window 204, for example, is positioned by default to "Stickers". Only stickers are displayed in the top bar, and the "Stickers" icon is automatically selected. Selecting the "Stickers" icon indicates the display of the sticker library. The icons in this resource window are categorized, for example, into local stickers, basic graphics, and the sticker library. Local stickers are, for example, stickers downloaded to the user's local device; basic graphics are, for example, stickers included with the subtitle template application; and the sticker library may include, for example, stickers saved by the user, popular stickers, etc. The sticker library may include stickers available online.
[0070] In some embodiments of this disclosure, stickers can be animated, including, for example, sticker entrance animations, exit animations, and looping animations. Stickers can be reused once within a time frame.
[0071] In some embodiments of this disclosure, the style panel 201 includes, for example, a subtitle layout area 231, which is used to acquire layout information of the initial text. The layout information is used to layout the initial text. The sub-controls in the subtitle layout area 231 are all layout sub-controls. As shown in Figure 2B, the subtitle layout area 231 includes sub-controls such as word count per line, number of lines per page, font, font size, style, character spacing, line spacing, and alignment. These sub-controls are examples of layout sub-controls. It should be noted that although the subtitle layout area in Figure 2B lists the above-mentioned layout types, this is not limiting to this disclosure. The subtitle layout area may include fewer or more layout types than those mentioned above, and those skilled in the art can set them according to their needs.
[0072] When any of the text content, word count per line, or line count per page in the text input box 211 changes, the "Update Caption" button 2311 can be highlighted and selected; when there is no change, the "Update Caption" button 2311 is grayed out and cannot be selected. Additionally, when the text input box 211 is empty, the "Update Caption" button 2311 can be highlighted.
[0073] In some embodiments of this disclosure, generating a text template based on initial text and text editing information includes: determining the number of characters in each line of text in the text template and the number of lines of text in each frame of the video editing draft based on layout information; and generating a text template including the initial text based on the number of characters in each line of text and the number of lines of text in each frame.
[0074] For example, Figure 2B includes sub-controls for "Words per Line" and "Lines per Page," which are used to set layout information. "Words per Line" defines how many words are in each line of the subtitle. Options for "Words per Line" include unlimited, intelligent line splitting, and word count, with unlimited being the default. Unlimited means there is no limit to the length of a single line of text. Word count means, for example, one word per line for English subtitles and one phrase per line for Chinese subtitles. Intelligent line splitting refers to recognizing the text content of the subtitle and intelligently splitting it into multiple lines of text.
[0075] In some embodiments of this disclosure, determining the number of characters in each line of text in the text template based on the layout information includes: in response to the layout information including intelligent splitting of the initial text, splitting the initial text into multiple lines of text.
[0076] For example, in response to the selection of intelligent line splitting in the layout sub-operation of the first child control, the initial text is split into multiple sub-contents, each of which is a line of text. Recognizing and intelligently splitting text content simplifies user operations and provides users with better line breaks, making the subtitle layout more aesthetically pleasing and improving the user experience.
[0077] If the smart line splitting option is selected, the loading process will be blocked. Users can cancel the smart line splitting process at any time. If the user chooses to cancel smart line splitting, the selection of smart line splitting will not take effect.
[0078] The number of lines per page defines how many lines of subtitles are displayed per page. For example, the selection range for the number of lines per page is 1 to 8 lines. This selection range can be modified in the backend configuration; that is, this disclosure does not limit the selection range for the number of lines per page. When the number of words per line is set to unlimited, the number of lines per page is unselectable, and the number of lines per page can only be 1. When the text content is divided into lines according to the number of words per line, if the number of lines is less than 'a', then 'a' lines minus the maximum number of lines (e.g., 8 lines) needs to be set to unselectable, where 'a' is less than the maximum number of lines and greater than 1.
[0079] For example, after the user sets the number of words per line and the number of lines per page, clicking the "Update Subtitles" button 2311 will display a word segmentation and layout loading pop-up window, which will generate the initial text layout based on the number of characters in each line of subtitles and the number of subtitle lines in each screen.
[0080] In some embodiments of this disclosure, the style control further includes a style sub-control area 241, where all sub-controls within the style sub-control area 241 are style sub-controls. For example, styles include appearance styles, which include styles such as text, fill, stroke, inner shadow, and outer shadow.
[0081] The text template generation process, based on input operations in the editing interface, also includes: determining the style of the subtitle content in response to configuration operations on the style sub-control, thereby generating a subtitle template based on the layout, subtitle content, and style.
[0082] Figure 2C shows a schematic diagram of a word segmentation and typesetting loading pop-up provided in at least one embodiment of the present disclosure.
[0083] As shown in Figure 2C, the word segmentation and typesetting loading pop-up displays "Word segmentation and typesetting in progress..." to inform the user that word segmentation and typesetting is currently in progress. The pop-up also includes a "Cancel" button, which the user can click to cancel the current operation and return to the original state. During the word segmentation and typesetting process, apart from the "Word segmentation and typesetting in progress..." pop-up, all other parts are grayed out to highlight "Word segmentation and typesetting in progress...".
[0084] Figure 2D shows a schematic diagram of a word segmentation and typesetting provided by at least one embodiment of the present disclosure.
[0085] As shown in Figure 2D, if the word count per line is set to intelligent line splitting and the number of lines per page is 2, the subtitle layout effect after line splitting can be previewed in the preview area. As shown in Figure 2D, after word splitting and layout, "Flexible editing" and "magical AI tools" are displayed in two separate lines in the player.
[0086] In some embodiments of this disclosure, each line of subtitles is automatically aligned. The original character size is maintained within the canvas; if the character exceeds the canvas size, it overflows, and the user can manually adjust it. The canvas, for example, refers to the display area of the player.
[0087] In some embodiments of this disclosure, the style of the text track is adjusted according to the layout after the layout is generated.
[0088] For example, adjusting the style of the text track according to the layout includes: based on the number of subtitle lines in each frame of the initial video as indicated by the layout, displaying the text corresponding to each of the at least one frame at its respective time position in the text track. In the text track, two adjacent texts are separated by a separator mark, which is positioned closer to the later text in the text track. This embodiment displays the subtitle layout in the text track, allowing users to see the subtitle layout promptly and conveniently.
[0089] As shown in Figure 2D, the style of the text track changes after successful intelligent line splitting. For example, a dividing line appears on the text track (e.g., the short dashed line in Figure 2D), which is used to separate the text between two adjacent frames. This dividing line is an example of a separator. For example, in the example of Figure 2B, two lines of text are displayed per page, so "team collaboration" and "magical AI tools" are located on two adjacent pages, and "team collaboration" and "magical AI tools" are displayed at their respective time positions in the two frames of the text track. In the text track, "team collaboration" and "magical AI tools" are separated by a dividing line to indicate that they belong to two different frames, and the dividing line is adjacent to "team collaboration".
[0090] As shown in Figure 2D, after successful intelligent line splitting, a loop icon 213 appears on the subtitle track. Users can click the loop icon 213 to cycle through the style attributes of the subtitles on that page. This loop icon 213 is also a keyboard shortcut, allowing users to directly apply the style attributes of the subtitles on that page to other pages. The subtitle style attributes include, for example, font, font size, style, character spacing, line spacing, and alignment, in the subtitle layout area 231, excluding the number of words per line and the number of lines per page.
[0091] As shown in Figure 2D, after the intelligent line splitting is successful, the "Update Subtitles" button 2311 is grayed out and cannot be selected.
[0092] As shown in Figure 2D, style attributes also include the appearance style of the subtitles. For example, the appearance style of the subtitles can be set in the style panel 201. For example, you can select a fancy text style. Fancy text styles include colorful fonts, various images and sound effects, etc., which are multi-layered texts for secondary creation of video content. Well-chosen fancy text can not only enhance the video scene, but also make seemingly mundane shots more interesting.
[0093] For example, the appearance style of subtitles can also include background, fill, stroke, inner shadow, and outer shadow. Users can set the color and opacity of the subtitles through the fill, and set the edge color, edge opacity, and edge thickness through the description. Inner or outer shadow settings include effects such as color, opacity, offset, shadow stroke, angle, and feathering. Background settings include effects such as color, opacity, rounded corners, height, width, vertical offset, and horizontal offset.
[0094] In some embodiments of this disclosure, the text visual effects include animation, the text editing information includes animation information, and the animation information includes animation mode and animation configuration information.
[0095] In some embodiments of this disclosure, the text template may include multiple animations, such as a first animation, a second animation, etc. Each animation may have different animation modes and configuration information.
[0096] In some embodiments of this disclosure, the editing interface 200 also includes animation controls.
[0097] Figure 3A illustrates a schematic diagram of an initial text animation provided by at least one embodiment of the present disclosure.
[0098] As shown in Figure 3A, after selecting the "Animation" option 301 in the editing interface 200, the animation mode operation panel 302 is displayed. The "Animation" option 301 is an example of an animation control. In some embodiments of this disclosure, the text editing information includes first animation information. Generating a text template based on the initial text and the text editing information includes: generating a first animation applied to the initial text based on the first animation information.
[0099] In some embodiments of this disclosure, the first animation information includes a first animation mode and first configuration information.
[0100] For example, this includes: in response to a first animation mode addition operation on the animation mode operation panel 302, obtaining the first animation mode selected by the first animation mode addition operation; the first animation mode indicates the scope of the first configuration information in the initial text. In some embodiments of this disclosure, the scope refers, for example, to the smallest unit applied to the initial text.
[0101] For example, animation modes include page-by-page, line-by-line, and word-by-word. Adding operations for the first animation mode include, for example, selecting the "Add Animation Mode" button 312 and selecting the animation model.
[0102] For example, in the animation mode operation panel 302, the "Add Animation Mode" button 312 can be selected. In response to clicking the "Add Animation Mode" button 312, a dynamic illustration card 322 pops up. For example, the dynamic illustration card 322 includes three dynamic illustrations, each representing one of three animation modes: page-by-page, line-by-line, and word-by-word.
[0103] The smallest unit of animation includes a single word, a single line of text, or a single frame.
[0104] Page-by-page refers to the smallest unit of animation being the subtitles on a single screen. For example, in the dynamic illustration of "page-by-page", "Good bye" and "guy" are two lines, located on the same page, and both lines are highlighted simultaneously.
[0105] Progressive animation means that the smallest unit of animation is a single line of text. For example, in the dynamic illustration of "progressive animation", "Good bye" and "guy" are two lines. The fact that "Good bye" and "guy" are highlighted in turn indicates that the animation mode is progressive.
[0106] Word-by-word means that the smallest unit of animation is a single word. For example, in the dynamic illustration of "word-by-word", "Good", "bye" and "guy" are three words. The fact that "Good", "bye" and "guy" are highlighted in sequence indicates that the animation mode is word-by-word.
[0107] For example, the first animation mode is one of line-by-line, word-by-word, or page-by-page.
[0108] For example, the configuration information includes at least one of the following: the color of the characters within the scope, the scaling ratio of the characters within the scope, stroke information, background information, glow information, and shadow information. For example, if the scope is line-by-line, the configuration information includes the color of each line and the scaling ratio of each line. If the scope is word-by-word, the configuration information includes the color of each word and the scaling ratio of each word.
[0109] For example, the color, scaling ratio, stroke information, background information, glow information, and shadow information of each word in the first animation are obtained, and the animation form of each word is generated based on this configuration information. In some embodiments of this disclosure, for example, the configuration information of each smallest unit in the first animation is obtained; and the first animation is generated based on the configuration information of each smallest unit.
[0110] In some embodiments of this disclosure, generating a text template based on initial text and text editing information further includes: in response to obtaining second animation information, determining whether the second animation mode in the second animation information is the same as the first animation mode; in response to the second animation mode being different from the first animation mode, generating a second animation applied to the initial text based on the second animation information.
[0111] For example, in response to a second animation mode addition operation on the animation mode operation panel 302, the second animation mode selected by the second animation mode addition operation is obtained; it is determined whether the second animation mode is the same as the first animation mode; in response to the second animation mode being different from the first animation mode, a second animation corresponding to the second animation mode addition operation is generated based on the second animation mode.
[0112] For example, if the second animation mode is page-by-page and the first animation mode is line-by-line, then the first animation mode and the second animation mode are different, and the second animation corresponding to page-by-page will be generated.
[0113] In some embodiments of this disclosure, multiple animation modes can be added. For example, after adding a step-by-step animation mode, a word-by-word animation mode can also be selected. In the subtitle track, the track for the later-added animation mode is listed above the track for the earlier-added animation mode.
[0114] In some embodiments of this disclosure, in response to the second animation mode being the same as the first animation mode, a prompt message is displayed to indicate that the first animation mode has been added.
[0115] In some embodiments of this disclosure, only one of the same animation mode is allowed to be added. If an animation mode that has already been added is added again, a message will appear indicating that the mode has already been added. For example, after adding the progressive animation mode, if progressive animation is selected again, a message will pop up saying "This mode has already been added".
[0116] In some embodiments of this disclosure, the editing track further includes an animation track for demonstrating the extent to which the animation is applied within the initial text.
[0117] Figure 3B shows a schematic diagram of a page with added animations according to at least one embodiment of the present disclosure.
[0118] As shown in Figure 3B, if "Line by Line" is selected in the dynamic illustration card 322 shown in Figure 3A, "Line by Line Animation" will be displayed in the animation mode operation panel 302, indicating that a line by line animation has been added. Furthermore, a new animation track corresponding to the line by line animation will be added to the editing track. In the animation track corresponding to the line by line animation, each line of text content will be displayed sequentially, and each line of text content will be separated by a dividing line.
[0119] In some embodiments of this disclosure, dragging and dropping to sort animation modes on the animation mode operation panel 302 and the editing track is supported, and the animation mode operation panel 302 and the editing track are linked to modify their states. For example, animation modes can be dragged and dropped on the editing track to adjust the order of multiple animation modes.
[0120] For example, if the editing track includes an animation track for each line of animation and an animation track for each word of animation, and the animation track for each word of animation is located above the animation track for each line of animation, then you can drag the animation track for each line of animation to move it below the animation track for each line of animation.
[0121] In some embodiments of this disclosure, it is supported to delete already added animation modes. For example, right-clicking in the area near "Step-by-Step Animation" will bring up a delete option. If the user selects to delete, a message will appear saying "After deletion, all animations in this mode will be deleted as well." After clicking "Confirm," the animation mode will be deleted.
[0122] In some embodiments of this disclosure, the editing track supports functions such as locking and hiding. For example, after locking the animation track corresponding to each animation, it cannot be operated on. If the animation track corresponding to each animation is hidden, it will be hidden and no longer displayed in the subtitle track area 203.
[0123] In some embodiments of this disclosure, the animation properties of the added animation mode can be set.
[0124] As shown in Figure 3B, the area near the animation mode includes animation attribute setting buttons. For example, the "+" button 332 in Figure 3B is an animation attribute setting button. Each animation mode corresponds to one animation attribute setting button. By selecting the animation attribute setting button, you can set the animation attributes of that animation mode.
[0125] For example, clicking the "+" button 332 displays the attribute selection card 342, where the user can select the animation attribute to set. For instance, selecting "Fill" in attribute selection card 342 allows setting the fill color of the text content in the "Step-by-Step Animation".
[0126] Figure 3C shows a schematic diagram of a page after setting the fill attribute, provided by at least one embodiment of the present disclosure.
[0127] As shown in Figure 3C, if the fill color of the text content in the "Step-by-Step Animation" is set to green after the operation described in Figure 3B, then the fill attribute color displayed on the page will also be "green". At this time, the user can drag the pointer in the editing track to preview the effect of this fill attribute. As shown in Figure 3C, when the user drags the pointer from 00:00 to a point before 00:02, the animation effect of the first line is displayed, with the first line's text "Flexible editing" displayed in green. If the user continues to drag the pointer forward, the second line's text "magical AI tools" will also be displayed in green. That is, the animation process can be seen by dragging the pointer.
[0128] As shown in Figure 3B, in addition to the fill attribute, the attribute selection card 342 can also include animation attributes such as position, stroke, background, glow, inner shadow, and outer shadow.
[0129] In some embodiments of this disclosure, the above-mentioned animation attributes may have default values. If a user selects an animation attribute but does not modify its setting value, the default value is retained. For example, the default color for the fill attribute is white, and the default opacity is 100%. In the position attribute, the default values for the X-axis coordinate, Y-axis coordinate, rotation angle, and position scaling are 0% and 100%, respectively. The default color for the stroke attribute is black, the default thickness is 50, and the default opacity is 0%. The default color for the background attribute is gray, the default opacity is 100%, the default rounded corners are 0%, the default height is 50%, the default width is 50%, the default vertical offset is 50%, and the default horizontal offset is 50%. The default color for the shadow attribute is black, the default opacity is 80%, the default offset is 5, the default angle is -29%, and the default feathering is 1%.
[0130] In some embodiments of this disclosure, if a user has set animation attributes, the user can restore the set attribute values to their default values by clicking the restore button.
[0131] In some embodiments of this disclosure, the attribute can be deleted by right-clicking and then clicking the delete button.
[0132] In some embodiments of this disclosure, if no animation attribute is added, the style attribute is retained; if an animation attribute is added, the animation attribute is retained, and the effect in the preview area is the superposition of the same attribute.
[0133] As shown in Figure 3C, after setting the animation properties for the word-by-word animation, you can click "Add Animation Mode" again to add another animation mode, such as word-by-word animation.
[0134] Figure 3D shows a schematic diagram of a page after adding word-by-word animation, provided by at least one embodiment of this disclosure.
[0135] As shown in Figure 3D, after adding word-by-word animation based on Figure 3C, the animation mode operation panel displays two animation modes: "Word-by-Word Animation" and "Action-by-Action Animation," with each mode displaying its animation properties. For example, if the fill property of the word-by-word animation is red, the first word "Flexible" will be displayed in red in the preview area. Correspondingly, an animation track for the word-by-word animation is also added to the subtitle track area 203.
[0136] As shown in Figure 3D, users can click the "+" button 332 again to add animation attributes.
[0137] Figure 3E shows a schematic diagram of adding a scaling attribute according to at least one embodiment of the present disclosure.
[0138] As shown in Figure 3E, the scaling attribute is a sub-attribute of the position attribute. Besides scaling, the position attribute can also include X-axis, Y-axis, and rotation attributes. The X-axis and Y-axis attributes define the text's position, while the rotation attribute defines the text's rotation angle. For example, using the width of a frame as the X-axis and the length of a frame as the Y-axis, the X-axis attribute might refer to the coordinates of the subtitle along the X-axis, and the Y-axis attribute might refer to the coordinates of the subtitle along the Y-axis.
[0139] The scaling property allows you to adjust the size of text, specifying how many times it is scaled from the default size. For example, if the scaling property is set to 100%, the default size is maintained; if the scaling property is set to 200%, the text size is twice the default size.
[0140] As shown in Figure 3E, a "Add Keyframe" button 342 can be included in the area near each attribute to add a keyframe for that attribute. For example, if the pointer on the subtitle track is at 00:00, clicking the "Add Keyframe" button for the scaling attribute will switch the button from an unselected state (e.g., a white diamond) to a selected state (e.g., a black diamond), and the scaling ratio will be 100%, indicating that a keyframe with a scaling ratio of 100% has been added at 00:00. In embodiments of this disclosure, a keyframe refers to the start and end frames of the animation, and the animation operates over the time period between the start and end frames. In the above embodiment, the frame at 00:00 is used as the start frame of the animation, and the size of the subtitle text content in the frame at 00:00 is the same as the default size.
[0141] Figure 3F shows a schematic diagram of adding a keyframe according to at least one embodiment of the present disclosure.
[0142] Based on the example described in Figure 3E, after using the frame at 00:00 as the starting frame of the animation, as shown in Figure 3F, the user can drag the clock hand to time m and set the scaling value of that time to 200%. The frame at time m then becomes the ending frame of the animation. In this ending frame, the size of the subtitle text content is twice the default size. That is, the animation from 00:00 to time m includes a change in the size of the subtitle text content from the default size to twice the default size.
[0143] As shown in Figure 3F, in the player's preview page, at time m, "Flexible" is enlarged to twice its default size and displayed in red according to the word-by-word animation fill property, while "editing" is displayed in green according to the animation fill property.
[0144] In some embodiments of this disclosure, the method further includes: adjusting a first animation in response to a track input operation performed in the subtitle track region. This embodiment enables convenient and targeted adjustment of the subtitle animation via a track input operation.
[0145] For example, track input operations include setting keyframe curves. A keyframe, for example, refers to a frame where subtitle attributes are set.
[0146] For example, in response to a selection operation on the animation track, the animation's speed curve is displayed; in response to an adjustment operation on the speed curve, the animation's playback speed is adjusted. This embodiment implements the adjustment operation of the animation speed, and only requires operation in the subtitle track area, making the operation simple, convenient, and targeted.
[0147] Figures 4A to 4C show schematic diagrams of setting a keyframe curve according to at least one embodiment of the present disclosure.
[0148] As shown in Figure 4A, when a keyframe at time m is selected, for example, by right-clicking the mouse, operation card 401 for that keyframe is displayed.
[0149] Operation card 401 includes a "Show Keyframe Animation" button. Selecting this button displays the state shown in Figure 4B in the subtitle track area. As shown in Figure 4B, the scaling attribute in the position properties is displayed in the subtitle track area. For example, the user can click the drop-down button 411 for this scaling attribute, and in response to the click of the drop-down button 411, an animation curve 421 from the start frame (00:00) to the end frame (m) is displayed (as shown in Figure 4C). For example, in the example of Figure 4C, the animation from the start frame (00:00) to the end frame (m) is a uniform change in the size of the text content from 100% to 200%.
[0150] For example, users can set the animation's curve by selecting the curve shortcut icon 431. For instance, if the curve shortcut icon 431 represents accelerated change, then the text content's size will change from the default size of 100% to 200% with accelerated change.
[0151] For example, users can add an intermediate keyframe between the start frame (00:00) and the end frame (m), and set the size of the text content in the intermediate keyframe to 50%. Then, from the start frame to the intermediate keyframe, the size of the text content changes from the default size of 100% to 50%, and from the intermediate keyframe to the end frame, the size of the text content changes from the default size of 50% to 200%.
[0152] In this embodiment, the user can adjust the animation curve via the subtitle track, such as adjusting the animation speed, for user convenience.
[0153] In some embodiments of this disclosure, the track input operation may also include setting the order of animation attributes, such as moving the stroke to the next layer and placing the glow at the bottom layer.
[0154] In some embodiments of this disclosure, after receiving input in the editing interface, the user can right-click to bring up an operation box, click "Apply to All and Preview" in the operation box, and a preview pop-up window will appear. The generated text template, including all text with the current animation effects applied, can be previewed in the preview pop-up window, and will automatically loop after the preview ends.
[0155] If the user is satisfied with the generated text template after previewing, they can click the "Publish" button in the text editing interface to publish the template. After clicking "Publish," the user will be guided to fill in information such as name, tags, originality, and authorization. If publication is successful, a success message will be displayed, and the review progress can be viewed on the user's profile page. If publication fails, a failure message will pop up, prompting the user to check their network connection and try again. Once the text template is published on the platform, it becomes available for users to select. Users can replace the initial text in the template with the target text, or replace the initial video in the template with their own video, thus obtaining a video editing draft.
[0156] In some embodiments of this disclosure, the editing interface may further include a material tool, in which a text template draft button is added. Unfinished text templates by the user can be stored as drafts, and when the user clicks the text template draft button, the draft is retrieved and displayed.
[0157] Figure 5 is a schematic block diagram of a video editing device provided in some embodiments of this disclosure. As shown in Figure 5, the video editing device 500 includes a display unit 110, an acquisition unit 120, and a generation unit 130. For example, the video editing device 500 can be applied to a user terminal, or to any device or system that requires previewing design materials; the embodiments of this disclosure do not limit this application.
[0158] The display unit 110 is configured to display an editing interface for a text template. The editing interface includes an editing area and a preview area. The editing area is used to edit the text template, and the preview area is used to preview the editing effect of the text template. For example, the display unit 110 can execute step S10 of the video editing method shown in Figure 1A.
[0159] The acquisition unit 120 is configured to acquire initial text and text editing information in response to an editing operation on the editing interface. The text editing information is used to indicate the text effects applied to the initial text. The text effects include text visual effects and / or text sound effects. For example, the acquisition unit 120 can execute step S20 of the video editing method shown in FIG1A.
[0160] The generation unit 130 is configured to generate a text template based on the initial text and text editing information. The text template is used to replace the initial text with target text to form text segments in the video editing draft. The text segments are generated based on the target text and text editing information. In the editing result corresponding to the video editing draft, the target text is presented as text effects. The generation unit 130 can execute step S30 of the video editing method shown in Figure 1A.
[0161] This device can generate text templates, allowing users to generate the text they need, such as subtitles, simply by calling the template. This improves the efficiency of creating text in videos and reduces the workload.
[0162] For example, the display unit 110, the acquisition unit 120, and the generation unit 130 can be hardware, software, firmware, or any feasible combination thereof. For example, the display unit 110, the acquisition unit 120, and the generation unit 130 can be dedicated or general-purpose circuits, chips, or devices, or a combination of a processor and memory. The embodiments of this disclosure do not limit the specific implementation of the display unit 110 and the generation unit 130.
[0163] It should be noted that in the embodiments of this disclosure, each unit of the video editing device 500 corresponds to each step of the aforementioned method. For details regarding the specific functions of the video editing device 500, please refer to the relevant description of the method above; further details will not be repeated here. The components and structure of the video editing device 500 shown in Figure 5 are merely exemplary and not restrictive. The video editing device 500 may also include other components and structures as needed.
[0164] Figure 6 is a schematic block diagram of an electronic device provided in some embodiments of this disclosure. As shown in Figure 6, the electronic device 600 includes a processor 210 and a memory 220. The memory 220 is used to store non-transitory computer-readable instructions (e.g., one or more computer program modules). The processor 210 is used to execute the non-transitory computer-readable instructions, which, when executed by the processor 210, can perform one or more steps in the video editing method described above. The memory 220 and the processor 210 can be interconnected via a bus system and / or other forms of connection mechanism (not shown).
[0165] For example, processor 210 can be a central processing unit (CPU), a digital signal processor (DSP), or other processing units with data processing and / or program execution capabilities, such as a field-programmable gate array (FPGA); for example, the central processing unit (CPU) can be an x86 or ARM architecture. Processor 210 can be a general-purpose processor or a special-purpose processor, and can control other components in electronic device 600 to perform desired functions.
[0166] For example, memory 220 may include any combination of one or more computer program products, which may include various forms of computer-readable storage media, such as volatile memory and / or non-volatile memory. Volatile memory may include, for example, random access memory (RAM) and / or cache memory. Non-volatile memory may include, for example, read-only memory (ROM), hard disk, erasable programmable read-only memory (EPROM), portable compact disc read-only memory (CD-ROM), USB memory, flash memory, etc. One or more computer program modules may be stored on the computer-readable storage medium, and processor 210 may run one or more computer program modules to implement various functions of electronic device 600. Various application programs and various data, as well as various data used and / or generated by the application programs, may also be stored in the computer-readable storage medium.
[0167] It should be noted that, in the embodiments of this disclosure, the specific functions and technical effects of the electronic device 600 can be referred to the description of the video editing method above, and will not be repeated here.
[0168] Figure 7 is a schematic block diagram of another electronic device provided in some embodiments of this disclosure. This electronic device 300 is, for example, suitable for implementing the video editing method provided in the embodiments of this disclosure. The electronic device 300 may be a user terminal, etc. It should be noted that the electronic device 300 shown in Figure 7 is merely an example and does not impose any limitations on the functionality and scope of use of the embodiments of this disclosure.
[0169] As shown in Figure 7, the electronic device 300 may include a processing unit (e.g., a central processing unit, a graphics processing unit, etc.) 310, which can perform various appropriate actions and processes according to a program stored in a read-only memory (ROM) 320 or a program loaded from a storage device 380 into a random access memory (RAM) 330. The RAM 330 also stores various programs and data required for the operation of the electronic device 300. The processing unit 310, the ROM 320, and the RAM 330 are interconnected via a bus 340. An input / output (I / O) interface 350 is also connected to the bus 340.
[0170] Typically, the following devices can be connected to I / O interface 350: input devices 360 including, for example, touchscreens, touchpads, keyboards, mice, cameras, microphones, accelerometers, gyroscopes, etc.; output devices 370 including, for example, liquid crystal displays (LCDs), speakers, vibrators, etc.; storage devices 380 including, for example, magnetic tapes, hard disks, etc.; and communication devices 390. Communication device 390 allows electronic device 300 to communicate wirelessly or wiredly with other electronic devices to exchange data. Although Figure 7 shows electronic device 300 with various devices, it should be understood that it is not required to implement or have all the devices shown, and electronic device 300 may alternatively implement or have more or fewer devices.
[0171] For example, according to embodiments of this disclosure, the video editing method shown in FIG1A can be implemented as a computer software program. For example, embodiments of this disclosure include a computer program product comprising a computer program carried on a non-transitory computer-readable medium, the computer program including program code for performing the video editing method described above. In such embodiments, the computer program can be downloaded and installed from a network via a communication device 390, or installed from a storage device 380, or installed from a ROM 320. When the computer program is executed by the processing device 310, it can perform the functions defined in the video editing method provided by embodiments of this disclosure.
[0172] At least one embodiment of this disclosure also provides a storage medium for storing non-transitory computer-readable instructions that, when executed by a computer, can implement the video editing method described in any embodiment of this disclosure. This storage medium enables the generation of text templates, allowing users to generate the required text, such as subtitles, simply by calling the text template, thus improving the efficiency of creating text in videos and reducing the workload.
[0173] Figure 8 is a schematic diagram of a storage medium provided in some embodiments of this disclosure. As shown in Figure 8, the storage medium 400 is used to store non-transitory computer-readable instructions 410. For example, when the non-transitory computer-readable instructions 410 are executed by a computer, one or more steps in the video editing method described above can be performed.
[0174] For example, the storage medium 400 can be used in the aforementioned electronic device 600. For example, the storage medium 400 can be the memory 220 in the electronic device 600 shown in FIG. 6. For example, for related descriptions of the storage medium 400, please refer to the corresponding description of the memory 220 in the electronic device 600 shown in FIG. 6, which will not be repeated here.
[0175] The video editing method, apparatus, electronic device, and storage medium provided in the embodiments of this disclosure have been described above with reference to Figures 1A to 8. The video editing method provided in the embodiments of this disclosure can generate text templates, allowing users to generate the required text, such as subtitles, simply by calling the text template, thus improving the efficiency of creating text in videos and reducing the workload of creating text in videos.
[0176] It should be noted that the storage medium (computer-readable medium) described above in this disclosure can be a computer-readable signal medium or a non-transitory computer-readable storage medium, or any combination of the two. A non-transitory computer-readable storage medium can be, for example,, but not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination thereof. More specific examples of a non-transitory computer-readable storage medium may include, but are not limited to: an electrical connection having one or more wires, a portable computer disk, a hard disk, random access memory (RAM), read-only memory (ROM), erasable programmable read-only memory (EPROM or flash memory), optical fiber, portable compact disk read-only memory (CD-ROM), optical storage device, magnetic storage device, or any suitable combination thereof. In this disclosure, a non-transitory computer-readable storage medium can be any tangible medium containing or storing a program that can be used by or in conjunction with an instruction execution system, apparatus, or device. In this disclosure, a computer-readable signal medium can include a data signal propagated in baseband or as part of a carrier wave, carrying computer-readable program code. Such transmitted data signals can take various forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination thereof. The computer-readable signal medium can also be any computer-readable medium other than a non-transitory computer-readable storage medium, which can send, propagate, or transmit a program for use by or in connection with an instruction execution system, apparatus, or device. The program code contained on the computer-readable medium can be transmitted using any suitable medium, including but not limited to: wires, optical fibers, RF (radio frequency), etc., or any suitable combination thereof.
[0177] In some implementations, clients and servers can communicate using any currently known or future-developed network protocol, such as Hypertext Transfer Protocol (HTTP), and can interconnect with digital data communication (e.g., communication networks) of any form or medium. Examples of communication networks include local area networks (LANs), wide area networks (WANs), the Internet (e.g., the Internet), and peer-to-peer networks (e.g., ad hoc peer-to-peer networks), as well as any currently known or future-developed networks.
[0178] The aforementioned computer-readable medium may be included in the aforementioned electronic device; or it may exist independently and not assembled into the electronic device.
[0179] The aforementioned computer-readable medium carries one or more programs that, when executed by the electronic device, cause the electronic device to perform one or more steps according to the video editing method described above. Computer program code for performing the operations of this disclosure can be written in one or more programming languages or combinations thereof, including but not limited to object-oriented programming languages such as Java, Smalltalk, and C++, as well as conventional procedural programming languages such as the "C" language or similar programming languages. The program code can be executed entirely on the user's computer, partially on the user's computer, as a standalone software package, partially on the user's computer and partially on a remote computer, or entirely on a remote computer or server. In cases involving remote computers, the remote computer can be connected to the user's computer via any type of network, such as a local area network (LAN) or a wide area network (WAN), or can be connected to an external computer (e.g., via the Internet using an Internet service provider).
[0180] The flowcharts and block diagrams in the accompanying drawings illustrate the architecture, functionality, and operation of possible implementations of systems, methods, and computer program products according to various embodiments of this disclosure. In this regard, each block in a flowchart or block diagram may represent a module, segment, or portion of code containing one or more executable instructions for implementing a specified logical function. It should also be noted that in some alternative implementations, the functions indicated in the blocks may occur in a different order than those indicated in the drawings. For example, two consecutively indicated blocks may actually be executed substantially in parallel, and they may sometimes be executed in reverse order, depending on the functions involved. It should also be noted that each block in the block diagrams and / or flowcharts, and combinations of blocks in the block diagrams and / or flowcharts, can be implemented using a dedicated hardware-based system that performs the specified function or operation, or using a combination of dedicated hardware and computer instructions.
[0181] The units described in the embodiments of this disclosure can be implemented in software or hardware. The names of the units are not, in some cases, intended to limit the specific unit.
[0182] The functions described above in this document can be performed, at least in part, by one or more hardware logic components. For example, exemplary types of hardware logic components that can be used, without limitation, include: Field Programmable Gate Arrays (FPGAs), Application-Specific Integrated Circuits (ASICs), Application Standard Products (ASSPs), System-on-Chip (SoCs), Complex Programmable Logic Devices (CPLDs), etc.
[0183] In this disclosure, a machine-readable medium can be a tangible medium that may contain or store a program for use by or in conjunction with an instruction execution system, apparatus, or device. A machine-readable medium can be a machine-readable signal medium or a machine-readable storage medium. Machine-readable media can be, but is not limited to, electronic, magnetic, optical, electromagnetic, infrared, or semiconductor systems, apparatus, or devices, or any suitable combination of the foregoing. More specific examples of machine-readable storage media include electrical connections based on one or more wires, portable computer disks, hard disks, random access memory (RAM), read-only memory (ROM), erasable programmable read-only memory (EPROM or flash memory), optical fiber, portable compact disk read-only memory (CD-ROM), optical storage devices, magnetic storage devices, or any suitable combination of the foregoing.
[0184] The above description is merely a partial embodiment of this disclosure and an explanation of the technical principles employed. Those skilled in the art should understand that the scope of this disclosure is not limited to technical solutions formed by specific combinations of the above-described technical features, but should also cover other technical solutions formed by arbitrary combinations of the above-described technical features or their equivalents without departing from the above-described concept. For example, technical solutions formed by substituting the above features with (but not limited to) technical features disclosed in this disclosure that have similar functions.
[0185] Furthermore, while the operations are described in a specific order, this should not be construed as requiring these operations to be performed in the specific order shown or in sequential order. Multitasking and parallel processing may be advantageous in certain environments. Similarly, while several specific implementation details are included in the above discussion, these should not be construed as limiting the scope of this disclosure. Certain features described in the context of individual embodiments may also be implemented in combination in a single embodiment. Conversely, various features described in the context of a single embodiment may also be implemented individually or in any suitable sub-combination in multiple embodiments.
[0186] Although the subject matter has been described using language specific to structural features and / or methodological logic, it should be understood that the subject matter defined in the appended claims is not necessarily limited to the specific features or actions described above. Rather, the specific features and actions described above are merely illustrative examples of implementing the claims.
Claims
1. A video editing method, comprising: displaying an editing interface of a text template, wherein the editing interface comprises an editing area and a preview area, the editing area is used for editing the text template, and the preview area is used for previewing an editing effect of the text template; in response to an editing operation on the editing interface, obtaining initial text and text editing information, the text editing information being used for indicating a text effect applied on the initial text, the text effect comprising a text visual effect and / or a text sound effect; generating the text template according to the initial text and the text editing information, wherein the text template is used for replacing the initial text with target text to form a text segment in a video editing draft, the text segment being generated based on the target text and the text editing information, and the target text is presented with the text effect in an editing result corresponding to the video editing draft.
2. The method of claim 1, wherein, the editing area presents an edit track of the text template, the edit track at least comprising a text track, and the text track is used for placing at least one text track segment formed based on the initial text.
3. The method of claim 2, wherein, in the text template, the initial text is displayed in an initial video, the initial video comprises at least one picture, the initial text comprises a plurality of characters, characters in the plurality of characters located in a same picture in the initial video are taken as one text track segment, and a time position of the at least one text track segment in the text track corresponds to a time position of at least one picture in the initial video in the text template respectively.
4. The method of claim 3, wherein, in the text track, two adjacent text track segments are separated by a separation mark, and the separation mark is close to a text track segment later in time among the two adjacent text track segments in the text track.
5. The method according to any one of claims 2-4, wherein, the text visual effect comprises animation, and the edit track further comprises an animation track, the animation track being used for displaying an application range of the animation in the initial text.
6. The method of claim 5, further comprising: in response to a selection operation on the animation track, displaying a speed curve of the animation; and in response to an adjustment operation on the speed curve, adjusting a playing speed of the animation. the text editing information comprises layout information of the initial text, and the generating the text template according to the initial text and the text editing information comprises:
7. The method according to any one of claims 1 to 6, wherein, determining a character number of each line of text in the text template and a number of text lines in each picture in the video editing draft according to the layout information; and generating the text template comprising the initial text according to the character number of each line of text and the number of text lines in each picture. the determining the character number of each line of text in the text template according to the layout information comprises:
8. The method of claim 7, wherein, in response to the layout information comprising intelligently splitting the initial text, splitting the initial text into multiple lines of text. the text editing information comprises first animation information, 9. The method of any one of claims 1-8, wherein, the generating the text template according to the initial text and the text editing information comprises: generate, according to the first animation information, a first animation applied to the initial text.
10. The method of claim 9, wherein, The animation information includes a first animation mode and first configuration information, generate, according to the first animation information, a first animation applied to the initial text, including: determining, based on the first animation mode, an effective range of the first configuration information in the initial text; and generating the first animation based on the effective range and the first configuration information.
11. The method of claim 10, wherein, The effective range includes a single word, a single subtitle, or a single picture.
12. The method of claim 10 or 11, wherein, The first configuration information includes at least one of: color of characters in the effective range, zoom ratio of characters in the effective range, stroke information, background information, light-emitting information, and shadow information.
13. The method according to any one of claims 10-12, wherein, The generating, according to the initial text and the text editing information, the text template further includes: in response to obtaining second animation information, determining whether a second animation mode in the second animation information is the same as the first animation mode; in response to the second animation mode being different from the first animation mode, generating, based on the second animation information, a second animation applied to the initial text.
14. The method of claim 13, wherein, The generating, according to the initial text and the text editing information, the text template further includes: in response to the second animation mode being the same as the first animation mode, displaying prompt information, wherein the prompt information is used to indicate that the animation of the first animation mode has been applied to the initial text.
15. A video editing apparatus, comprising: a display unit configured to display an editing interface of a text template, wherein the editing interface includes an editing area and a preview area, the editing area is used to edit the text template, and the preview area is used to preview an editing effect of the text template; an obtaining unit configured to, in response to an editing operation on the editing interface, obtain an initial text and text editing information, the text editing information being used to indicate a text effect applied to the initial text, the text effect including a text visual effect and / or a text sound effect; and a generating unit configured to generate the text template according to the initial text and the text editing information, wherein the text template is used to replace the initial text with a target text to form a text segment in a video editing draft, the text segment being generated based on the target text and the text editing information, and the target text being presented with the text effect in an editing result corresponding to the video editing draft.
16. An electronic device, comprising: a processor; a memory including one or more computer program instructions; wherein the one or more computer program instructions are stored in the memory and implemented by the processor to implement the video editing method of any one of claims 1-14.
17. A computer-readable storage medium having non-transitorily stored thereon computer-readable instructions, wherein, The computer readable instructions, when executed by the processor, implement the video editing method of any one of claims 1-14. The computer readable instructions, when executed by the processor, implement the video editing method of any one of claims 1-14.
Citation Information
Patent Citations
Subtitle making system
CN102082933A
Subtitle inserting system and method
CN105025378A
Video processing method and device, storage medium, and electronic device
CN108924622A
Method for automatically loading subtitles
CN112988005A
Caption using method in non-linear edit system
CN1417798A