Method and apparatus for content editing, and device and storage medium

By automatically identifying and highlighting the key subtitle content in multimedia works, the problem of manual adjustment of subtitle style in the prior art is solved, and editing efficiency is improved.

WO2025093008A1PCT designated stage expired Publication Date: 2025-05-08BEIJING ZITIAO NETWORK TECH CO LTD
View PDF 4 Cites 0 Cited by

Patent Information

Application Number
PCT/CN2024/129505
Authority / Receiving Office
WO · WO
Patent Type
Applications
Current Assignee / Owner
Priority Date
2023-11-03
Filing Date
2024-11-01
Publication Date
2025-05-08

AI Technical Summary

Technical Problem

When editing subtitles in multimedia works, the prior art needs to manually adjust the subtitle style to highlight key information, which consumes a lot of time and has a high operating threshold.

Method used

By obtaining text content associated with media content, a set of target characters in the text content is identified and a first display style is applied to highlight the target characters.

Benefits of technology

It automatically identifies the keyword characters in the text content and applies the corresponding display style, thereby improving the efficiency of content editing.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN2024129505_08052025_PF_FP_ABST
    Figure CN2024129505_08052025_PF_FP_ABST
Patent Text Reader

Abstract

The embodiments of the present disclosure relate to a method and apparatus for content editing, and a device and a storage medium. The method provided herein comprises: acquiring text content associated with media content, wherein the text content is used for being presented in an image of the media content; identifying a set of target characters in the text content, wherein the set of target characters is determined on the basis of the analysis of the text content; and applying a first display style to a set of target characters in target text. In this way, in the embodiments of the present disclosure, key characters in text content can be automatically identified, and a corresponding display style is applied, such that the content editing efficiency can be improved.
Need to check novelty before this filing date? Find Prior Art

Description

Content editing method, device, equipment and storage medium

[0001] This application claims priority to the Chinese invention patent application entitled “Method, apparatus, device and storage medium for content editing” filed on November 3, 2023, with application number 202311460206.8, the entire contents of which are incorporated by reference into this application. Technical Field

[0002] Example embodiments of the present disclosure generally relate to the field of computers, and more particularly, to a method, apparatus, device, and computer-readable storage medium for content editing. Background Art

[0003] With the development of computer technology, more and more users are beginning to share and access various types of works through the Internet. For multimedia works such as videos, text content such as subtitles can help creators convey information more effectively and help viewers understand the works.

[0004] Some traditional editing tools can help creators automatically generate subtitles. However, users still need to manually adjust the subtitle style to highlight certain key information they want to convey, which requires a lot of time and has a relatively high operational threshold.

[0005] Summary of the Invention

[0006] In a first aspect of the present disclosure, a method for content editing is provided. The method includes: obtaining text content associated with media content, the text content being used to be presented in an image of the media content; identifying a set of target characters in the text content, wherein the set of target characters is determined based on an analysis of the text content; and applying a first display style to the set of target characters in the target text.

[0007] In a second aspect of the present disclosure, a device for content editing is provided. The device includes: an acquisition module configured to acquire text content associated with media content, the text content being used to be presented in an image of the media content; an identification module configured to identify a set of target characters in the text content, wherein the set of target characters is determined based on an analysis of the text content; and an editing module configured to apply a first display style to the set of target characters in the target text.

[0008] In a third aspect of the present disclosure, an electronic device is provided. The device includes at least one processing unit; and at least one memory coupled to the at least one processing unit and storing instructions for execution by the at least one processing unit. When executed by the at least one processing unit, the instructions cause the device to perform the method of the first or second aspect.

[0009] In a fourth aspect of the present disclosure, a computer-readable storage medium is provided, wherein a computer program is stored on the computer-readable storage medium, and the computer program can be executed by a processor to implement the method of the first aspect or the second aspect.

[0010] It should be understood that the content described in this summary section is not intended to limit the key features or important features of the embodiments of the present disclosure, nor is it intended to limit the scope of the present disclosure. Other features of the present disclosure will become easily understood through the following description. BRIEF DESCRIPTION OF THE DRAWINGS

[0011] The above and other features, advantages and aspects of the embodiments of the present disclosure will become more apparent with reference to the following detailed description in conjunction with the accompanying drawings. In the accompanying drawings, the same or similar reference numerals represent the same or similar elements, wherein:

[0012] FIG1 shows a schematic diagram of an example environment in which embodiments according to the present disclosure may be implemented;

[0013] 2A to 2D illustrate example interfaces according to some embodiments of the present disclosure;

[0014] 3A and 3B illustrate example interfaces according to some embodiments of the present disclosure;

[0015] 4A and 4B illustrate example interfaces according to some embodiments of the present disclosure;

[0016] FIG5 shows a flowchart of an example process of content editing according to some embodiments of the present disclosure;

[0017] FIG6 shows a schematic structural block diagram of an apparatus for content editing according to some embodiments of the present disclosure; and

[0018] FIG7 shows a block diagram of an electronic device capable of implementing various embodiments of the present disclosure. DETAILED DESCRIPTION

[0019] The following describes embodiments of the present disclosure in more detail with reference to the accompanying drawings. Although certain embodiments of the present disclosure are shown in the accompanying drawings, it should be understood that the present disclosure can be implemented in various forms and should not be construed as limited to the embodiments described herein. Rather, these embodiments are provided to provide a more thorough and complete understanding of the present disclosure. It should be understood that the drawings and embodiments of the present disclosure are for illustrative purposes only and are not intended to limit the scope of protection of the present disclosure.

[0020] It should be noted that the titles of any section / subsection provided herein are not limiting. Various embodiments are described throughout this document, and any type of embodiment may be included under any section / subsection. Furthermore, the embodiments described in any section / subsection may be combined in any manner with any other embodiments described in the same section / subsection and / or in different sections / subsections.

[0021] In the description of the embodiments of the present disclosure, the term "including" and similar terms should be understood as open inclusion, that is, "including but not limited to". The term "based on" should be understood as "based at least in part on". The term "one embodiment" or "the embodiment" should be understood as "at least one embodiment". The term "some embodiments" should be understood as "at least some embodiments". Other explicit and implicit definitions may be included below. The terms "first", "second", etc. may refer to different or the same objects. Other explicit and implicit definitions may be included below.

[0022] The embodiments of the present disclosure may involve user data, data acquisition and / or use, etc. These aspects shall comply with the corresponding laws, regulations and relevant provisions. In the embodiments of the present disclosure, all data collection, acquisition, processing, processing, forwarding, use, etc. are carried out on the premise that the user is aware of and confirms them. Accordingly, when implementing the various embodiments of the present disclosure, the types, scope of use, and usage scenarios of the data or information that may be involved should be informed to the user and the user's authorization should be obtained in an appropriate manner in accordance with the relevant laws and regulations. The specific notification and / or authorization method may vary according to the actual situation and application scenario, and the scope of the present disclosure is not limited in this respect.

[0023] If this specification and the solutions in the examples involve the processing of personal information, such processing will be done only with a legitimate basis (such as with the consent of the subject of personal information or as necessary for the performance of a contract) and only within the prescribed or agreed scope. A user's refusal to process personal information other than that required for basic functions will not affect the user's use of basic functions.

[0024] As briefly mentioned above, text content such as subtitles is an important content element in media works. High-quality text content in media works can help viewers obtain information from multimedia works more efficiently.

[0025] Some traditional editing tools can help creators automatically generate text content such as subtitles. However, users still need to manually adjust the text style to highlight certain key information, which takes a lot of time.

[0026] Embodiments of the present disclosure provide a solution for content editing. According to the solution, text content associated with media content is obtained, the text content being used to be presented in an image of the media content; a set of target characters in the text content is identified, wherein the set of target characters is determined based on an analysis of the text content; and a first display style is applied to the set of target characters in the target text content.

[0027] In this way, the embodiments of the present disclosure can automatically identify key characters in text content and apply corresponding display styles, thereby improving content editing efficiency.

[0028] Various example implementations of this solution are described in detail below in conjunction with the accompanying drawings.

[0029] Sample Environment

[0030] FIG1 shows a schematic diagram of an example environment 100 in which embodiments of the present disclosure can be implemented. As shown in FIG1 , the example environment 100 may include an electronic device 110 .

[0031] In this example environment 100, electronic device 110 may run an application 120 that supports interface interaction. Application 120 may be any suitable type of application for interface interaction, examples of which may include, but are not limited to, work editing applications, video creation applications, and any other suitable applications with subtitle editing capabilities. User 140 may interact with application 120 via electronic device 110 and / or its attached devices.

[0032] In the environment 100 of FIG1 , if the application 120 is active, the electronic device 110 may present an interface 150 for supporting interface interaction through the application 120. Such an interface 150 may be used, for example, to edit various types of works.

[0033] In some embodiments, the electronic device 110 communicates with the server 130 to enable the provision of services for the application 120. The electronic device 110 can be any type of mobile terminal, fixed terminal or portable terminal, including a mobile phone, a desktop computer, a laptop computer, a notebook computer, a netbook computer, a tablet computer, a media computer, a multimedia tablet, a handheld computer, a portable game terminal, a VR / AR device, a personal communication system (PCS) device, a personal navigation device, a personal digital assistant (PDA), an audio / video player, a digital camera / camcorder, a positioning device, a television receiver, a radio broadcast receiver, an e-book device, a gaming device or any combination thereof, including accessories and peripherals of these devices or any combination thereof. In some embodiments, the electronic device 110 can also support any type of interface for the user (such as a "wearable" circuit, etc.).

[0034] The server 130 may be a standalone physical server, a server cluster or distributed system consisting of multiple physical servers, or a cloud server that provides basic cloud computing services such as cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communications, middleware services, domain name services, security services, content distribution networks, and big data and artificial intelligence platforms. For example, the server 130 may include a computing system / server such as a mainframe, an edge computing node, a computing device in a cloud environment, and the like. The server 130 may provide background services for the application 120 that supports virtual scenes in the electronic device 110.

[0035] A communication connection may be established between the server 130 and the electronic device 110. The communication connection may be established in a wired or wireless manner. The communication connection may include, but is not limited to, a Bluetooth connection, a mobile network connection, a Universal Serial Bus (USB) connection, a Wireless Fidelity (WiFi) connection, etc., and the embodiments of the present disclosure are not limited in this respect. In the embodiments of the present disclosure, the server 130 and the electronic device 110 may implement signaling interaction through the communication connection between the two.

[0036] It should be understood that the structure and function of the various elements in the environment 100 are described for illustrative purposes only and do not imply any limitation on the scope of the present disclosure.

[0037] Some example embodiments of the present disclosure will be described below with continued reference to the accompanying drawings.

[0038] Subtitle editing

[0039] An exemplary subtitle editing process according to an embodiment of the present disclosure will be described below with reference to the accompanying drawings.

[0040] Figure 2A illustrates an example interface 200A according to some embodiments of the present disclosure. Interface 200A may, for example, correspond to an editing interface. This interface may be provided, for example, by electronic device 110, as shown in Figure 1. As an example, electronic device 110 may utilize a locally deployed editing application to provide interface 200A. Alternatively, electronic device 110 may utilize a browser application to provide interface 200A for online editing.

[0041] As shown in Figure 2A, the interface 200A may include, for example, a playback area for media content 205. For example, taking Figure 2A as an example, the media content 205 may include, for example, video content.

[0042] In some embodiments, such media content 205 may also include, for example, appropriate media content having a certain length of time, such as dynamic pictures, electronic photo albums, audio, and the like.

[0043] 2A , the electronic device 110 may present corresponding text content 210 in association with the media content 205. In some embodiments, the electronic device 110 may automatically generate the text content 210 based on audio content associated with the media content 205, for example.

[0044] In some embodiments, text content 210 may include subtitle content associated with media content 205. Accordingly, interface 200A may correspond to an editor for media content 205. As shown in FIG2A , the editor may include multiple tracks, such as a video track, an audio track, and a subtitle track. In addition, the editor may also include a subtitle track corresponding to the subtitle content.

[0045] In some embodiments, the electronic device 110 may generate corresponding subtitle content 210 based on a transcription of the audio portion of the video content, for example.

[0046] 2A , the electronic device 110 may also display such text content 210 in the subtitle editing window of the editor. For example, the electronic device 110 may display text content 215-1, text content 215-2, text content 215-3, text content 215-4, text content 215-5, and text content 215-6 (individually or collectively referred to as text content 215).

[0047] In some embodiments, such text contents 215 - 1 to 215 - 6 may correspond to different sentences in the text content 210 .

[0048] Correspondingly, during the corresponding time period, the electronic device 200A can display the text content 210 corresponding to the text content 215. For example, the text content 215-1 corresponds to the text content 210 currently presented in the image of the media content 205.

[0049] In some embodiments, the electronic device 110 can also support, for example, the user to edit the text content 215. For example, the electronic device 110 can receive the user's editing operation on the text content 210 or the text content 215, and can correspondingly update the text content 210 and / or the text content 215.

[0050] In some embodiments, as shown in FIG. 2A, the electronic device 110 can identify a set of target characters 220 in the text content 215-1. The set of target characters 220 can be displayed in different styles in the media content 205, for example, to facilitate the audience to pay attention to such characters.

[0051] Taking FIG. 2A as an example, the set of target characters 220 can include two Chinese characters "漂亮". In some embodiments, the set of target characters 220 can include one or more characters, for example, one or more Chinese characters, one or more English words, one or more Arabic letters, etc.

[0052] In some embodiments, as shown in FIG. 2A, the electronic device 110 can highlight a set of target characters 220 in the text content 215-1 in the subtitle editor, so that the display style of the set of target characters 220 in the subtitle editing window is different from other characters in the subtitle content.

[0053] In some embodiments, the set of target characters 220 in the text content 215-1 can be automatically determined by the electronic device 110 and / or other appropriate electronic devices (for example, the server 130) based on the analysis of the text content 215-1.

[0054] As an example, the electronic device 110 can, for example, send the text content 215-1 to a target model deployed on an appropriate electronic device (for example, the electronic device 110 or the server 130), so that the target model can determine one or more target characters based on the analysis of the text content 215-1.

[0055] Such a target model can be implemented, for example, based on rules and / or any appropriate machine learning model. For example, such a target model can determine the set of target characters 220 based on word segmentation of the text content 215-1 and matching it with a preset vocabulary. As another example, the target model is trained based on training data, for example, and outputs the set of target characters 220 by processing feature representations of the text content 215-1. This disclosure is not intended to limit the specific manner in which the target model outputs the set of target characters 220.

[0056] 2A , the electronic device 110 may apply a first display style to the group of target characters 220 so that the group of characters 220 is presented in the image of the media content 205 based on the first display style. As shown in FIG2A , the display style of the group of target characters 220 may be different from the display style of other characters.

[0057] In some embodiments, the first display style applied to the group of target characters 220 may include a preset display style.

[0058] In other embodiments, the first display style applied to the group of target characters 220 may also include a display style determined based on user input. For example, the electronic device 110 may support the user to configure the display style to be applied to the target characters, so that the configured display style can be automatically applied to one or more characters identified as target characters.

[0059] In some embodiments, the group of target characters 220 may be, for example, key content in the text content 215-1, and the first display style applied to the group of target characters 220 may be such that the group of target characters 220 has a higher prominence. For example, taking subtitles as an example, the group of target characters in the subtitles of the media content 205 may have a larger font size, be displayed in bold, or be displayed in a more eye-catching color.

[0060] Based on the process described above, the embodiments of the present disclosure can automatically identify key content in text content and apply corresponding display styles, thereby improving the editing efficiency of the text content.

[0061] In some embodiments, the embodiments of the present disclosure may further support the user to adjust the identification status of characters, for example, to identify one or more characters as new target characters, or to cancel the identification of one or more existing target characters.

[0062] In some embodiments, in response to receiving an edit request from the user (also referred to as a second edit request), the electronic device 110 may present an edit window (also referred to as a second edit window) for adjusting the identification state.

[0063] Exemplarily, in the case of receiving a selection of the control 215 as shown in FIG. 2A, the electronic device 110 may present an interface 200B as shown in FIG. 2B. The interface 200B may include an editing window 230.

[0064] As shown in FIG. 2B, the editing window 230 displays a set of text elements corresponding to the text content, and each text element includes at least one character. Taking the text content 215-1 as an example, it may correspond to a set of text elements 235, which may include, for example, four text elements "wind", "scene", "nice", "pretty" and "bright".

[0065] In some embodiments, the set of text elements 235 may be determined based on word segmentation processing of the text content 215-1. Taking Chinese text as an example, each Chinese single character may correspond to a text element. For English text, for example, each English word may correspond to a text element.

[0066] It should be understood that the way such text elements are segmented can be appropriately selected according to the language type or category of the text content. For example, a single phrase in Chinese text can also be segmented into a text element. Again, for example, a single text element can correspond to a single character.

[0067] In some embodiments, the electronic device 110 can also distinguish, for example, text elements including target characters and text elements not including target characters in the set of text elements. For example, the electronic device 110 can display text elements including text elements corresponding to the target characters (for example, text elements "pretty" and "bright") in a first element style, and can display text elements corresponding to at least one other character except the set of target characters 220 in a different second element style (for example, text elements "wind", "scene" and "nice").

[0068] Furthermore, the electronic device 110 can also adjust the identification state of the characters included in at least one text element based on a preset operation on at least one text element in the set of text elements 235, and the identification state indicates whether the corresponding character is identified as a target character.

[0069] Specifically, as shown in FIG. 2C, the electronic device 110 can, for example, receive a first operation on the text elements "wind" and "scene" via the editing window 230. This first operation can be used to identify the characters included in the text elements "wind" and "scene" (that is, the character "wind" and the character "scene") as new target characters.

[0070] As an example, the user can, for example, identify the text elements "wind" and "scene" as new target characters by clicking on them respectively. As another example, the user can also, for example, use appropriate interaction operations such as dragging or selecting to identify the text elements "wind" and "scene" as new target characters at one time.

[0071] Correspondingly, as shown in FIG. 2C, the electronic device 110 can apply a third display style to the characters "wind" and "scene" to update the text content 210. In some embodiments, the third display style can be, for example, a preset display style, or can be a display style determined based on user input.

[0072] In some embodiments, the third display style can also be determined based on the display style already applied in the text content 215-1. Specifically, the electronic device 110 can, for example, determine the second display style to be applied to the characters "wind" and "scene" based on the style of the existing target characters (e.g., "drift") adjacent to the characters "wind" and "scene" in the text content 215-1.

[0073] Thus, the embodiments of the present disclosure can further maintain the unity of the display style.

[0074] Further, as shown in FIG. 2D, relative to the state shown in FIG. 2B, the electronic device 110 can also receive a second operation associated with the text element "drift" via the editing window 230. The second operation indicates canceling the identification of the original target character (i.e., the character "drift") included in the text element "drift".

[0075] As an example, the user can, for example, cancel the identification of the target character "drift" by clicking on the text element "drift" respectively. As another example, the user can also, for example, use appropriate interaction operations such as dragging or selecting to cancel the identification of the target characters corresponding to multiple text elements at one time.

[0076] Correspondingly, the electronic device 110 can apply the second display style to the second group of characters (i.e., the character "drift") to update the text content 210.

[0077] In some embodiments, the third display style is determined based on the style of non-target characters adjacent to the second group of characters in the text content. In some embodiments, the second display style is different from the first display style mentioned above, and it can be, for example, a preset display style, or can be a display style determined based on user input.

[0078] In some embodiments, the third display style may also be determined based on the display styles already applied in the text content 215-1. Specifically, the electronic device 110 may, for example, determine the third display style to be applied to the character "drift" based on the styles of non-target characters (e.g., "good") adjacent to the characters "wind" and "scene" in the text content 215-1.

[0079] Based on the processes described above, embodiments of the present disclosure can further support a user in modifying the automatically recommended target characters, thereby further improving the efficiency of content editing.

[0080] In still other embodiments, when the text content 215 as shown in FIG. 2A receives a preset editing operation, the electronic device 110 may, for example, trigger the identification of target characters in the edited text content 215 and then apply the corresponding display styles to such target characters. In some embodiments, for example, when the preset editing operation adds new content to the text content 215, the electronic device 110 triggers the re-identification of the target characters.

[0081] In some embodiments, embodiments of the present disclosure can also support the modification of the display styles applied to characters. As shown in FIG. 3A, the interface 300A may, for example, further provide a style entry 305. When a first editing request (e.g., selection of the style entry 305) is received, the electronic device 110 may present the interface 300B shown in FIG. 3B. The interface 300B may include an editing window 310.

[0082] As shown in FIG. 3B, the editing window 310 may be used to edit the display styles of a group of characters identified as target characters in the text content 215 mentioned above. Specifically, as shown in FIG. 3B, the editing window 310 may be used to edit style parameters associated with the display styles. Such style parameters may include, but are not limited to: appropriate parameters such as font, size, color, style information, etc. The present disclosure is not intended to limit the specific types of style parameters.

[0083] Further, the electronic device 110 may obtain updated parameters (e.g., updated font, updated color, etc.) via the editing window 310 and may accordingly apply the updated display styles (also referred to as the first updated styles) to the display styles of the group of target characters. For example, the electronic device 110 may globally adjust the display styles of all target characters in the entire text content 215. <H

[0084] In some embodiments, the embodiments of the present disclosure can also support adjusting the display style applied to one or more target characters. Specifically, the electronic device 110 can receive a preset operation for a target text element in a set of text elements. For example, the electronic device 110 can receive a long press operation (or a suitable operation such as a double click operation) for the text element 405, and can accordingly present an editing entry 410.

[0085] After receiving the selection of the editing entry 410, the electronic device 110 can present an interface 400B as shown in FIG. 4B. The interface 400B can include an editing window 415. As shown in FIG. 4B, the editing window 415 can be used to edit the display style of the characters (e.g., the character "亮") included in the text element 405.

[0086] Specifically, as shown in FIG. 4B, the editing window 310 can be used to edit style parameters associated with the display style. Such style parameters can include, but are not limited to: appropriate parameters such as font, size, color, style information, etc. The present disclosure is not intended to limit the specific type of style parameters.

[0087] Further, the electronic device 110 can obtain updated parameters (e.g., updated font, updated color, etc.) via the editing window 415, and can accordingly apply the updated display style (also referred to as the second updated style) to the characters (e.g., the character "亮") included in the text element 405. For example, the electronic device 110 can independently adjust the display style of single or multiple target characters in the entire text content 215.

[0088] Based on the process described above, the embodiments of the present disclosure can further support the user's unified modification or customization of the display style of all target characters, and can also support the user's independent modification or customization of the display style of single or a group of target characters. Thus, the embodiments of the present disclosure can further improve the efficiency of content editing.

[0089] In addition, it should be understood that the specific characters, specific styles, etc. mentioned in the above examples are all exemplary and are not intended to constitute a limitation to the present disclosure.

[0090] Example process

[0091] FIG. 5 shows a flowchart of an example process 500 for content editing according to some embodiments of the present disclosure. The process 500 can be implemented at the electronic device 110. The process 500 will be described below with reference to FIG. 1.

[0092] As shown in FIG. 5, at block 510, the electronic device 110 obtains text content associated with media content, and the text content is used to be presented in an image of the media content.

[0093] At block 520 , the electronic device 110 identifies a set of target characters in the text content, wherein the set of target characters is determined based on an analysis of the text content.

[0094] At block 530 , the electronic device 110 applies the first display style to a set of target characters in the target text.

[0095] In some embodiments, the textual content includes subtitle content associated with the media content.

[0096] In some embodiments, process 500 further includes presenting an editor for editing the media content, the editor including a plurality of tracks including at least a subtitle track corresponding to the subtitle content.

[0097] In some embodiments, process 500 further includes: displaying subtitle content associated with the media content in a subtitle editing window such that a set of target characters are displayed in a different style than other characters in the subtitle content in the subtitle editing window.

[0098] In some embodiments, process 500 further includes: in response to the first editing request, presenting a first editing window for editing style parameters associated with the first display style; determining a first update style via the first editing window; and applying the first update style to a set of target characters.

[0099] In some embodiments, process 500 also includes: in response to a second editing request, presenting a second editing window, the second editing window displaying a group of text elements corresponding to the text content, each text element including at least one character; and based on a preset operation for at least one text element in the group of text elements, adjusting the display style of one or more characters corresponding to at least one text element in the edited media content.

[0100] In some embodiments, a first text element and a second text element in a group of text elements are displayed in a second editing window in different element styles, the first text element corresponds to at least one target character in a group of target characters, and the second text element corresponds to at least one other character other than the group of target characters.

[0101] In some embodiments, based on a preset operation for at least one text element in a group of text elements, adjusting the display style of one or more characters corresponding to at least one text element in the edited media content includes: receiving a first operation associated with a first text element via a second editing window; and applying a second display style to at least one target character corresponding to the first text element, the second display style being different from the first display style.

[0102] In some embodiments, the second display style is determined based on a display style of a first character other than a group of target characters in the text content, the first character being adjacent to at least one target character.

[0103] In some embodiments, based on a preset operation for at least one text element in a group of text elements, adjusting the display style of one or more characters corresponding to at least one text element in the edited media content includes: receiving a second operation associated with a second text element via a second editing window; and applying a third display style to at least one other character corresponding to the second text element.

[0104] In some embodiments, the third display style is determined based on a display style of a second character in the group of target characters, the second character being adjacent to at least one other character.

[0105] In some embodiments, process 500 also includes: presenting a third editing window based on a third operation on a first text element in a group of text elements; determining a second update style via the third editing window; and applying the second update style to at least one target character corresponding to the first text element.

[0106] In some embodiments, a set of text elements is determined based on a word segmentation process of the text content.

[0107] In some embodiments, the first display style includes: a preset display style; or a display style determined based on user input.

[0108] In some embodiments, process 500 further includes: determining text content based on audio content associated with the media content; and / or obtaining text content based on an editing operation of the user.

[0109] In some embodiments, process 500 further includes: providing text content to a target model; and obtaining a set of target characters determined by the target model based on analysis of the text content.

[0110] Example devices and equipment

[0111] Embodiments of the present disclosure also provide corresponding apparatuses for implementing the above-described methods or processes. FIG6 shows a schematic structural block diagram of an example apparatus 600 for content editing according to certain embodiments of the present disclosure. Apparatus 600 may be implemented as or included in electronic device 110. Each module / component in apparatus 600 may be implemented by hardware, software, firmware, or any combination thereof.

[0112] As shown in Figure 6, the device 600 includes an acquisition module 610, which is configured to acquire text content associated with media content, and the text content is used to be presented in an image of the media content; an identification module 620, which is configured to identify a group of target characters in the text content, wherein the group of target characters is determined based on an analysis of the text content; and an editing module 630, which is configured to apply a first display style to a group of target characters in the target text.

[0113] In some embodiments, the textual content includes subtitle content associated with the media content.

[0114] In some embodiments, an editor for editing media content is presented, the editor comprising a plurality of tracks including at least a subtitle track corresponding to subtitle content.

[0115] In some embodiments, the device 600 further includes a content display module configured to: display subtitle content associated with the media content in the subtitle editing window, so that a display style of a group of target characters in the subtitle editing window is different from other characters in the subtitle content.

[0116] In some embodiments, the device 600 also includes a first update module configured to: present a first editing window in response to a first editing request, the first editing window being used to edit style parameters associated with the first display style; determine a first update style via the first editing window; and apply the first update style to a set of target characters.

[0117] In some embodiments, the device 600 also includes a second update module, which is configured to: present a second editing window in response to a second editing request, the second editing window displaying a group of text elements corresponding to the text content, each text element including at least one character; and adjust the display style of one or more characters corresponding to at least one text element in the edited media content based on a preset operation for at least one text element in the group of text elements.

[0118] In some embodiments, a first text element and a second text element in a group of text elements are displayed in a second editing window in different element styles, the first text element corresponds to at least one target character in a group of target characters, and the second text element corresponds to at least one other character other than the group of target characters.

[0119] In some embodiments, the second update module is further configured to: receive a first operation associated with the first text element via the second editing window; and apply a second display style to at least one target character corresponding to the first text element, the second display style being different from the first display style.

[0120] In some embodiments, the second display style is determined based on a display style of a first character other than a group of target characters in the text content, the first character being adjacent to at least one target character.

[0121] In some embodiments, the second updating module is further configured to: receive a second operation associated with the second text element via the second editing window; and apply a third display style to at least one other character corresponding to the second text element.

[0122] In some embodiments, the third display style is determined based on a display style of a second character in the group of target characters, the second character being adjacent to at least one other character.

[0123] In some embodiments, the device 600 also includes a third update module configured to: present a third editing window based on a third operation on a first text element in a group of text elements; determine a second update style via the third editing window; and apply the second update style to at least one target character corresponding to the first text element.

[0124] In some embodiments, a set of text elements is determined based on a word segmentation process of the text content.

[0125] In some embodiments, the first display style includes: a preset display style; or a display style determined based on user input.

[0126] In some embodiments, the apparatus 600 further includes a content acquisition module configured to: determine text content based on audio content associated with the media content; and / or acquire text content based on an editing operation of a user.

[0127] In some embodiments, the apparatus 600 further includes a content processing module configured to: provide text content to the target model; and obtain a set of target characters determined by the target model based on analysis of the text content.

[0128] FIG7 shows a block diagram of an electronic device 700 in which one or more embodiments of the present disclosure may be implemented. It should be understood that the electronic device 700 shown in FIG7 is merely exemplary and should not be construed as limiting the functionality and scope of the embodiments described herein. The electronic device 700 shown in FIG7 can be used to implement the electronic device 110 of FIG1 .

[0129] As shown in FIG7 , electronic device 700 is a general-purpose electronic device. Components of electronic device 700 may include, but are not limited to, one or more processors or processing units 710, memory 720, storage device 730, one or more communication units 740, one or more input devices 750, and one or more output devices 760. Processing unit 710 may be a real or virtual processor and is capable of performing various processes according to programs stored in memory 720. In a multi-processor system, multiple processing units execute computer-executable instructions in parallel to enhance the parallel processing capabilities of electronic device 700.

[0130] The electronic device 700 typically includes a plurality of computer storage media. Such media can be any accessible media that can be obtained by the electronic device 700, including but not limited to volatile and non-volatile media, removable and non-removable media. The memory 720 can be a volatile memory (e.g., registers, cache, random access memory (RAM)), a non-volatile memory (e.g., read-only memory (ROM), electrically erasable programmable read-only memory (EEPROM), flash memory), or some combination thereof. The storage device 730 can be a removable or non-removable medium and can include a machine-readable medium, such as a flash drive, a disk, or any other medium that can be used to store information and / or data and can be accessed within the electronic device 700.

[0131] The electronic device 700 may further include additional removable / non-removable, volatile / non-volatile storage media. Although not shown in FIG. 7 , a disk drive for reading from or writing to a removable, non-volatile disk (e.g., a “floppy disk”) and an optical drive for reading from or writing to a removable, non-volatile optical disk may be provided. In these cases, each drive may be connected to a bus (not shown) by one or more data media interfaces. The memory 720 may include a computer program product 725 having one or more program modules configured to perform various methods or actions of various embodiments of the present disclosure.

[0132] The communication unit 740 enables communication with other electronic devices via a communication medium. Additionally, the functions of the components of the electronic device 700 can be implemented as a single computing cluster or multiple computing machines that can communicate via a communication connection. Thus, the electronic device 700 can operate in a networked environment using a logical connection with one or more other servers, a network personal computer (PC), or another network node.

[0133] Input device 750 may be one or more input devices, such as a mouse, keyboard, or trackball. Output device 760 may be one or more output devices, such as a display, a speaker, or a printer. Electronic device 700 may also communicate with one or more external devices (not shown) via communication unit 740 as needed, such as storage devices, display devices, or the like, with one or more devices that allow a user to interact with electronic device 700, or with any device that allows electronic device 700 to communicate with one or more other electronic devices (e.g., a network card, a modem, etc.). Such communication may be performed via an input / output (I / O) interface (not shown).

[0134] According to an exemplary implementation of the present disclosure, a computer-readable storage medium is provided, on which computer-executable instructions are stored, wherein the computer-executable instructions are executed by a processor to implement the method described above. According to an exemplary implementation of the present disclosure, a computer program product is also provided, which is tangibly stored on a non-transitory computer-readable medium and includes computer-executable instructions, and the computer-executable instructions are executed by a processor to implement the method described above.

[0135] Various aspects of the present disclosure are described herein with reference to flowcharts and / or block diagrams of methods, apparatuses, devices, and computer program products implemented according to the present disclosure. It should be understood that each block of the flowcharts and / or block diagrams, and combinations of blocks in the flowcharts and / or block diagrams, can be implemented by computer-readable program instructions.

[0136] These computer-readable program instructions can be provided to a processing unit of a general-purpose computer, a special-purpose computer, or other programmable data processing device, thereby producing a machine, such that when these instructions are executed by the processing unit of the computer or other programmable data processing device, a device is generated that implements the functions / actions specified in one or more blocks in the flowchart and / or block diagram. These computer-readable program instructions can also be stored in a computer-readable storage medium, where these instructions cause the computer, programmable data processing device, and / or other device to operate in a specific manner. Thus, the computer-readable medium storing the instructions comprises an article of manufacture that includes instructions for implementing various aspects of the functions / actions specified in one or more blocks in the flowchart and / or block diagram.

[0137] Computer-readable program instructions can be loaded onto a computer, other programmable data processing apparatus, or other device so that a series of operational steps are performed on the computer, other programmable data processing apparatus, or other device to produce a computer-implemented process, thereby causing the instructions executed on the computer, other programmable data processing apparatus, or other device to implement the functions / actions specified in one or more boxes in the flowchart and / or block diagram.

[0138] The flow charts and block diagrams in the accompanying drawings show the possible architecture, functions and operations of the systems, methods and computer program products according to multiple implementations of the present disclosure. In this regard, each box in the flow chart or block diagram can represent a part for a module, program segment or instruction, and a part for a module, program segment or instruction comprises one or more executable instructions for realizing the logical function of the specification. In some alternative implementations, the functions marked in the box can also occur in a sequence different from that marked in the accompanying drawings. For example, two continuous boxes can actually be executed substantially in parallel, and they can sometimes be executed in the opposite order, depending on the functions involved. It should also be noted that each box in the block diagram and / or flow chart, and the combination of the boxes in the block diagram and / or flow chart can be realized by a special hardware-based system that performs the function or action of the specification, or can be realized by a combination of special hardware and computer instructions.

[0139] While various implementations of the present disclosure have been described above, the foregoing description is intended to be illustrative, not exhaustive, and not limited to the disclosed implementations. Many modifications and variations will be apparent to those skilled in the art without departing from the scope and spirit of the described implementations. The terminology used herein is selected to best explain the principles of the implementations, their practical applications, or improvements to existing technologies, or to enable others skilled in the art to understand the various implementations disclosed herein.

Claims

1. A method for content editing, comprising: Acquire text content associated with media content, wherein the text content is used to be presented in an image of the media content; identifying a set of target characters in the text content, wherein the set of target characters is determined based on an analysis of the text content; and A first display style is applied to the set of target characters in the target text. 2 . The method of claim 1 , wherein the text content comprises subtitle content associated with the media content.

3. The method according to claim 2, further comprising: An editor for editing the media content is presented, the editor comprising a plurality of tracks including at least a subtitle track corresponding to the subtitle content.

4. The method according to claim 2, further comprising: In a subtitle editing window of the editor, the subtitle content associated with the media content is displayed, so that a display style of the group of target characters in the subtitle editing window is different from other characters in the subtitle content.

5. The method according to claim 1, further comprising: In response to a first editing request, presenting a first editing window, wherein the first editing window is used to edit style parameters associated with the first display style; Determining a first update style via the first editing window; as well as The first update pattern is applied to the set of target characters.

6. The method according to claim 1, further comprising: In response to the second editing request, presenting a second editing window, wherein the second editing window displays a group of text elements corresponding to the text content, each text element including at least one character; as well as Based on a preset operation for at least one text element in the group of text elements, a display style of one or more characters corresponding to the at least one text element in the edited media content is adjusted.

7. A method according to claim 6, wherein a first text element and a second text element in the group of text elements are displayed in the second editing window with different element styles, the first text element corresponds to at least one target character in the group of target characters, and the second text element corresponds to at least one other character except the group of target characters.

8. The method according to claim 7, wherein based on a preset operation for at least one text element in the group of text elements, adjusting the display style of one or more characters corresponding to the at least one text element in the edited media content comprises: receiving, via the second editing window, a first operation associated with the first text element; as well as A second display style is applied to the at least one target character corresponding to the first text element, the second display style being different from the first display style.

9. The method according to claim 8, wherein the second display style is determined based on a display style of a first character in the text content except the group of target characters, the first character being adjacent to the at least one target character.

10. The method according to claim 7, wherein based on a preset operation for at least one text element in the group of text elements, adjusting the display style of one or more characters corresponding to the at least one text element in the edited media content comprises: receiving, via the second editing window, a second operation associated with the second text element; as well as A third display style is applied to the at least one other character corresponding to the second text element.

11. The method of claim 10, wherein the third display style is determined based on a display style of a second character in the group of target characters, the second character being adjacent to the at least one other character.

12. The method according to claim 7, further comprising: Based on a third operation on the first text element in the group of text elements, presenting a third editing window; Determining a second update style via the third editing window; as well as The second update style is applied to the at least one target character corresponding to the first text element.

13. The method according to claim 6, wherein the set of text elements is determined based on word segmentation processing of the text content.

14. The method according to claim 1, wherein the first display style comprises: Preset display styles; or Display style determined based on user input.

15. The method according to claim 1, further comprising: determining the text content based on audio content associated with the media content; and / or Based on the user's editing operation, the text content is obtained.

16. The method according to claim 1, further comprising: Providing the text content to the target model; as well as The set of target characters determined by the target model based on the analysis of the text content is obtained.

17. A device for content editing, comprising: An acquisition module, configured to acquire text content associated with the media content, wherein the text content is used to be presented in an image of the media content; The identification module is configured to identify a group of target characters in the text content, wherein the group of target characters The symbol is determined based on an analysis of the content of the text; as well as The editing module is configured to apply a first display style to the group of target characters in the target text.

18. An electronic device, comprising: at least one processing unit; as well as At least one memory, the at least one memory being coupled to the at least one processing unit and storing instructions for execution by the at least one processing unit, the instructions causing the electronic device to perform the method according to any one of claims 1 to 16 when executed by the at least one processing unit.

19. A computer-readable storage medium having a computer program stored thereon, wherein the computer program can be executed by a processor to implement the method according to any one of claims 1 to 16.

Citation Information

Patent Citations

  • Bullet screen processing method and device, electronic equipment and computer readable storage medium

    CN111294663A

  • Oral broadcasting video generation method and device, electronic equipment and storage medium

    CN113411655A

  • Subtitle processing method and device for video conference, electronic equipment and storage medium

    CN116156098A

  • Editing timed-text elements

    US20190379943A1