Media special effect processing method and device, equipment and medium
By acquiring and processing images in response to preset trigger events in the media processing system, adding decorative effects with preset styles, and generating dynamic media, the limitations of manual processing in the prior art are solved, and efficient and effective media special effects processing and dynamic media generation are achieved.
Patent Information
- Application Number
- CN202510112268.2
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2025-01-23
- Publication Date
- 2025-06-03
AI Technical Summary
The existing technology is difficult to promote among the general population and difficult to meet the needs of users, and it usually requires manual operation of professional software.
By responsive to the preset trigger event of the preset interface, the original image is acquired and special effects are processed, decorative effects with a preset style are added, the processed target image and preset interface are displayed, and the target dynamic media is generated based on the original and target images.
No manual processing is required, which improves the efficiency and effect of special effects processing, expands the processing path and improves the user experience.
Smart Images

Figure CN120091189A_ABST
Abstract
Description
Technical Field
[0001] Embodiments of the present disclosure relate to the technical field of media processing, and in particular, to a method, apparatus, device, and medium for media special effect processing. Background Art
[0002] With the continuous development and improvement of network technology and digital media technology, digital media such as images, videos, and animated pictures are increasingly applied in people's lives, providing entertainment services for people's lives, bringing a lot of convenience, and adding more fun to people's lives. In order to make the media more vivid, people usually need to perform special effect processing on the existing media to obtain more vivid and expressive media. For example, in some special scenarios of festivals, people usually hope to perform special effect processing on the existing media according to the corresponding festival atmosphere, so that the media after special effect processing has a stronger festival atmosphere effect. Currently, a media special effect processing solution is needed. Summary of the Invention
[0003] Embodiments of the present disclosure describe a method, apparatus, device, and medium for media special effect processing.
[0004] According to a first aspect, a method for media special effect processing is provided. The method includes: in response to a preset trigger event for a preset interface, obtaining an original image; performing special effect processing on the original image to add a decorative special effect with a preset style to the original image; displaying a target image after the special effect processing and a preset interface for the target image; and in response to a trigger operation for the preset interface, generating a target dynamic media based on the original image and the target image.
[0005] According to a second aspect, a media special effect processing apparatus is provided. The apparatus includes: an obtaining unit configured to obtain an original image in response to a preset trigger event for a preset interface; a processing unit configured to perform special effect processing on the original image to add a decorative special effect with a preset style to the original image; a display unit configured to display a target image after the special effect processing and a preset interface for the target image; and a generating unit configured to generate a target dynamic media based on the original image and the target image in response to a trigger operation for the preset interface.
[0006] According to a third aspect, a computer-readable storage medium is provided, on which a computer program is stored. When the computer program is executed on a computer, the computer is made to execute the method according to any one of the first aspect.
[0007] According to a fourth aspect, an electronic device is provided, including a memory and a processor. An executable code is stored in the memory. When the processor executes the executable code, the method according to any one of the first aspects is implemented.
[0008] According to a media special effect processing solution provided by an embodiment of the present disclosure, by responding to a preset trigger event for a preset interface, an original image is obtained, and the original image is subjected to special effect processing to add a decorative special effect with a preset style to the original image. The target image after the special effect processing and a preset interface for the target image are displayed. In response to a trigger operation for the preset interface, based on the original image and the target image, a target dynamic media is generated. Thus, based on the existing image, a target image with a decorative special effect of a preset style can be obtained, and a dynamic media that makes the target image have a dynamic effect is generated. It is not necessary to manually perform special effect processing on the image, which improves the efficiency of special effect processing on the image, improves the effect of special effect processing, expands the ways of special effect processing, and enhances the user experience. BRIEF DESCRIPTION OF THE DRAWINGS
[0009] Figure 1a FIG. is a schematic diagram of a media special effect processing interface shown according to an exemplary embodiment;
[0010] Figure 1b FIG. is a schematic diagram of another media special effect processing interface shown according to an exemplary embodiment;
[0011] Figure 1c FIG. is a schematic diagram of another media special effect processing interface shown according to an exemplary embodiment;
[0012] Figure 1d FIG. is a schematic diagram of another media special effect processing interface shown according to an exemplary embodiment;
[0013] Figure 2 FIG. is a schematic diagram of an exemplary system architecture to which the embodiments of the present disclosure are applied;
[0014] Figure 3 FIG. is a flowchart of a media special effect processing method shown according to an exemplary embodiment of the present disclosure;
[0015] Figure 4 FIG. is a block diagram of a media special effect processing device shown according to an exemplary embodiment of the present disclosure;
[0016] Figure 5 FIG. is a schematic block diagram of an electronic device provided according to an exemplary embodiment of the present disclosure. DETAILED DESCRIPTION
[0017] It is understandable that before using the technical solutions disclosed in the embodiments of the present disclosure, the types, usage scopes, usage scenarios, etc. of the personal information involved in the present disclosure should be informed to users in an appropriate manner and the authorization of users should be obtained in accordance with relevant laws and regulations.
[0018] For example, when responding to an active request from a user, a prompt message is sent to the user to clearly prompt the user that the operation requested by the user will require obtaining and using the user's personal information. Thus, the user can autonomously choose whether to provide personal information to software or hardware such as an electronic device, an application program, a server, or a storage medium that performs the operations of the technical solutions of the present disclosure according to the prompt message.
[0019] As an optional but non-limiting implementation manner, the manner of sending a prompt message to the user in response to receiving an active request from the user may be, for example, in the form of a pop-up window, and the prompt message may be presented in text in the pop-up window. In addition, the pop-up window may also carry a selection control for the user to choose "agree" or "disagree" to provide personal information to the electronic device.
[0020] It is understandable that the above process of notifying and obtaining user authorization is only illustrative and does not constitute a limitation on the implementation manner of the present disclosure. Other manners that meet relevant laws and regulations can also be applied to the implementation manner of the present disclosure.
[0021] The technical solutions provided by the present disclosure will be further described in detail below in conjunction with the accompanying drawings and embodiments. It is understandable that the specific embodiments described herein are only used to explain the related invention and do not limit the invention. In addition, it should be noted that for the convenience of description, only parts related to the relevant invention are shown in the accompanying drawings. It should be noted that, without conflict, the embodiments of the present disclosure and the features in the embodiments can be combined with each other.
[0022] With the continuous development and improvement of network technology and digital media technology, digital media such as images, videos, and animated gifs are increasingly applied to people's lives, providing entertainment services for people's lives, bringing a lot of convenience, and adding more fun to people's lives. In order to make the media more vivid, people usually need to perform special effects processing on the existing media to obtain more vivid and expressive media. For example, in some special scenarios of festivals, people usually hope to perform special effects processing on the existing media according to the corresponding festival atmosphere, so that the media after special effects processing has a stronger festival atmosphere effect. In the related art, generally, professional software is required to manually add special effects to the existing media. Therefore, it has great limitations and is difficult to be popularized among the general public and difficult to meet the needs of users.
[0023] A media special effect processing solution provided by the present disclosure obtains an original image by responding to a preset trigger event for a preset interface, performs special effect processing on the original image to add a decorative special effect with a preset style to the original image, displays the target image after the special effect processing and a preset interface for the target image, and generates a target dynamic media based on the original image and the target image in response to a trigger operation for the preset interface. Thus, a target image with a decorative special effect of a preset style can be obtained based on an existing image, and a dynamic media that gives the target image a dynamic effect is generated. There is no need for manual special effect processing of the image, which improves the efficiency of special effect processing of the image, improves the effect of special effect processing, expands the way of special effect processing, and enhances the user experience.
[0024] Figure 1a - Figure 1d FIG. is a schematic diagram of a media special effect processing interface shown according to an exemplary embodiment. The following combines Figure 1a - Figure 1d , and describes the present disclosure through the following specific application scenarios.
[0025] The specific application scenario of this embodiment can be that a media processing client is installed in the terminal device held by the user. First, the user enters a specified home page interface, and a specified button is displayed in a preset area of the home page interface. After the user clicks the specified button, an image selection interface can be output to the user, and the user can select an image from the local photo album through the image selection interface. After selecting the image and clicking the confirmation button, the media processing client can directly perform special effect processing on the image and display the processed image on the special effect processing interface.
[0026] For example, as Figure 1a shown, Image 101 is the image to be processed selected by the user. After the media processing client performs special effect processing on Image 101, the processed Image 102 can be displayed in the special effect processing interface 103. It should be noted that since the special effect processing is performed on Image 101 for the first time after Image 101 is selected, the media processing client can randomly select a special effect to process Image 101. Specifically, the media processing client can perform object recognition on Image 101 to determine that the object category of the main object included in the image is "person". Then, multiple text guidance information corresponding to "person" is obtained as candidate text guidance information, where one text guidance information corresponds to one special effect. The media processing client randomly selects a text guidance information from the multiple candidate text guidance information as the target text guidance information. For example, the target text guidance information can be "Most of the clothes turn into roses, and roses also grow around". Then, based on the target text guidance information, special effect processing is performed on Image 101 to obtain Image 102.
[0027] In the special effect processing interface 103, below the image 102, a regeneration button 104 is displayed. If the user is not satisfied with the effect of the first processing, they can click the regeneration button 104 to re-perform special effect processing on the image 101. When performing special effect processing on the image 101 again, the user can select their favorite special effect. Specifically, as Figure 1a shown, in the area 105 of the special effect processing interface 103, multiple options are displayed. Each option corresponds to a special effect, and each special effect corresponds to a text guidance message. The options can be displayed in the form of sample effect diagrams corresponding to their respective special effects. The user can select their favorite special effect based on the sample effect diagrams corresponding to the options. For example, the user can select the special effect corresponding to an option by clicking on the sample effect diagram corresponding to that option in the area 105. After the user selects a special effect, they can click the button 104 to re-perform special effect processing on the image 101 according to the selected special effect. Then, the newly processed image with the special effect is used to replace the image 102 displayed in the special effect processing interface 103.
[0028] When performing special effect processing on the image 101 again, the user can also customize their favorite special effect. Specifically, above the area 105 in the special effect processing interface 103, a custom effect button 106 is displayed. As Figure 1b shown, after the user clicks the button 106, the special effect processing interface 103 can output a custom area to the left of the image 102. A guidance information input box 107 is displayed in the custom area. The user can enter custom text-based guidance information in the guidance information input box 107. One custom guidance information can correspond to one custom special effect. Below the guidance information input box 107, an auxiliary button 108 is also displayed. The user can click the auxiliary button 108 to enable the media processing client to assist in generating text guidance information and automatically fill the assisted-generated text guidance information into the guidance information input box 107 for the user to modify and use.
[0029] Refer again to Figure 1a , above the special effect processing interface 103, an image replacement button 109 is displayed. In the area of the image replacement button 109, a thumbnail of the image 101 selected by the user can be displayed. When the user clicks the image replacement button 109, the media processing client can output an image selection interface to the user, and the user can select the image they want to replace through the image selection interface. As Figure 1cAs shown, the image 110 is the image re-selected by the user. Since it is the first time to perform special effect processing after selecting the image 110, therefore, a special effect can be directly randomly selected to process the image 110, and the processed image 111 is displayed in the special effect processing interface 103. Also, the thumbnail of the image 101 in the replacement button 109 can be updated to the thumbnail of the image 110.
[0030] Above the image 111, a video generation button 112 and a download button 113 are also displayed. When the user clicks the video generation button 112, the media processing client can generate a video, a gif, etc. based on the image 110 and the image 111. Specifically, the image 110 can be used as the first frame image, the image 111 can be used as the last frame image, multiple intermediate frame images are generated, and the image 110, the image 111, and the multiple intermediate frame images are combined to obtain the target video or the target gif. As Figure 1d shown, Figure 1d the multiple images shown in are some of the generated intermediate frame images. Among them, the arrows indicate the order of the intermediate frame images when generating the target video or the target gif. Finally, the user can click the download button 113 to download the generated image 111 and the video or gif generated based on the image 110 and the image 111.
[0031] It should be noted that the embodiment of FIG. 1 is described by taking the media processing client directly performing special effect processing on the original image and generating dynamic media as an example. In other embodiments, the media processing client can also transmit the original image, etc. to the media processing server deployed on the service platform through the network. The media processing server performs special effect processing based on the original image and generates dynamic media, and transmits the dynamic media to the media processing client through the network to provide the dynamic media to the user. For details, see Figure 2 .
[0032] Figure 2 is a schematic diagram of an exemplary system architecture for applying the embodiments of the present disclosure.
[0033] As Figure 2 shown, the system architecture 200 may include a terminal device 202, a network 203, and a server 204. It should be understood that Figure 2 the number or type of the terminal device, the network, and the server in are only illustrative. According to the implementation requirements, there can be any number or type of terminal devices, networks, and servers.
[0034] The network 203 is used to provide a medium for communication links between the terminal device and the server. The network 203 can include various connection types, such as wired, wireless communication links, or fiber optic cables, etc.
[0035] A media processing client is installed in the terminal device 202. The terminal device 202 can interact with the server through the network 203 to receive or send requests, information, etc. The terminal device 202 can be various electronic devices, including but not limited to smartphones, tablets, laptop portable computers, desktop computers, and smart wearable devices, etc.
[0036] A media processing server is deployed in the server 204. The server 204 can store, analyze, and process the received data, and can also send control commands or requests to the terminal device or other servers. The server can provide media processing services in response to the user's service request. It can be understood that one server can provide one or more services, and the same service can also be provided by multiple servers.
[0037] Based on Figure 2 The system architecture shown, in the embodiments of the present disclosure, the user 201 can input an original image through the terminal device 202. Then, the terminal device 202 can transmit the original image to the server 204 through the network 203. After receiving the original image, the server 204 can perform special effect processing on the original image to add decorative special effects with a preset style to the original image to obtain a target image. Then, the server 204 can return the target image to the terminal device 202 through the network 203 and display the target image to the user 201. In response to the trigger operation of the user 201 for the preset interface, the terminal device 202 can send an instruction to the server 204 through the network 203 to instruct the server 204 to generate a target dynamic media based on the original image and the target image. Finally, the server 204 can return the target dynamic media to the terminal device 202 through the network 203, so that the user 201 can view and save the target dynamic media through the terminal device 202.
[0038] The following will describe the present disclosure in detail with specific embodiments.
[0039] Figure 3 It is a flowchart of a media special effect processing method shown according to an exemplary embodiment. This method can be applied to the media processing client or the media processing server. In this embodiment, the media processing client is installed in the terminal device, and the terminal device can include but not limited to mobile terminal devices such as smartphones, smart wearable devices, tablets, laptop computers, and desktop computers, etc. The media processing server is deployed in the service platform, and the service platform can be implemented as any device, server, or device cluster with computing and processing capabilities. This method can include the following steps:
[0040] As Figure 3 shown, in step 301, in response to a preset trigger event for a preset interface, obtain an original image.
[0041] In this embodiment, the preset interface can be, for example, a specified home page interface, a specified pop-up window interface, or an image browsing interface, etc. This embodiment does not limit the specific category of the preset interface. The preset trigger event for the preset interface can be an event where the user opens the preset interface, or an event where the user triggers a preset interface in the preset interface (for example, clicks a specified button in the preset interface), or an event where the user browses a specified position in the preset interface, etc. This embodiment does not limit the specific setting of the preset trigger event.
[0042] Specifically, in one implementation, the preset interface is an image browsing interface, and the preset trigger event is to click a specified button in the preset interface. After detecting the preset trigger event for the preset interface, the image displayed in the current browsing interface can be obtained as the original image. For example, when the user browses the images in their own or their friends' photo albums, they can click the specified button in the browsing interface, and the media processing client or the media processing server can obtain the original image from the photo album browsed by the user.
[0043] In another implementation, the preset interface is a home page interface or a pop-up window interface, etc., and the preset trigger event is to click a specified button or browse to a specified position in the preset interface, etc. After detecting the preset trigger event for the preset interface, an image upload interface can be output to the user. The user can upload an image through the image upload interface, and the image uploaded by the user can be obtained as the original image.
[0044] It can be understood that the original image can also be obtained through other reasonable methods. This embodiment does not limit the specific method of obtaining the original image.
[0045] In step 302, special effects processing is performed on the original image.
[0046] In this embodiment, after obtaining the original image, special effects processing can be performed on the original image to add decorative special effects with a preset style to the original image. Among them, the preset style can be a style related to the festival atmosphere, for example, Spring Festival style, Christmas style, Valentine's Day style, etc. It can also be a style related to the current popular theme, for example, it can be the theme style of the currently popular movie, or the theme style of the current popular event, etc. It can also be a style related to the current season, etc., for example, spring style in spring, summer style in summer, etc. It can be understood that this embodiment does not limit the specific type of the preset style.
[0047] Among them, adding decorative special effects with a preset style to the original image can be adding some special decorations with a preset style to the main object in the original image. The main object in the original image can be a person, an animal, an object, or a building, etc. Taking the Valentine's Day style as an example, adding decorative special effects with the Valentine's Day style to the original image can be adding rose decorations to the main object in the original image. For example, adding a rose wreath to the head of the person in the original image, or adding a bouquet of roses to the hand of the person in the original image, etc.
[0048] Taking the Spring Festival style as another example, adding decorative special effects with the Spring Festival style to the original image can be adding Spring Festival element decorations to the main object in the original image. For example, adding Spring Festival element decorations to the clothing of the person in the original image, or adding gold ingots to the hand of the person in the original image.
[0049] In this embodiment, the original image can be processed with special effects in the following manner: First, the target text guiding information for guiding the special effect processing can be determined. Then, based on the target text guiding information, decorative special effects can be added to the original image. Specifically, in one implementation manner, a text guiding information can be randomly selected from the candidate text guiding information as the target text guiding information. In another implementation manner, the text guiding information specified by the user can also be obtained as the target text guiding information.
[0050] For example, after obtaining the original image, the original image can be processed with special effects at least once. Among them, when processing the original image with special effects for the first time, a text guiding information can be randomly selected from the candidate text guiding information as the target text guiding information. When processing the original image with special effects non-first time (i.e., reprocessing the original image), the text guiding information specified by the user can be obtained as the target text guiding information.
[0051] Specifically, a text guidance information can be randomly selected from the candidate text guidance information in the following manner as the target text guidance information: First, use an image recognition algorithm to detect and recognize the original image to determine the target object category corresponding to the target object included in the original image. Among them, the target object can be the main object in the original image, and the target object category corresponding to the target object can be one of multiple preset object categories. For example, the multiple object categories can include but are not limited to people, animals, buildings, objects, etc. Among them, one object category can correspond to at least one text guidance information, and different text guidance information corresponds to different decorative special effect effects. For example, the object category "person" can correspond to the following text guidance information: "The clothes are covered with rose decorations", "Wearing a rose wreath on the head", and "Holding a bunch of roses in the hand". Another example, the object category "animal" can correspond to the following text guidance information: "Rose bushes grow around the body", "Wearing a rose wreath on the head", and "The tail turns into a rose shape". Another example, the object category "building" can correspond to the following text guidance information: "Rose bushes grow around the building", "The roof is covered with roses", and "The road is covered with rose petals".
[0052] Next, a text guidance information can be randomly selected from at least one text guidance information corresponding to the target object category as the target text guidance information. For example, referring to the above examples, if the target object category is a person, a text guidance information can be randomly selected from "The clothes are covered with rose decorations", "Wearing a rose wreath on the head", and "Holding a bunch of roses in the hand" as the target text guidance information. Another example, if the target object category is an animal, a text guidance information can be randomly selected from "Rose bushes grow around the body", "Wearing a rose wreath on the head", and "The tail turns into a rose shape" as the target text guidance information. Another example, if the target object category is a building, a text guidance information can be randomly selected from "Rose bushes grow around the building", "The roof is covered with roses", and "The road is covered with rose petals" as the target text guidance information.
[0053] In this embodiment, in one implementation manner, the text guidance information specified by the user can be obtained in the following manner as the target text guidance information: First, the target object category corresponding to the target object included in the original image can be determined. Then, at least one text guidance information corresponding to the target object category can be determined as the candidate text guidance information. Output at least one candidate item, and each candidate item corresponds to a candidate text guidance information. Among them, the candidate item can be output in the form of text or in the form of an effect diagram. Finally, the text guidance information corresponding to the candidate item selected by the user from at least one candidate item is determined as the target text guidance information.
[0054] For example, referring to the above example, if the target object category is a person, then "the clothes are covered with rose decorations", "wearing a rose wreath on the head", and "holding a bunch of roses" can be used as candidate text guidance information. Output the alternatives corresponding to each of the above candidate text guidance information. For example, the alternative corresponding to "the clothes are covered with rose decorations" can be output in the form of its text content (i.e., the clothes are covered with rose decorations), or in the form of its example effect diagram (i.e., an image including a person with clothes covered with rose decorations). Another example, if the target object category is a building, then "rose bushes grow around the building", "the roof is covered with roses", and "the road is paved with rose petals" can be used as candidate text guidance information. Output the alternatives corresponding to each of the above candidate text guidance information. For example, the alternative corresponding to "the roof is covered with roses" can be the text content "the roof is covered with roses", or an example effect diagram of a building with the roof covered with roses.
[0055] In another implementation, the user-specified text guidance information can also be obtained in the following way as the target text guidance information: First, a preset custom interface can be displayed to the user, and the user's operations on the interface can be detected. When a trigger operation for the custom interface is detected, a text guidance information input box is provided to the user. The user can input user-defined text information through this input box. Finally, the text information input by the user in the text guidance information input box is obtained as the target text guidance information. Optionally, some available keyword label options can also be provided beside the input box so that the user can more conveniently input text information based on the keyword label options efficiently and accurately. Further optionally, the user-input custom text information can also be corrected to reduce the situation of special effect processing failure.
[0056] In this embodiment, after obtaining the target text guidance information, a decorative special effect can be added to the original image. Specifically, the original image and the target text guidance information can be input into a pre-trained adapter first. The adapter processes the original image and the target text guidance information to obtain target features, and the target features are input into a pre-trained large model, so that the large model generates a target image with a decorative special effect of a preset style added to the original image based on the target features.
[0057] In step 303, the target image after special effect processing and a preset interface for the target image are displayed, and in step 304, in response to a trigger operation for the preset interface, a target dynamic medium is generated based on the original image and the target image.
[0058] In this embodiment, after performing special effect processing on the original image, the target image after the special effect processing can be displayed in the special effect processing interface. The user can enter the browsing interface corresponding to the target image (such as clicking on the target image) to browse the target image. In the special effect processing interface, there may also be an interface for regenerating an image. The user can trigger this interface to perform special effect processing on the original image again.
[0059] In addition, in the special effect processing interface, a preset interface for the target image is also displayed. The media processing client can detect the user's operations in the special effect processing interface. When a trigger operation for the preset interface is detected, a target dynamic media can be generated based on the original image and the target image. The target dynamic media can be a video, a GIF, an animation, etc. This embodiment does not limit the specific form of the target dynamic media. Specifically, the original image can be used as the first frame image, and the target image can be used as the last frame image. Multiple intermediate images between the first frame image and the last frame image are generated. Then, the first frame image, the intermediate frame images, and the last frame image are combined into the target dynamic media. Thus, not only is special effect processing performed on the original image, but also the image after the special effect processing further has a dynamic effect. Among them, the intermediate frame images can be generated by frame interpolation or by using a network model. This embodiment does not limit the specific method for generating the intermediate frames.
[0060] A media special effect processing method provided by the present disclosure includes: in response to a preset trigger event for a preset interface, obtaining an original image, performing special effect processing on the original image to add a decorative special effect with a preset style to the original image, displaying the target image after the special effect processing and a preset interface for the target image, and in response to a trigger operation for the preset interface, generating a target dynamic media based on the original image and the target image. Thus, a target image with a decorative special effect of a preset style can be obtained based on the existing image, and a dynamic media that makes the target image have a dynamic effect is generated. There is no need for manual special effect processing of the image, which improves the efficiency of image special effect processing, improves the effect of special effect processing, expands the ways of special effect processing, and enhances the user experience.
[0061] It should be noted that although in the above embodiments, the operations of the method of the embodiments of the present disclosure are described in a specific order, this does not require or imply that these operations must be performed in that specific order, or that all the operations shown must be performed to achieve the desired result. On the contrary, the order of the steps depicted in the flowchart can be changed. Additionally or alternatively, some steps can be omitted, multiple steps can be combined into one step for execution, and / or one step can be decomposed into multiple steps for execution.
[0062] Corresponding to the foregoing embodiments of the media special effect processing method, the present disclosure also provides an embodiment of a media special effect processing device.
[0063] As Figure 4 shown, Figure 4 FIG. is a block diagram of a media special effect processing device according to an exemplary embodiment of the present disclosure. The device may include: an acquisition unit 401, a processing unit 402, a display unit 403, and a generation unit 404.
[0064] Among them, the acquisition unit 401 is configured to acquire an original image in response to a preset trigger event for a preset interface.
[0065] The processing unit 402 is configured to perform special effect processing on the original image to add a decorative special effect with a preset style to the original image.
[0066] The display unit 403 is configured to display the target image after special effect processing and a preset interface for the target image.
[0067] The generation unit 404 is configured to generate a target dynamic media based on the original image and the target image in response to a trigger operation for the preset interface.
[0068] In some embodiments, the acquisition unit 401 is configured to: output an image upload interface in response to a preset trigger event for a preset interface, and acquire the image uploaded through the image upload interface as the original image.
[0069] In other embodiments, the processing unit 402 includes: a determination sub-module and a special effect sub-unit (not shown in the figure).
[0070] Among them, the determination sub-unit is configured to determine target text guiding information for guiding special effect processing.
[0071] The special effect sub-unit is configured to add the decorative special effect to the original image based on the target text guiding information.
[0072] In other embodiments, the determination sub-unit is configured to: randomly select a text guiding information from the candidate text guiding information as the target text guiding information, or acquire the text guiding information specified by the user as the target text guiding information.
[0073] In other embodiments, after acquiring the original image, at least one special effect processing is performed on the original image. Among them, if the original image is subjected to special effect processing for the first time, a text guiding information is randomly selected from the candidate text guiding information as the target text guiding information. If the original image is not subjected to special effect processing for the first time, the text guiding information specified by the user is acquired as the target text guiding information.
[0074] In some other embodiments, the determining subunit may determine the target text guiding information for guiding special effect processing in the following manner: identifying the target object category corresponding to the target object included in the original image. Among them, one object category corresponds to at least one piece of text guiding information, and different pieces of text guiding information correspond to the effects of different decorative special effects. Randomly select one piece of text guiding information from at least one piece of text guiding information corresponding to the target object category as the target text guiding information.
[0075] In some other embodiments, the determining subunit may determine the target text guiding information for guiding special effect processing in the following manner: determining the target object category corresponding to the target object included in the original image. Among them, one object category corresponds to at least one piece of text guiding information, and different pieces of text guiding information correspond to the effects of different decorative special effects. Determine at least one piece of text guiding information corresponding to the target object category as the candidate text guiding information, and output at least one candidate. One candidate corresponds to one piece of candidate text guiding information. Determine the text guiding information corresponding to the candidate selected by the user from at least one candidate as the target text guiding information.
[0076] In some other embodiments, the determining subunit may determine the target text guiding information for guiding special effect processing in the following manner: displaying a preset custom interface, in response to a trigger operation on the custom interface, providing a text guiding information input box, and obtaining the text information input by the user in the text guiding information input box as the target text guiding information.
[0077] For the device embodiments, since they basically correspond to the method embodiments, the relevant parts can be referred to the partial descriptions of the method embodiments. The device embodiments described above are merely illustrative. The units described as separate components may or may not be physically separated, and the components shown as units may or may not be physical units, that is, they may be located in one place, or may be distributed to multiple network units. Some or all of the modules can be selected according to actual needs to achieve the purpose of the solution of the embodiments of the present disclosure. Those of ordinary skill in the art can understand and implement it without creative efforts.
[0078] Next, refer to Figure 5 , Figure 5A schematic block diagram of an electronic device provided by some embodiments of the present disclosure. The electronic device 920 is, for example, suitable for implementing the media special effect processing method provided by the embodiments of the present disclosure. The electronic device 920 may be a terminal device or the like, and may be used to implement a client or a server. The electronic device 920 may include, but is not limited to, mobile terminals such as mobile phones, laptop computers, digital broadcast receivers, PDAs (Personal Digital Assistants), PADs (Tablet Computers), PMPs (Portable Multimedia Players), in-vehicle terminals (such as in-vehicle navigation terminals), wearable electronic devices, etc., and fixed terminals such as digital TVs, desktop computers, smart home devices, etc. It should be noted that Figure 5 The illustrated electronic device 920 is merely an example and will not impose any limitations on the functions and usage scope of the embodiments of the present disclosure.
[0079] As Figure 5 shown, the electronic device 920 may include a processing device (such as a central processing unit, a graphics processing unit, etc.) 921, which may perform various appropriate actions and processes according to a program stored in a read-only memory (ROM) 922 or a program loaded from a storage device 928 into a random access memory (RAM) 923. In the RAM 923, various programs and data required for the operation of the electronic device 920 are also stored. The processing device 921, the ROM 922, and the RAM 923 are connected to each other through a bus 924. An input / output (I / O) interface 925 is also connected to the bus 924.
[0080] Generally, the following devices may be connected to the I / O interface 925: an input device 926 including, for example, a touch screen, a touchpad, a keyboard, a mouse, a camera, a microphone, an accelerometer, a gyroscope, etc.; an output device 927 including, for example, a liquid crystal display (LCD), a speaker, a vibrator, etc.; a storage device 928 including, for example, a magnetic tape, a hard disk, etc.; and a communication device 929. The communication device 929 may allow the electronic device 920 to communicate with other electronic devices wirelessly or wiredly to exchange data. Although Figure 5 the illustrated electronic device 920 has various devices, it should be understood that it is not required to implement or have all the illustrated devices, and the electronic device 920 may alternatively implement or have more or fewer devices. Figure 5 Each block shown in
[0081] According to an embodiment of the present disclosure, the above media special effect processing method can be implemented as a computer software program. For example, an embodiment of the present disclosure includes a computer program product, which includes a computer program carried on a non-transitory computer-readable medium, and the computer program includes program codes for executing the above media special effect processing method. In such an embodiment, the computer program can be downloaded and installed from a network through a communication device 929, or installed from a storage device 928, or installed from a ROM 922. When the computer program is executed by a processing device 921, the functions defined in the media special effect processing method provided by the embodiments of the present disclosure can be realized.
[0082] An embodiment of the present disclosure also provides a computer-readable storage medium, on which a computer program is stored. When the computer program is executed on a computer, the computer is made to execute the method provided by the present disclosure.
[0083] It should be noted that the computer-readable medium described in the embodiments of the present disclosure can be a computer-readable signal medium, a computer-readable storage medium, or any combination of the two. A computer-readable storage medium can be, for example, but not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination of the above. More specific examples of the computer-readable storage medium can include, but are not limited to: an electrical connection with one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the above. In the embodiments of the present disclosure, the computer-readable storage medium can be any tangible medium that contains or stores a program, and the program can be used by or combined with an instruction execution system, apparatus, or device. In the embodiments of the present disclosure, the computer-readable signal medium can include a data signal propagated in a baseband or as part of a carrier wave, which carries computer-readable program codes. Such a propagated data signal can take various forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination of the above. The computer-readable signal medium can also be any computer-readable medium other than the computer-readable storage medium, and the computer-readable signal medium can send, propagate, or transmit a program for use by or combined with an instruction execution system, apparatus, or device. The program codes contained on the computer-readable medium can be transmitted by any appropriate medium, including but not limited to: wires, optical cables, RF (Radio Frequency), etc., or any suitable combination of the above.
[0084] Computer program code for performing the operations of the embodiments of the present disclosure may be written in one or more programming languages or combinations thereof. The programming languages include object-oriented programming languages such as Java, Smalltalk, C++, and also include conventional procedural programming languages such as the "C" language or similar programming languages. The program code may be executed entirely on the user's computer, partially on the user's computer, executed as a stand-alone software package, partially on the user's computer and partially on a remote computer, or entirely on a remote computer or server. In the case of a remote computer, the remote computer may be connected to the user's computer through any type of network, including a local area network (LAN) or a wide area network (WAN), or may be connected to an external computer (e.g., connected through the Internet using an Internet service provider).
[0085] The various embodiments in the present disclosure are described in a progressive manner. For the same or similar parts among the various embodiments, reference may be made to each other. Each embodiment focuses on the differences from other embodiments. In particular, for the embodiments of the storage medium and the computing device, since they are basically similar to the method embodiments, the description is relatively simple, and the relevant parts may refer to the partial description of the method embodiments.
[0086] Those skilled in the art should be able to realize that in one or more of the above examples, the functions described in the embodiments of the present disclosure can be implemented by hardware, software, firmware, or any combination thereof. When implemented using software, these functions may be stored in a computer-readable medium or transmitted as one or more instructions or codes on a computer-readable medium.
[0087] The above-described specific implementation manners further elaborate on the objectives, technical solutions, and beneficial effects of the embodiments of the present disclosure. It should be understood that the above is only the specific implementation manners of the embodiments of the present disclosure and is not used to limit the protection scope of the present invention. Any modifications, equivalent replacements, improvements, etc. made on the basis of the technical solutions of the present disclosure shall be included in the protection scope of the present invention.
Claims
1. A method for processing media special effects, the method comprising: In response to a preset trigger event for a preset interface, acquiring an original image; Performing special effects processing on the original image to add decorative special effects with a preset style to the original image; Displaying the target image after the special effect processing and a preset interface for the target image; In response to a trigger operation on the preset interface, a target dynamic media is generated based on the original image and the target image.
2. The method according to claim 1, wherein: The acquiring of the original image in response to a preset triggering event for the preset interface includes: In response to a preset trigger event for a preset interface, outputting an image upload interface; The image uploaded through the image upload interface is obtained as the original image.
3. The method according to claim 1, wherein: The performing special effects processing on the original image comprises: Determining target text guiding information for guiding the special effect processing; Based on the target text guiding information, the decorative special effect is added to the original image.
4. The method according to claim 3, wherein: The determining of target text guiding information for guiding the special effect processing includes: Randomly select a text guide information from the candidate text guide information as the target text guide information; or Get the text guide information specified by the user as the target text guide information.
5. The method according to claim 4, wherein: After acquiring the original image, the original image is subjected to special effects processing at least once; wherein, if the special effects processing is performed on the original image for the first time, a text guide information is randomly selected from the text guide information to be selected as the target text guide information; if the special effects processing is not performed on the original image for the first time, the text guide information specified by the user is obtained as the target text guide information.
6. The method according to claim 3, wherein: The determining of target text guiding information for guiding the special effect processing includes: Identify a target object category corresponding to a target object included in the original image; wherein one object category corresponds to at least one text guide information; and different text guide information corresponds to different decorative special effects; A type of text guidance information is randomly selected from at least one type of text guidance information corresponding to the target object category as the target text guidance information.
7. The method according to claim 3, wherein: The determining of target text guiding information for guiding the special effect processing includes: Determine a target object category corresponding to a target object included in the original image; wherein one object category corresponds to at least one text guide information; and different text guide information corresponds to different decorative special effects; Determine at least one type of text guidance information corresponding to the target object category as the text guidance information to be selected; Output at least one alternative; one alternative corresponds to a type of text guide information to be selected; The text guiding information corresponding to the option selected by the user from the at least one option is determined as the target text guiding information.
8. The method according to claim 3, wherein: The determining of target text guiding information for guiding the special effect processing includes: Displays the preset custom interface; In response to a trigger operation on the custom interface, providing a text guidance information input box; The text information input by the user in the text guide information input box is obtained as the target text guide information.
9. A media special effects processing device, the device comprising: an acquisition unit, configured to acquire an original image in response to a preset trigger event for a preset interface; A processing unit configured to perform special effect processing on the original image to add a decorative special effect with a preset style to the original image; A display unit, configured to display the target image after the special effect processing and a preset interface for the target image; The generating unit is configured to generate target dynamic media based on the original image and the target image in response to a triggering operation on the preset interface.
10. A computer-readable storage medium having a computer program stored thereon, which, when executed in a computer, causes the computer to execute the method according to any one of claims 1 to 8.
11. An electronic device, comprising a memory and a processor, wherein the memory stores executable code, and when the processor executes the executable code, the method according to any one of claims 1 to 8 is implemented.
Citation Information
Patent Citations
Special effect processing method and device, electronic equipment and storage medium
CN117876208A
Video processing method and device, electronic equipment and storage medium
CN118018664A
Utilizing a diffusion prior neural network for text guided digital image editing
US20240362842A1
Techniques for generating a stylized media content item with a generative neural network
US20240412433A1
Cited By
Media special effect processing method and apparatus, device, and medium
WO2026157646A1