Image processing methods and apparatus, devices, storage media, and computer programs

The image processing method generates virtual avatars to replace real models in product photography, addressing the inefficiencies of traditional methods by providing cost-effective and time-saving product promotion.

JP2026517379APending Publication Date: 2026-05-29BEIJING YOUZHUJU NETWORK TECH CO LTD

Patent Information

Authority / Receiving Office
JP · JP
Patent Type
Applications
Current Assignee / Owner
BEIJING YOUZHUJU NETWORK TECH CO LTD
Filing Date
2024-08-20
Publication Date
2026-05-29

AI Technical Summary

Technical Problem

The high cost and time-consuming nature of photographing product effect diagrams using real models for e-commerce product promotion is a significant challenge.

Method used

An image processing method that identifies a target object in an image, generates a virtual avatar based on description information, and adds it to the image to create a processed result image, eliminating the need for real models.

Benefits of technology

Efficient and convenient generation of product effect diagrams using virtual avatars, reducing costs and time while maintaining effective product display.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 2026517379000001_ABST
    Figure 2026517379000001_ABST
Patent Text Reader

Abstract

This disclosure provides an image processing method and apparatus, device and storage medium, the method comprising: responding to a preset image processing request for an image to be processed, identifying a target object in an image to be processed; receiving target description information corresponding to the image to be processed; generating a target virtual avatar based on the target description information; and adding the target virtual avatar to the image to be processed to obtain a processed result image corresponding to the image to be processed. Embodiments of this disclosure generate a target virtual avatar for a target object based on target description information and process the image to be processed so that it becomes a processed result image that displays the target object based on the target virtual avatar.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] This application claims the priority of the Chinese patent with application number 202311055793.2 filed on August 21, 2023, and all the contents of the same application are incorporated herein by reference.

[0002] This disclosure relates to the field of data processing, and particularly to an image processing method, apparatus, device, and storage medium.

Background Art

[0003] With the development of e-commerce platforms, the number of people choosing online shopping is increasing. In order to display and promote products, merchants appeal the effects of products by photographing real models, for example, photographing product effect diagrams in which real models are wearing clothes and accessories to visually convey and promote products.

[0004] However, when photographing product effect diagrams with real models, the photographing cost is high and the cycle is long. Therefore, how to obtain more efficient and convenient effect diagrams for visually conveying products is an urgent problem to be solved currently.

Summary of the Invention

[0005] To solve the above technical problems, embodiments of this disclosure provide an image processing method.

[0006] In a first aspect, this disclosure provides an image processing method, and the method includes: identifying a target object in the image to be processed in response to a preset image processing requirement for the image to be processed; receiving target description information corresponding to the image to be processed, and generating a target virtual avatar based on the target description information; This includes adding the target virtual avatar to the image to be processed, and obtaining a processing result image corresponding to the image to be processed, wherein the target object in the image to be processed is displayed in the processing result image.

[0007] In one optional embodiment, the process further includes adding the target virtual avatar to the image to be processed and obtaining display area information in the image to be processed for the target before obtaining a processing result image corresponding to the image to be processed, Accordingly, adding the target virtual avatar to the image to be processed and obtaining a processing result image corresponding to the image to be processed is, This includes adding the target virtual avatar to the image to be processed based on the display area information of the target object in the image to be processed, and obtaining a processing result image corresponding to the image to be processed.

[0008] In one optional embodiment, based on the display area information of the target object in the image to be processed, the target virtual avatar is added to the image to be processed, and a processing result image corresponding to the image to be processed is obtained. Based on the display area information of the image to be processed for the target object, a portion of the image corresponding to the target object in the image to be processed is acquired. This includes performing a stitching process on the target virtual avatar and a portion of the images corresponding to the target object, and obtaining a processed result image corresponding to the images to be processed.

[0009] In one optional embodiment, obtaining display area information in the image to be processed for the target object is: The process includes performing a binarization process on the image to be processed based on the identification result for the target object, and obtaining a binary image corresponding to the image to be processed for marking the display area of ​​the target object in the image to be processed.

[0010] In one optional embodiment, based on the display area information of the target object in the image to be processed, the target virtual avatar is added to the image to be processed, and a processing result image corresponding to the image to be processed is obtained. The process includes performing a mask redraw on the image to be processed based on the binary image corresponding to the image to be processed and the target virtual avatar, and obtaining a processed result image corresponding to the image to be processed.

[0011] In one optional embodiment, the method is The process further includes inputting a binary image corresponding to the image to be processed, the processed result image, and the image to be processed into a pre-configured edge processing model, and then performing processing on the target edges in the processed result image using the pre-configured edge processing model to obtain an edge processing result image corresponding to the image to be processed.

[0012] In one optional embodiment, generating a target virtual avatar based on the target description information is: This includes inputting the target description information into a virtual avatar generation model, training the virtual avatar image sample having text description information, and then generating a target virtual avatar after processing by the virtual avatar generation model.

[0013] In one optional embodiment, before identifying the target object in the image to be processed, Perform image quality reduction processing on the image to be processed. and / or further includes performing sampling on pixel points in the length and width directions of the image to be processed.

[0014] In a second aspect, the present disclosure provides an image processing apparatus, the apparatus is An identification module for identifying a target object in an image to be processed in response to a pre-configured image processing request for an image to be processed, A generation module for receiving target description information corresponding to the image to be processed, and generating a target virtual avatar based on the target description information for describing the features of the target virtual avatar, The system includes an additional module for adding the target virtual avatar to the image to be processed and obtaining a processing result image corresponding to the image to be processed, for displaying the target object based on the target virtual avatar.

[0015] In a third aspect, the Disclosure provides a computer-readable storage medium in which instructions are stored, and when the instructions are executed by a terminal device, the terminal device enables the above-described method.

[0016] In a fourth aspect, the Disclosure provides an image processing device comprising a memory, a processor, and a computer program stored in the memory and executable by the processor, wherein the processor, when executing the computer program, realizes the above method.

[0017] In a fifth aspect, the Disclosure provides a computer program product which includes a computer program / instruction, and when the computer program / instruction is executed by a processor, the above method is realized. [Brief explanation of the drawing]

[0018] The drawings here are incorporated into the specification and constitute part of this specification, illustrating embodiments suitable for this disclosure, and are used together with the specification to interpret the principles of this disclosure.

[0019] To more clearly explain the technical solutions in the embodiments of the present disclosure or the prior art, the drawings necessary for describing the embodiments or the prior art are briefly described below. Obviously, those skilled in the art can obtain other drawings based on these drawings without creative efforts. [Figure 1] It is a flowchart of an image processing method according to an embodiment of the present disclosure. [Figure 2] It is a schematic diagram of an interaction interface for selecting an avatar feature tab according to an embodiment of the present disclosure. [Figure 3] It is a schematic diagram of a binary image according to an embodiment of the present disclosure. [Figure 4] It is a schematic diagram of an interaction interface according to an embodiment of the present disclosure. [Figure 5] It is a schematic diagram of another interaction interface according to an embodiment of the present disclosure. [Figure 6] It is a schematic diagram of data interaction according to an embodiment of the present disclosure. [Figure 7] It is a schematic diagram of the configuration of an image processing apparatus according to an embodiment of the present disclosure. [Figure 8] It is a schematic diagram of the configuration of an image processing device according to an embodiment of the present disclosure.

Modes for Carrying Out the Invention

[0020] To more clearly understand the above objects, features and advantages of the present disclosure, the technical solutions of the present disclosure are further described below. Unless there is a conflict, the embodiments of the present disclosure and the features in the embodiments can be combined with each other.

[0021] In the following description, many specific details are set forth to make the present disclosure thoroughly understandable. However, the present disclosure may be implemented in other forms different from those described herein. Obviously, the embodiments in the specification are only some embodiments of the present disclosure, not all embodiments.

[0022] In the field of e-commerce, there is a need to launch new products one after another, and to facilitate product display and promotion, products are often visually conveyed by photographing real people wearing the products in product effect images. However, this method is expensive and time-consuming, and there is a need for sellers and users to obtain product effect images efficiently and conveniently.

[0023] To this end, the embodiments of the present disclosure provide an image processing method which first identifies a target object in an image to be processed in response to a pre-set image processing request for an image to be processed, receives target description information corresponding to the image to be processed, generates a target virtual avatar based on the target description information for describing the features of the target virtual avatar, adds the target virtual avatar to the image to be processed, and obtains a processed result image corresponding to the image to be processed for displaying the target object based on the target virtual avatar. The embodiments of the present disclosure generate a target virtual avatar for the target object based on the target description information and process the image to be processed so that it becomes a processed result image that displays the target object based on the target virtual avatar. As can be seen from this, the embodiments of the present disclosure do not require photography using a real person model and generate effect diagrams for displaying the target object more efficiently and conveniently based on the image processing method, thereby meeting user needs.

[0024] Based on this, embodiments of the present disclosure provide an image processing method, which, referring to Figure 1, is a flowchart of the image processing method according to an embodiment of the present disclosure, and the method specifically includes the following:

[0025] In S101, in response to a pre-configured image processing request for the image to be processed, the target object in the image to be processed is identified.

[0026] Here, the image to be processed may be one that has been selected from a user album and uploaded, or one that has been obtained by taking a photograph and uploaded, and is not limited to the embodiments of this disclosure. The image to be processed displays a target object, which may be a displayable item, such as clothing, accessories, a backpack, or glasses.

[0027] The image processing method according to the embodiments of this disclosure can be used on the server side.

[0028] In embodiments of this disclosure, when a pre-configured image processing request for an image to be processed is received from the client side, the first step is to identify the target object in the image to be processed. Here, the pre-configured image processing request is used to request the server side to perform a pre-configured processing on the target object in the image to be processed.

[0029] In S102, target description information corresponding to the image to be processed is received, and a target virtual avatar is generated based on the target description information.

[0030] In the embodiments of this disclosure, the target description information is used to describe the features of the target virtual avatar and may include appearance feature description information, facial expression description information, etc. for the target virtual avatar. Furthermore, the target description information may include scene description information, environment description information, and scene feature information such as indoors and outdoors.

[0031] In one optional embodiment, the target description information may be user-customized feature information, specifically, the user can customize it by selecting an avatar features tab in the interaction interface.

[0032] As shown in Figure 2, it is an interaction interface for selecting avatar feature tabs according to an embodiment of the present disclosure. Specifically, the interaction interface displays multiple selectable facial feature tabs and multiple selectable scene feature tabs for the virtual avatar. The user can customize the avatar feature tabs of the target virtual avatar by clicking on the tabs in the interaction interface, and the user can customize different styles of virtual avatars according to actual needs, which can be used to enhance the display effect on the target object.

[0033] As shown in Figure 2, the avatar features tab currently selected by the user includes short hair, smiling, indoors, etc., and accordingly, the target description information corresponding to the image to be processed also includes short hair, smiling, indoors, etc.

[0034] In the embodiments of this disclosure, the client side receives an avatar feature tab selected by the user, and then sends the descriptive information corresponding to the selected avatar feature tab to the server side as target descriptive information for the image to be processed. The server side, after receiving the target descriptive information corresponding to the image to be processed, generates a target virtual avatar based on the target descriptive information. Here, the target virtual avatar may be generated by an artificial intelligence model and have a face and body that closely resembles a real person.

[0035] In one optional embodiment, to generate a target virtual avatar based on target description information, an artificial intelligence model can be used to generate a target virtual avatar that matches the description features of the target description information. Here, the artificial intelligence model may be a trained virtual avatar generation model, and a virtual avatar generation model is obtained by training it using virtual avatar image samples that have text description information.

[0036] In the application process of the virtual avatar generation model, target description information is input into the virtual avatar generation model, and after processing by the virtual avatar generation model, the target virtual avatar is generated.

[0037] In S103, the target virtual avatar is added to the image to be processed, and a processing result image corresponding to the image to be processed is obtained, in which the target object in the image to be processed is displayed, for displaying the target object based on the target virtual avatar.

[0038] In the embodiments of this disclosure, it is necessary to first obtain display area information in the image to be processed for the target object, then add the target virtual avatar to the image to be processed, and finally obtain a processed result image corresponding to the image to be processed.

[0039] Here, the display area information of the target image to be processed is used to mark the display position, display range, etc., of the target image to be processed.

[0040] In one optional embodiment, the display area information in the image to be processed for the target object may include coordinate information of the region occupied by the image to be processed for the target object, for example, coordinate information of edge pixel points in the region occupied by the image to be processed for the target object.

[0041] In another optional embodiment, the display position information of the target object in the image to be processed may be obtained by segmenting the target object and other content (also called background content) using an image segmentation model, generating a binary image (e.g., a grayscale image) using a binarization process, and using this information to mark the display area information of the target object in the image to be processed.

[0042] In the embodiments of this disclosure, based on the identification result for the target object, a binarization process is performed on the image to be processed to obtain a binary image corresponding to the image to be processed, and the binary image may be used to mark the display area of ​​the target object in the image to be processed.

[0043] For example, in a binary image, the pixel points in the region occupied by the target object can be set to pixel values ​​corresponding to black, and the pixel points in the image other than the target object can be set to pixel values ​​corresponding to white. Specifically, the embodiments of this disclosure do not limit the specific pixel values ​​set for the pixel points in the binary image.

[0044] As shown in Figure 3, it is a schematic diagram of a binary image according to an embodiment of the present disclosure, specifically a black and white grayscale image, where the white area is used to mark the display area in the image to be processed.

[0045] Based on the above embodiment, the target virtual avatar is added to the image to be processed based on the display area information of the target image to be processed, and a processed result image corresponding to the image to be processed is obtained.

[0046] In one optional embodiment, based on display area information in the image to be processed for the target object, a portion of the image corresponding to the target object is obtained from the image to be processed, and stitching is performed on the target virtual avatar and this portion of the image to obtain a processed result image corresponding to the image to be processed.

[0047] Here, the partial images corresponding to the target object are the smallest images in the image to be processed that include the target object. For example, if the image to be processed displays a model wearing a long skirt (the target object) and a background image (e.g., a live room scene), and the smallest image in the image to be processed that includes the target object is the target object itself, then the partial images corresponding to the target object refer to the images of the long skirt display area. Specifically, any image processing technique can be used to selectively perform stitching on the target virtual avatar and the partial images corresponding to the target object, and the embodiments of this disclosure are not limited thereto.

[0048] In another optional embodiment, a mask redraw is performed on the image to be processed based on a binary image corresponding to the image to be processed and a target virtual avatar, thereby obtaining a processed result image corresponding to the image to be processed. Here, the embodiment of the present disclosure uses a binary image as a mask and draws on the target virtual avatar within the area of ​​the binary image other than the display area of ​​the marked target object.

[0049] In the embodiments of this disclosure, in response to a pre-set image processing request for an image to be processed, the system identifies a target object in the image to be processed, receives target description information corresponding to the image to be processed, generates a target virtual avatar based on the target description information for describing the features of the target virtual avatar, adds the target virtual avatar to the image to be processed, and obtains a processed result image corresponding to the image to be processed for displaying the target object based on the target virtual avatar. The embodiments of this disclosure generate a target virtual avatar for the target object based on the target description information, and process the image to be processed so that it becomes a processed result image that displays the target object based on the target virtual avatar. As can be seen, the embodiments of this disclosure do not require photography using a real person model, and based on the image processing method, they generate effect diagrams for displaying the target object more efficiently and conveniently, thereby meeting user needs.

[0050] In actual applications, after stitching is performed on a target virtual avatar and some images corresponding to the target object, the stitching edges of the resulting processed image may exhibit undesirable edge fusion effects. Therefore, in order to further improve the quality of the processed image, the embodiments of this disclosure may also perform edge processing on the stitching edges of the processed image corresponding to the image to be processed.

[0051] Specifically, based on the display area information of the image to be processed for the target object, edge identification is performed on the target object in the image to be processed, and the edge identification result is obtained. Then, based on the edge identification result, edge processing is performed on the processed result image corresponding to the image to be processed, making the stitching edge effect between the target virtual avatar and the target object in the edge-processed processed result image more realistic.

[0052] In one optional embodiment, a binary image corresponding to the image to be processed, a processed result image, and the original image of the image to be processed are input to a pre-configured edge processing model. After processing by the pre-configured edge processing model, an edge-processed result image corresponding to the image to be processed is obtained. Specifically, the pre-configured edge processing model processes the edges of the target object in the processed result image, with the aim of making the stitching edge effect between the target virtual avatar and the target object more realistic.

[0053] In another optional embodiment, a neural network model can perform edge detection and identification on the target object in the image to be processed based on the display area information of the target object in the image to be processed, obtain the edge identification result, and then perform edge processing on the processing result diagram corresponding to the image to be processed based on the edge identification result. The specific type of neural network model is not limited.

[0054] In the embodiments of this disclosure, edge processing is performed on the processed result image corresponding to the image to be processed, thereby making the stitching edges between the target object and the target virtual avatar in the processed result image softer and smoother, bringing it closer to a real-life image and improving the generated image quality.

[0055] In practical applications, to improve image processing efficiency, image processing efficiency can be increased by reducing image quality through processes such as image quality reduction processing on the image to be processed and sampling the number of pixels in the length and width directions of the image. In the embodiment of this disclosure, when a pre-set image processing request for an image to be processed is received, image quality reduction processing is performed on the image to be processed, and / or sampling processing is performed on the pixels in the length and width directions of the image to be processed, thereby appropriately reducing the image quality of the image to be processed, reducing the computational load of image processing on the server side, increasing the image generation speed, and improving image processing efficiency.

[0056] To facilitate understanding of the above embodiment, the embodiment of this disclosure will be presented from the client-side perspective. As shown in Figure 4, it is a schematic diagram of the interaction interface according to the embodiment of this disclosure. First, the client-side receives target description information determined by the user based on the interface shown in Figure 2, and uploads it to the server-side. The user also uploads an image to be processed to the server-side via the interaction interface (let's say it's an image of a real person wearing clothes). After that, the interaction interface displays an interface state as shown in Figure 4, indicating prompt information, such as "Model generating," while waiting for the server-side to process the image.

[0057] The server-side, after receiving a pre-configured image processing request from the client-side, performs image quality reduction processing on the received image to be processed, identifies the target object, and performs segmentation processing on the target object (clothing) and background content in the image to be processed based on the identification result. For example, it performs binarization processing on the image to be processed to obtain a single black and white grayscale image, which is used to mark the display area information of the target object (clothing) in the image to be processed.

[0058] The server-side uses a black and white grayscale image as a mask, renders the target virtual avatar based on the target description information, and obtains a processing result image corresponding to the image to be processed; in other words, the processing result image of the target is displayed based on the target virtual avatar.

[0059] The server-side completes processing on the image to be processed and then sends the generated processing result image back to the client-side. The client-side, upon receiving the processing result image, displays it in an interaction interface, as shown in Figure 5, which is a schematic diagram of another interaction interface according to an embodiment of the present disclosure, and the interaction interface displays an image that shows the target object (clothing) based on the target virtual avatar (model).

[0060] When the server-side performs stitching on the target virtual avatar and target object based on the display area information of the image to be processed for the target object, the resulting processed image may have undesirable edge processing effects. To smooth the edges of the processed image, the server-side can further perform edge processing on the result and transmit it to the client-side. The client-side then displays the processed image on the interaction interface for user viewing.

[0061] To help users further understand the image processing method according to the embodiments of this disclosure, the embodiments of this disclosure provide a specific application method, which, as shown in Figure 6, is a data interaction diagram according to the embodiments of this disclosure, and specifically, the image processing method according to the embodiments of this disclosure can be realized based on data interaction between the client side and the server side. The server side is home to a trained first model and a second model.

[0062] First, the user selects an avatar features tab in the client-side interaction interface. The client then forms target description information based on the avatar features tab customized and selected by the user, and sends this target description information to the server. The user then uploads an image to be processed by the client (let's say an image of a real person wearing clothes) and sends a pre-configured image processing request for the image to be processed to the server.

[0063] After receiving a pre-configured image processing request from the client side, the server-side uses a first model to identify the target object (clothing) in the image to be processed. Based on the identification result for the target object, it performs a binarization process on the image to be processed to obtain a binary image corresponding to the image to be processed, which is used as a mask. To allow the user to feel the progress of the image processing and enhance the human-machine interaction experience, the server-side can first process the image based on the binary image corresponding to the image to be processed, and transmit an image showing only the target object to the client side. After receiving this image, the client side displays it.

[0064] To avoid the problem of slow execution speed due to the enormous amount of processing logic, and the possibility that the entire processing process may stop if one of the processing logics fails, interfaces may be provided separately for different logical units in the application process. For example, a separate interface may be provided to perform edge processing on the image to be processed. The embodiments of this disclosure are not limited to the arrangement of the server-side implementation architecture in actual applications.

[0065] As shown in Figure 6 above, the client side resends the image to be processed and the corresponding binary image to the server side via a single interface. The server side, after receiving the image to be processed and the corresponding binary image, performs edge processing on the image to be processed using a second model based on the corresponding binary image and the target virtual avatar, obtains a processed result image corresponding to the image to be processed, and finally sends the image to be processed back to the client side for display on the client side.

[0066] Based on embodiments of the above method, embodiments of the present disclosure further provide an image processing apparatus, referring to Figure 7, which is a schematic diagram of the configuration of an image processing apparatus according to embodiments of the present disclosure, and the apparatus is An identification module 701 for identifying a target object in an image to be processed in response to a pre-configured image processing request for an image to be processed, A generation module 702 receives target description information corresponding to the image to be processed, and generates a target virtual avatar based on the target description information for describing the features of the target virtual avatar, The system includes an additional module 703 for adding the target virtual avatar to the image to be processed and obtaining a processing result image corresponding to the image to be processed, for displaying the target object based on the target virtual avatar.

[0067] In one optional embodiment, the apparatus is The system further includes a first acquisition module that acquires display area information in the image to be processed of the target object, Accordingly, the additional module specifically, This method is used to add the target virtual avatar to the image to be processed based on the display area information of the target object in the image to be processed, and to obtain a processed result image corresponding to the image to be processed.

[0068] In one optional embodiment, the additional module is: A first acquisition submodule for acquiring a portion of the image corresponding to the target object in the image to be processed, based on the display area information of the target object in the image to be processed, It includes a first stitching submodule for performing stitching on the target virtual avatar and a portion of images corresponding to the target object, and obtaining a processed result image corresponding to the images to be processed.

[0069] In one optional embodiment, the first acquisition module specifically, Based on the identification result for the target object, a binarization process is performed on the image to be processed, and this is used to obtain a binary image corresponding to the image to be processed for marking the display area of ​​the target object in the image to be processed.

[0070] In one optional embodiment, the additional module is specifically: This method is used to obtain a processed result image corresponding to the image to be processed by performing a mask redraw on the image to be processed based on the binary image corresponding to the image to be processed and the target virtual avatar.

[0071] In one optional embodiment, the apparatus is The system further includes an edge processing module that inputs a binary image corresponding to the image to be processed, the processed result image, and the image to be processed into a pre-configured edge processing model, performs processing on the target edges in the processed result image using the pre-configured edge processing model, and then obtains an edge processing result image corresponding to the image to be processed.

[0072] In one optional embodiment, the generation module specifically, The target description information is input into a virtual avatar generation model, and after processing by the virtual avatar generation model obtained by training based on virtual avatar image samples having text description information, it is used to generate a target virtual avatar.

[0073] In one optional embodiment, the apparatus is An image processing module for performing image quality reduction processing on the image to be processed, and / or further includes an image sampling module for performing sampling on pixel points in the length and width directions of the image to be processed.

[0074] In the image processing method according to the embodiment of this disclosure, first, in response to a pre-set image processing request for an image to be processed, a target object in the image to be processed is identified, target description information corresponding to the image to be processed is received, a target virtual avatar is generated based on the target description information for describing the features of the target virtual avatar, the target virtual avatar is added to the image to be processed, and a processing result image corresponding to the image to be processed is obtained for displaying the target object based on the target virtual avatar. The embodiment of this disclosure generates a target virtual avatar for the target object based on the target description information, processes the image to be processed so that it becomes a processing result image that displays the target object based on the target virtual avatar, and as can be seen, the embodiment of this disclosure does not require photography using a real person model, and generates effect diagrams for displaying the target object more efficiently and conveniently based on the image processing method, thereby satisfying user needs.

[0075] In addition to the above-described method and apparatus, embodiments of the present disclosure further provide a computer-readable storage medium on which instructions are stored, and when the instructions are executed by a terminal device, the terminal device implements the image processing method described in embodiments of the present disclosure.

[0076] Embodiments of the present disclosure further provide a computer program product which includes a computer program / instruction, and when the computer program / instruction is executed by a processor, it realizes the image processing method described in the embodiment of the present disclosure.

[0077] Furthermore, embodiments of this disclosure further provide an image processing device, as shown in Figure 8, The image processing device may include a processor 801, a memory 802, an input device 803, and an output device 808. The number of processors 801 in the image processing device may be one or more, and in Figure 8, one processor is used as an example. In some embodiments of this disclosure, the processor 801, memory 802, input device 803, and output device 808 may be connected by a bus or by other means, where in Figure 8, a bus connection is used as an example.

[0078] Memory 802 may be used to store software programs and modules, and the processor 801 executes various functional applications and data processing of the image processing equipment by executing the software programs and modules stored in memory 802. Memory 802 may mainly include a program storage area and a data storage area, where the program storage area may store an operating system, an application program necessary for at least one function, etc. Memory 802 may also include high-speed random access memory, and may further include non-volatile memory, such as at least one magnetic disk memory device, a flash memory device, or other non-volatile solid-state memory devices. Input device 803 may be used to receive input numerical or character information and generate signal inputs related to user settings and function control of the image processing equipment.

[0079] Specifically, in this embodiment, the processor 801 loads executable files corresponding to the processes of one or more application programs into memory 802 according to the following instructions, and then executes the application programs stored in memory 802 by the processor 801, thereby realizing various functions of the image processing device.

[0080] In this specification, relational terms such as “first” and “second” are merely used to distinguish one entity or operation from another, and do not necessarily require or suggest that any such actual relationship or order exists between these entities or operations. Furthermore, the terms “include,” “incorporate,” or any other variations thereof are intended to cover non-exclusive inclusion, thereby meaning that a process, method, article, or apparatus that includes a set of elements includes not only those elements but also other elements not explicitly listed, or elements specific to such a process, method, article, or apparatus. Unless further limited, the elements limited by the phrase “include one…” do not preclude the existence of other identical elements in a process, method, article, or apparatus that includes such elements.

[0081] The above description is merely a set of specific embodiments of the Disclosure intended to enable those skilled in the art to understand or implement the Disclosure. Various modifications to these embodiments will be obvious to those skilled in the art, and the general principles defined herein can also be implemented in other embodiments without departing from the spirit or scope of the Disclosure. Therefore, the Disclosure is not limited to these embodiments described herein, but rather conforms to the broadest extent to which the principles and novel features disclosed herein are consistent.

Claims

1. An image processing method, In response to a pre-configured image processing request for an image to be processed, the system identifies the target object in the image to be processed. The process involves receiving target description information corresponding to the image to be processed, and generating a target virtual avatar based on the target description information for describing the features of the target virtual avatar. An image processing method comprising adding the target virtual avatar to the image to be processed, and obtaining a processing result image corresponding to the image to be processed for displaying the target object based on the target virtual avatar.

2. Before adding the target virtual avatar to the image to be processed and obtaining the processing result image corresponding to the image to be processed, The process further includes obtaining display area information in the image to be processed of the target object, Accordingly, adding the target virtual avatar to the image to be processed and obtaining a processing result image corresponding to the image to be processed is, The method according to claim 1, comprising adding the target virtual avatar to the image to be processed based on the display area information of the target object in the image to be processed, and obtaining a processing result image corresponding to the image to be processed.

3. Based on the display area information of the target object in the image to be processed, adding the target virtual avatar to the image to be processed and obtaining a processing result image corresponding to the image to be processed is: Based on the display area information of the image to be processed for the target object, a portion of the image corresponding to the target object in the image to be processed is acquired. The method according to claim 2, further comprising performing a stitching process on a portion of images corresponding to the target virtual avatar and the target object, and obtaining a processed result image corresponding to the images to be processed.

4. Obtaining display area information in the image to be processed of the target object is: The method according to claim 2, further comprising performing a binarization process on the image to be processed based on the identification result for the target object, and obtaining a binary image corresponding to the image to be processed for marking the display area of ​​the target object in the image to be processed.

5. Based on the display area information of the target object in the image to be processed, adding the target virtual avatar to the image to be processed and obtaining a processing result image corresponding to the image to be processed is: The method according to claim 4, comprising performing a mask redraw on the image to be processed based on the binary image corresponding to the image to be processed and the target virtual avatar, and obtaining a processed result image corresponding to the image to be processed.

6. The method according to claim 4, further comprising inputting a binary image corresponding to the image to be processed, the processed result image, and the image to be processed into a pre-configured edge processing model, and after processing the target edges in the processed result image by the pre-configured edge processing model, obtaining an edge processing result image corresponding to the image to be processed.

7. Generating a target virtual avatar based on the aforementioned target description information is: The method according to claim 1, further comprising inputting the target description information into a virtual avatar generation model, and generating a target virtual avatar after processing by the virtual avatar generation model obtained by training based on virtual avatar image samples having text description information.

8. Before identifying the target object in the image to be processed, Perform image quality reduction processing on the image to be processed. The method according to claim 1, further comprising and / or performing a sampling operation on pixel points in the length direction and width direction in the image to be processed.

9. An image processing device, An identification module for identifying a target object in an image to be processed in response to a pre-configured image processing request for an image to be processed, A generation module for receiving target description information corresponding to the image to be processed, and generating a target virtual avatar based on the target description information for describing the features of the target virtual avatar, An image processing apparatus comprising an additional module for adding the target virtual avatar to the image to be processed and obtaining a processing result image corresponding to the image to be processed for displaying the target object based on the target virtual avatar.

10. A computer-readable storage medium wherein a command is stored in the computer-readable storage medium, and when the command is executed by a terminal device, the terminal device is made to implement the method according to any one of claims 1 to 8.

11. An image processing device comprising memory, a processor, and a computer program stored in the memory and executable by the processor, wherein the processor, when executing the computer program, realizes the method according to any one of claims 1 to 8.