Method and device for generating media content, equipment and storage medium

CN121753072APending Publication Date: 2026-03-27BEIJING ZITIAO NETWORK TECH CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2024-07-26
Publication Date
2026-03-27

AI Technical Summary

Technical Problem

Traditional media platforms cannot meet users' needs for generating media content related to virtual objects. Users expect to generate high-quality media content based on virtual objects and clothing description text.

Method used

A method and apparatus for generating media content are provided. By presenting a generation interface associated with a virtual object, clothing description text is obtained, and target media content is generated based on the image data of the virtual object and the clothing description text, including a first visual element corresponding to the virtual object and a second visual element corresponding to the clothing description text.

Benefits of technology

It enables the generation of media content associated with virtual objects based on user-configured operations, thereby improving the user's interactive experience and the quality of media content.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN121753072A_ABST
    Figure CN121753072A_ABST
Patent Text Reader

Abstract

The embodiment of the invention relates to a media content generation method and device, equipment and a storage medium. The method provided by the invention comprises the following steps: presenting a generation interface associated with a virtual object, wherein the virtual object is created based on a configuration operation of a user; obtaining a clothing description text through the generation interface; and presenting a target media content, the target media content being generated based on the image data of the virtual object and the clothing description text, the target media content including a first visual element corresponding to the virtual object and a second visual element corresponding to the clothing description text. In this way, according to the embodiment of the invention, the media content associated with the clothing of the virtual object can be generated based on the virtual object created by the user according to the media generation request of the user. Therefore, the requirement of a user for generating media content based on the virtual object can be met, and the interaction experience of the user is improved.
Need to check novelty before this filing date? Find Prior Art

Description

Method, device, equipment and storage medium for generating media content TECHNICAL FIELD

[0001] Example embodiments of the present disclosure generally relate to the field of computers, and in particular, to a method, device, equipment and computer readable storage medium for generating media content. BACKGROUND

[0002] With the development of computer technology, the Internet has become an important platform for people to interact with information. For example, people use media platforms to view various media works and create virtual avatars. People expect to generate corresponding media content using virtual avatars.

[0003] SUMMARY

[0004] In a first aspect of the present disclosure, a method for generating media content is provided. The method comprises: presenting a generation interface associated with a virtual object, the virtual object being created based on configuration operations of a user; obtaining, via the generation interface, a clothing description text; and presenting target media content, the target media content being generated based on image data of the virtual object and the clothing description text, the target media content comprising a first visual element corresponding to the virtual object and a second visual element corresponding to the clothing description text.

[0005] In a second aspect of the present disclosure, a device for generating media content is provided. The device comprises: an interface presentation module configured to present a generation interface associated with a virtual object, the virtual object being created based on configuration operations of a user; a text obtaining module configured to obtain, via the generation interface, a clothing description text; and a content presentation module configured to present target media content, the target media content being generated based on image data of the virtual object and the clothing description text, the target media content comprising a first visual element corresponding to the virtual object and a second visual element corresponding to the clothing description text.

[0006] In a third aspect of the present disclosure, an electronic device is provided. The device comprises at least one processing unit; and at least one memory coupled to the at least one processing unit and storing instructions for execution by the at least one processing unit. The instructions, when executed by the at least one processing unit, cause the device to perform the method of the first aspect.

[0007] In a fourth aspect of the present disclosure, a computer readable storage medium is provided. The computer readable storage medium has stored thereon a computer program, the computer program being executable by a processor to implement the method of the first aspect.

[0008] It is to be understood that the content described in this Background section is not to be taken as an acknowledgement that this content is prior art to the present disclosure relative to any present or future application. The disclosure of other documents and art referred to in this Background section is not an admission that any of these documents or art is "prior art" to the present disclosure, with respect to any present or future application. BRIEF DESCRIPTION OF DRAWINGS

[0009] The above and other features, advantages and aspects of embodiments of the present disclosure will become more apparent by describing in detail some embodiments thereof with reference to the annexed drawings in which:

[0010] FIG. 1 shows a schematic diagram of an example environment in which embodiments according to the present disclosure can be implemented;

[0011] FIGS. 2A to 2H show example interfaces according to some embodiments of the present disclosure;

[0012] FIG. 3 shows a flowchart of an example process of generating media content according to some embodiments of the present disclosure;

[0013] FIG. 4 shows a schematic block diagram of an example media content generating apparatus according to some embodiments of the present disclosure; and

[0014] FIG. 5 shows a block diagram of an electronic device capable of implementing various embodiments of the present disclosure. DETAILED DESCRIPTION

[0015] Embodiments of the present disclosure will be described hereinafter with reference to the accompanying drawings. While certain embodiments of the present disclosure are shown in the drawings, it is understood that the present disclosure can be embodied in various forms and should not be construed as being limited to the embodiments set forth herein; rather, these embodiments are provided so that the present disclosure will be more thoroughly and completely understood. It is to be understood that the drawings and embodiments of the present disclosure are only for illustrative purposes and should not be construed as limiting the scope of protection of the present disclosure.

[0016] It is noted that the headings provided herein are not limitations of the embodiments described in the sections / sub-sections. Various embodiments are described throughout this document and any type of embodiment can be included under any section / sub-section. Further, embodiments described in any section / sub-section can be combined with any other embodiment described in the same section / sub-section and / or in a different section / sub-section in any manner.

[0017] In the description of the embodiments of the disclosure, the term "comprising" and similar terms thereof are understood to encompass open-ended inclusion, i.e., "comprising but not limited to". The term "based on" is understood to mean "based at least in part on". The term "one embodiment" or "the embodiment" is understood to mean "at least one embodiment". The term "some embodiments" is understood to mean "at least some embodiments". The following can also include other explicit and implicit definitions. The terms "first", "second", etc. can refer to different or same objects. The following can also include other explicit and implicit definitions.

[0018] In the embodiments of the disclosure, data of users, acquisition and / or use of data, etc. can be involved. These aspects all comply with corresponding laws and regulations and relevant provisions. In the embodiments of the disclosure, all collection, acquisition, processing, processing, forwarding, use, etc. of data are performed on the premise that users are aware of and confirm. Accordingly, in the implementation of the embodiments of the disclosure, the type, use range, use scenario, etc. of the data or information that can be involved should be informed to the user and the authorization of the user should be obtained through appropriate means according to relevant laws and regulations. The specific informing and / or authorization manner can vary according to actual conditions and application scenarios, and the scope of the disclosure is not limited in this aspect.

[0019] In the embodiments of the disclosure, if the scheme involves processing of personal information, the processing is performed on the premise of having a legal basis (for example, obtaining the consent of the subject of personal information, or being necessary for the performance of a contract, etc.) and is performed only within the prescribed or agreed range. If a user refuses to process personal information other than the necessary information required for basic functions, the user will not be affected in using the basic functions.

[0020] As mentioned above, with the development of the Internet, more and more people perform network activities in network platforms. For example, people use media platforms to view various media works and create virtual avatars. People expect to generate corresponding media content by using virtual avatars. However, the traditional media platforms cannot meet the needs of users.

[0021] Embodiments of the disclosure provide a scheme for generating media content. According to the scheme, a generation interface associated with a virtual object can be presented, the virtual object being created based on a configuration operation of a user; via the generation interface, a clothing description text is acquired; and target media content is presented, the target media content being generated based on image data of the virtual object and the clothing description text, the target media content including a first visual element corresponding to the virtual object and a second visual element corresponding to the clothing description text.

[0022] In this way, embodiments of the present disclosure are able to generate, for a user's media generation request, media content associated with the virtual object's costume based on the virtual object created by the user. Thus, the user's demand for generating media content based on the virtual object can be met, and the user's interactive experience can be improved.

[0023] Various example implementations of the solution are described in further detail below in conjunction with the accompanying drawings.

[0024] Example Environment

[0025] FIG. 1 illustrates a schematic diagram of an example environment 100 in which embodiments of the present disclosure can be implemented. As shown in FIG. 1, the example environment 100 can include an electronic device 110.

[0026] In this example environment 100, the electronic device 110 can run an application 120 that supports interface interaction. The application 120 can be any suitable type of application for interface interaction, examples of which can include, but are not limited to, a video application, a social application, or other suitable application. A user 140 can interact with the application 120 via the electronic device 110 and / or its attached devices.

[0027] In the environment 100 of FIG. 1, the electronic device 110 can present, through the application 120, an interface 150 for supporting interface interaction if the application 120 is in an active state.

[0028] In some embodiments, the electronic device 110 communicates with a server 130 to implement the provision of services to the application 120. The electronic device 110 can be any type of mobile terminal, fixed terminal, or portable terminal including a mobile handset, a desktop computer, a laptop computer, a notebook computer, a netbook computer, a tablet computer, a media computer, a multimedia tablet, a personal communication system (PCS) device, a personal navigation device, a personal digital assistant (PDA), an audio / video player, a digital camera / camcorder, a positioning device, a television receiver, a radio broadcast receiver, an electronic book device, a game device, or any combination thereof, including the accessories and peripherals of these devices, or any combination thereof. In some embodiments, the electronic device 110 can also support any type of interface to a user (such as "wearable" circuitry, etc.).

[0029] The server 130 can be a standalone physical server, a server cluster composed of multiple physical servers, or a distributed system, and can also be a cloud server providing cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communication, middleware services, domain name services, security services, content distribution networks, and basic cloud computing services such as big data and artificial intelligence platforms. The server 130 can include, for example, a computing system / server such as a mainframe, an edge computing node, a computing device in a cloud environment, and the like. The server 130 can provide background services for the application 120 in the electronic device 110 that supports content presentation.

[0030] A communication connection can be established between the server 130 and the electronic device 110. The communication connection can be established by wired or wireless means. The communication connection can include, but is not limited to, a Bluetooth connection, a mobile network connection, a universal serial bus connection, a wireless fidelity connection, and the like, and embodiments of the present disclosure are not limited in this regard. In embodiments of the present disclosure, the server 130 and the electronic device 110 can achieve signaling interaction through the communication connection therebetween.

[0031] It should be understood that the structure and function of the various elements in the environment 100 are described for illustrative purposes only, without implying any limitation on the scope of the present disclosure.

[0032] Example Interaction

[0033] The following will describe a process of generating media content according to embodiments of the present disclosure with reference to the accompanying drawings.

[0034] FIGS. 2A to 2H illustrate example interfaces 200A to 200H according to some embodiments of the present disclosure. The interfaces 200A to 200H can be provided by the electronic device 110 shown in FIG. 1.

[0035] In some embodiments, the electronic device 110 can present a generation interface 200A associated with a virtual object as shown in FIG. 2A. In some scenarios, such a virtual object can also be referred to as a digital avatar or a virtual avatar, for example.

[0036] Further, in some embodiments, such a virtual object can be created based on a configuration operation of a user. As an example, a user can be able to create a corresponding virtual object through a virtual object configuration platform, for example.

[0037] In some embodiments, such a virtual object can also include a first visual element of the virtual object and a capability (e.g., a conversation capability, a data processing capability, etc.) possessed by the virtual object. As an example, the first visual element of the virtual object can be generated based on a description text and / or image data input by a user, for example.

[0038] In some embodiments, the first visual element of the virtual object can be generated based on user inputted description text. As an example, such user inputted description text may, for example, include role setting information, knowledge information, image information, and the like, so as to enable the virtual object to have certain interactive ability and personality characteristics based on such configuration information. As an example, such user inputted description text may, for example, be “help me generate a middle-aged man”.

[0039] In some embodiments, such image data may, for example, be obtained based on user configuration operation. As an example, the generation interface 200A can include a preset control 210. The electronic device 110 can switch the interface 200A to the interface 200B as shown in FIG. 2B based on user triggering operation on the preset control 210. Further, the electronic device 110 can obtain image data of the virtual object based on user configuration operation by using the camera component 220 to capture an image or selecting an image from the album 222, so as to generate the first visual element of the virtual object.

[0040] In yet another embodiment, the electronic device 110 can generate the first visual element of the virtual object based on the combination of user inputted description text and user using the camera component 220 to capture an image or selecting an image from the album 222. In this way, the electronic device 110 can generate the first visual element of the virtual object that is more in line with user expectation.

[0041] In some embodiments, such first visual element of the virtual image may, for example, be presented in the form of three-dimensional solid, dynamic video or static image. For ease of understanding, the disclosure takes the first visual element of the virtual object in the form of three-dimensional solid as an example for description.

[0042] Further, for ease of description, the first visual element of the virtual object will be referred to as the target virtual object hereinafter.

[0043] In some embodiments, the electronic device 110 can generate a preset number (for example, 3) of target virtual objects for user to select the target virtual object in line with expectation.

[0044] In some embodiments, the generated target virtual object can be presented in a full-body avatar, for example. It can be appreciated that the electronic device 110 can obtain the target virtual object based on a single one or any combination of the three manners of taking an image with the camera component 220, selecting an image from the photo album 222, and inputting a description text by the user. Further, if the obtained target virtual object is a half-body avatar, the electronic device 110 can expand the half-body avatar of the target virtual object to a full-body avatar based on the description text input by the user or other description text. In order to perform the dressing operation in the subsequent dressing stage, the electronic device 110 can perform the dressing operation based on the full-body avatar of the target virtual object.

[0045] With continued reference to FIG. 2B, in some embodiments, the electronic device 110 can also view the historically generated target virtual objects based on the preset control 224. Further, the electronic device 110 can replace the currently generated target virtual object based on the user’s selection of the historically generated target virtual objects.

[0046] In some embodiments, referring to FIG. 2C, the electronic device 110 can present the target virtual object 230 in the generation interface 200C in response to the user’s completion of the creation of the target virtual object.

[0047] In some embodiments, the electronic device 110 can view the target virtual object 230 from various perspectives based on a preset operation (e.g., a swiping operation).

[0048] In some embodiments, the generation interface 200C can include a preset control 231, and the electronic device 110 can reselect the avatar data of the virtual object and / or re-input the description text to regenerate a new target virtual object in response to the user’s triggering operation on the preset control 231. As an example, the user can modify the input description text, for example, from “help me generate a middle-aged man” to “help me generate a middle-aged man wearing glasses”. In this way, the electronic device 110 can regenerate the target virtual object corresponding to the original description text or the description text modified by the user.

[0049] In some embodiments, the generation interface 200C can also include a preset control 232, and the electronic device 110 can present various types of clothing, such as a backpack, a shirt, pants, a skirt, a hat, and the like, in response to the user’s triggering operation on the preset control 232. In some embodiments, the electronic device 110 can perform the dressing operation for the target virtual object based on the user’s selection of the clothing in the wardrobe.

[0050] Further, in some embodiments, the electronic device 110 can present a wardrobe matching the target virtual object based on the information of the gender, style, etc. of the generated target virtual object. It can be appreciated that the wardrobe can be dynamically changed based on the information of the target virtual object to match the style of the target virtual object.

[0051] In some embodiments, the generation interface 200C can further include preset control 233 and preset control 234. The electronic device 110 can switch the interface 200C to the interface 200D as shown in FIG. 2D based on the triggering operation of the preset control 233 or the preset control 234 by the user.

[0052] In some embodiments, the electronic device 110 can generate the corresponding target expected clothing for the target virtual object 230 based on the operation of the user.

[0053] In some embodiments, the electronic device 110 can generate the target expected clothing for the target virtual object 230 based on the image captured by the camera component 242 or based on the image selected from the local album 244.

[0054] In the generation interface 200D, the electronic device 110 can switch the interface 200D to the interface 200E as shown in FIG. 2E based on the triggering operation of the preset control 240 by the user.

[0055] In some embodiments, the generation interface 200E can include an input control 250. In the input control 250, the user can input the clothing description text for the target virtual object 230. In some embodiments, such clothing description text includes but is not limited to the kind, color, texture, style, version of the clothing and the gender of the target virtual object wearing or wearing the clothing. As an example, such clothing description text can be a shirt, trousers, and sneakers, for example.

[0056] In some embodiments, the clothing description text input by the user can be an irregular or ambiguous natural language description. Further, the electronic device 110 can rewrite or expand the clothing description text input by the user based on a text expansion module to generate target text that is easy for the electronic device 110 to recognize, thereby facilitating the generation of target media content.

[0057] In some embodiments, the electronic device 110 can generate the corresponding second visual element (e.g., the target expected clothing of the target virtual object 230) based on the clothing description text.

[0058] In some embodiments, with continued reference to FIG. 2D, the electronic device 110 can further obtain the first media content via generating the interface 200D for generating the target media content. As an example, the electronic device 110 can combine the image captured by the camera component 242 or the image selected from the local photo album 244 with the user inputted clothing description text as the first media content, to generate the target expected clothing of the corresponding target virtual object 230.

[0059] In some embodiments, the electronic device 110 can obtain the first clothing object based on the above-mentioned obtained image, such as the first clothing object can include partial clothing information (e.g., a shirt) of the target virtual object 230. Further, the electronic device 110 can further supplement the remaining partial clothing information based on the user inputted clothing description text to obtain the second clothing object (e.g., pants). In this way, the electronic device 110 can still generate the complete clothing of the corresponding target virtual object 230 as the target expected clothing.

[0060] In some embodiments, such target media content can be the third visual content associated with the first clothing object, and the second visual content is associated with the second clothing object.

[0061] In some other embodiments, the electronic device 110 can further obtain the third clothing object (e.g., the complete clothing information of the target virtual object 230) based on the above-mentioned obtained image. Further, the electronic device 110 can further indicate to modify the partial clothing information of the third clothing object obtained based on the image based on the user inputted clothing description text, such as the style, style, color, etc. of the clothing.

[0062] For ease of description, the following is described by way of example of generating the target expected clothing based on the user inputted clothing description text.

[0063] In some embodiments, the interface 200E can further include a preset control 252. The electronic device can switch the interface 200E to the interface 200F as shown in FIG. 2F based on the user triggering operation on the preset control 252. In the generation interface 200F, the electronic device 110 can provide guidance information based on the user inputted content (e.g., clothing description text), such as the guidance information can be a suggestion about the clothing description text, such as “try to add adjectives, the effect is better”. In addition, such guidance information can also be the first set of candidate texts generated by the electronic device 110 based on the inputted content (e.g., clothing description text). Such first set of candidate texts can be the optimized candidate texts of the inputted clothing description text.

[0064] As an example, in the generation interface 200F, a component 260 can also be included. In the component 260, the electronic device 110 can present the first set of candidate texts. Such a first set of candidate texts can include two candidate texts, as an example, one candidate text is a man wearing a black shirt, white pants, and white sneakers; and another candidate text is a man wearing a white and black shirt, black suit pants, and white sneakers.

[0065] Further, in some embodiments, if the user is not satisfied with the currently generated candidate text, the electronic device 110 can regenerate the candidate text based on the user’s triggering operation on the preset control 264.

[0066] In some embodiments, the electronic device 110 can also present a second set of candidate texts based on the generation interface 200E or the generation interface 200F. As an example, such a second set of candidate texts can be, for example, a set of preset texts describing the clothing. In some embodiments, the electronic device 110 determines the clothing description text in response to the user’s selection of a target text in the second set of candidate texts. In addition, the electronic device 110 can also edit (e.g., add, modify, etc.) the target text in the second set of candidate texts, and further, the electronic device 110 can determine the edited target text as the clothing description text based on the user’s selection.

[0067] In some embodiments, the electronic device 110 can present the second visual element 270 corresponding to the clothing description text (e.g., the candidate text 262) in the interface 200G as shown in FIG. 2G based on the user’s selection of the candidate text 262 in the first set of candidate texts.

[0068] In some embodiments, the electronic device 110 can also present the target virtual object 230 and the second visual element 270 corresponding to the clothing description text in the interface 200G. In some embodiments, the interface 200G can also include a preset control 271, and the electronic device 110 can present the interface 200H as shown in FIG. 2H in response to the user’s triggering operation on the preset control 271. In the publishing interface 200H, the electronic device 110 can present the target media content. Such a target media content can be, for example, a dressing effect picture of the target virtual object 230 based on the second visual element 270 corresponding to the clothing description text.

[0069] In some embodiments, the presentation form of such a target media content includes but is not limited to the form of a video, the form of a static image, and the form of a three-dimensional solid.

[0070] In some embodiments, the electronic device 110 can further obtain a second media content via the generation interface 200G. In some embodiments, such a second media content can be, for example, the electronic device 110 obtaining a second media content uploaded by other users or a set of candidate media contents in the generation interface. In some embodiments, the electronic device 110 can provide a reference description text generated based on the second media content, such a reference description text is used to describe the clothing information in the second media content, i.e., the clothing description text.

[0071] As an example, the preset control 272 can also be included in the generation interface 200G. The electronic device 110 can provide various styles of second media content (e.g., outfit recommendations) based on the triggering operation of the user on the preset control 272. The electronic device 110 can present a dressing effect corresponding to the outfit in response to the user's selection of a certain second media content. Further, the electronic device 110 can also present the clothing description text corresponding to the second media content.

[0072] In some embodiments, the electronic device 110 can perform forwarding and saving operations on the target media content based on the user's operation.

[0073] In some embodiments, the electronic device 110 can also publish the target media content as a template in response to receiving a user's publishing request in the publishing interface 200H to share the target media content with other users. Further, in some embodiments, the electronic device 110 can display the above-mentioned input clothing description text on the viewing interface of the target media content. In some embodiments, the user can also set the visibility range of the target media content based on the electronic device 110.

[0074] In some embodiments, the electronic device 110 can present a re-generated target media content (or additional media content) based on the above-mentioned obtained virtual object appearance data and clothing description text in response to receiving a user's re-update request.

[0075] In this way, the embodiments of the present disclosure can apply the generated clothing to the created virtual object based on the user input clothing description text after receiving the user's media generation request, thereby generating the target media content. In this way, the present disclosure can provide the user with media content associated with the virtual object (e.g., virtual avatar), thereby improving the quality of the generated media content.

[0076] Example process

[0077] FIG. 3 illustrates a flowchart of an example process 300 of generating media content, in accordance with some embodiments of the present disclosure. Process 300 can be implemented at electronic device 110. Process 300 is described below with reference to FIG. 1.

[0078] As shown, at block 310, electronic device 110 presents a generation interface associated with a virtual object, the virtual object being created based on a configuration operation of a user.

[0079] At block 320, electronic device 110 obtains, via the generation interface, an apparel description text.

[0080] At block 330, electronic device 110 presents target media content, the target media content being generated based on the avatar data of the virtual object and the apparel description text, the target media content including a first visual element corresponding to the virtual object and a second visual element corresponding to the apparel description text.

[0081] In some embodiments, obtaining, via the generation interface, the apparel description text includes obtaining, via an input control in the generation interface, the apparel description text, wherein the apparel description text is provided to a text expansion module to generate a target text, the target text being provided for generating the target media content.

[0082] In some embodiments, process 300 further includes providing, based on input content in the input control, guidance information, the guidance information including suggestions on inputting the apparel description text, and / or a first set of candidate texts generated based on the input content.

[0083] In some embodiments, obtaining, via the generation interface, the apparel description text includes presenting, in the generation interface, a second set of candidate texts; and determining, based on a selection of a target text in the second set of candidate texts, the apparel description text.

[0084] In some embodiments, determining, based on the selection of the target text in the second set of candidate texts, the apparel description text includes receiving an editing operation for the target text; and determining, based on the edited target text, the apparel description text.

[0085] In some embodiments, process 300 further includes obtaining, via the generation interface, first media content, the first media content being provided for generating the target media content.

[0086] In some embodiments, the first media content corresponds to a first apparel object, the apparel description text corresponds to a second apparel object, the target media content further includes a third visual content associated with the first apparel object, and the second visual content is associated with the second apparel object.

[0087] In some embodiments, the apparel description text indicates a modification to a third apparel object in the first media content.

[0088] In some embodiments, obtaining the apparel description text via the generation interface includes: obtaining second media content via the generation interface; providing reference description text generated based on the second media content, the reference description text being used to describe apparel information in the second media content; and determining the apparel description text based on the reference description text.

[0089] In some embodiments, obtaining the second media content via the generation interface includes: obtaining the uploaded second media content; or receiving a selection of the second media content from a set of candidate media content in the generation interface.

[0090] In some embodiments, the process 300 further includes: in response to receiving the re-updating request, presenting additional media content generated based on the avatar data of the virtual object and the apparel description text.

[0091] In some embodiments, the process 300 further includes: in response to receiving the publishing request via the publishing interface, publishing the target media content such that a viewing interface of the target media content displays the apparel description text.

[0092] Example apparatuses and devices

[0093] Embodiments of the present disclosure also provide corresponding apparatuses for implementing the above-described methods or processes. FIG. 4 shows a schematic structural block diagram of an example generation media content apparatus 400 according to certain embodiments of the present disclosure. The apparatus 400 can be implemented as or included in the electronic device 110. Various modules / components in the apparatus 400 can be implemented by hardware, software, firmware, or any combination thereof.

[0094] As shown in FIG. 4, the apparatus 400 includes an interface presentation module 410 configured to present a generation interface associated with a virtual object, the virtual object being created based on a configuration operation of a user; a text obtaining module 420 configured to obtain, via the generation interface, apparel description text; and a content presentation module 430 configured to present target media content, the target media content being generated based on avatar data of the virtual object and the apparel description text, the target media content including a first visual element corresponding to the virtual object and a second visual element corresponding to the apparel description text.

[0095] In some embodiments, the text obtaining module 420 is further configured to obtain, via an input control in the generation interface, the apparel description text, wherein the apparel description text is provided to a text expansion module to generate target text, the target text being provided for generating the target media content.

[0096] In some embodiments, the apparatus 400 further includes an information providing module configured to provide, based on the input content in the input control, guidance information, the guidance information including: a suggestion on inputting the clothing description text, and / or a first set of candidate texts generated based on the input content.

[0097] In some embodiments, the text obtaining module 420 is further configured to present, in the generation interface, a second set of candidate texts; and determine, based on a selection of a target text in the second set of candidate texts, the clothing description text.

[0098] In some embodiments, the text obtaining module 420 is further configured to receive an editing operation on the target text; and determine, based on the edited target text, the clothing description text.

[0099] In some embodiments, the apparatus 400 further includes a content obtaining module configured to obtain, via the generation interface, first media content, the first media content being provided for generating the target media content.

[0100] In some embodiments, the first media content corresponds to a first clothing object, the clothing description text corresponds to a second clothing object, the target media content further includes third visual content associated with the first clothing object, and the second visual content is associated with the second clothing object.

[0101] In some embodiments, the clothing description text indicates a modification to a third clothing object in the first media content.

[0102] In some embodiments, the text obtaining module 420 is further configured to obtain, via the generation interface, second media content; provide a reference description text generated based on the second media content, the reference description text being used to describe clothing information in the second media content; and determine, based on the reference description text, the clothing description text.

[0103] In some embodiments, the text obtaining module 420 is further configured to obtain, via the generation interface, the second media content including: obtaining uploaded second media content; or receiving a selection of the second media content in a set of candidate media content in the generation interface.

[0104] In some embodiments, the apparatus 400 further includes a content generation module configured to, in response to receiving the re-updating request, present additional media content generated based on the avatar data of the virtual object and the clothing description text.

[0105] In some embodiments, the apparatus 400 further includes a content publishing module configured to, in response to receiving, via the publishing interface, a publishing request, publish the target media content such that a viewing interface of the target media content displays the clothing description text.

[0106] The modules included in the apparatus 400 can be implemented utilizing a variety of means, including software, hardware, firmware, or any combination thereof. In some embodiments, one or more of the units can be implemented using software and / or firmware, e.g., machine-executable instructions stored on a machine-readable medium. In addition or as an alternative, some or all of the modules in the apparatus 400 can be implemented, at least partially, by one or more hardware logic components. As an example and not by way of limitation, example types of hardware logic components that can be used include Field- programmable Gate Arrays (FPGAs), Application-specific Integrated Circuits (ASICs), Application-specific Standard Products (ASSPs), System-on-a-chip systems (SOCs), Complex Programmable Logic Devices (CPLDs), etc.

[0107] FIG. 5 illustrates a block diagram of an electronic device 500 in which one or more embodiments of the disclosure can be implemented. It should be understood that the electronic device 500 illustrated in FIG. 5 is exemplary only and should not be taken as limiting on the functionality and scope of the embodiments described herein. The electronic device 500 illustrated in FIG. 5 can be used to implement the electronic device 110 of FIG. 1.

[0108] As shown in FIG. 5, the electronic device 500 is in the form of a general electronic device. Components of the electronic device 500 can include, but are not limited to, one or more processors or processing units 510, a memory 520, a storage device 530, one or more communication units 540, one or more input devices 550, and one or more output devices 560. The processing unit 510 can be a real or virtual processor and capable of executing various processing according to programs stored in the memory 520. In a multi-processor system, multiple processing units execute computer-executable instructions in parallel to improve parallel processing capabilities of the electronic device 500.

[0109] The electronic device 500 typically includes a number of computer storage media. Such media can be any available media that is accessible by the electronic device 500 and includes both volatile and non-volatile media, removable and non-removable media. The memory 520 can be a volatile memory (e.g., registers, cache, random access memory (RAM)), a non-volatile memory (e.g., read-only memory (ROM), electrically erasable programmable read-only memory (EEPROM), flash memory), or some combination thereof. The storage device 530 can be a removable or non-removable medium and can include machine-readable media, such as a flash drive, a disk drive, or any other medium that can be used to store information and / or data and that can be accessed by the electronic device 500.

[0110] The electronic device 500 can further include additional detachable / non-detachable, volatile / non-volatile storage media. Although not shown in FIG. 5, a disk drive for reading from or writing to a detachable, non-volatile magnetic disk (e.g., a "floppy disk"), and an optical disk drive for reading from or writing to a detachable, non-volatile optical disk (e.g., a CD-ROM) can be provided. In these cases, each drive can be connected to the bus (not shown) by one or more data media interfaces. The memory 520 can include a computer program product 525 having one or more program modules configured to carry out the various methods or acts of the various embodiments of the present disclosure.

[0111] The communication unit 540 enables communication with other electronic devices through communication media. Additionally, the functionality of the components of the electronic device 500 can be implemented in a single computing cluster or a plurality of computer machines capable of communicating with one another through a communication connection. As such, the electronic device 500 can operate in a networked environment using logical connections to one or more other servers, network personal computers (PCs), or another network nodes in the networking environment.

[0112] The input device 550 can be one or more input devices, such as a mouse, a keyboard, a trackball, etc. The output device 560 can be one or more output devices, such as a display, a speaker, a printer, etc. The electronic device 500 can also communicate with one or more external devices (not shown) such as a storage device, a display device, etc., one or more devices that enable a user to interact with the electronic device 500, or any devices (e.g., a network card, a modem, etc.) that enable the electronic device 500 to communicate with one or more other electronic devices, through the communication unit 540, as needed. Such communication can be carried out via an input / output (I / O) interface (not shown).

[0113] According to an example implementation of the present disclosure, there is provided a computer-readable storage medium having computer-executable instructions stored thereon, where the computer-executable instructions are executed by a processor to implement the method described above. According to an example implementation of the present disclosure, there is also provided a computer program product tangibly stored on a non-transitory computer-readable medium and comprising computer-executable instructions, where the computer-executable instructions are executed by a processor to implement the method described above.

[0114] Various aspects of the disclosure are now described with reference to the drawings. In general, the drawings described below are diagrammatic and schematic representations of actual or conceptual structures and processes, and are not limiting of the scope of the present disclosure. In the drawings, the size and relative positioning of components can be exaggerated for clarity and / or descriptive purposes. Also, the drawings represent examples of apparatuses and / or methods in accordance with the present disclosure. In some instances, various aspects of the disclosure can be shown in a diagram, or by a series of diagrams, and can include a circuit, a flowchart, a table, a graph, a diagram, a schematic, or any combination thereof. These diagrams are meant to be exemplary and other implementations can be used. It is to be understood that the aspects of the present disclosure can be implemented in hardware, software, firmware, middleware, microcode, or any combination thereof. Furthermore, some aspects of the present disclosure can be implemented as a UDDI (Universal Description Discovery and Integration) service, a Web service, or any other suitable distributed computing architecture.

[0115] The computer readable program instructions can also be loaded onto a computer, other programmable data processing apparatus, or other device to cause a series of operational steps to be performed on the computer, other programmable apparatus or other device to produce a computer implemented process such that the instructions which execute on the computer or other programmable apparatus provide processes for implementing the functions / acts specified in the flowchart and / or block diagram block or blocks.

[0116] The computer readable program instructions can also be loaded onto a computer, other programmable data processing apparatus, or other device to cause a series of operational steps to be performed on the computer, other programmable apparatus or other device to produce a computer implemented process such that the instructions which execute on the computer or other programmable apparatus provide processes for implementing the functions / acts specified in the flowchart and / or block diagram block or blocks.

[0117] The computer readable program instructions can also be loaded onto a computer, other programmable data processing apparatus, or other device to cause a series of operational steps to be performed on the computer, other programmable apparatus or other device to produce a computer implemented process such that the instructions which execute on the computer or other programmable apparatus provide processes for implementing the functions / acts specified in the flowchart and / or block diagram block or blocks.

[0118] The implementations of the disclosure have been described above with the intent to be illustrative rather than limiting. Although being shown in only a few of the various implementations, the principle of each implementation can be extended to any other implementation. Some of the present implementations have also been described with the intent to be illustrative rather than restrictive. Many modifications and variations of the described implementations are possible in light of the above teachings. It is therefore contemplated that the application can encompass modifications and variations provided they come within the scope of the appended claims. It is also contemplated that the implementing specific electric circuitry such as, for example, application specific integrated circuits (ASICs) can be configured to implement one or more of the processes described herein. The choice of language in the claims is intended to be interpreted as limiting the scope of the claims to the full extent of the language.

Claims

1. A method of generating media content, comprising: presenting a generation interface associated with a virtual object, the virtual object being created based on a configuration operation of a user; acquiring, via the generation interface, a clothing description text; and presenting a target media content, the target media content being generated based on an appearance data of the virtual object and the clothing description text, the target media content including a first visual element corresponding to the virtual object and a second visual element corresponding to the clothing description text.

2. The method of claim 1, wherein acquiring, via the generation interface, a clothing description text comprises: acquiring, via an input control in the generation interface, the clothing description text, wherein the clothing description text is provided to a text expansion module to generate a target text, the target text being provided for generating the target media content.

3. The method of claim 2, further comprising: based on an input content in the input control, providing guidance information, the guidance information including a suggestion on inputting the clothing description text, and / or a first set of candidate texts generated based on the input content.

4. The method of claim 1, wherein acquiring, via the generation interface, a clothing description text comprises: presenting, in the generation interface, a second set of candidate texts; and based on a selection of a target text in the second set of candidate texts, determining the clothing description text.

5. The method of claim 4, wherein determining the clothing description text based on the selection of the target text in the second set of candidate texts comprises: receiving an editing operation on the target text; and based on the edited target text, determining the clothing description text.

6. The method of claim 1, further comprising: acquiring, via the generation interface, a first media content, the first media content being provided for generating the target media content.

7. The method of claim 6, wherein the first media content corresponds to a first clothing object, the clothing description text corresponds to a second clothing object, the target media content further includes a third visual content associated with the first clothing object, and the second visual content is associated with the second clothing object.

8. The method of claim 6, wherein the clothing description text indicates a modification to a third clothing object in the first media content.

9. The method of claim 1, wherein acquiring, via the generation interface, a clothing description text comprises: acquiring, via the generation interface, a second media content; providing a reference description text generated based on the second media content, the reference description text being used to describe clothing information in the second media content; and based on the reference description text, determining the clothing description text.

10. The method of claim 9, wherein acquiring, via the generation interface, a second media content comprises: acquiring the second media content that is uploaded; or receiving a selection of the second media content in a set of candidate media contents in the generation interface.

11. The method of claim 1, further comprising: ​ ​ ​ ​ ​ In response to receiving the re-updating request, presenting additional media content generated based on the avatar data of the virtual object and the clothing description text.

12. The method of claim 1, further comprising: In response to receiving a publishing request via the publishing interface, publishing the target media content such that a viewing interface of the target media content displays the clothing description text.

13. An apparatus for generating media content, comprising: an interface presenting module configured to present a generation interface associated with a virtual object, the virtual object being created based on a configuration operation of a user; a text obtaining module configured to obtain, via the generation interface, a clothing description text; and a content presenting module configured to present a target media content, the target media content being generated based on avatar data of the virtual object and the clothing description text, the target media content including a first visual element corresponding to the virtual object and a second visual element corresponding to the clothing description text.

14. An electronic device, comprising: at least one processing unit; and at least one memory coupled to the at least one processing unit and storing instructions for execution by the at least one processing unit, the instructions when executed by the at least one processing unit cause the electronic device to perform the method according to any one of claims 1-12.

15. A computer readable storage medium having stored thereon a computer program, the computer program being executable by a processor to implement the method according to any one of claims 1-12.