Method and device for generating media content, equipment and storage medium

By presenting visual content with users and virtual objects in the interactive interface and generating media content, the problem of enriching media content forms and improving interaction efficiency is solved, and more efficient user-to-virtual objects interaction and media content generation is achieved.

CN120547418APending Publication Date: 2025-08-26BEIJING ZITIAO NETWORK TECH CO LTD

Patent Information

Application Number
CN202510840587.5
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-06-21
Publication Date
2025-08-26

AI Technical Summary

Technical Problem

In the prior art, how to enrich the form of media content associated with virtual objects, improve the interaction efficiency between users and virtual objects and the efficiency of media content generation.

Method used

The first visual content corresponding to the first user and the second visual content corresponding to the virtual object are presented in the interactive interface, and media content, including the first visual content and the second visual content, is generated in response to the shooting operation.

Benefits of technology

It enriches the interaction form between users and virtual objects, improves the interaction efficiency between users and virtual objects and the efficiency of media content generation.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120547418A_ABST
    Figure CN120547418A_ABST
Patent Text Reader

Abstract

The embodiment of the invention relates to a media content generation method and device, equipment and a storage medium. The method comprises the steps that in an interaction interface, first visual content corresponding to a first user and second visual content corresponding to a virtual object are presented, and the virtual object is configured to be interacted by the first user and a second user in a combined mode; and in response to a shooting operation received in the interactive interface, generating media content, the media content including the first visual content and the second visual content. In this way, embodiments of the present disclosure can support generation of media content associated with visual content of a user and a virtual object.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] Example embodiments of the present disclosure generally relate to the field of computers, and more particularly, to methods, devices, apparatuses, and computer-readable storage media for generating media content. Background Art

[0002] As network technology matures, more and more users are engaging in interactive activities on online platforms. For example, users can view or share media content through online platforms. For example, users can share media content associated with virtual objects through online platforms. Therefore, how to enrich the forms of media content associated with virtual objects is worthy of attention. Summary of the Invention

[0003] In a first aspect of the present disclosure, a method for generating media content is provided. The method comprises: presenting, in an interactive interface, first visual content corresponding to a first user and second visual content corresponding to a virtual object, the virtual object being configured for joint interaction by the first user and the second user; and generating, in response to a capture operation received in the interactive interface, media content, the media content comprising the first visual content and the second visual content.

[0004] In a second aspect of the present disclosure, a device for generating media content is provided. The device includes: a presentation module configured to present, in an interactive interface, first visual content corresponding to a first user and second visual content corresponding to a virtual object, the virtual object being configured for joint interaction by the first user and the second user; and a generation module configured to generate media content in response to a capture operation received in the interactive interface, the media content including the first visual content and the second visual content.

[0005] In a third aspect of the present disclosure, an electronic device is provided. The device includes at least one processor; and at least one memory coupled to the at least one processor and storing instructions for execution by the at least one processor. When executed by the at least one processor, the instructions cause the device to perform the method of the first aspect.

[0006] In a fourth aspect of the present disclosure, a computer-readable storage medium is provided, wherein a computer program is stored on the computer-readable storage medium, and the computer program can be executed by a processor to implement the method of the first aspect.

[0007] It should be understood that the content described in this summary section is not intended to limit the key features or important features of the embodiments of the present disclosure, nor is it intended to limit the scope of the present disclosure. Other features of the present disclosure will become easily understood through the following description. BRIEF DESCRIPTION OF THE DRAWINGS

[0008] The above and other features, advantages and aspects of the embodiments of the present disclosure will become more apparent with reference to the following detailed description in conjunction with the accompanying drawings. In the accompanying drawings, the same or similar reference numerals represent the same or similar elements, wherein:

[0009] Figure 1 A schematic diagram illustrating an example environment in which embodiments according to the present disclosure may be implemented;

[0010] Figures 2A to 2G shows an example interface according to some embodiments of the present disclosure;

[0011] Figure 3 A flowchart illustrating an example process for generating media content according to some embodiments of the present disclosure is shown;

[0012] Figure 4 A schematic structural block diagram illustrating an example apparatus for generating media content according to some embodiments of the present disclosure is shown; and

[0013] Figure 5 A block diagram of an electronic device capable of implementing various embodiments of the present disclosure is shown. DETAILED DESCRIPTION

[0014] The following describes embodiments of the present disclosure in more detail with reference to the accompanying drawings. Although certain embodiments of the present disclosure are shown in the accompanying drawings, it should be understood that the present disclosure can be implemented in various forms and should not be construed as limited to the embodiments described herein. Rather, these embodiments are provided to provide a more thorough and complete understanding of the present disclosure. It should be understood that the drawings and embodiments of the present disclosure are for illustrative purposes only and are not intended to limit the scope of protection of the present disclosure.

[0015] It should be noted that the titles of any section / subsection provided herein are not limiting. Various embodiments are described throughout this document, and any type of embodiment may be included under any section / subsection. Furthermore, the embodiments described in any section / subsection may be combined in any manner with any other embodiments described in the same section / subsection and / or in different sections / subsections.

[0016] In the description of the embodiments of the present disclosure, the term "including" and similar terms should be understood as open inclusion, that is, "including but not limited to". The term "based on" should be understood as "based at least in part on". The term "one embodiment" or "the embodiment" should be understood as "at least one embodiment". The term "some embodiments" should be understood as "at least some embodiments". Other explicit and implicit definitions may be included below. The terms "first", "second", etc. may refer to different or the same objects. Other explicit and implicit definitions may be included below.

[0017] The embodiments of the present disclosure may involve user data, data acquisition and / or use, etc. These aspects shall comply with the corresponding laws, regulations and relevant provisions. In the embodiments of the present disclosure, all data collection, acquisition, processing, processing, forwarding, use, etc. are carried out on the premise that the user is aware of and confirms them. Accordingly, when implementing the various embodiments of the present disclosure, the types, scope of use, and usage scenarios of the data or information that may be involved should be informed to the user and the user's authorization should be obtained in an appropriate manner in accordance with the relevant laws and regulations. The specific notification and / or authorization method may vary according to the actual situation and application scenario, and the scope of the present disclosure is not limited in this respect.

[0018] If this specification and the solutions in the examples involve the processing of personal information, such processing will be done only with a legitimate basis (such as with the consent of the subject of personal information or as necessary for the performance of a contract) and only within the prescribed or agreed scope. A user's refusal to process personal information other than that required for basic functions will not affect the user's use of basic functions.

[0019] As mentioned above, with the increasing maturity of network technology, more and more users are engaging in interactive activities on online platforms. For example, users can view or share media content through online platforms. For example, users can share media content associated with virtual objects through online platforms. Therefore, how to enrich the forms of media content associated with virtual objects is worthy of attention.

[0020] Embodiments of the present disclosure provide a method for generating media content. According to this method, the present disclosure can present first visual content corresponding to a first user and second visual content corresponding to a virtual object in an interactive interface. The virtual object is configured to be jointly interacted with by the first user and the second user. Furthermore, the present disclosure can generate media content in response to a capture operation received in the interactive interface. The media content includes the first visual content and the second visual content.

[0021] In this way, the embodiments of the present disclosure can present the first visual content corresponding to the first user and the second visual content corresponding to the virtual object in the interactive interface, which is conducive to enriching the form of interaction between the user and the virtual object. Since such a virtual object is configured to be jointly interacted by the first user and the second user, it is conducive to improving the efficiency of interaction between the first user and the second user. In addition, the embodiments of the present disclosure can generate media content including the first visual content and the second visual content in response to the shooting operation received in the interactive interface, so as to support the user to take a photo with the virtual object, which is conducive to enriching the form of media content associated with the virtual object, and further conducive to improving the efficiency of interaction between the user and the virtual object.

[0022] Various example implementations of this solution are described in detail below with reference to the accompanying drawings.

[0023] Sample Environment

[0024] Figure 1 1 shows a schematic diagram of an example environment 100 in which embodiments of the present disclosure can be implemented. Figure 1 As shown, example environment 100 may include electronic device 110 .

[0025] In this example environment 100, electronic device 110 may run an application 120 that supports interface interaction. Application 120 may be any suitable type of application for interface interaction, examples of which may include, but are not limited to, media applications, photography applications, virtual object applications, or other suitable applications. User 140 may interact with application 120 via electronic device 110 and / or its attached devices.

[0026] exist Figure 1 In the environment 100 , if the application 120 is in an active state, the electronic device 110 may present an interface 150 for supporting interface interaction through the application 120 .

[0027] In some embodiments, the electronic device 110 communicates with the server 130 to enable the provision of services for the application 120. The electronic device 110 can be any type of mobile terminal, fixed terminal or portable terminal, including a mobile phone, a desktop computer, a laptop computer, a notebook computer, a netbook computer, a tablet computer, a media computer, a multimedia tablet, a handheld computer, a portable game terminal, a VR / AR device, a personal communication system (PCS) device, a personal navigation device, a personal digital assistant (PDA), an audio / video player, a digital camera / camcorder, a positioning device, a television receiver, a radio broadcast receiver, an e-book device, a gaming device or any combination thereof, including accessories and peripherals of these devices or any combination thereof. In some embodiments, the electronic device 110 can also support any type of interface for the user (such as a "wearable" circuit, etc.).

[0028] The server 130 may be a standalone physical server, a server cluster or distributed system consisting of multiple physical servers, or a cloud server that provides basic cloud computing services such as cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communications, middleware services, domain name services, security services, content distribution networks, and big data and artificial intelligence platforms. For example, the server 130 may include a computing system / server such as a mainframe, an edge computing node, a computing device in a cloud environment, and the like. The server 130 may provide background services for the application 120 that supports interface interaction in the electronic device 110.

[0029] A communication connection may be established between the server 130 and the electronic device 110. The communication connection may be established in a wired or wireless manner. The communication connection may include, but is not limited to, a Bluetooth connection, a mobile network connection, a Universal Serial Bus (USB) connection, a Wireless Fidelity (WiFi) connection, etc., and the embodiments of the present disclosure are not limited in this respect. In the embodiments of the present disclosure, the server 130 and the electronic device 110 may implement signaling interaction through the communication connection between the two.

[0030] It should be understood that the structure and function of the various elements in the environment 100 are described for illustrative purposes only and do not imply any limitation on the scope of the present disclosure.

[0031] Some example embodiments of the present disclosure will be described below with continued reference to the accompanying drawings.

[0032] Example Interaction

[0033] Figures 2A to 2G 200A to 200G according to some embodiments of the present disclosure. Figure 1 The electronic device 110 is provided as shown.

[0034] In some embodiments, as Figure 2AAs shown, the electronic device 110 can present an interface 200A. The interface 200A can be implemented as an interactive interface. Such an interactive interface may include a viewing interface of a first virtual image associated with the first user. As an example, the electronic device 110 can present a first visual content associated with the first user in the interface 200A. As an example, such a first visual content may include a first virtual image created based on the configuration operation of the first user, for example, the virtual image 205. Alternatively, such a first visual content may also correspond to the real image of the first user. For example, such a real image is determined based on a photo uploaded by the first user or an image content acquired by a camera component (e.g., a camera) carried by the electronic device 110.

[0035] As an example, a first virtual avatar associated with a first user may be created based on a configuration operation by the first user. As an example, the electronic device 110 may provide an upload portal to the first user. The electronic device 110 may present a creation interface (not shown) associated with creating a virtual avatar to the first user, and present the upload portal in the creation interface. The electronic device 110 may obtain a first image associated with the first user via the upload portal. For example, the electronic device 110 may present a set of image content associated with the first user (e.g., a set of image content in the first user's photo album) in response to triggering the upload portal. Further, the electronic device 110 may obtain a first image associated with the first user in response to the first user selecting at least one image in the set of image content. As an example, such a first image may include a photo of the first user or video content associated with the first user. As an example, such a first image may include one or more images associated with the first user. Such multiple images may, for example, correspond to different shooting angles. Alternatively, the electronic device 110 may obtain the first image associated with the first user based on a camera component (e.g., a camera).

[0036] Additionally, the electronic device 110 or the server 130 may create a first virtual image associated with the first user based on the first image. As an example, the electronic device 110 or the server 130 may process the first image using a pre-trained model to generate a first virtual image corresponding to the first image. As an example, such a first virtual image may be implemented as a three-dimensional image (or a three-dimensional visual model). As an example, the pre-trained model may be implemented as a generative model capable of generating a three-dimensional image based on image content. The present disclosure is not intended to limit the training process or specific implementation of the generative model.

[0037] In some embodiments, the first user can be associated with multiple virtual images (including the first virtual image). As an example, the multiple virtual images can be created based on the configuration operation of the first user. For example, the electronic device 110 can present an image portal, such as image portal 210, in the interface 200A. Further, the electronic device 110 can present at least one virtual image associated with the first user in response to triggering the image portal 210. Further, the electronic device 110 can determine the first virtual image (e.g., virtual image 205) presented in the interface 200A in response to receiving the first user's selection of a target virtual image among the at least one virtual images.

[0038] In some embodiments, the first visual content corresponding to the first user presented in interface 200A also includes at least one item of dress-up material associated with the first virtual avatar, such as dress-up material 208 (e.g., a top). For example, such at least one item of dress-up material may include virtual clothing (e.g., a coat, pants, and shoes), virtual accessories (e.g., a necklace, bracelet, and glasses), etc. As an example, such at least one item of dress-up material may be obtained by the first user through participating in an interactive activity. Alternatively, the at least one item of dress-up material may include dress-up material generated based on a second image uploaded by the first user. As an example, such a second image may include visual content corresponding to clothing and / or accessories. For example, the electronic device 110 or the server 130 may process the second image using a pre-trained model to generate dress-up material corresponding to the second image. As an example, the pre-trained model may be implemented as a generative model capable of generating dress-up material based on the second image. This disclosure is not intended to limit the training process or specific implementation of this generative model. As an example, such dress-up material may be three-dimensional dress-up material.

[0039] Additionally or alternatively, electronic device 110 may present a material entry associated with the dress-up material in interface 200A, such as material entry 215. Furthermore, electronic device 110 may present a set of dress-up materials associated with the first avatar in response to triggering material entry 215. Furthermore, electronic device 110 may present the at least one dress-up material associated with the first avatar in the interactive interface (e.g., interface 200A) in response to receiving a selection of at least one dress-up material from the set of dress-up materials.

[0040] Additionally or alternatively, such an interactive interface may be associated with a virtual object. Alternatively, the electronic device 110 may present a viewing interface associated with the virtual object to the first user. In some embodiments, descriptive information associated with the virtual object may be presented in the viewing interface of the virtual object. Such descriptive information may include the name identification, visual image, virtual level information and / or virtual resources associated with the virtual object, etc. Further, the electronic device 110 may provide a viewing entrance in the viewing interface of the virtual object. Further, the electronic device 110 may present an interactive interface (e.g., interface 200A) in response to triggering the viewing entrance.

[0041] As an example, such a virtual object can be configured to be jointly interacted with by a first user and a second user. For example, such a virtual object can be associated with both the first user and the second user. For example, the first user and the second user can be referred to as co-owners of the virtual object.

[0042] As an example, a virtual object can be associated with the first and second users in response to historical interactions between the first and second users satisfying preset conditions. For example, such historical interactions may include a set of interaction events in which the first and second users participated. The first and second users can obtain the virtual object in response to interaction attributes of the set of interaction events satisfying preset conditions. For example, interaction attributes may include interaction frequency, interaction content, etc.

[0043] For example, such a set of interaction events may include text message interactions, image message interactions, audio message interactions, video call interactions, voice call interactions, emoticon interactions, etc. in a conversation. For example, a virtual object may be associated with (or provided to) the first user and the second user in response to the number of consecutive days of conversation between the first user and the second user meeting a preset number of days (e.g., 3 days).

[0044] In some embodiments, such a set of interaction events may also include interaction events between the first user and the second user that occur independently of the session. For example, such interaction events may include a browsing event, a like event, a favorite event, a forwarding event, etc. of a work posted by the second user by the first user.

[0045] It should be understood that the acquisition and use of information on interactive events involved in the present disclosure are all carried out with the knowledge and authorization of relevant users (eg, the first user and the second user, etc.).

[0046] For example, the virtual object may include a virtual pet associated with the first user and the second user, such as a shared pet. Additionally, at least one attribute of the virtual object may be updated accordingly based on a set of interaction events between the current user and the target object, such that the virtual object may have different visual representations and / or different interaction capabilities.

[0047] As an example, the virtual object can be presented in a conversation interface and / or other associated interface with the first user and the second user. Alternatively, the virtual object can also be added to the virtual scene by the first user or the second user to interact, etc.

[0048] As an example, the electronic device 110 may present a second visual content corresponding to the virtual object in the interactive interface. As an example, the electronic device 110 may present a shooting portal associated with the virtual object in the interface 200A, such as the shooting portal 220. Furthermore, the electronic device 110 may present the second visual content corresponding to the virtual object in the interactive interface in response to a trigger (e.g., a click operation) on the shooting portal 220.

[0049] In some embodiments, as Figure 2B As shown, electronic device 110 may present interface 200B. For example, interface 200B may be implemented as an interactive interface. Electronic device 110 may present second visual content corresponding to the virtual object in interface 200B. The second visual content corresponding to the virtual object may include a visual image of the virtual object, such as visual image 225. As an example, the visual image of the virtual object is determined based on a configuration operation performed by the first user or the second user. For example, visual image 225 is determined by a configuration operation performed by the first user or the second user on the virtual object in a viewing interface or an associated interface (e.g., a conversation interface between the first user and the second user) for the virtual object.

[0050] Alternatively, the first user may be associated with multiple virtual objects. For example, the multiple virtual objects may include virtual object A and virtual object B, etc. For example, virtual object A may be configured to be jointly interacted with by the first user and the second user. Alternatively, virtual object B may be configured to be jointly interacted with by the first user and the third user. As an example, the electronic device 110 may present visual content corresponding to virtual object A in the interactive interface in response to the interactive interface being presented via the viewing entrance of the viewing interface associated with virtual object A. Alternatively, the electronic device 110 may present visual content corresponding to virtual object B in the interactive interface in response to the interactive interface being presented via the viewing entrance of the viewing interface associated with virtual object B.

[0051] Additionally or alternatively, the electronic device 110 may also present a third visual content corresponding to the second user in the interactive interface. Figure 2C As shown, the electronic device 110 can present an interface 200C. The interface 200C can be implemented as an interactive interface. The electronic device 110 can present a third visual content corresponding to the second user in the interface 200C. As an example, the third visual content can include a second virtual image corresponding to the second user, for example, the virtual image 230. Alternatively, such a second visual content can also correspond to the real image of the second user. For example, such a real image is determined based on a photo uploaded by the second user or an image content acquired by a camera component (e.g., a camera) carried by a client associated with the second user.

[0052] Alternatively, the electronic device 110 may present the second virtual image (e.g., virtual image 230) in the interactive interface (e.g., interface 200C) as the third visual content in response to the second user being associated with the second virtual image. Alternatively, the electronic device 110 may present an invitation control (not shown in the figure) associated with the second user in response to the second user not being associated with the virtual image. Further, the electronic device 110 may send an invitation message to the second user in response to triggering the invitation control, where the invitation message is configured to trigger the virtual image corresponding to the second user. For example, such an invitation message may include a creation entry associated with the virtual image. The second user may create a second virtual image corresponding to the second user via the creation entry in the invitation message. The process of the second user creating the second virtual image can refer to the above exemplary description of the process of the first user creating the first virtual image, which will not be repeated here.

[0053] In some embodiments, the electronic device 110 may generate media content in response to receiving a capture operation in the interactive interface. The media content may include the first visual content and / or the second visual content. Alternatively, such media content may also include third visual content corresponding to the second user.

[0054] As an example, the shooting operation can be associated with the photo mode or the video recording mode. For example, the electronic device 110 can present a first control (or referred to as a photo control) associated with the photo mode in the interface 200C, for example, control 232-1. The electronic device 110 can associate the shooting operation with the photo mode in response to the selection of the first control. Further, the electronic device 110 can generate picture content associated with the first visual content and the second visual content in response to the shooting operation being associated with the photo mode. Alternatively, such picture content can also include third visual content corresponding to the second user.

[0055] Alternatively, the electronic device 110 may further present a second control (or referred to as a recording control) associated with the recording mode in the interface 200C, for example, control 232-2. The electronic device 110 may associate the shooting operation with the recording mode in response to the selection of the second control. Further, the electronic device 110 may generate video content associated with the first visual content and the second visual content in response to the shooting mode being associated with the recording mode. Alternatively, such video content may further include third visual content corresponding to the second user.

[0056] As an example, the electronic device 110 may provide a third control in the interface 200C, such as the control 234. Furthermore, the electronic device 110 may receive a shooting operation in response to a trigger (eg, a click operation) on the third control.

[0057] In some embodiments, the electronic device 110 may present a set of shooting controls in an interactive interface. The electronic device 110 may determine a set of shooting parameters (e.g., the shooting angle, shooting position, and / or field of view of a virtual camera, etc.) via a set of shooting controls. Further, the electronic device 110 may generate media content based on a set of shooting parameters. As an example, such a set of shooting controls may be associated with a virtual camera that generates media content. As an example, such a set of shooting controls may include a first shooting control (not shown in the figure) for configuring the orientation of the virtual camera. Further, the electronic device 110 may obtain an interactive operation (e.g., a sliding operation) of the first user based on the first shooting control to adjust the orientation (or angle) of the virtual camera to capture media content at different angles. Alternatively, the electronic device 110 may adjust the orientation (or angle) of the virtual camera based on the interactive operation in response to receiving an interactive operation (e.g., a sliding operation) of the user in a blank area of ​​the interactive interface.

[0058] In some embodiments, such a set of shooting controls may include a second shooting control for configuring the position of the virtual camera, for example, control 235-1. Further, the electronic device 110 may obtain an interactive operation (for example, a sliding operation) of the first user based on the second shooting control to adjust the position of the virtual camera so as to capture media content at different positions.

[0059] In some embodiments, such a set of capture controls may include a third capture control for configuring the virtual camera's field of view (FOV). For example, the field of view may be referred to as the field of view angle. The field of view angle may refer to the angle between the observation point (e.g., the center of the camera lens) and the two edges of the observed object. The field of view may directly affect the scope and / or level of detail of the generated media content. Furthermore, electronic device 110 may receive an interaction (e.g., a swipe) from the first user based on the third capture control to adjust the virtual camera's field of view to capture media content with different fields of view. As an example, the virtual camera's field of view may be determined based on the virtual camera's focal length. As an example, such a third capture control may include control 235-2 (e.g., a focal length control). Furthermore, electronic device 110 may receive an interaction (e.g., a swipe) from the first user based on control 235-2 to adjust the virtual camera's field of view by adjusting the virtual camera's focal length to capture media content with different fields of view.

[0060] In some embodiments, the electronic device 110 may present a configuration panel in the interactive interface. Further, the electronic device 110 may adjust the display state of the first visual content and / or the second visual content in the interactive interface based on the configuration operation received via the configuration panel. Figure 2D As shown, electronic device 110 may present interface 200D. For example, electronic device 110 may present a set of configuration controls in interface 200D, such as configuration control 240-1, configuration control 240-2, configuration control 240-3, configuration control 240-4, and configuration control 240-5. Electronic device 110 may receive a configuration operation from the first user based on the set of configuration controls to adjust the display state of the first visual content and / or the second visual content in the interactive interface.

[0061] As an example, the electronic device 110 may present a configuration panel (e.g., also referred to as a first configuration panel), e.g., configuration panel 245, in response to a trigger (e.g., a click operation) on a configuration control 240-3 (e.g., a setting control). The electronic device 110 may present a set of display controls associated with the first visual content and / or the second visual content in the configuration panel 245. Further, the electronic device 110 may determine whether the first visual content and / or the second visual content is presented in the media content based on a first configuration operation received via a set of display controls. For example, a set of display controls may include a first display control associated with the first visual content, e.g., display control 250-1. For example, a set of display controls may include a second display control associated with the second visual content, e.g., display control 250-3. For example, a set of display controls may include a third display control associated with the third visual content, e.g., display control 250-2. Taking display control 250-1 as an example, the electronic device 110 may determine whether the display control 250-1 is in an on or off state in response to the display control receiving a first configuration operation (e.g., a click operation) from the user. For example, the electronic device 110 may determine that the first visual content is not presented in the generated media content in response to the display control 250-1 being in the open state. Alternatively, the electronic device 110 may determine that the first visual content is presented in the generated media content in response to the display control 250-1 being in the closed state.

[0062] Alternatively, the electronic device 110 may further present a gaze control, such as gaze control 250-4, in the configuration interface 245. Furthermore, the electronic device 110 may control the orientation of the first visual content, the second visual content, or the third visual content to point toward the virtual camera in response to the gaze control being activated based on a user operation.

[0063] As an example, Figure 2EAs shown, the electronic device 110 can present an interface 200E. Interface 200E can be implemented as an interactive interface. The electronic device 110 can present a configuration panel (e.g., also known as a second configuration panel), such as configuration panel 255, in response to a trigger (e.g., a click operation) on a configuration control 240-4 (e.g., a position control). As an example, the electronic device 110 can form a group of mobile components associated with the first visual content and / or the second visual content in the configuration panel 255. Further, the electronic device 110 can determine the display position of the first visual content and / or the second visual content in the generated media content based on the second configuration operation received via the group of mobile components. As an example, the electronic device 110 can present a group of content controls in the configuration panel 255. A group of content controls can, for example, include a first content control corresponding to the first visual content, such as control 260-1. A group of content controls can, for example, include a second content control corresponding to the second visual content, such as control 260-3. A group of content controls can, for example, include a third content control corresponding to the third visual content, such as control 260-2. Furthermore, the electronic device 110 may present a set of mobile components corresponding to the target content control in response to the selection of the target content control in the set of content controls. Taking the content control 260-1 as an example, the electronic device 110 may present a set of mobile controls with the content control 260-1 in response to the selection of the content control 260-1. For example, a set of mobile controls with the content control 260-1 may include at least one of the mobile control 265-1, the mobile control 265-2 and / or the mobile control 265-3. A set of mobile controls may, for example, correspond to different directions of movement (e.g., front, back, left, right, up and / or down). For example, the electronic device 110 may obtain the interactive operation of the first user (e.g., a click operation or a sliding operation) via the mobile control 265-1 to adjust the position of the first visual content in the media content with respect to the left and right directions.

[0064] Alternatively, the electronic device 110 may also present a set of size components associated with the first visual content and / or the second visual content in a configuration panel (e.g., a fourth configuration panel, not shown in the figure). Further, the electronic device 110 may determine the display size of the first visual content and / or the second visual content in the media content based on a third configuration operation (e.g., a sliding operation, a clicking operation, or an input operation (e.g., entering a number corresponding to the size)) received via the set of size components.

[0065] In some embodiments, the first visual content may include a first action associated with the first user. Alternatively, the second visual content may further include a second action associated with the virtual object. Alternatively, the media content may further include third visual content. The third visual content may include a fourth action associated with the second user. The presentation process of the fourth action may refer to the description of the presentation process of the first action below and will not be repeated here.

[0066] For example, Figure 2F As shown, the electronic device 110 can present an interface 200F. The interface 200F can be implemented as an interactive interface. For example, the electronic device 110 can present a configuration panel (also referred to as a fourth configuration panel), such as configuration panel 270, in response to a trigger (e.g., a click operation) on a configuration control 240-5 (e.g., an action control). For example, the electronic device 110 can present a group of content tags in the configuration panel 270. For example, a group of content tags can include, for example, a first content tag corresponding to a first visual content, such as tag 275-1. A group of content tags can include, for example, a second content tag corresponding to a second visual content, such as tag 275-3. A group of content tags can include, for example, a third content tag corresponding to a third visual content, such as tag 275-2. Further, the electronic device 110 can present a group of action controls corresponding to a target content tag in response to a selection of a target content tag in a group of content tags. Taking the first content tag as an example, in response to selecting the first content tag, the electronic device 110 may present a set of action controls associated with the first content tag in the configuration panel 270, for example, a first set of action controls associated with the first user. For example, the first set of action controls may include action control 280-1 and action control 280-2, etc. Further, in response to receiving a selection of the first action control (e.g., action control 280-1) in the first set of action controls, the electronic device 110 may trigger the first visual content to present a first action associated with the first action control.

[0067] Alternatively, the electronic device 110 may also present an input control in a configuration panel (e.g., a fifth configuration panel, not shown in the figure). Furthermore, the electronic device 110 may trigger the first visual content to present a first action associated with the input content in response to obtaining the input content of the first user via the input control. As an example, such input content may include an action description associated with the first visual content. As an example, such input content may include text content and / or voice content. As an example, the electronic device 110 may process the input content using a pre-trained model to generate action information associated with the output content. Further, the electronic device 110 may trigger the first visual content to present a first action associated with the input content based on the action information. As an example, such a pre-trained model may be implemented as a generative model that generates action information associated with the first virtual image based on the input content. The present disclosure is not intended to limit the training process or specific implementation of the generative model.

[0068] Alternatively, in response to the first action being triggered, the electronic device 110 may trigger the virtual object to perform a second action associated with the first action. For example, the first action may indicate a hand reaching motion of the first virtual avatar of the first user. The second action may indicate a jumping motion of the virtual object toward the hand position of the first virtual avatar.

[0069] Alternatively, the electronic device 110 may also trigger the virtual object to perform a second action associated with the second action control in response to receiving a selection of a second action control in a second group of action controls associated with the virtual object (e.g., a group of action controls associated with the label 275-3). The interaction process associated with the second action control can refer to the exemplary description associated with the first action control above and will not be repeated here.

[0070] Additionally or alternatively, the electronic device 110 may acquire image content of the first user via an image acquisition unit (e.g., a camera carried by the electronic device 110). As an example, such a process of acquiring image content of the first user may also be referred to as a motion capture process. Furthermore, the electronic device 110 may trigger the presentation of a third action corresponding to the image content in the first visual content based on the image content. As an example, such a third action may follow the action indicated in the acquired image content.

[0071] In some embodiments, the electronic device 110 can also adjust at least one of the virtual camera, first visual content, second visual content and / or third visual content to a preset state (e.g., a preset position, a preset orientation, etc.) in response to triggering the configuration control 240-1 (e.g., a reset control).

[0072] Alternatively, the electronic device 110 may trigger the configuration control 240-2 (e.g., hide the control) to stop presenting at least one control or component in the interactive interface to facilitate the user to view the visual content presented in the interactive interface, thereby improving the efficiency of generating media content.

[0073] In some embodiments, as Figure 2G As shown, the electronic device 110 can present an interface 200G. For example, the interface 200G can be implemented as a preview interface for media content. The electronic device 110 can present the generated media content, for example, media content 285, in the interface 200G. The electronic device 110 can publish a media content item corresponding to the media content in response to a request to publish the media content. As an example, the electronic device 110 can provide a save control, for example, control 290, in the interface 200G. Further, the electronic device 110 can store the media content 285 in a local media library (for example, a photo album) associated with the first user in response to triggering the save control.

[0074] As an example, the electronic device 110 may provide a publishing control, such as control 295, in the interface 200G. In response to receiving a trigger for the publishing control, the electronic device 110 may receive a publishing request for media content. Alternatively, the electronic device 110 may present a publishing interface associated with the media content in response to the publishing request for the media content. Furthermore, the electronic device 110 may obtain publishing information associated with the media content via the publishing interface to determine the media content item corresponding to the media content. As an example, such publishing information may indicate the publishing location, the name of the media content, and tag information of the media content. As an example, the viewing interface of the media content item may be configured to provide an interactive portal associated with the virtual object. For example, a user viewing the media content item may view the interactive interface associated with the virtual object via the interactive portal. In this way, embodiments of the present disclosure may support users viewing the interactive interface associated with the virtual object via the interactive portal in the viewing interface of the media content item, facilitating user generation of media content, thereby improving user interaction efficiency and the efficiency of generating media content associated with the virtual object.

[0075] Based on the process described above, in this way, the embodiments of the present disclosure can present the first visual content corresponding to the first user and the second visual content corresponding to the virtual object in the interactive interface, which is conducive to enriching the form of interaction between the user and the virtual object. Since such a virtual object is configured to be jointly interacted by the first user and the second user, it is conducive to improving the efficiency of interaction between the first user and the second user. In addition, the embodiments of the present disclosure can generate media content including the first visual content and the second visual content in response to the shooting operation received in the interactive interface, so as to support the user to take a photo with the virtual object, which is conducive to enriching the form of media content associated with the virtual object, and further conducive to improving the efficiency of interaction between the user and the virtual object.

[0076] Example Process

[0077] Figure 3 FIG. 3 is a flow chart illustrating an example process 300 for generating media content according to some embodiments of the present disclosure. The process 300 may be implemented at the electronic device 110. Figure 1 3. The process 300 is described below.

[0078] like Figure 3 As shown, in box 310, the electronic device 110 presents first visual content corresponding to the first user and second visual content corresponding to the virtual object in the interactive interface, and the virtual object is configured to be jointly interacted by the first user and the second user.

[0079] In block 320 , the electronic device 110 generates media content in response to a capture operation received in the interactive interface, where the media content includes first visual content and second visual content.

[0080] In some embodiments, the first visual content includes a first virtual image created based on a configuration operation of the first user.

[0081] In this way, the embodiments of the present disclosure support the generation of media content based on the first virtual image and virtual objects created by the first user, thereby enriching the generation method of media content and facilitating improving the efficiency of generating media content.

[0082] In some embodiments, the first virtual image is created based on the following process: providing an upload portal to the first user; obtaining a first image associated with the first user through the upload portal; and creating the first virtual image associated with the first user based on the first image.

[0083] In this way, the embodiments of the present disclosure can support users in creating a first virtual image, and can provide users with a corresponding upload portal to help users create a virtual image, thereby helping to improve the efficiency of users in creating a virtual image.

[0084] In some embodiments, the first visual content further includes at least one item of dressing material associated with the first virtual image.

[0085] In this way, the embodiments of the present disclosure can support the presentation of the dressing material associated with the first virtual image in the media content, thereby facilitating improving the quality of the generated media content.

[0086] In some embodiments, the at least one item of dressing material includes a dressing material generated based on the second image uploaded by the first user.

[0087] In this way, the embodiments of the present disclosure can support the generation of dressing materials based on images uploaded by users, thereby enriching the styles of dressing materials and meeting the diverse needs of users in using dressing materials.

[0088] In some embodiments, process 300 further includes: presenting, in the interactive interface, third visual content corresponding to the second user, wherein the generated media content further includes the third visual content.

[0089] In this way, the embodiments of the present disclosure can support the generation of media content based on the third visual content corresponding to the second user, thereby facilitating enriching the style of the generated media content and improving the quality and efficiency of generating the media content.

[0090] In some embodiments, presenting the third visual content corresponding to the second user includes: in response to the second user being associated with the second virtual image, presenting the second virtual image as the third visual content in the interactive interface.

[0091] In this way, the embodiments of the present disclosure can present a virtual image corresponding to the second user in the interactive interface in response to the second user being associated with the virtual image, and support the use of the virtual image as the third visual content for generating media content, thereby helping to improve the quality and efficiency of generating media content.

[0092] In some embodiments, process 300 further includes: in response to the second user not being associated with the virtual image, presenting an invitation control associated with the second user; and in response to triggering the invitation control, sending an invitation message to the second user, the invitation message being configured to trigger the virtual image corresponding to the second user.

[0093] In this way, embodiments of the present disclosure can provide an invitation control for inviting the second user to create an avatar in response to the second user not being associated with the avatar, thereby facilitating improved interaction efficiency between the first user and the second user.

[0094] In some embodiments, process 300 further includes: presenting a configuration panel in the interactive interface; and adjusting a display state of the first visual content and / or the second visual content in the interactive interface based on a configuration operation received via the configuration panel.

[0095] In this way, the embodiments of the present disclosure can receive the user's configuration operations via the configuration panel, thereby adjusting the display status of the first visual content and / or the second visual content in the interactive interface, which is conducive to improving the quality and efficiency of generating media content, and can meet the diverse needs of users for generating media content.

[0096] In some embodiments, based on the configuration operation received via the configuration panel, adjusting the display status of the first visual content and / or the second visual content in the interactive interface includes: presenting a set of display controls associated with the first visual content and / or the second visual content in the configuration panel; based on the first configuration operation received via a set of display controls, determining whether the first visual content and / or the second visual content is presented in the media content.

[0097] In this way, the embodiments of the present disclosure can support users to configure whether the first visual content and / or the second visual content is presented in the media content based on the display control, thereby helping to improve the efficiency of generating media content and meet the diverse needs of users.

[0098] In some embodiments, based on the configuration operation received via the configuration panel, adjusting the display status of the first visual content and / or the second visual content in the interactive interface includes: presenting a set of mobile components associated with the first visual content and / or the second visual content in the configuration panel; and based on the second configuration operation received via the set of mobile components, determining the display position of the first visual content and / or the second visual content in the media content.

[0099] In this way, the embodiments of the present disclosure can support users to configure the display position of the first visual content and / or the second visual content in the media content based on mobile controls, thereby helping to improve the efficiency of generating media content and meet the diverse needs of users.

[0100] In some embodiments, based on a configuration operation received via a configuration panel, adjusting the display status of the first visual content and / or the second visual content in the interactive interface includes: presenting a set of size components associated with the first visual content and / or the second visual content in the configuration panel; and based on a third configuration operation received via a set of size components, determining the display size of the first visual content and / or the second visual content in the media content.

[0101] In this way, the embodiments of the present disclosure can support users to configure the display size of the first visual content and / or the second visual content in the media content based on the size component, thereby helping to improve the efficiency of generating media content and meet the diverse needs of users.

[0102] In some embodiments, the first visual content includes a first action associated with the first user; and / or the second visual content includes a second action associated with the virtual object.

[0103] In this way, the embodiments of the present disclosure can support the presentation of first visual content corresponding to a first action associated with a first user and second visual content corresponding to a second action associated with a virtual object in media content, thereby helping to improve the quality and efficiency of generating media content and meet the diverse needs of users.

[0104] In some embodiments, based on the configuration operation received via the configuration panel, adjusting the display status of the first visual content and / or the second visual content in the interactive interface includes: presenting a first group of action controls associated with the first user in the configuration panel; and in response to receiving a selection of a first action control in the first group of action controls, triggering the first visual content to present a first action associated with the first action control.

[0105] In this way, the embodiments of the present disclosure can support the user to determine the first action indicated by the first visual content via the action control, thereby facilitating improving the efficiency of generating media content and the user's operation efficiency.

[0106] In some embodiments, based on the configuration operation received via the configuration panel, adjusting the display status of the first visual content and / or the second visual content in the interactive interface includes: presenting an input control in the configuration panel; and in response to obtaining the input content of the first user via the input control, triggering the first visual content to present a first action associated with the input content.

[0107] In this way, the embodiments of the present disclosure can support users to indicate the first action by inputting content, which is conducive to improving the user's operating efficiency, enriching the action forms associated with the first visual content, and meeting the diverse needs of users.

[0108] In some embodiments, process 300 also includes: in response to the first action being triggered, triggering the virtual object to perform a second action associated with the first action; or in response to receiving a selection of a second action control in a second group of action controls associated with the virtual object, triggering the virtual object to perform a second action associated with the second action control.

[0109] In this way, embodiments of the present disclosure can support triggering a virtual object to perform a second action associated with the first action after the first action is triggered, thereby enhancing the correlation between the first and second actions. Furthermore, embodiments of the present disclosure can support the user to determine the second action of the virtual object based on the action control associated with the virtual object, thereby improving user operation efficiency.

[0110] In some embodiments, the process 300 further includes: dynamically acquiring image content of the first user via an image acquisition unit; and triggering, based on the image content, presentation of a third action corresponding to the image content in the first visual content.

[0111] In this way, the embodiments of the present disclosure can support determining the third action presented by the first visual content through the collected image content of the user, thereby enriching the action form associated with the first visual content and improving the quality of the generated media content and the user's interaction efficiency.

[0112] In some embodiments, generating the media content includes: presenting a set of shooting controls in an interactive interface; determining a set of shooting parameters via the set of shooting controls; and generating the media content based on the set of shooting parameters.

[0113] In this way, the embodiments of the present disclosure can support users in determining shooting parameters for generating media content using a set of shooting controls, thereby facilitating improving the quality and efficiency of generating media content.

[0114] In some embodiments, a set of shooting controls is associated with a virtual camera that generates media content, and the set of shooting controls includes at least one of the following: a first shooting control for configuring the orientation of the virtual camera; a second shooting control for configuring the position of the virtual camera; and a third shooting control for configuring the field of view of the virtual camera.

[0115] In this way, the embodiments of the present disclosure can provide users with a variety of shooting controls to determine different shooting parameters, thereby facilitating improving the efficiency and quality of generating media content and meeting the diverse needs of users.

[0116] In some embodiments, process 300 further includes: in response to a publishing request for the media content, publishing a media content item corresponding to the media content, wherein the viewing interface of the media content item is configured to provide an interactive portal associated with the virtual object.

[0117] In this way, the embodiments of the present disclosure can support users in publishing media content, and improve interactive entrances associated with virtual objects in the viewing interface of media content items associated with the media content, thereby helping to improve the efficiency of interaction between users and the efficiency of other users in generating media content.

[0118] In some embodiments, generating media content includes: generating picture content associated with the first visual content and the second visual content in response to a shooting operation associated with a photo taking mode; or generating video content associated with the first visual content and the second visual content in response to a shooting operation associated with a video recording mode.

[0119] In this way, the embodiments of the present disclosure can provide users with a photo mode and a video mode, thereby meeting the user's needs for taking photos or videos.

[0120] Example devices and equipment

[0121] The embodiments of the present disclosure also provide corresponding devices for implementing the above methods or processes. Figure 4 1 shows a schematic structural block diagram of an example apparatus 400 for generating media content according to some embodiments of the present disclosure. Apparatus 400 may be implemented as or included in electronic device 110. Each module / component in apparatus 400 may be implemented by hardware, software, firmware, or any combination thereof.

[0122] like Figure 4 As shown, the device 400 includes: a presentation module 410, configured to present first visual content corresponding to the first user and second visual content corresponding to the virtual object in an interactive interface, and the virtual object is configured to be jointly interacted by the first user and the second user; and a generation module 420, configured to generate media content in response to a shooting operation received in the interactive interface, and the media content includes the first visual content and the second visual content.

[0123] In some embodiments, the first visual content includes a first virtual image created based on a configuration operation of the first user.

[0124] In some embodiments, the first virtual image is created based on the following process: providing an upload portal to the first user; obtaining a first image associated with the first user through the upload portal; and creating the first virtual image associated with the first user based on the first image.

[0125] In some embodiments, the first visual content further includes at least one item of dressing material associated with the first virtual image.

[0126] In some embodiments, the at least one item of dressing material includes a dressing material generated based on the second image uploaded by the first user.

[0127] In some embodiments, the apparatus 400 further includes a visual presentation module, which is configured to present, in the interactive interface, third visual content corresponding to the second user, wherein the generated media content further includes the third visual content.

[0128] In some embodiments, the visual presentation module is further configured to: in response to the second user being associated with the second virtual image, present the second virtual image in the interactive interface as the third visual content.

[0129] In some embodiments, the device 400 also includes an invitation module, which is configured to: present an invitation control associated with the second user in response to the second user not being associated with the virtual image; and send an invitation message to the second user in response to triggering the invitation control, the invitation message being configured to trigger the virtual image corresponding to the second user.

[0130] In some embodiments, the device 400 also includes an adjustment module, which is configured to: present a configuration panel in the interactive interface; and adjust the display status of the first visual content and / or the second visual content in the interactive interface based on the configuration operation received via the configuration panel.

[0131] In some embodiments, the adjustment module is further configured to: present a set of display controls associated with the first visual content and / or the second visual content in a configuration panel; and determine whether the first visual content and / or the second visual content is presented in the media content based on a first configuration operation received via a set of display controls.

[0132] In some embodiments, the adjustment module is further configured to: present a set of mobile components associated with the first visual content and / or the second visual content in a configuration panel; and determine a display position of the first visual content and / or the second visual content in the media content based on a second configuration operation received via the set of mobile components.

[0133] In some embodiments, the adjustment module is further configured to: present a set of size components associated with the first visual content and / or the second visual content in a configuration panel; and determine the display size of the first visual content and / or the second visual content in the media content based on a third configuration operation received via the set of size components.

[0134] In some embodiments, the first visual content includes a first action associated with the first user; and / or the second visual content includes a second action associated with the virtual object.

[0135] In some embodiments, the adjustment module is further configured to: present a first set of action controls associated with the first user in the configuration panel; and in response to receiving a selection of a first action control in the first set of action controls, trigger the first visual content to present a first action associated with the first action control.

[0136] In some embodiments, the adjustment module is further configured to: present an input control in the configuration panel; and in response to obtaining input content of the first user via the input control, trigger the first visual content to present a first action associated with the input content.

[0137] In some embodiments, the device 400 also includes a first trigger module, which is configured to: in response to the first action being triggered, trigger the virtual object to perform a second action associated with the first action; or in response to receiving a selection of a second action control in a second group of action controls associated with the virtual object, trigger the virtual object to perform a second action associated with the second action control.

[0138] In some embodiments, the device 400 further includes a second trigger module, which is configured to: dynamically acquire image content of the first user via the image acquisition unit; and based on the image content, trigger presentation of a third action corresponding to the image content in the first visual content.

[0139] In some embodiments, the generation module 420 is further configured to: present a set of shooting controls in the interactive interface; determine a set of shooting parameters via the set of shooting controls; and generate media content based on the set of shooting parameters.

[0140] In some embodiments, a set of shooting controls is associated with a virtual camera that generates media content, and the set of shooting controls includes at least one of the following: a first shooting control for configuring the orientation of the virtual camera; a second shooting control for configuring the position of the virtual camera; and a third shooting control for configuring the field of view of the virtual camera.

[0141] In some embodiments, the device 400 further includes a publishing module, which is configured to: publish a media content item corresponding to the media content in response to a publishing request for the media content, wherein the viewing interface of the media content item is configured to provide an interactive entrance associated with the virtual object.

[0142] In some embodiments, the generation module 420 is further configured to: generate picture content associated with the first visual content and the second visual content in response to the shooting operation being associated with the photo mode; or generate video content associated with the first visual content and the second visual content in response to the shooting operation being associated with the video recording mode.

[0143] The units included in the device 400 can be implemented in various ways, including software, hardware, firmware, or any combination thereof. In some embodiments, one or more units can be implemented using software and / or firmware, such as machine executable instructions stored on a storage medium. In addition to or as an alternative to machine executable instructions, some or all of the units in the device 400 can be implemented at least in part by one or more hardware logic components. By way of example and not limitation, exemplary types of hardware logic components that can be used include field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), systems on chip (SOCs), complex programmable logic devices (CPLDs), and the like.

[0144] Figure 5 1 shows a block diagram of an electronic device 500 in which one or more embodiments of the present disclosure may be implemented. Figure 5 The illustrated electronic device 500 is merely exemplary and should not be construed as limiting the functionality and scope of the embodiments described herein. Figure 5 The electronic device 500 shown can be used to implement Figure 1 electronic device 110.

[0145] like Figure 5 As shown, electronic device 500 is in the form of a general electronic device. Components of electronic device 500 may include, but are not limited to, one or more processing units or processors 510, memory 520, storage device 530, one or more communication units 540, one or more input devices 550, and one or more output devices 560. Processor 510 may be a real or virtual processor and is capable of performing various processes according to programs stored in memory 520. In a multi-processor system, multiple processors execute computer-executable instructions in parallel to increase the parallel processing capabilities of electronic device 500.

[0146] The electronic device 500 typically includes a plurality of computer storage media. Such media can be any accessible media that can be obtained by the electronic device 500, including but not limited to volatile and non-volatile media, removable and non-removable media. The memory 520 can be a volatile memory (e.g., registers, cache, random access memory (RAM)), a non-volatile memory (e.g., read-only memory (ROM), electrically erasable programmable read-only memory (EEPROM), flash memory), or some combination thereof. The storage device 530 can be a removable or non-removable medium and can include a machine-readable medium, such as a flash drive, a disk, or any other medium that can be used to store information and / or data and can be accessed within the electronic device 500.

[0147] The electronic device 500 may further include additional removable / non-removable, volatile / non-volatile storage media. Figure 5 As shown in FIG, a magnetic disk drive for reading from or writing to a removable, non-volatile magnetic disk (e.g., a "floppy disk") and an optical disk drive for reading from or writing to a removable, non-volatile optical disk may be provided. In these cases, each drive may be connected to a bus (not shown) by one or more data media interfaces. Memory 520 may include a computer program product 525 having one or more program modules configured to perform various methods or actions of various embodiments of the present disclosure.

[0148] The communication unit 540 enables communication with other electronic devices via a communication medium. Additionally, the functions of the components of the electronic device 500 can be implemented in a single computing cluster or multiple computing machines that can communicate via a communication connection. Thus, the electronic device 500 can operate in a networked environment using a logical connection with one or more other servers, a network personal computer (PC), or another network node.

[0149] Input device 550 may be one or more input devices, such as a mouse, keyboard, or trackball. Output device 560 may be one or more output devices, such as a display, a speaker, or a printer. Electronic device 500 may also communicate with one or more external devices (not shown) via communication unit 540 as needed, such as a storage device, a display device, or the like, with one or more devices that allow a user to interact with electronic device 500, or with any device that allows electronic device 500 to communicate with one or more other electronic devices (e.g., a network card, a modem, etc.). Such communication may be performed via an input / output (I / O) interface (not shown).

[0150] According to an exemplary implementation of the present disclosure, a computer-readable storage medium is provided, on which computer-executable instructions are stored, wherein the computer-executable instructions are executed by a processor to implement the method described above. According to an exemplary implementation of the present disclosure, a computer program product is also provided, which is tangibly stored on a non-transitory computer-readable medium and includes computer-executable instructions, and the computer-executable instructions are executed by a processor to implement the method described above.

[0151] Various aspects of the present disclosure are described herein with reference to flowcharts and / or block diagrams of methods, apparatuses, devices, and computer program products implemented according to the present disclosure. It should be understood that each block of the flowcharts and / or block diagrams, and combinations of blocks in the flowcharts and / or block diagrams, can be implemented by computer-readable program instructions.

[0152] These computer-readable program instructions can be provided to a processor of a general-purpose computer, a special-purpose computer, or other programmable data processing device, thereby producing a machine such that when these instructions are executed by the processor of the computer or other programmable data processing device, a device is generated that implements the functions / actions specified in one or more blocks in the flowchart and / or block diagram. These computer-readable program instructions can also be stored in a computer-readable storage medium, where these instructions cause the computer, programmable data processing device, and / or other device to operate in a specific manner. Thus, the computer-readable medium storing the instructions comprises an article of manufacture that includes instructions for implementing various aspects of the functions / actions specified in one or more blocks in the flowchart and / or block diagram.

[0153] Computer-readable program instructions can be loaded onto a computer, other programmable data processing apparatus, or other device so that a series of operational steps are performed on the computer, other programmable data processing apparatus, or other device to produce a computer-implemented process, thereby causing the instructions executed on the computer, other programmable data processing apparatus, or other device to implement the functions / actions specified in one or more boxes in the flowchart and / or block diagram.

[0154] The flow charts and block diagrams in the accompanying drawings show the possible architecture, functions and operations of the systems, methods and computer program products according to multiple implementations of the present disclosure. In this regard, each box in the flow chart or block diagram can represent a part for a module, program segment or instruction, and a part for a module, program segment or instruction comprises one or more executable instructions for realizing the logical function of the specification. In some alternative implementations, the functions marked in the box can also occur in a sequence different from that marked in the accompanying drawings. For example, two continuous boxes can actually be executed substantially in parallel, and they can sometimes be executed in the opposite order, depending on the functions involved. It should also be noted that each box in the block diagram and / or flow chart, and the combination of the boxes in the block diagram and / or flow chart can be realized by a special hardware-based system that performs the function or action of the specification, or can be realized by a combination of special hardware and computer instructions.

[0155] While various implementations of the present disclosure have been described above, the foregoing description is intended to be illustrative, not exhaustive, and not limited to the disclosed implementations. Many modifications and variations will be apparent to those skilled in the art without departing from the scope and spirit of the described implementations. The terminology used herein is selected to best explain the principles of the implementations, their practical applications, or improvements to existing technologies, or to enable others skilled in the art to understand the various implementations disclosed herein.

Claims

1. A method for generating media content, comprising: In an interactive interface, presenting first visual content corresponding to a first user and second visual content corresponding to a virtual object, the virtual object being configured for joint interaction by the first user and the second user; as well as In response to a capture operation received in the interactive interface, media content is generated, where the media content includes the first visual content and the second visual content. 2 . The method according to claim 1 , wherein the first visual content comprises a first virtual image created based on a configuration operation of the first user.

3. The method according to claim 2, wherein the first virtual image is created based on the following process: Providing an upload portal to the first user; acquiring a first image associated with the first user via the upload portal; and The first avatar associated with the first user is created based on the first image. The method according to claim 2 , wherein the first visual content further comprises at least one item of dressing material associated with the first virtual image. 5 . The method according to claim 4 , wherein the at least one item of dressing material comprises a dressing material generated based on the second image uploaded by the first user.

6. The method according to claim 1, further comprising: In the interactive interface, third visual content corresponding to the second user is presented, wherein the generated media content also includes the third visual content.

7. The method of claim 6, wherein presenting third visual content corresponding to the second user comprises: In response to the second user being associated with a second virtual image, the second virtual image is presented in the interactive interface as the third visual content.

8. The method according to claim 7, further comprising: In response to the second user not being associated with an avatar, presenting an invitation control associated with the second user; as well as In response to triggering the invitation control, an invitation message is sent to the second user, where the invitation message is configured to trigger the virtual image corresponding to the second user.

9. The method according to claim 1, further comprising: presenting a configuration panel in the interactive interface; as well as Based on the configuration operation received via the configuration panel, the display status of the first visual content and / or the second visual content in the interactive interface is adjusted.

10. The method according to claim 9, wherein adjusting the display state of the first visual content and / or the second visual content on the interactive interface based on the configuration operation received via the configuration panel comprises: presenting, in the configuration panel, a set of display controls associated with the first visual content and / or the second visual content; Based on a first configuration operation received via the set of display controls, it is determined whether the first visual content and / or the second visual content is presented in the media content.

11. The method according to claim 9, wherein adjusting the display state of the first visual content and / or the second visual content on the interactive interface based on the configuration operation received via the configuration panel comprises: presenting in the configuration panel a set of mobile components associated with the first visual content and / or the second visual content; as well as Based on a second configuration operation received via the set of movement components, a display position of the first visual content and / or the second visual content in the media content is determined.

12. The method according to claim 9, wherein adjusting the display state of the first visual content and / or the second visual content on the interactive interface based on the configuration operation received via the configuration panel comprises: presenting, in the configuration panel, a set of dimension components associated with the first visual content and / or the second visual content; as well as A display size of the first visual content and / or the second visual content in the media content is determined based on a third configuration operation received via the set of size components.

13. The method according to claim 9, wherein: The first visual content includes a first action associated with the first user; and / or The second visual content includes a second action associated with the virtual object.

14. The method according to claim 13, wherein adjusting the display state of the first visual content and / or the second visual content on the interactive interface based on the configuration operation received via the configuration panel comprises: presenting a first set of action controls associated with the first user in the configuration panel; as well as In response to receiving a selection of a first action control in the first set of action controls, the first visual content is triggered to present the first action associated with the first action control.

15. The method according to claim 13, wherein adjusting the display state of the first visual content and / or the second visual content on the interactive interface based on the configuration operation received via the configuration panel comprises: presenting input controls in the configuration panel; as well as In response to obtaining input content of the first user via the input control, the first visual content is triggered to present the first action associated with the input content.

16. The method according to claim 13, further comprising: In response to the first action being triggered, triggering the virtual object to perform the second action associated with the first action; or In response to receiving a selection of a second action control in a second set of action controls associated with the virtual object, the virtual object is triggered to perform the second action associated with the second action control.

17. The method according to claim 1, further comprising: Dynamically acquiring image content of the first user via an image acquisition unit; as well as Based on the image content, a third action corresponding to the image content is triggered to be presented in the first visual content.

18. The method of claim 1 , wherein generating media content comprises: Presenting a set of shooting controls in the interactive interface; determining a set of shooting parameters via the set of shooting controls; as well as The media content is generated based on the set of capture parameters.

19. The method of claim 18, wherein the set of capture controls is associated with a virtual camera that generates the media content, the set of capture controls comprising at least one of: a first shooting control for configuring the orientation of the virtual camera; a second capture control for configuring the position of the virtual camera; A third capture control is used to configure a field of view of the virtual camera.

20. The method of claim 1, further comprising: In response to a publishing request for the media content, a media content item corresponding to the media content is published, wherein a viewing interface of the media content item is configured to provide an interactive entrance associated with a virtual object.

21. The method of claim 1 , wherein generating media content comprises: In response to the shooting operation being associated with a photo taking mode, generating picture content associated with the first visual content and the second visual content; or In response to the capturing operation being associated with a video recording mode, video content associated with the first visual content and the second visual content is generated.

22. An apparatus for generating media content, comprising: a presentation module configured to present, in an interactive interface, first visual content corresponding to a first user and second visual content corresponding to a virtual object, the virtual object being configured to be jointly interacted with by the first user and the second user; as well as The generating module is configured to generate media content in response to a shooting operation received in the interactive interface, where the media content includes the first visual content and the second visual content.

23. An electronic device comprising: at least one processor; as well as At least one memory, the at least one memory being coupled to the at least one processor and storing instructions for execution by the at least one processor, the instructions causing the electronic device to perform the method according to any one of claims 1 to 21 when executed by the at least one processor.

24. A computer-readable storage medium having a computer program stored thereon, wherein the computer program is executable by a processor to implement the method according to any one of claims 1 to 21.

Citation Information

Patent Citations

  • Interaction method and device based on augmented reality

    CN113112614A

  • Media content generation method and device, electronic equipment and storage medium

    CN115098099A

  • Interaction method and device based on virtual object, equipment, storage medium and product

    CN116170398A

  • Interaction method and device, equipment and storage medium

    CN118450073A

  • Information interaction method and device, electronic equipment and storage medium

    CN118524078A

Cited By

  • Interface interaction method and device, equipment and storage medium

    CN121326183A