Content generation method and device, equipment and storage medium

By automatically displaying image materials of the display medium and target items through an interactive page, media content is generated, solving the problems of high cost and long time consumption in traditional manual production, and realizing efficient and flexible content generation.

CN121563641APending Publication Date: 2026-02-24BEIJING DAJIA INTERNET INFORMATION TECH CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Application Number
CN202511678892.5
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-11-14
Publication Date
2026-02-24

AI Technical Summary

Technical Problem

The creation of content for traditional product displays relies on manual labor, resulting in high costs, long processing times, and poor flexibility.

Method used

This provides an interactive page that includes a dialogue area and an editing area. After the user enters item information in the dialogue area, the page automatically displays images of the display medium and the target item. The editing area generates media content, replacing the manual shooting and post-editing process.

Benefits of technology

It reduces the production cost and time of item display content, improves the flexibility and efficiency of content production, and allows users to quickly synthesize media content without the need for professional tools, thus lowering the barrier to creation.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN121563641A_ABST
    Figure CN121563641A_ABST
Patent Text Reader

Abstract

The embodiment of the invention discloses a content generation method and device, equipment and a storage medium. According to the main technical scheme, the method comprises the steps of displaying an interaction page, wherein the interaction page comprises a dialogue area and an editing area; in response to article information of a target article input in the dialogue area, triggering a display request instruction for the target article, and displaying at least one image material in the dialogue area, the image material comprising a display carrier and an image of the target article, the display carrier comprising a main body carrier and / or a scene carrier; in response to a selection operation on a target image material in the at least one image material, displaying the target image material in the editing area; and in response to a content generation trigger instruction detected in the editing area, obtaining media content corresponding to the target article based on the target image material. The creation threshold can be obviously reduced, and the content production efficiency is improved.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] This application relates to the field of Internet application technology, and in particular to a content generation method, apparatus, device and storage medium. Background Technology

[0002] With the rapid development of e-commerce, online shopping has become an indispensable part of people's lives. When shopping online, the content displayed for a product plays a crucial role in the consumer's purchasing decision.

[0003] In the production of product display content, the traditional model relies heavily on manual labor, including model photography, material processing, and post-editing, resulting in high costs, long processing times, and poor flexibility. Summary of the Invention

[0004] In view of this, this application provides a content generation method, apparatus, device, and storage medium.

[0005] This application provides the following solution: Firstly, a content generation method is provided, which includes: Display an interactive page, which includes a dialog area and an editing area; In response to inputting item information of a target item in the dialog area and triggering a display request instruction for the target item, at least one image material is displayed in the dialog area. The image material is an image including a display carrier and the target item, and the display carrier includes a main carrier and / or a scene carrier. In response to a selection operation of a target image material in the at least one image material, the target image material is displayed in the editing area; In response to the detection of a content generation trigger command in the editing area, media content corresponding to the target item is obtained based on the target image material.

[0006] Optionally, the item information is the identification information of the target item; then, in response to inputting the item information of the target item in the dialog area and triggering a display request instruction for the target item, at least one image material is displayed in the dialog area, including: In response to inputting the identification information of the target item in the dialog area and triggering the display request instruction, the type of the target item and at least one corresponding item image are determined based on the identification information, and the item image and at least one carrier generation entry corresponding to the type are displayed in the dialog area; In response to the selection operation of the target carrier generation entry in the at least one carrier generation entry, at least one carrier image is displayed in the dialog area, the carrier image containing a display carrier corresponding to the target carrier generation entry; In response to a selection operation of a target item image in the at least one item image and a target carrier image in the at least one carrier image, the image material is displayed in the dialog area, the image material including the target item image and the target carrier image.

[0007] Optionally, the item information is a target item image of the target item; In response to inputting item information of a target item in the dialog area and triggering a display request instruction for the target item, at least one image asset is displayed in the dialog area, including: In response to inputting a target item image in the dialog area and triggering the display request instruction, the type of the target item is determined, and at least one carrier generation entry corresponding to the type is displayed in the dialog area; In response to the selection operation of the target carrier generation entry in the at least one carrier generation entry, at least one carrier image is displayed in the dialog area, the carrier image containing a display carrier corresponding to the target carrier generation entry; In response to a selection operation of a target carrier image in the at least one carrier image, the image material is displayed in the dialog area, the image material including the target item image and the target carrier image.

[0008] Optionally, the carrier generation entry includes: a main carrier entry and / or a scene carrier entry.

[0009] Optionally, the target carrier generation entry is the scene carrier entry, and the display carrier corresponding to the target carrier generation entry is the scene carrier; then The display of the image material in the dialogue area includes: The target item image is loaded into the target carrier image according to a preset posture to generate the image material, which is then displayed in the dialogue area.

[0010] Optionally, the target carrier generation entry is the main carrier entry, and the display carrier corresponding to the target carrier generation entry is the main carrier; then The display of the image material in the dialogue area includes: The target item image and the target carrier image are adapted to the usage scenario to generate the image material, which is then displayed in the dialogue area.

[0011] Optionally, the step of displaying at least one carrier image in the dialog area in response to a selection operation of a target carrier generation entry in the at least one carrier generation entry includes: In response to the selection operation of the target carrier generation entry, at least one scene carrier image with a preset style is generated; The method further includes: A style selection control is displayed in the dialog area, and the style selection control contains multiple style options; In response to the user's selection of a style option in the style selection control that is different from the preset style, a new scene carrier image that matches the style selected by the user is determined, and the new scene carrier image is updated and displayed in the dialog area.

[0012] Optionally, the step of displaying at least one carrier image in the dialog area in response to a selection operation of a target carrier generation entry in the at least one carrier generation entry includes: In response to the selection operation of the target carrier generation entry, a main carrier image with at least one preset attribute is generated; The method further includes: A subject attribute selection control is displayed in the dialog area, and the subject attribute selection control contains multiple attribute options; In response to a user's selection of an attribute option in the subject attribute selection control that is different from the preset attribute, a new subject carrier image that matches the attribute selected by the user is determined, and the new subject carrier image is updated and displayed in the dialog area.

[0013] Optionally, the method further includes: In response to the arrangement operation of the target image material in the editing area, update the existence state and / or arrangement order of the target image material.

[0014] Optionally, after displaying the target image material in the editing area, the method further includes: In response to triggering an action description input operation for any target image material in the editing area, an action description input pop-up window is displayed, which is used to input target action description text for the triggered target image material; The step of obtaining the media content corresponding to the target item based on the target image material includes: Based on the target image material and the target action description text of the triggered target image material, the media content corresponding to the target item is obtained, wherein the display of the target item in the media content conforms to the target action description text of the triggered target image material.

[0015] Optionally, the editing area may further include a text input area; The step of obtaining the media content corresponding to the target item based on the target image material includes: Based on the target image material and the text information of the target item in the text input area, the media content corresponding to the target item is obtained.

[0016] Optionally, the method further includes: In response to a text generation instruction for the target item input in the dialog area, a script for the target item is displayed in the dialog area, the script including content framework information; In response to a selection operation on the script, text information of the target item is displayed in the text input area, the text information being generated based on the content framework information.

[0017] Optionally, the dialog area may also display a refresh component corresponding to the script; The method further includes: In response to a trigger operation on the refresh component, a new script is displayed in the dialog area, the new script containing new content frame information; In response to the selection of the new script, the text information of the target item generated based on the new content framework information is displayed in the text input area.

[0018] Optionally, the method further includes: In response to an editing operation on the text information in the text input area, the content of the text information is updated.

[0019] Optionally, the editing area includes at least one composition mode control, each of which corresponds to a preset composition mode; then the presentation form of the media content is determined based on the composition mode corresponding to the selected composition mode control.

[0020] Optionally, the editing area further includes a prompt input area for inputting prompt information. This prompt information is a natural language description of the user's video generation requirements, and the natural language description includes at least one or more of the following: the action of the display carrier, camera movement, or scene atmosphere; and the selected compositing mode control corresponds to a video compositing mode. The process of obtaining the media content based on the target image material includes: The media content is obtained based on the video generation mode, the prompt information, and the target image material.

[0021] Optionally, the editing area may further include content wrapping controls; The process of obtaining the media content corresponding to the target item includes: In response to the selection of the content packaging control, preset enhancement effects are added to the media content during the generation of the media content. Secondly, a content generation apparatus is provided, the apparatus comprising: A page display unit is configured to display an interactive page, which includes a dialog area and an editing area; An event listening unit is configured to listen for events on the interactive page; The page display unit is further configured to, in response to inputting item information of a target item in the dialog area and triggering a display request instruction for the target item, display at least one image material in the dialog area, the image material being an image including a display carrier and the target item, the display carrier including a main carrier and / or a scene carrier; in response to a selection operation of a target image material among the at least one image material, display the target image material in the editing area; and in response to detecting a content generation instruction in the editing area, obtain media content corresponding to the target item based on the target image material. Thirdly, a computer-readable storage medium is provided having a computer program stored thereon, which, when executed, implements the steps of the method described in the first aspect.

[0022] Fourthly, an electronic device is provided, comprising: one or more processors; and a memory associated with the one or more processors, the memory being used to store program instructions that, when read and executed by the one or more processors, perform the steps of the method described in the first aspect.

[0023] Fifthly, a computer program product is provided, including a computer program that, when executed, implements the steps of the method described in the first aspect.

[0024] According to the specific embodiments provided in this application, the following technical effects are disclosed: 1) This application provides an interactive page containing a dialogue area and an editing area. After the user enters the item information of the target item in the dialogue area and triggers the display request instruction for the target item, image materials including the display carrier and the target item can be automatically displayed in the dialogue area. The display carrier includes the main carrier and / or the scene carrier. After the user selects the target image material, the media content corresponding to the target item is automatically generated through the content generation instructions in the editing area. This method replaces the traditional model that relies on manual processes such as model shooting, material processing, and post-editing. While reducing the production cost and shortening the production cycle of the item display content, it improves the flexibility of content production. Users can quickly synthesize the target item and the display carrier to obtain media content without professional tools, which significantly lowers the creative threshold and improves the efficiency of content production.

[0025] 2) This application proposes to generate image materials by inputting the identification information of the target item. This method can directly link to the item database, determine the accuracy of item type identification and image retrieval, and avoid mismatch problems caused by image recognition errors. Moreover, users do not need to upload images and can quickly trigger the process based on the identification information, reducing user operation steps.

[0026] 3) This application proposes to generate image materials by uploading images of target items. This method can trigger subsequent processes without inputting identification information, reducing the user's operational threshold and eliminating the need to rely on a database.

[0027] 4) This application achieves automatic adaptation of target items and scenes by pre-setting posture rules, which can avoid the tediousness of manual adjustment, ensure that the placement angle and proportion of target items in the scene conform to visual logic, and improve the efficiency and aesthetics of material generation.

[0028] 5) This application achieves dynamic switching of scene carriers through style selection controls, which not only retains the preset style as a quick reference, but also supports users to customize as needed, meet the display needs of the same item in different style scenes, and enhance the diversity of materials and the flexibility of scene adaptation.

[0029] 6) This application can quickly respond to users' needs for differentiated copywriting by setting the corresponding refresh component for the script in the dialogue area, and provide users with multiple copywriting directions; after the user selects a new script, the adapted copy generated based on the new framework can be automatically obtained, which not only avoids the tedious process of users coming up with copywriting from scratch, but also meets the copywriting adaptation needs of different scenarios (such as different platform promotion and different audience reach) through the dynamic update of the framework, greatly improving the flexibility and efficiency of copywriting generation, while reducing the operation threshold for non-professional users.

[0030] Of course, any product implementing this application does not necessarily need to achieve all of the advantages described above at the same time. Attached Figure Description

[0031] To more clearly illustrate the technical solutions in the embodiments of this application or the prior art, the drawings used in the embodiments will be briefly introduced below. Obviously, the drawings described below are only some embodiments of this application. For those skilled in the art, other drawings can be obtained based on these drawings without creative effort.

[0032] Figure 1 This is a schematic diagram of the system architecture applicable to the embodiments of this application.

[0033] Figure 2 A flowchart illustrating the content generation method provided in this application embodiment.

[0034] Figure 3 This is a schematic diagram of the initial interface of the interactive page provided in the embodiments of this application.

[0035] Figure 4 This is a schematic diagram of an interactive page in a scenario where a user inputs identification information, as provided in an embodiment of this application.

[0036] Figure 5 This is a schematic diagram of an interactive page in a user input target object image scenario provided in an embodiment of this application.

[0037] Figure 6 This is a schematic diagram of the editing area provided in an embodiment of this application.

[0038] Figure 7 A schematic block diagram of a content generation apparatus provided in an embodiment of this application.

[0039] Figure 8 A schematic block diagram of an electronic device provided in an embodiment of this application. Detailed Implementation

[0040] The technical solutions of the embodiments of this application will be clearly and completely described below with reference to the accompanying drawings. Obviously, the described embodiments are only some embodiments of this application, and not all embodiments. All other embodiments obtained by those skilled in the art based on the embodiments of this application are within the scope of protection of this application.

[0041] The terminology used in the embodiments of this application is for the purpose of describing particular embodiments only and is not intended to be limiting of this application. The singular forms “a,” “the,” and “the” used in the embodiments of this application and the appended claims are also intended to include the plural forms unless the context clearly indicates otherwise.

[0042] It should be understood that the term "and / or" used in this article is merely a description of the relationship between related objects, indicating that three relationships can exist. For example, A and / or B can represent: A existing alone, A and B existing simultaneously, and B existing alone. Additionally, the character " / " in this article generally indicates that the preceding and following related objects have an "or" relationship.

[0043] Depending on the context, the word "if" as used here can be interpreted as "when," "when," "in response to determination," or "in response to detection." Similarly, depending on the context, the phrase "if determination" or "if detection (of the stated condition or event)" can be interpreted as "when determination," "in response to determination," "when detection (of the stated condition or event)," or "in response to detection (of the stated condition or event)."

[0044] In the production of product display content, the traditional model relies heavily on manual labor, including model photography, material processing, and post-editing, resulting in high costs, long processing times, and poor flexibility.

[0045] In view of this, this application provides a new approach. To facilitate understanding of this application, the system architecture on which this application is based will first be described. Figure 1 This is a schematic diagram of the system architecture applicable to the embodiments of this application, such as... Figure 1 As shown, the system architecture may include: user terminal, terminal device and server terminal.

[0046] In this embodiment, the user terminal is set on the terminal device. The user terminal involved in this application can be a client running on the terminal device, a mini-program, or a web application running through a browser.

[0047] Terminal devices can include, but are not limited to, smart mobile terminals, wearable devices, and PCs (Personal Computers). Smart mobile devices can include devices such as mobile phones, tablets, laptops, PDAs (Personal Digital Assistants), and connected car terminals. Wearable devices can include devices such as smartwatches, smart glasses, smart bracelets, VR (Virtual Reality) devices, AR (Augmented Reality) devices, and mixed reality devices (i.e., devices that support both virtual and augmented reality), etc.

[0048] The user can interact with the server via the network. The server can be a single server, a server cluster consisting of multiple servers, or a cloud server. A cloud server, also known as a cloud computing server or cloud host, is a hosting product in the cloud computing service system, designed to address the shortcomings of traditional physical hosts and Virtual Private Servers (VPS) services, such as high management difficulty and weak service scalability.

[0049] One possible approach is for the user to trigger relevant commands for a target item through an interactive page, thereby obtaining the media content of the target item from the server.

[0050] It should be understood that Figure 1 The number of client terminals, terminal devices, and servers shown is merely illustrative. Depending on implementation needs, there can be any number of client terminals, terminal devices, and servers.

[0051] Figure 2 This is a flowchart of a content generation method provided in an embodiment of this application. This method can be... Figure 1 The user-side execution in the system shown is as follows. Figure 2 As shown, the method may include the following steps: Step 201: Display the interactive page, which includes a dialog area and an editing area. Step 202: In response to the input of the target item information in the dialog area and the triggering of a display request instruction for the target item, at least one image material is displayed in the dialog area. The image material is an image including the display carrier and the target item. The display carrier includes the main carrier and / or the scene carrier.

[0052] Step 203: In response to the selection operation of the target image material in at least one image material, display the target image material in the editing area. Step 204: In response to the detection of a content generation instruction in the editing area, obtain the media content corresponding to the target item based on the target image material.

[0053] As can be seen from the above process, this application provides an interactive page containing a dialogue area and an editing area. After the user enters the item information of the target item in the dialogue area and triggers the display request command for the target item, image materials including the display carrier and the target item can be automatically displayed in the dialogue area. The display carrier includes the main carrier and / or the scene carrier. After the user selects the target image material, the media content corresponding to the target item is automatically generated through the content generation command in the editing area. This method replaces the traditional model that relies on manual processes such as model shooting, material processing, and post-editing. While reducing the production cost and shortening the production cycle of the item display content, it improves the flexibility of content production. Users can quickly synthesize the target item and the display carrier to obtain media content without professional tools, which significantly lowers the creative threshold and improves the efficiency of content production.

[0054] The following describes in detail each step of the above process and the effects that can be further produced, with reference to the embodiments.

[0055] First, the above step 201, namely "displaying an interactive page, which includes a dialogue area and an editing area", will be described in detail with reference to the embodiments.

[0056] In this embodiment, the interactive page can be a page integrating an intelligent agent, where users can interact with the intelligent agent to generate media content related to objects. An intelligent agent is a software entity possessing autonomous perception, decision-making, and execution capabilities. In the field of artificial intelligence, intelligent agents are also translated as intelligent agents, proxies, agents, intelligent subjects, etc.

[0057] The interactive page may include, but is not limited to, a dialogue area and an editing area, both of which interact with the intelligent agent in real time to generate media content.

[0058] Interactive pages can adopt a visual layout with left and right columns or top and bottom sections, clearly dividing the dialogue area and the editing area. The two areas can be linked in real time to ensure the continuity of user operations.

[0059] It should be noted that the interactive page in this application embodiment can integrate only a single intelligent agent or multiple intelligent agents with different functions. When multiple intelligent agents with different functions are integrated, the corresponding intelligent agent can be determined first according to the user's needs, and then the user can interact with the determined intelligent agent to realize the user's needs. The following describes in detail step 202, namely, "in response to inputting item information of a target item in the dialog area and triggering a display request instruction for the target item, displaying at least one image material in the dialog area, wherein the image material is an image including a display carrier and the target item, and the display carrier includes a main carrier and / or a scene carrier," with reference to the embodiments.

[0060] In this embodiment of the application, after the user enters the item information of the target item in the dialogue area of ​​the interactive page, a display request instruction for the target item is triggered. Through the processing of the intelligent agent, at least one image material is displayed in the dialogue area. The image material refers to an image including the display carrier and the target item.

[0061] The display medium can include a main medium and / or a scene medium. The main medium is used to create spatial contact with the target item to present the target item in its usage state, such as a mannequin, a hand-touchable object, or a pet mannequin; while the scene medium is used to be in the same virtual space as the target item to present the scene effect of the target item, such as indoor scenes (Nordic living room, minimalist study), outdoor scenes (beach, street, forest), and virtual scenes (anime campus, cyberpunk city).

[0062] By combining the display medium with the target item, the usage scenario or appearance of the target item can be presented intuitively, thus replacing the traditional manual shooting of related display materials.

[0063] The aforementioned image materials can be generated and displayed directly based on a display request command, or a carrier image containing the display medium can be generated first and displayed in the dialog area, and then the image materials can be generated and displayed based on the user's selection. The second method can enhance the user's autonomy over the image materials.

[0064] As one possible approach, the item information input by the user is the identification information of the target item. The identification information is used to uniquely identify the item and may include, but is not limited to, item ID, barcode, etc.

[0065] In this scenario, when a user enters the identification information of a target item in the dialog area and triggers a display request command, the type of the target item is first determined based on the identification information, and at least one corresponding item image is retrieved and displayed in the dialog area. At the same time, based on the type of the target item, at least one carrier generation entry corresponding to the type of the target item is displayed in the dialog area. This entry can be presented in the form of an interactive card, a pagination tab, or a drop-down menu.

[0066] The user selects a target carrier generation entry from at least one carrier generation entry. At this time, at least one carrier image is displayed in the dialog area. The carrier image contains the display carrier corresponding to the target carrier generation entry. Then, the user selects a target item image from at least one displayed item image and a target carrier image from at least one carrier image. At least one image material is generated based on these two target images and displayed in the dialog area.

[0067] As another possible approach, the user inputs an image of the target item. In this case, when the user inputs the image of the target item in the dialog area, a display request instruction is triggered. Based on the image, the type of the target item is determined, and at least one carrier generation entry corresponding to that type is displayed in the dialog area. The user selects a target carrier generation entry from the at least one carrier generation entry, at which point at least one carrier image is displayed in the dialog area. This carrier image contains a display carrier corresponding to the target carrier generation entry. Then, the user selects a target carrier image from the at least one displayed carrier images. Based on the selected target carrier image and the target item image, at least one image asset is generated and displayed in the dialog area.

[0068] Furthermore, in order to more accurately reflect user needs, users can select the item category information when uploading images of the target item, thereby directly determining the type of the target item based on the item category information, thus improving recognition efficiency.

[0069] Item category information can be presented to the user via an item category selection box. This could include primary categories such as tops, trousers, shoes, and cosmetics, as well as secondary subcategories such as short-sleeved shirts, T-shirts, and jeans. Alternatively, the user can actively input item category information in a dialog area. This application does not limit the method of inputting item category information.

[0070] Through the two implementation methods described above, users can efficiently complete the operation from target object recognition to customized image material generation, meeting diverse needs such as object display and material production.

[0071] In the first method described above, the identification information of the target item can be directly linked to the item database to determine the accuracy of item type identification and image retrieval, avoiding mismatch problems caused by image recognition errors. Moreover, users do not need to upload images and can quickly trigger the process based on the identification information, reducing user operation steps. This method is suitable for scenarios with existing standardized item databases. The retrieved item images are pre-set standardized materials in the database, which can maintain style consistency and improve the professionalism of the materials when combined with the display carrier.

[0072] In the second method described above, users do not need to input identification information; they can simply upload an image of the target item to trigger the subsequent process, which lowers the operational threshold and does not rely on a database, making it applicable to a wider range of scenarios.

[0073] In practical applications, users can choose any implementation method to generate and display image materials according to their actual needs.

[0074] In the two feasible methods described above, the carrier generation entry is a functional entry point used to define the display format of the item. It can include the main carrier entry point and / or the scene carrier entry point, and common types include but are not limited to: Scene generation (entry point for scene carrier): Matches intelligent scene backgrounds to target items to enhance the display quality.

[0075] AI try-on (as the main entry point): Provides a virtual try-on platform to simulate the wearing effect of the target item on the human body.

[0076] Handheld Item (as the main carrier entry point): Generates a display carrier in which a person holds the target item, highlighting the details and feel of the target item.

[0077] When the target carrier generation entry is the scene carrier entry, the display carrier corresponding to the target carrier generation entry is the scene carrier; then when displaying image materials in the dialogue area, the target item image is loaded into the target carrier image according to the preset posture to generate image materials, and displayed in the dialogue area.

[0078] By setting up posture rules to automatically adapt target items to the scene, the tedious manual adjustment can be avoided, ensuring that the placement angle and proportion of the target items in the scene conform to visual logic (such as furniture items automatically fitting to the scene floor and ornaments naturally embedding into the tabletop), thus improving the efficiency and aesthetics of material generation.

[0079] Furthermore, when a user selects a target carrier generation entry from at least one carrier generation entry, at least one scene carrier image with a preset style is generated and displayed in the dialog area. Simultaneously, a style selection control is displayed in the dialog area, containing multiple style options. When the user selects a style option different from the preset styles from the style selection control, a new scene carrier image matching the user's selected style is determined, and the new scene carrier image is updated and displayed in the dialog area.

[0080] The style selection control enables dynamic switching of scene carriers, retaining preset styles for quick reference while also supporting user customization to meet the display needs of the same item in different style scenes (such as a water cup that can be quickly switched to "Nordic kitchen" or "Japanese tea room" scenes), enhancing the diversity of materials and the flexibility of scene adaptation.

[0081] When the target carrier generation entry is the main carrier entry, the display carrier corresponding to the target carrier generation entry is the main carrier; then when displaying image materials in the dialogue area, the target item image and the target carrier image are adapted to the usage scenario to generate image materials, which are then displayed in the dialogue area.

[0082] By using scene adaptation logic (such as the interaction angle between the object and the subject, and the fit calibration), the combination of the object and the subject carrier is made more in line with the real usage scenario (such as the alignment of the glasses with the bridge of the virtual face, and the natural grip angle between the handheld object and the palm), avoiding the incongruity of the synthesized image and enhancing the realism and immersion of the material.

[0083] Furthermore, when a user selects a target carrier generation entry from at least one carrier generation entry, a main carrier image with at least one preset attribute is generated and displayed in the dialog area. Simultaneously, a main attribute selection control containing multiple attribute options is displayed in the dialog area. When the user selects an attribute option from the main attribute selection control that differs from the preset attributes, a new main carrier image matching the user's selected attribute is determined, and the new main carrier image is updated and displayed in the dialog area.

[0084] The subject attribute selection control supports personalized customization of the carrier, providing a basic reference with preset attributes and allowing users to adjust as needed (such as switching the height and body shape of the human body in virtual try-on, or adjusting the skin color and posture of the hands in a handheld scene), so that the same item can be adapted to diverse subject characteristics, expanding the applicable scenarios of the material (such as product display for different groups of people) and increasing the user's freedom to customize the material.

[0085] The following describes step 203, namely, "in response to the selection operation of a target image material in at least one image material, display the target image material in the editing area," in detail with reference to the embodiments.

[0086] After at least one image is displayed in the dialog area, the user can select one or more target image assets from the at least one image asset in the dialog area.

[0087] Specifically, the selection of target image materials can be achieved through visual markers. For example, a selection button can be placed in a prominent position on the image material displayed in the dialog area. When the user clicks the selection button, the button switches to a "selected" state, and a border marker appears around the edge of the image material. Additionally, a deselect icon (×) can be placed in the upper right corner of the image material, allowing the user to deselect the image material by clicking it.

[0088] In addition to the visual marking methods mentioned above, the selection of target image materials can also be achieved through micro-animation effects (a slight upward floating effect occurs after a certain image material is selected) and text prompts (such as a prompt message popping up at the bottom of the dialog area: "Selected material XX").

[0089] It should be noted that the selection operation of the target item image and the target carrier image in step 202 can also be implemented based on the above principle.

[0090] By displaying the target image material in the editing area, a seamless connection can be achieved between the material generation stage and the content editing stage.

[0091] Furthermore, the target image materials displayed in the editing area can be further arranged. When a user's arrangement operation on the target image materials in the editing area is detected, the existence status and / or arrangement order of the target image materials will be updated in real time. The arrangement operation here may include, but is not limited to: adding target image materials, deleting a target image material, adjusting the arrangement order of target image materials, etc.

[0092] In the specific layout, the operation of adding target image materials can be set with an "Add Material" button in the editing area, for example, with an icon of "+" combined with an image style. When the user clicks it, two paths for adding the material can be provided for the user to choose from: Local import: A local file selection window pops up, supporting multiple image formats. Users can directly drag and drop local files to the blank space of the editing area to trigger quick import. Material library reuse: If users need to reuse historical materials generated in the dialog area, they can jump to the "Material History" panel via the button and select the generated materials to synchronize to the editing area with one click.

[0093] To delete a target image, a delete icon (such as "×") can be set in the upper right corner of each target image in the editing area. Users can click on it to delete the corresponding target image.

[0094] Adjusting the order of target image materials can be done in the editing area by long-pressing and dragging.

[0095] The specific layout methods of the above-mentioned various arrangement operations can be diverse, and there are no restrictions on them in the embodiments of this application.

[0096] The following describes step 204, namely, "in response to detecting a content generation instruction in the editing area, obtaining the media content corresponding to the target item based on the target image material," in detail with reference to the embodiments.

[0097] When a content generation instruction is detected in the editing area, such as setting a content generation button in the editing area, the agent generates corresponding media content based on the target image material when the user triggers the content generation button.

[0098] Furthermore, after displaying the target image material in the editing area, an action description input operation for any target image material can be triggered in the editing area, displaying an action description input pop-up window. In this pop-up window, the user inputs the target action description text for the triggered target image material. At this point, based on the target image material and the target action description text of the triggered target image material, the media content corresponding to the target item is obtained. The display of the target item in this media content should conform to the target action description text of the triggered target image material.

[0099] For example, a user selects a target image of a person holding a coffee cup in the editing area and triggers an action description input operation. In the pop-up window, the user enters the target action description text: "The coffee cup is slowly raised to the lips, the cup is slightly tilted, and a small amount of steam rises." Based on this target image and action description text, the intelligent agent generates dynamic media content: In the scene, the action of holding the coffee cup gradually transitions from the initial posture to the state of raising it to the lips, the cup tilts naturally with the action, and special effects simulate the dynamic effect of steam rising. This makes the display of the coffee cup perfectly match the user's input action description, presenting a coherent and logically consistent dynamic scene, allowing the media content to more vividly convey the usage scenario and details of the item.

[0100] By combining target image materials with user-defined action description text, the generated media content breaks through the limitations of static images. It can not only retain the basic visual connection between the object and the carrier, but also enhance the narrative and appeal of the content through action-oriented presentation, meet users' needs for dynamic display of objects and contextualized expression, and improve the expressiveness and customization accuracy of media content.

[0101] Furthermore, to ensure that the generated media content more accurately meets user needs, in addition to the target image material, it can also incorporate the textual information of the target item. Therefore, a text input area can be set in the editing area, where users can input the textual information of the target item. This allows the intelligent agent to generate media content corresponding to the target item based on the target image material and the textual information of the target item in the text input area.

[0102] Furthermore, the text information for the target item can be generated automatically, in addition to being based on user input, using the interactive area. Specifically, when an input command to generate text for the target item is detected in the dialog area, the script for the target item, which includes content framework information, is displayed in the dialog area; in response to a selection operation on the script, the text information for the target item generated based on the content framework information is displayed in the text input area of ​​the editing area.

[0103] The content framework information consists of pre-set structured templates and creative guidelines for the intelligent agent to manage different scenarios and item types. It typically comprises three parts: module information, text framework, and element prompts.

[0104] Depending on the promotional context of the product (such as e-commerce main images, short video narration, and product detail page descriptions), different functional modules can be broken down to ensure that the copywriting information covers the key information needed for user decision-making. Common modules include core selling point modules and scenario-adaptation modules; the copywriting framework further clarifies the order between these modules to avoid disorganized copywriting information; to avoid the module framework being too vague, each module will also include element prompts, that is, which specific dimensions the module should focus on. For example, the element prompts for the core selling point module in the lipstick copywriting framework could be [shade (e.g., mauve), texture (e.g., velvet), efficacy (e.g., long-lasting and non-slipping)], etc.

[0105] Furthermore, a refresh component corresponding to the script can be displayed in the dialog area. When the refresh component is triggered, a new script is displayed in the dialog area. This new script contains new content framework information. When the user selects the new script, the text information generated based on the new content framework information is displayed in the text input area.

[0106] By setting the corresponding refresh component for the script in the dialogue area, the system can quickly respond to users' needs for differentiated copywriting and provide them with diverse copywriting directions. After selecting a new script, users can automatically obtain adapted copy generated based on the new framework. This avoids the tedious process of users coming up with copywriting ideas from scratch and meets the copywriting adaptation needs of different scenarios (such as promotion on different platforms and reaching different audiences) through dynamic updates of the framework. This greatly improves the flexibility and efficiency of copywriting generation and lowers the operating threshold for non-professional users.

[0107] Furthermore, the text information displayed in the text input area, generated based on the content framework, is not fixed and unchangeable. Instead, it allows users to flexibly edit it according to their actual needs, enabling personalized adjustments to the text content. Specifically, the text input area can be in the form of an editable text box, allowing users to directly modify the text content, expression style, and key information—for example, adding or deleting core selling points (such as adding "supports wireless charging" to the headphone text), adjusting the tone (changing "buy now" to the more formal "welcome to purchase"), and refining detailed descriptions (refining "30 hours of battery life" to "8 hours of battery life on a single charge, 3 recharges with the charging case, for a total of 32 hours of battery life"). During the editing process, the system saves the input content in real time and provides "undo" and "redo" functions, as well as basic formatting tools (such as bolding and line breaks), facilitating user corrections and format optimization. This editable design retains the efficiency of script-based copy generation while allowing users to edit the text themselves to compensate for potential content deviations in automated generation. It combines "standardized generation" with "personalized customization," ensuring that the final copy is more aligned with the promotional needs of the target product and the user's expression habits.

[0108] Furthermore, the editing area may also include at least one composition mode control, each composition mode control corresponding to a preset composition mode; the composition modes here may include video composition mode and image carousel mode.

[0109] The so-called video compositing mode refers to a compositing method that strings together target image materials within an editing area according to a preset sequence logic to create dynamic video content. In this mode, transition animations, duration allocations, and audio adaptations are automatically added to the target image materials, ultimately outputting a smooth and coherent short video product, suitable for dynamic display scenarios such as product promotion and scene demonstrations.

[0110] Image carousel mode refers to a composite method that displays target image materials in a cyclical manner according to a set order, similar to the image carousel effect in PowerPoint. This method is suitable for scenarios that dynamically present static materials, such as multi-angle displays of items or pages with mixed text and images.

[0111] In this case, the presentation of the generated media content is determined based on the compositing mode corresponding to the selected compositing mode control.

[0112] Furthermore, the editing area may also include a prompt input area for inputting prompt information. This prompt information is a natural language description of the user's video generation needs, which includes at least one or more of the following: the movement of the display device, camera movement, or scene atmosphere. For example, the prompt information may include: slowly rotating a short-sleeved item in the frame to show details of the collar and cuffs (display device movement), panning the camera from left to right to show three short-sleeved shirts of different colors in sequence (camera movement), creating a refreshing and vibrant summer atmosphere (scene atmosphere), etc.

[0113] When the selected compositing mode control corresponds to the video compositing mode, media content can be obtained based on the prompt information and target image materials. The prompt information can influence the dynamic effects, virtual / real logic, and style tone of the media content.

[0114] Furthermore, the editing area can also include content wrapping controls. When a content wrapping control is selected, preset enhancement effects (such as stickers or sound effects) can be added to the media content during the media content generation process.

[0115] The methods described above (whether it's the text input area, the composition mode control, or the content packaging control) can all enable the generated media content to more accurately match user needs, thereby improving the quality of the generated media content. In practical applications, these methods can be flexibly and without restriction combined according to different creative goals and scenario requirements.

[0116] The following section uses a certain commodity as an example to introduce the implementation process of the content generation method provided in this application embodiment in practical applications.

[0117] like Figure 3 The image shows the initial interface of the interactive page. This page includes a dialog area and an editing area. It features two options: "Enter Product ID" and "Upload Locally." When the user selects "Enter Product ID," the corresponding input template is displayed in the input area of ​​the dialog area, such as "The product I want to change is: Enter Product Number." The user can then enter the identification information of the target product. When the user selects "Upload Locally," the corresponding upload button is displayed in the input area of ​​the dialog area, allowing the user to upload an image of the target product.

[0118] It should be noted that when the fill color of an element in the image is gray, it means that the content is currently selected. In addition, since the image materials in the current editing page are empty (no valid image materials have been imported or generated), the content generation control and content packaging control in the editing page are both in an untouchable state. The image uses gray text to visually indicate that these two controls cannot be operated at present and can only be enabled after the image materials are added or generated.

[0119] like Figure 4 The image shown is a schematic diagram of an interactive page in a scenario where a user inputs identification information. (Refer to...) Figure 4 In section (a), the target product type has been determined based on the user-input identifier information, and product images have been identified. Five product images are displayed, with the first two images marked with a recommendation tag. This recommendation tag can be determined based on multiple dimensions such as product attribute matching, user behavior data, scenario adaptation requirements, or algorithm scoring. The specific criteria can be flexibly set according to actual application scenarios (such as e-commerce product selection, marketing material screening, content creation assistance, etc.). Simultaneously, multiple carrier generation entry points are generated based on the target product type. These entry points are displayed in a drop-down menu format, including scene generation, AI try-on, and handheld product. (See reference...) Figure 4 In section (b), when the user selects scene generation, multiple preset style scene carrier images are displayed in the dialog area. Simultaneously, a style selection control is displayed above the scene carrier images, containing multiple style options in a drop-down menu. Users can switch between different style scene carrier images by toggling different style options. After the user selects the target product image and the target scene carrier image, they can click on the following... Figure 4 After clicking the "Load Product" button as shown in (b), the system generates corresponding image materials based on the selected target product image and target scene image, and displays them in the dialog area, such as... Figure 4 As shown in (c), after selecting product image 1, scene 1 and scene 2, product image 1 is loaded into scene 1 and scene 2 respectively according to the preset posture, generating image material 1 and image material 2, and displayed in the dialogue area.

[0120] like Figure 5 The image shown is a schematic diagram of an interactive page in a scenario where the user inputs an image of the target product. (Refer to...) Figure 5In the interactive page (a), the user uploads a target product image (product image 1) and enters the type of the target product image (shoes). The system then displays multiple carrier generation entry points corresponding to this type in the dialog area. These entry points are displayed as tabs, including scene generation, AI try-on, and handheld product. When the user selects AI try-on, multiple main carrier images with preset attributes are displayed in the dialog area. A main attribute selection control is displayed above the main carrier image, containing multiple attribute options in a drop-down menu. The user can switch between different attribute main carrier images by toggling different attribute options. After selecting a target main carrier image, the user can click on... Figure 5 After clicking the "Load Product" button as shown in (b), the system generates corresponding image materials based on the selected target product image and target subject image, and displays them in the dialog area, such as... Figure 5 As shown in (c), after selecting product image 1, model 1 and model 2, product image 1 is adapted to the usage scenario with model 1 and model 2 respectively to generate image material 1 and image material 2, which are then displayed in the dialogue area.

[0121] like Figure 6 The diagram shows the editing area. After a user selects a target image in the dialog area and generates a script for the product, the corresponding target image and script-generated text information are displayed in the editing area. The target image supports adding, deleting, and rearranging its order, and the text information can also be edited. The editing area also includes a selection of compositing modes, with two controls corresponding to image carousel mode and video compositing mode, allowing users to choose freely. The editing area also includes a prompt input area. When the user selects video compositing mode, this area is triggerable, allowing the user to input prompts. If the user selects image carousel mode, this area is not triggerable, and the user cannot input prompts. The editing area also includes a content wrapping control. When this control is selected, preset enhancement effects can be added to the generated media content. Additionally, for the target image displayed in the editing area, the user can trigger an action description input operation to display an action description pop-up window. This pop-up window is used to input target action description text for the triggered target image (not shown in the diagram). Finally, the editing area also includes a content generation control. When the user clicks this control, corresponding media content is generated based on the target image material, prompt information, compositing mode, and content packaging mentioned above. The foregoing has described specific embodiments of this specification. Other embodiments are within the scope of the appended claims. In some cases, the actions or steps recited in the claims may be performed in a different order than that shown in the embodiments and may still achieve the desired result. Furthermore, the processes depicted in the drawings do not necessarily require the specific or sequential order shown to achieve the desired result. In some embodiments, multitasking and parallel processing are possible or may be advantageous.

[0122] According to another embodiment, a content generation apparatus is provided. Figure 7 This is a schematic block diagram of a content generation apparatus provided in an embodiment of this application. The apparatus is disposed in... Figure 1 The user end in the illustrated architecture. For example... Figure 7 As shown, the device 700 mainly includes a page display unit 701 and an event listening unit 702. The main functions of each component are as follows: The page display unit 701 is configured to display an interactive page, which includes a dialog area and an editing area.

[0123] Event listening unit 702 is configured to listen for events on the interactive page.

[0124] The page display unit 701 is further configured to, in response to inputting item information of a target item in the dialog area and triggering a display request instruction for the target item, display at least one image material in the dialog area, the image material being an image including a display carrier and the target item, the display carrier including a main carrier and / or a scene carrier; in response to a selection operation of a target image material among the at least one image material, display the target image material in the editing area; and in response to detecting a content generation instruction in the editing area, obtain media content corresponding to the target item based on the target image material.

[0125] Optionally, the item information is the identification information of the target item; then the page display unit 701 is specifically configured as follows: In response to inputting the identification information of the target item in the dialog area and triggering the display request instruction, the type of the target item and at least one corresponding item image are determined based on the identification information, and the item image and at least one carrier generation entry corresponding to the type are displayed in the dialog area; In response to the selection operation of the target carrier generation entry in the at least one carrier generation entry, at least one carrier image is displayed in the dialog area, the carrier image containing a display carrier corresponding to the target carrier generation entry; In response to a selection operation of a target item image in the at least one item image and a target carrier image in the at least one carrier image, the image material is displayed in the dialog area, the image material including the target item image and the target carrier image.

[0126] Optionally, the item information is the target item image of the target item; then the page display unit 701 is specifically configured as follows: In response to inputting a target item image in the dialog area and triggering the display request instruction, the type of the target item is determined, and at least one carrier generation entry corresponding to the type is displayed in the dialog area; In response to the selection operation of the target carrier generation entry in the at least one carrier generation entry, at least one carrier image is displayed in the dialog area, the carrier image containing a display carrier corresponding to the target carrier generation entry; In response to a selection operation of a target carrier image in the at least one carrier image, the image material is displayed in the dialog area, the image material including the target item image and the target carrier image.

[0127] Optionally, the carrier generation entry includes: a main carrier entry and / or a scene carrier entry.

[0128] Optionally, the target carrier generation entry is the scene carrier entry, and the display carrier corresponding to the target carrier generation entry is the scene carrier; then the page display unit 701 is specifically configured as follows: The target item image is loaded into the target carrier image according to a preset posture to generate the image material, which is then displayed in the dialogue area.

[0129] Optionally, the target carrier generation entry is the main carrier entry, and the display carrier corresponding to the target carrier generation entry is the main carrier; then the page display unit 701 is specifically configured as follows: The target item image and the target carrier image are adapted to the usage scenario to generate the image material, which is then displayed in the dialogue area.

[0130] Optionally, the page display unit 701 is further configured to: In response to the selection operation of the target carrier generation entry, at least one scene carrier image with a preset style is generated; A style selection control is displayed in the dialog area, and the style selection control contains multiple style options; In response to the user's selection of a style option in the style selection control that is different from the preset style, a new scene carrier image that matches the style selected by the user is determined, and the new scene carrier image is updated and displayed in the dialog area.

[0131] Optionally, the page display unit 701 is further configured to: In response to the selection operation of the target carrier generation entry, a main carrier image with at least one preset attribute is generated; A subject attribute selection control is displayed in the dialog area, and the subject attribute selection control contains multiple attribute options; In response to a user's selection of an attribute option in the subject attribute selection control that is different from the preset attribute, a new subject carrier image that matches the attribute selected by the user is determined, and the new subject carrier image is updated and displayed in the dialog area.

[0132] Optionally, the page display unit 701 is further configured as follows: In response to the arrangement operation of the target image material in the editing area, update the existence state and / or arrangement order of the target image material.

[0133] Optionally, the page display unit 701 is further configured to: In response to triggering an action description input operation for any target image material in the editing area, an action description input pop-up window is displayed, which is used to input target action description text for the triggered target image material; Based on the target image material and the target action description text of the triggered target image material, the media content corresponding to the target item is obtained, wherein the display of the target item in the media content conforms to the target action description text of the triggered target image material.

[0134] Furthermore, the editing area also includes a text input area; Page display unit 701 is further configured as follows: Based on the target image material and the text information of the target item in the text input area, the media content corresponding to the target item is obtained.

[0135] Optionally, the page display unit 701 is further configured as follows: In response to a text generation instruction for the target item input in the dialog area, a script for the target item is displayed in the dialog area, the script including content framework information; In response to a selection operation on the script, text information of the target item is displayed in the text input area, the text information being generated based on the content framework information.

[0136] The dialog area also displays the refresh component corresponding to the script; Optionally, page display unit 701 is also configured as follows: In response to a trigger operation on the refresh component, a new script is displayed in the dialog area, the new script containing new content frame information; In response to the selection of the new script, the text information of the target item generated based on the new content framework information is displayed in the text input area.

[0137] Optionally, page display unit 701 is also configured as follows: In response to an editing operation on the text information in the text input area, the content of the text information is updated.

[0138] The editing area includes at least one composition mode control, and each composition mode control corresponds to a preset composition mode; the presentation form of the media content is determined based on the composition mode corresponding to the selected composition mode control.

[0139] The editing area also includes a prompt input area for inputting prompt information. This prompt information is a natural language description of the user's video generation needs, and the natural language description includes at least one or more of the following: the action of the display carrier, camera movement, or scene atmosphere; and the selected compositing mode control corresponds to a video compositing mode. Optionally, the page display unit 701 is further configured as follows: The media content is obtained based on the video generation mode, the prompt information, and the target image material.

[0140] The editing area also includes content packaging controls; Optionally, the page display unit 701 is further configured as follows: In response to the selection of the content packaging control, preset enhancement effects are added to the media content during the generation of the media content. The various embodiments in this specification are described in a progressive manner. Similar or identical parts between embodiments can be referred to mutually. Each embodiment focuses on describing the differences from other embodiments. In particular, for system or device embodiments, since they are basically similar to method embodiments, the description is relatively simple, and relevant parts can be referred to the description of the method embodiments. The system and device embodiments described above are merely illustrative. The units described as separate components may or may not be physically separate. The components shown as units may or may not be physical units; that is, they may be located in one place or distributed across multiple network units. Some or all of the modules can be selected to achieve the purpose of this embodiment according to actual needs. Those skilled in the art can understand and implement this without creative effort.

[0141] It should be noted that the user information (including but not limited to user device information, user personal information, etc.) and data (including but not limited to data used for analysis, data stored, data displayed, etc.) involved in this application are all information and data authorized by the user or fully authorized by all parties. Furthermore, the collection, use and processing of the relevant data must comply with the relevant laws, regulations and standards of the relevant countries and regions, and corresponding operation entry points are provided for users to choose to authorize or refuse.

[0142] In addition, embodiments of this application also provide a computer-readable storage medium storing a computer program thereon, which, when executed, implements the steps of the method described in any of the foregoing method embodiments.

[0143] And an electronic device, comprising: One or more processors; and A memory associated with the one or more processors, the memory being used to store program instructions that, when read and executed by the one or more processors, perform the steps of the method described in any of the foregoing method embodiments.

[0144] This application also provides a computer program product, including a computer program that, when executed, implements the steps of the method described in any of the foregoing method embodiments.

[0145] in, Figure 8This is a schematic block diagram of an electronic device provided in an embodiment of this application. Specifically, it may include a processor 810, a video display adapter 811, a disk drive 812, an input / output interface 813, a network interface 814, and a memory 820. The processor 810, video display adapter 811, disk drive 812, input / output interface 813, network interface 814, and memory 820 can be communicatively connected via a communication bus 830. The input / output interface 813 can also be referred to as an I / O interface 813.

[0146] The processor 810 can be implemented using a general-purpose CPU, microprocessor, application-specific integrated circuit (ASIC), or one or more integrated circuits to execute relevant programs and implement the technical solution provided in this application.

[0147] The memory 820 can be implemented in the form of ROM (Read Only Memory), RAM (Random Access Memory), static storage device, dynamic storage device, etc. The memory 820 can store the operating system 821 for controlling the operation of the electronic device 800, and the basic input / output system (BIOS) 822 for controlling the low-level operations of the electronic device 800. Additionally, it can store a web browser 823, a data storage management system 824, and a content generation device 700, etc. The aforementioned media file content generation device 700 can be the application program that specifically implements the aforementioned steps in this embodiment. In summary, when the technical solution provided in this application is implemented through software or firmware, the relevant program code is stored in the memory 820 and is called and executed by the processor 810.

[0148] The input / output interface 813 is used to connect input / output modules to realize information input and output. Input / output modules can be configured as components within the device (not shown in the figure) or externally connected to the device to provide corresponding functions. Input devices may include keyboards, mice, touchscreens, microphones, various sensors, etc., while output devices may include displays, speakers, vibrators, indicator lights, etc.

[0149] Network interface 814 is used to connect a communication module (not shown in the figure) to enable communication between this device and other devices. The communication module can communicate via wired means (such as USB, Ethernet cable, etc.) or wireless means (such as mobile network, WIFI, Bluetooth, etc.).

[0150] Bus 830 includes a pathway for transmitting information between various components of the device, such as processor 810, video display adapter 811, disk drive 812, input / output interface 813, network interface 814, and memory 820.

[0151] It should be noted that although the above-described device only shows the processor 810, video display adapter 811, disk drive 812, input / output interface 813, network interface 814, memory 820, bus 830, etc., in specific implementations, the device may also include other components necessary for normal operation. Furthermore, those skilled in the art will understand that the above-described device may only include the components necessary for implementing the solution of this application, and does not necessarily include all the components shown in the figures.

[0152] As can be seen from the above description of the embodiments, those skilled in the art can clearly understand that this application can be implemented by means of software plus necessary general-purpose hardware platforms. Based on this understanding, the technical solution of this application, in essence, or the part that contributes to the prior art, can be embodied in the form of a computer program product. This computer program product can be stored in a storage medium, such as ROM / RAM, magnetic disk, optical disk, etc., and includes several instructions to cause a computer device (which may be a personal computer, server, or network device, etc.) to execute the methods described in various embodiments or some parts of the embodiments of this application.

[0153] The technical solutions provided in this application have been described in detail above. Specific examples have been used to illustrate the principles and implementation methods of this application. The descriptions of the above embodiments are only for the purpose of helping to understand the methods and core ideas of this application. Furthermore, those skilled in the art will recognize that, based on the ideas of this application, there will be changes in the specific implementation methods and application scope. Therefore, the content of this specification should not be construed as a limitation of this application.

Claims

1. A content generation method, characterized in that, The method includes: Display an interactive page, which includes a dialog area and an editing area; In response to inputting item information of a target item in the dialog area and triggering a display request instruction for the target item, at least one image material is displayed in the dialog area. The image material is an image including a display carrier and the target item, and the display carrier includes a main carrier and / or a scene carrier. In response to a selection operation of a target image material in the at least one image material, the target image material is displayed in the editing area; In response to the detection of a content generation trigger command in the editing area, media content corresponding to the target item is obtained based on the target image material.

2. The method according to claim 1, characterized in that, The item information is the identification information of the target item; then, in response to the input of the item information of the target item in the dialog area and the triggering of a display request instruction for the target item, at least one image material is displayed in the dialog area, including: In response to inputting the identification information of the target item in the dialog area and triggering the display request instruction, the type of the target item and at least one corresponding item image are determined based on the identification information, and the item image and at least one carrier generation entry corresponding to the type are displayed in the dialog area; In response to the selection operation of the target carrier generation entry in the at least one carrier generation entry, at least one carrier image is displayed in the dialog area, the carrier image containing a display carrier corresponding to the target carrier generation entry; In response to a selection operation of a target item image in the at least one item image and a target carrier image in the at least one carrier image, the image material is displayed in the dialog area, the image material including the target item image and the target carrier image.

3. The method according to claim 1, characterized in that, The item information is the target item image of the target item; In response to inputting item information of a target item in the dialog area and triggering a display request instruction for the target item, at least one image asset is displayed in the dialog area, including: In response to inputting a target item image in the dialog area and triggering the display request instruction, the type of the target item is determined, and at least one carrier generation entry corresponding to the type is displayed in the dialog area; In response to the selection operation of the target carrier generation entry in the at least one carrier generation entry, at least one carrier image is displayed in the dialog area, the carrier image containing a display carrier corresponding to the target carrier generation entry; In response to a selection operation of a target carrier image in the at least one carrier image, the image material is displayed in the dialog area, the image material including the target item image and the target carrier image.

4. The method according to claim 2 or 3, characterized in that, The carrier generation entry points include: the main carrier entry point and / or the scene carrier entry point.

5. The method according to claim 4, characterized in that, The target carrier generation entry is the scene carrier entry, and the display carrier corresponding to the target carrier generation entry is the scene carrier; but The display of the image material in the dialogue area includes: The target item image is loaded into the target carrier image according to a preset posture to generate the image material, which is then displayed in the dialogue area.

6. The method according to claim 4, characterized in that, The target carrier generation entry is the main carrier entry, and the display carrier corresponding to the target carrier generation entry is the main carrier; but The display of the image material in the dialogue area includes: The target item image and the target carrier image are adapted to the usage scenario to generate the image material, which is then displayed in the dialogue area.

7. The method according to claim 5, characterized in that, The step of displaying at least one carrier image in the dialog area in response to a selection operation of a target carrier generation entry in the at least one carrier generation entry includes: In response to the selection operation of the target carrier generation entry, at least one scene carrier image with a preset style is generated; The method further includes: A style selection control is displayed in the dialog area, and the style selection control contains multiple style options; In response to the user's selection of a style option in the style selection control that is different from the preset style, a new scene carrier image that matches the style selected by the user is determined, and the new scene carrier image is updated and displayed in the dialog area.

8. The method according to claim 6, characterized in that, The step of displaying at least one carrier image in the dialog area in response to a selection operation of a target carrier generation entry in the at least one carrier generation entry includes: In response to the selection operation of the target carrier generation entry, a main carrier image with at least one preset attribute is generated; The method further includes: A subject attribute selection control is displayed in the dialog area, and the subject attribute selection control contains multiple attribute options; In response to a user's selection of an attribute option in the subject attribute selection control that is different from the preset attribute, a new subject carrier image that matches the attribute selected by the user is determined, and the new subject carrier image is updated and displayed in the dialog area.

9. The method according to claim 1, characterized in that, The method further includes: In response to the arrangement operation of the target image material in the editing area, update the existence state and / or arrangement order of the target image material.

10. The method according to claim 1, characterized in that, After displaying the target image material in the editing area, the method further includes: In response to triggering an action description input operation for any target image material in the editing area, an action description input pop-up window is displayed, which is used to input target action description text for the triggered target image material; The step of obtaining the media content corresponding to the target item based on the target image material includes: Based on the target image material and the target action description text of the triggered target image material, the media content corresponding to the target item is obtained, wherein the display of the target item in the media content conforms to the target action description text of the triggered target image material.

11. The method according to claim 1, characterized in that, The editing area also includes a text input area; The step of obtaining the media content corresponding to the target item based on the target image material includes: Based on the target image material and the text information of the target item in the text input area, the media content corresponding to the target item is obtained.

12. The method according to claim 11, characterized in that, The method further includes: In response to a text generation instruction for the target item input in the dialog area, a script for the target item is displayed in the dialog area, the script including content framework information; In response to a selection operation on the script, text information of the target item is displayed in the text input area, the text information being generated based on the content framework information.

13. The method according to claim 12, characterized in that, The dialog area also displays the refresh component corresponding to the script; The method further includes: In response to a trigger operation on the refresh component, a new script is displayed in the dialog area, the new script containing new content frame information; In response to the selection of the new script, the text information of the target item generated based on the new content framework information is displayed in the text input area.

14. The method according to claim 13, characterized in that, The method further includes: In response to an editing operation on the text information in the text input area, the content of the text information is updated.

15. The method according to claim 1, characterized in that, The editing area includes at least one composition mode control, and each composition mode control corresponds to a preset composition mode; the presentation form of the media content is determined based on the composition mode corresponding to the selected composition mode control.

16. The method according to claim 15, characterized in that, The editing area also includes a prompt input area for inputting prompt information. This prompt information is a natural language description of the user's video generation needs, and the natural language description includes at least one or more of the following: the action of the display carrier, camera movement, or scene atmosphere; and the selected compositing mode control corresponds to a video compositing mode. The process of obtaining the media content based on the target image material includes: The media content is obtained based on the video generation mode, the prompt information, and the target image material.

17. The method according to claim 1, characterized in that, The editing area also includes content packaging controls; The process of obtaining the media content corresponding to the target item includes: In response to the selection of the content packaging control, preset enhancement effects are added to the media content during the generation of the media content.

18. A content generation apparatus, characterized in that, The device includes: A page display unit is configured to display an interactive page, which includes a dialog area and an editing area; An event listening unit is configured to listen for events on the interactive page; The page display unit is further configured to, in response to inputting item information of a target item in the dialog area and triggering a display request instruction for the target item, display at least one image material in the dialog area, the image material being an image including a display carrier and the target item, the display carrier including a main carrier and / or a scene carrier; in response to a selection operation of a target image material among the at least one image material, display the target image material in the editing area; and in response to detecting a content generation instruction in the editing area, obtain media content corresponding to the target item based on the target image material.

19. A computer-readable storage medium having a computer program stored thereon, characterized in that, When the computer program is executed, it implements the steps of the method according to any one of claims 1 to 17.

20. An electronic device, characterized in that, include: One or more processors; as well as A memory associated with the one or more processors, the memory being used to store program instructions that, when read and executed by the one or more processors, perform the steps of the method according to any one of claims 1 to 17.

21. A computer program product, comprising a computer program, characterized in that, When the computer program is executed, it performs the steps of the method according to any one of claims 1 to 17.