Media content generation method and apparatus, electronic device, storage medium and program product
By displaying a candidate prompt information group and selecting the first prompt information to generate the target media content, the problem of high difficulty and high uncertainty in filling out prompts is solved, and the quality of generated content is improved.
Patent Information
- Application Number
- PCT/CN2024/107664
- Authority / Receiving Office
- WO · WO
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2024-03-08
- Filing Date
- 2024-07-25
- Publication Date
- 2025-09-11
AI Technical Summary
The existing prompt filling method is difficult and uncertain, resulting in low quality of generated content.
A media content generation method is provided, which displays a candidate prompt information group, allows a user to select a first prompt information, and generates target media content based on the prompt information, thereby reducing reliance on text input.
It reduces the difficulty and uncertainty of inputting prompt information and improves the quality of generated content.
Smart Images

Figure CN2024107664_12092025_PF_FP_ABST
Abstract
Description
Method, device, electronic device, storage medium and program product for generating media content
[0001] This application claims priority to the Chinese invention patent application entitled “Method, device, electronic device, storage medium and program product for generating media content” filed on March 8, 2024, with application number 202410269150.6, the entire contents of which are incorporated by reference into this application. Technical Field
[0002] The embodiments of the present disclosure relate to the field of computer technology, and in particular to a method, apparatus, electronic device, storage medium, and program product for generating media content. Background Art
[0003] Currently, AI-generated content (AIGC) can display an input box and generate content based on the prompts entered by the user. For example, in the text-to-image scenario, an image can be generated based on the description of the scene entered by the user in the input box.
[0004] However, the existing prompt filling method has the problems of high difficulty and high uncertainty in prompt filling, resulting in low quality of generated content.
[0005] Summary of the Invention
[0006] The embodiments of the present disclosure provide a method, apparatus, electronic device, storage medium, and program product for generating media content, so as to reduce the difficulty and uncertainty of filling in prompt information and improve the quality of generated content.
[0007] In a first aspect, an embodiment of the present disclosure provides a method for generating media content, comprising:
[0008] Displaying at least one prompt information group, wherein the prompt information group includes at least one candidate prompt information, and the candidate prompt information is used to describe content characteristics of target media content;
[0009] In response to an information selection operation on at least one of the candidate prompt information in the at least one prompt information group, determining first prompt information from the candidate prompt information;
[0010] The target media content generated based on the target prompt information is displayed; wherein the target prompt information includes the first prompt information.
[0011] In a second aspect, an embodiment of the present disclosure further provides a device for generating media content, including:
[0012] A prompt information display module, configured to display at least one prompt information group, wherein the prompt information group includes at least one candidate prompt information, and the candidate prompt information is used to describe content characteristics of target media content;
[0013] a prompt information determining module, configured to determine first prompt information from the candidate prompt information in response to an information selection operation on at least one of the candidate prompt information in the at least one prompt information group;
[0014] The media content display module is configured to display the target media content generated based on the target prompt information; wherein the target prompt information includes the first prompt information.
[0015] In a third aspect, an embodiment of the present disclosure further provides an electronic device, including:
[0016] one or more processors;
[0017] a memory for storing one or more programs,
[0018] When the one or more programs are executed by the one or more processors, the one or more processors implement the method for generating media content as described in the embodiment of the present disclosure.
[0019] In a fourth aspect, an embodiment of the present disclosure further provides a computer-readable storage medium having a computer program stored thereon, which, when executed by a processor, implements the method for generating media content as described in the embodiment of the present disclosure.
[0020] In a fifth aspect, the embodiments of the present disclosure further provide a computer program product. When the computer program product is executed by a computer, the computer implements the method for generating media content as described in the embodiments of the present disclosure. BRIEF DESCRIPTION OF THE DRAWINGS
[0021] The above and other features, advantages, and aspects of the various embodiments of the present disclosure will become more apparent with reference to the following detailed description in conjunction with the accompanying drawings. Throughout the drawings, the same or similar reference numerals represent the same or similar elements. It should be understood that the drawings are schematic and that the originals and elements are not necessarily drawn to scale.
[0022] FIG1 is a flow chart of a method for generating media content provided by an embodiment of the present disclosure;
[0023] FIG2 is a schematic diagram showing a prompt information group provided by an embodiment of the present disclosure;
[0024] FIG3 is a schematic diagram showing a target media content provided by an embodiment of the present disclosure;
[0025] FIG4 is a flow chart of another method for generating media content provided by an embodiment of the present disclosure;
[0026] FIG5 is a schematic diagram of an optional process of inputting prompt information provided by an embodiment of the present disclosure;
[0027] FIG6 is a structural block diagram of a device for generating media content provided by an embodiment of the present disclosure;
[0028] FIG7 is a schematic structural diagram of an electronic device provided by an embodiment of the present disclosure. DETAILED DESCRIPTION
[0029] The following describes embodiments of the present disclosure in more detail with reference to the accompanying drawings. Although certain embodiments of the present disclosure are shown in the accompanying drawings, it should be understood that the present disclosure can be implemented in various forms and should not be construed as limited to the embodiments described herein. Rather, these embodiments are provided to provide a more thorough and complete understanding of the present disclosure. It should be understood that the drawings and embodiments of the present disclosure are for illustrative purposes only and are not intended to limit the scope of protection of the present disclosure.
[0030] It should be understood that the various steps described in the method embodiments of the present disclosure may be performed in different orders and / or in parallel. In addition, the method embodiments may include additional steps and / or omit the steps shown. The scope of the present disclosure is not limited in this respect.
[0031] As used herein, the term "including" and its variations are open-ended, i.e., "including but not limited to." The term "based on" means "based, at least in part, on." The term "one embodiment" means "at least one embodiment," the term "another embodiment" means "at least one additional embodiment," and the term "some embodiments" means "at least some embodiments." Other terms are defined in the following description.
[0032] It should be noted that the concepts of "first" and "second" mentioned in this disclosure are only used to distinguish different devices, modules or units, and are not used to limit the order or interdependence of the functions performed by these devices, modules or units.
[0033] It should be noted that the modifications of "one" and "multiple" mentioned in the present disclosure are illustrative rather than restrictive, and those skilled in the art should understand that unless otherwise clearly indicated in the context, they should be understood as "one or more".
[0034] The names of the messages or information exchanged between multiple devices in the embodiments of the present disclosure are only used for illustrative purposes and are not used to limit the scope of these messages or information.
[0035] It is understandable that before using the technical solutions disclosed in the various embodiments of this disclosure, the type, scope of use, usage scenarios, etc. of the personal information involved in this disclosure should be informed to the user and the user's authorization should be obtained in an appropriate manner in accordance with relevant laws and regulations.
[0036] It is understandable that the data involved in this technical solution (including but not limited to the data itself, the acquisition or use of the data) must comply with the requirements of relevant laws, regulations and relevant provisions.
[0037] The media content generation method, device, electronic device, storage medium, and program product provided by the embodiments of the present disclosure display at least one prompt information group, wherein the prompt information group includes at least one candidate prompt information, and the candidate prompt information is used to describe the content characteristics of the target media content; in response to an information selection operation for at least one candidate prompt information in the at least one prompt information group, first prompt information is determined in the candidate prompt information; and target media content generated based on the target prompt information is displayed, wherein the target prompt information includes the first prompt information. The embodiments of the present disclosure utilize the above-mentioned technical solution to support users in selecting candidate prompt information from the prompt information group as prompt information for the media content to be generated, without relying entirely on the user to input through text input, which can reduce the difficulty and uncertainty of inputting prompt information and improve the quality of the generated media content.
[0038] FIG1 is a flow chart of a method for generating media content provided by an embodiment of the present disclosure. The method can be executed by a device for generating media content, wherein the device can be implemented by software and / or hardware and can be configured in an electronic device, typically a mobile phone or tablet computer. The method for generating media content provided by an embodiment of the present disclosure is applicable to AIGC scenarios, such as text-based image scenarios. As shown in FIG1 , the method for generating media content provided by this embodiment may include:
[0039] S101: Display at least one prompt information group, where the prompt information group includes at least one candidate prompt information, and the candidate prompt information is used to describe content characteristics of target media content.
[0040] The prompt information group can be understood as a grouping of candidate prompt information, and each prompt information group can include one or more candidate prompt information. There is no limit to the grouping method of the candidate prompt information. For example, different prompt information groups can correspond to different types of attribute information of the target content. Taking the description of the characters in the picture as an example, the candidate prompt information in different prompt information groups can describe different types of attribute information of the characters. The different types of attribute information can include, for example, clothing, hairstyle, accessories, posture and / or expression, etc. The candidate prompt information can be prompt information for the user to select, and the prompt information can be a description of the content characteristics, which is used to describe the content characteristics of the target media content to be generated. The type of candidate prompt information is not limited. For example, the candidate prompt information can be a word, a sentence and / or a paragraph, etc.
[0041] Figure 2 is a schematic diagram of a prompt information group 20 provided in an embodiment of the present disclosure. As shown in Figure 2, at least one prompt information group 20 can be displayed, and each prompt information group 20 can display at least one candidate prompt information for the user to select, so that the user can select one or more candidate prompt information to describe the content characteristics of the target media content to be generated.
[0042] The at least one prompt information group displayed can be set as needed. For example, the original media content used to generate the target media content can be ignored, and at least one pre-set prompt information group can be displayed to reduce the amount of calculation required to display the prompt information group and reduce the display delay of the prompt information group. The original media content used to generate the target media content can also be considered, and at least one prompt information group corresponding to the original media content can be displayed, so that the candidate prompt information in the displayed prompt information group has a high correlation with the original media content, thereby improving the practicality of the displayed prompt information group.
[0043] In some embodiments, for the case where original media content exists, at least one pre-set prompt information group can be displayed, or at least one prompt information group corresponding to the original media content can be displayed, such as when a user triggers the display of a certain prompt dimension, at least one prompt information group pre-set under the prompt dimension is displayed, or at least one prompt information group corresponding to the original media content under the prompt dimension is displayed; for the case where no original media content exists, at least one pre-set prompt information group can be displayed, such as when a user triggers the display of a certain prompt dimension, at least one prompt information group pre-set under the prompt dimension is displayed.
[0044] Exemplarily, when the original media content exists, at least one prompt information group corresponding to the original media content may be displayed. Optionally, the displaying of at least one prompt information group includes: displaying at least one prompt information group corresponding to the original media content.
[0045] In some embodiments, displaying at least one prompt information group corresponding to the original media content can be understood as the displayed prompt information group being associated with the original media content. When the original media content or the type of the original media content is different, the displayed at least one prompt information group may be different.
[0046] Optionally, different original media contents correspond to different prompt information groups; or, the prompt information groups corresponding to different original media contents contain different candidate prompt information; or, different original media contents correspond to different display orders, and the display order includes the display order of the at least one prompt information group and / or the display order of the candidate prompt information in the prompt information group.
[0047] For example, based on the original media content, at least one prompt information group to be displayed can be determined, the candidate prompt information included in the at least one prompt information group can be determined, the display order of the at least one prompt information group can be determined, and / or the display order of the candidate prompt information included in the at least one prompt information group in the corresponding prompt information group can be determined; and based on the above-mentioned determined content, at least one prompt information group corresponding to the original media content can be displayed. Thus, when the original media content provided by the user is different, different prompt information groups can be displayed, prompt information groups can be displayed in different display orders, different candidate prompt information can be displayed in the prompt information group, and / or the candidate prompt information included in the prompt information group can be displayed in different display orders, etc., so as to further improve the practicality of the displayed prompt information group and enrich the display method of the prompt information group.
[0048] S102: In response to an information selection operation on at least one candidate prompt information in the at least one prompt information group, determine first prompt information from the candidate prompt information.
[0049] In some embodiments, the information selection operation may be a triggering operation for selecting one or more candidate prompt information as the first prompt information. The information selection operation may be performed in any manner. For example, the information selection operation may be an operation for triggering a control or candidate prompt information, or an operation for performing a selection gesture. The first prompt information may be candidate prompt information determined for generating the target media content, such as the candidate prompt information selected or dragged by the information selection operation, or the candidate prompt information in a selected state.
[0050] In this embodiment, candidate prompt information for generating target media content may be determined based on the user's information selection operation.
[0051] Exemplarily, when receiving the user's information selection operation, the first prompt information can be determined in the candidate prompt information, such as determining the candidate prompt information corresponding to the information selection operation as the first prompt information, or determining the candidate prompt information in the selected state as the first prompt information.
[0052] In this embodiment, after determining the first prompt information, the determined first prompt information may or may not be displayed in the prompt information input area. Exemplarily, the determined first prompt information may be displayed in the prompt information input area, such as by inputting the first prompt information into the prompt information input area for the user to view. In this case, optionally, after determining the first prompt information from the candidate prompt information, the method further includes: displaying the first prompt information in the prompt information input area.
[0053] In some embodiments, the prompt information input area may be an area for inputting prompt information for the media content to be generated. There are various ways to input prompts in the prompt information input area, such as by selecting candidate prompt information or by entering characters.
[0054] S103: Display the target media content generated based on the target prompt information; wherein the target prompt information includes the first prompt information.
[0055] The target prompt information can be understood as the prompt information set by the user for generating the target media content, such as the prompt information input in the information input area. Exemplarily, the target prompt information may include the first prompt information and / or the second prompt information, depending on the actual setting of the user. The first prompt information may be the prompt information determined in the candidate prompt information by performing the information selection operation; the second prompt information may be the prompt information input by character input. The character input method may include but is not limited to voice input and / or text input. The text input method may include input through character editing or input through copying and pasting. This embodiment does not impose any restrictions on this.
[0056] The target media content may be media content generated based on the target prompt information, and the media content may be image media content, such as pictures and / or videos, etc. The target media content may be generated by the current application or a server corresponding to the current application.
[0057] Specifically, the target prompt information 31 set by the user can be obtained, such as obtaining the prompt information entered in the prompt information input area as the target prompt information 31, generating the target media content 30 based on the target prompt information 31, or generating the target media content 30 based on the target prompt information 31 through the server, and displaying the generated target media content 30, as shown in Figure 3.
[0058] In this embodiment, the target media content can be generated based on the target prompt information when a preset generation condition is met. The preset generation condition may include, for example, receiving a media content generation operation, confirming completion of the first prompt information, or the number of prompt information set by the user reaching a preset number, etc., and the specific conditions can be set as needed. In some embodiments, the media content generation operation can be a triggering operation for instructing the generation of new media content based on the prompt information set by the user. The execution method of the media content generation operation is not limited. For example, the media content generation operation can be an operation for triggering a media content generation control and / or executing a media content generation gesture.
[0059] There is no limit to the method of generating the target media content. For example, when the original media content does not exist, the original media content can be ignored and the target media content can be generated based on the target prompt information set by the user. When the original media content exists, the original media content can be considered and the target media content can be generated based on the original media content and the target prompt information set by the user. For example, the media content screen of the original media content can be processed based on the target prompt information set by the user to obtain the media content screen of the target media content, thereby obtaining the target media content.
[0060] In some implementations, the media content screen of the target media content is obtained by processing the media content screen of the original media content based on the target prompt information; and before displaying at least one prompt information group, the method further includes: acquiring the original media content to be processed.
[0061] For example, a shooting interface can be displayed, and the media content shot by the user in the shooting interface can be used as the original media content; or, a media content interface (such as the user's local photo album, etc.) can be displayed, and the media content selected by the user in the media content interface can be used as the original media content, and so on.
[0062] 3 , when the target media content 30 is displayed, the target prompt information 31 used when generating the target media content 30 may also be further displayed. Thus, the user can view the target media content 30 and the target prompt information 31 accordingly, and / or, if dissatisfied with the generated target media content 30, can regenerate new target media content 30 by modifying the target prompt information 31.
[0063] After the target media content 30 is generated, the user may be supported to save and / or publish the target media content 30. For example, a save control 32 (as shown in FIG3 ) and / or a publish control (not shown in FIG3 ) corresponding to the target media content 30 may be displayed, so that the user can save the target media content 30 by triggering the save control 32 and / or publish the target media content 30 by triggering the publish control.
[0064] In some embodiments, the method further includes: in response to an area selection operation performed on the original media content and / or the target media content, displaying the prompt information group associated with the selected area in the at least one prompt information group in a first state, and displaying the prompt information groups other than the prompt information group associated with the selected area in the at least one prompt information group in a second state.
[0065] In some embodiments, the area selection operation can be understood as a trigger operation for selecting a certain area on the original media content and / or the target media content. The execution method of the area selection operation is not limited. For example, the area selection operation may include an operation of making a frame selection on the media content screen of the original media content and / or the target media content, and / or an operation of selecting a certain area in the area list of the original media content and / or the target media content, etc. The selected area can be an area corresponding to the area selection operation, such as an area selected by the user by performing the area selection operation. The first state and the second state can be different states, such as the first state and the second state can correspond to different display styles or display orders, etc.
[0066] In the above embodiment, the prompt information group associated with the user's selected area on the original media content and / or the target media content and the prompt information groups other than the associated prompt information group of the at least one prompt information group can be displayed in different states, so that the user can quickly locate or filter out the prompt information group associated with his or her selected area in the at least one displayed prompt information group.
[0067] Specifically, when an area selection operation acting on the original media content and / or the target media content is received, the user's selected area can be determined, such as taking the area corresponding to the area selection operation as the selected area; and using the first state to display the prompt information group associated with the selected area in the above-mentioned at least one prompt information group, and using the second state to display the prompt information groups in the above-mentioned at least one prompt information group except the prompt information group associated with the selected area.
[0068] Exemplarily, in the case where an area selection operation is received before displaying at least one prompt information, when displaying at least one prompt information group, the prompt information group associated with the selected area in the at least one prompt information group can be displayed in a first state, and the prompt information group not associated with the selected area in the at least one prompt information group can be displayed in a second state.
[0069] In the case where an area selection operation is received after the at least one prompt information is displayed, when there is no selected area, the at least one prompt information group can be displayed in the second state; thus, after the area selection operation is received, the prompt information group associated with the selected area in the at least one prompt information group can be switched from the second state to the first state.
[0070] In some exemplary scenarios, the user may select at least one area on the original media content and / or the target media content, and the current application may display the prompt information group and / or candidate prompt information associated with the at least one area by moving it upwards and / or highlighting it, or may hide the prompt information group and / or candidate prompt information associated with the at least one area. For example, if a hat is presented on the original media content and / or the target media content, after detecting that the user has selected the hat area, the current application may move the prompt information group associated with the hat upwards and / or highlight it; or, after detecting that the user has selected the hair area, the current application may move the prompt information group associated with the hairstyle, hat, and / or headwear upwards and / or highlight it, and so on, to facilitate the user's selection.
[0071] The media content generation method provided in this embodiment displays at least one prompt information group, the prompt information group including at least one candidate prompt information, the candidate prompt information being used to describe the content characteristics of the target media content; in response to an information selection operation for at least one candidate prompt information in the at least one prompt information group, determines a first prompt information from the candidate prompt information; and displays target media content generated based on the target prompt information, the target prompt information including the first prompt information. This embodiment utilizes the above-mentioned technical solution to support users in selecting candidate prompt information from the prompt information group as prompt information for the media content to be generated, eliminating the need for users to input the information entirely through text input. This reduces the difficulty and uncertainty of inputting the prompt information and improves the quality of the generated media content.
[0072] In some embodiments, a plurality of prompt dimensions may be pre-set. Different prompt dimensions may correspond to different description dimensions of content features. For example, different prompt dimensions may correspond to different description themes of a picture, and the description themes may include, for example, the picture subject, the picture environment, the picture atmosphere and / or the picture style, etc. Each prompt dimension may include one or more prompt information groups, and each prompt information group may include at least one candidate prompt information. Exemplarily, the candidate prompt information included in the prompt information groups in different prompt dimensions may be used to describe the content features of the media content under different description dimensions. At this point, when displaying at least one prompt information group, one or more prompt information groups in the prompt dimension may be displayed.
[0073] Optionally, the at least one prompt information group belongs to the same prompt dimension, and different prompt dimensions correspond to different description dimensions of the content feature. That is, when displaying at least one prompt information group, at least one prompt information group in a certain prompt dimension may be displayed.
[0074] When displaying at least one prompt information group, different prompt information groups can be displayed in different sub-areas of the prompt information display area. Different prompt information groups can include different types of candidate prompt information, and the different types of candidate prompt information can be used to describe different types of attribute information in the content characteristics of the target media content. In this case, optionally, displaying at least one prompt information group includes: displaying different prompt information groups in different sub-areas of the prompt information display area; wherein different prompt information groups include different types of candidate prompt information, and the different types of candidate prompt information are used to describe different types of attribute information of the target media content.
[0075] In some embodiments, before displaying at least one prompt information group, it also includes: in response to a prompt information input operation, displaying dimension information of a prompt information input area and at least one prompt dimension, wherein the dimension information is used to trigger the display of at least one prompt information group in the prompt dimension corresponding to the dimension information.
[0076] In some embodiments, the prompt information input operation can be a triggering operation for indicating the display of the prompt information input area for inputting prompt information, such as an operation that triggers the prompt information input control or an operation in the prompt information area. Dimension information of a prompt dimension can be understood as the dimension information of the prompt dimension displayed with the prompt information input area. The dimension information can include a dimension identifier for the prompt dimension and can also include group information for at least some groups of prompt information within the prompt dimension. This group information can include, for example, a group identifier. In some embodiments, the dimension identifier can include, for example, a dimension name, and the group identifier can include, for example, a group name, etc. This embodiment does not limit this.
[0077] In the above embodiment, the user can instruct the current application to display at least one prompt information group in a prompt dimension by triggering the dimension information of the prompt dimension.
[0078] Specifically, when a prompt information input operation is received, a prompt information input area and dimension information of at least a portion of the prompt dimensions can be displayed, allowing the user to enter prompt information in the prompt information input area and / or view the prompt information groups under the prompt dimension by triggering the dimension information of a prompt dimension. Therefore, when it is detected that the user triggers the dimension information of a prompt dimension, at least one prompt information group in the prompt dimension can be displayed in the prompt information display area.
[0079] In some embodiments, the user may also switch at least one prompt information group displayed in the prompt information display area by performing a prompt dimension switching operation. In this case, the media content generation method may also optionally include: in response to the prompt dimension switching operation, displaying at least one prompt information group in the prompt dimension corresponding to the prompt dimension switching operation in the prompt information display area.
[0080] In some embodiments, the prompt dimension switching operation can be a triggering operation for indicating the prompt dimension of the prompt information group currently being displayed. The execution method of the prompt dimension switching operation is not limited, such as the prompt dimension switching operation can be an operation that triggers the prompt dimension switching control or an operation that triggers the dimension identifier of a certain prompt dimension. Optionally, the prompt information display area also displays dimension identifiers of at least some prompt dimensions, and the dimension identifiers are used to trigger the execution of the prompt dimension switching operation. The method of executing the prompt dimension switching operation by triggering the dimension identifier is not limited, such as the prompt dimension switching operation can be executed by clicking the dimension identifier, or the prompt dimension switching operation can be executed by sliding the dimension identifier to a preset trigger position.
[0081] Specifically, when a prompt dimension switching operation is received, the prompt dimension displayed in the prompt information display area can be switched in response to the prompt dimension switching operation, such as at least one prompt information group displayed in the prompt information display interface can be switched to a prompt information group in the prompt dimension corresponding to the prompt dimension switching operation.
[0082] In some embodiments, the user may also input the prompt information in the prompt information input area by text input. At this time, optionally, the target prompt information further includes second prompt information, and the second prompt information is input into the prompt information input area by character input.
[0083] In the case where prompt information (such as the second prompt information) has been input in the prompt information input area, when displaying at least one prompt information group, the at least one prompt information group can be pre-set or determined based on the input prompt information.
[0084] Optionally, the at least one prompt information group is a first prompt information group, and the first prompt information group is different from the second prompt information group to which the second prompt information belongs; or, the at least one prompt information group includes a first prompt information group and a second prompt information group, and the first prompt information group is displayed before the second prompt information group.
[0085] In some embodiments, the first prompt information group can be understood as a prompt information group that is different from the second prompt information group, such as a prompt information group other than the second prompt information group. The second prompt information group can be a prompt information group to which the second prompt information belongs, such as a prompt information group associated with the second prompt information. The number of second prompt information groups can be one or more. Taking the example of at least one prompt information group including five prompt information groups of clothing, hairstyle, accessories, posture and expression, if the second prompt information includes descriptive information of clothing and hairstyle, then both the clothing and hairstyle prompt information groups are the second prompt information group. In this case, the accessories, posture and expression prompt information groups can be the first prompt information group.
[0086] Taking the example of the second prompt information being entered in the prompt information input area, when displaying at least one prompt information group, the second prompt information group to which the second prompt information belongs can be determined, and at least one prompt information group other than the second prompt information group can be displayed, that is, only at least one first prompt information group is displayed, and the second prompt information group is not displayed; at least one first prompt information group and at least one second prompt information group can also be displayed, such as displaying at least one first prompt information group and displaying at least one second prompt information group after the at least one first prompt information group, so that the user can select the candidate prompt information in the prompt information group that has not yet been entered.
[0087] In some embodiments, the candidate prompt information in a selected state can be used as the first prompt information. In this case, optionally, in response to an information selection operation on at least one of the candidate prompt information in the at least one prompt information group, determining the first prompt information in the candidate prompt information includes: in response to the information selection operation on at least one of the candidate prompt information in the at least one prompt information group, displaying the target candidate prompt information in a selected state, the target candidate prompt information being the candidate prompt information corresponding to the information selection operation; and in response to a first confirmation operation, determining the candidate prompt information in the selected state as the first prompt information.
[0088] In some embodiments, the information selection operation can be a triggering operation for selecting a displayed candidate prompt. The information selection operation can be performed in any manner. For example, the information selection operation can be an operation of clicking or long pressing a candidate prompt. The target candidate prompt can be the candidate prompt corresponding to the information selection operation, i.e., the candidate prompt selected by the information selection operation. The first confirmation operation can be a triggering operation for confirming the selection of the selected candidate prompt, such as triggering a first confirmation control or performing a first confirmation gesture.
[0089] Exemplarily, after displaying at least one prompt information group, when a user's information selection operation is received for at least one candidate prompt information (such as target candidate prompt information) in the at least one prompt information group, such as when it is detected that the user triggers a candidate prompt information contained in a certain prompt information group, the candidate prompt information can be displayed as a selected state; and when the user's first confirmation operation is received, the candidate prompt information in the selected state is determined as the first prompt information.
[0090] In the above embodiment, the method of displaying the candidate prompt information as selected is not limited. For example, the candidate prompt information can be displayed as selected by adjusting the display style or display position of the candidate prompt information, or adding a method of displaying the candidate prompt information in a certain area.
[0091] Optionally, displaying the target candidate prompt information as a selected state includes: displaying the target candidate prompt information displayed in the first area as a preset display style, the first area being used to display the at least one prompt information group; and / or, adding the display of the target candidate prompt information in the second area, the second area being used to display the candidate prompt information in a selected state.
[0092] For example, when an information selection operation for target candidate prompt information is received, the target candidate prompt information can be displayed in a preset display style corresponding to the selected state; and / or, the target candidate prompt information can be added to be displayed in the second area, so that the target candidate prompt information is displayed in a selected state.
[0093] In some embodiments, the first area may be an area for displaying at least one prompt information group, and the second area may be an area for displaying candidate prompt information selected by the user. The first area and the second area may be different display areas, such as different subareas of the prompt information display area. The relative positional relationship between the first and second areas is not limited; for example, the second area may be located within or outside the first area. Selected candidate prompt information and unselected candidate prompt information may have different display styles, and the preset display style may be the display style of the selected candidate prompt information.
[0094] In the case where the second prompt information already exists, when the user selects candidate prompt information, the relationship between the candidate prompt information selected by the user and the existing second prompt information may be considered or not.
[0095] In some embodiments, the displaying of the target candidate prompt information as a selected state includes at least one of the following: if the target candidate prompt information meets a preset selection condition, the target candidate prompt information is displayed as a selected state; if the target candidate prompt information does not meet the preset selection condition, the selection confirmation information is displayed; in response to a second confirmation operation based on the selection confirmation information, the target candidate prompt information is displayed as a selected state, and the prompt information in the target area that causes the target candidate prompt information to not meet the preset selection condition is deleted, and the target area includes the prompt information input area and / or the second area; wherein the preset condition includes: there is no duplicate prompt information between the target candidate prompt information and the prompt information in the target area, and / or there is no contradictory prompt information between the target candidate prompt information and the prompt information in the target area.
[0096] In some embodiments, the second confirmation operation may be a triggering operation for indicating confirmation of the target candidate selection prompt information. The execution method of the second confirmation operation is not limited, such as the second confirmation operation may be an operation of triggering a second confirmation control or performing a second confirmation gesture.
[0097] Duplicate prompt information can be understood as repeated prompt information. Contradictory prompt information can be understood as contradictory prompt information. It should be noted that the repetition or contradiction here can be semantic repetition or contradiction. Taking hairstyle as an example, if the second prompt information already contains the relevant expression "short hair", at this time, if the candidate prompt information selected by the user is "short hair", then the candidate prompt information can be considered to be repeated with the second prompt information; if the candidate prompt information selected by the user is "long hair", then the candidate prompt information can be considered to be contradictory with the second prompt information.
[0098] In the above embodiment, when the candidate prompt information selected by the user is repeated with the prompt information already existing in the target area, the user can be prompted to improve the simplicity of the prompt information in the target area; and / or, when the candidate prompt information selected by the user is inconsistent with the prompt information already existing in the target area, the user can be prompted to avoid errors in generating target media content and improve the picture effect of the generated target media content.
[0099] For example, upon receiving an information selection operation for target candidate prompt information, it may be determined whether the target candidate prompt information satisfies a preset selection condition. If so, the target candidate prompt information is displayed as selected. If not, selection confirmation information is displayed, prompting the user through the selection confirmation information that the target candidate prompt information is duplicated or contradictory with the prompt information in the target area; and upon receiving a second confirmation operation based on the selection confirmation information, in response to the second confirmation operation, the target candidate prompt information is displayed as selected, and prompt information already existing in the target area that is duplicated or contradictory with the target candidate prompt information may be further deleted.
[0100] It should be noted that after prompting the user, the user can also delete the prompt information that causes the target candidate prompt information to not meet the preset selection conditions by performing a delete operation, which can be set specifically as needed.
[0101] FIG4 is a flow chart of another method for generating media content provided by an embodiment of the present disclosure. The solution in this embodiment can be combined with one or more optional solutions in the above embodiments. Optionally, before displaying the target media content generated based on the target prompt information, it also includes: in response to a polishing operation on the prompt information to be polished in the target prompt information, obtaining polishing prompt information corresponding to the prompt information to be polished, the polishing prompt information being obtained by polishing the prompt information to be polished; and replacing the prompt information to be polished with the polishing prompt information.
[0102] Accordingly, as shown in FIG4 , the method for generating media content provided in this embodiment may include:
[0103] S201: Display at least one prompt information group, where the prompt information group includes at least one candidate prompt information, and the candidate prompt information is used to describe content features of target media content.
[0104] S202: In response to an information selection operation on at least one candidate prompt information in the at least one prompt information group, determine first prompt information from the candidate prompt information.
[0105] S203: In response to a polishing operation on prompt information to be polished in target prompt information, obtain polishing prompt information corresponding to the prompt information to be polished, wherein the target prompt information includes the first prompt information, and the polishing prompt information is obtained by polishing the prompt information to be polished.
[0106] In some embodiments, the touch-up operation may be an operation for instructing to touch-up at least a portion of the target prompt information. The touch-up operation may be performed in any manner, such as triggering a touch-up control or performing a touch-up gesture.
[0107] The prompt information to be polished may be prompt information corresponding to the polishing operation, that is, prompt information to be polished in the target prompt information. The prompt information to be polished may be at least part of the prompt information in the target prompt information. Optionally, the prompt information to be polished is the prompt information in the target prompt information that is selected; or, the prompt information to be polished is the prompt information in the target prompt information that is not selected. For example, the prompt information selected by the user in the target prompt information may be used as the prompt information to be polished; or, the prompt information selected by the user in the target prompt information may be used as content that does not require polishing, that is, the prompt information in the target prompt information that is not selected by the user may be used as the prompt information to be polished.
[0108] The polishing prompt information can be understood as the prompt information obtained by polishing the polishing prompt information, that is, the prompt information after polishing. There is no limit to the method for polishing the polishing prompt information, such as modifying the polishing prompt information or expanding the polishing prompt information to supplement the image description dimensions not included in the prompt information input area.
[0109] In the above embodiment, the user may be supported to polish part or all of the set target prompt information to further improve the integrity of the target prompt information, thereby improving the quality of the generated target media content.
[0110] Specifically, when a polishing operation is received, the target prompt information corresponding to the polishing operation can be obtained. The target prompt information can be polished or polished by the server to obtain the polished prompt information corresponding to the target prompt information.
[0111] S204: Replace the prompt information to be polished with the polishing prompt information.
[0112] Specifically, after obtaining the polishing prompt information corresponding to the prompt information to be polished, the prompt information to be polished in the target prompt information may be replaced with the obtained polishing prompt information.
[0113] S205: Display the target media content generated based on the target prompt information.
[0114] Illustratively, after the prompt information to be polished in the target prompt information is replaced with the polished prompt information, target media content generated by the target prompt information replaced with the polished prompt information may be obtained and displayed.
[0115] FIG5 is a schematic diagram of an optional prompt information input process provided by an embodiment of the present disclosure. As shown in FIG3 and FIG5 , the method for generating media content provided by this embodiment can be described as follows:
[0116] The media content generation interface is displayed, and the original media content 50 set by the user for generating the target media content can be displayed in the media content generation interface, and a prompt information area 51 can be further displayed, as shown in the first figure in FIG5 .
[0117] When it is detected that the user triggers the prompt information area 51, it is determined that the prompt information input operation is received, and in response to the prompt information input operation, the prompt information input interface 52 is displayed, and the prompt information input area 53 and dimension information 54 of at least one prompt dimension are displayed in the prompt information input interface, as shown in the second figure in Figure 5.
[0118] When it is detected that a user triggers a certain dimension information 54 (such as the dimension information of the first prompt dimension) displayed in the prompt information input interface, it can be determined that the prompt information display operation is received, and in response to the prompt information display operation, the prompt information display area 21 is displayed (such as the prompt information display interface is displayed), and at least one prompt information group 20 under the prompt dimension is displayed in the prompt information display area 21, and the dimension identifier 22 of at least part of the prompt dimension can be further displayed in the prompt information display area 21, as shown in the third figure in Figure 5.
[0119] Thus, when it is detected that the user triggers the dimension identifier of a certain prompt dimension by clicking or sliding left or right, it can be determined that the prompt dimension switching operation has been received, and in response to the prompt dimension switching operation, at least one prompt information group in the prompt dimension corresponding to the dimension identifier is displayed in the prompt information display area. When it is detected that the user switches the prompt information group displayed in the prompt information display area by sliding up or down, the prompt information group can have an adsorption effect. For example, when a prompt information group moves to a position within a preset distance range from a preset position at the top of the prompt information display area, if the user releases his finger, the prompt information group can move to the preset position for display.
[0120] When it is detected that a user clicks on a displayed candidate prompt information, it can be determined that an information selection operation for the candidate prompt information is received, and in response to the information selection operation, the candidate prompt information is displayed as a selected state, such as displaying the candidate prompt information as a preset display style in the first area 55, and adding the candidate prompt information to be displayed in the second area 56, such as candidate prompt information 10, candidate prompt information 22 and candidate prompt information 31 shown in the fourth figure of Figure 5.
[0121] In addition, when the candidate prompt information selected by the user is displayed in the second area 56, a deselect control 561 corresponding to the candidate prompt information can be further displayed, so that the user can switch the candidate prompt information from a selected state to an unselected state by triggering the deselect control 561 corresponding to the candidate prompt information.
[0122] As shown in the fourth figure in FIG5 , a first selection confirmation control 562 may be displayed in the second area 56. Therefore, when it is detected that the user triggers the first selection confirmation control 562, it can be determined that the first confirmation operation has been received. In response to the first confirmation operation, the candidate prompt information in the selected state (i.e., the first prompt information) is determined, and the candidate prompt information in the selected state in the prompt information display interface 21 is input into the prompt information input area 52, as shown in the fifth figure in FIG5 .
[0123] As shown in FIG5 , a polishing control 57 may also be displayed in the prompt information input interface 52. When a user triggers this polishing control 57, it is determined that a polishing operation has been received. In response to this polishing operation, polished prompt information corresponding to the prompt information to be polished entered in the prompt information input area 52 is retrieved, and the prompt information to be polished displayed in the prompt information input area 52 is replaced with the polished prompt information (as shown in the sixth figure in FIG5 ). For example, when a user clicks on the polishing control 57, the input prompt is polished using a pre-set polishing model or polishing module. In some embodiments, the polishing process may utilize streaming processing; if the polishing fails or is abnormal, a prompt may be provided to the user.
[0124] 5 , a media content generation control 58 may also be displayed in the prompt information input interface 52. When it is detected that the user triggers the media content generation control 58, it can be determined that a media content generation operation has been received. In response to the media content generation operation, target media content 30 generated based on the target prompt information currently input in the prompt information input area 52 and the original media content is obtained and displayed, as shown in FIG3 .
[0125] It is understandable that when the prompt information input area is displayed, the user can also input the second prompt information in the prompt information input area by inputting characters.
[0126] The media content generation method provided in this embodiment uses a structured approach to display candidate prompts in dimensions and / or groups, such as displaying the subject, scene, and / or color of the prompt in a structured manner, and supports polishing the input prompt, which can improve the accuracy of the prompt filled in by the user and lower the threshold for filling in the prompt.
[0127] FIG6 is a block diagram of a media content generation device provided by an embodiment of the present disclosure. The device can be implemented by software and / or hardware, and can be configured in an electronic device, typically a mobile phone or a tablet computer, and can implement AIGC by executing a method for generating media content, such as implementing a text graph. As shown in FIG6 , the media content generation device provided by this embodiment may include: a prompt information display module 601, a prompt information determination module 602, and a media content display module 603, wherein,
[0128] The prompt information display module 601 is used to display at least one prompt information group, wherein the prompt information group includes at least one candidate prompt information, and the candidate prompt information is a screen description content to be selected;
[0129] a prompt information determining module 602, configured to input the first prompt information into a prompt information input area in response to an information selection operation on the first prompt information in the at least one prompt information group;
[0130] The media content display module 603 is configured to display target media content generated based on target prompt information in response to a media content generation operation, where the target prompt information is the prompt information input in the prompt information input area, and the target prompt information includes the first prompt information.
[0131] The media content generation device provided in this embodiment displays at least one prompt information group through a prompt information display module, wherein the prompt information group includes at least one candidate prompt information, and the candidate prompt information is used to describe the content characteristics of the target media content; the prompt information determination module responds to an information selection operation for at least one candidate prompt information in the at least one prompt information group, and determines the first prompt information in the candidate prompt information; and the media content display module displays the target media content generated based on the target prompt information, wherein the target prompt information includes the first prompt information. This embodiment utilizes the above-mentioned technical solution to support users in selecting candidate prompt information from the prompt information group as prompt information for the media content to be generated, without relying entirely on user input through text input. This can reduce the difficulty and uncertainty of inputting prompt information and improve the quality of the generated media content.
[0132] Optionally, the media content screen of the target media content is obtained by processing the media content screen of the original media content based on the target prompt information; the media content generation device may also include: a media content acquisition module, used to obtain the original media content to be processed before displaying at least one prompt information group; the prompt information display module 601 can be specifically used to: display at least one prompt information group corresponding to the original media content.
[0133] Optionally, different original media contents correspond to different prompt information groups; or, the prompt information groups corresponding to different original media contents contain different candidate prompt information; or, different original media contents correspond to different display orders, and the display order includes the display order of the at least one prompt information group and / or the display order of the candidate prompt information in the prompt information group.
[0134] Optionally, the at least one prompt information group belongs to the same prompt dimension, and different prompt dimensions correspond to different description dimensions of the content feature.
[0135] Optionally, the prompt information display module 601 can be specifically used to: display different prompt information groups in different sub-areas of the prompt information display area; wherein different prompt information groups include different types of candidate prompt information, and the different types of candidate prompt information are used to describe different types of attribute information of the target media content.
[0136] Furthermore, the media content generation device may also include: an input area display module, which is used to display the prompt information input area and dimensional information of at least one prompt dimension in response to the prompt information input operation before displaying at least one prompt information group, and the dimensional information is used to trigger the display of at least one prompt information group in the prompt dimension corresponding to the dimensional information.
[0137] Furthermore, the prompt information determining module 602 may also be configured to: after determining the first prompt information from the candidate prompt information, display the first prompt information in the prompt information input area.
[0138] Optionally, the target prompt information further includes second prompt information, and the second prompt information is input into the prompt information input area by character input.
[0139] Optionally, the at least one prompt information group is a first prompt information group, and the first prompt information group is different from the second prompt information group to which the second prompt information belongs; or, the at least one prompt information group includes a first prompt information group and a second prompt information group, and the first prompt information group is displayed before the second prompt information group.
[0140] Optionally, the prompt information determination module 602 includes: a selection unit for displaying the target candidate prompt information as a selected state in response to an information selection operation for at least one of the candidate prompt information in the at least one prompt information group, wherein the target candidate prompt information is the candidate prompt information corresponding to the information selection operation; and a determination unit for determining the candidate prompt information in the selected state as the first prompt information in response to a first confirmation operation.
[0141] Optionally, the selection unit is specifically used to: display the target candidate prompt information displayed in the first area as a preset display style, the first area is used to display the at least one prompt information group; and / or, add the target candidate prompt information to be displayed in the second area, the second area is used to display the candidate prompt information in the selected state.
[0142] Optionally, the selection unit is specifically used to perform at least one of the following: if the target candidate prompt information meets the preset selection condition, the target candidate prompt information is displayed as a selected state; if the target candidate prompt information does not meet the preset selection condition, the selection confirmation information is displayed; in response to a second confirmation operation based on the selection confirmation information, the target candidate prompt information is displayed as a selected state, and the prompt information in the target area that causes the target candidate prompt information to not meet the preset selection condition is deleted, and the target area includes the prompt information input area and / or the second area; wherein the preset condition includes: there is no duplicate prompt information between the target candidate prompt information and the prompt information in the target area, and / or there is no contradictory prompt information between the target candidate prompt information and the prompt information in the target area.
[0143] Furthermore, the media content generation device may further include: a polishing module for, before displaying the target media content generated based on the target prompt information, in response to a polishing operation on the prompt information to be polished in the target prompt information, obtaining polishing prompt information corresponding to the prompt information to be polished, wherein the polishing prompt information is obtained by polishing the prompt information to be polished; and a replacement module for replacing the prompt information to be polished with the polishing prompt information.
[0144] Optionally, the prompt information to be polished is the prompt information in the target prompt information that is in a selected state; or, the prompt information to be polished is the prompt information in the target prompt information that is not in a selected state.
[0145] Furthermore, the media content generation device may also include: a region selection module, which is used to display the prompt information group associated with the selected region in the at least one prompt information group in a first state in response to a region selection operation acting on the original media content and / or the target media content, and to display the prompt information groups in the at least one prompt information group other than the prompt information group associated with the selected region in a second state.
[0146] The media content generation device provided in the embodiments of the present disclosure can execute the media content generation method provided in any embodiment of the present disclosure, and has the corresponding functional modules and beneficial effects of executing the media content generation method. For technical details not fully described in this embodiment, please refer to the media content generation method provided in any embodiment of the present disclosure.
[0147] Reference is now made to FIG7 , which illustrates a schematic diagram of the structure of an electronic device (e.g., a terminal device) 700 suitable for implementing embodiments of the present disclosure. The terminal device in the embodiments of the present disclosure may include, but is not limited to, mobile terminals such as mobile phones, laptop computers, digital broadcast receivers, PDAs (personal digital assistants), PADs (tablet computers), PMPs (portable multimedia players), in-vehicle terminals (e.g., in-vehicle navigation terminals), and fixed terminals such as digital TVs and desktop computers. The electronic device illustrated in FIG7 is merely an example and should not limit the functionality or scope of use of the embodiments of the present disclosure.
[0148] As shown in Figure 7, electronic device 700 may include a processing device (e.g., a central processing unit, a graphics processing unit, etc.) 701, which can perform various appropriate actions and processes according to a program stored in a read-only memory (ROM) 702 or a program loaded from a storage device 708 into a random access memory (RAM) 703. Various programs and data required for the operation of electronic device 700 are also stored in RAM 703. Processing device 701, ROM 702, and RAM 703 are connected to each other via a bus 704. An input / output (I / O) interface 705 is also connected to bus 704.
[0149] Typically, the following devices may be connected to the I / O interface 705: an input device 706 including, for example, a touch screen, a touchpad, a keyboard, a mouse, a camera, a microphone, an accelerometer, a gyroscope, etc.; an output device 707 including, for example, a liquid crystal display (LCD), a speaker, a vibrator, etc.; a storage device 708 including, for example, a magnetic tape, a hard disk, etc.; and a communication device 709. The communication device 709 may allow the electronic device 700 to communicate with other devices wirelessly or by wire to exchange data. Although FIG. 7 shows the electronic device 700 with various devices, it should be understood that not all of the devices shown are required to be implemented or present. More or fewer devices may be implemented or present instead.
[0150] In particular, according to an embodiment of the present disclosure, the process described above with reference to the flowchart can be implemented as a computer software program. For example, an embodiment of the present disclosure includes a computer program product, which includes a computer program carried on a non-transitory computer-readable medium, and the computer program includes a program code for executing the method shown in the flowchart. In such an embodiment, the computer program can be downloaded and installed from the network through the communication device 709, or installed from the storage device 708, or installed from the ROM 702. When the computer program is executed by the processing device 701, the above-mentioned functions defined in the method of the embodiment of the present disclosure are performed.
[0151] It should be noted that the computer-readable medium mentioned above in the present disclosure may be a computer-readable signal medium or a computer-readable storage medium, or any combination of the two. A computer-readable storage medium may be, for example, but not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, device, or component, or any combination of the above. More specific examples of computer-readable storage media may include, but are not limited to: an electrical connection with one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the above. In the present disclosure, a computer-readable storage medium may be any tangible medium that contains or stores a program that can be used by or in conjunction with an instruction execution system, device, or component. In the present disclosure, a computer-readable signal medium may include a data signal propagated in baseband or as part of a carrier wave, which carries computer-readable program code. Such a propagated data signal may take a variety of forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination of the above. A computer-readable signal medium may also be any computer-readable medium other than a computer-readable storage medium that can transmit, propagate, or transport a program for use by or in conjunction with an instruction execution system, apparatus, or device. The program code contained on the computer-readable medium may be transmitted using any suitable medium, including but not limited to wires, optical cables, RF (radio frequency), etc., or any suitable combination thereof.
[0152] In some embodiments, the client and server can communicate using any currently known or later developed network protocol, such as HTTP (HyperText Transfer Protocol), and can be interconnected with any form or medium of digital data communication (e.g., a communication network). Examples of communication networks include a local area network ("LAN"), a wide area network ("WAN"), an internet (e.g., the Internet), and a peer-to-peer network (e.g., an ad hoc peer-to-peer network), as well as any currently known or later developed network.
[0153] The computer-readable medium may be included in the electronic device, or may exist independently without being incorporated into the electronic device.
[0154] The above-mentioned computer-readable medium carries one or more programs. When the above-mentioned one or more programs are executed by the electronic device, the electronic device: displays at least one prompt information group, wherein the prompt information group includes at least one candidate prompt information, and the candidate prompt information is used to describe the content characteristics of the target media content; in response to an information selection operation for at least one of the candidate prompt information in the at least one prompt information group, determines the first prompt information in the candidate prompt information; and displays the target media content generated based on the target prompt information; wherein the target prompt information includes the first prompt information.
[0155] Computer program code for performing the operations of the present disclosure may be written in one or more programming languages, or a combination thereof, including, but not limited to, object-oriented programming languages such as Java, Smalltalk, C++, and conventional procedural programming languages such as "C" or similar programming languages. The program code may be executed entirely on the user's computer, partially on the user's computer, as a stand-alone software package, partially on the user's computer and partially on a remote computer, or entirely on the remote computer or server. In cases involving a remote computer, the remote computer may be connected to the user's computer through any type of network, including a local area network (LAN) or a wide area network (WAN), or may be connected to an external computer (e.g., through the Internet using an Internet service provider).
[0156] The flowcharts and block diagrams in the accompanying drawings illustrate the possible implementation architecture, functions and operations of the systems, methods and computer program products according to various embodiments of the present disclosure. In this regard, each box in the flowchart or block diagram can represent a module, program segment, or a part of code, and the module, program segment, or a part of code contains one or more executable instructions for realizing the specified logical function. It should also be noted that in some alternative implementations, the functions marked in the box can also occur in a different order than that marked in the accompanying drawings. For example, two boxes represented in succession can actually be executed substantially in parallel, and they can sometimes be executed in the opposite order, depending on the functions involved. It should also be noted that each box in the block diagram and / or flowchart, and the combination of the boxes in the block diagram and / or flowchart, can be implemented with a dedicated hardware-based system that performs the specified function or operation, or can be implemented with a combination of dedicated hardware and computer instructions.
[0157] The units involved in the embodiments described in this disclosure may be implemented in software or hardware. The name of a module does not necessarily limit the unit itself.
[0158] The functions described above herein may be performed, at least in part, by one or more hardware logic components. For example, and without limitation, exemplary types of hardware logic components that may be used include: field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), systems on chip (SOCs), complex programmable logic devices (CPLDs), and the like.
[0159] In the context of the present disclosure, a machine-readable medium can be a tangible medium that can contain or store a program for use by or in conjunction with an instruction execution system, device or equipment. A machine-readable medium can be a machine-readable signal medium or a machine-readable storage medium. A machine-readable medium can include, but is not limited to, an electronic, magnetic, optical, electromagnetic, infrared, or semiconductor system, device or equipment, or any suitable combination of the foregoing. A more specific example of a machine-readable storage medium can include an electrical connection based on one or more lines, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the foregoing.
[0160] According to one or more embodiments of the present disclosure, Example 1 provides a method for generating media content, including:
[0161] Displaying at least one prompt information group, wherein the prompt information group includes at least one candidate prompt information, and the candidate prompt information is used to describe content characteristics of target media content;
[0162] In response to an information selection operation on at least one of the candidate prompt information in the at least one prompt information group, determining first prompt information from the candidate prompt information;
[0163] The target media content generated based on the target prompt information is displayed; wherein the target prompt information includes the first prompt information.
[0164] According to one or more embodiments of the present disclosure, Example 2 is the method according to Example 1, wherein the media content screen of the target media content is obtained by processing the media content screen of the original media content based on the target prompt information; and before presenting at least one prompt information group, the method further includes:
[0165] Obtaining the original media content to be processed;
[0166] The displaying of at least one prompt information group includes:
[0167] At least one prompt information group corresponding to the original media content is displayed.
[0168] According to one or more embodiments of the present disclosure, Example 3 is the method according to Example 2, wherein different original media contents correspond to different prompt information groups; or
[0169] The prompt information groups corresponding to different original media contents contain different candidate prompt information; or
[0170] Different original media contents correspond to different presentation orders, and the presentation order includes the presentation order of the at least one prompt information group and / or the presentation order of candidate prompt information in the prompt information group.
[0171] According to one or more embodiments of the present disclosure, Example 4 is based on the method described in Example 1, wherein the at least one prompt information group belongs to the same prompt dimension, and different prompt dimensions correspond to different description dimensions of the content feature.
[0172] According to one or more embodiments of the present disclosure, Example 5 is the method according to Example 4, wherein presenting at least one prompt information group includes:
[0173] Different prompt information groups are displayed in different sub-areas of the prompt information display area; wherein different prompt information groups include different types of candidate prompt information, and the different types of candidate prompt information are used to describe different types of attribute information of the target media content.
[0174] According to one or more embodiments of the present disclosure, Example 6 is the method according to Example 4, and further includes, before presenting at least one prompt information group:
[0175] In response to the prompt information input operation, the prompt information input area and dimension information of at least one prompt dimension are displayed, and the dimension information is used to trigger the display of at least one prompt information group in the prompt dimension corresponding to the dimension information.
[0176] According to one or more embodiments of the present disclosure, Example 7, according to the method of Example 1, further includes, after determining the first prompt information in the candidate prompt information:
[0177] The first prompt information is displayed in the prompt information input area.
[0178] According to one or more embodiments of the present disclosure, Example 8 is based on the method described in Example 1, wherein the target prompt information further includes second prompt information, and the second prompt information is input into the prompt information input area by character input.
[0179] According to one or more embodiments of the present disclosure, Example 9 is the method according to Example 8, wherein the at least one prompt information group is a first prompt information group, and the first prompt information group is different from the second prompt information group to which the second prompt information belongs; or
[0180] The at least one prompt information group includes a first prompt information group and a second prompt information group, wherein the first prompt information group is displayed before the second prompt information group.
[0181] According to one or more embodiments of the present disclosure, Example 10 is the method according to Example 8, wherein, in response to an information selection operation for at least one of the candidate prompt information in the at least one prompt information group, determining the first prompt information from the candidate prompt information includes:
[0182] In response to an information selection operation on at least one candidate prompt information in the at least one prompt information group, displaying target candidate prompt information as a selected state, the target candidate prompt information being the candidate prompt information corresponding to the information selection operation;
[0183] In response to the first confirmation operation, the selected candidate prompt information is determined as the first prompt information.
[0184] According to one or more embodiments of the present disclosure, Example 11 is the method according to Example 10, wherein displaying the target candidate prompt information as a selected state includes:
[0185] Displaying the target candidate prompt information displayed in the first area in a preset display style, the first area being used to display the at least one prompt information group; and / or
[0186] The target candidate prompt information is additionally displayed in the second area, and the second area is used to display the candidate prompt information in a selected state.
[0187] According to one or more embodiments of the present disclosure, Example 12 is the method according to Example 10, wherein displaying the target candidate prompt information as a selected state includes at least one of the following:
[0188] If the target candidate prompt information meets the preset selection condition, the target candidate prompt information is displayed as a selected state;
[0189] If the target candidate prompt information does not meet the preset selection condition, displaying selection confirmation information; in response to a second confirmation operation based on the selection confirmation information, displaying the target candidate prompt information as selected, and deleting the prompt information in the target area that causes the target candidate prompt information to not meet the preset selection condition, the target area including the prompt information input area and / or the second area;
[0190] The preset conditions include: there is no duplicate prompt information between the target candidate prompt information and the prompt information in the target area, and / or there is no contradictory prompt information between the target candidate prompt information and the prompt information in the target area.
[0191] According to one or more embodiments of the present disclosure, Example 13 is a method according to any one of Examples 1-12, and further includes, before presenting the target media content generated based on the target prompt information:
[0192] In response to a polishing operation on the prompt information to be polished in the target prompt information, obtaining polishing prompt information corresponding to the prompt information to be polished, wherein the polishing prompt information is obtained by polishing the prompt information to be polished;
[0193] The prompt information to be polished is replaced with the polishing prompt information.
[0194] According to one or more embodiments of the present disclosure, Example 14 is the method according to Example 13, wherein the prompt information to be polished is the prompt information in the target prompt information that is in a selected state; or
[0195] The prompt information to be polished is the prompt information that is not in a selected state in the target prompt information.
[0196] According to one or more embodiments of the present disclosure, Example 15 is a method according to any one of Examples 1-12, further comprising:
[0197] In response to an area selection operation performed on the original media content and / or the target media content, the prompt information group associated with the selected area in the at least one prompt information group is displayed in a first state, and the prompt information groups other than the prompt information group associated with the selected area in the at least one prompt information group are displayed in a second state.
[0198] According to one or more embodiments of the present disclosure, Example 16 provides a device for generating media content, including:
[0199] A prompt information display module, configured to display at least one prompt information group, wherein the prompt information group includes at least one candidate prompt information, and the candidate prompt information is used to describe content characteristics of target media content;
[0200] a prompt information determining module, configured to determine first prompt information from the candidate prompt information in response to an information selection operation on at least one of the candidate prompt information in the at least one prompt information group;
[0201] The media content display module is configured to display the target media content generated based on the target prompt information; wherein the target prompt information includes the first prompt information.
[0202] According to one or more embodiments of the present disclosure, Example 17 provides an electronic device, including:
[0203] one or more processors;
[0204] a memory for storing one or more programs,
[0205] When the one or more programs are executed by the one or more processors, the one or more processors implement the method for generating media content as described in any one of Examples 1-15.
[0206] According to one or more embodiments of the present disclosure, Example 18 provides a computer-readable storage medium having a computer program stored thereon, which, when executed by a processor, implements the method for generating media content as described in any one of Examples 1-15.
[0207] According to one or more embodiments of the present disclosure, Example 19 provides a computer program product. When the computer program product is executed by a computer, the computer implements the method for generating media content as described in any one of Examples 1-15.
[0208] The above description is merely a preferred embodiment of the present disclosure and an illustration of the technical principles employed. Those skilled in the art should understand that the scope of disclosure involved in the present disclosure is not limited to the technical solutions formed by the specific combination of the above-mentioned technical features, but also includes other technical solutions formed by any combination of the above-mentioned technical features or their equivalents without departing from the above-mentioned disclosed concepts. For example, a technical solution formed by replacing the above-mentioned features with (but not limited to) technical features with similar functions disclosed in this disclosure.
[0209] In addition, although each operation is described in a specific order, this should not be understood as requiring these operations to be performed in the specific order shown or in a sequential order. Under certain circumstances, multitasking and parallel processing may be advantageous. Similarly, although some specific implementation details have been included in the above discussion, these should not be interpreted as limiting the scope of the present disclosure. Some features described in the context of a separate embodiment can also be implemented in a single embodiment in combination. On the contrary, the various features described in the context of a single embodiment can also be implemented in multiple embodiments individually or in any suitable sub-combination mode.
[0210] Although the subject matter has been described in language specific to structural features and / or methodological logical acts, it should be understood that the subject matter defined in the appended claims is not necessarily limited to the specific features or acts described above. Rather, the specific features and acts described above are merely example forms of implementing the claims.
Claims
1. A method for generating media content, comprising: Displaying at least one prompt information group, wherein the prompt information group includes at least one candidate prompt information, and the candidate prompt information is used to describe content characteristics of target media content; In response to an information selection operation on at least one of the candidate prompt information in the at least one prompt information group, determining first prompt information from the candidate prompt information; The target media content generated based on the target prompt information is displayed; wherein the target prompt information includes the first prompt information.
2. The method according to claim 1, wherein the media content picture of the target media content is obtained by processing the media content picture of the original media content based on the target prompt information; Before presenting at least one prompt information group, the method further includes: Obtaining the original media content to be processed; The displaying of at least one prompt information group includes: At least one prompt information group corresponding to the original media content is displayed.
3. The method according to claim 2, wherein different original media contents correspond to different prompt information groups; or The prompt information groups corresponding to different original media contents contain different candidate prompt information; or Different original media contents correspond to different presentation orders, and the presentation order includes the presentation order of the at least one prompt information group and / or the presentation order of candidate prompt information in the prompt information group. 4 . The method according to claim 1 , wherein the at least one prompt information group belongs to the same prompt dimension, and different prompt dimensions correspond to different description dimensions of the content feature.
5. The method according to claim 4, wherein presenting at least one prompt information group comprises: Different prompt information groups are displayed in different sub-areas of the prompt information display area; wherein different prompt information groups include different types of candidate prompt information, and the different types of candidate prompt information are used to describe different types of attribute information of the target media content.
6. The method according to claim 4, wherein before presenting at least one prompt information group, the method further comprises: In response to the prompt information input operation, the prompt information input area and dimension information of at least one prompt dimension are displayed, and the dimension information is used to trigger the display of at least one prompt information group in the prompt dimension corresponding to the dimension information.
7. The method according to claim 1, wherein after determining the first prompt information in the candidate prompt information, the method further comprises: The first prompt information is displayed in the prompt information input area. 8 . The method according to claim 1 , wherein the target prompt information further includes second prompt information, and the second prompt information is input into the prompt information input area by character input.
9. The method according to claim 8, wherein the at least one prompt information group is a first prompt information group, and the first prompt information group is different from the second prompt information group to which the second prompt information belongs; or The at least one prompt information group includes a first prompt information group and a second prompt information group, wherein the first prompt information group is displayed before the second prompt information group.
10. The method according to claim 8, wherein in response to an information selection operation on at least one of the candidate prompt information in the at least one prompt information group, determining the first prompt information from the candidate prompt information comprises: In response to information about at least one candidate prompt information in the at least one prompt information group A selection operation is performed to display the target candidate prompt information as a selected state, wherein the target candidate prompt information is the candidate prompt information corresponding to the information selection operation; In response to the first confirmation operation, the selected candidate prompt information is determined as the first prompt information.
11. The method according to claim 10, wherein displaying the target candidate prompt information as a selected state comprises: Displaying the target candidate prompt information displayed in a first area in a preset display style, wherein the first area is used to display the at least one prompt information group; and / or The target candidate prompt information is additionally displayed in the second area, and the second area is used to display the candidate prompt information in a selected state.
12. The method according to claim 10, wherein displaying the target candidate prompt information as a selected state comprises at least one of the following: If the target candidate prompt information meets the preset selection condition, the target candidate prompt information is displayed as a selected state; If the target candidate prompt information does not meet the preset selection condition, displaying selection confirmation information; in response to a second confirmation operation based on the selection confirmation information, displaying the target candidate prompt information as selected, and deleting the prompt information in the target area that causes the target candidate prompt information to not meet the preset selection condition, the target area including the prompt information input area and / or the second area; in, The preset conditions include: there is no duplicate prompt information between the target candidate prompt information and the prompt information in the target area, and / or there is no contradictory prompt information between the target candidate prompt information and the prompt information in the target area.
13. The method according to any one of claims 1 to 12, wherein before presenting the target media content generated based on the target prompt information, the method further comprises: In response to a polishing operation on the prompt information to be polished in the target prompt information, obtaining polishing prompt information corresponding to the prompt information to be polished, wherein the polishing prompt information is obtained by polishing the prompt information to be polished; The prompt information to be polished is replaced with the polishing prompt information.
14. The method according to claim 13, wherein the prompt information to be polished is the prompt information in the selected state in the target prompt information; or The prompt information to be polished is the prompt information that is not in a selected state in the target prompt information.
15. The method according to any one of claims 1 to 12, further comprising: In response to an area selection operation performed on the original media content and / or the target media content, the prompt information group associated with the selected area in the at least one prompt information group is displayed in a first state, and the prompt information groups other than the prompt information group associated with the selected area in the at least one prompt information group are displayed in a second state.
16. A device for generating media content, comprising: A prompt information display module, configured to display at least one prompt information group, wherein the prompt information group includes at least one candidate prompt information, and the candidate prompt information is used to describe content characteristics of target media content; a prompt information determining module, configured to determine first prompt information from the candidate prompt information in response to an information selection operation on at least one of the candidate prompt information in the at least one prompt information group; The media content display module is configured to display the target media content generated based on the target prompt information; wherein the target prompt information includes the first prompt information.
17. An electronic device comprising: at least one processor; as well as a memory communicatively connected to the at least one processor; wherein, The memory stores a computer program executable by the at least one processor. The computer program is executed by the at least one processor to enable the at least one processor to perform the method for generating media content according to any one of claims 1 to 15.
18. A computer-readable storage medium storing computer instructions, wherein the computer instructions are configured to enable a processor to implement the method for generating media content according to any one of claims 1 to 15 when executed.
19. A computer program product, comprising a computer program, wherein when the computer program is executed by a processor, the computer program implements the method for generating media content according to any one of claims 1 to 15.
Citation Information
Patent Citations
Event prompt processing method and device, equipment and medium
CN115220613A
Image generation method and device, electronic equipment and storage medium
CN116524052A
Image generation method and device, terminal and storage medium
CN117392254A
Image generation method and device and storage medium
CN117475031A
Media content generation method and device, electronic equipment, storage medium and program product
CN118034562A