Image generation method and device, electronic equipment and readable storage medium

Through an image generation method and device, artificial intelligence technology is used to generate new images based on the images and text selected by the user, which solves the problem of insufficient number of images acquired by users, realizes intelligence and automation of image acquisition, and improves the efficiency of users to acquire images.

CN120070669APending Publication Date: 2025-05-30VIVO MOBILE COMM CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Application Number
CN202510151549.9
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2025-02-11
Publication Date
2025-05-30

AI Technical Summary

Technical Problem

When users share images through social media, they often face the problem of insufficient number of images, which makes it difficult to express stories or emotions effectively. The existing image editing and stitching tools lack intelligence and automation, making it difficult to meet users' needs to quickly obtain new images.

Method used

An image generation method and device are provided, by receiving user input, determining selected images and text, and using artificial intelligence technology to generate new images to ensure that the image content corresponds to text, thereby realizing the intelligence and automation of image acquisition.

Benefits of technology

Without the need for user to manually edit and process images, new images corresponding to the text can be automatically generated based on pre-provided images and text, which improves the efficiency of users to acquire images and meets users' needs for quick and convenient acquisition of new images.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN120070669A_ABST
    Figure CN120070669A_ABST
Patent Text Reader

Abstract

The invention discloses an image generation method and device, electronic equipment and a readable storage medium, and belongs to the technical field of artificial intelligence. The image generation method comprises the steps of receiving first input of a user; in response to the first input, determining at least one first image selected by the user; receiving a second input of the user; in response to the second input, determining a first text; receiving a third input of the user; and in response to the third input, displaying at least one first image and at least one second image, the at least one second image being generated from the at least one first image and the first text, the image content of the at least one first image and the image content of the at least one second image corresponding to the first text.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] This application belongs to the technical field of artificial intelligence, and particularly relates to an image generation method and apparatus, an electronic device, and a readable storage medium. Background Art

[0002] With the rapid development of mobile Internet technology, social media has become an indispensable part of people's daily lives. In particular, the sharing function of various social media has become an important channel for people to record their lives and share their emotions.

[0003] When users share images through the sharing functions of various social media, they often face the problem of insufficient image quantity, which will limit users' ability to express stories or emotions through images and requires users to obtain more images to supplement a more complete story line.

[0004] However, when users obtain new images through existing picture editing and splicing tools, they need to rely on manual operations by users. The image acquisition lacks intelligence and automation, making it difficult to meet users' needs for quickly and conveniently obtaining new images and reducing the efficiency of users' image acquisition. Summary of the Invention

[0005] The purpose of the embodiments of this application is to provide an image generation method and apparatus, an electronic device, and a readable storage medium, which can improve the intelligence and automation of image acquisition, meet users' needs for quickly and conveniently obtaining new images, and improve the efficiency of users' image acquisition.

[0006] In a first aspect, the embodiments of this application provide an image generation method, which includes: receiving a first input from a user; in response to the first input, determining at least one first image selected by the user; receiving a second input from the user; in response to the second input, determining a first text; receiving a third input from the user; in response to the third input, displaying at least one first image and at least one second image, where at least one second image is generated based on at least one first image and the first text, and the image content of at least one first image and the image content of at least one second image correspond to the first text.

[0007] In a second aspect, an embodiment of the present application provides an image generation device, which includes: a receiving unit for receiving a first input from a user; a processing unit for determining at least one first image selected by the user in response to the first input; the receiving unit is further configured to receive a second input from the user; the processing unit is further configured to determine a first text in response to the second input; the receiving unit is further configured to receive a third input from the user; a display unit for displaying at least one first image and at least one second image in response to the third input, where at least one second image is generated based on at least one first image and the first text, and the image content of at least one first image and the image content of at least one second image correspond to the first text.

[0008] In a third aspect, an embodiment of the present application provides an electronic device, which includes a processor and a memory. The memory stores a program or instruction that can run on the processor. When the program or instruction is executed by the processor, the steps of the image generation method in the first aspect are implemented.

[0009] In a fourth aspect, an embodiment of the present application provides a readable storage medium, on which a program or instruction is stored. When the program or instruction is executed by a processor, the steps of the image generation method in the first aspect are implemented.

[0010] In a fifth aspect, an embodiment of the present application provides a chip, which includes a processor and a communication interface. The communication interface is coupled to the processor, and the processor is configured to run a program or instruction to implement the steps of the image generation method in the first aspect.

[0011] In a sixth aspect, an embodiment of the present application provides a computer program product, which is stored in a storage medium. The program product is executed by at least one processor to implement the steps of the image generation method in the first aspect.

[0012] In the image generation method provided by the embodiment of the present application, a first input from a user is received; in response to the first input, at least one first image selected by the user is determined; a second input from the user is received; in response to the second input, a first text is determined; a third input from the user is received; in response to the third input, at least one first image and at least one second image are displayed, where at least one second image is generated based on at least one first image and the first text, and the image content of at least one first image and the image content of at least one second image correspond to the first text. Through the above image generation method, based on the known first image and the first text, a new second image corresponding to the first text is generated. In this way, there is no need for the user to manually edit and process the image. According to the image and text provided by the user in advance, a new image corresponding to the text can be automatically generated, realizing the intelligence and automation of image acquisition, meeting the user's need to quickly and conveniently obtain new images, and improving the efficiency of the user to obtain images. Description of the Drawings

[0013] Figure 1 One of the schematic flowcharts of the image generation method provided by the embodiment of the present application;

[0014] Figure 2 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0015] Figure 3 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0016] Figure 4 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0017] Figure 5 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0018] Figure 6 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0019] Figure 7 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0020] Figure 8 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0021] Figure 9 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0022] Figure 10 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0023] Figure 11 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0024] Figure 12 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0025] Figure 13 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0026] Figure 14 One of the schematic diagrams of the operation interface of the image generation method provided by the embodiment of the present application;

[0027] Figure 15The second schematic flowchart of the image generation method provided by the embodiment of the present application;

[0028] Figure 16 The structural block diagram of the image generation device provided by the embodiment of the present application;

[0029] Figure 17 The structural block diagram of the electronic device provided by the embodiment of the present application;

[0030] Figure 18 The schematic hardware structure diagram of the electronic device provided by the embodiment of the present application. Specific embodiments

[0031] Next, the technical solutions in the embodiments of the present application will be clearly described in conjunction with the accompanying drawings in the embodiments of the present application. Obviously, the described embodiments are some, but not all, of the embodiments of the present application. Based on the embodiments in the present application, all other embodiments obtained by those of ordinary skill in the art belong to the scope of protection of the present application.

[0032] The terms "first", "second", etc. in the specification and claims of the present application are used to distinguish similar objects, rather than to describe a specific order or sequence. It should be understood that such terms can be interchanged under appropriate circumstances so that the embodiments of the present application can be implemented in an order other than those illustrated or described herein, and the objects distinguished by "first", "second", etc. generally belong to the same category, and the number of objects is not limited. For example, the first object can be one or multiple. In addition, "and / or" in the specification and claims means at least one of the connected objects, and the character " / " generally represents an "or" relationship between the associated objects before and after.

[0033] Next, the image generation method provided by the embodiment of the present application will be described in detail in conjunction with the accompanying drawings, through specific embodiments and their application scenarios.

[0034] As Figure 1 shown, the embodiment of the present application provides an image generation method, which may include the following S102 to S112:

[0035] S102: Receive the first input of the user.

[0036] The image generation method proposed by the embodiment of the present application is executed by an electronic device, which may specifically be a smart electronic device such as a smart phone, a tablet computer, a laptop computer, and a smart watch, and no specific limitation is made here.

[0037] Among them, the above first input is a touch input of the user to the electronic device, and the touch input can specifically be a single - click input, a long - press input, a double - click input, a slide input, an input along a preset trajectory, etc. Those skilled in the art can set the specific form of the above first input according to the actual situation, and no specific limitation is made here.

[0038] S104: In response to the first input, determine at least one first image selected by the user.

[0039] Among them, the above first input is used for the user to select the first image.

[0040] Furthermore, the above first image can be an image locally stored in the electronic device, and the above first image can also be an image that the electronic device has uploaded to the cloud, and no specific limitation is made here.

[0041] Specifically, in the image generation method provided in the embodiments of the present application, before generating a new image, the electronic device can receive and respond to the user's first input, select at least one first image for generating the new image, so as to provide image materials for subsequent generation of the new image.

[0042] Exemplarily, as Figure 2 shown, the user opens the image application interface 202 of the electronic device. The image application interface 202 includes multiple images, and the user can browse and select the required images within the image application interface 202. Further, as Figure 3 shown, the electronic device receives a single - click input or a drag - and - drop input of the user to the multiple first images 204 in the image application interface 202. The electronic device responds to this input and selects the first image 204 selected by the user. Further, the electronic device receives a single - click input of the user to the first control 206 in the image application interface 202. The electronic device responds to this single - click input, as Figure 4 shown, enters the image editing interface 208, and in the image editing interface 208, in the order in which the user selects the first images 204, sequentially displays the multiple first images 204 selected by the user.

[0043] S106: Receive the second input of the user.

[0044] Among them, the above second input is a touch input of the user to the electronic device, and the touch input can specifically be a single - click input, a long - press input, a double - click input, a slide input, an input along a preset trajectory, etc. Those skilled in the art can set the specific form of the above second input according to the actual situation, and no specific limitation is made here.

[0045] S108: In response to the second input, determine the first text.

[0046] Among them, the first text can specifically be the story line framework that the user wants to express through the image, and this story line framework defines the theme, emotional trend, and key nodes of the story.

[0047] For example, the above-mentioned first text can be "a happy journey from departure to homecoming", "the growth and changes of a cat".

[0048] In the actual application process, the above-mentioned first text can be manually input by the user, and the above-mentioned first text can also be selected by the user from at least one stored text, and no specific limitation is made here.

[0049] Specifically, in the image generation method provided in the embodiments of the present application, before generating a new image, the electronic device can also receive and respond to the second input of the user, determine the first text for generating the new image, so as to provide text materials for subsequent generation of the new image.

[0050] Exemplarily, as Figure 4 shown, the image editing interface 208 includes a plurality of text labels, each text label corresponds to a first text, that is, each text label corresponds to a story line framework, and each story line framework contains a brief description of the theme, emotional trend, and key nodes. Further, as Figure 5 shown, the electronic device receives a click input of the user on the first text label 210, and the electronic device responds to this click input, as Figure 6 shown, selects the first text label 210, and determines the first text corresponding to the first text label 210. Or, as Figure 4 shown, the image editing interface 208 also includes a second control 212, and the electronic device can receive and respond to the input of the user on the second control 212 to determine the first text manually input by the user.

[0051] S110: Receive the third input of the user.

[0052] Among them, the above-mentioned third input is a touch input of the user to the electronic device, and this touch input can specifically be a click input, a long press input, a double click input, a slide input, an input along a preset trajectory, etc. Those skilled in the art can set the specific form of the above-mentioned third input according to the actual situation, and no specific limitation is made here.

[0053] S112: In response to the third input, display at least one first image and at least one second image.

[0054] Among them, the above-mentioned third input is a confirmation input. When the electronic device receives the third input, it means that the user has completed the selection or input of the first image and the first text, and starts to generate a new image.

[0055] Further, at least one second image is generated based on at least one first image and the first text.

[0056] Further, the image content of at least one first image and the image content of at least one second image correspond to the first text, and the storyline corresponding to the first text can be clearly expressed through at least one first image and at least one second image.

[0057] Specifically, in the image generation method provided in the embodiments of the present application, after the user has completed the selection or input of the first image and the first text, the user can perform a third input on the electronic device. In response to this third input, based on AI (Artificial Intelligence) technology, the electronic device generates at least one second image whose image content corresponds to the first text according to at least one first image and the first text, and simultaneously displays the first image and the generated second image, so that the user can view each image used to describe the storyline for preview and editing. Among them, the storyline corresponding to the first text can be clearly expressed through at least one first image and at least one second image.

[0058] Among them, by displaying the first image and the generated second image, the user can preview all the first images and second images, that is, the user can preview all the image content used to express the storyline, so as to ensure the coherence and satisfaction of the storyline expression.

[0059] Further, during the process of generating the second image, the electronic device can automatically match or suggest corresponding image description information according to the image content of each second image. After displaying the second image, the user can view the image description information of each second image through a touch input on the second image, such as a long press input, so as to evaluate the satisfaction of the image description information of each second image.

[0060] Among them, in the case where the user is not satisfied with the preview content, such as the generated image or the image description information, the user can return to the previous step and re-select or input the image material and text material for generating the second image. Or, the user can use an editing tool to fine-tune the first image and the generated second image.

[0061] In the actual application process, the user can also set the total number of images required to express the storyline, that is, set the total number of the first image and the second image, so as to set the number of second images to be generated.

[0062] For example, if the user sets the total number of images required to express the storyline to be 9, and the number of the first images selected by the user is 4, the electronic device generates 5 more second images, so as to clearly express the storyline through these 9 images.

[0063] In addition, after the second image is generated, the user can evaluate the satisfaction of the generated second image. If the user is not satisfied with the generated second image, the user can, through touch input on the electronic device, re-select or input the image material and text material used to generate the second image, and continue to generate a new second image until a second image that meets the user's needs is obtained.

[0064] Exemplarily, the user sets the total number of images required to express the story line to 9. As Figure 6 shown, the image editing interface 208 further includes a third control 214. After the user selects 4 first images 204 and selects the first text corresponding to the first text label 210, the user can perform a click input on the third control 214. In response to this click input, the electronic device, as Figure 7 shown, based on the 4 first images 204 and the first text corresponding to the first text label 210, the AI generates 5 second images 216 and automatically fills and displays these 5 second images 216 after the 4 first images 204. Among them, the story line corresponding to the first text can be clearly expressed through these 4 first images 204 and 5 second images 216. On this basis, the user can evaluate the satisfaction of the generated second images 216. If the user is not satisfied with the generated second images 216, as Figure 8 shown, the user can, through touch input on the electronic device, select the first text corresponding to the second text label 218. Further, the user can perform a click input on the third control 214 again. In response to this click input, the electronic device, as Figure 9 shown, regenerates the second images 216 based on the 4 first images 204 and the first text corresponding to the second text label 218 until the generated second images 216 meet the user's requirements. At this time, as Figure 9 shown, the image editing interface 208 further includes a fourth control 220. The user can, through touch input on the fourth control 220, save the first images 204 and second images 216 that the user is satisfied with to local or cloud storage.

[0065] The above image generation method provided by the embodiments of the present application receives a first input from a user; in response to the first input, determines at least one first image selected by the user; receives a second input from the user; in response to the second input, determines a first text; receives a third input from the user; in response to the third input, displays at least one first image and at least one second image, where at least one second image is generated based on at least one first image and the first text, and the image content of at least one first image and the image content of at least one second image correspond to the first text. Through the above image generation method, based on the known first image and first text, a new second image corresponding to the first text is generated. In this way, without the user manually editing and processing the image, a new image corresponding to the text can be automatically generated according to the image and text provided by the user in advance, realizing the intelligence and automation of image acquisition, meeting the user's need to quickly and conveniently obtain a new image, and improving the efficiency of the user to obtain an image.

[0066] In the embodiments of the present application, the above image generation method may specifically further include the following S126 to S132:

[0067] S126: Receive a fourth input from the user.

[0068] Among them, the above fourth input is a touch input of the user to the electronic device, and the touch input may specifically be a click input, a long press input, a double click input, a slide input, an input along a preset trajectory, etc. Those skilled in the art can set the specific form of the above fourth input according to the actual situation, and no specific limitation is made here.

[0069] Further, the fourth input is used to perform an overall preview of the first image and the second image.

[0070] S128: In response to the fourth input, display a second text.

[0071] Among them, the number of characters of the second text is more than that of the first text. Specifically, the second text is obtained by refining and supplementing the first text. Compared with the first text, the second text is richer in content and describes and expresses the story more completely and clearly.

[0072] Specifically, in the image generation method provided in the embodiments of the present application, when the user is satisfied with the first image and the generated second image, the first image and the second image can also be shared. Among them, before sharing the first image and the second image, the user can perform an overall preview of the first image and the second image through a fourth input to the electronic device. At this time, while displaying at least one first image and at least one second image, the electronic device will also supplement and expand the first text to generate and display the second text. In this way, it is possible to perform text matching on the first image and the second image to display the content to be shared in a vivid and illustrated manner for the user.

[0073] S130: Receive a fifth input from the user.

[0074] Among them, the above-mentioned fifth input is a touch input of the user to the electronic device, and this touch input can specifically be a click input, a long-press input, a double-click input, a slide input, an input along a preset trajectory, etc. Those skilled in the art can set the specific form of the above-mentioned fifth input according to the actual situation, and no specific limitation is made here.

[0075] Furthermore, the fifth input is used to share the first image and the second image so as to share the first image and the second image to a social media or other platforms for others to view.

[0076] S132: In response to the fifth input, share the first image and at least one of the second images, and copy the second text.

[0077] Specifically, in the image generation method provided in the embodiments of the present application, after previewing and displaying the first image, the second image, and the second text, the user can share the first image and the second image through a fifth input to the electronic device. At this time, the electronic device will respond to this fifth input, quickly share at least one first image and at least one second image to a social media or other platforms, and copy the second text to the clipboard for subsequent pasting use.

[0078] In the above embodiments provided by the present application, receive the fourth input of the user; in response to the fourth input, display the second text, and the number of characters of the second text is more than that of the first text; receive the fifth input of the user; in response to the fifth input, share the first image and at least one second image, and copy the second text. In this way, the preview and quick sharing of the image are realized, the steps of the user for sharing operations are reduced, and the sharing efficiency is improved.

[0079] In the embodiments of the present application, the above-mentioned fifth input includes a first sub-input and a second sub-input, and the above-mentioned S132 may specifically include S132a and S132b described below:

[0080] S132a: Display at least one sharing control in response to a first sub - input.

[0081] Wherein, the first sub - input is used to share a first image and a second image so as to share the first image and the second image to a social media or other platforms for others to view.

[0082] Furthermore, one sharing control corresponds to one application.

[0083] Furthermore, at least one sharing control includes a first sharing control, and the first sharing control corresponds to a first application.

[0084] Specifically, in the image generation method provided in the embodiments of the present application, the user can share the first image and the second image through a first sub - input to an electronic device. At this time, the electronic device will display at least one sharing control in response to the first sub - input.

[0085] S132b: In response to a second sub - input of the user to the first sharing control, share at least one first image and at least one second image to the first application in the display order of at least one first image and at least one second image, and copy the second text.

[0086] Specifically, in the image generation method provided in the embodiments of the present application, after displaying at least one sharing control, the electronic device can receive a second sub - input of the user to the first sharing control among at least one sharing control. The first sharing control corresponds to the first application. The electronic device responds to the second sub - input and automatically shares at least one first image and at least one second image to the first application in the current display order of the first image and the second image, and copies the second text to the clipboard for subsequent pasting use.

[0087] In the above - mentioned embodiments provided by the present application, the fifth input includes a first sub - input and a second sub - input. In response to the first sub - input, at least one sharing control is displayed. At least one sharing control includes a first sharing control, and the first sharing control corresponds to a first application. In response to a second sub - input of the user to the first sharing control, at least one first image and at least one second image are shared to the first application in the display order of at least one first image and at least one second image, and the second text is copied. In this way, through the input to the sharing control, the images that meet the requirements are directly shared to the application corresponding to the sharing control, without the user entering the specific application for operation, reducing the steps of the user's sharing operation and improving the sharing efficiency.

[0088] In the embodiments of the present application, the above - mentioned image generation method further includes S134 below, the fifth input includes a third sub - input, and S132 can specifically include S132c below:

[0089] S134: Determine a second application in response to a second input.

[0090] Specifically, in the image generation method provided in the embodiments of the present application, the user can also, through a second input to the electronic device, pre-set the sharing target, i.e., the second application, of the first image and the subsequently generated second image.

[0091] S132c: In response to a third sub-input, share at least one first image and at least one second image to the second application in the display order of the at least one first image and the at least one second image, and copy the second text.

[0092] Specifically, in the image generation method provided in the embodiments of the present application, when the user pre-sets the sharing target, i.e., the second application, of the first image and the subsequently generated second image, when the user shares the first image and the second image, the electronic device can receive and respond to the third sub-input of the user, and directly share at least one first image and at least one second image to the second application in the current display order of the first image and the second image, and copy the second text to the clipboard for subsequent pasting use.

[0093] In the above embodiments provided by the present application, a second application is determined in response to a second input; in response to a third sub-input, at least one first image and at least one second image are shared to the second application in the display order of the at least one first image and the at least one second image, and the second text is copied. In this way, one-key sharing of images is achieved, without the user having to select the corresponding sharing control for sharing operations, improving the sharing efficiency.

[0094] In the embodiments of the present application, the step of displaying at least one first image and at least one second image may specifically include S114 below, and the step of displaying the second text may specifically include S116 below:

[0095] S114: Display at least one first image and at least one second image in a first order.

[0096] Wherein, the first order corresponds to the first text. That is, the display order of the first image and the second image is consistent with the occurrence order of the key nodes in the story line framework corresponding to the first text. By viewing the first image and the second image in the first order in sequence, the story process described in the first text can be understood.

[0097] Specifically, in the image generation method provided in the embodiments of the present application, after generating at least one second image, the electronic device may, according to the image contents of the first image and the second image, and based on the occurrence order of the key nodes in the story line framework corresponding to the first text, perform intelligent sorting on the first image and the second image, and display at least one first image and at least one second image in the first order, so as to perform intelligent layout on the first image and the second image, improving the coherence and aesthetics of the display of the first image and the second image and making it easy to read.

[0098] S116: Display the second text according to the first text, the first order, the image contents of at least one first image, and the image contents of at least one second image.

[0099] Wherein, the second text is generated according to the first text, the image contents of at least one first image, and the image contents of at least one second image. The second text is obtained by refining and supplementing the first text. Compared with the first text, the second text is richer in content and more complete and clear in describing and expressing the story.

[0100] Furthermore, the second text corresponds to the first order. That is, the occurrence order of the key nodes in the story line framework corresponding to the second text is the same as the display order of the first image and the second image. By viewing the first image and the second image in the first order in sequence, the story process described by the second text can be understood.

[0101] Specifically, in the image generation method provided in the embodiments of the present application, when the user is satisfied with the first image and the generated second image, the first image and the second image can also be shared. Among them, before sharing the first image and the second image, the user can perform an overall preview of the first image and the second image through a fourth input to the electronic device. At this time, while displaying at least one first image and at least one second image in the first order, the electronic device will also supplement and expand the first text according to the first text, the image contents of at least one first image, and the image contents of at least one second image, and generate and display the second text. In this way, text matching can be performed on the first image and the second image to display the content to be shared in a vivid and illustrated manner for the user.

[0102] Exemplarily, as Figure 10 shown, the image editing interface 208 further includes a fifth control 222. When the user is satisfied with the first image 204 and the generated second image 216, the user can click on the fifth control 222, and the electronic device responds to this input, as Figure 11As shown, enter the image preview interface 224, display at least one first image 204 and at least one second image 216 in the image preview interface 224 in the first order, and supplement and expand the first text according to the first text, the image content of at least one first image 204 and at least one second image 216, generate and display a second text 226 in the image preview interface 224.

[0103] In the above embodiments provided by the present application, at least one first image and at least one second image are displayed in the first order, and the first order corresponds to the first text; according to the first text, the first order, the image content of at least one first image and the image content of at least one second image, the second text is displayed; wherein, the second text corresponds to the first order. In this way, intelligent layout and text matching are performed on the first image and the second image, improving the coherence and aesthetics of the display of the first image and the second image, being easy to read, and facilitating the display of the content to be shared with the user in a vivid and illustrated manner.

[0104] In the embodiment of the present application, after the above S128, the above image generation method may specifically further include the following S118 and S120:

[0105] S118: Receive a sixth input from the user.

[0106] Wherein, the above sixth input is a touch input of the user to the electronic device, and the touch input may specifically be a click input, a long press input, a double click input, a slide input, an input along a preset trajectory, etc. Those skilled in the art can set the specific form of the above sixth input according to the actual situation, and no specific limitation is made here.

[0107] Further, the sixth input is used to adjust the display order of at least one first image and at least one second image from the above first order to a second order.

[0108] S120: In response to the sixth input, adjust the second text to a third text.

[0109] Wherein, the third text corresponds to the second order. That is, the occurrence order of the key nodes in the story line framework corresponding to the third text is consistent with the adjusted display order of the first image and the second image. By viewing the first image and the second image in sequence according to the adjusted second order, the story process described in the third text can be understood.

[0110] Specifically, in the image generation method provided in the embodiments of the present application, after previewing and displaying the first image, the second image, and the second text, the user can, through a sixth input to the electronic device, adjust the display order of at least one first image and at least one second image from the above-mentioned first order to the second order. At this time, the electronic device will synchronously adjust the content of the second text based on the adjusted display order of the first image and the second image, so as to adjust the second text to a third text whose key node occurrence order conforms to the second order.

[0111] Exemplarily, as Figure 12 shown, after displaying the first image 204 and the second image 216 in the first order and displaying the second text 226, the user can adjust the display order of the first image 204 and the second image 216 through a drag input on the first image 204 or the second image 216. At this time, as Figure 13 shown, while the electronic device displays the first image 204 and the second image 216 in the adjusted second order, it will also synchronously update and display the corresponding third text 228 according to the occurrence order of the image content of the first image 204 and the second image 216 after adjusting the display order.

[0112] In the above-mentioned embodiments provided by the present application, after displaying the second text, a sixth input from the user is received, and the sixth input is used to adjust the display order of at least one first image and at least one second image to the second order; in response to the sixth input, the second text is adjusted to a third text, and the third text corresponds to the second order. In this way, it is convenient for the user to automatically synchronously adjust the text corresponding to the image by adjusting the image display order, without the user having to adjust both, reducing the user's operation steps.

[0113] In the embodiments of the present application, after the above-mentioned S128, the image generation method may specifically further include the following S122 and S124:

[0114] S122: Receive a seventh input from the user.

[0115] Among them, the above-mentioned seventh input is a touch input by the user on the second text, and the touch input may specifically be a click input, a long-press input, a double-click input, a slide input, an input along a preset trajectory, etc. Those skilled in the art can set the specific form of the above-mentioned seventh input according to the actual situation, and no specific limitation is made here.

[0116] Further, the seventh input is used to adjust the second text to a fourth text.

[0117] S124: In response to the seventh input, display at least one first image and at least one second image in the third order.

[0118] Among them, the third order corresponds to the fourth text. That is, the display order of the adjusted first image and second image is consistent with the occurrence order of the key nodes in the story line framework corresponding to the fourth text. By viewing the first image and the second image in sequence according to the third order, the story process described in the fourth text can be understood.

[0119] Specifically, in the image generation method provided in the embodiments of the present application, after previewing and displaying the first image, the second image, and the second text, the user can adjust the second text to the fourth text by means of a seventh input to the second text. At this time, the electronic device will synchronously adjust the display order of the first image and the second image based on the occurrence order of the key nodes in the story line framework corresponding to the fourth text, so as to adjust the display order of the first image and the second image to a third order that conforms to the fourth text.

[0120] Exemplarily, as Figure 12 shown, after displaying the first image 204 and the second image 216 in the first order and displaying the second text 226, the electronic device can receive a sixth input from the user to the second text 226. In response to this sixth input, as Figure 14 shown, it adjusts the display to the fourth text 230, and synchronously adjusts the display order of the first image 204 and the second image 216 to the third order based on the occurrence order of the key nodes in the story line framework corresponding to the fourth text 230.

[0121] In the above embodiments provided by the present application, after displaying the second text, a seventh input from the user is received, and the seventh input is used to adjust the second text to the fourth text; in response to the seventh input, at least one first image and at least one second image are displayed in the third order, and the third order corresponds to the fourth text. In this way, it is convenient for the user to automatically synchronously adjust the image display order by adjusting the text, without the user having to adjust both, reducing the user's operation steps.

[0122] In the embodiments of the present application, the above S108 may specifically include the following S108a or S108b:

[0123] S108a: In response to the user's selection input to the first text label, determine the first text corresponding to the first text label.

[0124] Among them, in the image generation method provided in the embodiments of the present application, multiple text labels may be pre-stored in the electronic device, each text label corresponding to a first text, that is, each text label corresponds to a story line framework, and each story line framework includes a theme, an emotional trend, and a brief description of key nodes.

[0125] Further, the above second input may be the user's selection input to any text label among the multiple text labels.

[0126] Specifically, in the image generation method provided in the embodiments of the present application, before generating a new image, the electronic device may receive and respond to a user's selection input of a first text label among multiple text labels, and determine the first text corresponding to the first text label as the first text for generating the new image, so as to provide text material for subsequent generation of the new image.

[0127] S108b: In response to the user's text input, determine the first text input by the user.

[0128] Among them, the above second input may also be the user's text input to the electronic device, and this text input is used to input the above first text.

[0129] Specifically, in the image generation method provided in the embodiments of the present application, before generating a new image, the electronic device may receive and respond to the user's text input, and determine the first text input by the user as the first text for generating the new image, so as to provide text material for subsequent generation of the new image.

[0130] In the above embodiments provided by the present application, in response to the user's selection input of the first text label, determine the first text corresponding to the first text label; or in response to the user's text input, determine the first text input by the user. In this way, based on the user's selection or text input, determine the text material for subsequent generation of the new image, ensuring that the text material meets the user's requirements and can improve the user's satisfaction with the generated second image.

[0131] In the embodiments of the present application, the above S112 may specifically include the following S112a to S112c:

[0132] S112a: In response to a third input, analyze the image information of at least one first image and the text information corresponding to the first text.

[0133] Among them, the image information includes at least one of the following items or a combination thereof: image color, image theme, image object.

[0134] Furthermore, the text information includes at least one of the following items or a combination thereof: theme information, emotional trend information, key node information.

[0135] Specifically, in the image generation method provided in the embodiments of the present application, after the user has completed the selection or input of the first image and the first text, the user may perform a third input on the electronic device. The electronic device responds to this third input, uses an AI algorithm to analyze the image information of each first image, and analyzes the text information corresponding to the first text, so as to determine information such as the image color, image theme, and image object of each first image, and determine information such as the theme information, emotional trend information, and key node information of the story line described by the first text.

[0136] Among them, the above AI algorithms include but are not limited to: deep learning algorithms, image recognition algorithms, etc., and no specific limitations are made here.

[0137] S112b: Generate at least one second image according to the image information and text information.

[0138] Specifically, in the image generation method provided in the embodiments of the present application, after analyzing the image information and text information, based on the AI image generation technology, at least one second image is generated according to the image information and text information.

[0139] S112c: Display at least one first image and at least one second image.

[0140] Specifically, in the image generation method provided in the embodiments of the present application, after generating at least one second image, the first image and the generated second image are displayed so that the user can view each image used to describe the story line for preview and editing by the user.

[0141] Exemplarily, user A plans to share the experience of a family trip, but only has 3 photos taken at different scenic spots. Based on the above image generation method provided in the embodiments of the present application, user A selects the above 3 photos as the first images and inputs "A happy journey from departure to return home" as the first text. The electronic device automatically generates 6 second images according to the colors and scenic features of the 3 first images and the prompts of the first text. The image content of the 6 second images includes interesting things during the trip, interactions among family members, etc. At the same time, the electronic device matches a corresponding text description to each second image. After user A successfully shares the above 9 images and their corresponding story lines, a large number of likes and comments are obtained.

[0142] Exemplarily, user B wants to record the growth process of his own cat from infancy to adulthood, but only has 3 photos of the cat at different times. Based on the above image generation method provided in the embodiments of the present application, user B selects the above 3 photos as the first images and selects "The growth and changes of the cat" as the first text. The electronic device analyzes the changes in the body shape, hair, etc. of the cat in the 3 photos and combines the prompts of the first text to automatically generate 6 second images reflecting the key stages of the cat's growth, including the cat's first time learning to climb a tree, the cat's first time catching a mouse, etc. At the same time, the electronic device matches a corresponding text description to each second image. After user B successfully shares the above 9 images, friends leave messages one after another saying that they are touched by the cat's growth story.

[0143] In the above embodiments provided by the present application, in response to a third input, the image information of at least one first image and the text information corresponding to the first text are analyzed; at least one second image is generated according to the image information and the text information; at least one first image and at least one second image are displayed; wherein, the image information includes at least one of the following items or a combination thereof: image color, image theme, image object, and the text information includes at least one of the following items or a combination thereof: theme information, emotional trend information, key node information. In this way, it is ensured that the generated new image is consistent and coherent with the initial image provided by the user in terms of visual style, emotional expression, and storyline.

[0144] In summary, as Figure 15 shown, when the user sets the total number of the first image and the second image to 9, the image generation method provided by the embodiments of the present application may specifically include the following S302 to S318:

[0145] S302: The user selects at least one first image.

[0146] S304: The user inputs the first text.

[0147] S306: Analyze the first image and the first text.

[0148] S308: Generate a second image according to the analysis result.

[0149] S310: Judge: The total number of the first image and the second image is 9. If so, execute S312. If not, execute S308.

[0150] S312: Generate a second text.

[0151] S314: Preview and display the first image, the second image, and the second text.

[0152] S316: Judge: The user confirms that they are satisfied with the preview content. If so, execute S318. If not, execute S308.

[0153] S318: Share the first image and the second image.

[0154] Specifically, in the image generation method provided in the embodiments of the present application, when the user sets the total number of the first image and the second image to 9, the user can select at least one first image from the local storage or the cloud of the electronic device, and determine the first text by selection or input, so as to describe the content that the user wants to express through the image. Further, the electronic device analyzes the first image and the first text based on the AI algorithm, and according to the analysis result, the AI generates a second image whose image content is related to the first text. Among them, in the process of generating the second image, the electronic device will judge whether the total number of the first image and the second image reaches 9. If not, it will continue to generate a new second image. If so, it will enter the next process. Further, after the total number of the first image and the second image reaches 9, the electronic device will display the first image and the second image, and display the generated second text according to the first image, the second image and the first text for the user to preview. The user can view the second image generated by the AI to confirm whether it meets the expectation, so as to evaluate the satisfaction of the second image, and at the same time evaluate the satisfaction of the second text. If the user is not satisfied with the preview content, the user can edit or adjust the preview content until the user is satisfied with the preview content and confirms that the editing of the preview content is completed. When the electronic device receives the confirmation input that the user is satisfied with the preview content, it supports sharing the preview content so that the user can share the first image and the second image.

[0155] That is to say, the image generation method provided in the embodiments of the present application provides a solution based on artificial intelligence technology to automatically generate new images related to and coherent with a small number of images and a preset story line provided by the user, so that the user can share a complete picture-and-text combined story on the social media platform. In this way, on the one hand, the user can quickly generate a creative and attractive picture-and-text story without spending a lot of time looking for or editing images, greatly improving the sharing efficiency and fun; on the other hand, the complete and coherent story line can attract more attention and resonance, promote social interaction, and facilitate enhancing the emotional connection between users. Further, the image generation method combining images with a story line can greatly enhance the fun of taking pictures, and at the same time can greatly improve the beauty of the images shared by users. Moreover, since the sharing process of the user is greatly simplified, the willingness of the user to share can be improved, and the user can be promoted to perform the operations of taking pictures and sharing.

[0156] In the actual application process, the application scenarios of the above image generation method are not limited to image sharing, and can also be applied to picture-and-text integration. For example, the above image generation method can be applied to note editing so that the user can read a complete story every time.

[0157] The image generation method provided by the embodiments of this application may be executed by an image generation device. In the embodiments of this application, taking the image generation device as the executor of the above image generation method as an example, the image generation device provided by the embodiments of this application is described.

[0158] As Figure 16 shown, the embodiments of this application provide an image generation device 500, which may include the following receiving unit 502, processing unit 504, and display unit 506.

[0159] The receiving unit 502 is configured to receive a first input from the user;

[0160] The processing unit 504 is configured to determine at least one first image selected by the user in response to the first input;

[0161] The receiving unit 502 is further configured to receive a second input from the user;

[0162] The processing unit 504 is further configured to determine a first text in response to the second input;

[0163] The receiving unit 502 is further configured to receive a third input from the user;

[0164] The display unit 506 is configured to display at least one first image and at least one second image in response to the third input, where at least one second image is generated based on at least one first image and the first text, and the image content of at least one first image and the image content of at least one second image correspond to the first text.

[0165] The image generation device 500 provided by the embodiments of this application receives a first input from the user; in response to the first input, determines at least one first image selected by the user; receives a second input from the user; in response to the second input, determines a first text; receives a third input from the user; and in response to the third input, displays at least one first image and at least one second image, where at least one second image is generated based on at least one first image and the first text, and the image content of at least one first image and the image content of at least one second image correspond to the first text. Through the above image generation device 500, new second images corresponding to the first text are generated based on the known first images and the first text. In this way, without the user manually operating and editing the image, new images corresponding to the text can be automatically generated according to the images and text provided by the user in advance, realizing the intelligence and automation of image acquisition, meeting the user's need to quickly and conveniently obtain new images, and improving the efficiency of the user to obtain images.

[0166] In an embodiment of the present application, the receiving unit 502 is further configured to: receive a fourth input from a user; the display unit 506 is further configured to: in response to the fourth input, display a second text, where the number of characters in the second text is greater than the number of characters in the first text; the receiving unit 502 is further configured to: receive a fifth input from the user; the processing unit 504 is further configured to: in response to the fifth input, share the first image and at least one second image, and copy the second text.

[0167] In the above embodiment provided by the present application, it receives a fourth input from the user; in response to the fourth input, it displays a second text, where the number of characters in the second text is greater than the number of characters in the first text; it receives a fifth input from the user; in response to the fifth input, it shares the first image and at least one second image, and copies the second text. In this way, preview and quick sharing of images are realized, the steps for the user to perform the sharing operation are reduced, and the sharing efficiency is improved.

[0168] In an embodiment of the present application, the fifth input includes a first sub-input and a second sub-input, and the display unit 506 is further configured to: in response to the first sub-input, display at least one sharing control, where the at least one sharing control includes a first sharing control, and the first sharing control corresponds to a first application; the processing unit 504 is specifically configured to: in response to a second sub-input from the user to the first sharing control, share the at least one first image and the at least one second image to the first application in the display order of the at least one first image and the at least one second image, and copy the second text.

[0169] In the above embodiment provided by the present application, the fifth input includes a first sub-input and a second sub-input. In response to the first sub-input, it displays at least one sharing control, where the at least one sharing control includes a first sharing control, and the first sharing control corresponds to a first application; in response to a second sub-input from the user to the first sharing control, it shares the at least one first image and the at least one second image to the first application in the display order of the at least one first image and the at least one second image, and copies the second text. In this way, by inputting to the sharing control, the images that meet the requirements are directly shared to the application corresponding to the sharing control, without the user entering the specific application to perform operations, reducing the steps for the user to perform the sharing operation and improving the sharing efficiency.

[0170] In an embodiment of the present application, the processing unit 504 is further configured to: in response to a second input, determine a second application; the fifth input includes a third sub-input, and the processing unit 504 is specifically configured to: in response to the third sub-input, share the at least one first image and the at least one second image to the second application in the display order of the at least one first image and the at least one second image, and copy the second text.

[0171] In the above embodiments provided by the present application, in response to a second input, a second application is determined; in response to a third sub-input, at least one first image and at least one second image are shared with the second application in the display order of the at least one first image and the at least one second image, and the second text is copied. In this way, one-key sharing of images is achieved, without the user having to select the corresponding sharing control for sharing operations, improving the sharing efficiency.

[0172] In an embodiment of the present application, the display unit 506 is specifically configured to: display at least one first image and at least one second image in a first order, the first order corresponding to the first text; display the second text according to the first text, the first order, the image content of the at least one first image, and the image content of the at least one second image; wherein, the second text corresponds to the first order.

[0173] In the above embodiments provided by the present application, at least one first image and at least one second image are displayed in a first order, the first order corresponding to the first text; the second text is displayed according to the first text, the first order, the image content of the at least one first image, and the image content of the at least one second image; wherein, the second text corresponds to the first order. In this way, intelligent layout and text matching are performed on the first image and the second image, enhancing the coherence and aesthetics of the display of the first image and the second image, being easy to read, and facilitating the display of the content to be shared with the user in a vivid and illustrated manner.

[0174] In an embodiment of the present application, after the second text is displayed, the receiving unit 502 is further configured to: receive a sixth input from the user, the sixth input being used to adjust the display order of the at least one first image and the at least one second image to a second order; the processing unit 504 is further configured to: in response to the sixth input, adjust the second text to a third text, the third text corresponding to the second order.

[0175] In the above embodiments provided by the present application, after the second text is displayed, a sixth input from the user is received, the sixth input being used to adjust the display order of the at least one first image and the at least one second image to a second order; in response to the sixth input, the second text is adjusted to a third text, the third text corresponding to the second order. In this way, it is convenient for the user to automatically synchronously adjust the text corresponding to the image by adjusting the image display order, without the user having to adjust both, reducing the user's operation steps.

[0176] In an embodiment of the present application, after the second text is displayed, the receiving unit 502 is further configured to: receive a seventh input from the user, the seventh input being used to adjust the second text to a fourth text; the display unit 506 is further configured to: in response to the seventh input, display at least one first image and at least one second image in a third order, the third order corresponding to the fourth text.

[0177] In the above embodiments provided by the present application, after the second text is displayed, the seventh input of the user is received, and the seventh input is used to adjust the second text to the fourth text; in response to the seventh input, at least one first image and at least one second image are displayed in the third order, and the third order corresponds to the fourth text. In this way, it is convenient for the user to automatically synchronously adjust the image display order by adjusting the text, without the user having to adjust both, reducing the user's operation steps.

[0178] In the embodiments of the present application, the processing unit 504 is specifically configured to: in response to the user's selection input for the first text label, determine the first text corresponding to the first text label; or in response to the user's text input, determine the first text input by the user.

[0179] In the above embodiments provided by the present application, in response to the user's selection input for the first text label, determine the first text corresponding to the first text label; or in response to the user's text input, determine the first text input by the user. In this way, based on the user's selection or text input, the text material for generating the new image subsequently is determined, ensuring that the text material meets the user's requirements and can improve the user's satisfaction with the generated second image.

[0180] In the embodiments of the present application, the processing unit 504 is further configured to: in response to the third input, analyze the image information of at least one first image and the text information corresponding to the first text; generate at least one second image according to the image information and the text information; the display unit 506 is specifically configured to: display at least one first image and at least one second image; wherein, the image information includes at least one of the following or a combination thereof: image color, image theme, image object, and the text information includes at least one of the following or a combination thereof: theme information, emotional tendency information, key node information.

[0181] In the above embodiments provided by the present application, in response to the third input, analyze the image information of at least one first image and the text information corresponding to the first text; generate at least one second image according to the image information and the text information; display at least one first image and at least one second image; wherein, the image information includes at least one of the following or a combination thereof: image color, image theme, image object, and the text information includes at least one of the following or a combination thereof: theme information, emotional tendency information, key node information. In this way, it is ensured that the generated new image is consistent and coherent with the initial image provided by the user in terms of visual style, emotional expression, and storyline.

[0182] The image generation device 500 in the embodiments of the present application may be an electronic device or a component in an electronic device, such as an integrated circuit or a chip. The electronic device may be a terminal or other devices other than terminals. Exemplarily, the electronic device may be a mobile phone, a tablet computer, a laptop computer, a handheld computer, a vehicle-mounted electronic device, a Mobile Internet Device (MID), an augmented reality (AR) / virtual reality (VR) device, a robot, a wearable device, an ultra-mobile personal computer (UMPC), a netbook, or a personal digital assistant (PDA), etc. It may also be a server, a Network Attached Storage (NAS), a personal computer (PC), a television (TV), a teller machine, or a self-service machine, etc. The embodiments of the present application do not make specific limitations.

[0183] The image generation device 500 in the embodiments of the present application may be a device with an operating system. The operating system may be an Android operating system, an iOS operating system, or other possible operating systems. The embodiments of the present application do not make specific limitations.

[0184] The image generation device 500 provided in the embodiments of the present application can implement Figure 1 and Figure 15 each process implemented by the method embodiments. To avoid repetition, it will not be elaborated here.

[0185] Optionally, as Figure 17 shown, the embodiments of the present application further provide an electronic device 600, including a processor 602 and a memory 604. A program or instruction that can run on the processor 602 is stored on the memory 604. When the program or instruction is executed by the processor 602, it implements each step of the above image generation method embodiment and can achieve the same technical effect. To avoid repetition, it will not be elaborated here.

[0186] It should be noted that the electronic devices in the embodiments of the present application include the above-mentioned mobile electronic devices and non-mobile electronic devices.

[0187] Figure 18 A schematic diagram of the hardware structure of an electronic device for implementing the embodiments of the present application.

[0188] The electronic device 700 includes, but is not limited to, components such as a radio frequency unit 701, a network module 702, an audio output unit 703, an input unit 704, a sensor 705, a display unit 706, a user input unit 707, an interface unit 708, a memory 709, and a processor 710.

[0189] Those skilled in the art can understand that the electronic device 700 may further include a power source (such as a battery) for powering each component. The power source can be logically connected to the processor 710 through a power management system, so as to implement functions such as management of charging, discharging, and power consumption management through the power management system. Figure 18 The structure of the electronic device shown does not limit the electronic device. The electronic device may include more or fewer components than shown, or combine certain components, or have different component arrangements, which will not be elaborated here.

[0190] Among them, the user input unit 707 is used to receive a first input from the user.

[0191] The processor 710 is used to determine at least one first image selected by the user in response to the first input.

[0192] The user input unit 707 is further used to receive a second input from the user.

[0193] The processor 710 is further used to determine a first text in response to the second input.

[0194] The user input unit 707 is further used to receive a third input from the user.

[0195] The display unit 706 is used to display at least one first image and at least one second image in response to the third input. The at least one second image is generated based on the at least one first image and the first text, and the image content of the at least one first image and the image content of the at least one second image correspond to the first text.

[0196] In an embodiment of the present application, a first input of a user is received; in response to the first input, at least one first image selected by the user is determined; a second input of the user is received; in response to the second input, a first text is determined; a third input of the user is received; in response to the third input, at least one first image and at least one second image are displayed, where the at least one second image is generated based on the at least one first image and the first text, and the image content of the at least one first image and the image content of the at least one second image correspond to the first text. In an embodiment of the present application, based on the known first image and the first text, a new second image corresponding to the first text is generated. In this way, without the user manually editing and processing the image, a new image corresponding to the text can be automatically generated according to the image and text provided by the user in advance, realizing the intelligence and automation of image acquisition, meeting the user's need to quickly and conveniently obtain a new image, and improving the efficiency of the user to obtain an image.

[0197] Optionally, the user input unit 707 is further configured to: receive a fourth input of the user; the display unit 706 is further configured to: in response to the fourth input, display a second text, where the number of characters of the second text is more than the number of characters of the first text; the user input unit 707 is further configured to: receive a fifth input of the user; the processor 710 is further configured to: in response to the fifth input, share the first image and the at least one second image, and copy the second text.

[0198] In the above embodiment provided by the present application, a fourth input of the user is received; in response to the fourth input, a second text is displayed, where the number of characters of the second text is more than the number of characters of the first text; a fifth input of the user is received; in response to the fifth input, the first image and the at least one second image are shared, and the second text is copied. In this way, the preview and quick sharing of the image are realized, the steps of the user for sharing operations are reduced, and the sharing efficiency is improved.

[0199] Optionally, the fifth input includes a first sub-input and a second sub-input, the display unit 706 is further configured to: in response to the first sub-input, display at least one sharing control, where the at least one sharing control includes a first sharing control, and the first sharing control corresponds to a first application; the processor 710 is specifically configured to: in response to a second sub-input of the user on the first sharing control, share the at least one first image and the at least one second image to the first application in the display order of the at least one first image and the at least one second image, and copy the second text.

[0200] In the above embodiments provided by the present application, the fifth input includes a first sub-input and a second sub-input. In response to the first sub-input, at least one sharing control is displayed. The at least one sharing control includes a first sharing control, and the first sharing control corresponds to a first application. In response to a second sub-input of the user on the first sharing control, at least one first image and at least one second image are shared to the first application in the display order of the at least one first image and the at least one second image, and the second text is copied. In this way, through the input to the sharing control, the images that meet the requirements are directly shared to the application corresponding to the sharing control, without the user entering the specific application for operation, reducing the steps of the user's sharing operation and improving the sharing efficiency.

[0201] Optionally, the processor 710 is further configured to: in response to a second input, determine a second application. The fifth input includes a third sub-input. Specifically, the processor 710 is configured to: in response to the third sub-input, share at least one first image and at least one second image to the second application in the display order of the at least one first image and the at least one second image, and copy the second text.

[0202] In the above embodiments provided by the present application, in response to a second input, a second application is determined; in response to a third sub-input, at least one first image and at least one second image are shared to the second application in the display order of the at least one first image and the at least one second image, and the second text is copied. In this way, one-key sharing of images is achieved, without the user selecting the corresponding sharing control for sharing operation, improving the sharing efficiency.

[0203] Optionally, the display unit 706 is specifically configured to: display at least one first image and at least one second image in a first order, and the first order corresponds to the first text; display the second text according to the first text, the first order, the image content of the at least one first image, and the image content of the at least one second image; wherein, the second text corresponds to the first order.

[0204] In the above embodiments provided by the present application, at least one first image and at least one second image are displayed in a first order, and the first order corresponds to the first text; the second text is displayed according to the first text, the first order, the image content of the at least one first image, and the image content of the at least one second image; wherein, the second text corresponds to the first order. In this way, intelligent layout and text matching are performed on the first image and the second image, improving the coherence and aesthetics of the display of the first image and the second image, being easy to read, and facilitating the display of the content to be shared to the user in a vivid and illustrated manner.

[0205] Optionally, after the second text is displayed, the user input unit 707 is further configured to: receive a sixth input from the user, where the sixth input is used to adjust the display order of at least one first image and at least one second image to a second order; the processor 710 is further configured to: in response to the sixth input, adjust the second text to a third text, where the third text corresponds to the second order.

[0206] In the above embodiments provided by the present application, after the second text is displayed, a sixth input from the user is received, where the sixth input is used to adjust the display order of at least one first image and at least one second image to a second order; in response to the sixth input, the second text is adjusted to a third text, where the third text corresponds to the second order. In this way, it is convenient for the user to automatically synchronously adjust the text corresponding to the image by adjusting the image display order, without the user having to adjust both, reducing the user's operation steps.

[0207] Optionally, after the second text is displayed, the user input unit 707 is further configured to: receive a seventh input from the user, where the seventh input is used to adjust the second text to a fourth text; the display unit 706 is further configured to: in response to the seventh input, display at least one first image and at least one second image in a third order, where the third order corresponds to the fourth text.

[0208] In the above embodiments provided by the present application, after the second text is displayed, a seventh input from the user is received, where the seventh input is used to adjust the second text to a fourth text; in response to the seventh input, at least one first image and at least one second image are displayed in a third order, where the third order corresponds to the fourth text. In this way, it is convenient for the user to automatically synchronously adjust the image display order by adjusting the text, without the user having to adjust both, reducing the user's operation steps.

[0209] Optionally, the processor 710 is specifically configured to: in response to a selection input of the user for the first text label, determine a first text corresponding to the first text label; or in response to a text input of the user, determine a first text input by the user.

[0210] In the above embodiments provided by the present application, in response to a selection input of the user for the first text label, a first text corresponding to the first text label is determined; or in response to a text input of the user, a first text input by the user is determined. In this way, based on the user's selection or text input, the text material for generating a new image is determined, ensuring that the text material meets the user's requirements and can improve the user's satisfaction with the generated second image.

[0211] Optionally, the processor 710 is further configured to: in response to a third input, analyze the image information of at least one first image and the text information corresponding to the first text; generate at least one second image according to the image information and the text information; the display unit 706 is specifically configured to: display at least one first image and at least one second image; wherein, the image information includes at least one of the following or a combination thereof: image color, image theme, image object, and the text information includes at least one of the following or a combination thereof: theme information, emotional tendency information, key node information.

[0212] In the above embodiments provided by the present application, in response to a third input, analyze the image information of at least one first image and the text information corresponding to the first text; generate at least one second image according to the image information and the text information; display at least one first image and at least one second image; wherein, the image information includes at least one of the following or a combination thereof: image color, image theme, image object, and the text information includes at least one of the following or a combination thereof: theme information, emotional tendency information, key node information. In this way, it is ensured that the generated new image is consistent and coherent with the initial image provided by the user in terms of visual style, emotional expression, and plot.

[0213] It should be understood that in the embodiments of the present application, the input unit 704 may include a Graphics Processing Unit (GPU) 7041 and a microphone 7042. The graphics processor 7041 processes the image data of static pictures or videos obtained by an image capture device (such as a camera) in a video capture mode or an image capture mode. The display unit 706 may include a display panel 7061, and the display panel 7061 may be configured in the form of a liquid crystal display, an organic light emitting diode, etc. The user input unit 707 includes at least one of a touch panel 7071 and other input devices 7072. The touch panel 7071 is also called a touch screen. The touch panel 7071 may include two parts: a touch detection device and a touch controller. The other input devices 7072 may include, but are not limited to, a physical keyboard, function keys (such as volume control keys, switch keys, etc.), a trackball, a mouse, a joystick, which will not be elaborated herein.

[0214] The memory 709 can be used to store software programs and various data. The memory 709 may mainly include a first storage area for storing programs or instructions and a second storage area for storing data. Among them, the first storage area may store an operating system, application programs or instructions required for at least one function (such as a sound playback function, an image playback function, etc.). In addition, the memory 709 may include volatile memory or non-volatile memory, or the memory 709 may include both volatile and non-volatile memory. Among them, the non-volatile memory may be a read-only memory (ROM), a programmable read-only memory (PROM), an erasable programmable read-only memory (EPROM), an electrically erasable programmable read-only memory (EEPROM), or a flash memory. The volatile memory may be a random access memory (RAM), a static random access memory (SRAM), a dynamic random access memory (DRAM), a synchronous dynamic random access memory (SDRAM), a double data rate synchronous dynamic random access memory (DDR SDRAM), an enhanced synchronous dynamic random access memory (ESDRAM), a synch link dynamic random access memory (SLDRAM), and a direct rambus random access memory (DRRAM). The memory 709 in the embodiments of the present application includes, but is not limited to, these and any other suitable types of memory.

[0215] The processor 710 may include one or more processing units; optionally, the processor 710 integrates an application processor and a modem processor. Among them, the application processor mainly processes operations related to the operating system, user interface, and application programs, etc., and the modem processor mainly processes wireless communication signals, such as a baseband processor. It can be understood that the above modem processor may not be integrated into the processor 710 either.

[0216] The embodiments of the present application also provide a readable storage medium. A program or instructions are stored on the readable storage medium. When the program or instructions are executed by a processor, each process of the above image generation method embodiments is implemented, and the same technical effects can be achieved. To avoid repetition, it will not be elaborated here.

[0217] Among them, the processor is the processor in the electronic device in the above embodiment. The readable storage medium includes computer-readable storage media such as computer read-only memory ROM, random access memory RAM, magnetic disk, or optical disc, etc.

[0218] Another embodiment of the present application provides a chip, which includes a processor and a communication interface. The communication interface is coupled to the processor. The processor is used to run programs or instructions to implement each process of the above embodiment of the image generation method, and can achieve the same technical effect. To avoid repetition, it will not be elaborated here.

[0219] It should be understood that the chip mentioned in the embodiments of the present application may also be referred to as a system-on-chip, system chip, chip system, or system-on-chip, etc.

[0220] The embodiments of the present application provide a computer program product, which is stored in a storage medium. The program product is executed by at least one processor to implement each process of the above embodiment of the image generation method, and can achieve the same technical effect. To avoid repetition, it will not be elaborated here.

[0221] It should be noted that in this article, the term "including", "comprising" or any other variant thereof is intended to cover non-exclusive inclusion, so that a process, method, article or device including a series of elements not only includes those elements, but also includes other elements not expressly listed, or further includes elements inherent to such process, method, article or device. Without further limitation, an element defined by the statement "including a..." does not exclude the existence of additional identical elements in the process, method, article or device including the element. In addition, it should be pointed out that the scope of the methods and devices in the embodiments of the present application is not limited to performing functions in the order shown or discussed, and may also include performing functions in a substantially simultaneous manner or in a reverse order according to the functions involved. For example, the described methods may be performed in an order different from that described, and various steps may be added, omitted, or combined. Additionally, features described with reference to certain examples may be combined in other examples.

[0222] Through the description of the above embodiments, those skilled in the art can clearly understand that the methods in the above embodiments can be implemented by means of software plus a necessary general hardware platform. Of course, it can also be implemented by hardware, but in many cases the former is a better implementation method. Based on such an understanding, the technical solution of the present application, in essence, or the part that contributes to the prior art, can be embodied in the form of a computer software product. The computer software product is stored in a storage medium (such as ROM / RAM, magnetic disk, optical disc), and includes several instructions for causing a terminal (which may be a mobile phone, computer, server, or network device, etc.) to execute the methods of the various embodiments of the present application.

[0223] The embodiments of the present application have been described above in conjunction with the accompanying drawings. However, the present application is not limited to the above specific embodiments. The above specific embodiments are merely illustrative rather than restrictive. Under the inspiration of the present application, those of ordinary skill in the art can also make many forms without departing from the purpose of the present application and the scope protected by the claims, and all of them fall within the protection scope of the present application.

Claims

1. An image generation method, characterized in that: include: receiving a first input from a user; In response to the first input, determining at least one first image selected by the user; receiving a second input from the user; In response to the second input, determining a first text; receiving a third input from the user; In response to the third input, at least one of the first images and at least one of the second images are displayed, wherein the at least one of the second images is generated based on the at least one of the first images and the first text, and image contents of the at least one of the first images and the at least one of the second images correspond to the first text.

2. The image generation method according to claim 1, characterized in that: Also includes: receiving a fourth input from the user; In response to the fourth input, displaying a second text, the second text having a greater number of characters than the first text; receiving a fifth input from the user; In response to the fifth input, the first image and at least one of the second images are shared, and the second text is copied.

3. The image generation method according to claim 2, characterized in that: The fifth input includes a first sub-input and a second sub-input, and in response to the fifth input, sharing the first image and at least one of the second images, and copying the second text, comprises: In response to the first sub-input, display at least one sharing control, the at least one sharing control including a first sharing control, the first sharing control corresponding to a first application; In response to the second sub-input of the user to the first sharing control, at least one first image and at least one second image are shared to the first application in a display order of at least one first image and at least one second image, and the second text is copied.

4. The image generation method according to claim 2, characterized in that: Also includes: In response to the second input, determining a second application; The fifth input includes a third sub-input, and in response to the fifth input, sharing the first image and at least one of the second images, and copying the second text, comprises: In response to the third sub-input, at least one of the first images and at least one of the second images are shared to the second application in a display order of at least one of the first images and at least one of the second images, and the second text is copied.

5. The image generation method according to claim 2, characterized in that: The displaying of at least one of the first images and at least one of the second images comprises: displaying at least one of the first images and at least one of the second images in a first order, the first order corresponding to the first text; The displaying of the second text comprises: displaying the second text according to the first text, the first sequence, the image content of at least one of the first images, and the image content of at least one of the second images; Among them, the second text corresponds to the first order.

6. The image generation method according to claim 2, characterized in that: After displaying the second text, the image generation method further includes: receiving a sixth input from the user, wherein the sixth input is used to adjust a display order of at least one of the first images and at least one of the second images to a second order; In response to the sixth input, the second text is adjusted to a third text, the third text corresponding to the second order.

7. The image generation method according to claim 2, characterized in that: After displaying the second text, the image generation method further includes: receiving a seventh input from the user, where the seventh input is used to adjust the second text to a fourth text; In response to the seventh input, at least one of the first images and at least one of the second images are displayed in a third order, the third order corresponding to the fourth text.

8. The image generation method according to any one of claims 1 to 7, characterized in that: The step of determining the first text in response to the second input includes: In response to the user's selection input of a first text label, determining the first text corresponding to the first text label; or In response to the text input by the user, the first text input by the user is determined.

9. The image generation method according to any one of claims 1 to 7, characterized in that: The step of displaying at least one of the first image and at least one of the second image in response to the third input comprises: In response to the third input, analyzing at least one of image information of the first image and text information corresponding to the first text; generating at least one second image according to the image information and the text information; displaying at least one of the first images and at least one of the second images; The image information includes at least one of the following or a combination thereof: image color, image theme, image object; the text information includes at least one of the following or a combination thereof: theme information, emotional trend information, key node information.

10. An image generating device, characterized in that: include: A receiving unit, configured to receive a first input from a user; a processing unit, configured to determine, in response to the first input, at least one first image selected by the user; The receiving unit is further configured to receive a second input from the user; The processing unit is further configured to determine a first text in response to the second input; The receiving unit is further configured to receive a third input from the user; A display unit is used to display at least one of the first images and at least one of the second images in response to the third input, wherein the at least one of the second images is generated based on the at least one of the first images and the first text, and the image content of the at least one of the first image and the at least one of the second images corresponds to the first text.

11. The image generating device according to claim 10, characterized in that: The receiving unit is also used for: receiving a fourth input from the user; The display unit is also used for: In response to the fourth input, displaying a second text, the second text having a greater number of characters than the first text; The receiving unit is also used for: receiving a fifth input from the user; The processing unit is also used for: In response to the fifth input, the first image and at least one of the second images are shared, and the second text is copied.

12. The image generating device according to claim 11, characterized in that: The fifth input includes a first sub-input and a second sub-input, and the display unit is further used for: In response to the first sub-input, display at least one sharing control, the at least one sharing control including a first sharing control, the first sharing control corresponding to a first application; The processing unit is specifically used for: In response to the second sub-input of the user to the first sharing control, at least one first image and at least one second image are shared to the first application in a display order of at least one first image and at least one second image, and the second text is copied.

13. The image generating device according to claim 11, characterized in that: The processing unit is also used for: In response to the second input, determining a second application; The fifth input includes a third sub-input, and the processing unit is specifically configured to: In response to the third sub-input, at least one of the first images and at least one of the second images are shared to the second application in a display order of at least one of the first images and at least one of the second images, and the second text is copied.

14. The image generating device according to claim 11, characterized in that: The display unit is specifically used for: displaying at least one of the first images and at least one of the second images in a first order, the first order corresponding to the first text; displaying the second text according to the first text, the first sequence, the image content of at least one of the first images, and the image content of at least one of the second images; Among them, the second text corresponds to the first order.

15. The image generating device according to claim 11, characterized in that: After displaying the second text, the receiving unit is further used for: receiving a sixth input from the user, wherein the sixth input is used to adjust a display order of at least one of the first images and at least one of the second images to a second order; The processing unit is also used for: In response to the sixth input, the second text is adjusted to a third text, the third text corresponding to the second order.

16. The image generating device according to claim 11, characterized in that: After displaying the second text, the receiving unit is further used for: receiving a seventh input from the user, where the seventh input is used to adjust the second text to a fourth text; The display unit is also used for: In response to the seventh input, at least one of the first images and at least one of the second images are displayed in a third order, the third order corresponding to the fourth text.

17. The image generating device according to any one of claims 10 to 16, characterized in that: The processing unit is specifically used for: In response to the user's selection input of a first text label, determining the first text corresponding to the first text label; or In response to the text input by the user, the first text input by the user is determined.

18. The image generating device according to any one of claims 10 to 16, characterized in that: The processing unit is also used for: In response to the third input, analyzing at least one of image information of the first image and text information corresponding to the first text; generating at least one second image according to the image information and the text information; The display unit is specifically used for: displaying at least one of the first images and at least one of the second images; The image information includes at least one of the following or a combination thereof: image color, image theme, image object; the text information includes at least one of the following or a combination thereof: theme information, emotional trend information, key node information.

19. An electronic device, characterized in that: The method comprises a processor and a memory, wherein the memory stores a program or instruction that can be run on the processor, and when the program or instruction is executed by the processor, the steps of the image generating method according to any one of claims 1 to 9 are implemented.

20. A readable storage medium, characterized in that: The readable storage medium stores a program or an instruction, and when the program or the instruction is executed by a processor, the steps of the image generation method according to any one of claims 1 to 9 are implemented.