Method, apparatus, electronic device, and readable storage medium for generating facial expression images

The method and apparatus automate facial expression generation from initial media content using user inputs, simplifying the expression-making process and integrating it seamlessly with media content editing and sharing, enhancing user experience.

JP2026516090APending Publication Date: 2026-05-19BEIJING ZITIAO NETWORK TECH CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Authority / Receiving Office
JP · JP
Patent Type
Applications
Current Assignee / Owner
BEIJING ZITIAO NETWORK TECH CO LTD
Filing Date
2024-04-25
Publication Date
2026-05-19

AI Technical Summary

Technical Problem

Existing methods for generating facial expressions in media content are complex and cumbersome, requiring multiple steps and interrupting the user's workflow during the process of publishing or communicating.

Method used

A method and apparatus that allow users to generate facial expressions automatically based on initial media content, using inputs such as clicks, voice commands, or gestures, and display corresponding expressions to simplify the expression-making process, enabling seamless integration with media content editing and sharing.

Benefits of technology

Facilitates easy and efficient creation of facial expressions, allowing users to publish or communicate without interruption, by automating the expression generation and simplifying the workflow for publishing or sending messages.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 2026516090000001_ABST
    Figure 2026516090000001_ABST
Patent Text Reader

Abstract

This application discloses a method, apparatus, electronic device and readable storage medium for generating facial expression images, relating to the field of communications technology. The method for generating facial expression images includes receiving a first input when initial media content is acquired, the first input being used to trigger the generation of a first facial expression, the first facial expression being a facial expression image generated based on the initial media content, the facial expression image being used to edit and / or transmit target content, the initial media content including captured content and / or uploaded content, and displaying target information in response to the first input, the target information relating to the first facial expression.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] [Cross - Reference to Related Applications] This application claims the priority of a Chinese patent application with application number 202310505294.2 filed on May 6, 2023, and this application is incorporated herein by reference. This application relates to the field of communication technologies, specifically to a method for generating expression images, an apparatus for generating expression images, an electronic device, and a readable storage medium.

Background Art

[0002] With the continuous development of electronic communication technologies, users often use various expressions to comment on works or communicate with other users in the process of using various Internet platforms.

Summary of the Invention

Problems to be Solved by the Invention

[0003] The purpose of the embodiments of this application is to provide a method for generating expression images, an apparatus, an electronic device, and a readable storage medium. When a user uploads or takes a media content as a personal work, this application can simply add this media content as an expression, realizing the simplification of the user's expression - making process.

Means for Solving the Problems

[0004] According to a first aspect, an embodiment of the present application provides a method for generating an expression image, wherein on the first user side, the method for generating an expression image includes receiving a first input when initial media content is acquired, the first input being used to trigger the generation of a first expression, the first expression being an expression image generated based on the initial media content, the expression image being used to edit and / or transmit target content, the initial media content including captured content and / or uploaded content, and displaying target information in response to the first input, the target information relating to the first expression.

[0005] According to a second aspect, an embodiment of the present application provides an expression image generation device, on the first user side, the expression image generation device includes a receiving module for receiving a first input when initial media content is acquired, the first input being used to trigger the generation of a first expression, the first expression being an expression image generated based on the initial media content, the expression image being used to edit and / or transmit target content, and the initial media content being captured content and / or uploaded content; and a display module for displaying target information in response to the first input, the display module being related to the first expression.

[0006] According to a third aspect, an embodiment of the present application provides an electronic device comprising a processor, memory, and a program or instruction stored in the memory and executable on the processor, wherein when the program or instruction is executed by the processor, the steps of the method of the first aspect are realized.

[0007] According to a fourth aspect, an embodiment of the present application provides a readable storage medium on which a program or instruction is stored, and when this program or instruction is executed by a processor, the steps of the facial expression image generation method of the first aspect are realized.

[0008] According to a fifth aspect, an embodiment of the present application provides a chip comprising a processor and a communication interface, the communication interface being coupled with the processor, the processor being used to execute a program or instruction and to implement a step of the facial image generation method of the first aspect.

[0009] According to the sixth aspect, an embodiment of the present application provides a computer program product which is stored in a storage medium and is executed by at least one processor to realize the facial expression image generation method of the first aspect.

[0010] According to the seventh aspect, an embodiment of the present application provides a computer program which, when executed by a processor, implements the steps of the facial expression image generation method of the first aspect. [Effects of the Invention]

[0011] In the embodiments of this application, when a user publishes a work or chats with another user, after uploading initial media content or taking a picture of initial media content, the first user can respond to the first input by automatically generating a corresponding first facial expression based on the initial media content and displaying target information to prompt the user. [Brief explanation of the drawing]

[0012] [Figure 1] This is a flowchart of the method for generating facial expression images according to the embodiment of this application. [Figure 2] This is one of the schematic diagrams of a media content generation interface according to several embodiments of this application. [Figure 3] This is the second schematic diagram of a media content generation interface according to some embodiments of this application. [Figure 4] This is the third schematic diagram of a media content generation interface according to some embodiments of this application. [Figure 5] This is the fourth schematic diagram of a media content generation interface according to several embodiments of this application. [Figure 6] This is a schematic diagram of an expression generation interface according to several embodiments of this application. [Figure 7] This is a block diagram of a facial expression image generation device according to an embodiment of the present application. [Figure 8] This is a block diagram of an electronic device according to an embodiment of this application. [Figure 9] This is a schematic diagram of the hardware structure of the electronic device according to the embodiment of this application. [Modes for carrying out the invention]

[0013] In the following, the technical concepts in the embodiments of this application will be clearly described, with reference to the drawings of the embodiments. It is clear that the embodiments described are only some, and not all, embodiments of this application. All other embodiments obtained by those skilled in the art based on the embodiments of this application are all within the scope of protection of this application.

[0014] The terms "first," "second," etc., used in the specification and claims of this application are intended to distinguish similar subjects and not to describe a specific order or sequence. It should be understood that these terms are interchangeable where appropriate, so that the embodiments of this application may be carried out in an order other than those illustrated or described herein, and that the subjects distinguished by "first," "second," etc., are generally of the same kind and do not limit the number of subjects; for example, the first subject may be one or more. Furthermore, "and / or" in the specification and claims indicates at least one of the connected subjects, and the letter " / " generally indicates that the preceding and succeeding related subjects are in an "or" relationship.

[0015] In the following, with reference to FIGS. 1 to 9, a method for generating an expression image according to an embodiment of the present application, an apparatus for generating an expression image, an electronic device, and a readable storage medium will be described in detail by way of specific embodiments and their application scenarios.

[0016] An embodiment of the present application provides a method for generating an expression image. FIG. 1 shows a flowchart of the method for generating an expression image according to an embodiment of the present application. As shown in FIG. 1, for the method for generating an expression image, on the first user side, the method for generating an expression image includes the following.

[0017] Step 102: When initial media content is obtained, receive a first input, where the first input is used to trigger the generation of a first expression, and the first expression is an expression image generated based on the initial media content. The expression image is used to edit and / or transmit target content. The initial media content includes captured content and / or uploaded content. In an embodiment of the present application, the initial media content may include pictures or video content. The initial media content may be media content uploaded by the user on the first user side, or media content captured by the user on the first user side.

[0018] Exemplarily, the user uploads a video file to the platform, determines the video file uploaded by the user as the initial media content, and after this video file is uploaded, determines this video content as the obtained initial media content.

[0019] Exemplarily, the user captures image data using the shooting function of the platform and determines the captured image data as the initial media content.

[0020] In an embodiment of the present application, the first input is used to trigger the generation of a first expression. The first expression is an expression image generated based on initial media content. The expression image is used to edit and / or transmit target content. The initial media content includes captured content and / or uploaded content. In an embodiment of the present application, the target content includes communication message information when a user conducts an instant communication session with another user. The target content may further include comment information published by the user on the works of the platform. The target content may further include character information and image information used in the user's personal works.

[0021] Exemplarily, when chatting with another user, the user can send the first expression to the chat dialog box. When editing a text message, the user can insert the first expression into the text message and send it together within the chat session page.

[0022] Exemplarily, when commenting on a media work on the platform, the user can choose to send the first expression to the comment column.

[0023] Exemplarily, when uploading a video work or an image work, the user can insert the first expression into the image or character information of the video work or the image and text work.

[0024] In an embodiment of the present application, the first input may be a click input on a control in an interface where the user has initial media content, or a voice command input by the user, or a specific gesture input by the user. Specifically, it can be determined according to the actual usage requirements, and the embodiments of the present application are not limited thereto.

[0025] The specific gesture in the embodiments of this application may be any one of the following: single-click gesture, swipe gesture, drag gesture, pressure-recognition gesture, long-press gesture, area-change gesture, two-point press gesture, or double-click gesture. The click input in the embodiments of this application may be a single-click input, a double-click input, or any number of clicks, and may also be a long-press input or a short-press input.

[0026] Step 104: In response to the first input, the target information is displayed, where the target information relates to the first facial expression.

[0027] For example, target information includes facial expression addition success prompt information, which is used to give a prompt indicating that the corresponding first facial expression has already been generated based on the initial media content and that the memory of this first facial expression has already been completed.

[0028] For example, target information includes facial expression images, and after generating a first facial expression based on the initial media content, the user can preview the generated first facial expression by displaying the facial expression image corresponding to the first facial expression.

[0029] In the embodiments of this application, when a user publishes a work or chats with another user, after uploading or capturing initial media content, the first user can respond to the first input by automatically generating a corresponding first facial expression based on the initial media content and displaying target information to prompt the user. This allows the user to easily create a corresponding facial expression based on the image acquired in the process of publishing a personal work or sending an instant communication message, thereby providing a simple facial expression generation path and simplifying the steps of facial expression creation for the user.

[0030] In some embodiments of this application, a method for generating facial image further includes, after displaying target information, generating target media content on a media content generation interface based on initial media content, and, in response to a target media content determination operation, publishing the target media content and / or transmitting the target media content to at least one second user.

[0031] In the embodiments of this application, after displaying target information related to a first facial expression, the user can continue editing the initial media content to generate target media content, and after the user has made a decision on the target media content, the target media content can be published as a personal work or transmitted to at least one second user.

[0032] Here, the media content generation interface displays the initial media content uploaded or captured by the user. Within the media generation interface, the user can edit the initial media content and, in response to a decision, upload and publish or send to other users the completed target media content based on the initial media content.

[0033] Specifically, after adding the first facial expression based on the initial media content, the user can perform editing operations on the initial media content in the media content generation interface, such as adding stickers, adjusting filters or special effects, and adjusting the cutting size, and the edited initial media content becomes the target media content.

[0034] Figure 2 shows one schematic diagram of a media content generation interface according to several embodiments of the present application. As shown in Figure 2, after the user shoots a video work 202, the media content generation interface displays a preview of this video work 202, and after the first facial expression is created based on the first input and completed, facial expression addition completion information 204 is displayed. The user edits this video work 202 to obtain the edited video work 202, and the user can upload and publish the edited video work 202 as target media content by clicking the "Publish" button 206. Here, the facial expression addition completion information is target information, the video work 202 is the initial media content, the edited video work 202 is the target media content, and the user clicking the "Publish" button 206 is the user's decision operation regarding the target media content.

[0035] In embodiments of this application, the ability to generate a first facial expression based on initial media content in the process of a user publishing a work or sending a session message is provided, and after generating the first facial expression, the user can continue editing the initial media content to obtain target media content, and after the user performs a decision operation on the target media content, publish this target content or send this target content to other users, and ensure that the facial expression creation process does not interrupt the user's progress in uploading or sending the relevant target media content.

[0036] In some embodiments of this application, when initial media content is acquired, receiving a first input includes presenting the initial media content to a media content generation interface and confirming receipt of the first input in response to a trigger operation on the media content generation interface.

[0037] In the embodiments of this application, the first input is a trigger operation on the user's media content generation interface. This first input may be a trigger operation on an operation control in the media content generation interface, or it may be a trigger operation on the initial media content in the media content generation interface. After receiving the corresponding trigger operation, the system determines that the first input has been received and continues to respond to this first input by performing actions such as displaying subsequent target information.

[0038] Here, the media content generation interface displays the initial media content uploaded or captured by the user, and the user can trigger an operation in the media generation interface to generate a first facial expression based on the initial media content using a first input.

[0039] Figure 3 shows a second schematic diagram of a media content generation interface according to some embodiments of the present application. As shown in Figure 3, after the user has filmed a video work 302, the media content generation interface displays a preview of the video work 302. The user then long-presses the video work 302 in the media content generation interface to trigger an operation to generate a first facial expression based on the video work 302. After the first facial expression is produced and completed, facial expression addition completion information 304 is displayed. Here, the video work 302 is the initial media content, the user's long-press operation on the media content generation interface is the first input, and the facial expression addition completion information is the target information. In some other embodiments, the media content generation interface presents a facial expression addition control, and the first input may further be a trigger operation for the facial expression addition control.

[0040] In the embodiments of this application, the initial media content can be displayed on the media content generation interface, and the user can perform a first input to trigger the generation of a first facial expression based on the initial media content by triggering an operation on the media content generation interface, thereby enabling the user to generate the corresponding first facial expression with simple operation when the initial media content is displayed, and further simplifying the user's facial expression creation operation flow.

[0041] In some embodiments of this application, a first input includes a first sub-input and a second sub-input, and receiving the first input when initial media content is obtained includes receiving a first sub-input for a first control in the media content generation interface when the initial media content is displayed in the media content generation interface, displaying a second control in the content editing interface in response to the first sub-input, and receiving a second sub-input for the second control, where the second control includes an expression addition control.

[0042] In the embodiments of this application, both the first sub-input and the second sub-input are sub-inputs of the first input. The first control includes a content share button for sharing initial media content, and the second control includes an expression add control for adding the generated first expression to the expression favorites.

[0043] Figure 4 shows the third schematic diagram of a media content generation interface according to some embodiments of the present application. As shown in Figure 4, after the user has filmed a video 402, the video 402 is displayed in the media content generation interface, and a content share button 404 is further displayed in the media generation interface. After the user clicks the content share button 404, a floating window 406 pops up in the media generation interface, and an expression add button 408 is displayed in the floating window 406. When the user clicks the expression add button 408, it can trigger an action that generates a corresponding expression image based on the video 402 and adds this expression image to the expression favorites. Here, the video 402 is the initial media content, the content share button 404 is the first control, the expression add button 408 is the second control, the user clicking the content share button 404 is the first sub-input, and the user clicking the expression add button 408 is the second sub-input.

[0044] In some other embodiments, the trigger for the first sub-input in the user's media content generation interface may also be a pre-configured trigger operation on the media content generation interface, for example, a long press operation on the screen area of ​​the media generation interface, after which the second control appears.

[0045] In the embodiments of this application, a user can easily generate an expression image based on initial media content by sequentially executing a first sub-input and a second sub-input for a first control and a second control in the media content generation interface, and can also add the expression image to their favorites, thereby ensuring ease of operation for the user in the media content generation interface.

[0046] In some embodiments of this application, the first facial expression includes either a dynamic facial expression image or a static facial expression image.

[0047] In the embodiments of this application, the first facial expression may be a dynamic facial expression image or a static facial expression image.

[0048] In embodiments of this application, the first facial expression may be stored as a dynamic facial expression image in GIF format, and the first facial expression may be convertible between a static facial expression image and a dynamic facial expression image, that is, a frame image from a dynamic facial expression image may be cropped as a static facial expression image. Exemplaryly, when the initial media content is a dynamic image or video, the user interface may display a button to switch between static and dynamic facial expressions and / or support the user in selecting the image frame that generates the first facial expression.

[0049] For example, when the first facial expression is a dynamic facial expression image, and the user sends this first facial expression to a dialog box, this first facial expression is dynamic in the dialog box. When the first facial expression is a static facial expression image, this first facial expression remains fixed in the dialog box.

[0050] In the embodiments of this application, dynamic facial expression images can be generated from acquired initial media content, and static facial expression images can also be generated, improving the flexibility of users to create their own facial expressions, and allowing users to choose whether to create dynamic or static facial expressions according to their actual needs.

[0051] In some embodiments of this application, the initial media content includes a first video, and before displaying target information, it includes generating a first facial expression based on the first video, where the first facial expression is a dynamic facial expression image.

[0052] In the embodiments of this application, the initial media content includes a first video, which may be an uploaded video or a recorded video. When generating a first facial expression using the first video, the first video can be converted into a dynamic image, and this dynamic image can be directly used as a dynamic facial expression image corresponding to the first facial expression. Exemplarily, in the process of converting the first video into a dynamic facial expression image, some video frames in the first video are deleted to reduce the amount of data in the generated dynamic facial expression image.

[0053] For example, in the process of converting a first video into a dynamic facial expression image, the interval between at least two video frames in the first video is adjusted to adjust the effect of playback on the generated dynamic facial expression image.

[0054] For example, in the process of converting a first video into a dynamic facial expression image, the playback order between at least two video frames in the first video is adjusted to adjust the display content in the generated dynamic facial expression image.

[0055] In the embodiments of this application, when the initial media content acquired is a video, it is possible to generate corresponding dynamic facial expressions based on the video, thereby simplifying the operation flow required when a user creates dynamic facial expressions.

[0056] In some embodiments of this application, the initial media content includes a first video, and before displaying target information, it includes generating a first facial expression based on the first video, where the first facial expression is a static facial expression image.

[0057] In the embodiments of this application, by extracting a target video frame from a first video and generating a first facial expression using the target video frame as a static facial expression image, the first facial expression generated based on the first video may be a dynamic or static facial expression. The user can flexibly choose whether to create a dynamic or static facial expression based on the first video, thereby simplifying the user's facial expression creation process and simultaneously improving the user's flexibility in selecting facial expressions.

[0058] For example, a target video frame is automatically extracted from the first video, where the target video frame is the first, last, or middle video frame of the first video.

[0059] In some embodiments of this application, generating a first facial expression based on a first video includes determining a first dynamic image as a first facial expression if the video length of the first video is less than or equal to a predetermined length, wherein the first dynamic image is a dynamic image obtained by transforming the first video.

[0060] In the embodiments of this application, by comparing the video length of the first video with a preset length, when the video length of the first video is less than or equal to the preset length, the first dynamic image is directly converted based on the complete first video, and this first dynamic image is determined to be the first facial expression.

[0061] Here, the predetermined range of length values ​​is from 5 seconds to 10 seconds.

[0062] For example, if the preset length is 10 seconds and the video length of the first video is identified as 8 seconds, the format of the first video is directly converted to a GIF image, and this GIF image is stored in the expression favorites as the first expression, the expression favorites may be an expression memory area, and this memory area matches the user's account, for example, the expression favorites may be an expression panel.

[0063] In the embodiments of this application, by comparing the video length of a first video with a preset length, and automatically performing a format conversion on the first video whose video length does not exceed the preset length, it is possible to obtain a first facial expression, ensuring that the amount of data of the obtained first facial expression is relatively small, and the user does not need to perform processing such as cutting on the manually acquired video to complete the production of the corresponding facial expression, further simplifying the user's operation steps when creating facial expressions.

[0064] In some embodiments of this application, generating a first facial expression based on a first video includes determining a second video in the first video if the video length of the first video is greater than a preset length, wherein the video length of the second video is less than or equal to a preset length, and determining a second dynamic image as the first facial expression, wherein the second dynamic image is a dynamic image obtained by transforming the second video.

[0065] In the embodiments of this application, the second video is a portion of the first video, the first video contains all the video content in the second video, and the video length of the first video is greater than the video length of the second video. In the embodiments of this application, by comparing the video length of the first video with a preset length, the first video is cut when the video length of the first video is greater than the preset length, the first video is cut into a second video whose video length is less than or equal to the preset length, the second video is converted into a second dynamic image, and the second dynamic image is determined to be the first facial expression.

[0066] In some possible embodiments, the first video can be automatically cut based on a preset length.

[0067] For example, a target video frame segment is automatically cut from the first video to a predetermined length, and this target video frame segment becomes the second video. Here, the selection position of the target video frame segment may be predetermined by the user, and the selection position of the target video frame segment includes the beginning segment of the first video, the end segment of the first video, or the middle segment of the first video.

[0068] In the embodiments of this application, the video length of the first video is compared with a preset length, and when the video length of the first video exceeds the preset length, the second video is cut from the first video. By determining the second dynamic image obtained by conversion based on the second video as the first facial expression, it is possible to ensure that even when the video length of the first video is too long, the amount of data obtained for the first facial expression is still relatively small.

[0069] In some embodiments of this application, determining a second video in a first video includes displaying initial media content and a third control, receiving a second input to the third control, wherein the third control includes a video cutting control, and cutting out the second video in the first video in response to the second input.

[0070] In the embodiments of this application, a third control is used to cut a first video in the initial media content, and a second input is used to trigger a cutting operation on the first video. Here, the user can select a target frame segment in the first video by performing a second input to the third control, and cut a second video from the first video based on the selected target frame segment.

[0071] Figure 5 shows the fourth schematic diagram of a media content generation interface according to some embodiments of the present application, and as shown in Figure 5, the media content generation interface displays a first video 502 and a video cutting control 506. The user cuts a second video 504 from the first video 502 by dragging a floating label 508 in the video cutting control 506. A preview image 510 of the video is displayed at the position associated with the video cutting control 506, and the user determines the video fragment that needs to be cut (for example, the preview image is currently displayed above the floating label 508 in Figure 5) by checking the preview image 510 at the corresponding position when dragging the floating label 508. Here, the video cutting control is a third control, and the user dragging the floating label in the video cutting control is a second input.

[0072] In the embodiments of this application, the user can manually cut the first video by operating a third control, thereby obtaining a second video, and the content of the first facial expressions generated based on the second video is manually selected by the user, ensuring that the resulting first facial expressions meet the user's actual needs.

[0073] In some embodiments of this application, the method for generating facial expression images further includes generating a first facial expression based on initial media content.

[0074] In the embodiments of this application, a first facial expression is generated based on initial media content displayed on the media content generation interface before displaying target information. The initial media content may be video content, image content, or text content, and is media content uploaded or photographed by the user, and is not specifically limited to any particular type.

[0075] For example, when the initial media content is video content, a static first expression can be generated by extracting video frames from the video content, and furthermore, a portion of the video content can be converted into a dynamic image to obtain a dynamic first expression.

[0076] For example, when the initial media content is an image, the image can be directly used as the first expression. If the image content is a static image, the first expression is a static expression; if the image content is a dynamic image, the first expression is a dynamic image. When there are multiple images, it is possible to choose to combine the multiple images to obtain a dynamic first expression.

[0077] For example, when the initial media content is text content, a text image can be generated based on the text content, and this text image will be the primary expression.

[0078] In the embodiments of this application, a first facial expression is generated based on initial media content, and it is possible to ensure that different types of initial media content can generate corresponding first facial expressions, thereby allowing the user to flexibly select the type of initial media content uploaded or captured, and thereby produce different types of first facial expressions.

[0079] In some embodiments of this application, generating a first facial expression based on initial media content includes displaying target information and a fourth control, wherein the target information includes a facial expression image generated based on the initial media content; receiving a third input to the fourth control, wherein the third input is used to edit the facial expression image; and generating a first facial expression based on the edited facial expression image.

[0080] In the embodiments of this application, the fourth control is for editing the generated facial expression image, and the user can perform editing operations on the facial expression image, such as inserting stickers or text, by performing the third input to this fourth control.

[0081] In the embodiments of this application, in the process of generating a first facial expression based on initial media content, a facial expression image generated based on the initial content and a fourth control for editing this facial expression image are displayed. The user can edit the facial expression image using the displayed fourth control, and after editing, decide that the facial expression image is the first facial expression, and then add the first facial expression to the facial expression favorites.

[0082] For example, editing an expression image includes at least one of the following: inserting text into the expression image, inserting stickers into the expression image, adjusting the color parameters of the expression image, adjusting the size parameters of the expression image, and adjusting the playback speed of the expression image.

[0083] To make it easier to understand, after receiving the first input, the system can jump to the facial expression generation interface to display target information, which may include initial media content or intermediate media content already edited in the media content generation interface. The user's editing operations in the facial expression generation interface generate the first facial expression. After generating the first facial expression, the system can return to the media content generation interface and continue editing the initial media content or the media content before entering the facial expression generation interface to generate the target media content.

[0084] Figure 6 shows a schematic diagram of an expression generation interface according to several embodiments of the present application. As shown in Figure 6, the expression generation interface displays an expression image 602 and an image editing control 604. After the user clicks the text insertion option in the image editing control 604, a text input box 606 is displayed. The user completes the editing by inserting text into the expression image by entering text content 608 into the text input box 606, after which the user clicks the "Confirm" button 610. Here, the user clicking the text insertion option in the image editing control 604 and entering text into the text input box 606 is a third input, and the image editing control 604 is a fourth control.

[0085] In the embodiments of this application, before storing the facial expression image generated based on the initial media content as the first facial expression, the user can edit the facial expression image using a fourth control, and the edited facial expression image can be used as the first facial expression. This allows the user to perform secondary editing on the facial expression image generated based on the initial media content, further improving the flexibility of facial expression creation.

[0086] In some embodiments of this application, the initial media content includes at least one preset image, and before displaying target information, the at least one preset image is determined to be a first expression, the number of first expressions matching the number of preset images.

[0087] In the embodiments of this application, the pre-set image may be an image uploaded by the user or an image that has been captured. When the initial media content is a pre-set image, the pre-set image is directly stored in the favorite expression as the first expression.

[0088] For example, when there are multiple pre-set images, each pre-set image is designated as a first expression, and then multiple first expressions are stored in the expression favorites. The number of first expressions is the same as the number of pre-set images, and each first expression corresponds one-to-one with a pre-set image.

[0089] For example, the pre-set image may be a static image, and the first facial expression corresponding to the pre-set image is a static facial expression. The pre-set image may also be a dynamic image, and the first facial expression corresponding to the pre-set image is a dynamic facial expression. For example, the dynamic image may be a Live Photo, and this dynamic image may be converted to a GIF image and used as the first facial expression.

[0090] In the embodiment of this application, when the initial media content consists of multiple pre-set images, each of the pre-set images is designated as a corresponding first facial expression. This allows multiple first facial expressions to be added to the favorite facial expressions in a single step, thereby achieving the effect of producing a large number of facial expressions.

[0091] In some embodiments of this application, the number of preset images is at least two, and determining at least one preset image as the first expression includes determining a target image as the first expression in response to a selection operation for a target image among at least two preset images, wherein the target image is at least one of the at least two preset images.

[0092] In the embodiments of this application, when the initial media content consists of multiple pre-configured images, the user can select multiple pre-configured images through a selection operation, and the selected target image is determined as the first facial expression. Here, the user can select one or more target images from the multiple pre-configured images.

[0093] As an example, a media content generation interface displays multiple pre-configured images, and the user clicks on a target image among the multiple pre-configured images. The pre-configured image clicked after the click is designated as the target image, and a selection indicator is displayed on the target image. Here, the click input performed by the user when clicking on the target image is a selection operation.

[0094] In the embodiments of this application, when the initial media content includes at least two pre-set images, the at least two pre-set images are displayed, and the user can choose to generate a portion of the initial media content into the corresponding first facial expression by selecting a target image from among them through a selection operation.

[0095] In some embodiments of this application, the initial media content includes at least two pre-set images, and before displaying target information, the at least two pre-set images are combined into a first facial expression, where the first facial expression is a dynamic facial expression image.

[0096] In the embodiments of this application, when the initial media content consists of at least two pre-set images, the at least two pre-set images can be combined, and the combined image is designated as the first facial expression. The facial expression image obtained by combining at least two pre-set images is a dynamic facial expression image.

[0097] For example, by synthesizing at least two pre-defined images based on the arrangement order of at least two pre-defined images to generate a dynamic facial expression image, the display order of the at least two pre-defined images in the dynamic facial expression image matches the arrangement order.

[0098] In the embodiments of this application, the user can set the display time length of each pre-set image in the synthesized dynamic facial expression image. For example, the pre-set images include a first image and a second image, where the display time length of the first image in the dynamic facial expression image is 1 second and the display time length of the second image is 3 seconds.

[0099] In the embodiments of this application, when the initial media content consists of at least two pre-set images, the at least two pre-set images can be combined into a dynamic facial expression image, and the resulting dynamic facial expression image can be designated as the first facial expression. This eliminates the need for the user to combine multiple pre-set images in other applications, and allows for the rapid generation of the combined dynamic facial expression image.

[0100] In some embodiments of this application, the initial media content includes at least two pre-set images, and before displaying target information, the at least two pre-set images are combined into a first facial expression, where the first facial expression is a static facial expression image.

[0101] In the embodiments of this application, when the initial media content consists of at least two pre-set images, the at least two pre-set images can be combined in a tile-like manner. For example, the area ratio occupied by each pre-set image in the static facial expression image obtained by combining them can be freely set by the user.

[0102] In some embodiments of this application, the pre-set image includes either a pre-set static image or a pre-set dynamic image. In some embodiments, the pre-set dynamic image may include a dynamic image or a video, etc.

[0103] In the embodiments of this application, when the initial media content is a pre-set static image or a pre-set dynamic image, a first facial expression can be generated based on this pre-set image, thereby improving the flexibility of the user's facial expression creation.

[0104] For example, a dynamic first facial expression can be obtained by combining multiple pre-set static images.

[0105] For example, a static first facial expression can be obtained by cropping an image frame from a pre-defined dynamic image.

[0106] In some embodiments of this application, in response to a first input, a first facial expression is stored in a local memory area after displaying target information, and / or the first facial expression is added to a facial expression panel, the facial expression panel being used to generate target content.

[0107] In the embodiments of this application, after generating a first facial expression based on initial media content, this first facial expression can be saved, which may be in a local storage area or uploaded to a server, whereupon the server binds this first facial expression to a user account and adds it to the facial expression panel corresponding to the account.

[0108] Specifically, the user can pre-set the memory location of the first facial expression, thereby choosing to store the first facial expression locally and / or upload and save it to the server.

[0109] In the embodiments of this application, the convenience of the user using the first facial expression can be improved by saving the generated first facial expression in a local memory area and / or adding it to the facial expression panel.

[0110] In some embodiments of the present application, the target content includes an expression image, and the method further includes generating the target content based on a first expression in an expression panel, and transmitting and / or presenting the target content to at least one second user in response to a target content determination operation.

[0111] In the embodiments of this application, when a user edits target content, the user can edit target content with a first facial expression, publish this target content, and / or transmit it to at least one second user.

[0112] For example, the target content includes social chat information that a user needs to send to a second user.

[0113] For example, target content includes video works published and uploaded by users.

[0114] For example, target content includes comment information on video works posted by other users.

[0115] In the method for generating facial expression images according to the embodiments of this application, the execution unit may be a facial expression image generation device. In the embodiments of this application, the facial expression image generation device according to the embodiments of this application will be described, with the example that the facial expression image generation device performs the facial expression image generation method.

[0116] In some embodiments of this application, an expression image generation device is provided, and Figure 7 shows a block diagram of an expression image generation device according to an embodiment of this application, and as shown in Figure 7, with respect to the expression image generation device 700, on the first user side, the expression image generation device 700 is A receiving module 702 for receiving a first input when initial media content is acquired, wherein the first input is used to trigger the generation of a first facial expression, the first facial expression is a facial expression image generated based on the initial media content, the facial expression image is used to edit and / or transmit target content, and the receiving module 702 includes captured content and / or uploaded content, A display module 704 for displaying target information in response to a first input, the display module 704 including a display module 704 for displaying target information relating to a first facial expression.

[0117] In the embodiments of this application, when a user publishes a work or chats with another user, after uploading or capturing initial media content, the first user can respond to the first input by automatically generating a corresponding first facial expression based on the initial media content and displaying target information to prompt the user. This allows the user to easily create a corresponding facial expression based on the image acquired in the process of publishing a personal work or sending an instant communication message, thereby providing a simple facial expression generation path and simplifying the steps of facial expression creation for the user.

[0118] In some embodiments of this application, the display module 704 is used to generate target media content based on initial media content in the media content generation interface after displaying target information. The facial expression image generation device 700 is A presentation module for announcing target media content in response to a target media content determination operation. and / or include a transmission module for sending target media content to at least one second user.

[0119] In embodiments of this application, the ability to generate a first facial expression based on initial media content in the process of a user publishing a work or sending a session message is provided, and after generating the first facial expression, the user can continue editing the initial media content to obtain target media content, and after the user performs a decision operation on the target media content, publish this target content or send this target content to other users, and ensure that the facial expression creation process does not interrupt the user's progress in uploading or sending the relevant target media content.

[0120] In some embodiments of this application, the display module 704 is used to present initial media content to the media content generation interface. The facial expression image generation device 700 is It includes a decision module for confirming the acceptance of a first input in response to a trigger operation on the media content generation interface.

[0121] In the embodiments of this application, the initial media content can be displayed on the media content generation interface, and the user can perform a first input to trigger the generation of a first facial expression based on the initial media content by triggering an operation on the media content generation interface, thereby enabling the user to generate the corresponding first facial expression with simple operation when the initial media content is displayed, and further simplifying the user's facial expression creation operation flow.

[0122] In some embodiments of this application, the first input includes a first sub-input and a second sub-input. The receiving module 702 is used to receive a first sub-input to the first control in the media content generation interface when displaying initial media content in the media content generation interface. The display module 704 is used to display a second control on the content editing interface in response to a first sub-input. The receiving module 702 is used to receive a second sub-input to a second control, where the second control includes an expression addition control.

[0123] In the embodiments of this application, a user can easily generate an expression image based on initial media content by sequentially executing a first sub-input and a second sub-input for a first control and a second control in the media content generation interface, and can also add the expression image to their favorites, thereby ensuring ease of operation for the user in the media content generation interface.

[0124] In some embodiments of this application, the first facial expression includes either a dynamic facial expression image or a static facial expression image.

[0125] In the embodiments of this application, dynamic facial expression images can be generated from acquired initial media content, and static facial expression images can also be generated, improving the flexibility of users to create their own facial expressions, and allowing users to choose whether to create dynamic or static facial expressions according to their actual needs.

[0126] In some embodiments of this application, the initial media content includes a first video. The facial expression image generation device 700 is It includes a generation module for generating a first facial expression based on a first video, where the first facial expression is a dynamic facial expression image.

[0127] In the embodiments of this application, when the initial media content acquired is a video, it is possible to generate corresponding dynamic facial expressions based on the video, thereby simplifying the operation flow required when a user creates dynamic facial expressions.

[0128] In some embodiments of this application, the initial media content includes a first video. The generation module is used to generate a first facial expression based on a first video, where the first facial expression is a static facial expression image.

[0129] In the embodiments of this application, a target video frame is extracted from a first video, and a first facial expression is generated using the target video frame as a static facial expression image. The first facial expression generated based on the first video may be a dynamic or static facial expression, and the user can flexibly choose whether to create a dynamic or static facial expression based on the first video. This simplifies the user's facial expression creation process while simultaneously improving the user's flexibility in selecting facial expressions.

[0130] In some embodiments of this application, a determination module is used to determine a first dynamic image as a first facial expression when the video length of a first video is less than or equal to a preset length, and the first dynamic image is a dynamic image obtained by transforming the first video.

[0131] In the embodiments of this application, by comparing the video length of a first video with a preset length, and automatically performing a format conversion on the first video whose video length does not exceed the preset length, it is possible to obtain a first facial expression, ensuring that the amount of data of the obtained first facial expression is relatively small, and the user does not need to perform processing such as cutting on the manually acquired video to complete the production of the corresponding facial expression, further simplifying the user's operation steps when creating facial expressions.

[0132] In some embodiments of this application, the determination module is used to determine a second video in the first video when the video length of the first video is greater than a preset length, and the video length of the second video is less than or equal to the preset length. The decision module is used to determine the second dynamic image as the first facial expression, and the second dynamic image is the dynamic image obtained by the transformation of the second video.

[0133] In the embodiments of this application, the video length of the first video is compared with a preset length, and when the video length of the first video exceeds the preset length, the second video is cut from the first video. By determining the second dynamic image obtained by conversion based on the second video as the first facial expression, it is possible to ensure that even when the video length of the first video is too long, the amount of data obtained for the first facial expression is still relatively small.

[0134] In some embodiments of this application, the display module 704 is used to display initial media content and a third control. The receiving module is used to receive a second input to a third control, and the third control includes a video cutting control. The facial expression image generation device 700 is It includes a cropping module for cropping the second video from the first video in response to a second input.

[0135] In the embodiments of this application, the user can manually cut the first video by operating a third control, thereby obtaining a second video, and the content of the first facial expression generated based on the second video is manually selected by the user, ensuring that the resulting first facial expression matches the user's actual needs.

[0136] In some embodiments of this application, a generation module is used to generate a first facial expression based on initial media content.

[0137] In the embodiments of this application, a first facial expression is generated based on initial media content, and it is possible to ensure that different types of initial media content can generate corresponding first facial expressions, thereby allowing the user to flexibly select the type of initial media content uploaded or captured, and thereby produce different types of first facial expressions.

[0138] In some embodiments of this application, the display module 704 is used to display target information and a fourth control, the target information includes an expression image generated based on the initial media content, The receiving module 702 receives a third input to the fourth control, and the third input is used to edit the facial expression image. The generation module is used to generate the first facial expression based on the edited facial expression image.

[0139] In the embodiments of this application, before storing the facial expression image generated based on the initial media content as the first facial expression, the user can edit the facial expression image using a fourth control, and the edited facial expression image can be used as the first facial expression. This allows the user to perform secondary editing on the facial expression image generated based on the initial media content, further improving the flexibility of facial expression creation.

[0140] In some embodiments of this application, a determination module is used to determine at least one preset image as a first expression, and the number of first expressions matches the number of preset images.

[0141] In the embodiment of this application, when the initial media content is a pre-set image, each of the pre-set images is designated as a corresponding first facial expression. This allows multiple first facial expressions to be added to the favorite facial expressions in a single step when the initial media content consists of multiple pre-set images, thereby achieving the effect of producing a large number of facial expressions.

[0142] In some embodiments of this application, the number of pre-set images is at least two. The decision module is used to determine the target image as the first facial expression in response to a selection operation on the target image among at least two pre-set images, the target image being at least one of the at least two pre-set images.

[0143] In the embodiments of this application, when the initial media content includes at least two pre-set images, the at least two pre-set images are displayed, and the user can choose to generate a portion of the initial media content into the corresponding first facial expression by selecting a target image from among them through a selection operation.

[0144] In some embodiments of this application, the initial media content includes at least two pre-set images. The facial expression image generation device 700 is It includes a synthesis module for combining at least two pre-defined images into a first facial expression, where the first facial expression is a dynamic facial expression image.

[0145] In the embodiments of this application, when the initial media content consists of at least two pre-set images, the at least two pre-set images can be combined into a dynamic facial expression image, and the resulting dynamic facial expression image is designated as the first facial expression. The user does not need to combine multiple pre-set images in other applications, and the combined dynamic facial expression image can be generated quickly.

[0146] In some embodiments of this application, the initial media content includes at least two pre-set images. The synthesis module is used to synthesize at least two pre-defined images into a first facial expression, where the first facial expression is a static facial expression image.

[0147] In the embodiments of this application, when the initial media content consists of at least two pre-set images, the at least two pre-set images can be combined in a tile-like manner. For example, the area ratio occupied by each pre-set image in the static facial expression image obtained by combining them can be freely set by the user.

[0148] In some embodiments of this application, the pre-set image includes either a pre-set static image or a pre-set dynamic image.

[0149] In the embodiments of this application, when the initial media content is a pre-set static image or a pre-set dynamic image, a first facial expression can be generated based on this pre-set image, thereby improving the flexibility of the user's facial expression creation.

[0150] In the embodiments of this application, the convenience of the user using the first facial expression can be improved by saving the generated first facial expression in a local memory area and / or adding it to the facial expression panel.

[0151] In some embodiments of this application, the generation module is used to generate target content based on a first facial expression in the facial expression panel. The transmission module is used to transmit and / or publish the target content to at least one second user in response to a target content determination operation.

[0152] In the embodiments of this application, when a user edits target content, the user can edit target content with a first facial expression, publish this target content, and / or transmit it to at least one second user.

[0153] The facial expression image generation apparatus in the embodiments of this application may be an electronic device, or a component of an electronic device, such as an integrated circuit or a chip. This electronic device may be a terminal, or other device other than a terminal. Exemplary examples of electronic devices include mobile phones, tablet PCs, notebook computers, palmtop computers, in-vehicle electronic devices, mobile internet devices (MIDs), augmented reality (AR) / virtual reality (VR) devices, robots, wearable devices, ultra-mobile personal computers (UMPCs), netbooks, or personal digital assistants (PDAs), and may also include servers, network-attached storage (NAS), personal computers (PCs), televisions (TVs), cabinet machines, or self-service machines, and the embodiments of this application are not specifically limited.

[0154] The facial expression image generation apparatus in the embodiments of this application may be an apparatus having an operating system. This operating system may be the Android® operating system, the iOS operating system, or any other possible operating system, and the embodiments of this application are not specifically limited.

[0155] The facial expression image generation apparatus according to the embodiment of this application can implement each of the processes realized by the embodiment of the above method, and to avoid repetition of the explanation, it will not be explained further here.

[0156] Selectively, embodiments of the present application further provide electronic devices, and Figure 8 shows a block diagram of an electronic device according to an embodiment of the present application, as shown in Figure 8, the electronic device 800 includes a processor 802 and a memory 804, the memory 804 stores a program or instruction that can be executed on the processor 802, and when this program or instruction is executed by the processor 802, each step of the embodiment of the above method can be realized and the same technical effect can be achieved, and to avoid repetition of the explanation, it will not be explained any further here.

[0157] It should be explained that the electronic devices in the embodiments of this application include the mobile electronic devices and non-mobile electronic devices described above.

[0158] Figure 9 is a schematic diagram of the hardware structure of an electronic device that realizes an embodiment of the present application.

[0159] This electronic device 900 includes, but is not limited to, components such as a radio frequency unit 901, a network module 902, an audio output unit 903, an input unit 904, a sensor 905, a display unit 906, a user input unit 907, an interface unit 908, a memory 909, and a processor 910.

[0160] As those skilled in the art will understand, the electronic device 900 may further include power supplies (e.g., batteries) for supplying power to each component, and the power supplies may be logically connected to the processor 910 by a power management system, thereby enabling functions such as charge / discharge management and power consumption management by the power management system. The electronic device structure shown in Figure 9 does not constitute a limitation on the electronic device, and the electronic device may include more or fewer components than those shown, or combinations of some components, or different arrangements of components, which will not be described further here.

[0161] Here, the processor 910 is used to receive a first input when media content is acquired, the first input is used to trigger the generation of a first facial expression, the first facial expression is a facial expression image generated based on the media content, the facial expression image is used to edit and / or transmit target content, the media content includes captured content and / or uploaded content. The display unit 906 is used to display target information in response to a first input, where the target information relates to a first facial expression.

[0162] In the embodiments of this application, after a user uploads or photographs initial media content, the first user can respond to the first input by automatically generating a corresponding first facial expression based on the initial media content and displaying target information to prompt the user. This allows the user to easily create corresponding facial expressions based on their published personal works during the process of publishing their personal works, simplifying the user's facial expression creation steps, further improving the rate of repeated access to the user's personal works, and expanding the user-supplied pool of facial expressions.

[0163] In some embodiments of this application, the processor 910 is used to generate target media content based on initial media content in the media content generation interface after displaying target information. The processor 910 is used to announce the target media content in response to the operation to determine the target media content. and / or processor 910 is used to transmit the target media content to at least one second user.

[0164] In the embodiments of this application, after generating a first facial expression based on initial media content, the user can continue editing the initial media content to obtain target media content, and after the user performs a decision operation on the target media content, publish this target content or send this target content to other users, ensuring that the facial expression creation process does not interrupt the user's progress in uploading or sending the corresponding target media content.

[0165] In some embodiments of this application, the display unit 906 is used to present initial media content to the media content generation interface. The processor 910 is used to confirm the acceptance of a first input in response to a trigger operation on the media content generation interface.

[0166] In the embodiments of this application, the initial media content can be displayed on the media content generation interface, and the user can perform a first input to trigger the generation of a first facial expression based on the initial media content by triggering an operation on the media content generation interface, thereby enabling the user to generate the corresponding first facial expression with simple operation when the initial media content is displayed, and further simplifying the user's facial expression creation operation flow.

[0167] In some embodiments of this application, the first input includes a first sub-input and a second sub-input. The processor 910 is used to receive a first sub-input to the first control in the media content generation interface when displaying initial media content in the media content generation interface. The display unit 906 is used to display a second control on the content editing interface in response to a first sub-input. Processor 910 is used to receive a second sub-input to a second control, where the second control includes an expression addition control.

[0168] In the embodiments of this application, a user can easily generate an expression image based on initial media content by sequentially executing a first sub-input and a second sub-input for a first control and a second control in the media content generation interface, and can also add the expression image to their favorites, thereby ensuring ease of operation for the user in the media content generation interface.

[0169] In some embodiments of this application, the first facial expression includes either a dynamic facial expression image or a static facial expression image.

[0170] In the embodiments of this application, dynamic facial expression images can be generated from acquired initial media content, and static facial expression images can also be generated, improving the flexibility of users to create their own facial expressions, and allowing users to choose whether to create dynamic or static facial expressions according to their actual needs.

[0171] In some embodiments of this application, the initial media content includes a first video. Processor 910 is used to generate a first facial expression based on a first video, where the first facial expression is a dynamic facial expression image.

[0172] In the embodiments of this application, when the initial media content acquired is a video, it is possible to generate corresponding dynamic facial expressions based on the video, thereby simplifying the operation flow required when a user creates dynamic facial expressions.

[0173] In some embodiments of this application, the initial media content includes a first video. Processor 910 is used to generate a first facial expression based on a first video, where the first facial expression is a static facial expression image.

[0174] In the embodiments of this application, a target video frame is extracted from a first video, and a first facial expression is generated using the target video frame as a static facial expression image. The first facial expression generated based on the first video may be a dynamic or static facial expression, and the user can flexibly choose whether to create a dynamic or static facial expression based on the first video. This simplifies the user's facial expression creation process while simultaneously improving the user's flexibility in selecting facial expressions.

[0175] In some embodiments of this application, the processor 910 is used to determine a first dynamic image as a first facial expression when the video length of the first video is less than or equal to a preset length, and the first dynamic image is a dynamic image obtained by the transformation of the first video.

[0176] In the embodiments of this application, by comparing the video length of a first video with a preset length, and automatically performing a format conversion on the first video whose video length does not exceed the preset length, it is possible to obtain a first facial expression, ensuring that the amount of data of the obtained first facial expression is relatively small, and the user does not need to perform processing such as cutting on the manually acquired video to complete the production of the corresponding facial expression, further simplifying the user's operation steps when creating facial expressions.

[0177] In some embodiments of this application, the processor 910 is used to determine a second video in the first video when the video length of the first video is greater than a preset length, and the video length of the second video is less than or equal to the preset length. Processor 910 is used to determine the second dynamic image as the first facial expression, and the second dynamic image is the dynamic image obtained by the conversion of the second video.

[0178] In the embodiments of this application, the video length of the first video is compared with a preset length, and when the video length of the first video exceeds the preset length, the second video is cut from the first video. By determining the second dynamic image obtained by conversion based on the second video as the first facial expression, it is possible to ensure that even when the video length of the first video is too long, the amount of data obtained for the first facial expression is still relatively small.

[0179] In some embodiments of this application, the display unit 906 is used to display initial media content and a third control. Processor 910 is used to receive a second input to a third control, the third control including video cutting control, Processor 910 is used to cut out the second video from the first video in response to the second input.

[0180] In the embodiments of this application, the user can manually cut the first video by operating a third control, thereby obtaining a second video, and the content of the first facial expression generated based on the second video is manually selected by the user, ensuring that the resulting first facial expression matches the user's actual needs.

[0181] In some embodiments of this application, the processor 910 is used to generate a first facial expression based on initial media content.

[0182] In the embodiments of this application, a first facial expression is generated based on initial media content, and it is possible to ensure that different types of initial media content can generate corresponding first facial expressions, thereby allowing the user to flexibly select the type of initial media content uploaded or captured, and thereby produce different types of first facial expressions.

[0183] In some embodiments of this application, the display unit 906 is used to display target information and a fourth control, the target information includes an expression image generated based on initial media content, Processor 910 is used to receive a third input to the fourth control, and the third input is used to edit the facial expression image. Processor 910 is used to generate a first facial expression based on the edited facial expression image.

[0184] In the embodiments of this application, before storing the facial expression image generated based on the initial media content as the first facial expression, the user can edit the facial expression image using a fourth control, and the edited facial expression image can be used as the first facial expression. This allows the user to perform secondary editing on the facial expression image generated based on the initial media content, further improving the flexibility of facial expression creation.

[0185] In some embodiments of this application, the processor 910 is used to determine at least one preset image as a first facial expression, and the number of first facial expressions matches the number of preset images.

[0186] In the embodiment of this application, when the initial media content is a pre-set image, each of the pre-set images is designated as a corresponding first facial expression. This allows multiple first facial expressions to be added to the favorite facial expressions in a single step when the initial media content consists of multiple pre-set images, thereby achieving the effect of producing a large number of facial expressions.

[0187] In some embodiments of this application, the number of pre-set images is at least two. The processor 910 is used to determine the target image as the first facial expression in response to a selection operation on a target image among at least two pre-set images, where the target image is at least one of the at least two pre-set images.

[0188] In the embodiments of this application, when the initial media content includes at least two pre-set images, the at least two pre-set images are displayed, and the user can choose to generate a portion of the initial media content into the corresponding first facial expression by selecting a target image from among them through a selection operation.

[0189] In some embodiments of this application, the initial media content includes at least two pre-set images. The processor 910 is used to synthesize at least two pre-set images into a first facial expression, where the first facial expression is a dynamic facial expression image.

[0190] In the embodiments of this application, when the initial media content consists of at least two pre-set images, the at least two pre-set images can be combined into a dynamic facial expression image, and the resulting dynamic facial expression image is designated as the first facial expression. The user does not need to combine multiple pre-set images in other applications, and the combined dynamic facial expression image can be generated quickly.

[0191] In some embodiments of this application, the initial media content includes at least two pre-set images. The processor 910 is used to synthesize at least two pre-set images into a first facial expression, where the first facial expression is a static facial expression image.

[0192] In the embodiments of this application, when the initial media content consists of at least two pre-set images, the at least two pre-set images can be combined in a tile-like manner. For example, the area ratio occupied by each pre-set image in the static facial expression image obtained by combining them can be freely set by the user.

[0193] In some embodiments of this application, the pre-set image includes either a pre-set static image or a pre-set dynamic image.

[0194] In the embodiments of this application, when the initial media content is a pre-set static image or a pre-set dynamic image, a first facial expression can be generated based on this pre-set image, thereby improving the flexibility of the user's facial expression creation.

[0195] In the embodiments of this application, the convenience of the user using the first facial expression can be improved by saving the generated first facial expression in a local memory area and / or adding it to the facial expression panel.

[0196] In some embodiments of this application, the processor 910 is used to generate target content based on a first facial expression in the facial expression panel. The processor 910 is used to transmit the target content to at least one second user and / or to announce the target content in response to a target content determination operation.

[0197] In the embodiments of this application, when a user edits target content, the user can edit target content with a first facial expression, publish this target content, and / or transmit it to at least one second user.

[0198] It should be understood that, in the embodiments of this application, the input unit 904 may include a graphics processing unit (GPU) 9041 and a microphone 9042, the graphics processor 9041 processing static image or video image data obtained by an image capture device (e.g., a camera) in video capture mode or image capture mode. The display unit 906 may include a display panel 9061, which may be configured in the form of a liquid crystal display, organic light-emitting diodes, etc. The user input unit 907 includes at least one of a touch panel 9071 and other input devices 9072. The touch panel 9071 is also called a touchscreen. The touch panel 9071 may include two parts: a touch detection device and a touch controller. The other input devices 9072 may include, but are not limited to, a physical keyboard, function keys (e.g., volume control buttons, switch buttons, etc.), a trackball, a mouse, or an operating lever, and will not be described further here.

[0199] Memory 909 may be used to store software programs and various data. Memory 909 may include a first storage area that mainly stores programs or instructions and a second storage area that stores data, where the first storage area can store an operating system, an application program or instructions necessary for at least one function (e.g., audio playback function, image playback function, etc.). Memory 909 may include volatile memory or non-volatile memory, or it may include both volatile and non-volatile memory. Here, non-volatile memory may be read-only memory (ROM), programmable read-only memory (PROM), erasable programmable read-only memory (Erasable PROM, EPROM), electrically erasable programmable read-only memory (Electrically EPROM, EEPROM), or flash memory. The volatile memory may be Random Access Memory (RAM), Static Random Access Memory (Static RAM, SRAM), Dynamic Random Access Memory (DRAM), Synchronous Dynamic Random Access Memory (Synchronous DRAM, SDRAM), Double Data Rate Synchronous Dynamic Random Access Memory (Double Data Rate SDRAM, DDRSDRAM), Enhanced Synchronous Dynamic Random Access Memory (Enhanced SDRAM, ESDRAM), Synch-link Dynamic Random Access Memory (Synch-link DRAM, SLDRAM), and Direct Rambus Random Access Memory (DRRAM). The memory 909 in the embodiments of this application includes, but is not limited to, these and any other suitable types of memory.

[0200] The processor 910 may include one or more processing units. Selectively, the processor 910 integrates an application processor and a modem processor, where the application processor primarily handles operations related to the operating system, user interface, and application programs, and the modem processor primarily handles wireless communication signals, such as a baseband processor. To be clear, the above-mentioned modem processor does not have to be integrated into the processor 910.

[0201] The embodiments of this application further provide a readable storage medium in which a program or instruction is stored, and when this program or instruction is executed by a processor, each process of the embodiments of the facial expression image generation method described above can be realized and the same technical effects can be achieved. To avoid repetition of the description, no further explanation is provided here.

[0202] Here, the processor is the processor in the electronic device in the above embodiment. The readable storage medium includes computer-readable storage media such as computer read-only memory (ROM), random access memory (RAM), magnetic disk, or optical disk.

[0203] Embodiments of this application further provide a chip comprising a processor and a communication interface, the communication interface being coupled with the processor, the processor executing a program or instructions and used to implement each process of the embodiment of the facial expression image generation method described above and achieving the same technical effects, which will not be described further here in order to avoid repetition of the description.

[0204] It should be understood that the chips referred to in the embodiments of this application may also be called system-level chips, system chips, chip systems, or system-on-a-chip, etc.

[0205] The embodiments of this application provide a computer program product which is stored in a storage medium and is executed by at least one processor to realize each process of the embodiments of the facial expression image generation method described above and achieve the same technical effects. To avoid repetition of the description, no further explanation is provided here.

[0206] It should be noted that, in this specification, the terms “include,” “incorporate,” or any other variation thereof are intended to cover non-exclusive “include,” thereby including not only those elements but also other elements not explicitly listed, or elements specific to such process, method, article, or apparatus. Unless otherwise specified, an element limited by the phrase “includes one of…” is not excluded from the existence of other identical elements in a process, method, article, or apparatus containing that element. It should also be noted that the scope of methods and apparatus in embodiments of this application is not limited to performing functions in the order illustrated or discussed, but may include performing functions in a manner that is essentially simultaneous or in reverse order based on the functions involved, and methods described in a different procedure than those described, for example, may be performed, and various steps may be added, omitted, or combined. Furthermore, features described by reference to some examples may be combined with other examples.

[0207] As will be clearly evident to those skilled in the art from the above description of the embodiments, the methods of the above embodiments can be implemented in the form of software and a necessary general-purpose hardware platform. Of course, they may also be implemented in hardware, but in many cases the former is a more preferred embodiment. With this understanding in mind, the technical invention of this application may be embodied in substance or in part in the form of a computer software product, which is stored on a single storage medium (e.g., ROM / RAM, magnetic disk, optical disk), contains a certain number of instructions, thereby enabling a terminal (which may be a mobile phone, computer, server, or network device, etc.) to execute the methods of each embodiment of this application.

[0208] While embodiments of this application have been described above, accompanied by drawings, this application is not limited to the specific embodiments described above. The specific embodiments described above are merely illustrative and not restrictive. Those skilled in the art can, by the suggestion of this application, make many forms, as long as they do not deviate from the spirit and claims of this application, and all of these fall within the scope of protection of this application.

Claims

1. A method for generating facial expression images, wherein on the first user side, the method for generating facial expression images is: When initial media content is acquired, a first input is received, the first input is used to trigger the generation of a first facial expression, the first facial expression is a facial expression image generated based on the initial media content, the facial expression image is used to edit and / or transmit target content, and the initial media content includes captured content and / or uploaded content. A method for generating an expression image, comprising displaying target information in response to the first input, wherein the target information is related to the first expression.

2. The aforementioned method, After displaying the target information, the media content generation interface generates target media content based on the initial media content, A method for generating an expression image according to claim 1, further comprising, in response to the operation of determining the target media content, publishing the target media content and / or transmitting the target media content to at least one second user.

3. When the initial media content is acquired, the first input is received. The initial media content is presented to the media content generation interface, A method for generating an expression image according to claim 2, comprising confirming the receipt of a first input in response to a trigger operation on the media content generation interface.

4. The aforementioned first input includes a first sub-input and a second sub-input, and when initial media content is acquired, the first input is received. When the initial media content is displayed on the media content generation interface, the first sub-input is received for the first control in the media content generation interface. In response to the first sub-input, a second control is displayed on the content editing interface, A method for generating an expression image according to claim 3, comprising receiving a second sub-input to the second control, wherein the second control includes an expression addition control.

5. The method for generating an expression image according to any one of claims 1 to 4, wherein the first expression includes either a dynamic expression image or a static expression image.

6. The aforementioned initial media content includes the first video, Before displaying the target information mentioned above, A method for generating an expression image according to any one of claims 1 to 5, comprising generating the first expression based on the first video, wherein the first expression is a dynamic expression image.

7. Generating the first facial expression based on the first video is, The method for generating an expression image according to claim 6, comprising determining a first dynamic image as the first expression when the video length of the first video is less than or equal to a predetermined length, wherein the first dynamic image is a dynamic image obtained by the conversion of the first video.

8. Generating the first facial expression based on the first video is, If the video length of the first video is greater than a predetermined length, a second video is determined in the first video, wherein the video length of the second video is less than or equal to the predetermined length. A method for generating an expression image according to claim 6, comprising determining a second dynamic image as the first expression, wherein the second dynamic image is a dynamic image obtained by the conversion of the second video.

9. Determining the second video in the first video as described above is Displaying the initial media content and the third control, Receiving a second input to the third control, wherein the third control includes a video cutting control, A method for generating an expression image according to claim 8, comprising cutting out the second video from the first video in response to the second input.

10. The method for generating an expression image according to any one of claims 1 to 5, further comprising generating the first expression based on the initial media content.

11. The above-mentioned generation of the first facial expression based on the initial media content is, Displaying the target information and the fourth control, wherein the target information includes an expression image generated based on the initial media content, The fourth control receives a third input, the third input being used to edit the facial expression image, A method for generating an expression image according to claim 10, comprising generating the first expression based on the edited expression image.

12. The initial media content includes at least one pre-set image, Before displaying the target information mentioned above, A method for generating facial expression images according to any one of claims 1 to 5, comprising determining at least one of the pre-set images as the first facial expression, wherein the number of the first facial expressions matches the number of the pre-set images.

13. The number of pre-set images is at least two. Determining at least one of the aforementioned pre-set images as the first facial expression is, A method for generating an expression image according to claim 12, comprising determining the target image as the first expression in response to a selection operation for a target image among at least two of the pre-set images, wherein the target image is at least one of the at least two of the pre-set images.

14. The initial media content includes at least two pre-set images, Before displaying the target information mentioned above, A method for generating an expression image according to any one of claims 1 to 5, comprising combining the at least two pre-set images with the first expression, wherein the first expression is a dynamic expression image.

15. The aforementioned pre-set image is A method for generating an expression image according to any one of claims 12 to 14, comprising either a pre-set static image or a pre-set dynamic image.

16. After displaying the target information in response to the aforementioned first input, The first facial expression is stored in a local memory area, and / or A method for generating an expression image according to any one of claims 1 to 14, comprising adding the first expression to an expression panel, wherein the expression panel is used to generate the target content.

17. The target content includes the facial expression image, and the method is The target content is generated based on the first facial expression in the facial expression panel, A method for generating an expression image according to claim 16, further comprising transmitting the target content to at least one second user and / or publishing the target content in response to a target content determination operation.

18. A facial expression image generation device, wherein on the first user side, the facial expression image generation device is A receiving module for receiving a first input when initial media content is acquired, wherein the first input is used to trigger the generation of a first facial expression, the first facial expression is a facial expression image generated based on the initial media content, the facial expression image is used to edit and / or transmit target content, and the initial media content includes captured content and / or uploaded content. A facial expression image generating apparatus, comprising a display module for displaying target information in response to the first input, wherein the target information relates to the first facial expression.

19. Memory in which a program or instruction is stored, An electronic device comprising a processor for realizing the steps of the facial expression image generation method described in any one of claims 1 to 17 when executing the program or instruction.

20. A readable storage medium that stores a program or instruction that, when executed by a processor, implements the steps of the facial expression image generation method described in any one of claims 1 to 17.

21. A computer program product comprising a computer program that, when executed by a coprocessor, implements the steps of the method for generating facial expression images described in any one of claims 1 to 17.

22. A computer program that, when executed by a processor, implements the steps of the method for generating facial expression images according to any one of claims 1 to 17.