Emoji generation method and apparatus, work posting method and apparatus, and device and storage medium
By generating multi-themed emoticon resources based on target images, the problem of time-consuming and difficult to personalize traditional emoticon resources is solved, efficient and customized expression generation is achieved, and the efficiency of information interaction is improved.
Patent Information
- Application Number
- PCT/CN2024/138610
- Authority / Receiving Office
- WO · WO
- Patent Type
- Applications
- Current Assignee / Owner
- Priority Date
- 2023-12-19
- Filing Date
- 2024-12-11
- Publication Date
- 2025-06-26
AI Technical Summary
The production of traditional expression resources relies on professional teams, and it is time-consuming and difficult to provide personalized expressions, which cannot meet users' needs to express information conveniently.
A method of emoticon generation is provided, by receiving expression generation requests associated with the target image, generating expression resources based on the target image, including expression resources of multiple preset themes, associated with a specific style.
It improves the efficiency of expression resources generation, can quickly generate expression resources for different themes associated with specific styles, improves the efficiency of information interaction, and provides higher quality customized expressions.
Smart Images

Figure CN2024138610_26062025_PF_FP_ABST
Abstract
Description
Method, device, equipment and storage medium for generating expressions and publishing works
[0001] This application claims priority to the Chinese invention patent application entitled “Methods, devices, equipment and storage media for expression generation and work publication” filed on December 19, 2023, with application number 202311754645.X. The entire contents of that application are incorporated by reference into this application. Technical Field
[0002] Example embodiments of the present disclosure generally relate to the field of computers, and more particularly, to methods, devices, apparatuses, and computer-readable storage media for generating and publishing expressions. Background Art
[0003] With the development of computer technology, the internet has become a vital platform for information exchange. Throughout this process, various emojis have become a crucial medium for social expression and information exchange. High-quality emojis can help users more easily and effectively convey their desired messages. Summary of the Invention
[0004] In a first aspect of the present disclosure, a method for generating an expression is provided. The method includes: receiving an expression generation request associated with a target image; and providing a first set of expression resources generated based on the target image, wherein the set of expression resources is associated with a first style and includes multiple expression resources corresponding to multiple preset themes.
[0005] In a second aspect of the present disclosure, a method for publishing a work is provided. The method includes: presenting a viewing interface for a first work, the first work comprising an image work or a video work; presenting a set of expression resources generated based on the first work upon selection of a third entry in the viewing interface; and publishing a second work corresponding to the at least one expression resource upon selection of at least one expression resource in the set of expression resources.
[0006] In a third aspect of the present disclosure, an expression generation device is provided. The device includes: a request receiving module configured to receive an expression generation request associated with a target image; and an expression providing module configured to provide a first set of expression resources generated based on the target image, wherein the set of expression resources is associated with a first style and includes multiple expression resources corresponding to multiple preset themes.
[0007] In a fourth aspect of the present disclosure, a work publishing device is provided. The device includes: an interface presentation module configured to present a viewing interface for a first work, where the first work includes a picture work or a video work; an expression presentation module configured to present a set of expression resources generated based on the first work, based on a selection of a third entry in the viewing interface; and a work publishing module configured to publish a second work corresponding to at least one expression resource in the set of expression resources, based on a selection of at least one expression resource.
[0008] In a fifth aspect of the present disclosure, an electronic device is provided. The device includes at least one processing unit; and at least one memory coupled to the at least one processing unit and storing instructions for execution by the at least one processing unit. When executed by the at least one processing unit, the instructions cause the device to perform the method of the first aspect or the second aspect.
[0009] In a sixth aspect of the present disclosure, a computer-readable storage medium is provided, wherein a computer program is stored on the computer-readable storage medium, and the computer program can be executed by a processor to implement the method of the first aspect or the second aspect.
[0010] It should be understood that the content described in this summary section is not intended to limit the key features or important features of the embodiments of the present disclosure, nor is it intended to limit the scope of the present disclosure. Other features of the present disclosure will become easily understood through the following description. BRIEF DESCRIPTION OF THE DRAWINGS
[0011] The above and other features, advantages and aspects of the embodiments of the present disclosure will become more apparent with reference to the following detailed description in conjunction with the accompanying drawings. In the accompanying drawings, the same or similar reference numerals represent the same or similar elements, wherein:
[0012] FIG1 shows a schematic diagram of an example environment in which embodiments according to the present disclosure may be implemented;
[0013] 2A to 2D illustrate example interfaces according to some embodiments of the present disclosure;
[0014] 3A and 3B illustrate example interfaces according to further embodiments of the present disclosure;
[0015] FIG4A shows a flowchart of an example process of emoticon generation according to some embodiments of the present disclosure;
[0016] FIG4B shows a flowchart of an example process of work publication according to some embodiments of the present disclosure;
[0017] FIG5A shows a schematic structural block diagram of an example expression generating apparatus according to some embodiments of the present disclosure;
[0018] FIG5B shows a schematic structural block diagram of an exemplary work publishing apparatus according to some embodiments of the present disclosure; and
[0019] FIG6 illustrates a block diagram of an electronic device capable of implementing various embodiments of the present disclosure. DETAILED DESCRIPTION
[0020] The following describes embodiments of the present disclosure in more detail with reference to the accompanying drawings. Although certain embodiments of the present disclosure are shown in the accompanying drawings, it should be understood that the present disclosure can be implemented in various forms and should not be construed as limited to the embodiments described herein. Rather, these embodiments are provided to provide a more thorough and complete understanding of the present disclosure. It should be understood that the drawings and embodiments of the present disclosure are for illustrative purposes only and are not intended to limit the scope of protection of the present disclosure.
[0021] It should be noted that the titles of any section / subsection provided herein are not limiting. Various embodiments are described throughout this document, and any type of embodiment may be included under any section / subsection. Furthermore, the embodiments described in any section / subsection may be combined in any manner with any other embodiments described in the same section / subsection and / or in different sections / subsections.
[0022] In the description of the embodiments of the present disclosure, the term "including" and similar terms should be understood as open inclusion, that is, "including but not limited to". The term "based on" should be understood as "based at least in part on". The term "one embodiment" or "the embodiment" should be understood as "at least one embodiment". The term "some embodiments" should be understood as "at least some embodiments". Other explicit and implicit definitions may be included below. The terms "first", "second", etc. may refer to different or the same objects. Other explicit and implicit definitions may be included below.
[0023] The embodiments of the present disclosure may involve user data, data acquisition and / or use, etc. These aspects shall comply with the corresponding laws, regulations and relevant provisions. In the embodiments of the present disclosure, all data collection, acquisition, processing, processing, forwarding, use, etc. are carried out on the premise that the user is aware of and confirms them. Accordingly, when implementing the various embodiments of the present disclosure, the types, scope of use, and usage scenarios of the data or information that may be involved should be informed to the user and the user's authorization should be obtained in an appropriate manner in accordance with the relevant laws and regulations. The specific notification and / or authorization method may vary according to the actual situation and application scenario, and the scope of the present disclosure is not limited in this respect.
[0024] If this specification and the solutions in the examples involve the processing of personal information, such processing will be done only with a legitimate basis (such as with the consent of the subject of personal information or as necessary for the performance of a contract) and only within the prescribed or agreed scope. A user's refusal to process personal information other than that required for basic functions will not affect the user's use of basic functions.
[0025] When people interact with each other online, they expect to use high-quality emoticons to conveniently express their desired messages. Traditional emoticon resources rely on professional teams to produce them, which is time-consuming and difficult to provide personalized emoticons for users.
[0026] Embodiments of the present disclosure provide an expression generation scheme. According to the scheme, an expression generation request associated with a target image may be received. Furthermore, a first set of expression resources generated based on the target image may be provided, wherein the set of expression resources is associated with a first style and includes multiple expression resources corresponding to multiple preset themes.
[0027] In this way, the embodiments of the present disclosure can efficiently generate expression resources of different themes associated with a specific style, thereby improving the generation efficiency of expression resources.
[0028] Various example implementations of this solution are described in detail below in conjunction with the accompanying drawings.
[0029] Sample Environment
[0030] FIG1 shows a schematic diagram of an example environment 100 in which embodiments of the present disclosure can be implemented. As shown in FIG1 , the example environment 100 may include an electronic device 110 .
[0031] In this example environment 100, electronic device 110 may run an application 120 that supports interface interaction. Application 120 may be any suitable type of application for interface interaction, examples of which may include, but are not limited to, video applications, social applications, or other suitable applications. User 140 may interact with application 120 via electronic device 110 and / or its attached devices.
[0032] In the environment 100 of FIG. 1 , if the application 120 is in an active state, the electronic device 110 may present an interface 150 for supporting interface interaction through the application 120 .
[0033] In some embodiments, the electronic device 110 communicates with the server 130 to enable the provision of services for the application 120. The electronic device 110 can be any type of mobile terminal, fixed terminal or portable terminal, including a mobile phone, a desktop computer, a laptop computer, a notebook computer, a netbook computer, a tablet computer, a media computer, a multimedia tablet, a handheld computer, a portable game terminal, a VR / AR device, a personal communication system (PCS) device, a personal navigation device, a personal digital assistant (PDA), an audio / video player, a digital camera / camcorder, a positioning device, a television receiver, a radio broadcast receiver, an e-book device, a gaming device or any combination thereof, including accessories and peripherals of these devices or any combination thereof. In some embodiments, the electronic device 110 can also support any type of interface for the user (such as a "wearable" circuit, etc.).
[0034] The server 130 may be a standalone physical server, a server cluster or distributed system consisting of multiple physical servers, or a cloud server that provides basic cloud computing services such as cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communications, middleware services, domain name services, security services, content distribution networks, and big data and artificial intelligence platforms. For example, the server 130 may include a computing system / server such as a mainframe, an edge computing node, a computing device in a cloud environment, and the like. The server 130 may provide background services for the application 120 that supports virtual scenes in the electronic device 110.
[0035] A communication connection may be established between the server 130 and the electronic device 110. The communication connection may be established in a wired or wireless manner. The communication connection may include, but is not limited to, a Bluetooth connection, a mobile network connection, a Universal Serial Bus (USB) connection, a Wireless Fidelity (WiFi) connection, etc., and the embodiments of the present disclosure are not limited in this respect. In the embodiments of the present disclosure, the server 130 and the electronic device 110 may implement signaling interaction through the communication connection between the two.
[0036] It should be understood that the structure and function of the various elements in the environment 100 are described for illustrative purposes only and do not imply any limitation on the scope of the present disclosure.
[0037] Some example embodiments of the present disclosure will be described below with continued reference to the accompanying drawings.
[0038] Example Interaction
[0039] The following describes an example process for generating emojis according to embodiments of the present disclosure in conjunction with Figures 2A to 2D. Figures 2A to 2D illustrate example interfaces 200A to 200D according to some embodiments of the present disclosure. Interfaces 200A to 200D may be provided by the electronic device 110 shown in Figure 1.
[0040] 2A , the interface 200A may be a conversation interface, which may provide, for example, an input field 205 . For example, the user may utilize the input field 205 to input text content, voice content, expression content, image content, etc. into the conversation.
[0041] As an example, the electronic device 110 may further present a panel (also referred to as an input control) at the bottom of the interface 200A, which may provide an entry 210 for generating an expression.
[0042] After receiving a selection for entry 210, electronic device 110 may present a set of candidate styles in the panel, for example, candidate style 215-1, candidate style 215-2, and candidate style 215-3 (individually or collectively referred to as candidate styles 215). In some embodiments, the set of candidate styles 215 may include one or more preset styles.
[0043] In some embodiments, such candidate styles 215 may correspond to different styles of expressions, for example, different picture styles (e.g., cartoon style, anthropomorphic style, etc.). As an example, the electronic device 110 may also present a corresponding preview image associated with each candidate style 215 to intuitively represent its corresponding style.
[0044] In some scenarios, such candidate styles 215 may also correspond to different expression themes. For example, style 1 may correspond to animation A, and style 2 may correspond to film and television work B.
[0045] For example, after receiving a selection of the candidate style 215-1 (also referred to as the first style 215-1), the electronic device 110 may present an interface 200B as shown in FIG2B . The interface 200B may also be referred to as an image acquisition interface or a shooting interface.
[0046] 2B , the electronic device 110 may provide a viewfinder in the interface 200B for capturing an image using a camera of the electronic device 110. As an example, the electronic device 110 may capture a target image 220 for expression generation through a camera.
[0047] In some embodiments, the electronic device 110 may further provide, for example, instruction information in the viewfinder to guide the user to capture the target image 220 corresponding to the preset part. For example, the electronic device 110 may display a viewfinder area represented by a dotted line in the viewfinder, where the viewfinder area may have a shape corresponding to the preset part (e.g., face, hand, etc.), to guide the user to capture an image corresponding to the preset part (e.g., facial image, hand image, etc.).
[0048] In some embodiments, such indication information may be determined based on the selected first style 215 - 1 . As an example, different styles may correspond to different framing guidance information, for example, to guide the user to photograph different preset parts.
[0049] In some embodiments, when the content currently viewed by the user does not match the instruction information, the electronic device 110 may further provide text information to guide the user to capture the target image 220 corresponding to the preset part.
[0050] Furthermore, upon receiving a trigger for the capture control 225, the electronic device 110 may capture the corresponding target image 220 for use in generating the expression resource. As an example, the electronic device 110 may also capture using different cameras of the electronic device 110 based on a trigger for the switch control 235.
[0051] As shown in FIG2B , electronic device 110 may further provide an upload control 230 in interface 200B. Electronic device 110 may obtain a target image uploaded by the user based on the user's selection of upload control 230. For example, with the user's knowledge and authorization, electronic device 110 may receive the user's selection of a local image as the target image.
[0052] Additionally, as shown in FIG2B , the electronic device 110 may also present a set of candidate styles 215 in the interface 200B, wherein the first style 215-1 may be selected, for example. The user may also switch to another candidate style 215 by, for example, clicking on the other candidate styles 215, or the user may switch to a different style by performing a sliding operation in a preset direction (e.g., sliding left) in the interface 200B.
[0053] Although the above example triggers presentation of the interface 200B by selecting a style in FIG. 2A , it should be understood that the electronic device 110 may also present the interface 200B as shown in FIG. 2B based on other appropriate operations.
[0054] For example, the electronic device 110 may present an interface 200B (also referred to as a shooting interface or a work creation interface) based on a current user's shooting request or work creation request. The electronic device 110 may, for example, initially present the set of candidate styles 215 in the interface 200B, or may present the set of candidate styles 215 based on a preset operation of the user.
[0055] Furthermore, after the user triggers an expression generation request by photographing or uploading a target image, the electronic device 110 may accordingly present an interface 200C as shown in FIG2C . The interface 200C may also be referred to as an expression preview interface.
[0056] As shown in FIG2C , electronic device 110 may provide a set of expression resources generated based on acquired target image 220, such as expression resource 240-1, expression resource 240-2, expression resource 240-3, and expression resource 240-4 (individually or collectively referred to as expression resources 240). Such expression resources 240 may include, but are not limited to, static images, dynamic images, videos, three-dimensional models, and the like.
[0057] In some embodiments, the set of expression resources 240 may include multiple expression resources corresponding to multiple preset themes. Such preset themes may correspond to the content of the expression resources, for example.
[0058] As an example, the emoticon resource 240 - 1 , the emoticon resource 240 - 2 , the emoticon resource 240 - 3 , and the emoticon resource 240 - 4 may correspond to different holiday greeting themes.
[0059] In some embodiments, the expression resource 240 may further include corresponding text content. Additionally, the electronic device 110 may further receive a user's editing operation on the text content and update the text content in the expression resource 240 accordingly.
[0060] In some embodiments, such text content can be used to express the theme corresponding to the expression resource 240. Alternatively, such text content can also be used as an identifier of the expression resource 240.
[0061] In some embodiments, the set of expression resources 240 may be generated by a target model based on the received target image 220. It should be understood that the target model may be implemented based on any appropriate machine learning technology, examples of which may include but are not limited to image generation models. Furthermore, the target model may be deployed locally on the electronic device 110 or on another appropriate remote device, and may be accessed or called by the electronic device 110 through appropriate means.
[0062] In some embodiments, multiple expression resources 240 can be generated by providing a plurality of different preset guide words to the target model. For example, expression resource 240-1 can be generated based on a first guide word corresponding to a first theme, and the first guide word can be used to guide the target model to generate media content corresponding to a first style indicated by the first guide word based on the target image as an expression resource. Conversely, expression resource 240-2 can be generated based on a second guide word corresponding to a second theme, and the second guide word can be used to guide the target model to generate media content corresponding to a second style indicated by the second guide word based on the target image as an expression resource.
[0063] In some embodiments, the set of expression resources 240 generated by the target model may have visual content corresponding to the received target image 220. For example, when the target image 220 includes a specific part (such as a face or a hand), the set of expression resources 240 may have visual content corresponding to the specific part (such as different facial expressions or different hand movements).
[0064] In some embodiments, the electronic device 110 may also provide a corresponding update entry associated with a single emoticon resource 240, such as an update icon in the upper left corner of the emoticon resource 240-1. Upon selecting the update entry, the electronic device 110 may trigger the regeneration of another emoticon resource corresponding to the theme, and may display the other emoticon resource to replace the original emoticon resource 240-1.
[0065] In some embodiments, the electronic device 110 may further provide an entry (not shown) for triggering the regeneration of the entire set of expression resources 240. For example, if the user is not satisfied with the set of expression resources 240, the electronic device 110 may provide a new set of regenerated expression resources based on the user's triggering of the entry.
[0066] As an example, the electronic device 110 may trigger the regeneration and provision of expression resources based on a user's sliding operation on the interface 200C (eg, when the sliding amplitude is greater than a threshold).
[0067] As another example, the electronic device 110 may also present emoticon resources corresponding to other themes based on a user's upward swipe operation on the interface 200C. Alternatively, the electronic device 110 may also provide for the regeneration and provision of emoticon resources that were hidden due to the upward swipe operation on the interface 200C.
[0068] In some embodiments, the electronic device 110 may also provide a second set of expression resources corresponding to another style (e.g., a second style) based on a preset operation of the user. Similar to the first set of expression resources 240, the second set of expression resources may include multiple expression resources generated based on the target image 220 and corresponding to multiple preset themes.
[0069] For example, electronic device 110 may receive a left swipe operation from the user on interface 200C, and display a second set of emoticon resources corresponding to the second style in interface 200C. Alternatively, electronic device 110 may provide a style selection control similar to interface 200B to present the second set of emoticon resources based on the user's selection of the second style.
[0070] 2C , the electronic device 110 may further provide a publish control 245. The electronic device 110 may, for example, receive a user's selection of one or more target expression resources in the set of expression resources 240 and publish the work with the selected one or more target expression resources based on the selection of the publish control 245.
[0071] As an example, the published works may include picture works, atlas works, video works, etc. generated by one or more expression resources.
[0072] As another example, when the interface 200C is also associated with providing one or more groups of expression resources corresponding to other styles, the electronic device 110 can, for example, receive the user's selection of multiple target expression resources in the multiple groups of expression resources, and publish the corresponding works based on the selection of the publishing control 245.
[0073] 2C , the electronic device 110 may further provide an add control 250. Similarly, the electronic device 110 may receive a user's selection of one or more target expression resources in the set of expression resources 240, and add the one or more target expression resources to the expression resource library for the user based on the selection of the add entry 250.
[0074] As an example, the expression resource library may include a user's local expression library or a cloud expression library, so as to support the user to interact by selecting an added expression resource from the expression resource library.
[0075] As another example, when the interface 200C is also associated with providing one or more groups of expression resources corresponding to other styles, the electronic device 110 can, for example, receive the user's selection of multiple target expression resources from the multiple groups of expression resources, and add the corresponding expression resources to the expression resource library based on the selection of the add control 250.
[0076] As an example, as shown in Figure 2D, when the user adds expression resources 240-1 to expression resources 240-4 to the expression resource library, the electronic device 110 can present the expressions added to the expression resource library, for example, expression resources 240-1 to expression resources 240-4, based on the user's expression input request.
[0077] Furthermore, based on the user's selection of the expression resource 240-1, the electronic device 110 may send the expression resource 240-1 to the current input environment, for example, a conversation with "User B." It should be understood that such an input environment is merely exemplary, and embodiments of the present disclosure may support any other input environment suited for expression input, such as document editing, comment input, or editing of media content.
[0078] Furthermore, electronic device 110 may also provide a creation portal 255 to support users in capturing or uploading images to create other new emoticons. For example, electronic device 110 may present interface 200B discussed with reference to FIG. 2B upon selecting creation portal 225. The emoticon generation process can be referred to above and will not be further elaborated here.
[0079] Based on the process discussed above, embodiments of the present disclosure can enable users to automatically create emoticons of corresponding styles and themes by uploading or taking images, thereby improving the efficiency of emoticon generation. In addition, by generating and providing such customized emoticons, embodiments of the present disclosure further improve the quality of generated emoticons, thereby helping to improve the efficiency of information interaction.
[0080] 3A and 3B illustrate example interfaces 300A and 300B according to some embodiments of the present disclosure. Interfaces 300A and 300B may be provided by the electronic device 110 shown in FIG1 .
[0081] As shown in Figure 3A, interface 300A may be, for example, a viewing interface for a work 310. The work 310 may include, for example, an image work (eg, a single image work, a collection of images) or a video work.
[0082] Furthermore, the electronic device 110 may present the panel 305 based on a preset operation of the user (eg, a sharing operation). In the panel 305 , the electronic device 110 may provide an entry 315 for generating an expression based on the work 310 .
[0083] In some embodiments, upon receiving a selection of the entry 315, the electronic device 110 may present an interface 300B as shown in Figure 3B. As shown in Figure 3B, the electronic device 110 may provide an expression preview panel 325 in the interface 300B.
[0084] In some embodiments, electronic device 110 may provide a set of expression resources generated based on work 310 in expression preview panel 325, such as expression resource 320-1, expression resource 320-2, expression resource 320-3, and expression resource 320-4 (individually or collectively referred to as expression resources 320). Such expression resources 320 may include, but are not limited to, static images, dynamic images, videos, 3D models, etc.
[0085] In some embodiments, the set of expression resources 320 may include expression resources extracted from the screen of the work 310. As an example, the electronic device 110 and / or other appropriate devices may analyze the screen content of the work 310 and extract static images, dynamic images, videos, etc. as the expression resources 320.
[0086] Taking work 310 as a video work as an example, the electronic device 110 and / or other appropriate devices can analyze the pictures of the video work and extract the corresponding video clips based on a preset time window, and can crop the pictures of the video clips to generate corresponding expression resources 320.
[0087] For example, the electronic device 110 and / or other appropriate devices may select segments in which characters' facial expressions change significantly from a video work, and crop the corresponding areas of the characters' faces as corresponding expression resources 320 .
[0088] In some other embodiments, the set of expression resources 325 may be expression resources generated by a target model based on target screen content in the work 310 .
[0089] As an example, the work 110 may be an image work, and the electronic device 110 may determine the image work itself or specific picture content in the image work (for example, picture content corresponding to a specific part) as the target picture content, and provide it to the target model for the generation of the expression resource 320.
[0090] As another example, the work 110 may be a video work or an atlas work, and the electronic device 110 may determine one or more specific images in the video work or atlas work (for example, image frames in a video or a single picture in an atlas) as target screen content, or may determine one or more specific screen parts in the video work or atlas work (for example, screen parts corresponding to specific parts of an object) as target screen content and provide them to the target model for generating expression resources 320.
[0091] In some embodiments, the electronic device 110 may also provide the user with a set of candidate screen contents determined based on the work, and may receive a selection of target screen content from the set of candidate screen contents for use in generating the expression resource 320 .
[0092] Regarding the generation process of each expression resource 320 in the expression preview panel 325 and the corresponding interaction process, reference may be made to the contents described above with reference to FIG. 2 c , and the present disclosure will not elaborate on them here.
[0093] In some embodiments, when the set of expression resources 320 includes expression resources extracted from the image of the work 310, the electronic device 110 may also trigger, based on the user's selection of an expression resource (e.g., expression resource 320-1), the generation of a new set of expression resources based on the expression resource 320-1. For example, the expression resource 320-1 may be used with the target image described above to generate a new set of expression resources.
[0094] Continuing with FIG3B , in some embodiments, electronic device 110 may further provide, for example, a publish control 330. Similar to the description above with reference to FIG2c , electronic device 110 may publish a work corresponding to one or more selected expression resources 320 based on a selection of publish control 330. For example, the published work may include a picture work, an atlas work, a video work, and the like generated from one or more expression resources.
[0095] In some embodiments, the electronic device 110 may further provide an add control 335. Similar to the above description with reference to FIG2C, the electronic device 110 may add the selected one or more emoticon resources to the user's emoticon resource library based on the selection of the add control 335.
[0096] Based on the process described above, the embodiments of the present disclosure can support users to create corresponding expressions based on published works, and can publish new works based on the generated expressions, thereby improving the efficiency of user work creation.
[0097] Example Process
[0098] FIG4A shows a flow chart of an example expression generation process 400A according to some embodiments of the present disclosure. The process 400A may be implemented at the electronic device 110. The process 400A is described below with reference to FIG1.
[0099] As shown in FIG. 4A , at block 410 , the electronic device 110 receives an expression generation request associated with a target image.
[0100] In block 420 , the electronic device 110 provides a first set of expression resources generated based on the target image, wherein the set of expression resources is associated with a first style, and the first set of expression resources includes multiple expression resources corresponding to multiple preset themes.
[0101] In some embodiments, receiving an expression generation request associated with a target image includes: presenting a session interface, the session interface including an input control; receiving a selection of a first style in the input control, presenting an image acquisition interface; and in response to acquiring the target image via the image acquisition interface, receiving an expression generation request associated with the target image.
[0102] In some embodiments, the target image includes: an image captured via an image acquisition interface, or an image uploaded via an image acquisition interface.
[0103] In some embodiments, receiving the expression generation request associated with the target image includes: presenting a work creation interface; and in response to acquiring the target image via the work creation interface, receiving the expression generation request associated with the target image.
[0104] In some embodiments, receiving an expression generation request associated with a target image includes: presenting a viewing interface of a first work; and receiving an expression generation request associated with the target image based on a selection of a first entry in the viewing interface, wherein the target image is determined based on the first work.
[0105] In some embodiments, the first work is an image work including at least one image, and the target image is determined based on the at least one image; or the first work is a video work, and the target image is determined based on at least one video frame of the video work.
[0106] In some embodiments, process 400A further includes receiving a selection of a first style from a set of preset styles.
[0107] In some embodiments, process 400A further includes: presenting a shooting interface for shooting a target image, the shooting interface including instruction information, and the instruction information is used to guide the user to shoot a target image corresponding to a preset part.
[0108] In some embodiments, the shooting interface further includes a style selection component for receiving a user's selection of a first style from a group of preset styles.
[0109] In some embodiments, the first group of expression resources includes at least a first expression resource, and the first expression resource includes text content.
[0110] In some embodiments, the process 400A further includes: updating the text content in the first emoticon resource based on an editing operation on the text content.
[0111] In some embodiments, process 400A further includes: providing an update entry associated with a second expression resource in the first set of expression resources; and updating the second expression resource based on a selection of the update entry, wherein the updated second expression resource is an expression resource regenerated based on the target image.
[0112] In some embodiments, the first set of expression resources is presented in the expression preview interface, and process 400A further includes: based on a preset operation for the expression preview interface, providing a second set of expression resources corresponding to the second style in the expression preview interface, wherein the second set of expression resources is generated based on the target image.
[0113] In some embodiments, process 400A further includes: receiving a first selection for a first group of target expression resources in the first group of expression resources; and based on the first selection, adding the first group of target expression resources to the expression resource library for the current user or publishing works corresponding to the first group of target expression resources.
[0114] In some embodiments, adding a first group of target expression resources to the expression library or publishing works corresponding to the first group of target expression resources based on the first selection includes: receiving a second selection for a second group of target expression resources in a third group of expression resources generated based on the target image; and adding the first group of target expressions and the second group of target expressions to the expression library or publishing works corresponding to the first group of target expression resources and the second group of target expression resources.
[0115] In some embodiments, the plurality of expression resources included in the first group of expression resources are generated by providing a plurality of preset guide words to the target model, where the plurality of preset guide words correspond to a plurality of preset themes.
[0116] FIG4B shows a flowchart of an example work publishing process 400B according to some embodiments of the present disclosure. Process 400B may be implemented at electronic device 110. Process 400B is described below with reference to FIG1.
[0117] As shown in FIG. 4B , in block 430 , the electronic device 110 presents a viewing interface of a first work, where the first work includes a picture work or a video work.
[0118] In block 440 , the electronic device 110 presents a set of expression resources generated based on the first work based on the selection of the third entry in the viewing interface.
[0119] In block 450 , the electronic device 110 publishes a second work corresponding to the at least one expression resource based on the selection of the at least one expression resource from the set of expression resources.
[0120] In some embodiments, a set of expression resources includes at least one of the following: expression resources extracted from a screen of a first work; and expression resources generated by a target model based on target screen content in the first work.
[0121] In some embodiments, the process 400B further includes: presenting a set of candidate screen contents in the first work; and receiving a selection of target screen content from the set of candidate screen contents.
[0122] Example devices and equipment
[0123] Embodiments of the present disclosure also provide corresponding apparatuses for implementing the above-described methods or processes. FIG5A illustrates a schematic block diagram of an example expression generation apparatus 500A according to certain embodiments of the present disclosure. Apparatus 500A may be implemented as or included in electronic device 110. Each module / component in apparatus 500A may be implemented in hardware, software, firmware, or any combination thereof.
[0124] As shown in Figure 5A, the device 500A includes a request receiving module 510, which is configured to receive an expression generation request associated with a target image; and an expression providing module 520, which is configured to provide a first group of expression resources generated based on the target image, wherein a group of expression resources is associated with a first style, and the first group of expression resources includes multiple expression resources corresponding to multiple preset themes.
[0125] In some embodiments, the request receiving module 510 is further configured to: present a conversation interface, the conversation interface including an input control; receive a selection of a first style in the input control, present an image acquisition interface; and in response to acquiring a target image via the image acquisition interface, receive an expression generation request associated with the target image.
[0126] In some embodiments, the target image includes: an image captured via an image acquisition interface, or an image uploaded via an image acquisition interface.
[0127] In some embodiments, the request receiving module 510 is further configured to: present a work creation interface; and in response to obtaining a target image via the work creation interface, receive an expression generation request associated with the target image.
[0128] In some embodiments, the request receiving module 510 is further configured to: present a viewing interface of the first work; and based on a selection of a first entry in the viewing interface, receive an expression generation request associated with a target image, wherein the target image is determined based on the first work.
[0129] In some embodiments, the first work is an image work including at least one image, and the target image is determined based on the at least one image; or the first work is a video work, and the target image is determined based on at least one video frame of the video work.
[0130] In some embodiments, the apparatus 500A further includes a style selection module configured to receive a selection of a first style from a set of preset styles.
[0131] In some embodiments, the device 500A further includes an image capturing module configured to present a capturing interface for capturing a target image, wherein the capturing interface includes instruction information for guiding a user to capture a target image corresponding to a preset part.
[0132] In some embodiments, the shooting interface further includes a style selection component for receiving a user's selection of a first style from a group of preset styles.
[0133] In some embodiments, the first group of expression resources includes at least a first expression resource, and the first expression resource includes text content.
[0134] In some embodiments, the apparatus 500A further includes a text editing module configured to update the text content in the first expression resource based on an editing operation on the text content.
[0135] In some embodiments, the device 500A also includes a regeneration module configured to: provide an update entry in association with a second expression resource in the first group of expression resources; and update the second expression resource based on a selection of the update entry, wherein the updated second expression resource is an expression resource regenerated based on the target image.
[0136] In some embodiments, a first group of expression resources is presented in an expression preview interface, and the device 500A also includes an expression preview module, which is configured to: based on a preset operation for the expression preview interface, provide a second group of expression resources corresponding to the second style in the expression preview interface, and the second group of expression resources is generated based on the target image.
[0137] In some embodiments, the device 500A also includes a first publishing module, which is configured to: receive a first selection for a first group of target expression resources in the first group of expression resources; and based on the first selection, add the first group of target expression resources to the expression resource library for the current user or publish works corresponding to the first group of target expression resources.
[0138] In some embodiments, the first publishing module is further configured to: receive a second selection of a second group of target expression resources in a third group of expression resources generated based on the target image; and add the first group of target expressions and the second group of target expressions to the expression library or publish works corresponding to the first group of target expression resources and the second group of target expression resources.
[0139] In some embodiments, the plurality of expression resources included in the first group of expression resources are generated by providing a plurality of preset guide words to the target model, where the plurality of preset guide words correspond to a plurality of preset themes.
[0140] Embodiments of the present disclosure also provide corresponding apparatuses for implementing the aforementioned methods or processes. Figure 5B shows a schematic structural block diagram of a work publishing apparatus 500B according to certain embodiments of the present disclosure. Apparatus 500B may be implemented as or included in electronic device 110. Each module / component in apparatus 500B may be implemented using hardware, software, firmware, or any combination thereof.
[0141] As shown in Figure 5B, the device 500B includes an interface presentation module 530, which is configured to present a viewing interface of a first work, where the first work includes a picture work or a video work; an expression presentation module 540, which is configured to present a group of expression resources generated based on the first work based on a selection of a third entry in the viewing interface; and a work publishing module, which is configured to: based on a selection of at least one expression resource in a group of expression resources, publish a second work corresponding to at least one expression resource.
[0142] In some embodiments, a set of expression resources includes at least one of the following: expression resources extracted from a screen of a first work; and expression resources generated by a target model based on target screen content in the first work.
[0143] In some embodiments, the apparatus 500B further includes a content selection module configured to: present a set of candidate screen contents in the first work; and receive a selection of target screen content from the set of candidate screen contents.
[0144] FIG6 shows a block diagram of an electronic device 600 in which one or more embodiments of the present disclosure may be implemented. It should be understood that the electronic device 600 shown in FIG6 is merely exemplary and should not be construed as limiting the functionality and scope of the embodiments described herein. The electronic device 600 shown in FIG6 can be used to implement the electronic device 110 of FIG1 .
[0145] As shown in FIG6 , electronic device 600 is a general-purpose electronic device. Components of electronic device 600 may include, but are not limited to, one or more processors or processing units 610, memory 620, storage device 630, one or more communication units 640, one or more input devices 650, and one or more output devices 660. Processing unit 610 may be a real or virtual processor and is capable of performing various processes according to programs stored in memory 620. In a multi-processor system, multiple processing units execute computer-executable instructions in parallel to enhance the parallel processing capabilities of electronic device 600.
[0146] The electronic device 600 typically includes a plurality of computer storage media. Such media can be any accessible media that can be obtained by the electronic device 600, including but not limited to volatile and non-volatile media, removable and non-removable media. The memory 620 can be a volatile memory (e.g., registers, cache, random access memory (RAM)), a non-volatile memory (e.g., read-only memory (ROM), electrically erasable programmable read-only memory (EEPROM), flash memory), or some combination thereof. The storage device 630 can be a removable or non-removable medium and can include a machine-readable medium, such as a flash drive, a disk, or any other medium that can be used to store information and / or data and can be accessed within the electronic device 600.
[0147] The electronic device 600 may further include additional removable / non-removable, volatile / non-volatile storage media. Although not shown in FIG6 , a disk drive for reading or writing from a removable, non-volatile disk (e.g., a “floppy disk”) and an optical drive for reading or writing from a removable, non-volatile optical disk may be provided. In these cases, each drive may be connected to a bus (not shown) by one or more data media interfaces. The memory 620 may include a computer program product 625 having one or more program modules configured to perform various methods or actions of various embodiments of the present disclosure.
[0148] The communication unit 640 enables communication with other electronic devices via a communication medium. Additionally, the functions of the components of the electronic device 600 can be implemented in a single computing cluster or multiple computing machines that can communicate via a communication connection. Thus, the electronic device 600 can operate in a networked environment using a logical connection with one or more other servers, a network personal computer (PC), or another network node.
[0149] The input device 650 may be one or more input devices, such as a mouse, keyboard, or trackball. The output device 660 may be one or more output devices, such as a display, a speaker, or a printer. The electronic device 600 may also communicate with one or more external devices (not shown) through the communication unit 640 as needed, such as a storage device, a display device, or the like, with one or more devices that allow a user to interact with the electronic device 600, or with any device that allows the electronic device 600 to communicate with one or more other electronic devices (e.g., a network card, a modem, etc.). Such communication may be performed via an input / output (I / O) interface (not shown).
[0150] According to an exemplary implementation of the present disclosure, a computer-readable storage medium is provided, on which computer-executable instructions are stored, wherein the computer-executable instructions are executed by a processor to implement the method described above. According to an exemplary implementation of the present disclosure, a computer program product is also provided, which is tangibly stored on a non-transitory computer-readable medium and includes computer-executable instructions, and the computer-executable instructions are executed by a processor to implement the method described above.
[0151] Various aspects of the present disclosure are described herein with reference to flowcharts and / or block diagrams of methods, apparatuses, devices, and computer program products implemented according to the present disclosure. It should be understood that each block of the flowcharts and / or block diagrams, and combinations of blocks in the flowcharts and / or block diagrams, can be implemented by computer-readable program instructions.
[0152] These computer-readable program instructions can be provided to a processing unit of a general-purpose computer, a special-purpose computer, or other programmable data processing device, thereby producing a machine, such that when these instructions are executed by the processing unit of the computer or other programmable data processing device, a device is generated that implements the functions / actions specified in one or more blocks in the flowchart and / or block diagram. These computer-readable program instructions can also be stored in a computer-readable storage medium, where these instructions cause the computer, programmable data processing device, and / or other device to operate in a specific manner. Thus, the computer-readable medium storing the instructions comprises an article of manufacture that includes instructions for implementing various aspects of the functions / actions specified in one or more blocks in the flowchart and / or block diagram.
[0153] Computer-readable program instructions can be loaded onto a computer, other programmable data processing apparatus, or other device so that a series of operational steps are performed on the computer, other programmable data processing apparatus, or other device to produce a computer-implemented process, thereby causing the instructions executed on the computer, other programmable data processing apparatus, or other device to implement the functions / actions specified in one or more boxes in the flowchart and / or block diagram.
[0154] The flow charts and block diagrams in the accompanying drawings show the possible architecture, functions and operations of the systems, methods and computer program products according to multiple implementations of the present disclosure. In this regard, each box in the flow chart or block diagram can represent a part for a module, program segment or instruction, and a part for a module, program segment or instruction comprises one or more executable instructions for realizing the logical function of the specification. In some alternative implementations, the functions marked in the box can also occur in a sequence different from that marked in the accompanying drawings. For example, two continuous boxes can actually be executed substantially in parallel, and they can sometimes be executed in the opposite order, depending on the functions involved. It should also be noted that each box in the block diagram and / or flow chart, and the combination of the boxes in the block diagram and / or flow chart can be realized by a special hardware-based system that performs the function or action of the specification, or can be realized by a combination of special hardware and computer instructions.
[0155] While various implementations of the present disclosure have been described above, the foregoing description is intended to be illustrative, not exhaustive, and not limited to the disclosed implementations. Many modifications and variations will be apparent to those skilled in the art without departing from the scope and spirit of the described implementations. The terminology used herein is selected to best explain the principles of the implementations, their practical applications, or improvements to existing technologies, or to enable others skilled in the art to understand the various implementations disclosed herein.
Claims
1. A method for generating an expression, comprising: receiving an expression generation request associated with a target image; as well as A first group of expression resources generated based on the target image is provided, wherein the group of expression resources is associated with a first style, and the first group of expression resources includes a plurality of expression resources corresponding to a plurality of preset themes.
2. The method of claim 1, wherein receiving an expression generation request associated with a target image comprises: Presenting a conversation interface, wherein the conversation interface includes an input control; receiving a selection of the first style in the input control, and presenting an image acquisition interface; as well as In response to acquiring the target image via the image acquisition interface, receiving the expression generation request associated with the target image.
3. The method according to claim 2, wherein the target image comprises: an image captured via the image acquisition interface, or The image is uploaded via the image acquisition interface.
4. The method of claim 1 , wherein receiving an expression generation request associated with a target image comprises: Present the work creation interface; as well as In response to acquiring the target image via the work creation interface, receiving the expression generation request associated with the target image.
5. The method of claim 1 , wherein receiving an expression generation request associated with a target image comprises: Presenting a viewing interface of the first work; as well as Based on a selection of a first entry in the viewing interface, receiving the expression generation request associated with the target image, wherein the target image is determined based on the first work.
6. The method according to claim 5, wherein: The first work is an image work including at least one image, and the target image is determined based on the at least one image; or The first work is a video work, and the target image is determined based on at least one video frame of the video work.
7. The method according to claim 1, further comprising: A selection of the first style from a set of preset styles is received.
8. The method according to claim 1, further comprising: A shooting interface for shooting the target image is presented, wherein the shooting interface includes instruction information, and the instruction information is used to guide the user to shoot the target image corresponding to the preset part. 9 . The method according to claim 8 , wherein the shooting interface further comprises a style selection component for receiving the user's selection of the first style from a group of preset styles. 10 . The method according to claim 1 , wherein the first group of expression resources at least includes a first expression resource, and the first expression resource includes text content.
11. The method according to claim 10, further comprising: Based on the editing operation on the text content, the text content in the first expression resource is updated.
12. The method according to claim 1, further comprising: Associated with the second expression resource in the first group of expression resources, providing an update entry; as well as Based on the selection of the update entry, the second expression resource is updated, wherein the updated second expression resource is an expression resource regenerated based on the target image.
13. The method according to claim 1, wherein the first set of expression resources is presented in an expression preview interface, the method further comprising: Based on a preset operation for the expression preview interface, a second group of expression resources corresponding to a second style is provided in the expression preview interface, where the second group of expression resources is generated based on the target image.
14. The method according to claim 1, further comprising: receiving a first selection for a first group of target expression resources in the first group of expression resources; as well as Based on the first selection, the first group of target expression resources is added to an expression resource library for the current user or works corresponding to the first group of target expression resources are published.
15. The method according to claim 14, wherein adding the first group of target expression resources to an expression resource library or publishing works corresponding to the first group of target expression resources based on the first selection comprises: receiving a second selection of a second group of target expression resources in a third group of expression resources generated based on the target image; as well as Add the first group of target expressions and the second group of target expressions to the expression resource library or publish the works corresponding to the first group of target expression resources and the second group of target expression resources.
16. The method according to claim 1, wherein the plurality of expression resources included in the first group of expression resources are generated by providing a plurality of preset guide words to a target model, and the plurality of preset guide words correspond to the plurality of preset themes.
17. A method for publishing a work, comprising: Presenting a viewing interface of a first work, where the first work includes a picture work or a video work; Based on the selection of the third entry in the viewing interface, presenting a group of expression resources generated based on the first work; as well as Based on the selection of at least one expression resource in the group of expression resources, a second work corresponding to the at least one expression resource is published.
18. The method according to claim 17, wherein the set of expression resources comprises at least one of the following: Expression resources extracted from the screen of the first work; Expression resources generated by the target model based on the target screen content in the first work.
19. The method according to claim 18, further comprising: presenting a set of candidate screen contents in the first work; as well as A selection of the target picture content from the set of candidate picture contents is received.
20. An expression generating device, comprising: a request receiving module configured to receive an expression generation request associated with a target image; as well as The expression providing module is configured to provide a first group of expression resources generated based on the target image, wherein the group of expression resources is associated with a first style, and the first group of expression resources includes multiple expression resources corresponding to multiple preset themes.
21. A work publishing device, comprising: An interface presentation module, configured to present a viewing interface of a first work, wherein the first work includes a picture work or a video work; An expression presentation module, configured to present a group of expression resources generated based on the first work based on a selection of a third entry in the viewing interface; as well as The work publishing module is configured to publish a second work corresponding to at least one expression resource based on the selection of at least one expression resource in the group of expression resources.
22. An electronic device, comprising: at least one processing unit; as well as At least one memory, the at least one memory being coupled to the at least one processing unit and storing instructions for execution by the at least one processing unit, the instructions, when executed by the at least one processing unit, causing the electronic device to perform a method according to any one of claims 1 to 16 or 17 to 19.
23. A computer-readable storage medium having a computer program stored thereon, the computer program being executable by a processor to implement the method according to any one of claims 1 to 16 or 17 to 19.
Citation Information
Patent Citations
Expression generation and work release method and device, equipment and storage medium
CN117745885A
Facial expression generation method and device, terminal and storage medium
CN107977928A
Expression generation method and device and storage medium
CN111612876A
Media content processing method and device, equipment and storage medium
CN115269886A
Expression image generation method and device, electronic equipment and readable storage medium
CN116543079A