METHOD, APPARATUS, ELECTRONIC DEVICE, STORAGE MEDIUM AND COMPUTER PROGRAM FOR PROCESSING MEDIA CONTENT
The method and apparatus enhance media content processing by generating personalized expression objects from user-selected media content, improving interaction and resource utilization.
Patent Information
- Application Number
- JP2024568015
- Authority / Receiving Office
- JP · JP
- Patent Type
- Patents
- Current Assignee / Owner
- Priority Date
- 2022-08-15
- Filing Date
- 2023-08-14
- Publication Date
- 2025-10-27
- Estimated Expiration
- 2043-08-14
AI Technical Summary
Existing media content processing methods lack the ability to optimize and generate personalized expression objects based on media content, limiting user interaction and expression options.
A method and apparatus that display media content on a preset page, determine target content, and generate expression objects such as emojis or stickers based on user selection, allowing placement in an expression selection panel for personalized use.
Enriches user expression options, improves interaction experience, and enhances the utilization of media content resources by enabling personalized facial expression object generation during media consumption.
Smart Images

Figure 0007760760000001 
Figure 0007760760000002 
Figure 0007760760000003
Abstract
Description
[Technical Field]
[0001] CROSS-REFERENCE TO RELATED APPLICATIONS This application claims priority from a Chinese patent application filed with the State Intellectual Property Office of the People's Republic of China on August 15, 2022, bearing application number 202210977422.9, the entire contents of which are incorporated herein by reference.
[0002] TECHNICAL FIELD Embodiments of the present disclosure relate to the field of computer technology, such as media content processing methods, apparatus, devices and storage media. [Background technology]
[0003] With the rapid development of Internet technology, communication between users has become more and more convenient, and users can perform various information interactions through applications.
[0004] Interactions based on facial expressions, such as emojis, are a way of expressing emotions through images. Compared to text information, facial expressions can more vividly express a user's emotions, and have found wide application. Facial expressions are often provided by applications, and users can use the facial expressions in the applications to chat, comment, and so on. Summary of the Invention [Problem to be solved by the invention]
[0005] The embodiments of the present disclosure provide a media content processing method, apparatus, storage medium, and device that can optimize a media content processing method and generate an expression object based on the media content. [Means for solving the problem]
[0006] According to a first aspect, an embodiment of the present disclosure comprises: Displaying media content, including images and / or videos, in the target media work on a preset page of the current application; determining at least one target media content among the media content; generating at least one target expression object based on the at least one target media content in response to an expression object generation command for the at least one target media content; The at least one target facial expression object is placed in a facial expression selection panel of the current application, thereby providing a method for processing media content.
[0007] According to a second aspect, an embodiment of the present disclosure comprises: a media content display module configured to display media content, including images and / or videos, of the target media work on a preset page of the current application; a target content determination module configured to determine at least one target media content in the media content; an expression object generation module configured to generate at least one target expression object based on the at least one target media content in response to an expression object generation command for the at least one target media content; The at least one target facial expression object is arranged on a facial expression selection panel of the current application, and the apparatus for processing media content further provides.
[0008] According to a third aspect, an embodiment of the present disclosure comprises: one or more processors; a storage device configured to store one or more programs; The present invention further provides an electronic device in which the one or more programs, when executed by the one or more processors, cause the one or more processors to implement a method for processing media content according to an embodiment of the present disclosure.
[0009] According to a fourth aspect, an embodiment of the present disclosure further provides a storage medium including computer-executable instructions for, when executed by a computer processor, performing a method of media content processing according to an embodiment of the present disclosure. [Brief explanation of the drawings]
[0010] In the drawings, the same or similar reference numerals indicate the same or similar elements, and it should be understood that the drawings are schematic and that parts and elements are not necessarily drawn to scale. [Figure 1] 1 is a flowchart of a method for processing media content according to an embodiment of the present disclosure. [Figure 2] FIG. 1 is a schematic diagram of an interface according to an embodiment of the present disclosure. [Figure 3] 10 is a flowchart of another method for processing media content according to an embodiment of the present disclosure. [Figure 4] 10 is a flowchart of yet another method for processing media content according to an embodiment of the present disclosure. [Figure 5] FIG. 10 is a schematic diagram of another interface according to an embodiment of the present disclosure. [Figure 6] 10 is a flowchart of another method for processing media content according to an embodiment of the present disclosure. [Figure 7] FIG. 1 is a schematic diagram of interface interactions according to an embodiment of the present disclosure. [Figure 8] FIG. 1 is a structural schematic diagram of a media content processing device according to an embodiment of the present disclosure; [Figure 9] 1 is a structural schematic diagram of an electronic device according to an embodiment of the present disclosure. DETAILED DESCRIPTION OF THE INVENTION
[0011] Hereinafter, embodiments of the present disclosure will be described with reference to the drawings. It should be understood that the drawings and embodiments of the present disclosure are for illustrative purposes only and are not intended to limit the scope of protection of the present disclosure.
[0012] It should be understood that the steps described in the method embodiments of the present disclosure may be performed in different orders and / or in parallel.
[0013] As used herein, the term "comprises" and variations thereof are open-ended and mean "including but not limited to." The term "based on" means "based at least in part on." The term "in one embodiment" refers to "at least one embodiment," the term "in another embodiment" refers to "at least one other embodiment," and the term "in some embodiments" refers to "at least some embodiments." Relevant definitions of other terms are provided below.
[0014] It should be noted that the concepts of "first," "second," etc. referred to in this disclosure are merely intended to distinguish between different devices, modules, or units, and do not limit the order or interdependence of the functions performed by these devices, modules, or units.
[0015] It should be noted that the modifications "one" and "multiple" referred to in this disclosure are exemplary rather than limiting, and one of ordinary skill in the art should understand "one or more" unless the context clearly dictates otherwise.
[0016] The names of messages or information exchanged between devices in the embodiments of the present disclosure are for illustrative purposes only and do not limit the scope of these messages or information.
[0017] It should be understood that before using any of the technical solutions disclosed in each embodiment of the present disclosure, the type, scope of use, and usage scenarios of the personal information disclosed herein should be notified to users in an appropriate manner in accordance with relevant laws and regulations, and user approval should be obtained.
[0018] For example, in response to receiving an active request from a user, presentation information is sent to the user to clearly indicate to the user that the requested operation requires the acquisition and use of the user's personal information, allowing the user to independently choose, based on the presentation information, whether to provide the personal information to software or hardware, such as an electronic device, application, server, or storage medium, that performs the operation of the technical solution of the present disclosure.
[0019] As an optional but non-limiting implementation, the method of transmitting the presentation information to the user in response to receiving an active request from the user may be, for example, a pop-up window, which may display the presentation information in text, and the pop-up window may further include a selection control for the user to select "agree" or "disagree" to providing personal information to the electronic device.
[0020] It should be understood that the above notification and user permission acquisition process is merely exemplary and does not limit the implementation manner of the present disclosure, and that methods that comply with other relevant laws may also be applied to the implementation manner of the present disclosure.
[0021] It should be understood that the data related to this technical solution (including but not limited to the data itself, the acquisition or use of the data) should comply with the requirements of relevant laws and regulations.
[0022] FIG. 1 is a flowchart of a media content processing method according to an embodiment of the present disclosure, which is applied to the case of media content processing. The method may be performed by a media content processing apparatus, which may be implemented in the form of software and / or hardware, and may optionally be implemented by an electronic device, which may be a mobile terminal such as a mobile phone, a smart watch, a tablet computer, or a personal digital assistant, or may be a device such as a personal computer (PC) terminal or a server.
[0023] As shown in FIG. 1, the method includes step 101 of displaying media content, including images and / or videos, in a target media work on a preset page of a current application.
[0024] In an embodiment of the present disclosure, the current application may be a preset application, and the preset page may be a page in the preset application, which may provide a media work display function and an expression object generation function. For example, when a user needs to use the expression object generation function, the user may open a preset page in the preset application. The media work includes one or more media contents, which may include images and / or videos and may further include audio, etc., but are not limited thereto. The images may include still images and / or dynamic images such as Graphics Interchange Format (GIF). The images or video frames of the videos may include image content and / or text content, etc.
[0025] In the embodiments of the present disclosure, the target media work may be understood as the media work to which the media content currently displayed on the preset page belongs. The media content currently displayed on the preset page may include all or part of the media content in the target media work, and the display format may be the same as or different from the display format of the target media work. For example, the media content may be displayed as a thumbnail, or the image size or ratio may be reduced, relative to the display format of the target media work.
[0026] Before distributing a media work, a media work distributor may set attribute information for each media content in the media work, and the attribute information may be used to indicate whether the corresponding media content is permitted to be used to generate an expression object by other users. The attribute information corresponding to the media content displayed on the preset page indicates that it is permitted to be used to generate an expression object, that is, the function of generating an expression object based on the media content in the media work has been fully authorized by the media work distributor.
[0027] 2 is a schematic diagram of an interface according to an embodiment of the present disclosure. As shown in FIG. 2, a preset page 201 displays media content 202. The displayed media content is a portion of the media content in a target media work. The fourth media content, which is not fully shown, can be switched to display different media content by inputting operations such as sliding left or right. The target media work may include only one media content, for example, a single image, and in this case, all or part of the media content may be displayed.
[0028] In step 102, at least one target media content is determined in the media content.
[0029] In an embodiment of the present disclosure, the target media content may be understood as media material for generating an expression object, and the determination of the target media content may be automatically determined by the current application or independently determined by the user.
[0030] Optionally, determining at least one target media content from the media content includes determining at least one selected media content as the at least one target media content in response to a selection operation on the media content. The advantage of this configuration is that it allows a user to more freely select target media content to achieve personalized customization of facial expression objects.
[0031] For example, when media contents are displayed on the preset page, a selection control, such as a check box or a positioning cursor, corresponding to each media content may be displayed, and the user can select the media content to be used to generate the expression target by inputting a selection operation through the selection control. After receiving the user's selection operation for the media content, the selected media content is determined as the target media content in response to the user's selection operation.
[0032] As shown in FIG. 2, a check box 203 may be displayed for the media content, and the user can select the media content to which the check box belongs by checking the check box, for example, the first media content in FIG. 2 is selected.
[0033] Alternatively, if the target media work contains only one media content, the media content may be automatically determined as the target media content.
[0034] In step 103, in response to the expression object generation command for the at least one target media content, at least one target expression object is generated based on the at least one target media content, and the at least one target expression object is placed on an expression selection panel of the current application.
[0035] For example, after receiving a command to generate an expression object for a target media content, a media resource corresponding to the target media content, such as image data or video frame data, may be obtained, and the obtained media resource may undergo related operations such as image processing, format conversion, or encoding to generate a corresponding target expression object. The expression object may be an emoji, an expression sticker, or the like.
[0036] Optionally, a preset generation control may be displayed on the preset page, and the user can input an expression target generation command by triggering the preset generation control. As shown in Fig. 2, an "Add" button 204, which is a preset generation control, is displayed, and when the user clicks the "Add" button 204, a target expression target can be generated.
[0037] In an embodiment of the present disclosure, the target expression target is arranged in an expression selection panel of the current application, and the user can select and apply the target expression target based on the expression selection panel. Optionally, the generated target expression target may be used for information interactions, such as sending an instant message including the target expression target or posting a comment message including the target expression target. The generated target expression target may be added to an expression library of a preset application, and may thereby be displayed in the expression selection panel of the preset application so that it can be selected by the user. For example, the target expression target may be added to a custom expression set in the expression library. Optionally, the target expression target may be used for information interactions between the current user (i.e., the user who triggered the generation of the target expression target) and the distributor of the target media work, improving the interaction experience between them.
[0038] A media content processing method according to an embodiment of the present disclosure displays media content, including images and / or videos, in a target media work on a preset page of a current application, determines at least one target media content from the media content, and generates a target expression object based on the target media content in response to an expression object generation command for the target media content, where the target expression object is arranged in an expression selection panel of the current application. By adopting the above technical solution, a user can generate an expression object based on media content in a media work during the process of viewing a media work, and arrange the generated expression object in the expression selection panel of the current application, thereby satisfying the user's needs for personalized expression object generation, enriching the style of expression objects, and allowing the user to select and apply expression objects in the expression selection panel more personalized options, improving the user experience, while enriching the uses of media content in the media work, and improving the utilization rate of media content resources.
[0039] In some embodiments, the target media piece includes a target image piece, the target image piece including at least one image, i.e., one or more images. Illustratively, the target image piece further includes audio, which may serve as background music when the target image piece is displayed, and the images are played in a predetermined order, possibly looping.
[0040] FIG. 3 is a flowchart of another media content processing method according to an embodiment of the present disclosure, which is described based on the above optional embodiment, taking the target media work as an example that the target image work, and the method may include the following steps:
[0041] In step 301, at least one image in the target image composition is displayed on a preset page of the current application.
[0042] For example, if the target image work includes multiple images, the images in the target image work may be displayed one by one or in batches, and the number of images displayed in the batches may be less than or equal to the total number of images in the target image work, and the order in which the images are displayed may match the order in which the images are displayed when the target image work is displayed.
[0043] Illustratively, as shown in FIG. 2, the media content 202 may be an image in a target image production.
[0044] In step 302, in response to a selection operation on at least one image, the selected at least one image is determined as at least one target image.
[0045] For example, as shown in FIG. 2, the target image may be determined based on the user's selection operation on a check box.
[0046] In step 303, in response to an expression object generation command for at least one target image, at least one target expression object is generated based on the at least one target image.
[0047] Optionally, one target expression object includes one or more target images.
[0048] For example, each target image may generate one target expression target by itself, or two or more target images may be integrated to generate one target expression target, for example, a target expression target of a dynamic image effect may be generated based on two or more target images.
[0049] Optionally, the target expression object includes a dynamic expression object, and the dynamic expression object is generated from a dynamic image in the at least one target image and / or a plurality of still images in the at least one target image.
[0050] Optionally, in order to meet the user's different facial expression object generation needs, the preset page may display preset generation controls corresponding to two generation methods, such as a first preset generation control "Add Each" button and a second preset generation control "Add All" button.
[0051] According to the media content processing method of the embodiment of the present disclosure, in the process of viewing an image work, a user can independently select images in the image work, and generate one or more facial expression objects based on the one or more images selected by the user, thereby meeting the user's needs for generating personalized facial expression objects, enriching the style of facial expression objects, and improving the user experience, while enriching the uses of images in the image work and improving the utilization rate of image resources.
[0052] In some embodiments, the target media piece comprises a target video piece, the target video piece comprising a plurality of video frames.
[0053] 4 is a flowchart of yet another method for processing media content according to an embodiment of the present disclosure, which is described based on the above-mentioned preferred embodiment, taking the target media work as a target video work as an example. The method may include the following steps:
[0054] In step 401, video progress information corresponding to a target video work is displayed on a preset page of a current application.
[0055] Exemplarily, the video progress information may be a progress bar, a video frame sequence, video chapter information, etc. If it is a video frame sequence, the displayed video frame sequence may include all or some of the video frames (which may be thumbnail views) in the target video work, and the order of the video frames in the video frame sequence may match the playback order of the video frames in the target video work.
[0056] FIG. 5 is a schematic diagram of another interface according to an embodiment of the present disclosure, displaying a video frame sequence 502 of a video production on a preset page 501.
[0057] In step 402, at least one target video frame set is determined in response to a video frame selection operation on the video progress information, each target video frame set including at least one video frame.
[0058] For example, a certain number of single video frames (which may be understood as one or more video snapshots) in the target video work may be selected and used to generate the expression object. One or more sets of consecutive video frames (which may be understood as one or more video segments) in the target video work may be selected and used to generate the expression object. The video frame selection operation may be a selection operation for a single video frame or a batch selection operation for multiple video frames.
[0059] For example, for video segment selection, this step includes determining a start video frame and an end video frame in response to a start frame selection operation and an end frame selection operation on the video progression information, and determining at least one target video frame set based on the start video frame and the end video frame, where each target video frame set includes one start video frame, one end video frame, and zero or at least one intermediate video frame, and the intermediate video frame is located between the corresponding start video frame and the corresponding end video frame in the video progression information. The advantage of this configuration is that by conveniently selecting a video segment, a corresponding expression target can be quickly generated based on the video segment. Optionally, the target video frame set may include only at least one intermediate video frame.
[0060] The start frame selection operation is used to select a start video frame, and the end frame selection operation is used to select an end video frame, and a start frame selection indicator and an end frame selection indicator may be associated with and displayed in the video progress information, and the user can determine the start video frame by adjusting the position pointed to by the start frame selection indicator, and determine the end video frame by adjusting the position pointed to by the end frame selection indicator.
[0061] For example, the start frame selection indicator and the end frame selection indicator may be displayed in pairs, and each pair of indicators corresponds to one target video frame set. For one target video frame set, if there is no intermediate video frame between the start video frame and the end video frame, the target video frame set includes two video frames, i.e., the start video frame and the end video frame.
[0062] 5, a video frame sequence 502 is displayed with an associated video frame selection box 503, the left boundary of which may be understood as a start frame selection indicator and the right boundary of which may be understood as an end frame selection indicator. A user can adjust the range of selected video frames by dragging the left or right boundary of the video frame selection box 503.
[0063] In step 403, in response to an expression object generation command for the at least one target video frame set, at least one target expression object is generated based on the at least one target video frame set.
[0064] Optionally, one target expression object includes one or more target video frame sets.
[0065] For example, each target video frame set may generate one target expression object by itself, or two or more target video frame sets may be combined to generate one target expression object.
[0066] Optionally, the target expression object includes a dynamic expression object, the dynamic expression object being generated by a target video frame set of the at least one target video frame set.
[0067] Optionally, the preset page may display preset generation controls corresponding to two generation methods, such as a third preset generation control "Generate Individually" button and a fourth preset generation control "Generate Integratedly" button, to meet the user's different facial expression object generation needs.
[0068] According to the media content processing method of the embodiment of the present disclosure, in the process of a user viewing a video work, the user can independently select video frames in the video work, and generate one or more facial expression objects based on one or more sets of video frames selected by the user, thereby meeting the user's needs for generating personalized facial expression objects, enriching the style of facial expression objects, and improving the user experience, while enriching the uses of video content in the video work and improving the utilization rate of video resources.
[0069] In some embodiments, the method further includes generating at least one preview expression object based on the at least one target media content and displaying the at least one preview expression object before generating the at least one target expression object based on the at least one target media content. The advantage of this configuration is that a preview function can be provided before generating an expression object, allowing a user to preview the effect of the expression object in advance, avoiding repeated modifications and improving efficiency of expression object generation.
[0070] Optionally, if there are multiple preview expression targets, one or more preview expression targets may be displayed.
[0071] For example, the preview expression target may be displayed on a preset page or in a target display area other than the preset page. For example, as shown in Figures 2 and 5, target display areas may be set above the preset page 201 and above the preset page 501, and the preview expression target may be displayed in the target display area.
[0072] In some embodiments, the method further includes receiving an editing operation for the at least one preview expression object, and generating at least one target expression object based on the at least one target media content includes generating at least one target expression object based on the at least one target media content and an editing result of the editing operation. An advantage of this configuration is that by allowing a user to edit based on the preview expression object, the generated target expression object can be more tailored to the user's needs. Editing operations may include, for example, adding text, adding stamps, or adjusting size.
[0073] In some embodiments, the method further includes displaying the preset page in response to a first preset trigger operation on the target display page of the target media work in the current application before displaying the media content of the target media work on the preset page of the current application. The advantage of this configuration is that it allows a user to conveniently enter the preset page in the process of browsing the target media work.
[0074] Alternatively, the first preset trigger operation may be a trigger operation for a preset entrance of a preset page, or may be an operation for triggering entry into a preset page that acts on the preset page, such as a double-click operation.
[0075] Optionally, the preset entrance may be an entrance control on the target display page, such as a "convert to facial expression" button, and triggering the preset entrance may trigger the display of the preset page.
[0076] Alternatively, the size of the preset page may be the same as or different from the size of the target display page. If the size of the preset page is smaller than the size of the target display page, the preset page may be overlapped on top of the target display page. In this case, the target media work may continue to be displayed on the target display page, and the user may continue to view the target media work during the process of setting the facial expression.
[0077] In an embodiment of the present disclosure, a target media work may be tagged with a work tag, which may be used to indicate, for example, the type of media work to which the target media work belongs or a topic related to the media work to which the target media work belongs. Exemplarily, when a distributor of a target media work distributes the target media work, the distributor may add a work tag to the target media work. The preset indicators may include work tags related to facial expressions, which may be referred to as preset facial expression tags, and may include tags such as #emoji#, #facialexpression#, #competeimage#, #facialexpression#, or #newfacialexpression#.
[0078] In some embodiments, the step of displaying the target media work on the target display page further includes: displaying a preset control display area on the target display page in response to a second preset trigger operation; determining whether the target media work has a preset indicator; and if the target media work has the preset indicator, displaying a preset entrance of the preset page at a first preset display position in the preset control display area. The advantage of this configuration is that the display position of the preset entrance can be flexibly determined based on whether the target media work has a preset indicator, and if the target media work has a preset indicator, it will be displayed at the preset display position.
[0079] Exemplarily, the preset control display area may include preset interaction controls, such as a watch together control, a like control, a share control, and a comment control, and may also include preset function controls, such as a save control. The first preset display position may be a fixed position preset in the preset control display area, or may be a relative position determined based on the display positions of the preset interaction controls and / or preset function controls in the preset control display area, and the relative positional relationship between the first preset display position and the current display position may be preset.
[0080] In some embodiments, after determining whether the target media work has a preset indicator, if the target media work does not have the preset indicator, the method further includes displaying a preset entrance of the preset page at a second preset display position in the preset control display area, where the display priority of the first preset display position is higher than the display priority of the second preset display position. The advantage of this setting is that if the target media work has a preset indicator, it indicates that the target media work is more likely to be used to generate facial expressions, and by displaying the preset entrance at a position with a higher display priority, it is easier for the user to trigger the preset entrance. If the target media work does not have a preset indicator, the display priority may be slightly lower, allowing the preset control display area to be used rationally for displaying controls.
[0081] 6 is a flowchart of another media content processing method according to an embodiment of the present disclosure, which will be described based on the alternative method in the above embodiment. The method includes the following steps:
[0082] In step 601, during the process of displaying a target media work on a target display page of a current application, a preset control display area is displayed on the target display page in response to a second preset trigger operation.
[0083] FIG. 7 is a schematic diagram of interface interaction according to an embodiment of the present disclosure. For example, take the target media work as an example, the target image work is displayed on the target display page 701. When displaying the second image in the target image work, the user inputs a long press operation (preset trigger operation) on the target display page 701, and a preset control display area 702 is displayed on the target display page 701.
[0084] In step 602, it is determined whether the target media work has a preset mark, and if yes, step 603 is executed; if the target media work does not have a preset mark, step 604 is executed.
[0085] Optionally, step 602 may be performed before displaying the preset control display area on the target display page, and after displaying the preset control display area, the display position of the preset entrance can be directly determined based on the judgment result of step 602, thereby improving the display speed of the preset entrance.
[0086] As shown in FIG. 7, if the target image work has a #emoji# tag (preset indicator), step 603 can be performed.
[0087] In step 603, the preset entry of the preset page is displayed at the first preset display position in the preset control display area, and step 605 is executed.
[0088] For example, the display priority of the preset display position may be determined based on the relative positional relationship between the preset display position and the display position of the preset interaction control and / or the preset function control in the preset control display area.
[0089] 7, the preset entry is the add expression button 703, and the preset function control is the save button 704. If the target image work has a preset indicator, the add expression button 703 is displayed before the save button 704. If the target image work does not have an indicator, the add expression button 703 is displayed after the save button 704. For example, the add expression button 703 and the save button 704 can be swapped, or the add expression button 703 can be replaced with another control, such as a view together control, and the add expression button 703 can be placed after the save button 704, and the add expression button 703 can be displayed after the user inputs a left slide operation.
[0090] In step 604, a preset entrance of a preset page is displayed in a second preset display position in the preset control display area, and the display priority of the first preset display position is higher than the display priority of the second preset display position.
[0091] In step 605, a preset page is displayed in response to a trigger operation on a preset entry of the preset page.
[0092] As shown in FIG. 7, when the user clicks on the Add Expression button 703, a Preset page 705 is displayed.
[0093] In step 606, the media content in the target media work is displayed on a preset page.
[0094] As shown in FIG. 7, a preset page 705 displays a plurality of images in the target image work, and a check box is displayed for each image.
[0095] In step 607, in response to a selection operation on the media content, the selected at least one media content is determined as at least one target media content.
[0096] As shown in FIG. 7, the user clicks the checkbox of the second image containing the moon pattern to determine the second image as the target media content.
[0097] In step 608, generate at least one preview expression object based on the at least one target media content, and display the at least one preview expression object.
[0098] 7, a preview expression object is generated based on the second image and displayed in the target display area 706. Optionally, if the user continues to select another image, a preview expression object can be generated and displayed based on the current newly selected image. Optionally, the user can switch between displaying different preview expression objects by triggering media content in the preset page.
[0099] In step 609, an edit operation for at least one preview expression object is received.
[0100] As shown in FIG. 7, if the user wants to further edit the facial expression target, he / she can click the edit button to edit, for example, add the text “good night”, and the edited result can be displayed in the target display area in real time.
[0101] In step 610, in response to an expression object generation command for the at least one target media content, at least one target expression object is generated based on the at least one target media content and an editing result of the editing operation.
[0102] As shown in FIG. 7, if the user is satisfied with the current preview effect, he or she may instruct the generation of an expression object by clicking the Add button and inputting an expression object generation command.
[0103] For example, if the user is satisfied with the initial preview expression object, there is no need to edit it, and the user may directly click the Add button to input an expression object generation command.
[0104] In other embodiments, when the user is allowed to select multiple images or multiple video frames, the preset control display area 702 may include a "Generate Composite" button for combining multiple images or multiple video frames into one expression target and a "Generate Each" button for combining the expression targets individually, so that the user can select different expression target synthesis results as needed.
[0105] In step 611, at least one target expression object is placed in the expression selection panel of the current application.
[0106] For example, after the target expression object is successfully generated, the target expression object may be added to the expression library of the current application, such as by placing the target expression object in the expression selection panel of the current application, and the user may return to the target display page to continue displaying the target media work. Also, a notification of successful addition may be displayed on the target display page, such as "Expression added successfully" in FIG. 7.
[0107] According to the media content processing method of the embodiment of the present disclosure, when a user is viewing a media work, the user can conveniently trigger the display of a control display area, and determine the display position of a preset entrance of a preset page for generating an expression object based on whether the media work has a preset expression tag. When the user wants to generate an expression object based on the media content in the media work, the user can trigger the preset entrance to enter the preset page, select media content on the preset page, and view a preview expression object. The user can further edit based on the preview expression object, and further create a personalized expression object that better meets their needs and add it to the expression library, which makes it easy to search for and use the expression object in the expression library later, and is advantageous to improving the interaction experience with the interaction object.
[0108] FIG. 8 is a structural schematic diagram of a media content processing device according to an embodiment of the present disclosure. As shown in FIG. 8, the device includes a media content display module 801, a target content determination module 802, and an expression object generation module 803.
[0109] The media content display module 801 is configured to display media content including images and / or videos in a target media work on a preset page, the target content determination module 802 is configured to determine at least one target media content in the media content, and the expression object generation module 803 is configured to generate a target expression object based on the at least one target media content in response to an expression object generation command.
[0110] A media content processing device according to an embodiment of the present disclosure displays media content, including images and / or videos, in a target media work on a preset page of a current application, determines at least one target media content from the media content, and generates a target expression object based on the target media content in response to an expression object generation command for the target media content, where the target expression object is arranged in an expression selection panel of the current application. By adopting the above technical solution, a user can generate an expression object based on media content in a media work during the process of viewing a media work, and arrange the generated expression object in the expression selection panel of the current application, thereby meeting the user's needs for personalized expression object generation, enriching the style of expression objects, and allowing more personalized choices when the user selects and applies an expression object in the expression selection panel, improving the user experience, while enriching the uses of media content in the media work, and improving the utilization rate of media content resources.
[0111] Optionally, the target content determination module is configured to determine, in response to a selection operation on the media content, at least one selected media content as at least one target media content.
[0112] Optionally, the target media work includes a target image work including at least one image, the media content display module is configured to display at least one image in the target image work on a preset page of a current application, and the target content determination module is configured to determine the selected at least one image as at least one target image in response to a selection operation on the at least one image.
[0113] Optionally, the target image creation includes dynamic and / or still images.
[0114] Optionally, the expression object generation module 803 is configured to generate at least one target expression object based on the target media content by generating at least one target expression object including one or more target images of the at least one target image based on the at least one target image.
[0115] Optionally, the target media work includes a target video work including a plurality of video frames, the media content display module is configured to display video progress information corresponding to the target video work on a preset page of a current application, and the target content determination module is configured to determine at least one target video frame set in response to a video frame selection operation on the video progress information, each target frame set including at least one video frame.
[0116] Optionally, the expression object generation module 803 is configured to generate at least one target expression object based on the at least one target media content by generating at least one target expression object based on the at least one target video frame set, where one target expression object includes one or more target video frame sets of the at least one target expression object.
[0117] Optionally, the at least one target expression object includes a dynamic expression object, and the dynamic expression object is generated from at least one of a dynamic image in the at least one target image, a plurality of still images in the at least one target image, and a target video frame set in the at least one target video frame set.
[0118] Optionally, the target content determination module includes a video frame determination unit configured to determine a start video frame and an end video frame in response to a start frame selection operation and an end frame selection operation on the video progress information, and a video frame set determination unit configured to determine at least one target video frame set based on the start video frame and the end video frame.
[0119] Optionally, the device further includes a preview expression object generation module configured to generate at least one preview expression object based on the at least one target media content and display the at least one preview expression object before generating at least one target expression object based on the at least one target media content.
[0120] Optionally, the device further includes an editing operation receiving module configured to receive an editing operation on the at least one preview expression object, and the expression object generation module 803 is configured to generate at least one target expression object based on the at least one target media content by generating at least one target expression object based on the at least one target media content and an editing result of the editing operation.
[0121] Optionally, the device further includes a preset page display module configured to display the preset page in response to a first preset trigger operation on the preset page in a target display page of the target media work of the current application before displaying media content in the target media work on the preset page of the current application.
[0122] Optionally, the device further includes: a display area display module configured to display a preset control display area on the target display page in response to a second preset trigger operation during the process of displaying the target media work on the target display page; a preset indicator determination module configured to determine whether the target media work has a preset indicator; and a first preset entrance display module configured to display a preset entrance of the preset page at a first preset display position in the preset control display area if the target media work has the preset indicator.
[0123] Optionally, the device further includes a second preset entrance display module configured to, after determining whether the target media work has a preset indicator, if the target media work does not have the preset indicator, display a preset entrance of the preset page at a second preset display position in the preset control display area, wherein the display priority of the first preset display position is higher than the display priority of the second preset display position.
[0124] The media content processing device according to the embodiments of the present disclosure can execute the media content processing method according to any embodiment of the present disclosure, and has functional modules and effects according to the executed method.
[0125] Each unit and module included in the above device is simply divided according to functional logic, but is not limited to the above division, as long as it can realize the appropriate function. Furthermore, the specific names of each functional unit are intended to distinguish them from each other, and do not limit the scope of protection of the embodiments of the present disclosure.
[0126] FIG. 9 is a structural schematic diagram of an electronic device according to an embodiment of the present disclosure. Hereinafter, reference will be made to FIG. 9 , which illustrates a structural schematic diagram of an electronic device (e.g., a terminal device or server in FIG. 9 ) 900 for implementing an embodiment of the present disclosure. The terminal device in the embodiment of the present disclosure may include, but is not limited to, mobile terminals such as mobile phones, laptops, digital broadcast receivers, personal digital assistants (PDAs), tablet PCs (Portable Android Devices, PADs), portable media players (PMPs), and in-vehicle terminals (e.g., in-vehicle navigation terminals), as well as fixed terminals such as digital televisions (TVs) and desktop computers. The electronic device illustrated in FIG. 9 is merely an example and does not impose any limitations on the functionality and scope of use of the embodiment of the present disclosure.
[0127] 9, electronic device 900 may include a processing unit (e.g., a central processing unit, a graphics processor, etc.) 901 that can perform various appropriate operations and processes based on programs stored in read-only memory (ROM) 902 or loaded from storage device 908 into random access memory (RAM) 903. RAM 903 also stores various programs and data necessary for the operation of electronic device 900. Processing unit 901, ROM 902, and RAM 903 are interconnected by bus 904. Input / output (I / O) interface 905 is also connected to bus 904.
[0128] Generally, the following devices may be connected to the I / O interface 905: input devices 906 including, for example, a touchscreen, touchpad, keyboard, mouse, camera, microphone, accelerometer, gyroscope, etc.; output devices 907 including, for example, a liquid crystal display (LCD), speaker, vibrator, etc.; storage devices 908 including, for example, a magnetic tape, hard disk, etc.; and communication devices 909. The communication devices 909 may allow the electronic device 900 to communicate wirelessly or via wires with other devices to exchange data. While FIG. 9 illustrates the electronic device 900 with a variety of devices, it should be understood that it is not required to implement or include all of the devices shown. More or fewer devices may alternatively be implemented or included.
[0129] According to embodiments of the present disclosure, the processes described with reference to the flowcharts above may be implemented as a computer software program. For example, embodiments of the present disclosure include a computer program product including a computer program embodied in a non-transitory computer-readable medium, the computer program including program code for performing the methods illustrated in the flowcharts. In such embodiments, the computer program may be downloaded and installed from a network via the communication device 909, installed from the storage device 908, or installed from the ROM 902. When the computer program is executed by the processing device 901, it performs the functions defined in the methods of the embodiments of the present disclosure.
[0130] The names of messages or information exchanged between devices in the embodiments of the present disclosure are for illustrative purposes only and do not limit the scope of these messages or information.
[0131] The electronic device according to the embodiment of the present disclosure belongs to the same inventive concept as the media content processing method according to the above embodiment, and the above embodiment may be referred to for technical details not described in detail in this embodiment, and this embodiment has the same effects as the above embodiment.
[0132] An embodiment of the present disclosure provides a computer storage medium having a computer program stored thereon, the computer storage medium implementing the method for processing media content according to the above embodiment when the program is executed by a processor.
[0133] The computer-readable medium of the present disclosure may be a computer-readable transmission medium or a computer-readable storage medium, or any combination thereof. The computer-readable storage medium may be, for example, but is not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination thereof. More specific examples of the computer-readable storage medium include, but are not limited to, an electrical connection having one or more conductors, a portable computer magnetic disk, a hard disk, RAM, ROM, an Erasable Programmable Read-Only Memory (EPROM), a flash memory, an optical fiber, a portable Compact Disc Read-Only Memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the above. In the present disclosure, the computer-readable storage medium may be any tangible medium that contains or stores a program, and the program may be used in or in connection with an instruction execution system, apparatus, or device. In the present disclosure, the computer-readable signal medium may include a data signal propagating in baseband or as part of a carrier wave, with computer-readable program code borne thereon. Such propagated data signals may take various forms, including, but not limited to, electromagnetic signals, optical signals, or any suitable combination of the above. A computer-readable signal medium may be any computer-readable medium other than a computer-readable storage medium, which is capable of transmitting, propagating, or transmitting a program for use in or in connection with an instruction execution system, apparatus, or device. Program code contained in a computer-readable medium may be transmitted over any suitable medium, including, but not limited to, electrical wire, optical cable, radio frequency (RF), or the like, or any suitable combination thereof.
[0134] In some embodiments, clients and servers may communicate using any now known or future developed network protocol, such as HyperText Transfer Protocol (HTTP), and may interconnect any form or medium of digital data communication (e.g., a communications network). Examples of communications networks include local area networks (LANs), wide area networks (WANs), international networks (e.g., the Internet), and end-to-end networks (e.g., ad hoc end-to-end networks), as well as any now known or future developed networks.
[0135] The computer-readable medium may be included in the electronic device, or may exist independently of the electronic device.
[0136] The computer-readable medium includes one or more programs, and when the one or more programs are executed by the electronic device, the electronic device performs the following operations: displaying media content, including images and / or videos, in a target media work on a preset page of a current application; determining at least one target media content from the media content; and generating at least one target expression object based on the at least one target media content in response to an expression object generation command for the at least one target media content, wherein the at least one target expression object is placed on an expression selection panel of the current application.
[0137] Computer program code for carrying out the operations of the present disclosure may be written in one or more programming languages, or a combination thereof. Such programming languages include, but are not limited to, object-oriented programming languages such as Java, Smalltalk, and C++, as well as conventional procedural programming languages, including "C" or similar programming languages. The program code may run entirely on the user's computer, partially on the user's computer, as a separate software package, partially on the user's computer and partially on a remote computer, or entirely on a remote computer or server. When remote computers are involved, the remote computers may be connected to the user's computer via any type of network (including a LAN or WAN) or may be connected to an external computer (e.g., via the Internet using an Internet Service Provider).
[0138] The flowcharts and block diagrams in the figures illustrate possible system architectures, functions, and operations of systems, methods, and computer program products according to various embodiments of the present disclosure. In this regard, each block in the flowcharts or block diagrams may represent a module, program segment, or portion of code, including one or more executable instructions for implementing a given logical function. It should be noted that in some alternative implementations, the functions noted in the blocks may occur in an order different from that noted in the figures. For example, two blocks shown in succession may actually be executed essentially in parallel, or may be executed in the reverse order, as determined by such functionality. It should be noted that each block in the block diagrams and / or flowcharts, and combinations of blocks in the block diagrams and / or flowcharts, may be implemented in a system using dedicated hardware that performs a given function or operation, or may be implemented using a combination of dedicated hardware and computer instructions.
[0139] The units mentioned in the embodiments of the present disclosure may be implemented in a software manner or a hardware manner. The names of modules may not necessarily be limited to the modules themselves. For example, a target content determination module may be described as "a module for determining at least one target media content in the media content."
[0140] The functions described herein above may be performed, at least in part, by one or more hardware logic components. For example, and without limitation, exemplary types of hardware logic components that may be used include Field Programmable Gate Arrays (FPGAs), Application Specific Integrated Circuits (ASICs), Application Specific Standard Products (ASICs), and the like. This includes ASSPs (Assembled Specific Standard Parts), System on Chip (SOC), and Complex Programmable Logic Devices (CPLD).
[0141] In the context of this disclosure, a machine-readable medium may be a tangible medium containing or storing a program for use in or in connection with an instruction execution system, apparatus, or device. A machine-readable medium may be a machine-readable signal medium or a machine-readable storage medium. A machine-readable medium includes, but is not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any suitable combination thereof. More specific examples of machine-readable storage media include one or more wire-based electrical connections, a portable computer disk, a hard disk, RAM, ROM, EPROM, or flash memory, optical fiber, a portable CD-ROM, an optical storage device, a magnetic storage device, or any suitable combination of the above.
[0142] According to one or more embodiments of the present disclosure, a method for processing media content is provided, comprising: displaying media content, including images and / or videos, in a target media work on a preset page of a current application; determining at least one target media content from the media content; and generating at least one target expression object based on the at least one target media content in response to an expression object generation command for the at least one target media content, wherein the at least one target expression object is arranged on an expression selection panel of the current application.
[0143] According to one or more embodiments of the present disclosure, determining at least one target media content in the media content includes determining at least one selected media content as at least one target media content in response to a selection operation on the media content.
[0144] According to one or more embodiments of the present disclosure, the target media work includes a target image work including at least one image, displaying media content in the target media work on a preset page of the current application includes displaying at least one image of the target image work on a preset page of the current application, and determining the selected at least one media content as at least one target media content in response to a selection operation on the media content includes determining the selected at least one image as at least one target image in response to a selection operation on the at least one image.
[0145] According to one or more embodiments of the present disclosure, the at least one target image composition includes a dynamic image and / or a still image.
[0146] According to one or more embodiments of the present disclosure, generating at least one target expression object based on the target media content includes generating at least one target expression object based on the at least one target image, and one target expression object includes one or more target images of the at least one target image.
[0147] According to one or more embodiments of the present disclosure, the target media work includes a target video work including a plurality of video frames, displaying media content in the target media work on a preset page of the current application includes displaying video progress information corresponding to the target video work on a preset page of the current application, and determining at least one selected media content as at least one target media content in response to a selection operation on the media content includes determining at least one target video frame set in response to a video frame selection operation on the video progress information, each target video frame set including at least one video frame.
[0148] According to one or more embodiments of the present disclosure, generating at least one target expression object based on the at least one target media content includes generating at least one target expression object based on the at least one target video frame set, and one target expression object includes one or more target video frame sets of the at least one target video frame set.
[0149] According to one or more embodiments of the present disclosure, the at least one target expression object includes a dynamic expression object, and the dynamic expression object is generated from at least one of a dynamic image in the at least one target image, a plurality of still images in the at least one target image, and a target video frame set in the at least one target video frame set.
[0150] According to one or more embodiments of the present disclosure, determining at least one target video frame set in response to a video frame selection operation on the video progress information includes determining a start video frame and an end video frame in response to a start frame selection operation and an end frame selection operation on the video progress information, and determining at least one target video frame set based on the start video frame and the end video frame.
[0151] According to one or more embodiments of the present disclosure, before generating at least one target expression object based on the at least one target media content, the method further includes generating at least one preview expression object based on the at least one target media content and displaying the at least one preview expression object.
[0152] According to one or more embodiments of the present disclosure, the method further includes receiving an editing operation for the at least one preview expression object, and generating at least one target expression object based on the at least one target media content includes generating at least one target expression object based on the at least one target media content and an editing result of the editing operation.
[0153] According to one or more embodiments of the present disclosure, before displaying media content in the target media work on a preset page of the current application, the method further includes displaying the preset page on a target display page of the target media work in the current application in response to a first preset trigger operation on the preset page.
[0154] According to one or more embodiments of the present disclosure, in the process of displaying the target media work on the target display page, the process further includes: displaying a preset control display area on the target display page in response to a second preset trigger operation; determining whether the target media work has a preset indicator; and if the target media work has the preset indicator, displaying a preset entrance of the preset page at a first preset display position in the preset control display area.
[0155] According to one or more embodiments of the present disclosure, after determining whether the target media work has a preset indicator, if the target media work does not have the preset indicator, the method further includes displaying a preset entrance of the preset page at a second preset display position in the preset control display area, and the display priority of the first preset display position is higher than the display priority of the second preset display position.
[0156] According to one or more embodiments of the present disclosure, there is provided a media content processing device including: a media content display module configured to display media content, including images and / or videos, in a target media work on a preset page of a current application; a target content determination module configured to determine at least one target media content in the media content; and an expression object generation module configured to generate at least one target expression object based on the at least one target media content in response to an expression object generation command for the at least one target media content, wherein the at least one target expression object is arranged on an expression selection panel of the current application.
[0157] According to one or more embodiments of the present disclosure, there is further provided an electronic device comprising one or more processors and a storage device configured to store one or more programs, the one or more programs, when executed by the one or more processors, causing the one or more processors to implement a method of media content processing according to an embodiment of the present disclosure.
[0158] According to one or more embodiments of the present disclosure, there is further provided a storage medium including computer-executable instructions that, when executed by a computer processor, perform a method for media content processing according to an embodiment of the present disclosure.
[0159] Although operations are described in a particular order, it should not be understood that these operations require that they be performed in the particular order or chronological order shown. In certain environments, multitasking and parallel processing may be advantageous. Similarly, although the above includes specific implementation details, these should not be construed as limiting the scope of the disclosure. Some features that are described in the context of a single embodiment may also be implemented in combination in a single embodiment. Conversely, various features that are described in the context of a single embodiment may also be implemented in multiple embodiments alone or in any suitable subcombination.
[0160] Although the present subject matter has been described in language specific to structural features and / or methodological acts, it should be understood that claimed subject matter is not necessarily limited to the particular features or acts described above. The particular features and acts so described merely implement example forms of the claims.
Claims
1. 1. A method of media content processing, comprising: Displaying one or more media contents, including images and / or videos, of a target media work on a preset page of a current application, the target media work being a work distributed to the current application by a distributor; determining at least one target media content from the media content; and generating at least one target expression object based on the at least one target media content in response to an expression object generation command for the at least one target media content; The at least one target facial expression target is placed in a facial expression selection panel of the current application; A method in which attribute information or work tags of the target media work are set by the distributor, and the attribute information is used to indicate whether or not the media content of the target media work is allowed to be used to generate the target expression object.
2. 2. The method of claim 1, wherein determining at least one target media content from the media content comprises determining at least one selected media content as the at least one target media content in response to a selection operation on the media content.
3. the target media production includes a target image production including at least one image; Displaying media content in the target media work on a preset page of the current application includes displaying at least one image of the target media work on a preset page of the current application; 3. The method of claim 2, wherein determining at least one selected media content as at least one target media content in response to a selection operation on the media content comprises determining at least one selected image as at least one target image in response to a selection operation on the at least one image.
4. The method of claim 3 , wherein at least one of the target image compositions comprises a dynamic image and / or a still image.
5. generating at least one target expression object based on the target media content includes generating at least one target expression object based on the at least one target image; The method according to claim 3 , wherein one target expression target includes one or more target images among the at least one target image.
6. the target media work comprises a target video work including a plurality of video frames; Displaying the media content of the target media work on the preset page of the current application includes displaying video progress information corresponding to the target video work on the preset page of the current application; determining at least one selected media content as at least one target media content in response to a selection operation on the media content includes determining at least one target video frame set in response to a video frame selection operation on the video progress information; The method of claim 4 , wherein each target video frame set includes at least one video frame.
7. generating at least one target expression object based on the at least one target media content includes generating at least one target expression object based on the at least one target video frame set; The method of claim 6 , wherein one target expression target includes one or more target video frame sets of the at least one target video frame set.
8. the at least one target expression object includes a dynamic expression object; 7. The method of claim 6, wherein the dynamic expression target is generated from at least one of a dynamic image in the at least one target image, a plurality of still images in the at least one target image, and a target video frame set in the at least one target video frame set.
9. Determining at least one target video frame set in response to a video frame selection operation on the video progress information includes: determining a start video frame and an end video frame in response to a start frame selection operation and an end frame selection operation on the video progress information; and determining at least one target video frame set based on the starting video frame and the ending video frame.
10. 2. The method of claim 1, further comprising, before generating at least one target expression object based on the at least one target media content, generating at least one preview expression object based on the at least one target media content and displaying the at least one preview expression object.
11. receiving an edit operation for the at least one preview expression object; The method of claim 10 , wherein generating at least one target facial expression object based on the at least one target media content comprises generating at least one target facial expression object based on the at least one target media content and an editing result of the editing operation.
12. 2. The method of claim 1, further comprising: displaying a preset page in a target display page of a target media work in a current application in response to a first preset trigger operation on the preset page before displaying media content in the target media work in a preset page of the current application.
13. displaying a preset control display area on the target display page in response to a second preset trigger operation during the step of displaying the target media piece on the target display page; determining whether the target media piece is marked with a preset tag; 13. The method of claim 12, further comprising: displaying a preset entrance of the preset page in a first preset display position in the preset control display area in response to determining that the target media piece is marked with the preset indicator.
14. and after determining whether the target media work has a preset indicator, in response to a determination result that the target media work does not have the preset indicator, displaying a preset entrance of the preset page at a second preset display position in the preset control display area; The method of claim 13 , wherein a display priority of the first preset display position is higher than a display priority of the second preset display position.
15. 1. A media content processing device, comprising: a media content display module configured to display one or more media contents, including images and / or videos, of a target media work on a preset page of a current application, the target media work being a work distributed to the current application by a distributor; a target content determination module configured to determine at least one target media content from the media content; an expression object generation module configured to generate at least one target expression object based on the at least one target media content in response to an expression object generation command for the at least one target media content; The at least one target facial expression target is placed in a facial expression selection panel of the current application; The attribute information or work tag of the target media work is set by the distributor, and the attribute information is used to indicate whether or not the media content of the target media work is allowed to be used to generate the target expression object.
16. Apparatus according to claim 15, comprising a module for carrying out the method according to any one of claims 2 to 14.
17. at least one processor; a storage device configured to store at least one program; An electronic device, wherein said at least one program, when executed by said at least one processor, causes said at least one processor to implement the method of any one of claims 1 to 14.
18. A storage medium comprising computer-executable instructions for performing the method of any one of claims 1 to 14 when executed by a computer processor.
19. A computer program comprising: A computer program which, when executed by a processor, causes the processor to carry out the method of any one of claims 1 to 14.
Citation Information
Patent Citations
Expression generation method and device, computer equipment and storage medium
CN114693827A