Method, apparatus, device, storage medium and program product for media content processing

CN122601946APending Publication Date: 2026-08-18BEIJING ZITIAO NETWORK TECH CO LTD
View PDF 0 Cites 0 Cited by

Patent Information

Application Number
CN202610813197.3
Authority / Receiving Office
CN · China
Patent Type
Applications(China)
Current Assignee / Owner
Filing Date
2026-06-05
Publication Date
2026-08-18

AI Technical Summary

Benefits of technology

[0010] This approach allows media content applying the first effect to display the first part of the first information while hiding the second part before publication, and to display the complete first information after publication. On one hand, this allows users to obtain the complete first information after publication without additional action, providing a phased interactive method for question-and-answer or result-based gameplay. On the other hand, since the display control of the first information is uniformly driven by configuration information corresponding to the effect package, it reduces storage overhead during cross-stage processing of media content, reduces the number of data transfers required for state synchronization, and reduces the computational overhead of compositing during the publication stage.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN122601946A_ABST
    Figure CN122601946A_ABST
Patent Text Reader

Abstract

Methods, apparatuses, devices, storage media, and program products for media content processing are provided. The method includes presenting a first media content, wherein the first media content presents a first portion of first information and does not present a second portion of the first information; receiving a first operation, the first operation indicating to publish the first media content; and presenting a second media content, the second media content being obtained by publishing the first media content, the second media content presenting the first portion and the second portion of the first information.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] The examples in this article generally relate to the field of computers, and in particular to methods, apparatuses, electronic devices, computer-readable storage media, and computer program products for media content processing. Background Technology

[0002] With the rapid development of computer technology, more and more applications and platforms are designed to provide users with various services. For example, users can edit, create, and publish various types of media content, such as media content with special effects, in applications that provide media interaction services. Media content can also be referred to as media, multimedia, media items, etc. Media content can include one or more of various types such as video, images, image sets, text, audio, and animation. Summary of the Invention

[0003] In a first aspect, a media content processing method is provided. The method includes: presenting first media content, wherein the first media content displays a first portion of first information and does not display a second portion of the first information; receiving a first operation, the first operation instructing the publication of the first media content; and presenting second media content, the second media content being obtained by publishing the first media content, the second media content displaying the first and second portions of the first information.

[0004] In a second aspect, a method for generating special effects packages is provided. The method includes: presenting a first interface for configuring special effects packages, the special effects packages corresponding to a first special effect; receiving first configuration information for enabling media content applying the first special effect to: display a first portion of the first information without displaying a second portion of the first information before publication, and display both the first and second portions after publication; and adding the first configuration information to the special effects package.

[0005] In a third aspect, a media content processing apparatus is provided. The apparatus includes: a first media presentation module configured to present first media content, wherein the first media content displays a first portion of first information and does not display a second portion of the first information; a first operation receiving module configured to receive a first operation instructing the publication of the first media content; and a second media presentation module configured to present second media content obtained by publishing the first media content, wherein the second media content displays both the first and second portions of the first information.

[0006] In a fourth aspect, a special effects package generation apparatus is provided. The apparatus includes: a first interface presentation module configured to present a first interface for configuring a special effects package, the special effects package corresponding to a first special effect; a configuration information receiving module configured to receive first configuration information, the first configuration information being used to cause media content applying the first special effect to: display a first part of the first information before publication but not a second part of the first information, and display both the first part and the second part after publication; and a configuration information adding module configured to add the first configuration information to the special effects package.

[0007] In a fifth aspect, an electronic device is provided. The device includes at least one processor; and at least one memory coupled to the at least one processor and storing instructions for execution by the at least one processor. When executed by the at least one processor, the instructions cause the device to perform the methods of the first or second aspect.

[0008] In a sixth aspect, a computer-readable storage medium is provided. The computer-readable storage medium stores computer-executable instructions that can be executed by a processor to implement the methods of the first or second aspect.

[0009] In a seventh aspect, a computer program product is provided, which is tangibly stored in a computer storage medium and includes computer-executable instructions that, when executed by a device, cause the device to perform the method of the first aspect or the second aspect.

[0010] This approach allows media content applying the first effect to display the first part of the first information while hiding the second part before publication, and to display the complete first information after publication. On one hand, this allows users to obtain the complete first information after publication without additional action, providing a phased interactive method for question-and-answer or result-based gameplay. On the other hand, since the display control of the first information is uniformly driven by configuration information corresponding to the effect package, it reduces storage overhead during cross-stage processing of media content, reduces the number of data transfers required for state synchronization, and reduces the computational overhead of compositing during the publication stage.

[0011] It should be understood that the content described in this section is not intended to limit the key or important features of the examples in this article, nor is it intended to restrict the scope of the solution. Other features will become readily apparent from the following description. Attached Figure Description

[0012] The above and other features, advantages, and aspects of the various examples herein will become more apparent when taken in conjunction with the accompanying drawings and the following detailed description. In the accompanying drawings, the same or similar reference numerals denote the same or similar elements, wherein: Figure 1 A schematic diagram of an example environment based on some scenarios is shown; Figures 2A to 2F Example interfaces for media content processing based on various scenarios are shown; Figures 3A to 3C Example interfaces generated based on effects packs for various scenarios are shown; Figure 4 A schematic diagram illustrating the signaling flow for media content processing under certain circumstances is shown. Figure 5 A flowchart illustrating example methods for media content processing under several scenarios is shown; Figure 6 The flowchart illustrates example methods for generating special effects packages based on several scenarios; Figure 7 A schematic structural block diagram of an example device for media content processing is shown, based on some scenarios. Figure 8 A schematic structural block diagram of an example apparatus for generating special effects packs under certain scenarios is shown; and Figure 9 A block diagram of an electronic device capable of implementing multiple illustrative scenarios is shown. Detailed Implementation

[0013] The examples in this document will now be described in more detail with reference to the accompanying drawings. While some examples are shown in the drawings, it should be understood that solutions can be implemented in various forms and should not be construed as limited to the examples presented herein. Rather, these examples are provided to provide a more thorough and complete understanding of the solutions. It should be understood that the drawings and examples in this document are for illustrative purposes only and are not intended to limit the scope of protection of the solutions.

[0014] In the description of the examples in this document, the term "including" and similar terms should be understood as open inclusion, i.e., "including but not limited to". The term "based on" should be understood as "at least partially based on". The term "an example" or "the example" should be understood as "at least one example". The term "some examples" should be understood as "at least some examples". Other explicit and implicit definitions may also be included below. The terms "first", "second", etc., may refer to different or the same objects. Other explicit and implicit definitions may also be included below.

[0015] It should be noted that, unless explicitly stated otherwise, performing a step in response to A does not mean that the step is performed immediately after A, but may include one or more intermediate steps.

[0016] The examples in this document may involve user data, data acquisition, and / or use. All of these aspects comply with relevant laws, regulations, and rules. In the examples, all data collection, acquisition, processing, manipulation, forwarding, and use are conducted with the user's knowledge and confirmation. Accordingly, when implementing each example, the type, scope of use, and usage scenarios of any data or information that may be involved should be communicated to the user and their authorization obtained through appropriate means, in accordance with relevant laws and regulations. The specific methods of notification and / or authorization can vary depending on the actual situation and application scenario; the scope of the solution is not limited in this regard.

[0017] In this manual and the sample solutions, any processing of personal information will be conducted only under legal grounds (such as obtaining the consent of the data subject or being necessary for the performance of a contract) and will only be carried out within the scope stipulated or agreed upon. A user's refusal to process personal information beyond what is necessary for basic functions will not affect the user's use of basic functions.

[0018] As used in this article, "media content" refers to content processed or presented by a computer, which may include, but is not limited to, video content, image content, audio content, or a combination thereof. In some examples, media content may be video generated by applying effects to captured footage.

[0019] As used in this article, the term "effects" refers to processing that can be applied to captured visuals or media content to alter its presentation, such as, but not limited to, stickers, filters, interactive features, or combinations thereof. The term "effects package" as used in this article refers to a data set that carries resources and / or configuration information related to effects, typically including, but not limited to, the material resources used by the effects, configuration information, and calling logic.

[0020] The term "Supplemental Enhancement Information (SEI)" as used in this document refers to metadata that can be transmitted along with the video stream to carry additional information and typically does not directly affect the video decoding process. SEI can be, for example, supplemental enhancement information as defined in video coding standards such as H.264 and H.265. In some examples, SEI can be used to carry business semantics related to occlusion.

[0021] The term “question-answering process” as used in this article refers to an interactive process presented in media content that includes at least one question and corresponding candidate answers.

[0022] As mentioned above, users can edit, create, and publish various types of media content and / or effects within the application. In media content creation and sharing scenarios, it's often desirable to control the display of information carried by the media content, ensuring that some information is displayed at certain times and not at others. There are several possible implementation methods for controlling the phased display of information carried by media content. Some solutions involve directly writing the occlusion effect into the pixels of the media content to obtain occluded media content during the shooting or editing stages, and then attempting to restore or regenerate unoccluded media content after publication. However, once the occlusion effect is written into the pixels, the subsequent restoration of the original content becomes complex, often requiring additional saving of the original content, proxy content, or restoration data to generate unoccluded media content based on the saved content. Furthermore, such solutions typically require maintaining multiple sets of materials or multiple streams of media content between different stages such as shooting, editing, drafting, and publishing, making cross-stage state synchronization difficult. In other related solutions, the terminal device maintains occlusion resources, timing information, and display logic separately, and transmits and manages business information such as occlusion area, appearance timing, and prompts between different interfaces such as shooting, editing, and drafting. In such solutions, occlusion-related business information is separated from the media content itself, making it prone to loss during cross-interface and cross-stage transmission. Furthermore, additional adaptation may be required during editing processes such as overlaying speed changes, cropping, splitting, and applying special effects, making the process more complex.

[0023] In view of this, an improved media content processing scheme is proposed. According to this scheme, on the consumer user side, the terminal device presents first media content, which displays the first part of the first information but hides the second part; the terminal device receives a first operation instructing the publication of the first media content; and the terminal device presents second media content obtained by publishing the first media content, which displays both the first and second parts of the first information. This allows the media content to present different display states before and after publication—that is, displaying the first part of the first information while hiding the second part before publication, and displaying the complete first information after publication. This allows users to obtain the complete first information after publication without additional action, thus providing a phased display interaction method for question-and-answer or result-based gameplay.

[0024] An improved scheme for generating special effects packages is also proposed. According to this scheme, on the effects creator's side, the terminal device presents a first interface for configuring the special effects package, which corresponds to a first special effect. The terminal device receives first configuration information, which is used to ensure that media content applying the first special effect displays the first part of the first information but not the second part before publication, and displays both the first and second parts after publication. The terminal device adds the first configuration information to the special effects package. Thus, since the display control of the first information is uniformly driven by the configuration information corresponding to the special effects package, the first media content shares the same underlying content carrier across the creation, editing, and publication stages. Compared to schemes that maintain multiple media contents before and after occlusion or transmit business states separately between multiple stages, the proposed scheme reduces the storage overhead of media content during cross-stage processing, reduces the number of data transfers required for state synchronization across stages and interfaces, and reduces the compositing computation overhead during the publication stage, thereby improving the stability of the processing chain.

[0025] The following describes various examples of this scheme in further detail with reference to the accompanying drawings.

[0026] Figure 1 A schematic diagram of an example environment 100 based on some scenarios is shown. For example... Figure 1 As shown, the example environment 100 may include a terminal device 120 on the special effects creator side, a server 130, and a terminal device 150 on the consumer user side (which may include terminal devices 150-1, 150-2, ..., 150-N, where N is a positive integer, and these one or more terminal devices may be referred to individually or collectively as terminal device 150).

[0027] In this example environment 100, terminal devices 120 and 150 can run applications that support media content processing and / or effects package generation. The application can be any suitable type of application for media content processing and / or effects package generation, including but not limited to multimedia applications, video applications, or other suitable applications with shooting, editing, and publishing functions.

[0028] User 110 can be the creator of the special effects package (i.e., the special effects maker). User 110 can interact with the special effects creation application running on terminal device 120 via terminal device 120 and / or its attached devices to configure and generate the special effects package, and publish the special effects package to the media content application. User 140 can be a consumer user (i.e., the consumer user) who uses the special effects corresponding to the special effects package to create videos. Users 140-1, 140-2, ..., 140-N can correspond to terminal devices 150-1, 150-2, ..., 150-N respectively, and one or more users can be referred to individually or collectively as User 140. User 140 can also interact with the media content application running on terminal device 150 via the corresponding terminal device 150 and / or its attached devices to create, edit, and publish media content, and can use the special effects in the content application to generate media content. In this document, the special effects creation application and the media content application can be two independent applications, or two business modules of the same application.

[0029] It should be noted that user 110 can also apply special effects created by other users, and user 140 can also create special effects. This article only uses the example of special effects created by user 110 and consumed by user 140 for illustrative purposes.

[0030] In some cases, if the application is active, terminal device 120 can present an interface for configuring effects packages through the application, and terminal device 150 can present an interface for media content processing through the application.

[0031] In some scenarios, terminal devices 120 and 150 communicate with server 130 to provide services for the application. In other scenarios, terminal devices 120 and / or 150 may communicate with server 130 via network 132. Server 130 may also provide functions such as application management, configuration, and maintenance. Server 130 may also store data such as special effects packages and media content.

[0032] Terminal devices 120 and 150 can be any type of mobile terminal, fixed terminal, or portable terminal, including mobile phones, desktop computers, laptop computers, notebook computers, netbook computers, tablet computers, media computers, multimedia tablets, handheld computers, portable gaming terminals, VR / AR devices, personal communication system (PCS) devices, personal navigation devices, personal digital assistants (PDAs), audio / video players, digital cameras / camcorders, positioning devices, television receivers, radio receivers, e-book devices, gaming devices, or any combination thereof, including accessories and peripherals of these devices or any combination thereof. In some cases, terminal devices 120 and 150 may also support any type of user-facing interface (such as "wearable" circuitry).

[0033] Server 130 can be a standalone physical server, a server cluster or distributed system composed of multiple physical servers, or a cloud server providing basic cloud computing services such as cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communication, middleware services, domain name services, security services, content delivery networks, and big data and artificial intelligence platforms. Server 130 may include, for example, computing systems / servers such as mainframes, edge computing nodes, computing devices in a cloud environment, etc. Server 130 can provide backend services for applications in terminal devices 120 and 150 that support media content processing and / or special effects package generation.

[0034] A communication connection can be established between server 130 and terminal device 120 and / or terminal device 150. This communication connection can be established via wired or wireless means. The communication connection can include, but is not limited to, Bluetooth, mobile network, Universal Serial Bus (USB), and Wireless Fidelity (WiFi) connections. In some cases, server 130 and terminal device 120 and / or terminal device 150 can exchange signaling signals through the corresponding communication connection.

[0035] It should be understood that the structure and function of the various elements in environment 100 are described for illustrative purposes only and do not imply any limitation on the scope of the scheme. Example environment 100 is used to illustrate the application environment of the scheme, and users 110 and 140 are used to distinguish between the two types of operating subjects: the special effects creator side and the consuming user side.

[0036] The examples will continue to be described with reference to the accompanying figures. To better explain the examples in this article, further details will be provided below. Figures 2A to 2F Examples of media content processing, combined with Figures 3A to 3C This describes an example of how special effects packages are generated. Figures 2A to 3C The example interfaces 200A to 300C shown are merely examples; various interface designs are possible in practice. The graphic elements within the interface can have different arrangements and visual representations, one or more elements can be omitted or replaced, and one or more other elements may also be present. The examples in this document are not limited in this respect.

[0037] Figures 2A to 2F Example interfaces 200A to 200F are shown for media content processing according to certain scenarios. Example interfaces 200A to 200F can, for example, be handled by... Figure 1 The terminal device shown is provided by 150. See below for reference. Figures 2A to 2F This describes the process of processing media content.

[0038] First, during the shooting, editing, and drafting stages of the media content, terminal device 150 can present the first media content. The first media content displays a first portion of the first information but does not display a second portion of the first information. In some cases, the first media content may be media content obtained based on a first special effect. In some examples, the first media content includes a second portion of the first information, but terminal device 150 obscures the second portion by overlaying a visual element on top of it, thus preventing the second portion from being displayed. The term "visual element" refers to an element that can be overlaid on the second portion to obscure it, such as, but not limited to, obscuring images, color blocks, blurred areas, or other suitable forms of elements. Such visual elements can have any suitable visual style; for example, they may include mosaics. Such visual elements may also be referred to as obscuring elements. In other examples, the first media content may not include the second portion of the first information, and terminal device 150 can display the first media content normally. Terminal device 150 can receive a first operation (which may also be referred to as a publishing operation) instructing the publication of the first media content and publish the first media content. The media content obtained from publishing the first media content can be referred to as the second media content. The terminal device 150 can display the second media content, which displays the first part and the second part of the first information.

[0039] refer to Figure 2ATerminal device 150 can present an example interface 200A during the media content shooting phase. Terminal device 150 can present first media content 201 and recording control 202 in the example interface 200A. In some cases, terminal device 150 can present the captured image in the example interface 200A. Terminal device 150 can, in response to receiving a request to apply a first special effect, load the special effect resource corresponding to the first special effect and apply the first special effect to the captured image. Terminal device 150 can then, in response to a second operation instructing the shooting of media content, shoot the image with the first special effect applied to obtain the first media content 201. As an example only, the second operation may include a touch operation on the recording control 202. In some cases, the first media content can be obtained based on the content presented on the shooting interface during shooting. For example, the content presented on the shooting interface during shooting is encoded as the first media content.

[0040] In some examples, the first effect may include a question-and-answer process. In this case, the first media content may include the content of the question-and-answer process and the corresponding result information (e.g., a result overview), result details, etc. The terminal device 150 may, in response to a request to apply the first effect, overlay at least one first question and at least one candidate answer for each first question onto the image captured by the camera, based on the first effect. For example, the terminal device 150 may sequentially present at least one first question and the candidate answer associated with each first question. As an example only, at least one first question may include... Figure 2A The question 210 shown can include candidate answer 211 and candidate answer 212, for example.

[0041] In some cases, terminal device 150 can receive selection information indicating the selection of at least one of candidate answers 211 and 212. For example, this selection information may include the posture information of an object in the captured image. This posture information may be the object's hand posture information or the object's overall posture information, and terminal device 150 can determine the answer selected by the user based on the object's posture information. For example, terminal device 150 can determine the answer selected by the user from at least one candidate answer based on the direction of the object's fingers and the direction of the object's body. As another example, the selection information may include touch operation information on the screen of terminal device 150, which may include any appropriate operation such as clicking, long pressing, or hovering. Terminal device 150 can determine the answer touched by the user as the answer selected by the user. As yet another example, the selection information may include operation information on terminal device controls of terminal device 150, which may be physical controls (also called hardware controls). For example, device controls may be buttons on terminal device 150. Different answers may correspond to different buttons, and terminal device 150 can determine the corresponding answer based on the button pressed by the user. For example, the selection information may also include voice input information, and the terminal device 150 may perform semantic analysis on the voice input information to determine the answer selected by the user.

[0042] In this way, users can be provided with multiple answer selection options, allowing them to select answers without interrupting the shooting process. At the same time, since answer selection is completed on-site during shooting through methods such as recognizing the pose of objects in the captured image, detecting screen touch, or recognizing voice, compared to jumping to a separate selection interface and then returning to shooting, the number of interface switching during shooting can be reduced, thereby reducing the page rendering overhead and the computational overhead required to reload the captured image when the terminal device 150 switches interfaces.

[0043] Furthermore, the terminal device 150 can also present the first part of the first information based on the first special effect. In some examples, the first information may include result information corresponding to the question-and-answer process, which may be determined based on at least one first question in the question-and-answer process and the answer selected by the user for at least one first question. In some examples, the first information may be result details information corresponding to the result information, which may include an explanation and analysis of the result information. It is understood that both the result information and the result details information can be pre-configured by the creator of the first special effect, who may, for example, pre-configure multiple result information and result details information corresponding to each result information for the question-and-answer process. In this way, the presentation of the question-and-answer process and the first part can be uniformly driven by the first special effect, reducing the additional operations performed by the user to obtain the question and answer and the first part; at the same time, since the presentation of the question and answer and the first part is uniformly implemented through the same first special effect, compared with loading and driving each presentation content separately, the number of processing logics that the terminal device 150 needs to load and schedule can be reduced, thereby reducing processing overhead.

[0044] like Figure 2B As shown, after the user selects an answer to at least one first question, terminal device 150 can present example interface 200B. Example interface 200B includes panel 220. Terminal device 150 can present result information corresponding to the question-and-answer process in panel 220. If the result information is first information, terminal device 150 can overlay visual elements on top of the second part of result 220 to obscure the second part, or it can choose not to present the second part. Terminal device 150 can, for example, respond to a trigger on panel 220, or respond to the presentation duration of panel 220 reaching an appropriate threshold, presenting... Figure 2C The example interface 200C is shown. Example interface 200C includes panel 230, as shown... Figure 2C The result details information shown is displayed on panel 230. This document uses the result details information as the first piece of information as an example for illustrative purposes. Terminal device 150 may display the first part of the result details information but not the second part. For example... Figure 2C As shown, the terminal device 150 can overlay visual elements (such as visual elements 241, 242, and 243 shown in the figure) onto the result details information presented on the panel 230, so that the first part of the result details 230 is displayed while the second part is not displayed.

[0045] In some cases, terminal device 150 may overlay a visual element onto the second part of the first information, this visual element being used to occlude the entire content of the second part. In some examples, the second part may include multiple sub-parts, and terminal device 150 may overlay visual elements onto each of the multiple sub-parts, that is, terminal device 150 may present multiple visual elements. The multiple sub-parts of the second part may include at least two first sub-parts that are in the same video frame (i.e., in the same time interval) but in different positions on the screen. In this case, terminal device 150 may simultaneously present multiple visual elements in the video frame to occlude at least two first sub-parts respectively. Figure 2C As shown, the terminal device 150 can display visual elements 241, 242, and 243 to respectively occlude three first sub-parts located in different areas of the same screen. In this way, multiple areas of the result details 230 can be occluded separately in the same screen.

[0046] In other cases, the multiple sub-parts of the second part may include at least two second sub-parts corresponding to different time intervals of the first media content 201 (i.e., these at least two sub-parts may be located in different video frames). In this case, the terminal device 150 may overlay visual elements on different video frames to obscure at least two second sub-parts respectively. In this way, the display of the first information can be segmented and controlled in the time dimension, allowing users to see the corresponding first part in different time intervals before publication.

[0047] In some cases, prior to receiving the first operation, the terminal device 150 may present a prompt message instructing the presentation of the complete first information after the first media content 201 is published. The prompt message may include any appropriate content, including but not limited to text, symbols, images, icons, etc. For example, see reference... Figure 2C Terminal device 150 can display a prompt message 250 on the example interface 200C, which may include the text "Hidden content can be unlocked after submission". In this way, users can be prompted before publication that they will be able to obtain complete first information after publication, so that users do not need to repeatedly search between different interfaces to learn about the display rules.

[0048] Terminal device 150 can display an editing interface for the first media content in response to the completion of its acquisition. Terminal device 150 can also determine that the acquisition of the first media content is complete in response to receiving another touch operation on the recording control 202, or in response to the duration of the first media content (i.e., the recording duration) reaching a predetermined duration. (Reference) Figure 2DExample interface 200D illustrates an example of an editing interface. In example interface 200D, terminal device 150 can present first media content 201, which displays a first part of first information (e.g., result details information) and does not display a second part. As shown, the first media content 201 may include a panel 230, which includes result details information. Terminal device 150 can obscure the second part of the result details information by overlaying visual elements 241, 242, and 243 on top of it.

[0049] The example interface 200D may also include multiple editing controls (such as music, settings, clips, text, filters, etc.) for editing the first media content 201, and a control 203 for instructing the publication of the first media content 201. In some cases, the terminal device 150 may respond to receiving an editing operation on the first media content in the editing interface and present the edited first media content in the editing interface. The edited first media content still displays the first part and does not display the second part. In this way, the second part of the first information can be kept occluded during the editing stage, so that the user can maintain a consistent display state without having to repeatedly set the occlusion during the editing process. At the same time, since the occlusion display during the editing stage is uniformly driven by the configuration information bound to the first media content 201, rather than inserting an independent occlusion track or maintaining an occlusion timeline separately on the editing side, the number of tracks that need to be maintained during the editing stage can be reduced compared to maintaining occlusion resources separately on the editing side, thereby reducing the memory usage and rendering pipeline processing overhead of the terminal device 150 during the editing stage. In the example interface 200D, the terminal device 150 can still display a prompt message 250 to remind the user to publish the first media content in order to view the complete first information.

[0050] In some examples, terminal device 150 may respond to a touch operation on control 203, determine that a first operation has been received, and then publish the first media content. In some examples, during the shooting or editing stage, terminal device 150 may respond to a save operation on the first media content, saving the first media content as a first draft. Terminal device 150 may, for example, present the first draft in a drafts folder. Terminal device 150 may respond to a viewing operation on the first draft, presenting an editing interface for the first draft, which is also the editing interface for the first media content (e.g., example interface 200D).

[0051] Terminal device 150 can, in response to the successful publication of the first media content, determine second media content based on the first media content and present the second media content. Terminal device 110 can, for example, display a content presentation interface and present the published media content within that interface. (See reference) Figure 2E and Figure 2F Terminal device 110 can, for example, respond to the publication of first media content by presenting example interfaces 200E and 200F. Terminal device 110 can also present second media content 205 in example interfaces 200E and 200F, which is obtained by publishing first media content 201. In example interfaces 200E and 200F, terminal device 150 can present a progress bar 204. The progress bar 204 can not only indicate the current playback progress and total duration of the second media content 205, but can also switch to continue playing the second media content at the playback progress indicated by the user interaction based on user interaction. Taking the first special effect indicating a question-and-answer process as an example, in the second media content 205, the answer selected by the user (e.g., candidate answer 211) can be highlighted (e.g., by adding an underline, highlighting, bolding a border, or any other appropriate form).

[0052] like Figure 2F As shown, in the example interface 200F, the second media content 205 can present the complete first information. That is, both the first and second parts of the result details information are displayed. The terminal device 150 can display the second part, for example, by canceling the presentation of visual elements, or by adding the second part to the first information that is missing a second part. In this way, the second media content 205 displays the complete first information after publication, allowing users and other users to directly see the complete first information when browsing the information stream, which helps to increase users' enthusiasm for publishing media content.

[0053] Figures 3A to 3C Example interfaces 300A to 300C, generated based on special effects packages for various scenarios, are shown. Example interfaces 300A to 300C can, for example, be generated by... Figure 1 The terminal device 120 shown is provided. See below for reference. Figures 3A to 3C This describes the process of configuring special effects packages on the effects creator's side.

[0054] Terminal device 120 can present a first interface for configuring an effects package, which may correspond to the aforementioned first effects. Terminal device 120 can receive first configuration information for the effects package via the first interface. This first configuration information may, for example, be used to cause media content applying the first effects to display a first portion of the first information but not a second portion before publication, and to display both the first and second portions after publication. In some examples, the first configuration information may include occlusion information indicating a visual element to be overlaid on the second portion to occlude it. In some examples, the first configuration information may directly include the second portion, indicating that the second portion should not be added to the media content before publication (this allows the media content to directly exclude the second portion before publication), and that the second portion should be added to the first information after publication.

[0055] In some examples, the first configuration information may indicate that the second part includes multiple sub-parts, which include at least two first sub-parts and / or at least two second sub-parts. The at least two first sub-parts may correspond to different areas of the same frame in the media content, and the at least two second sub-parts may correspond to different time intervals of the media content. That is, different first sub-parts may be located at different positions within the same video frame, and different second sub-parts may be located in different video frames.

[0056] In some examples, the terminal device 120 can also receive second configuration information for the effects package and add the second configuration information to the effects package. This second configuration information can instruct the display of a question-and-answer process in the media content. In some examples, the second configuration information can instruct at least one question involved in the question-and-answer process and candidate answers for each of the at least one question. In some examples, the second configuration information can instruct at least one result information corresponding to the question-and-answer process, where different result information can correspond to different answers. For example, if at least one question includes five questions, and each question includes candidate answers (such as candidate answer A and candidate answer B), then the second configuration information can configure different result information based on the combination of the answers corresponding to each of the five questions (the answer is the answer selected by the user from the candidate answers). For example, if the user selects candidate answer A for all five answers, the question-and-answer process can correspond to result information 1; if the user selects candidate answer B for all five answers, the question-and-answer process can correspond to result information 2, where result information 1 is different from result information 2. In some examples, the second configuration information can also instruct first information corresponding to the question-and-answer process. For example, the second configuration information can instruct the first information to be result information or result details information corresponding to the result information. In some examples, at least one question, candidate answer, result information, result details, etc., indicated by the second configuration information will be overlaid on the captured screen.

[0057] Figure 3A The example interface 300A shown illustrates an example of the first interface. For example... Figure 3A As shown, the example interface 300A may include area 310 and area 320. The terminal device 120 may present multiple editing controls for editing the effects package in area 310, including at least controls 311 and 312. In some examples, area 310 may include control 313, and the terminal device 120 may present more editing controls in response to touch operations on control 313. Area 320 may be used to preview effects 330. It is understood that when the effects indicate a question-and-answer process, the effects may include multiple video frames, and the terminal device 120 may present at least one question, result information, result details information, etc., in each of the multiple video frames. Taking result details information as the first information as an example, the example interface 300A may be used to configure the first information, and the effect 330 presented in the example interface 300A may correspond to a video frame including result details information 331. Of course, the terminal device 120 may also present a configuration interface for other content related to the effects, such as a configuration interface for questions and candidate answers in a question-and-answer process.

[0058] In some examples, terminal device 120 may respond to touch operations on control 311 in example interface 300A, and present Figure 3B The example interface 300B is shown. Example interface 300B may also include region 340. Configuration region 340 can be used to receive initial configuration information for visual elements (e.g., mosaic). Terminal device 120 can receive configurations for the shape (e.g., rounded rectangle, rectangle, ellipse, or other shapes), fill style (e.g., solid fill style), color, transparency (e.g., opaque or semi-transparent), and timing of appearance of visual elements via configuration region 340. In this way, multiple attributes of visual elements can be configured via configuration region 340, allowing effects creators to complete the setting of visual elements without manually drawing occlusions frame by frame. Furthermore, since visual elements are parameterized via configuration information, compared to pre-generating and storing independent occlusion materials for each occlusion style, the number of occlusion materials required to be stored in the effects package can be reduced, thereby reducing the storage overhead of the effects package.

[0059] In some examples, terminal device 120 can receive configurations of the position, size, etc. of visual elements via region 320. For example... Figure 3CAs shown, terminal device 120 can overlay visual element 332 on top of result details information 331 in region 320 based on configuration information received in region 340 and user interaction received in region 320, thereby obscuring the second part. In some cases, region 320 may include icon 333, which can indicate the location of the user's touch operation. Terminal device 120 can determine the overlay position and / or size of visual element 332 based on the position indicated by icon 333, for example. For example, if the user starts a long press at one position in result details information 331 and ends the long press at another position, terminal device 120 can construct a rectangle using the coordinates corresponding to these two positions, and determine the area where the rectangle is located as the area where the visual element is to be presented, thereby determining the size and style of the visual element based on the first configuration information.

[0060] In some cases, in response to receiving an operation on control 312, terminal device 120 generates an effects package based on the received configuration information (such as first configuration information, second configuration information, etc.), and the effects package contains the configuration information. In this way, the first configuration information is added to the effects package corresponding to the first effects, so that media content applying the effects package can display the first part of the first information without displaying the second part before publication, and display both the first and second parts after publication. The effects creator does not need to set the occlusion logic one by one on the consumer user side.

[0061] Figure 4 A schematic diagram of signaling flow 400 is shown, illustrating media content processing under certain conditions. See below for reference. Figure 4 This describes the process of shooting, editing, and publishing the first media content, involving the interaction between the terminal device 150, the media processing unit 401, and the special effects package 402. The special effects package 402 can be, for example, a special effects package created and published by the aforementioned terminal device 120. The media processing unit 401 is used to generate and edit media content. In some cases, the media processing unit 401 can be deployed on the server 130, or locally on the terminal device 150, or it can be implemented collaboratively by the server 130 and the terminal device 150. That is, the terminal device 150 can independently perform the processing of the first and second media content, the server 130 can perform some or all of the processing and return the processing results to the terminal device 150, or the terminal device 150 and the server 130 can share the corresponding processing; the scope of the solution is not limited in this respect.

[0062] During the shooting phase 403, the terminal device 150 sends (411) an effects package loading request and (412) a video recording request to the media processing unit 401. For example, the terminal device 110 may send an effects package loading request indicating the first effects to the media processing unit 401 in response to the selection of the first effects, and may send a video recording request in response to a touch operation on the recording control. The media processing unit 401 may send (413) an effects application request to the effects package 402 in response to receiving the effects package loading request to indicate the application of the first effects. The effects package 402 may return (414) relevant information about the effects to the media processing unit 401 based on the effects application request. The relevant information about the effects may include configuration information corresponding to the first effects, for example, it may include first configuration information (which may include occlusion information or a second part) and second configuration information. Taking the effects indicating a question-and-answer process as an example, the second configuration information may indicate at least one question involved in the question-and-answer process, candidate answers for each of the at least one question, and result information, result details information, etc., corresponding to the question-and-answer process. The first configuration information could indicate, for example, that the first information is result information or result details information, the first part and / or the second part of the first information, occlusion information, etc. The occlusion information could indicate the visual elements used to occlude the second part.

[0063] The media processing unit 401 can process the currently presented video frame based on the relevant information of the (415) special effects. For example, during the shooting stage, the media processing unit 401 can input the underlying video frame and the resources corresponding to the visual elements as two-way texture inputs to the graphics processing unit (GPU). The GPU performs a blending operation based on the transparency information of the visual elements, so that the transparent areas present the captured image and the occluded areas present the visual elements. Thus, the occlusion is only reflected in the preview rendering result without changing the original pixels of the underlying video. Of course, in some examples, the media processing unit 401 can be a GPU.

[0064] The media processing unit 401 can determine (416) the first media content and its supplemental enhancement information (SEI). In some cases, the media processing unit 401 can write occlusion information (which can indicate visual elements) or the second part of the first information (i.e., the part that is not displayed) into the supplemental enhancement information, so that the occlusion information or the second part is uniformly carried along with the first media content. Compared with the method of transmitting occlusion-related business information separately between multiple interfaces, this can reduce the number of data transmissions required for cross-interface and cross-stage transmissions and reduce state synchronization overhead.

[0065] The effects package 402 can also send a message (417) to the terminal device 150, which can then present a prompt (418) indicating that the complete first information will be presented after the first media content is published. In response to an operation indicating the end of recording (e.g., a touch operation on the recording control), the terminal device 150 sends a request to the media processing unit 401 (419) to end recording. The media processing unit 401 can then return (420) the first media content, including supplementary enhancement information, to the terminal device 150. For example, in response to receiving the first media content, the terminal device 150 can present (421) an editing interface for the first media content.

[0066] During the editing phase 404 of the first media content, the terminal device 150 may send (422) a request to the media processing unit 401 to load the first media content 422. Based on this request, the media processing unit 401 may process (423) the currently presented video frame based on supplementary enhancement information. In some cases, in response to the first media content being in an editing state, the media processing unit 401 may parse the supplementary enhancement information to obtain occlusion information therein, and overlay rendered visual elements on the second part of the first information based on the occlusion information, thereby presenting the first media content on the editing interface, displaying the first part but not the second part. For example, the media processing unit 401 may parse the supplementary enhancement information to overlay a mosaic on the second part of the result details information of the question-and-answer process. In some cases, the media processing unit 401 may control whether to overlay the corresponding visual elements based on whether the current playback time is within the time interval corresponding to the occlusion. In some cases, in response to the first media content being in an editing state, the media processing unit 401 can parse the supplementary enhancement information to obtain the second part therein, and remove the second part from the first information based on the occlusion information so as to present the first media content on the editing interface, display the first part and not display the second part.

[0067] In some cases, terminal device 150 can also save the first media content carrying supplementary enhancement information as a first draft, which the user can receive and view through corresponding operations. As mentioned above, terminal device 150 can display an editing interface for the first media content in response to viewing the first draft. Therefore, the logic for viewing the first draft is the same as that of editing stage 404 described above, and will not be repeated here. Since the display control related to occlusion is stored and restored uniformly with the first media content through supplementary enhancement information, the second part remains occluded after the draft is restored. Compared to maintaining the occlusion state separately for the draft, this reduces the amount of state data that needs to be transmitted when the draft is restored, thereby reducing the storage overhead of terminal device 150.

[0068] During the first media content publishing phase 405, terminal device 150 may receive (424) a first operation instructing the publishing of the first media content. In response to the first operation, terminal device 150 may send (425) a media content publishing instruction to media processing unit 401. Media processing unit 401, based on the publishing instruction, generates (426) second media content based on the first media content. In some cases, media processing unit 401 may filter out occlusions related to occlusion control in the first media content, preventing them from participating in the final output (i.e., removing visual elements used to occlude the second part), and encode the underlying video content to obtain the second media content.

[0069] The media processing unit 401 can return (427) the second media content to the terminal device 150 in response to the generation of the second media content. The terminal device 150 can present (428) the second media content in response to obtaining the second media content, and the second media content displays the first part and the second part of the first information. In this way, the user can obtain the complete first information after publication without additional operation; at the same time, since the second media content without occlusion can be output by simply filtering and supplementing the enhanced information during the publication stage, without the need to restore or re-synthesize pixel-level occlusion, compared with the method of maintaining multiple media content before and after occlusion, the pixel rewriting during the publication stage can be reduced and the synthesis calculation overhead during the publication stage can be reduced.

[0070] In some cases, when exporting the effects package corresponding to the first effect, the occluded area can be recalculated according to the specified resolution and filled with corresponding pixels. This is then encoded as an image resource such as Portable Network Graphics (PNG) and saved in the effects package. Transparent parts correspond to non-occluded areas, and opaque parts correspond to occluded areas; the specific style can be determined by the occlusion material. The occlusion information recorded in the supplementary enhancement information can have various storage formats. In some cases, the supplementary enhancement information stores the complete data of the resource corresponding to the visual element, which is directly restored and overlaid in the editing interface. In other cases, the supplementary enhancement information stores the resource identifier or resource index of the resource corresponding to the visual element, which is mapped to the corresponding local or cloud resource in the editing interface before being overlaid. In still other cases, the supplementary enhancement information stores the relative path of the resource corresponding to the visual element in the application sandbox, which is parsed in the editing interface and the resource in the corresponding path is read and displayed.

[0071] Figure 5 A flowchart of an example method 500 for media content processing based on several scenarios is shown. Method 500 can be... Figure 1 The terminal device 150 shown is used for execution, but it can also be executed by other suitable devices. For ease of discussion, please refer to [reference needed]. Figure 1 Let's describe method 500.

[0072] In frame 510, terminal device 150 presents first media content, wherein the first media content displays a first part of the first information but does not display a second part of the first information.

[0073] In frame 520, terminal device 150 receives a first operation, which instructs the release of first media content.

[0074] In frame 530, terminal device 150 presents second media content, which is obtained by publishing first media content. The second media content displays the first part and the second part of the first information.

[0075] In this way, users see a first media content that displays partial information before publication, and obtain the complete first information through a second media content after publication. This provides a phased interactive method for question-and-answer or result-based gameplay. At the same time, since the first and second media content are displayed and controlled based on the same underlying content carrier, compared to maintaining multiple media content before and after obscuring, it can reduce storage overhead and reduce the amount of data that needs to be transmitted for state synchronization.

[0076] In some cases, presenting the primary media content includes: presenting visual elements that are overlaid on the secondary part of the primary information to obscure the secondary part; specific interactive presentations can be found above. Figure 2C The example described illustrates how this method enables occlusion of the second part without altering the underlying content pixels. Furthermore, compared to directly writing the occlusion to the pixels, it reduces pixel rewriting and lowers the image processing computational overhead of the terminal device by 150%.

[0077] In some cases, method 500 further includes: presenting a prompt message before receiving the first operation, the prompt message indicating that the complete first information will be presented after the first media content is published; specific interactive presentation can be found above. Figure 2C The example described illustrates this. In this way, users can learn the rules for unlocking the full content after publication before it is even released; simultaneously, since the notification information is presented on the same interface as the primary media content, there is no need to load a separate notification interface, reducing the number of interface loads and lowering the rendering overhead of the terminal device by 150%.

[0078] In some cases, the primary media content includes the content displaying the question-and-answer process, and the primary information includes the result information corresponding to the question-and-answer process; for specific interactive presentation, please refer to the above. Figures 2A to 2CThe example described illustrates this. This approach combines the question-and-answer process with the result information, enhancing the engaging nature of the phased presentation. Furthermore, since the question-and-answer process and result information are uniformly carried by the same primary media content, the processing overhead required for the terminal device 150 to load each piece of content separately is reduced.

[0079] In some cases, method 500 further includes acquiring the first media content by: receiving a second operation, the second operation instructing the capture of media content; during the capture, presenting the captured image on a capture interface, and presenting at least one first question and at least one candidate answer for each of the first questions on the capture interface, the at least one first question and candidate answers being presented on the captured image; receiving selection information, the selection information instructing the selection of at least one answer from the candidate answers; presenting a first portion of the first information on the captured image; and acquiring the first media content based on the content presented on the capture interface during the capture; specific interactive presentation can be found above. Figures 2A to 2C The example described illustrates this. In this way, users can complete the Q&A and presentation of the first part on the spot during the shooting process; at the same time, compared with jumping to a separate interface for selection, it can reduce the number of interface switching and reduce the rendering overhead of the terminal device by 150%.

[0080] In some cases, the selection information corresponding to choosing at least one of the candidate answers includes at least one of the following: the pose information of the object in the captured image, the touch operation information of the screen, the operation information of the terminal device controls, and the voice input information; for specific interactive presentation, please refer to the above. Figure 2A The example described illustrates this. In this way, users can select answers using a variety of input methods, enhancing the flexibility of the interaction; simultaneously, since answer selection is directly based on the input collected during the shooting process, it reduces the processing overhead required by the terminal device 150 for additional input collection.

[0081] In some cases, the second operation instruction generates media content based on the first special effect, and presenting at least one first question and candidate answer on the shooting interface includes: based on the first special effect, overlaying at least one first question and candidate answer onto the captured image, and presenting the first part of the first information includes: based on the first special effect, presenting the first part of the first information on the captured image; for specific interactive presentation, please refer to the above. Figures 2A to 2C The example described illustrates this. In this way, question-and-answer and information presentation can be driven by a unified first effect; simultaneously, compared to loading and scheduling multiple processing logics separately, the number of processing logics that the terminal device 150 needs to load and schedule can be reduced, thereby reducing processing overhead.

[0082] In some cases, presenting the first media content includes: presenting the first media content in an editing interface, wherein the presented first media content displays a first part of the first information and does not display a second part; and method 500 further includes: presenting the edited first media content in an editing interface, wherein the edited first media content displays a first part and does not display a second part; specific interactive presentation can be found above. Figure 2D The example described. In this way, users can maintain a consistent occlusion presentation during the editing phase as before publication; at the same time, since the occlusion display of the second part is driven by configuration information uniformly bound to the first media content rather than additionally maintaining occlusion tracks, the number of tracks required to be maintained during the editing phase can be reduced, thereby reducing the memory usage of the terminal device 150.

[0083] In some scenarios, the first media content is saved as a first draft. Presenting the first media content includes: receiving a third operation that instructs the user to view the first draft; and responding to the third operation by presenting the first media content. In this way, the user can maintain consistent occlusion presentation after the draft is restored. Furthermore, because the draft and its occlusion control semantics are stored and restored uniformly, compared to maintaining the occlusion state separately for the draft, the amount of state data required to be transmitted during draft restoration is reduced, thereby lowering the storage overhead of the terminal device 150.

[0084] In some cases, the second part includes multiple sub-parts, which include at least one of the following: at least two first sub-parts, each corresponding to a different area of ​​the same frame in the first media content; at least two second sub-parts, each corresponding to a different time interval of the first media content; for specific interactive presentation, please refer to the above. Figure 2C The example described. In this way, the second part can be segmented and occluded in both the image area dimension and the time dimension; at the same time, since the occlusion of multiple sub-parts is driven by a unified display control semantic, the memory overhead occupied by the terminal device 150 in maintaining occlusion tracks can be reduced compared to maintaining independent occlusion tracks for each sub-part.

[0085] In some cases, the first media content is generated by applying a first special effect to the captured footage. Method 500 further includes: in response to a request to apply the first special effect, obtaining a special effect package corresponding to the first special effect, the special effect package including first configuration information, the first configuration information being used to hide a second part of the first information before publication and to display the second part after publication; and applying the special effect package to the captured footage to obtain the first media content; specific interactive presentation can be found above. Figure 4The example described illustrates this. In this way, display control before and after release can be uniformly carried over via the effects package; simultaneously, compared to configuring occlusion logic separately on the consumer side, it reduces configuration processing on the consumer side, thereby lowering the processing overhead of the terminal device 150.

[0086] In some cases, the first configuration information includes occlusion information, which indicates visual elements used to overlay on the second part to occlude it. Presenting the first media content includes: based on the occlusion information, overlaying visual elements on the second part; specific interactive presentation can be found above. Figure 2A as well as Figure 4 The example described illustrates this. In this way, the visual elements used for occlusion can be flexibly specified via configuration information; at the same time, since occlusion is implemented based on overlaying visual elements without rewriting the underlying pixels, pixel rewriting and image processing computational overhead can be reduced.

[0087] In some cases, the occlusion information is included in the supplementary enhancement information corresponding to the primary media content. Method 500 further includes: writing the occlusion information into the supplementary enhancement information during the image capture process; specific interactive presentation can be found above. Figure 4 The example described illustrates this. In this way, occlusion information is carried uniformly along with the primary media content; at the same time, compared to transmitting occlusion business information separately across multiple interfaces, it can reduce the number of data transmissions required for cross-interface and cross-stage transmission, thereby reducing state synchronization overhead.

[0088] In some cases, overlaying visual elements on top of the second part based on occlusion information includes: responding to the first media content being in an editing state, parsing supplementary enhancement information to obtain occlusion information, and rendering visual elements in the first media content based on the occlusion information. For specific interactive presentation details, please refer to the above. Figure 2D as well as Figure 4 The example described. In this way, occlusion rendering can be restored based on a unified occlusion semantic during the editing stage; at the same time, since the occlusion rendering is uniformly driven by supplementary enhancement information without the need for additional maintenance of occlusion tracks, the memory usage of the terminal device 150 during the editing stage can be reduced.

[0089] In some cases, method 500 further includes: in response to a draft saving operation for the first media content, storing the first media content carrying supplementary enhancement information as a first draft; and overlaying visual elements on top of the second part based on occlusion information, including: in response to an operation of viewing the first draft, parsing the supplementary enhancement information carried in the first draft to obtain occlusion information; and rendering visual elements in the first media content based on the obtained occlusion information. This approach ensures consistency in occlusion presentation during draft saving and restoration; simultaneously, since the occlusion control semantics are stored and restored uniformly along with the first media content, the amount of state data required to be transmitted during draft restoration is reduced, thereby reducing the storage overhead of the terminal device 150.

[0090] In some cases, presenting the second media content includes: in response to the first operation, filtering supplementary enhancement information in the first media content and encoding the filtered first media content to obtain the second media content; specific interactive presentation can be found above. Figure 4 The described publishing stage. In this way, unobstructed secondary media content can be directly output; at the same time, compared with methods that restore or re-composite pixel-level occlusions, pixel rewriting and compositing computational overhead in the publishing stage can be reduced.

[0091] Figure 6 A flowchart of an example method 600 for generating special effects packages is shown, based on several scenarios. Method 600 can be generated by... Figure 1 The terminal device 120 shown is used for execution, but it can also be executed by other suitable devices. For ease of discussion, please refer to... Figure 1 To describe method 600.

[0092] In frame 610, terminal device 120 presents a first interface, which is used to configure special effects packages, and the special effects packages correspond to the first special effects.

[0093] In frame 620, terminal device 120 receives first configuration information, which enables media content to apply a first effect: displaying a first part of the first information without displaying a second part of the first information before publication, and displaying both the first and second parts after publication.

[0094] In box 630, terminal device 120 adds the first configuration information to the effects package.

[0095] In this way, special effects creators can set the display control before and after release through the configuration interface without having to set the occlusion logic one by one on the consumer side. At the same time, since the display control is uniformly carried through the first configuration information in the special effects package, compared with the method of configuring and maintaining the occlusion logic separately on the consumer side, it can reduce the configuration processing required on the consumer side to implement the display control, thereby reducing the processing overhead of the consumer side terminal device.

[0096] In some cases, the first configuration information includes occlusion information, which indicates the visual elements superimposed on the second part to occlude it; specific configuration presentation can be found above. Figure 3B as well as Figure 3C The example described illustrates this. In this way, effects creators can flexibly configure the visual elements used for occlusion; at the same time, since occlusion is based on overlaying visual elements without rewriting the underlying pixels, it can reduce pixel rewriting during the processing of media content applying this effects package and reduce image processing computational overhead.

[0097] In some cases, the first configuration information indicates that the second part includes multiple sub-parts, and the multiple sub-parts include at least one of the following: at least two first sub-parts, each corresponding to a different area of ​​the same frame in the media content; and at least two second sub-parts, each corresponding to a different time interval of the media content. For specific configuration presentation, please refer to the above. Figure 3B The example described illustrates this. In this way, occlusion can be configured separately in the image region dimension and the time dimension; at the same time, since the occlusion of multiple sub-parts is driven by a unified first configuration information, the number of configurations that need to be maintained can be reduced compared to maintaining independent configurations for each sub-part.

[0098] In some cases, the first configuration information also includes a second part, indicating that the second part should be displayed after the media content is published. In this way, the second part and its display timing can be uniformly carried through the first configuration information; at the same time, compared to obtaining the second part separately on the consumer side, it can reduce additional requests on the consumer side and reduce the number of data transfers between the end and the cloud.

[0099] In some cases, method 600 further includes: receiving second configuration information, the second configuration information instructing the display of the question-and-answer process in the media content, the first information including result information corresponding to the question-and-answer process, and adding the second configuration information to the effects package. In this way, the presentation of the question-and-answer process can be uniformly configured via the second configuration information; at the same time, since the question-and-answer related content is driven by the unified second configuration information, the processing overhead required for the consumer user to load each presentation item separately can be reduced.

[0100] In some cases, the second configuration information indicates the following items to be overlaid on the captured screen: at least one first question, at least one candidate answer for each of the first questions, and first information. In this way, the overlay presentation of questions and answers and information can be uniformly specified via the second configuration information; at the same time, compared with loading and configuring each presentation item separately, the amount of presentation logic that the consumer-side terminal device needs to load can be reduced, thereby reducing processing overhead.

[0101] A corresponding apparatus for implementing the above methods or processes is also provided.

[0102] Figure 7 A schematic structural block diagram of an example device 700 for media content processing according to some scenarios is shown. Device 700 may be implemented as or included in... Figure 1 The terminal device 150 shown. The various modules / components in the device 700 can be implemented by hardware, software, firmware, or any combination thereof.

[0103] like Figure 7 As shown, the device 700 includes: a first media presentation module 710 configured to present first media content, wherein the first media content displays a first portion of first information and does not display a second portion of the first information; a first operation receiving module 720 configured to receive a first operation, the first operation instructing the publication of the first media content; and a second media presentation module 730 configured to present second media content, the second media content being obtained by publishing the first media content, the second media content displaying a first portion and a second portion of the first information.

[0104] In some cases, the first media presentation module 710 can also be configured to present visual elements that are superimposed on the second part of the first information to obscure the second part.

[0105] In some cases, the device 700 further includes a prompt message presentation module configured to present a prompt message before receiving the first operation, the prompt message indicating that the complete first information is presented after the first media content is published.

[0106] In some cases, primary media content includes content that displays the question-and-answer process, and primary information includes result information corresponding to the question-and-answer process.

[0107] In some cases, the device 700 further includes a content acquisition module configured to acquire first media content by: receiving a second operation instructing the capture of media content; during capture, presenting the captured image on a capture interface and presenting at least one first question and at least one candidate answer for each of the first questions on the capture interface, the at least one first question and candidate answers being presented on the captured image; receiving selection information instructing the selection of at least one answer from the candidate answers; presenting a first portion of the first information on the captured image; and acquiring the first media content based on the content presented on the capture interface during capture.

[0108] In some cases, the selected information includes at least one of the following: the pose information of objects in the captured image, the touch operation information of the screen, the operation information of terminal device controls, and the voice input information.

[0109] In some cases, the second operation instruction generates media content based on the first effect, and the content acquisition module can also be configured to: overlay at least one first question and candidate answer on the captured screen based on the first effect, and present the first part of the first information on the captured screen based on the first effect.

[0110] In some cases, the first media presentation module 710 may also be configured to present first media content in an editing interface, the editing interface being used to edit the first media content, the presented first media content displaying a first part and not displaying a second part, and the device 700 further includes an editing content presentation module configured to present the edited first media content in an editing interface, wherein the edited first media content displays a first part and not displaying a second part.

[0111] In some cases, the first media content is saved as a first draft, and the first media presentation module 710 can also be configured to: receive a third operation, the third operation instructing the view of the first draft; and in response to the third operation, present the first media content.

[0112] In some cases, the second part includes multiple sub-parts, which include at least one of the following: at least two first sub-parts, each corresponding to a different area of ​​the same frame in the first media content, and at least two second sub-parts, each corresponding to a different time interval of the first media content.

[0113] In some cases, the first media content is generated by applying a first special effect to the captured image. The device 700 further includes: an effect package acquisition module, configured to acquire an effect package corresponding to the first special effect in response to an application request for the first special effect. The effect package includes first configuration information, which is used to make the media content with the applied effect hide a second part of the first information before publication and display the second part after publication; and an effect package application module, configured to apply the effect package to the captured image to obtain the first media content.

[0114] In some cases, the first configuration information includes occlusion information, which indicates visual elements used to overlay on the second part to occlude the second part. The first media presentation module 710 can also be configured to overlay and present visual elements on the second part based on the occlusion information.

[0115] In some cases, the occlusion information is included in supplemental enhancement information corresponding to the first media content. The device 700 also includes an occlusion information writing module configured to write the occlusion information into the supplemental enhancement information during the acquisition of the image.

[0116] In some cases, the first media presentation module 710 can also be configured to: parse supplementary enhancement information to obtain occlusion information in response to the first media content being in an editing state; and render visual elements in the first media content based on the occlusion information.

[0117] In some cases, the device 700 further includes: a first draft storage module configured to store the first media content carrying supplementary enhancement information as a first draft in response to a draft saving operation for the first media content; and the first media presentation module 710 may also be configured to: in response to an operation of viewing the first draft, parse the supplementary enhancement information carried in the first draft to obtain occlusion information; and render visual elements in the first media content based on the obtained occlusion information.

[0118] In some cases, the second media presentation module 730 may also be configured to: in response to the first operation, filter supplementary enhancement information in the first media content and encode the filtered first media content to obtain the second media content.

[0119] Figure 8 A schematic structural block diagram of an example device 800 for generating special effects packages according to some scenarios is shown. Device 800 can be implemented as or included in... Figure 1 The terminal device 120 shown. The various modules / components in the device 800 can be implemented by hardware, software, firmware, or any combination thereof.

[0120] like Figure 8As shown, the device 800 includes: a first interface presentation module 810 configured to present a first interface, the first interface being used to configure an effects package, the effects package corresponding to a first effects; a configuration information receiving module 820 configured to receive first configuration information, the first configuration information being used to cause media content applying the first effects to: display a first part of the first information before publication but not a second part of the first information, and display the first part and the second part after publication; and a configuration information adding module 830 configured to add the first configuration information to the effects package.

[0121] In some cases, the first configuration information includes occlusion information that indicates a visual element used to overlay the second part to occlude it.

[0122] In some cases, the first configuration information indicates that the second part includes multiple sub-parts, and the multiple sub-parts include at least one of the following: at least two first sub-parts, each corresponding to a different area of ​​the same frame in the media content, and at least two second sub-parts, each corresponding to a different time interval of the media content.

[0123] In some cases, the first configuration information also includes a second part, and instructs that the second part be displayed after the media content is published.

[0124] In some cases, the configuration information receiving module 820 can also be configured to: receive second configuration information, the second configuration information indicating the display of a question-and-answer process in the media content, the first information including result information corresponding to the question-and-answer process; the configuration information adding module 830 can also be configured to add the second configuration information to the special effects package.

[0125] In some cases, the second configuration information indicates the following items to be overlaid on the captured image: at least one first question, at least one candidate answer for each of the first questions, and first information.

[0126] It should be understood that each module in device 700 and / or device 800 can be implemented in software, hardware, firmware, or any combination thereof. For example, in some cases, one or more modules can be implemented using software and / or firmware, such as machine-executable instructions stored on a storage medium. In addition to machine-executable instructions, or as an alternative, each module can be implemented partially or entirely based on hardware, such as a Field Programmable Gate Array (FPGA), an Application Specific Integrated Circuit (ASIC), an Application Specific Standard Product (ASSP), a System on Chip (SOC), a Complex Programmable Logic Device (CPLD), etc.

[0127] Figure 9 A block diagram of an electronic device 900 capable of implementing various scenarios of the present disclosure is shown. It should be understood that... Figure 9 The electronic device 900 shown is merely exemplary and should not be construed as limiting the functionality and scope of the situation described herein. The electronic device 900 may, for example, be used to implement the terminal devices 120 and 150 described above, or related functions thereof.

[0128] like Figure 9 As shown, electronic device 900 is in the form of a general-purpose electronic device. Components of electronic device 900 may include, but are not limited to, one or more processing units or processors 910, memory 920, storage devices 930, one or more communication units 940, one or more input devices 950, and one or more output devices 960. Processor 910 may be a physical or virtual processor and is capable of performing various processes according to programs stored in memory 920. In a multiprocessor system, multiple processors execute computer-executable instructions in parallel to improve the parallel processing capability of electronic device 900.

[0129] Electronic device 900 typically includes multiple computer storage media. Such media can be any accessible media that is accessible to electronic device 900, including but not limited to volatile and non-volatile media, removable and non-removable media. Memory 920 can be volatile memory (e.g., registers, cache, random access memory (RAM)), non-volatile memory (e.g., read-only memory (ROM), electrically erasable programmable read-only memory (EEPROM), flash memory), or some combination thereof). Storage device 930 can be removable or non-removable media and can include machine-readable media, such as flash drives, disks, or any other media capable of storing information and / or data and accessible within electronic device 900.

[0130] Electronic device 900 may further include additional removable / non-removable, volatile / non-volatile storage media. Although not explicitly stated... Figure 9 As shown, disk drives for reading from or writing to removable, non-volatile disks (e.g., "floppy disks") and optical disk drives for reading from or writing to removable, non-volatile optical disks can be provided. In these cases, each drive can be connected to a bus (not shown) via one or more data media interfaces. Memory 920 may include computer program product 925 having one or more program modules configured to perform various methods or actions for the various scenarios described herein.

[0131] The communication unit 940 enables communication with other electronic devices via a communication medium. Additionally, the functionality of the components of the electronic device 900 can be implemented using a single computing cluster or multiple computing machines capable of communicating via communication connections. Therefore, the electronic device 900 can operate in a networked environment using logical connections to one or more other servers, network personal computers (PCs), or another network node.

[0132] Input device 950 can be one or more input devices, such as a mouse, keyboard, trackball, etc. Output device 960 can be one or more output devices, such as a monitor, speaker, printer, etc. Electronic device 900 can also communicate with one or more external devices (not shown) via communication unit 940 as needed. External devices include storage devices, display devices, etc., and can communicate with one or more devices that enable user interaction with electronic device 900, or with any device (such as a network card, modem, etc.) that enables electronic device 900 to communicate with one or more other electronic devices. Such communication can be performed via input / output (I / O) interface (not shown).

[0133] A computer-readable storage medium is provided that stores computer-executable instructions thereon, wherein the computer-executable instructions are executed by a processor to implement the methods described above. According to this disclosure, a computer program product is also provided, which is tangibly stored on a non-transitory computer-readable medium and includes computer-executable instructions that are executed by a processor to implement the methods described above.

[0134] Flowcharts and / or block diagrams of methods, apparatuses, devices, and computer program products referenced herein describe various aspects of this disclosure. It should be understood that each block of a flowchart and / or block diagram, and combinations of blocks in flowcharts and / or block diagrams, can be implemented by computer-readable program instructions.

[0135] These computer-readable program instructions can be provided to a processor of a general-purpose computer, a special-purpose computer, or other programmable data processing apparatus to produce a machine such that, when executed by the processor of the computer or other programmable data processing apparatus, they create means for implementing the functions / actions specified in one or more blocks of the flowchart and / or block diagram. These computer-readable program instructions can also be stored in a computer-readable storage medium that causes a computer, programmable data processing apparatus, and / or other device to operate in a particular manner; thus, the computer-readable medium storing the instructions comprises an article of manufacture that includes instructions for implementing aspects of the functions / actions specified in one or more blocks of the flowchart and / or block diagram.

[0136] Computer-readable program instructions can be loaded onto a computer, other programmable data processing apparatus, or other device to cause a series of operational steps to be performed on the computer, other programmable data processing apparatus, or other device to produce a computer-implemented process, thereby causing the instructions that execute on the computer, other programmable data processing apparatus, or other device to perform the functions / actions specified in one or more boxes of a flowchart and / or block diagram.

[0137] The flowcharts and block diagrams in the accompanying figures illustrate the architecture, functionality, and operation of possible implementations of systems, methods, and computer program products under various scenarios. In this respect, each block in a flowchart or block diagram may represent a module, segment, or portion of an instruction, which contains one or more executable instructions for implementing the specified logical function. In some alternative implementations, the functions marked in the blocks may occur in a different order than those shown in the figures. For example, two consecutive blocks may actually be executed substantially in parallel, and they may sometimes be executed in reverse order, depending on the functions involved. It should also be noted that each block in the block diagrams and / or flowcharts, and combinations of blocks in the block diagrams and / or flowcharts, can be implemented using a dedicated hardware-based system that performs the specified function or action, or using a combination of dedicated hardware and computer instructions.

[0138] Various examples have been described above. The foregoing descriptions are exemplary and not exhaustive, nor are they limited to the disclosed implementations. Many modifications and variations will be apparent to those skilled in the art without departing from the scope and spirit of the described implementations. The terminology used herein is chosen to best explain the principles, practical applications, or improvements to technology in the market, or to enable others skilled in the art to understand the various implementations disclosed herein.

Claims

1. A media content processing method, comprising: Presenting first media content; wherein the first media content displays a first part of the first information but does not display a second part of the first information; Receive a first operation, the first operation instructing the publication of the first media content; and The second media content is presented, which is obtained by publishing the first media content, and the second media content displays the first part and the second part of the first information.

2. The method according to claim 1, wherein presenting the first media content includes: A visual element is presented, which is superimposed on the second part of the first information to obscure the second part.

3. The method according to claim 1, further comprising: Before receiving the first operation, a prompt message is presented, indicating that the complete first information will be presented after the first media content is published.

4. The method of claim 1, wherein the first media content further comprises: The content of the question-and-answer process is displayed, and the first information includes the result information corresponding to the question-and-answer process.

5. The method according to claim 4, further comprising obtaining the first media content by means of: Receive a second operation, which instructs you to capture media content; During the filming, The captured image is displayed on the shooting interface, and at least one first question and candidate answers for each of the at least one first question are displayed on the shooting interface, with the at least one first question and the candidate answers displayed on the captured image; Receive selection information, the selection information indicating the selection of at least one answer from the candidate answers; The first portion of the first information is displayed on the captured image; as well as The first media content is obtained based on the content displayed on the shooting interface during the shooting process.

6. The method according to claim 5, wherein, The selection information includes at least one of the following: The captured image contains the pose information of the objects in the scene. Touch operation information on the screen, where the screen is the screen of the terminal device; Operation information for terminal device controls; Input information via voice.

7. The method of claim 1, wherein presenting the first media content comprises: The first media content is presented in an editing interface, which is used to edit the first media content. The presented first media content displays the first part but not the second part. The method further includes: The edited first media content is presented in the editing interface, wherein the edited first media content displays the first part but not the second part.

8. The method according to claim 1, wherein the first media content is saved as a first draft, and presenting the first media content includes: Receive a third operation, which instructs you to view the first draft; as well as In response to the third operation, the first media content is presented.

9. The method of claim 1, wherein the second portion comprises a plurality of sub-parts, the plurality of sub-parts comprising at least one of the following: At least two first sub-parts, each corresponding to a different area of ​​the same frame in the first media content. At least two second sub-parts, each corresponding to a different time interval of the first media content.

10. A method for generating special effects packages, comprising: A first interface is presented, which is used to configure a special effects package, and the special effects package corresponds to a first special effect. Receive first configuration information, the first configuration information being used to cause media content applying the first effect to: display a first part of the first information without displaying a second part of the first information before publication, and display both the first part and the second part after publication; as well as Add the first configuration information to the effects package.

11. The method of claim 10, wherein the first configuration information includes occlusion information, the occlusion information indicating a visual element for overlaying the second portion to occlude the second portion.

12. The method of claim 10, wherein the first configuration information indicates that the second part comprises a plurality of sub-parts, the plurality of sub-parts comprising at least one of the following: At least two first sub-parts, each corresponding to a different area of ​​the same frame in the media content. At least two second sub-parts, each corresponding to a different time interval of the media content.

13. The method of claim 10, wherein the first configuration information includes the second portion and indicates that the second portion is displayed after the media content is published.

14. The method of claim 10, further comprising: Receive second configuration information, the second configuration information indicating that a question-and-answer process is displayed in the media content, the first information including result information corresponding to the question-and-answer process; as well as Add the second configuration information to the effects package.

15. The method of claim 14, wherein the second configuration information indicates the following items to be overlaid on the captured image: At least one first question, Each of the at least one candidate answer to the first question. The first piece of information.

16. A media content processing apparatus, comprising: The first media presentation module is configured to present first media content; wherein the first media content displays a first part of the first information and does not display a second part of the first information. A first operation receiving module is configured to receive a first operation, wherein the first operation instructs the publication of the first media content; and The second media presentation module is configured to present second media content, which is obtained by publishing the first media content. The second media content displays the first part and the second part of the first information.

17. A special effects package generation device, comprising: The first interface presentation module is configured to present a first interface, which is used to configure a special effects package, and the special effects package corresponds to a first special effect. The configuration information receiving module is configured to receive first configuration information, which is used to make the media content applying the first special effect: display a first part of the first information before publication but not a second part of the first information, and display both the first part and the second part after publication; as well as The configuration information adding module is configured to add the first configuration information to the effects package.

18. An electronic device, comprising: At least one processor; as well as At least one memory coupled to the at least one processor and storing instructions for execution by the at least one processor, the instructions, when executed by the at least one processor, causing the electronic device to perform the method according to any one of claims 1 to 9, or to perform the method according to any one of claims 10 to 15.

19. A computer-readable storage medium having stored thereon computer-executable instructions, said computer-executable instructions being executable by a processor to implement the method according to any one of claims 1 to 9, or to implement the method according to any one of claims 10 to 15.

20. A computer program product comprising computer-executable instructions, wherein the computer-executable instructions, when executed by a processor, implement the method according to any one of claims 1 to 9, or perform the method according to any one of claims 10 to 15.