Methods and apparatus for creating media content, electronic devices, and computer programs

The method and apparatus automate the conversion between media content modalities, reducing the workload and improving efficiency by allowing creators to generate content in a main modal and convert it to different formats.

JP7841120B2Active Publication Date: 2026-04-06TENCENT TECHNOLOGY (SHENZHEN) CO LTD
View PDF 3 Cites 0 Cited by

Patent Information

Authority / Receiving Office
JP · JP
Patent Type
Patents
Current Assignee / Owner
Filing Date
2023-08-03
Publication Date
2026-04-06

AI Technical Summary

Technical Problem

Media content creators face increased workload and decreased efficiency due to the need to create the same content multiple times for different media modalities on various platforms.

Method used

A method and apparatus that supports conversion between multiple types of media content modalities, allowing creators to generate content in a main modal and convert it to different submodal formats using an application program based on modal conversion algorithms.

Benefits of technology

The technical solution effectively addresses the challenge of reducing the workload and enhancing the efficiency of media content creation by automating the conversion process between media content modalities.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure 0007841120000001
    Figure 0007841120000001
  • Figure 0007841120000002
    Figure 0007841120000002
  • Figure 0007841120000003
    Figure 0007841120000003
Patent Text Reader

Abstract

The present application provides a method and apparatus for creating media content, a device, and a storage medium applicable to technical fields such as multimedia, cloud technology, etc. The method includes the steps of: a terminal device displays a main modal editing interface in response to a trigger operation; a terminal device generates a main modal media content in response to an editing operation on the main modal editing interface; a terminal device converts the main modal media content into a target sub-modal media content in response to a modal conversion operation; and a terminal device displays the generated main modal media content and the target sub-modal media content. That is, the present application converts the main modal media content into at least one sub-modal media content different from the main modal, and the entire modal conversion process is performed by an application program based on its own modal conversion algorithm, so that the media content creator does not need to be involved, and furthermore, the workload of the media content creator is reduced, and the creation efficiency of the media content is improved.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] This application claims priority based on a Chinese patent application filed with the China National Intellectual Property Administration on September 2, 2022, with an application number of 202211074165.4 and an invention title of "Method and Apparatus, Device, and Storage Medium for Creating Media Content", and incorporates all of its content by reference into this application.

[0002] Embodiments of this application relate to the field of multimedia technologies, and in particular, to a method and apparatus, device, and storage medium for creating media content.

Background Art

[0003] With the rapid development of multimedia technologies, various modal media contents are spreading on social media, enriching the lives of viewers.

[0004] Social media platforms support various different media modalities. For example, some support videos, some support documents, some support images, and of course, there are also those that support multiple types of media modalities such as videos, documents, and images. To meet the requirements of different media platforms, creators of media content usually need to create the same content multiple times to generate media content of different media modalities. This causes an increase in the workload of media content creators and a decrease in the creation efficiency of media content.

Summary of the Invention

Problems to be Solved by the Invention

[0005] Embodiments of this application provide a method and apparatus, device, and storage medium for creating media content to reduce the workload of media content creators and improve the creation efficiency of media content.

Means for Solving the Problems

[0006] In the first aspect, an embodiment of the present application is a method for creating media content applicable to a terminal device, The steps include: displaying the main modal editing interface in response to a trigger operation, The steps include generating media content for the main modal in response to editing operations on the main modal editing interface, The steps include: converting the media content of the main modal to the media content of the target submodal in response to a modal conversion operation; The steps include displaying the media content of the generated main modal and the media content of the target submodal, It provides methods for creating media content, including [specific examples of media content creation methods].

[0007] In a second aspect, the embodiment of the present application is a media content creation device applied to a terminal device, A first display unit that displays the main modal editing interface in response to a trigger operation, A processing unit that generates media content for the main modal in response to editing operations on the main modal editing interface, A conversion unit that converts the media content of the main modal to the media content of the target submodal in response to a modal conversion operation, A second display unit that displays the media content of the generated main modal and the media content of the target submodal, The present invention provides a media content creation device equipped with the following features.

[0008] In a third aspect, an embodiment of the present application includes a processor and a memory for storing a computer program, the processor calling and executing the computer program stored in the memory, thereby providing an electronic device that performs the method of the first aspect.

[0009] In a fourth aspect, an embodiment of the present application provides a computer-readable storage medium storing a computer program that causes a computer to execute the method of the first aspect.

[0010] In a fifth aspect, an embodiment of the present application provides a chip that implements a method in any aspect of the first aspect described above or in each implementation thereof. Specifically, the chip includes a processor that calls and executes a computer program from memory, thereby causing a device on which the chip is mounted to execute a method in any aspect of the first aspect described above or in each implementation thereof.

[0011] In a sixth aspect, an embodiment of the present application provides a computer program product that, when executed on a computer, causes the computer to perform any aspect of the first aspect described above or the method in each implementation thereof.

[0012] In the seventh aspect, an embodiment of the present application provides a computer program that, when executed on a computer, causes the computer to execute any aspect of the first aspect described above or the method in each of its implementations. [Effects of the Invention]

[0013] As described above, in this application, the terminal device displays a main modal editing interface in response to a trigger operation, generates media content for the main modal in response to editing operations on the main modal editing interface, converts the media content for the main modal to media content for a target submodal in response to a modal conversion operation, and displays the generated media content for both the main modal and the target submodal. In other words, in the embodiment of this application, conversion between media content of multiple types of modals is supported, and a media content creator can create media content for the main modal using this application program, and then convert this media content for the main modal to media content for at least one submodal different from the main modal by inputting a modal conversion operation. The entire modal conversion process is performed by the application program based on its own modal conversion algorithm, eliminating the need for media content creators to be involved, and further reducing the workload of media content creators and improving the efficiency of media content creation.

[0014] To more clearly explain the technical aspects of the embodiments of this application, the following drawings, which are necessary for describing the embodiments, are briefly introduced below. The drawings described below represent only some embodiments of the present invention, and it will be obvious to those skilled in the art that other drawings can be obtained from these drawings without requiring any creative work. [Brief explanation of the drawing]

[0015] [Figure 1] This is a schematic diagram illustrating an application scenario related to the embodiment of the present invention. [Figure 2] This is a flowchart of a method for creating media content provided in one embodiment of the present invention. [Figure 3] This is a schematic diagram of the interface according to an embodiment of the present invention. [Figure 4]It is a schematic diagram of the media content interface of the generated main modal. [Figure 5A] It is a schematic diagram of the sub-modal corresponding to the main modal. [Figure 5B] It is a schematic diagram of the sub-modal corresponding to the main modal. [Figure 5C] It is a schematic diagram of the sub-modal corresponding to the main modal. [Figure 5D] It is a schematic diagram of the sub-modal corresponding to the main modal. [Figure 6A] It is a schematic diagram of the conversion process from the main modal to the sub-modal. [Figure 6B] It is a schematic diagram of the conversion process from the main modal to the sub-modal. [Figure 6C] It is a schematic diagram of the conversion process from the main modal to the sub-modal. [Figure 7A] It is a schematic diagram of another conversion process from the main modal to the sub-modal. [Figure 7B] It is a schematic diagram of another conversion process from the main modal to the sub-modal. [Figure 7C] It is a schematic diagram of another conversion process from the main modal to the sub-modal. [Figure 7D] It is a schematic diagram of another conversion process from the main modal to the sub-modal. [Figure 7E] It is a schematic diagram of another conversion process from the main modal to the sub-modal. [Figure 7F] It is a schematic diagram of another conversion process from the main modal to the sub-modal. [Figure 8] It is a schematic flowchart of the method for creating media content provided in an embodiment of the present application. [Figure 9] It is a schematic diagram of the interaction between the creation application program according to the embodiment of the present application and the creator of the media content, and within the creation application program. [Figure 10] It is a schematic configuration diagram of the media content creation device provided in an embodiment of the present application. [Figure 11] This is a schematic block diagram of the electronic device provided in the embodiment of the present application. [Modes for carrying out the invention]

[0016] The following describes the technical aspects of embodiments of the present invention clearly and completely with reference to the drawings of the embodiments of the present invention, but it is clear that the embodiments described are only a part of the present invention, not all of it. Other embodiments that a person skilled in the art could obtain based on the embodiments of the present invention without requiring any creative work should also be included within the scope of protection of the present invention.

[0017] Furthermore, terms such as “First,” “Second,” etc., in the specification and claims of the present invention and in the drawings above are for distinguishing similar subjects and are not intended to describe a specific order or priority. It should be understood that the data used herein are interchangeable, where appropriate, in order to enable the embodiments of the present invention described herein to be implemented in an order other than that shown or described herein. Also, the terms “includes,” “has,” and any variations thereof are intended to cover non-exclusively included items. For example, processes, methods, systems, products, and servers that include a series of steps or units do not have to be limited to the steps or units explicitly shown, and may include steps or units not explicitly shown for these processes, methods, products, or servers, or other steps or units inherent to them.

[0018] To facilitate understanding of the embodiments of this application, we will first explain the relevant concepts mentioned in the embodiments.

[0019] Metacreation: In the embodiments of this application, metacreation refers to a creation mode that simultaneously considers the content to be published, such as video, audio, charts, and documents, within a single creation flow. Because multiple types of media modals are involved, the metacreation in the embodiments of this application may also be called multimodal creation.

[0020] Metacreation Project: In embodiments of this application, a metacreation project refers to a metacreation project that combines a custom structured data file with corresponding accessible multimedia material files. The metacreation project includes metadata describing the project structure and accessible paths to all cited material media files. Here, the metadata describing the project structure can be understood as media material related to the media content of the current modal.

[0021] Main Modal and Submodal: A key feature of the meta-creation project is the distinction between the main modal and submodals. One meta-creation project in the embodiment of this application supports one main modal and multiple submodals. The media types of the submodals do not overlap with those of the main modal. For example, if the main modal is a video project file, its submodals may be documents, audio, or pictures.

[0022] Intelligent Modal Transformation: Supports metacreation projects and the mutual transformation of modals, providing standardized modal transformation rules by offering transformation algorithm logic for multiple types of modals.

[0023] An Application Programming Interface (API) is a set of predefined functions designed to allow application programs and developers to access a set of processes based on certain software or hardware. It does not require direct access to source code or a deep understanding of the internal workings.

[0024] A Software Development Kit (SDK) is a collection of related documentation, presentation examples, and tools that support the development of a particular type of software.

[0025] The media content creation method according to the embodiment of this application can also be combined with cloud technology. For example, by combining it with cloud storage in cloud technology, the generated media content can be stored in the cloud. The following describes the details of cloud technology.

[0026] Cloud technology refers to hosting technology that integrates a set of resources, such as hardware, software, and networks, within a wide area network or local area network to enable data computation, storage, processing, and sharing.

[0027] Cloud computing is a computing mode that distributes computing tasks across a resource pool consisting of numerous computers, allowing various application systems to obtain computing power, storage capacity, and information services as needed. The network providing these resources is called the "cloud." Resources in the "cloud" are infinitely scalable for users, can be acquired at any time, used as needed, and expanded at any time, with charges based on usage.

[0028] Cloud storage is a new concept that extends and develops the concept of cloud computing. A distributed cloud storage system (hereinafter referred to as a storage system) is a storage system that uses functions such as cluster applications, grid technology, and distributed storage file systems to integrate and coordinate a large number of various types of storage devices (storage devices are also called storage nodes) in a network through application software or application interfaces, and jointly provides data storage and business access functions to the outside.

[0029] The following diagrams illustrate the application scenarios of the embodiments of this application.

[0030] Figure 1 is a schematic diagram of an application scenario according to an embodiment of the present invention. As shown in Figure 1, it includes a terminal device 101 and a server 102.

[0031] Terminal device 101 includes, but is not limited to, desktop computers, notebook computers, smartphones, tablet computers, IoT devices, and portable wearable devices. IoT devices may include smart speakers, smart TVs, smart air conditioners, and smart in-car devices. Portable wearable devices may include smartwatches, smart bracelets, and head-mounted devices. Terminal device 101 is often equipped with a display device, which may be a display, display screen, touchscreen, etc., and the touchscreen may be a touch command screen, touch command panel, etc.

[0032] Server 102 may be one or more. If there are multiple servers 102, there are at least two servers to provide different services and / or at least two servers to provide the same service, for example in a load-balanced manner, but this is not limited to the embodiments of the present application. Here, Server 102 may be an independent physical server, a server cluster or distributed system composed of multiple physical servers, or a cloud server providing basic cloud computing services such as cloud services, cloud databases, cloud computing, cloud functions, cloud storage, network services, cloud communications, middleware services, domain name services, security services, CDN (Content Delivery Network), and big data and artificial intelligence platforms. Server 102 may also be a node in a blockchain.

[0033] The terminal device 101 and the server 102 can be connected directly or indirectly by wired communication or wireless communication, but there are no restrictions in this application.

[0034] In the embodiment of this invention, the terminal device 101 has an application program for creating media content installed. The media content creator launches this application program by triggering the application program icon on the desktop of the terminal device 101. When this application program is launched, the "Create New Metacreation" option is displayed. When the media content creator triggers this "Create New Metacreation" option, the application program displays a creation container via the terminal device 101. The media content creator can perform creation within this creation container and further generate media content for the main modal. Next, the media content creator can input a modal conversion operation on this application program. The application program converts the generated main modal media content into media content for the target submodal in accordance with the modal conversion operation input by the media content creator, and displays the generated main modal media content and target submodal media content in the creation container for the media content creator to refer to.

[0035] Furthermore, the application program can transmit the generated main modal media content and / or target submodal media content to the server via the terminal device 101, thereby enabling content storage or publication.

[0036] The application scenarios of the embodiments of this application include, but are not limited to, those shown in Figure 1.

[0037] Social media platforms support a variety of media modals; some support video, others documents, and still others images. Of course, some support multiple types of media modals, such as video, documents, and images. To meet the requirements of different media platforms, media content creators typically need to create the same content multiple times, for example, by using different media tools to create separate media content for each modal. This increases the workload for media content creators and reduces the efficiency of content creation.

[0038] To solve the aforementioned problems, the embodiment of the present invention provides a method for creating media content that supports conversion between multiple types of modal media content, such as conversion from video modal to picture, document, or audio modal, from picture modal to video, document, or audio modal, from document modal to video, picture, or audio modal, and from audio modal to video, picture, or document modal. This allows a media content creator to create media content in a main modal using this application program, and then, by inputting a modal conversion operation, convert this main modal media content into media content in at least one submodal different from the main modal. Since the entire modal conversion process is performed by the application program based on its own modal conversion algorithm, the media content creator does not need to be involved, further reducing the workload of the media content creator and improving the efficiency of media content creation.

[0039] The technical aspects of the embodiments of the present application will be described in detail below through several examples. The following embodiments may be combined with each other, and detailed explanations of the same or similar concepts or processes may be omitted in some embodiments.

[0040] Figure 2 is a flowchart of a method for creating media content provided in one embodiment of the present invention.

[0041] The execution entity in the embodiments of this application is a device having a media content creation function, for example, a media content creation device (simply referred to as a creation device). In some embodiments, this creation device may be a terminal device, for example, the terminal device shown in Figure 1. In some embodiments, this creation device may be an application program installed on the terminal device. The following description will use the case where the execution entity is a terminal device as an example.

[0042] As shown in Figure 2, this method includes the following steps S201 to S204.

[0043] In S201, the terminal device displays the main modal editing interface in response to a trigger operation.

[0044] In the embodiment of the present invention, as shown in Figure 1, an application program for creating media content (hereinafter simply referred to as the creation application program) is installed on the terminal device.

[0045] In some embodiments, the display screen of the terminal device is a touchscreen. This allows the creator of media content to interact with the terminal device via this touchscreen, for example, to interact with the creation application program on the terminal device via the touchscreen.

[0046] In some embodiments, the display screen of the terminal device is not a touchscreen, and in this case, the terminal device further includes mechanical keys. This allows the media content creator to interact with the terminal device via these mechanical keys, for example, to interact with the creation application program on the terminal device via the mechanical keys.

[0047] In some embodiments, the terminal device also supports voice control functionality. This allows media content creators to interact with the terminal device by voice, for example, by voice, with their creation application program on the terminal device.

[0048] In some embodiments, the terminal device also supports gesture control. This allows media content creators to interact with the terminal device through gestures, for example, by interacting with the creation application program on the terminal device through gestures.

[0049] In the embodiments of this application, there are no restrictions on the specific means of interaction between the media content creator and the terminal device.

[0050] For example, as shown in Figure 3, when the creation application program is installed on a terminal device, an icon for the creation application program is generated on the terminal device's desktop. When a media content creator needs to create media, they can trigger the creation application program icon on the terminal device's desktop and, for example, launch the application program on the terminal device by tapping this icon.

[0051] In one example, as shown in Figure 3, the display interface of the creation application program after startup includes a "Create New Meta Creation" option. This "Create New Meta Creation" option is used to create the main media content. For example, the creator of the media content can trigger this "Create New Meta Creation" option by, for example, tapping it, and the creation application program on the terminal device will display the main modal editing interface in response to the media content creator's trigger operation on the "Create New Meta Creation" option.

[0052] In some embodiments, the display interface of the creation application program after startup includes, in addition to the Meta Creation New Create option, a Main Modal Selection option that allows the media content creator to select a main modal, for example, by selecting a video or a picture as the main modal. In this embodiment, the media content creator (i.e., the user) selects a main modal, then triggers the Meta Creation New Create option, causing the creation application program to jump to the main modal editing interface.

[0053] In S202, the terminal device generates media content for the main modal in response to editing operations on the main modal editing interface.

[0054] The creation application program on the terminal device displays a main modal editing interface in response to a trigger operation by the media content creator on the displayed MetaCreation New Creation option. This allows the media content creator to create media content within this main modal editing interface, for example, by editing video content, picture content, audio content, or document content.

[0055] The modal in the embodiment of the present application can be understood as a media form, such as video, audio, picture, or document.

[0056] In some embodiments, to facilitate the creation of media content, the main modal editing interface includes multiple media creation tools, such as tools for editing, deleting, correcting, and inserting.

[0057] In some embodiments, different creation templates can be set for different modal media content to further facilitate the creation process for media content creators. For example, a video creation template, an audio creation template, a picture creation template, and a document creation template can be set. This allows media content creators to select different creation templates as needed.

[0058] In some embodiments, different creation tools can be set for different types of creation templates. For example, a video creation template may include multiple tools for video creation, an audio creation template may include multiple tools for audio creation, a picture creation template may include multiple tools for picture creation, and a document creation template may include multiple tools for document creation.

[0059] In the embodiments of this application, one creation is referred to as one media creation project, and a media creation project is also referred to as a metacreation project. For the sake of explanation, the media creation project in the embodiments of this application will be referred to as the first media creation project.

[0060] In the embodiments of this application, media content created by a media content creator through editing in the main modal editing interface is referred to as the media content of the main modal. For example, if a media content creator selects video as the main modal, the media content creator performs video editing in the main modal editing interface, and the resulting media content of the main modal is video content. If a media content creator selects audio as the main modal, the media content creator performs audio editing in the main modal editing interface, and the resulting media content of the main modal is audio content. If a media content creator selects a picture as the main modal, the media content creator performs picture editing in the main modal editing interface, and the resulting media content of the main modal is picture content. If a media content creator selects a document as the main modal, the media content creator performs document editing in the main modal editing interface, and the resulting media content of the main modal is document content.

[0061] In other words, in some embodiments, the main modal may be one of the following: video, audio, picture, document, etc.

[0062] In the embodiments of this application, video, audio, pictures, and documents were described as examples. However, the media modals mentioned in the embodiments of this application include, but are not limited to, video, audio, pictures, and documents. Other new modals may also be used, and there are no limitations in the embodiments of this application.

[0063] In some embodiments, after the media content for the main modal of the first media creation project is generated, the creation container displays the generated main modal media content.

[0064] In one possible implementation, the creation container displays the media content of the generated main modal as a floating icon.

[0065] Optionally, this floating icon may include a marker representing the form of the main modal. For example, if the main modal is video, this floating icon will include a camera marker; if the main modal is audio, this floating icon will include a sound marker; if the main modal is a picture, this floating icon will include a picture marker; and if the main modal is a document, this floating icon will include a document marker.

[0066] For example, as shown in Figure 4, if the main modal is a document, the creation container displays a floating icon representing the generated main modal content, and this floating icon includes a document marker. By tapping the floating icon in the creation container, the media content creator can switch to the main modal editing interface and re-edit the media content of this main modal.

[0067] In some embodiments, the creation container displays the media content of the generated main modal, as well as at least one of the following: a cover image, title, update date, additional features entry, etc., that corresponds to the media content of the main modal.

[0068] In the embodiments of this invention, if the modal for the media content of the main modal is different, the corresponding cover image is also different.

[0069] For example, if the media content of the main modal is video content, the cover image may be the first image of the video content, or one image specified by the creator of the media content. If the media content of the main modal is audio content, the cover image may be the sound wave image corresponding to the first frame of the audio content, or one sound wave image corresponding to one frame of the audio content, as specified by the creator of the media content. If the media content of the main modal is picture content, the cover image may be one of the pictures in the picture content, or one picture specified by the creator of the media content. If the media content of the main modal is document content, the cover image may be the first page of the document with the title of the document content, or one page of the document, as specified by the creator of the media content.

[0070] In the embodiment of this invention, if the modal for the media content of the main modal is different, the corresponding title will also be different.

[0071] For example, if the media content of the main modal is video content, the corresponding title can be "Create Video". Also, for example, if the media content of the main modal is audio content, the corresponding title can be "Create Audio". Also, for example, if the media content of the main modal is picture content, the corresponding title can be "Create Picture". Also, for example, if the media content of the main modal is document content, the corresponding title can be "Create Document".

[0072] Here, the update date can be understood as the most recent time the media content of the main modal was updated.

[0073] For example, if the media content of the main modal is video content, then, referring to Figure 4, the creation container displays a cover page corresponding to this media content of the main modal, and on this cover page, a floating icon representing the generated media content of the main modal is displayed. Tapping this floating icon corresponding to the media content of the main modal jumps to the editing interface for the media content of the main modal, where the current media content of the main modal can be viewed and edited. Optionally, the creation container may also display a theme corresponding to the media content of the main modal, such as "Create Video," and the latest update time of the media content of the main modal, such as "2022-03-31." Optionally, the creation container may also display an entry point for additional functions. For example, in Figure 4, the entry point for additional functions is indicated by the icon "···," and tapping this icon displays a dropdown list of additional functions, from which the desired function can be selected and executed.

[0074] In some embodiments, the above S202 includes the following steps S202-A1 and S202-A2.

[0075] In S202-A1, the terminal device generates media content for the main modal in response to editing operations on the main modal editing interface, determines N types of submodals corresponding to the main modal, and the N types of submodals are different from the main modal, where N is a positive integer.

[0076] In S202-A2, the terminal device displays the media content of the main modal and the conversion waiting icons for N types of submodals in the same creation container.

[0077] In the embodiment of the present invention, the creation application program on the terminal device not only generates media content for the main modal in response to editing operations performed by the media content creator on the main modal editing interface, but also determines N types of submodals corresponding to this main modal. Here, the N types of submodals are different from the main modal. Next, the terminal device displays the media content of the main modal and conversion waiting icons for the N types of submodals in the creation container. Here, the conversion waiting icons indicate that these three types of submodals have not yet been converted.

[0078] In some embodiments, S202-A2 includes the step of the terminal device displaying media content for a main modal in a first area of ​​a creation container and displaying N types of submodal conversion waiting icons in a second area of ​​the creation container, where the first area is larger than the second area.

[0079] Optionally, the sizes of the N submodal conversion waiting icons displayed in the second area may be the same.

[0080] For example, as shown in Figure 5A, if the main modal is a document, then the N types of submodals corresponding to this main modal are video, picture, and audio. In this case, the creation container displays icons for the media content of the generated main modal, as well as icons for the three types of submodals (video, picture, and audio) awaiting conversion.

[0081] As shown in Figure 5B, if the main modal is a video, then the N types of submodals corresponding to this main modal are picture, document, and audio. In this case, the terminal device displays in the creation container icons for the media content of the generated main modal, as well as icons for the three types of submodals (picture, document, and audio) awaiting conversion.

[0082] As shown in Figure 5C, if the main modal is a picture, then the N types of submodals corresponding to this main modal are video, document, and audio. In this case, the creation container displays icons for the media content of the generated main modal, as well as icons for the three types of submodals (video, document, and audio) awaiting conversion.

[0083] As shown in Figure 5D, if the main modal is audio, then the N types of submodals corresponding to this main modal are video, picture, and document. In this case, the terminal device displays in the creation container icons for the media content of the generated main modal, as well as icons for the three types of submodals (video, picture, and document) awaiting conversion.

[0084] Although the above explanation uses N=3 as an example, the number of submodals corresponding to different main modals may vary. Furthermore, the type and number of submodals corresponding to main modals may be specified by the media content creator.

[0085] In addition to generating the media content for the main modal of the first media creation project according to the steps described above, perform the following step S203.

[0086] In S203, the terminal device converts the media content of the main modal to the media content of the target submodal in response to a modal conversion operation.

[0087] In the embodiments of this invention, the media content creator only needs to create and generate media content for the main modal in the main modal editing interface, and does not need to generate media content for other modals. Media content for other modals can be obtained by modal conversion from the generated main modal media content, further reducing the workload of the media content creator and improving the efficiency of media content creation.

[0088] In the embodiments of this invention, there are no restrictions on the specific means by which the creator of media content inputs a modal conversion operation to the creation application program.

[0089] Method 1: The creator of the media content inputs a conversion command to the creation application program, which is used to instruct the creation application program to convert the media content of the current main modal to the media content of the target submodal. As a result, the creation application program on the terminal device converts the media content of the main modal to the media content of the target submodal according to this conversion command. In other words, the modal conversion operation in Method 1 is a conversion command input by the creator of the media content.

[0090] Method 2: As shown in Figures 5A to 5D, the creation container contains the media content of the generated main modal and conversion waiting icons for N types of submodals corresponding to this main modal. By triggering the conversion waiting icons of the submodals, the conversion of the media content of these submodals can be achieved. Based on this, step S203 above includes the following steps S203-A1 and S203-A2.

[0091] In S203-A1, the terminal device converts the media content of the main modal to the media content of the target submodal in response to a trigger operation on the conversion waiting icon of the target submodal among the N types of submodals.

[0092] In S203-A2, the terminal device replaces the target submodal conversion waiting icon in the creation container with the media content of the target submodal.

[0093] In this method 2, the media content creator triggers the target submodal's conversion waiting icon among the N submodal conversion waiting icons, and the terminal device converts the main modal's media content to the target submodal's media content in response to the trigger operation on the target submodal's conversion waiting icon among the N types of submodals. Next, the terminal device replaces the target submodal's conversion waiting icon in the creation container with the target submodal's media content.

[0094] To illustrate with an example, using the document creation process shown in Figure 5A above, the creation container contains the document content, which is the media content of the generated main modal, and three submodal conversion waiting icons: a video conversion waiting icon, a picture conversion waiting icon, and an audio conversion waiting icon. As shown in Figure 6A, we assume that the creator of the media content taps the video conversion waiting icon, one of the three submodal conversion waiting icons, to trigger the process. The creation application program on the terminal device converts the document content into a video submodal in response to the trigger operation on the video conversion waiting icon. Figure 6B shows the waiting interface indicating that the conversion is in progress, and Figure 6C shows the conversion success. At this point, the video submodal is in the converted state, meaning that the video submodal conversion waiting icon in the creation container has been replaced with the video submodal's media content.

[0095] In this method 2, the modal transformation operation can be understood as a trigger operation by the media content creator on the transformation pending icon of the target submodal.

[0096] This method 2 enables the conversion of media content in the main modal to media content in the submodal by triggering a conversion waiting icon in the submodal. The entire process is simple and time-saving, further reducing the workload of media content creators and improving the efficiency of media content creation.

[0097] Method 3: The editing interface for the media content of the main modal includes a modal conversion option, which enables the conversion of the modal. Based on this, S203 above includes the following steps S203-B1 to S203-B4.

[0098] In the S203-B1, the terminal device displays the main modal editing interface, including modal conversion options, in response to taps on media content in the main modal.

[0099] In S203-B2, the terminal device displays N types of submodal conversion waiting icons in response to a trigger operation for modal conversion options.

[0100] In S203-B3, the terminal device, upon triggering an action on the conversion waiting icon of the target submodal among the N types of submodals, converts the media content of the main modal to the media content of the target submodal and jumps to the submodal editing interface, which includes an editing completion option.

[0101] In S203-B4, the terminal device replaces the target submodal conversion waiting icon in the creation container with the media content of the target submodal in response to a trigger operation for the edit completion option.

[0102] In this method 3, media content for the main modal is generated according to the steps described above, and this generated media content for the main modal is displayed in the creation container. The creator of the media content taps the generated media content for the main modal in the creation container, and the creation application program on the terminal device jumps to the main modal editing interface in response to the trigger operation on the media content for the main modal. The creator of the media content can re-edit the media content for the main modal in this main modal editing interface.

[0103] Furthermore, this main modal editing interface includes a modal conversion option, which the media content creator can trigger. The creation application program on the terminal device displays conversion waiting icons for N types of submodals corresponding to this main modal, in response to the trigger operation on the modal conversion option. This allows the media content creator to trigger these N types of submodal conversion waiting icons to perform submodal conversion. For example, the media content creator taps the conversion waiting icon for the target submodal among the N types of submodal conversion waiting icons, and the creation application program on the terminal device converts the media content of the main modal to the media content of the target submodal in response to the trigger operation on the target submodal conversion waiting icon. After the conversion to the target submodal's media content is successful, the creation application program on the terminal device jumps to the submodal editing interface, where the media content creator can edit the media content of the currently generated target submodal. This submodal editing interface includes an "edit complete" option. When the media content creator taps this option, the creation application program on the terminal device, in response to the trigger operation of selecting this option, replaces the target submodal conversion waiting icon in the creation container with the media content of the target submodal, i.e., displays the generated media content of the target submodal in the creation container.

[0104] To illustrate with an example, using the document creation process shown in Figure 5A above, the creation container contains the document content, which is the media content of the generated main modal, and three submodal conversion waiting icons: a video conversion waiting icon, a picture conversion waiting icon, and an audio conversion waiting icon. As shown in Figure 7A, we assume that the creator of the media content triggers the process by tapping the document content, which is the media content of the main modal. In response to the tap operation on the document media content, the creation application program on the terminal device jumps to the document editing interface, where the creator of the media content can re-edit the document media content.

[0105] As shown in Figure 7B, this document editing interface includes a modal conversion option, which the media content creator can trigger. The creation application program on the terminal device, in response to the trigger operation on the modal conversion option, displays the interface shown in Figure 7C, specifically, icons for three types of submodal conversion waiting for the document: conversion to video, conversion to picture, and conversion to audio. This allows the media content creator to trigger these three submodal conversion waiting icons; for example, the media content creator taps the video submodal conversion waiting icon. The creation application program on the terminal device, in response to the trigger operation on the video submodal conversion waiting icon, converts the document media content to video submodal media content. Exemplarily, Figure 7D shows a conversion waiting interface indicating that conversion is in progress.

[0106] After the conversion of the video submodal to media content is successful, the creation application program on the terminal device jumps to the submodal editing interface shown in Figure 7E, where the media content creator can edit the media content of the currently generated video submodal. This submodal editing interface includes an editing complete option, and when the media content creator taps this editing complete option, the creation application program on the terminal device displays the interface shown in Figure 7F in response to this trigger operation, that is, it replaces the video submodal conversion waiting icon in the creation container with the media content of the video submodal.

[0107] In some embodiments, when generating media content for the main modal, the media content creator creates the content using the main modal editing interface and generates the media content for the main modal. This main modal editing interface includes conversion options, and the media content creator can perform modal conversion using this interface. The specific modal conversion process is the same as described in S203-B2 to S203-B4 above, so a detailed explanation is omitted here.

[0108] In the embodiments of this application, there are no restrictions on the specific conversion means for converting the media content of the main modal to the media content of the target submodal.

[0109] In some embodiments, the media content of the main modal is converted to the media content of a target submodal based on the media content of the main modal. For example, if the main modal is a video and the submodal is a picture, the video media content is converted to the media content of the picture modal based on the image frames contained in the video media content. For example, all the images contained in the video media content are converted to one or more pictures.

[0110] In some embodiments, in S201 above, in addition to generating media content for the main modal in response to editing operations on the main modal editing interface, a project file for the main modal's media content is generated, which includes materials referenced by the main modal's media content and paths to access those materials. In this case, the step of converting the main modal's media content to the target submodal's media content in S203 above includes the step of converting the main modal's media content to the target submodal's media content based on the main modal's media content and the project file in response to a modal conversion operation.

[0111] In other words, in this embodiment, the media content of the main modal is converted to the media content of the target submodal based on the media content of the main modal and the project file of the main modal's media content.

[0112] As an example, in this embodiment, in addition to the media content of the target submodal, a project file for the media content of this target submodal is generated.

[0113] In the embodiments of this application, no specific type of media modal is selected.

[0114] In some embodiments, the embodiments of the present invention include four types of media modals: video, audio, picture, and document, in which case, when modal conversion is performed, twelve types of conversion logic are generated for each of the four types.

[0115] The conversion logic relating to the embodiment of this application will be described below.

[0116] Example 1: A project to be converted to video. Specifically, it converts a picture, document, or audio into video, and builds a project file by extracting visual elements from the media content of the input modal (i.e., the main modal) and the project file, and then constructs a video project by combining it with the cited source files using data segments with timestamp tokens. The video intelligent algorithm is primarily responsible for extracting and combining screen layers, text layers, and audio layers from the input source to generate video media content and project files for the video media content.

[0117] For example, if the main modal is a picture and the target submodal is a video, the terminal device invokes a pre-configured picture-to-video conversion algorithm in response to a trigger on the video submodal by the creator of the media content, and uses this picture-to-video conversion algorithm to convert the picture media content into media content for the video modal. In the embodiments of the present invention, there are no restrictions on the specific type of picture-to-video conversion algorithm. In one possible implementation, the terminal device uses the picture-to-video conversion algorithm to extract material from the screen layer in the picture, such as objects such as people, animals, buildings, and landscapes in the picture, and creates a video based on the material from the screen layer, for example, with at least one visual element as one video frame, and then converts this picture into a video. In another possible implementation, the terminal device uses the picture-to-video conversion algorithm to trim the picture based on the size requirements of the video frames in a pre-configured video template, add transition effects between video frames based on the video template, apply filters to the video frames, optionally add an opening or ending, music, etc., and then converts this picture into a video. At the same time, this picture and its project file will be used as the project file for this video media content.

[0118] Furthermore, for example, if the main modal is a document and the target submodal is a video, the terminal device invokes a pre-configured document-to-video conversion algorithm in response to a trigger on the video submodal by the media content creator, and converts the document media content into video media content using this document-to-video conversion algorithm. In the embodiments of this application, there are no restrictions on the specific type of document-to-video conversion algorithm. In one possible implementation, the terminal device extracts the material of the text layer in the document using the document-to-video conversion algorithm and generates picture and text video content based on this text layer material. Optionally, the terminal device may apply special effects, filter stickers, etc., to the picture and text video content using the document-to-video conversion algorithm, where the beautification effects such as special effects and filter stickers may be defaulted by a template or selected by the media content creator themselves. Simultaneously, this document and its project file become the project file for this video media content.

[0119] Furthermore, for example, if the main modal is audio and the target submodal is video, the terminal device invokes a pre-configured audio-to-video conversion algorithm in response to a trigger on the video submodal by the media content creator, and converts the audio media content into media content for the video modal using this audio-to-video conversion algorithm. In the embodiments of this application, there are no restrictions on the specific type of audio-to-video conversion algorithm. In one possible implementation, the terminal device extracts the audio layer material from the audio content using the audio-to-video conversion algorithm and generates audio-video content based on this audio layer material. Optionally, the terminal device may add special effects, filter stickers, subtitles, etc., to the audio-video content using the audio-to-video conversion algorithm, where these beautification effects such as special effects, filter stickers, and subtitles may be defaulted by a template or selected by the media content creator themselves. Simultaneously, this audio and the project file of this audio become the project file for this video media content.

[0120] Example 2: A project to convert to audio. That is, extract audio material from the media content and project file of the input modal, or generate audio media content and a project file of audio media content based on text.

[0121] For example, if the main modal is a picture and the target submodal is audio, the terminal device, in response to a trigger on the audio submodal by the media content creator, invokes a pre-configured picture-to-audio conversion algorithm and converts the picture media content into media content for the audio modal using this algorithm. In the embodiments of the present invention, there are no restrictions on the specific type of picture-to-audio conversion algorithm. In one possible implementation, the terminal device uses the picture-to-audio conversion algorithm to extract text content from the picture, for example, extracting characters in the picture as text content, and then creates audio based on the text content, for example, converting the text content in the picture into speech, and further forming audio media content. Simultaneously, this picture and its project file become the project file for this audio media content.

[0122] Furthermore, for example, if the main modal is a document and the target submodal is audio, the terminal device invokes a pre-configured document-to-audio conversion algorithm in response to a trigger on the audio submodal by the media content creator, and converts the document media content into audio media content using this document-to-audio conversion algorithm. In the embodiments of the present application, there are no restrictions on the specific type of document-to-audio conversion algorithm. In one possible implementation, the terminal device extracts the material of the text layer in the document using the document-to-audio conversion algorithm and generates audio content based on this material of the text layer. For example, the text content in the document is converted into a speech form, and further forms audio media content. At the same time, this document and its project file become the project file for this audio media content.

[0123] Furthermore, for example, if the main modal is video and the target submodal is audio, the terminal device invokes a pre-configured video-to-audio conversion algorithm in response to a trigger on the audio submodal by the media content creator, and converts the video media content into media content for the audio modal using this video-to-audio conversion algorithm. In the embodiments of the present application, there are no restrictions on the specific type of video-to-audio conversion algorithm. In one possible implementation, the terminal device extracts the audio layer material from the video content using the video-to-audio conversion algorithm and generates audio content based on this audio layer material. For example, it extracts subtitles or documents and audio information from the video, converts this information into audio form to obtain audio media content. At the same time, this video and its project file become the project file for this audio media content.

[0124] Example 3: A project to be converted to a picture. Specifically, screen text elements are extracted from the input modal's media content and project file, mapped to a picture intelligent template, and picture media content and a project file for the picture media content are generated.

[0125] For example, if the main modal is a video and the target submodal is a picture, the terminal device, in response to a trigger on the picture submodal by the media content creator, invokes a pre-configured video-to-picture conversion algorithm and uses this video-to-picture conversion algorithm to convert the video media content into media content for the picture modal. In the embodiments of the present invention, there are no restrictions on the specific type of video-to-picture conversion algorithm. In one possible implementation, the terminal device uses the video-to-picture conversion algorithm to map the video frames contained in the video media content to a picture intelligent template corresponding to the video-to-picture conversion algorithm, and integrates them into one or more pictures. At the same time, this video and its project file become the project file for this picture media content.

[0126] Furthermore, for example, if the main modal is a document and the target submodal is a picture, the terminal device invokes a pre-configured document-to-picture conversion algorithm in response to a trigger on the picture submodal by the creator of the media content, and converts the document media content into picture media content using this document-to-picture conversion algorithm. In the embodiments of the present application, there are no restrictions on the specific type of document-to-picture conversion algorithm. In one possible implementation, the terminal device uses the document-to-picture conversion algorithm to extract the material of the text layer in the document, maps this material of the text layer to a picture intelligent template corresponding to the document-to-picture conversion algorithm, and integrates it into one or more pictures. At the same time, this document and its project file become the project file for this picture media content.

[0127] Furthermore, for example, if the main modal is audio and the target submodal is a picture, the terminal device invokes a pre-configured audio-to-picture conversion algorithm in response to a trigger on the picture submodal by the media content creator, and converts the audio media content into picture media content using this audio-to-picture conversion algorithm. In the embodiments of the present application, there are no restrictions on the specific type of audio-to-picture conversion algorithm. In one possible implementation, the terminal device extracts character elements from the audio using the audio-to-picture conversion algorithm, maps these character elements to a picture intelligent template corresponding to the audio-to-picture conversion algorithm, and integrates them into one or more pictures. Simultaneously, this audio and its project file become the project file for the picture media content.

[0128] Example 4: A project to convert to a document. Specifically, it extracts text-picture content from the input modal's media content and project file, and constructs unstyled document media content and a project file of document media content according to a logical sequence.

[0129] For example, if the main modal is a video and the target submodal is a document, the terminal device, in response to a trigger on the document submodal by the media content creator, invokes a pre-configured video-to-document conversion algorithm and uses this algorithm to convert the video media content into media content for the document modal. In the embodiments of this application, there are no restrictions on the specific type of video-to-document conversion algorithm. In one possible implementation, the terminal device uses the video-to-document conversion algorithm to extract text information from the video, for example, text information from text pictures in video frames and text information from video subtitles, and maps this text information to a document intelligent template to generate document media content. Simultaneously, this video and its project file become the project file for this document media content.

[0130] For example, if the main modal is a picture and the target submodal is a document, the terminal device, in response to a trigger on the document submodal by the media content creator, invokes a pre-configured picture-to-document conversion algorithm and uses this algorithm to convert the picture media content into media content for the document modal. In the embodiments of this application, there are no restrictions on the specific type of picture-to-document conversion algorithm. In one possible implementation, the terminal device uses the picture-to-document conversion algorithm to extract text information from the picture, for example, extract character information from the picture, map the text information to a document intelligent template, and generate document media content. Simultaneously, this picture and its project file become the project file for this document media content.

[0131] For example, if the main modal is audio and the target submodal is a document, the terminal device, in response to a trigger on the document submodal by the media content creator, invokes a pre-configured audio-to-document conversion algorithm, and uses this algorithm to convert the audio media content into media content for the document modal. In the embodiments of this application, there are no restrictions on the specific type of audio-to-document conversion algorithm. In one possible implementation, the terminal device uses the audio-to-document conversion algorithm to convert the audio information in the audio into text information, maps the text information to a document intelligent template, and generates document media content. Simultaneously, this audio and its project file become the project file for this document media content.

[0132] As can be seen from the above, in the embodiments of the present application, when converting the media content of the main modal to the media content of the target submodal, content corresponding to the target submodal is extracted from the media content of the main modal, and then the media content of the target submodal is generated based on the extracted content corresponding to the target submodal. For example, if the main modal is a video and the target submodal is a picture, at least one picture is extracted from the video, and then picture media content is generated based on the extracted at least one picture. Also, for example, if the main modal is a video and the target submodal is a document, text information is extracted from the video, and then document media content is generated based on the extracted text information.

[0133] In the embodiments of the present invention, the creation application program automatically invokes the conversion logic according to the input modal currently initiating the intelligent conversion and the target output modal. If there are multiple modes for a certain type of logic, the media content creator can either actively select one of the logics or implicitly execute one of them. For example, in a project to convert to a picture, there are two logics: one to convert to a long picture and one to convert to a short picture. The media content creator can either select one of these two logics to perform the picture conversion, or they can set one of the logics as the default to perform the picture conversion.

[0134] Following the steps described above, the terminal device converts the media content of the main modal to the media content of the target submodal in response to the modal conversion operation, and then performs the following step S204.

[0135] In S204, the terminal device displays the media content of the generated main modal and the media content of the target submodal.

[0136] In some embodiments, the terminal device displays the media content of the generated main modal and the media content of the target submodal in the same creation container.

[0137] In some implementations, both the media content of the main modal and the media content of the target submodal are displayed in the creation container as floating icons.

[0138] In the embodiments of this invention, both the media content of the generated main modal and the media content of the target submodal can be re-edited.

[0139] Based on this, in some embodiments, the method of the embodiment of the present application further includes the following steps 11 and 12.

[0140] In step 11, the terminal device displays a submodal editing interface containing several types of primary editing tools in response to a trigger operation on the media content of the target submodal.

[0141] In step 12, the terminal device displays the edited media content of the target submodal in response to editing the media content of the target submodal using multiple types of first editing tools.

[0142] Specifically, the media content creator taps the media content of the target submodal in the creation container, and the creation application program on the terminal device displays a submodal editing interface containing multiple types of first editing tools in response to the trigger operation on the media content of the target submodal. The media content creator can edit the media content of the target submodal using these multiple types of first editing tools, and the creation application program on the terminal device displays the edited media content of the target submodal in response to the media content creator editing the media content of the target submodal using the multiple types of first editing tools.

[0143] In some embodiments, the multiple types of first editing tools in the submodal editing interface described above include a new creation tool, which is used to create the media content of the current target submodal as the media content of the main modal of another creation project. Based on this, in the submodal editing interface, when the creator of the media content triggers this new creation tool, the creation application program on the terminal device creates the media content of the target submodal as the media content of the main modal of a second media creation project in response to the trigger operation on the new creation tool.

[0144] In other words, by triggering the New creation tool in the submodal editing interface, the media content of the submodal in the first media creation project can be newly created as the media content of the main modal in the second media creation project.

[0145] In some embodiments, the method of the embodiment of the present application further includes the following steps 21 and 22.

[0146] In step 21, the terminal device displays a main modal editing interface for the main modal media content, including several types of second editing tools, in response to a trigger operation on the main modal media content.

[0147] In step 22, the terminal device displays the edited media content of the main modal, depending on the editing of the media content of the main modal using multiple types of second editing tools.

[0148] Specifically, the media content creator taps the media content in the main modal within the creation container, and the creation application program on the terminal device displays a main modal editing interface containing multiple types of second editing tools in response to the trigger operation on the media content in the main modal. The media content creator can edit the media content in the main modal using these multiple types of second editing tools. The creation application program on the terminal device displays the edited media content in the main modal in response to the media content creator editing the media content in the main modal using these multiple types of second editing tools.

[0149] In some embodiments, the multiple types of second editing tools in the main modal editing interface described above include a copy tool, which is used to create the media content of the current main modal as the media content of the main modal of another creation project. Based on this, when the creator of the media content triggers this copy tool in the main modal editing interface, the creation application program on the terminal device creates the media content of the main modal of the first media creation project as the media content of the main modal of the third media creation project in response to the trigger operation on the copy tool.

[0150] In other words, by triggering the copy tool in the main modal editing interface, the media content of the main modal in the first media creation project can be copied as the media content of the main modal in the third media creation project.

[0151] Here, the third media creation project may be the same as the second media creation project described above, or it may be different from the second media creation project described above, but there are no restrictions in the embodiments of this application.

[0152] In some embodiments, after generating the media content for the main modal and the target submodal of the first media creation project according to the steps described above, the media content for the main modal and the target submodal of the first media creation project can be stored in a cloud storage file.

[0153] In some embodiments, the creation application program of the embodiments of the present invention further provides operation options, including operations such as renaming, deleting, sharing, and moving. The creator of media content can operate on at least one media content from among the media content of the main modal and the media content of the target submodal by operations included in the operation options. For example, the creator of media content triggers a target operation in the operation options, and the creation application program on the terminal device, in response to the trigger to the target operation in the operation options, performs a target operation, including renaming, deleting, sharing, or moving, on at least one media content from among the media content of the main modal and the media content of the target submodal.

[0154] For example, a media content creator can delete at least one media content from the main modal and target submodal by triggering a delete operation. Alternatively, a media content creator can rename at least one media content from the main modal and target submodal by triggering a rename operation. Alternatively, a media content creator can share at least one media content from the main modal and target submodal by triggering a share operation. Alternatively, a media content creator can move the position of at least one media content from the main modal and target submodal within the creation container by triggering a move operation.

[0155] In some embodiments, to maintain consistency between the terminal and the cloud, a target operation is performed on at least one media content from the main modal media content and the target submodal media content on the terminal side. Then, the target operation performed on at least one media content from the main modal media content and the target submodal media content is synchronized with a cloud storage file, so that a target operation is performed on at least one media content in the cloud storage file. For example, after receiving a target operation, the cloud performs a target operation on at least one media content from the main modal media content and the target submodal media content in the same manner as on the terminal side, maintaining consistency between the content stored on both sides.

[0156] In some embodiments, the above operation options are placed in the additional functionality options of the creation container.

[0157] In some embodiments, the target operation described above is a copy operation, which has substantially the same functionality as the copy tool in the main modal editing interface described above. For example, when a media content creator triggers a copy operation in this operation option, the creation application program, in response to this trigger, copies at least one media content from the main modal media content and the target submodal media content as the main modal media content of a new media creation project. For example, it copies the main modal media content of a first media creation project as the main modal media content of a third media creation project, and copies the target submodal media content of a first media creation project as the main modal media content of a second media creation project.

[0158] In some embodiments, the creation application program of the embodiment of the present invention further includes an export option, in which case the method of the embodiment of the present invention further includes the step of the creation application program exporting at least one media content from the main modal media content and the target submodal media content of a first media creation project in response to a trigger operation to the export option of the media content creator.

[0159] In this embodiment, the creation application program supports content export of metacreation projects, and the content export exposes APIs based on the export of metacreation project files, as well as video, audio, picture, and document exports, and supports format conversion for exporting and synchronizing metacreations as a single media file.

[0160] Depending on the dimensions of the export objective, embodiments of the present invention can export as a locally stored media file, or they can call an interface to asynchronously render and export on the service side and obtain the target stored file.

[0161] From the perspective of export logic, embodiments of the present invention support the encapsulation of four types of media in common formats. For example, video files support export to encapsulated files such as mp4 in multiple types of video encoding formats. Document files support export to TXT plain text, or documents such as Word or PDF, or to HTML files.

[0162] Optionally, the above export options may be placed in the Additional Features options for the creation container.

[0163] In some embodiments, the creation application program of the embodiments of the present invention further includes a publish option, in which case the method of the embodiments of the present invention further includes the step of the creation application program publishing at least one media content from the main modal media content and the target submodal media content of a first media creation project to a third-party platform in response to a trigger operation to the publish option by the creator of the media content.

[0164] In other words, in the embodiments of the present invention, the content of the created first media creation project can be published to an internet platform. For example, at least one of the media content from the main modal and the target submodal of the first media creation project can be published to a third-party platform via the shared interface of the front-end SDK of the creation application program or the public interface of the back-end API.

[0165] Optionally, the above publishing options may be placed in the additional features options for the creation container.

[0166] An embodiment of the present invention provides a method for creating media content, comprising the steps of: displaying a main modal editing interface in response to a trigger operation; generating media content for a main modal in response to editing operations on the main modal editing interface; converting the media content for a main modal to media content for a target submodal in response to a modal conversion operation; and displaying the generated media content for the main modal and the media content for the target submodal. That is, the embodiment of the present invention supports conversion between media content of multiple types of modals, and a media content creator can create media content for a main modal using this application program, and then convert this media content for a main modal to media content for at least one submodal different from the main modal by inputting a modal conversion operation. The entire modal conversion process is performed by the application program based on its own modal conversion algorithm, eliminating the need for media content creator involvement, and further reducing the workload of media content creators and improving the efficiency of media content creation.

[0167] The above describes the process of creating media content according to the embodiment of this application. Below, we will describe the overall flow of creating media content according to the embodiment of this application.

[0168] Figure 8 is a schematic flowchart of a media content creation method provided in one embodiment of the present invention, and Figure 9 is a schematic diagram of the interaction between the creation application program and the media content creator, and within the creation application program, in a terminal device according to an embodiment of the present invention.

[0169] As shown in Figures 8 and 9, the method of the embodiment of the present application includes the following steps S301 to S307.

[0170] In S301, you will create the media content for the main modal of the first media creation project.

[0171] Specifically, the creation application program on the terminal device displays a main modal editing interface in response to a trigger operation by the media content creator on the displayed meta-creation new creation option, and the creation application program on the terminal device generates media content for the main modal in response to an editing operation by the media content creator on the main modal editing interface.

[0172] As shown in Figure 9, the creation application program of the embodiment of the present invention includes a UI interface, a local metacreation SDK, and several types of APIs. The local metacreation SDK is mainly used for creating metacreation projects, implementing intelligent modal transformations, and updating / merging metacreation projects.

[0173] In the embodiments of this application, if the media modal is video, audio, picture, and document, the creation application program includes a video API, an audio API, a picture API, and a document API, and optionally the creation application program may further include a business API.

[0174] When generating media content for the main modal of the first media creation project, the media content creator triggers the "Create New Meta Creation" option on the UI interface and selects the main modal. The UI interface displays the main modal editing interface in response to the media content creator's actions. The media content creator creates the media content for the main modal in this main modal editing interface. The local meta creation SDK creates the media content for the main modal and the corresponding project file by calling the creation project algorithm for the main modal. In the embodiment of this application, the media content for the main modal is referred to as the meta creation draft.

[0175] For example, as shown in Figure 9, if the main modal is a video, the local SDK creates a video project via the video API. If the main modal is audio, the local SDK creates an audio project via the audio API. If the main modal is a picture, the local SDK creates a picture project via the picture API. If the main modal is a document, the local SDK creates a document project via the document API.

[0176] The aforementioned project creation includes the creation of the main modal media content and the meta-creation project file.

[0177] In some embodiments, the media content of the main modal is contained within a metacreation project file. That is, the metacreation project file contains the media content of the created main modal, the materials related to the main modal's media content, and the storage paths for those materials.

[0178] For the specific implementation process of S301 described above, please refer to the explanations in S201 and S202 above.

[0179] In some embodiments, as shown in Figure 9, the creation application program can, after generating the media content of the main modal of the first media creation project, call a business API to create a new cloud disk metacreation, that is, store the media content of the main modal of the first media creation project in a cloud storage file, for example, on a cloud disk.

[0180] In S302, you edit the media content of the main modal.

[0181] Specifically, the process includes the steps of: a creation application program on a terminal device displays a main modal editing interface for the media content of the main modal, which includes multiple types of second editing tools, in response to a trigger operation on the media content of the main modal; and a creation application program on a terminal device displays the edited media content of the main modal in response to editing the media content of the main modal using the multiple types of second editing tools.

[0182] In other words, when a media content creator selects to edit the media content in the main modal from the UI interface, they directly invoke the editing tools for the corresponding modal and enter the editing state. For example, as shown in Figure 9, if the main modal is a video, the editing tools for the video modal are invoked by calling the video API, and the video modal editing interface is entered.

[0183] In S303, you edit the media content of the submodal.

[0184] When a media content creator selects to edit a submodal from the UI interface, they must first make a business logic decision. If they select to edit a submodal for the first time, they must generate the submodal based on the main modal in order to enter the editing state for that submodal.

[0185] Specifically, it is determined whether media content exists for the selected submodal. If it does not exist, S304 is executed. If media content exists for the selected submodal, it is determined whether the media content of the current main modal has been updated. If it has been updated, S303-A is executed; otherwise, S303-B and S303-C are executed.

[0186] In S303-A, the application program created on the terminal device generates submodal media content from the main modal media content.

[0187] For example, the creation application program on the terminal device, in response to a trigger operation on the conversion waiting icon of the target submodal among the N types of submodals, converts the media content of the main modal to the media content of the target submodal, and then replaces the conversion waiting icon of the target submodal in the creation container with the media content of the target submodal. Next, the following S303-B and S303-C are executed.

[0188] In S303-B, the creation application program on the terminal device displays a submodal editing interface containing multiple types of primary editing tools in response to a trigger operation on the media content of the target submodal.

[0189] In S303-C, the creation application program on the terminal device displays the edited media content of the target submodal in response to editing the media content of the target submodal using multiple types of first editing tools.

[0190] For example, if the main modal is a video, as shown in Figure 9, the intelligent modal conversion module in the local SDK implements a video-to-audio conversion project by calling the audio API, a video-to-picture conversion project by calling the picture API, and a video-to-document conversion project by calling the document API.

[0191] S304 converts the content of the first media creation project into the content of another media creation project.

[0192] In one example, a creation application program on a terminal device displays a submodal editing interface containing multiple types of first editing tools in response to a trigger operation on the media content of the target submodal, and these multiple types of first editing tools include a new creation tool. Then, in response to a trigger operation on the new creation tool, the creation application program on the terminal device creates the media content of the target submodal as the media content of the main modal of a second media creation project.

[0193] In other words, the media content of the submodal in the first media creation project is newly created as the media content of the main modal in the other media creation project.

[0194] In another example, a creation application program on a terminal device displays a main modal editing interface for the main modal's media content, including several types of second editing tools, in response to a trigger operation on the main modal's media content, with a copy tool being one of the several types of second editing tools. Then, in response to a trigger operation on the copy tool, the creation application program on the terminal device copies the main modal's media content as the main modal's media content for a third media creation project.

[0195] In other words, the media content of the main modal of the first media creation project is newly created as the media content of the main modal of the other media creation project.

[0196] S305 is used to manipulate the content of the first media creation project.

[0197] For example, an application program created on a terminal device performs a target operation, including rename, delete, share, or move, on at least one of the media contents from the main modal media content and the target submodal media content, in response to a trigger for a target operation in the operation options.

[0198] In some embodiments, the media content of the main modal and the media content of the target submodal of the first media creation project are stored in a cloud storage file. Based on this, target operations are synchronized with the cloud storage file so that target operations are performed on at least one media content in the cloud storage file, thereby maintaining consistency between the media content stored on the terminal and the cloud.

[0199] In some embodiments, if the operation options include a copy operation, the creation application program on the terminal device copies at least one of the media content from the main modal and the target submodal as the media content for the main modal of a new media creation project, in response to a trigger for the copy operation.

[0200] S306 exports the content of the first media creation project.

[0201] For example, a creation application program on a terminal device exports at least one media content from the main modal and target submodal of the first media creation project in response to a trigger operation on the export option.

[0202] In some implementations, when exporting media content, the corresponding project file is also exported along with it.

[0203] For example, as shown in Figure 9, when exporting a video file, the creation application program on the terminal device exports the video file by calling the video API through the user interface of the creation application program in response to a trigger operation on the export option. The video file includes video media content and a corresponding project file.

[0204] In S307, we will release the content of the first media creation project.

[0205] For example, a creation application program on a terminal device, in response to a trigger operation on the publishing option, publishes at least one media content from the main modal media content and the target submodal media content of the first media creation project to a third-party platform.

[0206] In the embodiments of this application, through a metacreation flow, creators can create works for different social media platforms simultaneously based on a given theme, and can quickly complete the entire generation, editing, and publishing flow using intelligent templates. Furthermore, based on the intelligent transformation tool of metacreation, the embodiments of this application effectively combine multiple types of creation tools and interfaces according to the type of input / output source and the purpose of operation, and further maximize the decoupling of tools and workflows. In addition, the design of associating main modals and submodals in the embodiments of this application helps creators decompose creation elements from dimensions such as copywriting, audio, visual, and viewing, respectively, achieving relative separation between corresponding project files and cited materials, and enabling the more effective generation of creation templates that are not dependent on specific materials.

[0207] Figures 2 through 9 are merely examples of the present application and should not be understood as limiting the present application.

[0208] While preferred embodiments of the present application have been described in detail with reference to the drawings, the present application is not limited to the specific details of the embodiments described above. Within the scope of the technical concept of the present application, various simple modifications can be made to the technical embodiments of the present application, and all such simple modifications should be included within the scope of protection of the present application. For example, each specific technical feature described in the specific embodiments described above can be combined in any appropriate manner as long as they do not contradict each other, and in order to avoid unnecessary duplication, the present application does not separately describe each possible combination method. Furthermore, for example, various different embodiments of the present application can be combined in any way as long as they do not contradict the concept of the present application, and should be considered to be disclosed in the present application as well.

[0209] The embodiments of the method of the present application have been described in detail above with reference to Figures 10 to 11. Now, embodiments of the apparatus of the present application will be described in detail below.

[0210] Figure 10 is a schematic diagram of a media content creation apparatus provided in one embodiment of the present invention.

[0211] As shown in Figure 10, the media content creation device 10 is applied to terminal equipment. A first display unit 110 that displays the main modal editing interface in response to a trigger operation, A processing unit 120 generates media content for the main modal in response to editing operations on the main modal editing interface, A conversion unit 130 converts the media content of the main modal to the media content of the target submodal in response to a modal conversion operation, The system includes a second display unit 140 that displays the media content of the generated main modal and the media content of the target submodal.

[0212] In some embodiments, the generation unit 120 specifically generates media content for the main modal in response to the editing operation on the main modal editing interface, determines N types of submodals corresponding to the main modal, where the N types of submodals are different from the main modal and N is a positive integer. The second display unit 140 further displays the media content for the main modal and the conversion waiting icons for the N types of submodals in the same creation container.

[0213] In some embodiments, the modal conversion operation is a trigger operation on the conversion waiting icon of the target submodal, and the conversion unit 130 specifically converts the media content of the main modal to the media content of the target submodal in response to the trigger operation on the conversion waiting icon of the target submodal among the N types of submodals, and replaces the conversion waiting icon of the target submodal in the creation container with the media content of the target submodal.

[0214] In some embodiments, the modal conversion operation is a trigger operation to a modal conversion option, and the conversion unit 130 specifically displays the main modal editing interface including the modal conversion option in response to a tap operation on the media content of the main modal, and displays conversion waiting icons for the N types of submodals in response to a trigger operation to the modal conversion option, and converts the media content of the main modal to the media content of the target submodal in response to a trigger operation to the conversion waiting icon of the target submodal among the N types of submodals, jumps to a submodal editing interface including an editing complete option, and further replaces the conversion waiting icon of the target submodal in the creation container with the media content of the target submodal in response to a trigger operation to the editing complete option.

[0215] In some embodiments, the processing unit 120 specifically generates media content for the main modal and a project file for the media content of the main modal in response to editing operations on the main modal editing interface, the project file containing material referenced by the media content of the main modal and paths to access the material. The conversion unit 130 specifically converts the media content of the main modal to media content for the target submodal based on the media content of the main modal and the project file in response to the modal conversion operation.

[0216] In some embodiments, the second display unit 140 specifically displays the media content of the main modal in a first area of ​​the creation container and displays the conversion waiting icons for the N types of submodals in a second area of ​​the creation container, the first area being larger than the second area.

[0217] In some embodiments, the processing unit 120 further displays a submodal editing interface including a plurality of first editing tools in response to a trigger operation on the media content of the target submodal, and displays the edited media content of the target submodal in response to editing the media content of the target submodal using the plurality of first editing tools.

[0218] In some embodiments, the plurality of first editing tools include a new creation tool, and the processing unit 120 further creates the media content of the target submodal as the media content of the main modal of a second media creation project in response to a trigger operation on the new creation tool.

[0219] In some embodiments, the processing unit 120 further displays a main modal editing interface for the media content of the main modal, which includes a plurality of second editing tools, in response to a trigger operation on the media content of the main modal, and displays the edited media content of the main modal in response to editing the media content of the main modal with the plurality of second editing tools.

[0220] In some embodiments, the multiple types of second editing tools include a copy tool, and the processing unit 120 further copies the media content of the main modal as the media content of the main modal of a third media creation project in response to a trigger operation on the copy tool.

[0221] In some embodiments, the processing unit 120 further stores the media content of the main modal and the media content of the target submodal of the first media creation project in a cloud storage file.

[0222] In some embodiments, the processing unit 120 further performs a target operation, including rename, delete, share, or move, on at least one of the media contents from the main modal media content and the target submodal media content, in response to a trigger for a target operation in the operation options.

[0223] In some embodiments, the processing unit 120 further synchronizes the target operation to the cloud storage file so that the target operation is performed on the at least one media content in the cloud storage file.

[0224] In some embodiments, the processing unit 120 specifically copies at least one of the media content from the main modal and the target submodal as the media content for the main modal of a new media creation project, in response to a trigger for the copy operation.

[0225] In some embodiments, the processing unit 120 further publishes at least one of the media content from the main modal and the target submodal of the first media creation project to a third-party platform in response to a trigger operation to the publish option.

[0226] In some embodiments, the processing unit 120 further exports at least one media content from the main modal media content and the target submodal media content of the first media creation project in response to a trigger operation to the export option.

[0227] It should be understood that the embodiments of the apparatus and the embodiments of the method can correspond to each other, and similar descriptions can refer to the embodiments of the method. To avoid redundancy, detailed descriptions are omitted here. Specifically, the apparatus 10 shown in Figure 10 can perform the embodiments of the method described above, and the above and other operations and / or functions of each module in the apparatus 10 are for realizing the embodiments of the method described above, and for the sake of brevity, detailed descriptions are omitted here.

[0228] The apparatus of the embodiment of the present application has been described above with reference to the drawings, from the perspective of a functional module. It should be understood that if this functional module can be realized in hardware form, it can also be realized by instructions in software form, and furthermore, by a combination of hardware and software modules. Specifically, each step of the embodiment of the method in the embodiment of the present application can be completed by integrated logic circuits and / or instructions in software form in the hardware of a processor, and the steps of the method disclosed in accordance with the embodiment of the present application can be executed and completed by a hardware decode processor, or by a combination of hardware and software modules in a decode processor. Optionally, the software module can be placed in a storage medium mature in the art, such as random access memory, flash memory, read-only memory, programmable read-only memory, electrically erasable and rewritable programmable memory, or registers. This storage medium is placed in memory, and the processor reads the information in memory and, together with its hardware, completes the steps of the embodiment of the method described above.

[0229] Figure 11 is a schematic block diagram of an electronic device provided in an embodiment of the present application, which may be the terminal device described above.

[0230] As shown in Figure 11, this electronic device 40 is The system may include a memory 41 and a processor 42, the memory 41 storing a computer program and transmitting the program code to the processor 42. In other words, the processor 42 can call and execute a computer program from the memory 41 to implement the method in the embodiment of the present application.

[0231] For example, this processor 42 can execute an embodiment of the method described above in accordance with the instructions in this computer program.

[0232] In some embodiments of the present application, this processor 42 is This includes, but is not limited to, general-purpose processors, digital signal processors (DSPs), application-specific integrated circuits (ASICs), field programmable gate arrays (FPGAs) or other programmable logic devices, discrete gate or transistor logic devices, and discrete hardware components.

[0233] In some embodiments of the present application, this memory 41 is This includes, but is not limited to, volatile memory and / or non-volatile memory. Here, non-volatile memory may be read-only memory (ROM), programmable read-only memory (PROM), erasable programmable read-only memory (Erasable PROM, EPROM), electrically erasable programmable read-only memory (Electrically Erasable EPROM, EEPROM), or flash memory. Volatile memory may be random access memory (RAM) that functions as an external cache. This is an illustrative and non-restrictive description, but many forms of RAM are available, such as static random access memory (Static RAM, SRAM), dynamic random access memory (DRAM), synchronous dynamic random access memory (Synchronous DRAM, SDRAM), double data rate synchronous dynamic random access memory (Double Data Rate SDRAM, DDR SDRAM), enhanced synchronous dynamic random access memory (Enhanced SDRAM, ESDRAM), synch-link dynamic random access memory (synch-link DRAM, SLDRAM), and direct Rambus random access memory (Direct Rambus RAM, DR RAM).

[0234] In some embodiments of the present application, the computer program can be divided into one or more modules, which are stored in the memory 41 and executed by the processor 42 to complete the method provided in the present application. The one or more modules may be a series of computer program instruction segments that can accomplish a specific function and are used to describe the execution process of the computer program in the video production device.

[0235] As shown in Figure 11, this electronic device 40 is This may further include a transceiver 43 that can be connected to the processor 42 or memory 41.

[0236] Here, the processor 42 can control the transceiver 43 to communicate with other devices, specifically by transmitting information or data to other devices or receiving information or data transmitted from other devices. The transceiver 43 may include a transmitter and a receiver. The transceiver 43 may further include antennas, and the number of antennas may be one or more.

[0237] It should be understood that each component in this video production device is connected via a bus system, which includes a power bus, a control bus, and a status signal bus, in addition to a data bus.

[0238] The present invention further provides a computer storage medium storing a computer program that, when executed by a computer, causes the computer to perform the method of the embodiment of the above method. Alternatively, embodiments of the present invention further provide a computer program product that, when executed by a computer, includes instructions that cause the computer to perform the method of the embodiment of the above method.

[0239] When implemented using software, all or part of the implementation can be in the form of a computer program product. This computer program product includes one or more computer instructions. When this computer program instruction is loaded into a computer and executed, all or part of the flow or function described according to the embodiments of this application is generated. This computer may be a general-purpose computer, a dedicated computer, a computer network, or other programmable device. This computer instruction may be stored in a computer-readable storage medium or transmitted from one computer-readable storage medium to another. For example, this computer instruction may be transmitted from one website, computer, server, or data center to another website, computer, server, or data center via a wired connection (e.g., coaxial cable, optical fiber, digital subscriber line (DSL)) or wireless connection (e.g., infrared, radio, microwave, etc.). This computer-readable storage medium may be any available medium accessible to a computer, or a data storage device such as a server or data center that integrates one or more available media. These usable media can be magnetic media (e.g., floppy disks, hard disks, magnetic tapes), optical media (e.g., digital video discs (DVDs)), or semiconductor media (e.g., solid state disks (SSDs)).

[0240] Those skilled in the art will recognize that the modules and algorithmic steps of each example described in accordance with the embodiments disclosed herein can be implemented by electronic hardware or by a combination of computer software and electronic hardware. Whether these functions are performed in hardware or software depends on the specific application and design constraints of the technical embodiment. While experts in the art may implement the described functions using different methods for each specific application, such implementations should not be considered beyond the scope of this application.

[0241] In some embodiments provided herein, it should be understood that the presented systems, apparatus, and methods may be implemented in other forms. For example, the embodiments of the apparatus described above are merely schematic, and the division of modules, for instance, is merely a division of logical functions, and other division methods may be possible in actual implementation. For example, multiple modules or components may be combined or integrated into another system, or some features may be ignored or not performed. In other words, the shown or described coupling or direct coupling or communication connection between them may be an indirect coupling or communication connection via some interface, apparatus, or module, and may be in the form of electrical, mechanical, or other.

[0242] Modules described as separate components may or may not be physically separated, and components shown as modules may or may not be physical modules, that is, they may be located in one place or distributed across multiple network units. Some or all of these modules can be selected as needed to achieve the objectives of this embodiment. Furthermore, each functional module in the various embodiments of this application may be integrated into a single processing module, the individual modules may exist physically separately, or two or more modules may be integrated into a single module.

[0243] The above are merely specific embodiments of the present application, but the scope of protection of this application is not limited thereto. Any modifications or substitutions that a person skilled in the art could easily conceive of within the scope of the technical information disclosed herein should also be included within the scope of protection of this application. Therefore, the scope of protection of this application should be in accordance with the scope of protection of its claims. [Explanation of symbols]

[0244] 10 Manufacturing device 40 Electronic equipment 41 memory 42 processors 43 transceivers 101 Terminal equipment 102 Servers 110 First display unit 120 processing units 120 generation units 130 Conversion Unit 140 Second display unit

Claims

1. A method for creating multimodal media content applicable to terminal devices, The steps include: displaying the main modal editing interface in response to a trigger operation, The steps include generating media content for the main modal in response to editing operations on the main modal editing interface, A step of converting the media content of the main modal to the media content of the target submodal in response to a modal conversion operation, wherein the type of media content of the main modal is different from the type of media content of the target submodal. The steps include: playing the generated media content of the main modal and the media content of the target submodal, or displaying icons for the media content of the main modal and the media content of the target submodal; Methods that include...

2. The step of generating media content for the main modal in response to editing operations on the main modal editing interface is as follows: A step of generating media content for the main modal and determining N types of submodals corresponding to the main modal in response to the editing operation on the main modal editing interface, wherein the N types of submodals are different from the main modal and N is a positive integer; The steps include displaying the media content of the main modal and the conversion waiting icons of the N types of submodals in the same creation container. The method according to claim 1, including the method described in claim 1.

3. The modal transformation operation includes a trigger operation on the transformation waiting icon of the target submodal, The step of converting the media content of the main modal to the media content of the target submodal in response to a modal conversion operation is: The steps include: converting the media content of the main modal to the media content of the target submodal in response to a trigger operation on the conversion waiting icon of the target submodal among the N types of submodals; The steps include replacing the conversion waiting icon for the target submodal in the creation container with the media content of the target submodal. The method according to claim 2, including the method described in claim 2.

4. The modal transformation operation described above is a trigger operation on the modal transformation option, The step of converting the media content of the main modal to the media content of the target submodal in response to a modal conversion operation is: The steps include: displaying the main modal editing interface, including the modal conversion options, in response to a tap operation on the media content of the main modal; The steps include: displaying conversion waiting icons for the N types of submodals in response to a trigger operation on the modal conversion option; In response to a trigger operation on the conversion waiting icon of the target submodal among the N types of submodals, the media content of the main modal is converted to the media content of the target submodal, and the user jumps to a submodal editing interface that includes an editing completion option. In response to the trigger operation for the editing completion option, the steps include replacing the conversion waiting icon for the target submodal in the creation container with the media content of the target submodal, and The method according to claim 2, including the method described in claim 2.

5. The step of generating media content for the main modal in response to editing operations on the main modal editing interface is as follows: A step of generating media content for the main modal and a project file for the media content for the main modal in response to editing operations on the main modal editing interface, wherein the project file includes material referenced by the media content for the main modal and paths to access the material. The step of converting the media content of the main modal to the media content of the target submodal in response to a modal conversion operation is: The modal transformation operation includes the step of transforming the media content of the main modal into the media content of the target submodal based on the media content of the main modal and the project file, in response to the modal transformation operation. The method according to any one of claims 1 to 4.

6. The step of displaying the media content of the main modal and the conversion waiting icons of the N types of submodals in the same creation container is: The steps include displaying the media content of the main modal in a first area of ​​the creation container, and displaying the conversion waiting icons for the N types of submodals in a second area of ​​the creation container, wherein the first area is larger than the second area. The method according to any one of claims 2 to 4, including the method described in any one of claims 2 to 4.

7. The steps include: displaying a submodal editing interface, which includes multiple types of first editing tools, in response to a trigger operation on the media content of the target submodal; The steps include: displaying the edited media content of the target submodal in response to editing the media content of the target submodal using the aforementioned multiple types of first editing tools; and The method according to any one of claims 1 to 4, further comprising:

8. The aforementioned multiple types of first editing tools include a new creation tool, The aforementioned method, The process further includes the step of creating the media content of the target submodal as the media content of the main modal of a second media creation project in response to a trigger operation on the aforementioned creation tool, The method according to claim 7.

9. In response to a trigger operation on the media content of the main modal, the steps include displaying a main modal editing interface for the media content of the main modal, which includes multiple types of second editing tools; The steps include: displaying the edited media content of the main modal in response to editing the media content of the main modal using the aforementioned multiple types of second editing tools; and The method according to any one of claims 1 to 4, further comprising:

10. The aforementioned multiple types of second editing tools include a copy tool, The aforementioned method, The process further includes the step of copying the media content of the main modal as the media content of the main modal of a third media creation project in response to a trigger operation on the copy tool. The method according to claim 9.

11. The method according to any one of claims 1 to 4, further comprising the step of storing the media content of the main modal and the media content of the target submodal in a cloud storage file.

12. The method according to any one of claims 1 to 4, further comprising the step of performing a target operation, including rename, delete, share, or move, on at least one media content from among the media content of the main modal and the media content of the target submodal, in response to a trigger for a target operation in the operation options.

13. The method according to claim 12, further comprising the step of synchronizing the target operation to the cloud storage file so that the target operation is performed on the at least one media content in the cloud storage file.

14. The aforementioned target operation is a copy operation, The step of performing a target operation on at least one media content from among the media content of the main modal and the media content of the target submodal in response to a trigger for a target operation in the operation options is: The process includes, in response to a trigger for the copy operation, copying at least one media content from the main modal and the target submodal as the main modal media content of a new media creation project. The method according to claim 12.

15. The method according to any one of claims 1 to 4, further comprising the step of publishing at least one media content from the main modal and the target submodal to a third-party platform in response to a trigger operation on the publishing option.

16. The method according to any one of claims 1 to 4, further comprising the step of exporting at least one media content from the media content of the main modal and the media content of the target submodal in response to a trigger operation to the export option.

17. A device for creating multimodal media content, applicable to terminal devices, A first display unit that displays the main modal editing interface in response to a trigger operation, A processing unit that generates media content for the main modal in response to editing operations on the main modal editing interface, A conversion unit that converts the media content of the main modal to the media content of the target submodal in response to a modal conversion operation, wherein the type of media content of the main modal and the type of media content of the target submodal are different. A second display unit that displays the media content of the generated main modal and the media content of the target submodal. A media content creation device equipped with the following features.

18. It includes a processor and memory for storing computer programs, The processor performs the method according to any one of claims 1 to 4 by calling and executing a computer program stored in the memory. electronic equipment.

19. A computer program configured to cause a computer to perform the method described in any one of claims 1 to 4.

Citation Information

Patent Citations

  • Information display method and device, electronic equipment and storage medium

    CN113473204A

  • Multimedia data processing apparatus, multimedia data processing program, and data structure of multimedia content data

    JP2006197618A

  • Interactive video generation

    JP2019154045A