Interactive multimedia content processing method and apparatus, device, medium and product

Through generative models and machine learning technology, branches and images of interactive multimedia content are generated according to the original story, which solves the problem of low generation efficiency in existing technologies and realizes efficient creation and content expansion with user participation.

WO2025184856A1PCT designated stage Publication Date: 2025-09-11BEIJING ZITIAO NETWORK TECH CO LTD
View PDF 6 Cites 0 Cited by

Patent Information

Application Number
PCT/CN2024/080501
Authority / Receiving Office
WO · WO
Patent Type
Applications
Current Assignee / Owner
Filing Date
2024-03-07
Publication Date
2025-09-11

AI Technical Summary

Technical Problem

The existing interactive multimedia content generation process is complex and difficult to produce efficiently through user-generated content. It requires cooperation from multiple professional fields, resulting in low generation efficiency.

Method used

A generative model is used to generate content text, images, and interactive multimedia content for story branches based on the content text of the original story. Machine learning is used to understand the plot development and expand branch nodes, allowing users to participate in creation and modification.

Benefits of technology

It improves the efficiency of generating interactive multimedia content, reduces resource consumption, enables ordinary users to participate in creation, and increases the diversity of content and the flexibility of creation.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN2024080501_12092025_PF_FP_ABST
    Figure CN2024080501_12092025_PF_FP_ABST
Patent Text Reader

Abstract

The present invention relates to the technical field of computers, and relates to an interactive multimedia content processing method and apparatus, a device, a medium and a product. The interactive multimedia content processing method of the present invention comprises: on the basis of content text of an original story, generating content text of one or more story branches; on the basis of the content text of the original story and the content text of the one or more story branches, generating images, wherein the images include images of characters and images of scenes; and on the basis of the content text of the original story, the content text of the one or more story branches and the images, generating interactive multimedia content.
Need to check novelty before this filing date? Find Prior Art

Description

Interactive multimedia content processing method, device, equipment, medium and product Technical Field

[0001] The present disclosure relates to the field of computer technology, and in particular to a method, apparatus, device, medium, and product for processing interactive multimedia content. Background Art

[0002] Interactive multimedia content integrates multiple elements, such as images, sounds, and text, and provides users with an interactive interface. Examples include interactive stories, interactive movies, interactive TV series, and games. Compared to traditional novels and movies, interactive multimedia content is more engaging and engaging, allowing users to manipulate and interact, and influence the plot's development.

[0003] Interactive multimedia content is generally produced by professional designers and developers, and requires a series of processes such as plot creation, drawing by artists, coding by developers or organizing actors for filming, and post-processing to be completed.

[0004] Summary of the Invention

[0005] This summary is provided to briefly introduce concepts that will be described in detail in the detailed description below. This summary is not intended to identify key features or essential features of the claimed technical solution, nor is it intended to limit the scope of the claimed technical solution.

[0006] According to some embodiments of the present disclosure, a method for processing interactive multimedia content is provided, including: generating content texts of one or more story branches based on the content text of the original story; generating images based on the content text of the original story and the content texts of one or more story branches, wherein the images include images of characters and images of scenes; generating interactive multimedia content based on the content text of the original story, the content texts of one or more story branches, and the images.

[0007] According to other embodiments of the present disclosure, a device for processing interactive multimedia content is provided, including: a text generation module, configured to generate content texts of one or more story branches based on the content text of the original story; an image generation module, configured to generate images based on the content text of the original story and the content text of one or more story branches, wherein the images include images of characters and images of scenes; and a multimedia generation module, configured to generate interactive multimedia content based on the content text of the original story, the content text of one or more story branches, and the images.

[0008] According to some further embodiments of the present disclosure, an electronic device is provided, comprising: a processor; and a memory coupled to the processor, for storing instructions, which, when executed by the processor, causes the processor to execute the method for processing interactive multimedia content of any embodiment of the present disclosure.

[0009] According to some further embodiments of the present disclosure, a computer-readable storage medium is provided, on which a computer program is stored. When the program is executed by a processor, the method for processing interactive multimedia content in any embodiment of the present disclosure is performed.

[0010] According to some further embodiments of the present disclosure, a computer program product is provided, comprising: instructions, which, when executed by a processor, implement the method for processing interactive multimedia content according to any embodiment of the present disclosure.

[0011] Other features, aspects and advantages of the present disclosure will become apparent from the following detailed description of exemplary embodiments of the present disclosure with reference to the accompanying drawings. BRIEF DESCRIPTION OF THE DRAWINGS

[0012] The preferred embodiments of the present disclosure are described below with reference to the accompanying drawings. The drawings described herein are used to provide a further understanding of the present disclosure. Each of the drawings, together with the following detailed description, is included in this specification and forms a part of the specification to explain the present disclosure. It should be understood that the drawings described below only relate to some embodiments of the present disclosure and do not constitute a limitation of the present disclosure. In the drawings:

[0013] FIG1 is a schematic flow chart showing a method for processing interactive multimedia content according to some embodiments of the present disclosure;

[0014] FIG2 is a schematic diagram showing a story line of some embodiments of the present disclosure;

[0015] FIG3A is a schematic diagram showing an image of a person according to some embodiments of the present disclosure;

[0016] FIG3B is a schematic diagram showing an image of a display scene according to some embodiments of the present disclosure;

[0017] FIG3C shows a schematic diagram of displaying audio according to some embodiments of the present disclosure;

[0018] FIG4 is a schematic diagram showing a preview interface according to some embodiments of the present disclosure;

[0019] FIG5 is a schematic diagram illustrating a character's details page according to some embodiments of the present disclosure;

[0020] FIG6 is a schematic flow chart showing a method for processing interactive multimedia content according to other embodiments of the present disclosure;

[0021] FIG7A is a schematic diagram illustrating a character selection interface according to some embodiments of the present disclosure;

[0022] FIG7B is a schematic diagram showing a selection interface according to some embodiments of the present disclosure;

[0023] FIG7C is a schematic diagram illustrating a display process of interactive multimedia content according to some embodiments of the present disclosure;

[0024] FIG8 is a schematic structural diagram of an apparatus for processing interactive multimedia content according to some embodiments of the present disclosure;

[0025] FIG9 is a schematic structural diagram of an electronic device according to some embodiments of the present disclosure;

[0026] FIG10 is a schematic diagram showing the structure of a computer system according to some embodiments of the present disclosure.

[0027] It should be understood that, for ease of description, the dimensions of the various parts shown in the drawings are not necessarily drawn to scale. The same or similar reference numerals are used throughout the drawings to indicate the same or similar parts. Therefore, once an item is defined in one drawing, it may not be discussed further in subsequent drawings. DETAILED DESCRIPTION

[0028] The following will be combined with the accompanying drawings in the embodiments of the present disclosure to clearly and completely describe the technical solutions in the embodiments of the present disclosure. However, it is obvious that the embodiments described are only some embodiments of the present disclosure, rather than all embodiments. The following description of the embodiments is actually only illustrative and is in no way intended to limit the present disclosure and its application or use. It should be understood that the present disclosure can be implemented in various forms and should not be construed as being limited to the embodiments described herein.

[0029] It should be understood that the various steps described in the method embodiments of the present disclosure can be performed in different orders and / or performed in parallel. In addition, the method embodiments may include additional steps and / or omit the steps shown. The scope of the present disclosure is not limited in this respect. Unless otherwise specifically stated, the relative arrangement, numerical expressions and numerical values ​​of the parts and steps set forth in these embodiments should be interpreted as being merely exemplary and do not limit the scope of the present disclosure.

[0030] As used in this disclosure, the term "include" and its variations are intended to be open-ended terms that include at least the following elements / features but do not exclude other elements / features, i.e., "including but not limited to." Furthermore, the term "comprise" and its variations are intended to be open-ended terms that include at least the following elements / features but do not exclude other elements / features, i.e., "including but not limited to." Therefore, "include" and "include" are synonymous. The term "based on" means "based, at least in part, on."

[0031] Reference throughout this specification to "one embodiment," "some embodiments," or "an embodiment" means that a particular feature, structure, or characteristic described in connection with the embodiment is included in at least one embodiment of the present invention. For example, the term "one embodiment" means "at least one embodiment," the term "another embodiment" means "at least one additional embodiment," and the term "some embodiments" means "at least some embodiments." Furthermore, the appearances of the phrases "in one embodiment," "in some embodiments," or "in an embodiment" in various places throughout this specification are not necessarily all referring to the same embodiment, but may.

[0032] It should be noted that the concepts of "first" and "second" mentioned in this disclosure are only used to distinguish different devices, modules, or units, and are not used to limit the order or interdependence of the functions performed by these devices, modules, or units. Unless otherwise specified, concepts such as "first" and "second" are not intended to imply that the objects described in such a manner must be in a given order in time, space, ranking, or any other manner.

[0033] It should be noted that the modifications of "one" and "multiple" mentioned in the present disclosure are illustrative rather than restrictive, and those skilled in the art should understand that unless otherwise clearly indicated in the context, they should be understood as "one or more".

[0034] The names of the messages or information exchanged between multiple devices in the embodiments of the present disclosure are only used for illustrative purposes and are not used to limit the scope of these messages or information.

[0035] The following detailed description of the embodiments of the present disclosure is provided in conjunction with the accompanying drawings, but the present disclosure is not limited to these specific embodiments. The following specific embodiments may be combined with each other, and the same or similar concepts or processes may not be described in detail in some embodiments. In addition, in one or more embodiments, specific features, structures, or characteristics may be combined in any suitable manner that will be apparent to those skilled in the art from this disclosure.

[0036] Currently, interactive multimedia content features multiple storylines, all created in advance by the author. This content is considered professionally produced content (PGC), and its creation process is challenging and often involves collaboration among professionals from multiple fields. It is also difficult to create user-generated content (UGC), as its production efficiency is relatively low.

[0037] In order to improve the efficiency of generating interactive multimedia content, the present disclosure proposes a method for processing interactive multimedia content, which will be described below with reference to Figures 1 to 7C.

[0038] Figure 1 is a flow chart of some embodiments of the method for processing interactive multimedia content disclosed herein. As shown in Figure 1 , the method of this embodiment includes steps S102 to S106.

[0039] In step S102 , content texts of one or more story branches are generated based on the content text of the original story.

[0040] The original story can be input by the authoring user. For example, an authoring interface for displaying interactive multimedia content is displayed on the client, and an input area is displayed in the authoring interface to receive the original story input by the authoring user. The original story can be input as a whole or in chapters, etc., without limitation.

[0041] The original story can also be published and used to generate interactive multimedia content with the consent of the original user. The original story can include a story line, and based on the story line of the original story, one or more story lines and the content text of each story line can be generated.

[0042] In some embodiments, the original story includes multiple characters, and based on the content text of the original story, one or more story branch content texts are generated for each of the multiple characters. For viewers of the interactive multimedia content, they can select any of the multiple characters to watch (experience) the plot corresponding to that character.

[0043] For example, a generative model is used to generate the content text of one or more story branches based on the content text of the original story. Generative models include, for example, models that generate based on text or models that generate based on images, and the output of the generative model may include text, images, or a combination of the two. Of course, the input or output of the generative model may also be data of other modalities, such as audio, video, or a combination of multiple types of data. The generative model can be a single-modal model, such as a model that generates text based on text (referred to as a "text-to-text model") or a model that generates images based on images (referred to as a "image-to-image model"); or, the generative model can also be a cross-modal model, that is, a model whose input and output belong to different modalities, such as a model that generates images based on text (referred to as a "text-to-image model"); or, the input of the generative model may include multiple modalities, and the output may also include multiple modalities.

[0044] In step S104 , an image is generated based on the content text of the original story and the content text of one or more story branches.

[0045] For example, images include images of people and images of scenes. Since people and scenes are the basic components of each picture in interactive multimedia content, each frame of the interactive multimedia content can be generated based on the images of people and scenes. For example, if the interactive multimedia content is an interactive story (interactive novel), interactive TV series, interactive movie, etc., the generated image can be an image of each frame of the interactive multimedia content.

[0046] For example, a generative model is used to generate images based on the content text of the original story and the content text of one or more story branches.

[0047] In step S106 , interactive multimedia content is generated based on the content text of the original story, the content text of one or more story branches, and images.

[0048] For example, a generative model can be used to generate interactive multimedia content based on the original story text, the content text of one or more story branches, and images. The genre of the interactive multimedia content can be set to generate interactive multimedia content of the corresponding genre.

[0049] The above steps S102 to S106 may be performed using the same generative model. For example, the content text of the original story is input into the generative model, and the interactive multimedia content is output.

[0050] The method of the above embodiment is based on the content text of the original story, and the content text of one or more story branches is expanded. Based on the content text of the original story and the content text of one or more story branches, an image is generated, and then interactive multimedia content is generated. Through the method of the above embodiment, only the content text of the original story can be expanded to form the overall content text of the interactive multimedia content, and further interactive multimedia content can be generated, thereby improving the efficiency of creating interactive multimedia content and saving resources. Ordinary users can also participate in the creation and development of the story to better interact with the characters and create their own unique stories, which improves the feasibility of creating interactive multimedia content and can convert more creative ideas into interactive multimedia products.

[0051] In some embodiments, the method for processing interactive multimedia content also includes: displaying the main story line and one or more story branches of the original story, wherein the main story line includes multiple main line nodes, and each of the one or more story branches includes multiple branch nodes.

[0052] Machine learning models can be used to understand the content of the original story and extract multiple key storyline nodes to form the main storyline. For example, key plot points that influence the development of the original story can be extracted. Based on the content of each key plot point, a summary of each key plot point can be generated to form each key storyline node. Alternatively, each key storyline node can be generated based on a brief description such as the title of each key plot point. The main storyline nodes are then linked according to the development of the original story to form the main storyline.

[0053] For example, a main storyline and one or more subplots can be displayed using a flowchart or other format. Connecting lines can be set between related mainline nodes, between mainline nodes and subplot nodes, and between subplot nodes, connecting them in the order of plot development. Displaying the main storyline and subplots allows creative users to more clearly and accurately determine the relationship between generated subplots and the main storyline, facilitating confirmation or adjustment.

[0054] In some embodiments, one or more story branches include one or more story branches corresponding to each of one or more characters, and displaying the main story line of the original story and the one or more story branches includes at least one of the following: for each character, displaying the main story line of the original story and the one or more story branches corresponding to the character; in response to a selection operation of a first target character among the one or more characters, displaying the main story line of the original story and the one or more story branches corresponding to the first target character.

[0055] One or more story branches corresponding to multiple characters can be displayed in the same interface or in different interfaces, without limitation.

[0056] As shown in Figure 2, for character one, a main storyline and multiple branch storylines are displayed. Each chapter in the main storyline can be used as a mainline node. Some mainline nodes can derive branch story nodes, while some mainline nodes do not. The main storyline corresponds to one ending, and the endings corresponding to the branch storylines are different from those corresponding to the main storyline. The endings corresponding to different branch storylines can be the same or different. The last mainline node of the main storyline displays the corresponding ending information, and the last branch node of each branch storyline displays the corresponding ending information. In response to the authoring user's selection operation for a character other than character one, one or more branch storylines corresponding to the main storyline and other characters can be displayed.

[0057] In some embodiments, for each main line node other than the last main line node among multiple main line nodes, the identification of the chapter corresponding to the main line node is displayed, and when a branch line node is derived from the main line node, option information corresponding to the main line node is displayed; ending information corresponding to the last main line node among multiple main line nodes is displayed; for each branch line node other than the last branch line node among multiple branch line nodes of each story branch line, when the branch node is a branch node derived from other nodes or other branch nodes are derived from the branch node, option information corresponding to the branch node is displayed, wherein the other nodes include main line nodes or branch nodes other than the branch node; ending information corresponding to the last branch node among multiple branch nodes of each story branch line is displayed.

[0058] For example, the chapter identifier is a chapter number, etc., and the option information corresponding to the main line node can be generated by a machine learning model based on the content of the chapter corresponding to the main line node. The option information corresponding to the main line node can be used to describe the options provided to the viewing user for interactive operations before the chapter corresponding to the main line node begins. The main line node can derive a branch node, that is, the content of the branch node is expanded based on the content of the main line node, and the branch node can further derive other branch nodes. The option information corresponding to the main line node and the branch nodes derived from it is the different choices made by the characters in the story based on the same event. The option information corresponding to the branch node and the branch nodes derived from it is the different choices made by the characters in the story based on the same event. The ending information may include: at least one of the name of the ending chapter and a brief description (abstract) of the ending, but is not limited to the examples given.

[0059] As shown in Figure 2, each main line node in the main story line, except for the last main line node, displays the chapter identifier corresponding to the main line node. Of course, the last main line node can also display the corresponding chapter identifier, and is not limited to displaying only the corresponding ending information. The main line node corresponding to the second chapter derives a branch node. The main line node corresponding to the second chapter displays the corresponding option information "Go to Mars with B", and the derived branch node displays the corresponding option information "Do not go to Mars with B". Based on the derived branch node, subsequent branch nodes can be further written (expanded). For example, after the derived branch node, the branch node corresponding to the ending information is further written.

[0060] The information displayed by each node (main line node or branch line node) is not limited to the information shown in Figure 2 and can be configured according to actual needs as long as the authoring user can clearly and accurately understand the relationship and content between each node.

[0061] In some embodiments, using a generative model, the content text corresponding to each branch node can be expanded based on each branch node.

[0062] In response to a selection operation on a target node among various nodes (main line nodes or branch line nodes), content text corresponding to the target node is displayed.

[0063] In some embodiments, in response to a modification operation on the option information of a first target main line node among multiple main line nodes, the modified option information of the first target main line node and the modified option information of the branch line nodes derived from the first target main line node are displayed; and / or, in response to a modification of the content text of a first target branch line node among multiple branch line nodes in one or more story branches, the modified content text of the first target branch node is displayed.

[0064] By using a generative model, the content text corresponding to each branch node can be expanded based on each branch node. The creative user can modify each branch node, the option information of each branch node, and the content text corresponding to each branch node, and can also modify the option information of the main line node. For example, in response to modifying the option information of the first target main line node, it is determined whether the option information of the branch node derived from it needs to be modified. If it needs to be modified, the modified option information of the branch node derived from the first target main line node is generated based on the modified option information of the branch node derived from the first target main line node. Further, it is determined whether the story branch where the branch node derived from the first target main line node is located needs to be modified. If it needs to be modified, the branch nodes after the branch node derived from the first target main line node on the story branch are generated based on the modified option information of the branch node derived from the first target main line node. Furthermore, the content text of each branch node on the story branch is also generated.

[0065] For example, in response to a modification to the content text of a first target branch node, a modified content text of the first target branch node is generated. Based on the modified content text of the first target branch node, the content text of other branch nodes and other branch nodes on the story branch where the first target branch node is located are generated. Each branch node on the story branch where the modified first target branch node is located can be displayed, and the content text of each branch node can also be displayed.

[0066] The authoring user can also perform operations such as deleting and adding branch nodes. In some embodiments, in response to the deletion operation of the second target branch node, the story branch where the second target branch node is located is deleted, or in response to the deletion operation of the target branch, the target branch is deleted. For example, in response to the operation of adding a branch node to the second target main line node, the option information of the second target main line node and the option information of the branch node derived from the second target main line node are generated based on the content of the second target main line node. The content text of the branch node is generated based on the option information of the branch node derived from the second target main line node, and further, the content text of other branch nodes on the story branch is expanded based on the content text of the branch node.

[0067] In some embodiments, generating content texts of one or more story branches based on the content text of the original story includes: determining one or more main line nodes in the main story line of the original story based on the content text of the original story; deriving one or more branch nodes as one or more initial branch nodes based on each main line node in the one or more main line nodes; generating multiple decoded branch nodes in a story line based on each initial branch node in the one or more initial branch nodes, and expanding the content of the multiple branch nodes to generate content texts of one or more story branches.

[0068] For example, a generative model is used to generate one or more story branches and the content text of each story branch based on the content text of the original story.

[0069] In some embodiments, one or more story branches include one or more story branches corresponding to each of one or more characters, and determining one or more main line nodes in the main story line of the original story based on the content text of the original story includes, for each character: based on the content text of the original story, selecting one or more key chapters associated with the character's story from multiple chapters of the main story line and one or more key plots in each key chapter of the one or more key chapters; based on the one or more key plots in each key chapter, generating one or more main line nodes associated with the character in the main story line of the original story.

[0070] For example, a generative model is used to understand the content of the original story, determine the main storyline, key chapters, and key plots, and further generate a summary (brief description information) of the main line nodes based on the content of the key plots.

[0071] In some embodiments, deriving one or more branch nodes based on each main line node in one or more main line nodes includes: determining option information for each main line node based on the key plot corresponding to each main line node, wherein the option information is used to describe the choice of the character associated with the main line node; for each main line node, generating one or more option information different from the option information of the main line node based on the option information of the main line node, and one or more branch nodes corresponding to the one or more option information.

[0072] For example, a generative model is used to understand the key plot corresponding to each main line node to determine whether it is suitable for deriving branch nodes. If it is suitable for deriving branch nodes, option information for one or more branch nodes is generated based on the content of the chapter or key plot corresponding to the main line node.

[0073] In some embodiments, expanding the content of multiple branch nodes includes: for each branch node among the multiple branch nodes, expanding the content of the branch node based on the content text of the main line node before the branch node and at least one node in the branch node, and information about the characters in the story branch to which the branch node belongs, wherein the character information includes at least one of the character's setting information and the relationship information between the characters.

[0074] For example, character setting information includes, but is not limited to, name, gender, age, personality, appearance, and abilities. For example, using a generative model, for each branch node, the content of the branch node is expanded based on the content text of the main node preceding the branch node and at least one of the branch nodes, as well as information about the characters in the branch story to which the branch node belongs.

[0075] In some embodiments, the method for processing interactive multimedia content further includes at least one of the following: associating and displaying the description text of the characters in the content text of the original story with the images of the characters; associating and displaying the description text of the scenes in the content text of the original story with the images of the scenes.

[0076] As shown in Figure 3A, the character description text is displayed in association with the character image. The main characters in the original story and story branches include A, B, C, and D. Multiple characters can be displayed in the same interface, or each character can be displayed in a separate interface. There is no limitation here.

[0077] As shown in Figure 3B, the scene description text is associated with the scene image. The original story and story branches include multiple scenes, which can be displayed in the same interface or each scene can be displayed in a separate interface.

[0078] In some embodiments, there are multiple characters and multiple scenes, and the method for processing interactive multimedia content further includes at least one of the following: in response to an update operation on an image of a second target character among the multiple characters, regenerating and displaying an image of the second target character; for each of the multiple characters, displaying multiple images of the character, and in response to a selection operation on a first target image among the multiple images of the character, determining the first target image as the image of the character; in response to an update operation on an image of a first target scene among the multiple scenes, regenerating and displaying an image of the first target scene; for each of the multiple scenes, displaying multiple images of the scene, and in response to a selection operation on a second target image among the multiple images of the scene, determining the second target image as the image of the scene.

[0079] If the authoring user is dissatisfied with the image of a character or scene, they can regenerate the image of the character or scene. For example, a regeneration control corresponding to each character can be set, and in response to the triggering operation of the regeneration control corresponding to the character, the image of the character can be regenerated. For each scene, a regeneration control corresponding to the scene can be set, and in response to the triggering operation of the regeneration control corresponding to the scene, the image of the scene can be regenerated. For each character or scene, multiple images can be generated, and the authoring user can select them.

[0080] The authoring user can also modify the description information of the character or scene, and in response to the modification of the description information of the character or scene, regenerate the image of the character or scene. Controls can be set to modify the style, specific parts, clothing, etc. of the character. For example, for each character, in response to modifying the character's style to the target style, the image of the target style is regenerated. For example, the Chinese style can be modified to anime style, etc. For example, the character's hairstyle, clothing, appearance, etc. can be modified by triggering the corresponding control, which will not be repeated here. Controls can be set to modify the style of the scene, specific buildings in the scene, etc., and in response to the triggering of the control, the image of the scene is regenerated, which will not be repeated here.

[0081] In some embodiments, generating an image based on the content text of the original story and the content text of one or more story branches includes: extracting descriptive text of each of multiple characters from the content text of the original story and the content text of one or more story branches; generating an image of each character based on the descriptive text of each character; extracting descriptive text of each of multiple scenes from the content text of the original story and the content text of one or more story branches; and generating an image of each scene based on the descriptive text of each scene.

[0082] For example, using a generative model, we can extract the descriptive text of each character from the original story and the content of one or more story branches, and generate an image for each character. For example, using a generative model, we can extract the descriptive text of each scene from the original story and the content of one or more story branches, and generate an image for each scene.

[0083] In some embodiments, the image also includes a map, and the method for processing interactive multimedia content also includes: displaying the map; responding to a selection operation of a target location in the map; associating an image of a scene corresponding to the target location and a description text of the scene corresponding to the target location.

[0084] If the original story includes descriptions of multiple locations and the positional relationships between the locations, a map can be generated, drawn according to the positional relationships between the multiple locations. The map can be displayed in association with the descriptive text of the multiple locations. For example, corresponding selection controls can be set for the multiple locations in the map, and in response to triggering the selection control for a target location, an image of the scene corresponding to the target location and the descriptive text of the scene corresponding to the target location are displayed.

[0085] In some embodiments, description texts of multiple locations are extracted from the content text of the original story and the content text of one or more story branches, wherein the description texts of the multiple locations include the orientation relationship of the multiple locations; and a map is generated based on the description texts of the multiple locations.

[0086] For example, a generative model is used to identify and extract the description text of the locations in the content text of the original story and the content text of one or more story branches, thereby generating a map.

[0087] In some embodiments, the method for processing interactive multimedia content also includes: generating audio based on the content text of the original story and the content text of one or more story branches, wherein the audio includes the voice audio of the characters and the background sound audio of the scene, wherein, generating interactive multimedia content based on the content text of the original story, the content text of one or more story branches and images includes: generating interactive multimedia content based on the content text of the original story, the content text, images and audio of one or more story branches.

[0088] The audio in interactive multimedia content can also be automatically generated using generative models. For example, the background audio of a scene includes the soundtrack, the sounds of characters and objects in the background, etc.

[0089] In some embodiments, generating audio based on the content text of the original story and the content text of one or more story branches includes: extracting setting information of each of multiple characters from the content text of the original story and the content text of one or more story branches; generating sound audio for each character based on the setting information of each character; extracting description text for each of multiple scenes from the content text of the original story and the content text of one or more story branches, and determining the story style; generating background sound audio for each scene based on the description text and story style of each scene.

[0090] The timbre and pitch of each character's audio can reflect their gender, age, personality, and other set information. Therefore, a generative model can be used to extract this set information and generate audio for each character based on it. Audio samples for each character can be generated first, and then, based on the character's dialogue content, audio of each character's dialogue can be generated. The dialogue content influences the character's emotions, which in turn affects characteristics such as timbre, pitch, loudness, and speaking speed. Therefore, audio for each character's dialogue needs to be generated based on the dialogue content.

[0091] Based on the description of the sounds in the scene, sound effects or music materials for different scenes can be generated as background sound audio (for example, the background music in a restaurant should be noisy but cheerful, while the background music when entering a battle should be tense and exciting).

[0092] The background audio for each scene can be generated solely based on the scene's description text. The story's style can also influence the soundtrack. For example, if the story's style is Chinese, a Chinese-style soundtrack can be generated. Different scenes may include not only the soundtrack but also the sounds of characters or objects. For example, the background sounds of people hawking or cars. Generative models can be used to generate audio for these background sounds, including the soundtrack, characters, and objects, based on the scene's description text.

[0093] In some embodiments, there are multiple characters and multiple scenes, and the method for processing interactive multimedia content also includes at least one of the following: associating and displaying the voice audio of multiple characters with the description texts of multiple characters, and playing the first target audio in response to a selection operation of the first target audio in the voice audio of the multiple characters; associating and displaying the background sound audio of multiple scenes with the description texts of the multiple scenes, and playing the second target audio in response to a selection operation of the second target audio in the background sound audio of the multiple scenes.

[0094] For each character, the character and the character's audio can be displayed in association, or the character's settings information can be displayed in association with the character's audio, or the character's description text can be displayed in association with the character's audio, without limitation. For each character, multiple audio recordings of the character can be displayed, and in response to a confirmation operation on a particular audio recording from the multiple audio recordings, the audio recording is determined to be the character's audio recording.

[0095] As shown in FIG3C , the description text of each scene is displayed in association with the background sound audio, and the authoring user can play the audio by clicking or other means.

[0096] In some embodiments, in response to the regeneration or modification operation of the character's voice audio or the background sound audio of the scene, the character's voice audio or the background sound audio of the scene is regenerated.

[0097] If the author is not satisfied with the generated audio, they can regenerate it. Controls for modifying the audio style and other aspects can be set to modify the generated audio. An audio library and audio selection controls can be set. When the audio selection controls are triggered, multiple audio options are displayed. When a particular audio is selected, it is used as the soundtrack for the current scene.

[0098] Creators can also upload their own audio to serve as character voices or background sounds for scenes. For example, if a creator uploads a character's voice, the character's dialogue will be generated based on the character's voice.

[0099] In each of the above embodiments, one or more story branches can be automatically generated based on the original story, generating a story framework with multiple endings. Based on the textual content of the story framework, the complete story content can be further generated. Furthermore, images and audio of characters and scenes can be automatically generated based on the textual content of the original story and story branches. Ultimately, interactive multimedia content is generated based on the textual content, images, and audio of the original story and story branches, thereby improving the efficiency of generating interactive multimedia content.

[0100] In the process of generating interactive multimedia content, images and audio of story lines, characters, and scenes can be displayed. Creative users can adjust and modify this information, incorporate more of their own creativity, and create their own unique stories, which improves the feasibility and flexibility of creating interactive multimedia content.

[0101] After the interactive multimedia content is generated, a preview interface of the interactive multimedia content may be displayed for the authoring user to preview or modify the content. For example, the preview interface of the interactive multimedia content may be displayed in an interactive multimedia content production tool.

[0102] In some embodiments, a preview interface of the interactive multimedia content is displayed, wherein the preview interface includes controls corresponding to each of multiple chapters of the interactive multimedia content; in response to a triggering operation on a control of a target chapter among the multiple chapters, a preview interface corresponding to the target chapter is displayed.

[0103] For example, the preview interface corresponding to the target chapter includes at least one of the picture of the target chapter, the image of the character, the image of the scene, and the audio. While displaying the preview interface, the audio corresponding to the current page in the preview interface can be played.

[0104] As shown in FIG4 , the preview interface of interactive multimedia content includes a first selection area 401, a main display area 402, a time sequence display area 403, and a second selection area 404. The first selection area 401 can be used to display controls for each chapter, the name of each chapter, and a summary (introduction information) of each chapter. In FIG4 , "XXXX" is used to indicate that text can be displayed.

[0105] In response to selecting the target chapter, the interactive multimedia content screen is displayed in the main display area 402, as shown in FIG4 . For example, the interactive multimedia content screen includes images of characters and scenes, text, etc. The interactive multimedia content screen can be the screen presented to the viewing user during playback of the interactive multimedia content. A chapter can include multiple screens, and the screens can be switched for display in response to triggering a switch control or a preset switching operation.

[0106] As shown in Figure 4, the timing display area 403 includes multiple tracks corresponding to images of characters, audio and scenes. According to the time sequence of display, the display bar of the character, the display bar of the audio and the image of the scene are displayed in the corresponding tracks. As shown in Figure 4, the currently displayed picture is the first scene picture, the character C and the first audio (the audio of the character's voice or the audio of the background sound). In response to the first selection operation (which can be used as a preset switching operation) of the display bar of the character, the display bar of the audio or the image of the scene, the switched picture is displayed. For example, the creative user can click on the images of different scenes to display the picture corresponding to the image of the clicked scene. For another example, the creative user can click on the display bar of the character or different positions of the display bar of the character to display the corresponding picture.

[0107] For example, in response to a second selection operation on a character display bar, an audio display bar, or a scene image, a character modification interface, an audio modification interface, or a scene modification interface is displayed. The character image, description information, etc., can be modified (or regenerated), the audio can be modified (or regenerated), or the scene image can be modified (or regenerated), which will not be described in detail here.

[0108] As shown in FIG4 , the main display area 402 may also include multiple option controls, such as a canvas control, a storyline control, and a worldview control. In response to triggering the canvas control, a screen of interactive multimedia content is displayed. In response to triggering the storyline control, storylines (including main and side stories) corresponding to one or more characters are displayed. For example, the main and side stories corresponding to the different characters shown in FIG2 may be displayed. In response to triggering the worldview control, a map is displayed. The specific method for displaying a map can be referred to in the aforementioned embodiment and will not be described in detail here.

[0109] In some embodiments, the method for processing interactive multimedia content further includes: displaying controls for one or more characters, wherein the one or more characters include one or more characters appearing in the target chapter, or the one or more characters include one or more characters in the interactive multimedia content; and in response to a triggering operation on a control for a third target character among the one or more characters, displaying a details interface for the third target character, wherein the details interface includes at least one of the following: setting information, image, character relationship information, audio and sound, chapter information, and ability information of the third target character. The setting information, image, character relationship information, audio and sound, chapter information, and ability information of the third target character can be displayed on the same page or on different pages, without limitation.

[0110] As shown in FIG4 , the second selection area 404 includes multiple option controls, such as a character option control, a scene option control, a sound effect option control, a text option control, and an interaction option control. In response to triggering a character option control, controls for one or more characters are displayed. As shown in FIG4 , controls for characters in this chapter can be displayed separately from controls for other characters. For each character, an avatar, name, and brief description of the character can be displayed. In response to triggering a control for a third target character, a detailed interface for the third target character is displayed.

[0111] As shown in Figure 5, for any character, the character details interface includes an image of the character and may also include multiple option controls, such as a basic information control, a timbre control, a character relationship control, an appearance chapter control, and an ability control. In response to triggering the basic information control, the character's set information is displayed, such as name, gender, age, personality, appearance, and background, but not limited to the examples listed. In response to triggering the timbre control, the character's voice audio is displayed and can be played. In response to triggering the character relationship control, the relationship between the character and other characters is displayed, for example, in the form of a network diagram, topology diagram, text, etc., but not limited to the examples listed. In response to triggering the appearance chapter control, the character's appearance chapter is displayed, for example, the chapter name, summary, and other information may also be displayed. In response to triggering the ability control, the character's ability information is displayed, such as information about the weapon used and skill information, but not limited to the examples listed.

[0112] In some embodiments, the method for processing interactive multimedia content also includes: displaying controls for one or more scenes, wherein the one or more scenes include one or more scenes in the target chapter or the one or more scenes include one or more scenes in the interactive multimedia content; and displaying an image of the second target scene in response to a triggering operation on a control for a second target scene in the one or more scenes.

[0113] As shown in FIG4 , a scene selection control is displayed in the second selection area 404. In response to triggering the scene selection control, controls for one or more scenes are displayed. For example, the scene controls can be displayed as thumbnails of the scenes, and can also display the scene's name, a brief description, etc. Controls for the scenes in the current chapter can be displayed separately from controls for other scenes. In response to triggering the control for the second target scene, an image of the second target scene or a screen containing interactive multimedia content of the second target scene is displayed in the main display area.

[0114] In some embodiments, the method for processing interactive multimedia content also includes: displaying controls for one or more audios, wherein the one or more audios include one or more audios in a target chapter or one or more scenes include one or more audios in the interactive multimedia content; and playing the third target audio in response to a triggering operation on a control for a third target audio in the one or more audios.

[0115] As shown in FIG4 , the second selection area 404 displays an audio selection control. In response to triggering the audio selection control, controls for one or more audio items are displayed. Keywords, a brief description of the audio items, etc. can also be displayed. Controls for the current chapter's audio items can be displayed separately from controls for other audio items. In response to triggering the control for a third target audio item, the third target audio item is played, and the main display area 404 displays the interactive multimedia content corresponding to the third target audio item.

[0116] In some embodiments, the preview interface corresponding to the target chapter includes images of characters, images of scenes, and text in the target chapter, and the method for processing interactive multimedia content further includes: displaying the edited text in response to an editing operation on the text.

[0117] As shown in FIG4 , the main display area 402 displays text, such as character lines, scene descriptions, etc., and in response to an edit operation on the text, the edited text is displayed. The authoring user can edit and modify the content text, such as character lines, scene descriptions, etc.

[0118] As shown in FIG4 , a text selection control is displayed in the second selection area 404. In response to triggering the text selection control, controls for one or more texts are displayed. Controls for the text in the current chapter can be displayed separately from controls for other texts. In response to triggering a control for a target text, the interactive multimedia content corresponding to the target text can be displayed in the main display area 404.

[0119] As shown in FIG4 , the second selection area 404 displays an interactive option control. In response to triggering the interactive option control, one or more interactive controls are displayed. The controls for the current chapter's interactions can be displayed separately from the controls for other interactions. In response to triggering the control for the target interaction, the main display area 404 displays the interactive multimedia content corresponding to the target interaction. The screen may display multiple options that require the user to select, or actions that require the user to perform.

[0120] An animation selection control can also be displayed in the second selection area 404. In response to triggering the animation selection control, controls for one or more animations are displayed. Controls for the animation for this chapter can be displayed separately from controls for other animations. In response to triggering a control for a target animation, the target animation can be displayed in the main display area 404. The interactive multimedia content can include some animations, which can be previewed.

[0121] In the second selection area, there is no restriction on the display form of character controls, scene controls, audio controls, text controls, interactive controls, etc., as long as different controls can be distinguished and serve as identification.

[0122] The method of the above embodiment can generate interactive multimedia content or a sample of interactive multimedia content after the content text, image, and audio of the interactive multimedia content, and can display a preview interface of the interactive multimedia content to facilitate the creative user to adjust various details such as the story structure, story line, text, image, audio, etc. to generate accurate and appropriate interactive multimedia content.

[0123] Some embodiments of the use process of generated interactive multimedia content will be described below with reference to FIG. 6 .

[0124] Fig. 6 is a flow chart of another embodiment of the method for processing interactive multimedia content disclosed herein. As shown in Fig. 6 , the method of this embodiment includes steps S602 to S608.

[0125] In step S602 , a character selection interface for interactive multimedia content is displayed.

[0126] For example, in response to a start-up operation of the interactive multimedia content, a character selection interface for the interactive multimedia content is displayed. The character selection interface may include multiple character images, and may also include at least one of a name and a description, without limitation to the examples given. As shown in FIG7A , the character selection interface includes multiple character images and may include a selection control.

[0127] In step S604 , in response to a selection operation on a character in the interactive multimedia content, the interactive multimedia content is displayed according to the selected character.

[0128] For example, in response to triggering a selection control, multiple character options are displayed, and in response to a triggering operation on a target character option (i.e., a character selection operation), the selected character is determined. Because different characters can correspond to different story lines, content corresponding to different story lines in the interactive multimedia content can be displayed for different characters. When a user viewing the interactive multimedia content selects a character (role), a story from that character's perspective is generated, allowing the user to experience a story completely different from that of other characters.

[0129] As shown in FIG7B , the interactive multimedia content is presented in the form of scenes, character dialogues, narrations, or character conversations, making text reading more interesting.

[0130] In step S606 , in response to the plot development of the interactive multimedia content reaching a plot point where one or more story branches are generated, a selection interface is displayed.

[0131] For example, the selection interface includes multiple option information, namely, option information of the main line node on the main story line and option information of the derived branch line nodes.

[0132] In step S608, in response to a triggering operation on target option information among the plurality of option information, a screen of a main story line or a branch story line corresponding to the target option information is displayed.

[0133] If the target option information is the option information corresponding to the main line node, the picture of the main line of the story corresponding to the target option information is displayed; if the target option information is the option information corresponding to the branch line node, the picture of the branch line of the story corresponding to the target option information is displayed.

[0134] Steps S606-S608 can be replaced with the following steps: displaying a selection interface in response to the interactive multimedia content reaching a plot point where one or more story branches are generated; receiving input selection information; identifying the selection information and determining the story branch corresponding to the selection information; and displaying the screen corresponding to the story branch. The viewing user can select the next plot by inputting the selection information. A machine learning model can be used to identify and understand the selection information and match the story branch associated with the selection information.

[0135] As shown in FIG7C , the selection interface may include options that require the viewing user to select, such as “use props”, “defense”, “escape”, etc. If the viewing user selects different options, plots on different story branches may be triggered.

[0136] Based on the method of the above embodiment, the interactive multimedia content is displayed with richer screen content and stronger interactivity, thereby improving the user's reading experience.

[0137] The present disclosure also provides a device for processing interactive multimedia content, which will be described below in conjunction with FIG. 8 .

[0138] FIG8 is a structural diagram of some embodiments of the interactive multimedia content processing device disclosed herein. As shown in FIG8 , the processing device 80 of this embodiment includes: a text generation module 810 , an image generation module 820 , and a multimedia generation module 830 .

[0139] The text generation module 810 is configured to generate content texts of one or more story branches based on the content text of the original story.

[0140] The image generation module 820 is configured to generate images based on the content text of the original story and the content text of one or more story branches, wherein the images include images of characters and images of scenes.

[0141] The multimedia generation module 830 is configured to generate interactive multimedia content based on the content text of the original story, the content text of one or more story branches, and images.

[0142] In some embodiments, the processing device 80 further includes: a display module 840 configured to display a main story line and one or more branch story lines of the original story, wherein the main story line includes multiple main line nodes, and each of the one or more branch story lines includes multiple branch line nodes.

[0143] In some embodiments, one or more story branches include one or more story branches corresponding to each of one or more characters, and the display module 840 is configured to perform at least one of the following: for each character, display the main story line of the original story and the one or more story branches corresponding to the character; in response to a selection operation of a first target character among the one or more characters, display the main story line of the original story and the one or more story branches corresponding to the first target character.

[0144] In some embodiments, the display module 840 is configured to perform at least one of the following: for each main line node other than the last main line node among multiple main line nodes, display the identification of the chapter corresponding to the main line node, and when a branch line node is derived from the main line node, display the option information corresponding to the main line node; display the ending information corresponding to the last main line node among multiple main line nodes; for each branch line node other than the last branch line node among multiple branch line nodes of each story branch line, display the option information corresponding to the branch line node when the branch node is a branch node derived from other nodes or other branch nodes are derived from the branch node, wherein the other nodes include main line nodes or branch nodes other than the branch node; display the ending information corresponding to the last branch node among multiple branch nodes of each story branch line.

[0145] In some embodiments, the display module 840 is further configured to perform at least one of the following: in response to a modification operation on the option information of a first target main line node among multiple main line nodes, display the modified option information of the first target main line node and the modified option information of the branch nodes derived from the first target main line node; in response to a modification of the content text of the first target branch node among multiple branch nodes in one or more story branches, display the modified content text of the first target branch node.

[0146] In some embodiments, the text generation module 810 is configured to determine one or more main line nodes in the main story line of the original story based on the content text of the original story; derive one or more branch line nodes based on each main line node in the one or more main line nodes as one or more initial branch line nodes; generate multiple branch line nodes in a story line based on each initial branch node in the one or more initial branch line nodes, and expand the content of the multiple branch nodes to generate content text of one or more story branches.

[0147] In some embodiments, one or more story branches include one or more story branches corresponding to each of one or more characters, and the text generation module 810 is configured to: for each character, based on the content text of the original story, select one or more key chapters associated with the character's story and one or more key plots in each of the one or more key chapters from multiple chapters of the main story line; based on the one or more key plots in each key chapter, generate one or more main line nodes associated with the character in the main story line of the original story.

[0148] In some embodiments, the text generation module 810 is configured to determine the option information of each main line node based on the key plot corresponding to each main line node, wherein the option information is used to describe the choice of the character associated with the main line node; for each main line node, based on the option information of the main line node, generate one or more option information different from the option information of the main line node, and one or more branch nodes corresponding to the one or more option information.

[0149] In some embodiments, the text generation module 810 is configured to expand the content of each branch node among multiple branch nodes based on the content text of the main line node before the branch node and at least one node in the branch node, and the information of the characters in the story branch to which the branch node belongs, wherein the character information includes at least one of the character's setting information and the relationship information between the characters.

[0150] In some embodiments, the display module 840 is further configured to perform at least one of the following: displaying the description text of the characters in the content text of the original story in association with the images of the characters; and displaying the description text of the scenes in the content text of the original story in association with the images of the scenes.

[0151] In some embodiments, there are multiple characters and multiple scenes, and the image generation module 820 is further configured to perform at least one of the following: in response to an update operation on the image of a second target character among the multiple characters, regenerate and display the image of the second target character; for each of the multiple characters, display multiple images of the character, and in response to a selection operation on a first target image among the multiple images of the character, determine the first target image as the image of the character; in response to an update operation on the image of a first target scene among the multiple scenes, regenerate and display the image of the first target scene; for each of the multiple scenes, display multiple images of the scene, and in response to a selection operation on a second target image among the multiple images of the scene, determine the second target image as the image of the scene.

[0152] In some embodiments, the image generation module 820 is configured to extract descriptive text for each of a plurality of characters from the content text of the original story and the content text of one or more story branches; generate an image of each character based on the descriptive text of each character; extract descriptive text for each of a plurality of scenes from the content text of the original story and the content text of one or more story branches; and generate an image of each scene based on the descriptive text of each scene.

[0153] In some embodiments, the image also includes a map, and the display module 840 is further configured to display the map; in response to a selection operation of a target location in the map, an image of a scene corresponding to the target location and a description text of the scene corresponding to the target location are associated and displayed.

[0154] In some embodiments, the image also includes a map, and the image generation module 820 is further configured to extract descriptive texts of multiple locations from the content text of the original story and the content text of one or more story branches, wherein the descriptive texts of the multiple locations include the directional relationship between the multiple locations; and generate a map based on the descriptive texts of the multiple locations.

[0155] In some embodiments, the processing device 80 also includes: an audio generation module 850, which is configured to generate audio based on the content text of the original story and the content text of one or more story branches, wherein the audio includes the voice audio of the characters and the background sound audio of the scene; and a multimedia generation module 830 is configured to generate interactive multimedia content based on the content text of the original story, the content text of one or more story branches, images and audio.

[0156] In some embodiments, there are multiple characters and multiple scenes, and the display module 840 is configured to perform at least one of the following: associating and displaying the voice audio of multiple characters with the description texts of multiple characters, and playing the first target audio in response to a selection operation of the first target audio in the voice audio of multiple characters; associating and displaying the background sound audio of multiple scenes with the description texts of multiple scenes, and playing the second target audio in response to a selection operation of the second target audio in the background sound audio of multiple scenes.

[0157] In some embodiments, the audio generation module 850 is configured to extract setting information of each of multiple characters from the content text of the original story and the content text of one or more story branches; generate sound audio for each character based on the setting information of each character; extract description text of each of multiple scenes from the content text of the original story and the content text of one or more story branches, and determine the story style; generate background sound audio for each scene based on the description text and story style of each scene.

[0158] In some embodiments, the display module 840 is configured to display a preview interface of the interactive multimedia content, wherein the preview interface includes controls corresponding to each of multiple chapters of the interactive multimedia content; in response to a triggering operation on a control of a target chapter among the multiple chapters, a preview interface corresponding to the target chapter is displayed.

[0159] In some embodiments, the display module 840 is configured to display controls for one or more characters, wherein the one or more characters include one or more characters appearing in the target chapter, or the one or more characters include one or more characters in the interactive multimedia content; in response to a triggering operation on the control of a third target character among the one or more characters, a details interface of the third target character is displayed, wherein the details interface includes: at least one of the setting information, image, character relationship information, sound audio, appearance chapter information, and ability information of the third target character.

[0160] In some embodiments, the display module 840 is configured to display controls for one or more scenes, wherein the one or more scenes include one or more scenes in a target chapter or the one or more scenes include one or more scenes in interactive multimedia content; and in response to a triggering operation on a control for a second target scene in the one or more scenes, an image of the second target scene is displayed.

[0161] In some embodiments, the display module 840 is configured to display controls for one or more audios, wherein the one or more audios include one or more audios in a target chapter or one or more scenes include one or more audios in interactive multimedia content; and in response to a triggering operation on a control for a third target audio in the one or more audios, the third target audio is played.

[0162] In some embodiments, the preview interface corresponding to the target chapter includes images of characters, images of scenes, and text in the target chapter, and the display module 840 is configured to display the edited text in response to an editing operation on the text.

[0163] In some embodiments, the processing device 80 further includes: the display module 840 is further configured to display a character selection interface for the interactive multimedia content; in response to a character selection operation in the interactive multimedia content, the interactive multimedia content is displayed according to the selected character.

[0164] In some embodiments, the display module 840 is also configured to display a selection interface in response to the plot development of the interactive multimedia content reaching a plot node that generates one or more story branches, wherein the selection interface includes multiple option information; in response to a triggering operation on target option information among the multiple option information, display a screen of the main story line or story branch corresponding to the target option information.

[0165] In some embodiments, the display module 840 is also configured to display a selection interface in response to the plot development of the interactive multimedia content reaching a plot point that generates one or more story branches; the processing device 80 also includes: an input module 860, configured to receive input selection information, identify the selection information, and determine the story branch corresponding to the selection information; the display module 840 is also configured to display a screen of the story branch corresponding to the selection information.

[0166] The above functions of displaying the character selection interface, the selection interface, etc. on the viewing user side may also be implemented by other devices or clients.

[0167] It should be noted that the above-mentioned units (modules) are merely logical modules divided according to the specific functions they implement, and are not intended to limit specific implementation methods. For example, they can be implemented in software, hardware, or a combination of software and hardware. In actual implementation, the above-mentioned units can be implemented as independent physical entities, or can also be implemented by a single entity (for example, a processor (CPU or DSP, etc.), an integrated circuit, etc.). In addition, the above-mentioned units are shown with dotted lines in the drawings to indicate that these units may not actually exist, and the operations / functions they implement can be implemented by the processing circuit itself.

[0168] In addition, although not shown, the device may also include a memory that can store various information generated by the device and the various units contained in the device during operation, programs and data used for operation, data to be sent by the communication unit, etc. The memory can be volatile memory and / or non-volatile memory. For example, the memory can include but is not limited to random access memory (RAM), dynamic random access memory (DRAM), static random access memory (SRAM), read-only memory (ROM), and flash memory. Of course, the memory can also be located outside the device. Optionally, although not shown, the device may also include a communication unit that can be used to communicate with other devices. In one example, the communication unit can be implemented in an appropriate manner known in the art, for example, including communication components such as an antenna array and / or a radio frequency link, various types of interfaces, communication units, etc. This will not be described in detail here. In addition, the device may also include other components not shown, such as a radio frequency link, a baseband processing unit, a network interface, a processor, a controller, etc. This will not be described in detail here.

[0169] Some embodiments of the present disclosure also provide an electronic device. Figure 9 shows a block diagram of some embodiments of the electronic device of the present disclosure. For example, in some embodiments, the electronic device 90 can be various types of devices, for example, including but not limited to mobile terminals such as mobile phones, laptops, digital broadcast receivers, PDAs (personal digital assistants), PADs (tablet computers), PMPs (portable multimedia players), vehicle-mounted terminals (such as vehicle-mounted navigation terminals), etc., and fixed terminals such as digital TVs, desktop computers, etc. For example, the electronic device 90 may include a display panel for displaying data and / or execution results utilized in the scheme of the present disclosure. For example, the display panel can be of various shapes, such as a rectangular panel, an elliptical panel, or a polygonal panel. In addition, the display panel can be not only a flat panel, but also a curved panel or even a spherical panel.

[0170] As shown in FIG9 , the electronic device 90 of this embodiment includes a memory 91 and a processor 92 coupled to the memory 91. It should be noted that the components of the electronic device 90 shown in FIG9 are merely exemplary and non-limiting. The electronic device 90 may also include other components as required by actual applications. The processor 92 may control the other components in the electronic device 90 to perform desired functions.

[0171] In some embodiments, the memory 91 is configured to store one or more computer-readable instructions. When the processor 92 is configured to execute the computer-readable instructions, the computer-readable instructions, when executed by the processor 92, implement a method according to any of the above-described embodiments. The specific implementation and related explanations of each step of the method can be found in the above-described embodiments, and any repetitive details are omitted here.

[0172] For example, the processor 92 and the memory 91 may communicate with each other directly or indirectly. For example, the processor 92 and the memory 91 may communicate with each other via a network. The network may include a wireless network, a wired network, and / or any combination of wireless networks and wired networks. The processor 92 and the memory 91 may also communicate with each other via a system bus, which is not limited in this disclosure.

[0173] For example, the processor 92 can be embodied as various appropriate processors, processing devices, etc., such as a central processing unit (CPU), a graphics processing unit (GPU), a network processor (NP), etc.; it can also be a digital signal processor (DSP), an application-specific integrated circuit (ASIC), a field programmable gate array (FPGA) or other programmable logic devices, discrete gate or transistor logic devices, discrete hardware components. The central processing unit (CPU) can be an X86 or ARM architecture, etc. For example, the memory 91 can include any combination of various forms of computer-readable storage media, such as volatile memory and / or non-volatile memory. The memory 91 can include, for example, a system memory, which stores, for example, an operating system, an application, a boot loader (Boot Loader), a database, and other programs. Various applications and various data can also be stored in the storage medium.

[0174] In addition, according to some embodiments of the present disclosure, when various operations / processes according to the present disclosure are implemented through software and / or firmware, the programs constituting the software can be installed from a storage medium or a network to a computer system having a dedicated hardware structure, such as the computer system (or electronic device) 100 shown in Figure 10. When the various programs are installed, the computer system can perform various functions, including functions such as those described above. Figure 10 is a block diagram showing an example structure of a computer system that can be used in embodiments of the present disclosure.

[0175] In Figure 10, a central processing unit (CPU) 1001 performs various processes according to a program stored in a read-only memory (ROM) 1002 or a program loaded from a storage part 1008 to a random access memory (RAM) 1003. In the RAM 1003, data required when the CPU 1001 performs various processes, etc., is also stored as needed. The central processing unit is merely exemplary and may also be other types of processors, such as the various processors described above. The ROM 1002, RAM 1003, and storage part 1008 may be various forms of computer-readable storage media, as described below. It should be noted that although ROM 1002, RAM 1003, and storage device 1008 are shown separately in Figure 10, one or more of them may be combined or located in the same or different memory or storage modules.

[0176] The CPU 1001, the ROM 1002, and the RAM 1003 are connected to one another via a bus 1004. An input / output interface 1005 is also connected to the bus 1004.

[0177] The following components are connected to the input / output interface 1005: an input portion 1006, such as a touch screen, touchpad, keyboard, mouse, image sensor, microphone, accelerometer, gyroscope, etc.; an output portion 1007, including a display, such as a cathode ray tube (CRT), liquid crystal display (LCD), speaker, vibrator, etc.; a storage portion 1008, including a hard disk, magnetic tape, etc.; and a communication portion 1009, including a network interface card, such as a LAN card, modem, etc. The communication portion 1009 allows communication processing to be performed via a network, such as the Internet. It will be readily understood that although FIG10 shows that the various devices or modules in the computer system 100 communicate via the bus 1004, they may also communicate via a network or other means, where the network may include a wireless network, a wired network, and / or any combination of wireless and wired networks.

[0178] A drive 1010 is also connected to the input / output interface 1005 as needed. A removable medium 1011 such as a magnetic disk, an optical disk, a magneto-optical disk, a semiconductor memory, etc. is mounted on the drive 1010 as needed so that a computer program read therefrom is installed in the storage section 1008 as needed.

[0179] In the case of realizing the above-described series of processing by software, a program constituting the software can be installed from a network such as the Internet or a storage medium such as the removable medium 1011 .

[0180] According to an embodiment of the present disclosure, the process described above with reference to the flowchart can be implemented as a computer software program. For example, an embodiment of the present disclosure includes a computer program product, which includes a computer program carried on a computer-readable medium, and the computer program includes program code for executing the method shown in the flowchart. In such an embodiment, the computer program can be downloaded and installed from the network through the communication device 1009, or installed from the storage device 1008, or installed from the ROM 1002. When the computer program is executed by the CPU 1001, the above-mentioned functions defined in the method of the embodiment of the present disclosure are performed.

[0181] It should be noted that in the context of the present disclosure, a computer-readable medium may be a tangible medium that may contain or store a program for use by or in conjunction with an instruction execution system, apparatus, or device. A computer-readable medium may be a computer-readable signal medium or a computer-readable storage medium or any combination thereof. A computer-readable storage medium may be, for example, but not limited to, an electrical, magnetic, optical, electromagnetic, infrared, or semiconductor system, apparatus, or device, or any combination thereof. More specific examples of computer-readable storage media may include, but are not limited to, an electrical connection with one or more wires, a portable computer disk, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or flash memory), an optical fiber, a portable compact disk read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination thereof. In the present disclosure, a computer-readable storage medium may be any tangible medium that contains or stores a program that may be used by or in conjunction with an instruction execution system, apparatus, or device. In the present disclosure, a computer-readable signal medium may include a data signal propagated in baseband or as part of a carrier wave, which carries a computer-readable program code. Such propagated data signals may take a variety of forms, including but not limited to electromagnetic signals, optical signals, or any suitable combination thereof. A computer-readable signal medium may also be any computer-readable medium other than a computer-readable storage medium, which may send, propagate, or transmit a program for use by or in conjunction with an instruction execution system, apparatus, or device. The program code contained on the computer-readable medium may be transmitted using any suitable medium, including but not limited to: wires, optical cables, RF (radio frequency), etc., or any suitable combination thereof.

[0182] The computer-readable medium may be included in the electronic device, or may exist independently without being incorporated into the electronic device.

[0183] In some embodiments, a computer program is further provided, comprising: instructions, which, when executed by a processor, cause the processor to perform any of the methods of the above embodiments. For example, the instructions may be embodied as computer program codes.

[0184] In embodiments of the present disclosure, computer program code for performing the operations of the present disclosure can be written in one or more programming languages ​​or combinations thereof, including but not limited to object-oriented programming languages, such as Java, Smalltalk, C++, and conventional procedural programming languages, such as "C" language or similar programming languages. The program code can be executed entirely on the user's computer, partially on the user's computer, as an independent software package, partially on the user's computer and partially on a remote computer, or entirely on a remote computer or server. In situations involving a remote computer, the remote computer can be connected to the user's computer via any type of network (including a local area network (LAN) or a wide area network (WAN)), or can be connected to an external computer (e.g., using an Internet service provider to connect via the Internet).

[0185] The flowcharts and block diagrams in the accompanying drawings illustrate the possible implementation architecture, functions and operations of the systems, methods and computer program products according to various embodiments of the present disclosure. In this regard, each box in the flowchart or block diagram can represent a module, program segment, or a part of code, and the module, program segment, or a part of code contains one or more executable instructions for realizing the specified logical function. It should also be noted that in some alternative implementations, the functions marked in the box can also occur in a different order than that marked in the accompanying drawings. For example, two boxes represented in succession can actually be executed substantially in parallel, and they can sometimes be executed in the opposite order, depending on the functions involved. It should also be noted that each box in the block diagram and / or flowchart, and the combination of the boxes in the block diagram and / or flowchart, can be implemented with a dedicated hardware-based system that performs the specified function or operation, or can be implemented with a combination of dedicated hardware and computer instructions.

[0186] The modules, components, or units described in the embodiments of the present disclosure may be implemented in software or hardware. The names of the modules, components, or units do not necessarily limit the modules, components, or units themselves.

[0187] The functions described above herein may be performed at least in part by one or more hardware logic components. For example, and without limitation, exemplary hardware logic components that may be used include: field programmable gate arrays (FPGAs), application specific integrated circuits (ASICs), application specific standard products (ASSPs), systems on chip (SOCs), complex programmable logic devices (CPLDs), and the like.

[0188] According to some embodiments of the present disclosure, a method for processing interactive multimedia content is provided, including: generating content texts of one or more story branches based on the content text of the original story; generating images based on the content text of the original story and the content texts of one or more story branches, wherein the images include images of characters and images of scenes; generating interactive multimedia content based on the content text of the original story, the content texts of one or more story branches, and the images.

[0189] In some embodiments, the method further includes: displaying a main story line and one or more branch story lines of the original story, wherein the main story line includes a plurality of main story line nodes, and each of the one or more branch story lines includes a plurality of branch story nodes.

[0190] In some embodiments, one or more story branches include one or more story branches corresponding to each of one or more characters, and displaying the main story line of the original story and the one or more story branches includes at least one of the following: for each character, displaying the main story line of the original story and the one or more story branches corresponding to the character; in response to a selection operation of a first target character among the one or more characters, displaying the main story line of the original story and the one or more story branches corresponding to the first target character.

[0191] In some embodiments, displaying the main story line and one or more story branches of the original story includes at least one of the following: for each main line node other than the last main line node among multiple main line nodes, displaying the identification of the chapter corresponding to the main line node, and when a branch line node is derived from the main line node, displaying option information corresponding to the main line node; displaying ending information corresponding to the last main line node among multiple main line nodes; for each branch node other than the last branch node among multiple branch nodes of each story branch line, when the branch node is a branch node derived from other nodes or other branch nodes are derived from the branch node, displaying option information corresponding to the branch node, wherein the other nodes include main line nodes or branch nodes other than the branch node; displaying ending information corresponding to the last branch node among multiple branch nodes of each story branch line.

[0192] In some embodiments, the method further includes at least one of the following: in response to a modification operation on the option information of a first target main line node among multiple main line nodes, displaying the modified option information of the first target main line node and the modified option information of the branch nodes derived from the first target main line node; in response to a modification of the content text of a first target branch node among multiple branch nodes in one or more story branches, displaying the modified content text of the first target branch node.

[0193] In some embodiments, generating content texts of one or more story branches based on the content text of the original story includes: determining one or more main line nodes in the main story line of the original story based on the content text of the original story; deriving one or more branch nodes from each of the one or more main line nodes as one or more initial branch nodes; generating multiple branch nodes in a story line based on each of the one or more initial branch nodes, and expanding the content of the multiple branch nodes to generate content texts of one or more story branches.

[0194] In some embodiments, one or more story branches include one or more story branches corresponding to each of one or more characters, and determining one or more main line nodes in the main story line of the original story based on the content text of the original story includes, for each character: based on the content text of the original story, selecting one or more key plots in each key chapter from multiple chapters of the main story line; based on the one or more key plots in each key chapter, generating one or more main line nodes associated with the character in the main story line of the original story.

[0195] In some embodiments, deriving one or more branch nodes based on each main line node in one or more main line nodes includes: determining option information for each main line node based on the key plot corresponding to each main line node, wherein the option information is used to describe the choice of the character associated with the main line node; for each main line node, generating one or more option information different from the option information of the main line node based on the option information of the main line node, and one or more branch nodes corresponding to the one or more option information.

[0196] In some embodiments, expanding the content of multiple branch nodes includes: for each branch node among the multiple branch nodes, expanding the content of the branch node based on the content text of the main line node before the branch node and at least one node in the branch node, and information about the characters in the story branch to which the branch node belongs, wherein the character information includes at least one of the character's setting information and the relationship information between the characters.

[0197] In some embodiments, the method further includes at least one of the following: displaying the description text of a character in the content text of the original story in association with the image of the character; and displaying the description text of a scene in the content text of the original story in association with the image of the scene.

[0198] In some embodiments, there are multiple characters and multiple scenes, and the processing method further includes at least one of the following: in response to an update operation on an image of a second target character among the multiple characters, regenerating and displaying an image of the second target character; for each of the multiple characters, displaying multiple images of the character, and in response to a selection operation on a first target image among the multiple images of the character, determining the first target image as the image of the character; in response to an update operation on an image of a first target scene among the multiple scenes, regenerating and displaying an image of the first target scene; for each of the multiple scenes, displaying multiple images of the scene, and in response to a selection operation on a second target image among the multiple images of the scene, determining the second target image as the image of the scene.

[0199] In some embodiments, generating an image based on the content text of the original story and the content text of one or more story branches includes: extracting descriptive text of each of multiple characters from the content text of the original story and the content text of one or more story branches; generating an image of each character based on the descriptive text of each character; extracting descriptive text of each of multiple scenes from the content text of the original story and the content text of one or more story branches; and generating an image of each scene based on the descriptive text of each scene.

[0200] In some embodiments, the image also includes a map, and the processing method further includes: displaying the map; in response to a selection operation of a target location in the map, associating and displaying an image of a scene corresponding to the target location and a description text of the scene corresponding to the target location.

[0201] In some embodiments, the image also includes a map, and generating the image based on the content text of the original story and the content text of one or more story branches includes: extracting description texts of multiple locations from the content text of the original story and the content text of one or more story branches, wherein the description texts of the multiple locations include the orientation relationship of the multiple locations; and generating a map based on the description texts of the multiple locations.

[0202] In some embodiments, the method further includes: generating audio based on the content text of the original story and the content text of one or more story branches, wherein the audio includes the voice audio of the characters and the background sound audio of the scene, wherein, generating interactive multimedia content based on the content text of the original story, the content text of one or more story branches and images includes: generating interactive multimedia content based on the content text of the original story, the content text, images and audio of one or more story branches.

[0203] In some embodiments, there are multiple characters and multiple scenes, and the processing method further includes at least one of the following: associating and displaying the voice audio of multiple characters with the description texts of multiple characters, and playing the first target audio in response to a selection operation of the first target audio in the voice audio of the multiple characters; associating and displaying the background sound audio of multiple scenes with the description texts of the multiple scenes, and playing the second target audio in response to a selection operation of the second target audio in the background sound audio of the multiple scenes.

[0204] In some embodiments, generating audio based on the content text of the original story and the content text of one or more story branches includes: extracting setting information of each of multiple characters from the content text of the original story and the content text of one or more story branches; generating sound audio for each character based on the setting information of each character; extracting description text for each of multiple scenes from the content text of the original story and the content text of one or more story branches, and determining the story style; generating background sound audio for each scene based on the description text and story style of each scene.

[0205] In some embodiments, the method further includes: displaying a preview interface of the interactive multimedia content, wherein the preview interface includes controls corresponding to each of multiple chapters of the interactive multimedia content; and in response to a triggering operation on a control of a target chapter among the multiple chapters, displaying a preview interface corresponding to the target chapter.

[0206] In some embodiments, the method further includes: displaying controls for one or more characters, wherein the one or more characters include one or more characters appearing in the target chapter, or the one or more characters include one or more characters in the interactive multimedia content; in response to a triggering operation on the control of a third target character among the one or more characters, displaying a details interface of the third target character, wherein the details interface includes: at least one of the third target character's setting information, image, character relationship information, sound audio, appearance chapter information, and ability information.

[0207] In some embodiments, the method further includes: displaying controls for one or more scenes, wherein the one or more scenes include one or more scenes in a target chapter or the one or more scenes include one or more scenes in interactive multimedia content; and displaying an image of the second target scene in response to a triggering operation on a control for a second target scene in the one or more scenes.

[0208] In some embodiments, the method further includes: displaying controls for one or more audios, wherein the one or more audios include one or more audios in a target chapter or one or more scenes include one or more audios in interactive multimedia content; and playing the third target audio in response to a triggering operation on a control for a third target audio among the one or more audios.

[0209] In some embodiments, the preview interface corresponding to the target chapter includes images of characters, images of scenes, and text in the target chapter, and the processing method further includes: displaying the edited text in response to an editing operation on the text.

[0210] In some embodiments, the method further includes: displaying a character selection interface for the interactive multimedia content; and displaying the interactive multimedia content according to the selected character in response to a selection operation on a character in the interactive multimedia content.

[0211] In some embodiments, the method further includes: in response to the plot development of the interactive multimedia content reaching a plot point that generates one or more story branches, displaying a selection interface, wherein the selection interface includes multiple option information; in response to a triggering operation on target option information among the multiple option information, displaying a screen of the main story line or story branch corresponding to the target option information.

[0212] In some embodiments, the method further includes: displaying a selection interface in response to the plot development of the interactive multimedia content reaching a plot point that generates one or more story branches; receiving input selection information; identifying the selection information and determining the story branch corresponding to the selection information; and displaying a screen of the story branch corresponding to the selection information.

[0213] According to other embodiments of the present disclosure, a device for processing interactive multimedia content is provided, including: a text generation module, configured to generate content texts of one or more story branches based on the content text of the original story; an image generation module, configured to generate images based on the content text of the original story and the content texts of one or more story branches, wherein the images include images of characters and images of scenes; and a multimedia generation module, configured to generate interactive multimedia content based on the content text of the original story, the content texts of one or more story branches, and the images.

[0214] According to some further embodiments of the present disclosure, an electronic device is provided, comprising: a memory; and a processor coupled to the memory, wherein the processor is configured to execute a method for processing interactive multimedia content as in any embodiment of the present disclosure based on instructions stored in the memory.

[0215] According to some further embodiments of the present disclosure, a computer-readable storage medium is provided, on which a computer program is stored. When the program is executed by a processor, the method for processing interactive multimedia content according to any embodiment of the present disclosure is implemented.

[0216] According to some further embodiments of the present disclosure, a computer program is provided, comprising: instructions, which, when executed by a processor, cause the processor to execute the method for processing interactive multimedia content according to any one of the embodiments of the present disclosure.

[0217] According to some embodiments of the present disclosure, a computer program product is provided, comprising instructions, which, when executed by a processor, implement the method for processing interactive multimedia content according to any one of the embodiments of the present disclosure.

[0218] The above descriptions are merely some embodiments of the present disclosure and an illustration of the technical principles employed. Those skilled in the art should understand that the scope of disclosure involved in the present disclosure is not limited to the technical solutions formed by a specific combination of the above-mentioned technical features, but also encompasses other technical solutions formed by any combination of the above-mentioned technical features or their equivalents without departing from the above-mentioned disclosed concepts. For example, a technical solution formed by replacing the above-mentioned features with (but not limited to) technical features with similar functions disclosed in the present disclosure.

[0219] In the description provided herein, numerous specific details are set forth. However, it is understood that embodiments of the present invention may be practiced without these specific details. In other cases, well-known methods, structures, and techniques are not presented in detail in order not to obscure the understanding of the description.

[0220] In addition, although each operation is described in a specific order, this should not be understood as requiring these operations to be performed in the specific order shown or in a sequential order. Under certain circumstances, multitasking and parallel processing may be advantageous. Similarly, although some specific implementation details have been included in the above discussion, these should not be interpreted as limiting the scope of the present disclosure. Some features described in the context of a separate embodiment can also be implemented in a single embodiment in combination. On the contrary, the various features described in the context of a single embodiment can also be implemented in multiple embodiments individually or in any suitable sub-combination mode.

[0221] Although some specific embodiments of the present disclosure have been described in detail by way of examples, those skilled in the art will appreciate that the above examples are for illustrative purposes only and are not intended to limit the scope of the present disclosure. Those skilled in the art will appreciate that modifications may be made to the above embodiments without departing from the scope and spirit of the present disclosure. The scope of the present disclosure is defined by the appended claims.

Claims

1. A method for processing interactive multimedia content, comprising: Generate content texts for one or more story branches based on the content text of the original story; generating an image based on the content text of the original story and the content text of the one or more story branches, wherein the image includes an image of a character and an image of a scene; Interactive multimedia content is generated based on the content text of the original story, the content text of the one or more story branches, and the image.

2. The processing method according to claim 1, further comprising: The main story line of the original story and the one or more branch story lines are displayed, wherein the main story line includes a plurality of main story line nodes, and each of the one or more branch story lines includes a plurality of branch line nodes.

3. The processing method according to claim 2, wherein: The one or more story branches include one or more story branches corresponding to each of the one or more characters, and the display of the main story line of the original story and the one or more story branches includes at least one of the following: For each character, display the main storyline of the original story and one or more story branches corresponding to the character; In response to a selection operation on a first target character among the one or more characters, a main story line of the original story and one or more branch story lines corresponding to the first target character are displayed.

4. The processing method according to claim 2 or 3, wherein: The display of the main story line of the original story and the one or more branch story lines includes at least one of the following: For each main line node except the last main line node among the plurality of main line nodes, an identifier of the chapter corresponding to the main line node is displayed, and if a branch line node is derived from the main line node, option information corresponding to the main line node is displayed; Displaying the ending information corresponding to the last main line node among the multiple main line nodes; For each branch node other than the last branch node among the multiple branch nodes of each story branch, if the branch node is a branch node derived from another node or other branch nodes are derived from the branch node, display option information corresponding to the branch node, wherein the other nodes include main line nodes or branch nodes other than the branch node; Display the ending information corresponding to the last branch node among the multiple branch nodes of each story branch.

5. The processing method according to claim 4, further comprising at least one of the following: In response to a modification operation on option information of a first target mainline node among the plurality of mainline nodes, displaying the modified option information of the first target mainline node and the modified option information of branch nodes derived from the first target mainline node; In response to modification of the content text of a first target branch node among a plurality of branch nodes in one or more story branches, the modified content text of the first target branch node is displayed.

6. The processing method according to any one of claims 1 to 5, wherein: Generating content texts of one or more story branches based on the content text of the original story includes: Determining one or more main storyline nodes in the main storyline of the original story based on the content text of the original story; Derived one or more branch nodes from each of the one or more main line nodes as one or more initial branch nodes; Based on each of the one or more initial branch nodes, multiple branch nodes in a story line are generated, and the content of the multiple branch nodes is expanded to generate content text of the one or more story branches.

7. The processing method according to claim 6, wherein: The one or more story branches include one or more story branches corresponding to each of one or more characters, and determining one or more mainline nodes in the main storyline of the original story based on the content text of the original story includes, for each character: Selecting, from the plurality of chapters of the main storyline according to the content text of the original story, one or more key chapters associated with the story of the character and one or more key plots in each of the one or more key chapters; One or more main line nodes associated with the characters in the main story line of the original story are generated according to one or more key plots in each key chapter.

8. The processing method according to claim 7, wherein: The deriving one or more branch nodes from each of the one or more main line nodes includes: Determining option information for each main line node according to the key plot corresponding to each main line node, wherein the option information is used to describe the selection of the character associated with the main line node; For each main line node, one or more option information different from the option information of the main line node and one or more branch line nodes corresponding to the one or more option information are generated according to the option information of the main line node.

9. The processing method according to any one of claims 6 to 8, wherein: The expanding the content of the multiple branch nodes includes: For each of the multiple branch nodes, the content of the branch node is expanded based on the content text of the main line node before the branch node and at least one of the branch nodes, and the information of the characters in the story branch to which the branch node belongs, wherein the information of the characters includes at least one of the setting information of the characters and the relationship information between the characters.

10. The treatment method according to any one of claims 1 to 9, further comprising at least one of the following: Associating and displaying the description text of the character in the content text of the original story with the image of the character; The description text of the scene in the content text of the original story is associated with the image of the scene and displayed.

11. The processing method according to claim 10, wherein: There are multiple characters and multiple scenes, and the processing method further includes at least one of the following: In response to an update operation on an image of a second target person among the plurality of persons, regenerating and displaying the image of the second target person; For each of the plurality of characters, displaying a plurality of images of the character, and in response to a selection operation on a first target image among the plurality of images of the character, determining the first target image as the image of the character; In response to an update operation on an image of a first target scene among the multiple scenes, regenerate and display the image of the first target scene; For each of a plurality of scenes, a plurality of images of the scene are displayed, and in response to a selection operation on a second target image among the plurality of images of the scene, the second target image is determined as the image of the scene.

12. The processing method according to any one of claims 1 to 11, wherein: Generating an image based on the content text of the original story and the content text of the one or more story branches includes: Extracting description text of each of the multiple characters from the content text of the original story and the content text of the one or more story branches; Generate an image of each character according to the description text of each character; Extracting description text for each of the multiple scenes from the content text of the original story and the content text of the one or more story branches; An image of each scene is generated according to the description text of each scene.

13. The processing method according to any one of claims 1 to 12, wherein: The image also includes a map, and the processing method further includes: displaying the map; In response to the selection operation of the target location in the map, the scene corresponding to the target location is displayed in association. A description text of the scene corresponding to the image and the target location.

14. The processing method according to any one of claims 1 to 13, wherein: The image further includes a map, and generating the image based on the content text of the original story and the content text of the one or more story branches includes: Extracting description texts of a plurality of locations from the content text of the original story and the content texts of the one or more story branches, wherein the description texts of the plurality of locations include positional relationships between the plurality of locations; The map is generated according to the description texts of the plurality of places.

15. The processing method according to any one of claims 1 to 14, further comprising: Generate audio based on the content text of the original story and the content text of the one or more story branches, wherein the audio includes the voice audio of the characters and the background sound audio of the scene, The step of generating interactive multimedia content based on the content text of the original story, the content text of the one or more story branches, and the image includes: Interactive multimedia content is generated based on the content text of the original story, the content text of the one or more story branches, the image and the audio.

16. The processing method according to claim 15, wherein: There are multiple characters and multiple scenes, and the processing method further includes at least one of the following: Associating and displaying the voice audios of a plurality of characters with the description texts of the plurality of characters, and in response to a selection operation of a first target audio among the voice audios of the plurality of characters, playing the first target audio; The background sound audios of a plurality of scenes are displayed in association with the description texts of the plurality of scenes, and in response to a selection operation of a second target audio in the background sound audios of the plurality of scenes, the second target audio is played.

17. The treatment method according to claim 15 or 16, wherein: Generating audio according to the content text of the original story and the content text of the one or more story branches includes: Extracting setting information for each of the multiple characters from the content text of the original story and the content text of the one or more story branches; Generating audio of each character according to the setting information of each character; Extracting description text of each of the multiple scenes from the content text of the original story and the content text of the one or more story branches, and determining a story style; Generate background sound audio for each scene according to the description text of each scene and the story style.

18. The processing method according to any one of claims 1 to 17, further comprising: Displaying a preview interface of the interactive multimedia content, wherein the preview interface includes a control corresponding to each of the multiple chapters of the interactive multimedia content; In response to a triggering operation on a control of a target chapter among the multiple chapters, a preview interface corresponding to the target chapter is displayed.

19. The processing method according to claim 18, further comprising: Displaying a control for one or more characters, wherein the one or more characters include one or more characters appearing in the target chapter, or the one or more characters include one or more characters in the interactive multimedia content; In response to a triggering operation on a control of a third target character among the one or more characters, a detail interface of the third target character is displayed, wherein the detail interface includes at least one of: setting information, image, character relationship information, sound audio, appearance chapter information, and ability information of the third target character.

20. The processing method according to claim 18 or 19, further comprising: a control for displaying one or more scenes, wherein the one or more scenes include one or more scenes in the target chapter or the one or more scenes include one or more scenes in the interactive multimedia content; In response to a triggering operation on a control of a second target scene among the one or more scenes, an image of the second target scene is displayed.

21. The processing method according to any one of claims 18 to 20, further comprising: A control for displaying one or more audios, wherein the one or more audios include one or more audios in the target chapter or the one or more scenes include one or more audios in the interactive multimedia content; In response to a triggering operation on a control of a third target audio among the one or more audios, the third target audio is played.

22. The treatment method according to any one of claims 18 to 21, wherein: The preview interface corresponding to the target chapter includes images of characters, images of scenes, and text in the target chapter, and the processing method further includes: In response to an editing operation on the text, the edited text is displayed.

23. The processing method according to any one of claims 1 to 22, further comprising: Displaying a character selection interface for the interactive multimedia content; In response to a selection operation on a character in the interactive multimedia content, the interactive multimedia content is displayed according to the selected character.

24. The processing method according to claim 23, further comprising: In response to the plot development of the interactive multimedia content reaching a plot point generating the one or more story branches, displaying a selection interface, wherein the selection interface includes a plurality of option information; In response to a triggering operation on target option information among the plurality of option information, a screen of a main story line or a branch story line corresponding to the target option information is displayed.

25. The processing method according to claim 23 or 24, further comprising: In response to the plot development of the interactive multimedia content reaching a plot point that generates the one or more story branches, displaying a selection interface; Receive input selection information; Identifying the selection information and determining a story branch corresponding to the selection information; A screen of a story branch corresponding to the selection information is displayed.

26. A device for processing interactive multimedia content, comprising: A text generation module is configured to generate content texts of one or more story branches based on the content text of the original story; an image generation module configured to generate an image based on the content text of the original story and the content text of the one or more story branches, wherein the image includes an image of a character and an image of a scene; The multimedia generation module is configured to generate interactive multimedia content based on the content text of the original story, the content text of the one or more story branches and the image.

27. An electronic device comprising: processor; as well as A memory coupled to the processor, for storing instructions, wherein when the instructions are executed by the processor, the processor executes the method for processing interactive multimedia content according to any one of claims 1 to 25.

28. A computer-readable storage medium having a computer program stored thereon, wherein: When the program is executed by a processor, the steps of the method for processing interactive multimedia content according to any one of claims 1 to 25 are implemented.

29. A computer program product comprising: Instructions, when executed by a processor, implement the steps of the method for processing interactive multimedia content according to any one of claims 1 to 25.

30. A computer program comprising: Instructions, when executed by a processor, implement the steps of the method for processing interactive multimedia content according to any one of claims 1 to 25.

Citation Information

Patent Citations

  • Interactive multimedia content processing method and device, electronic equipment and storage medium

    CN113014985A

  • Media content display method and device, equipment and storage medium

    CN116704078A

  • Method and system for automatically generating story video in meta universe

    CN117177003A

  • Story video generation method and device, storage medium and equipment

    CN117332118A

  • Generation method and device of interactive multimedia content, electronic equipment and storage medium

    CN117633258A