Video generation method and device, computer device, and storage medium

By displaying editing instructions and target video scripts on the video creation page, the problem of users having to conceive and collect book list content themselves in existing technologies is solved, thus simplifying the video generation process.

CN116668786BActive Publication Date: 2026-05-29BEIJING ZITIAO NETWORK TECH CO LTD

Patent Information

Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
BEIJING ZITIAO NETWORK TECH CO LTD
Filing Date
2023-05-31
Publication Date
2026-05-29

Smart Images

  • Figure CN116668786B_ABST
    Figure CN116668786B_ABST
Patent Text Reader

Abstract

The present disclosure provides a video generation method and device, computer equipment and a storage medium, wherein the method comprises: in response to generating a recommended video for a target book list, displaying a video creation page; the video creation page comprises editing instruction information corresponding to each video constituent dimension, and the editing instruction information is used to guide the addition of target video materials under each video constituent dimension; in response to receiving target video materials under each video constituent dimension, integrating the target video materials by using a target video script to generate a recommended video corresponding to the target book list; the target video script indicates target video frames associated with each video constituent dimension in the recommended video, and the display position of target video materials corresponding to the video constituent dimension in the target video frames.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] This disclosure relates to the field of video creation technology, and more specifically, to a method, apparatus, computer device, and storage medium for generating video. Background Technology

[0002] When creating videos, such as recommendation videos for book lists, users can collect various materials needed to introduce the book list themselves in some ways. Then, they can import these materials as pictures and text through video editing software, so that the video editing software can incorporate this information into the video frames to create a recommendation video that includes book list introduction information.

[0003] However, the tools available to users under the above methods are relatively limited. Video editing software can only provide video editing functions. Users also need to conceive the video structure for creating the video and collect and organize the book list content to be arranged in the video before they can use the video editing software to create an introductory video. The operation is also relatively complicated. Summary of the Invention

[0004] This disclosure provides at least one method, apparatus, computer device, and storage medium for generating video.

[0005] In a first aspect, embodiments of this disclosure provide a video generation method, comprising: displaying a video creation page in response to generating a recommended video for a target book list; the video creation page including editing instruction information corresponding to multiple video composition dimensions, the editing instruction information being used to guide the addition of target video materials under each video composition dimension; in response to receiving the target video materials under each of the video composition dimensions, integrating the target video materials using a target video script to generate a recommended video corresponding to the target book list; the target video script indicating the target video frames associated with each of the video composition dimensions in the recommended video, and the display position of the target video materials corresponding to the video composition dimensions in the target video frames.

[0006] In one optional implementation, after displaying editing instructions on the video creation page, the method further includes: in response to a triggering operation on any editing instruction, obtaining at least one first reference material associated with the target book list under the video composition dimension corresponding to the editing instruction; and in response to a selection operation on the first reference material, using the selected first reference material as the target video material under the video composition dimension.

[0007] In one optional implementation, after displaying editing instructions on the video creation page, the method further includes: in response to receiving input information under any editing instructions, determining multiple book description dimensions corresponding to the target book list based on the video composition dimension corresponding to the editing instructions, and determining a target book description dimension from the multiple book description dimensions based on the input information; obtaining book information under the target book description dimension from the target book list, and displaying the book information in association with the input information.

[0008] In one optional implementation, in response to receiving target video materials under each of the video composition dimensions, the method further includes: displaying identification information corresponding to multiple video scripts respectively, and in response to a selection operation of the first identification information, generating a first preview video based on the video script corresponding to the first identification information and the target video materials under each video composition dimension; displaying the first preview video, and in response to a switching selection operation of the second identification information, switching to displaying a second preview video under the video script corresponding to the second identification information.

[0009] In one optional implementation, displaying identification information corresponding to multiple video scripts further includes: determining the book list type of the target book list, and determining a video style matching the target book list based on the book list type; the book list type of the target book list is determined based on the book types corresponding to multiple books in the target book list; from the multiple video scripts, determining a target video script that matches the video style, and recommending and displaying the identification information corresponding to the target video script.

[0010] In one optional implementation, the video composition dimension includes a background audio dimension. The target video material under the background audio dimension is obtained as follows: in response to receiving an editing operation corresponding to the editing instruction information under the background audio dimension, a selected audio segment to be recorded is determined from the background audio under the recommended video; the currently generated preview video is played, and when the video segment corresponding to the audio segment to be recorded is played in the preview video, the recorded audio for the audio segment to be recorded is received; the recorded audio is used to update the audio segment to be recorded to obtain the video material under the background audio dimension.

[0011] In one optional implementation, after generating the recommended video corresponding to the target book list, the method further includes: playing the recommended video on the video recommendation page, and marking the video segments in the recommended video that are associated with each book in the target book list in the progress bar of the recommended video; in response to the selection operation of any of the video segments, jumping to play the recommended video from the starting playback position of the video segment in the recommended video, and displaying the reading information of the books associated with the video segment.

[0012] Secondly, embodiments of this disclosure also provide a video generation apparatus, comprising: a display module, configured to display a video creation page in response to generating a recommended video for a target book list; the video creation page includes editing instruction information corresponding to multiple video composition dimensions, the editing instruction information being used to guide the addition of target video materials under each video composition dimension; and a generation module, configured to, in response to receiving the target video materials under each of the video composition dimensions, integrate the target video materials using a target video script to generate a recommended video corresponding to the target book list; the target video script indicating the target video frames associated with each of the video composition dimensions in the recommended video, and the display position of the target video materials corresponding to the video composition dimensions in the target video frames.

[0013] In one optional implementation, after displaying editing instructions on the video creation page, the display module is further configured to: in response to a trigger operation on any editing instructions, obtain at least one first reference material associated with the target book list under the video composition dimension corresponding to the editing instructions; and in response to a selection operation on the first reference material, use the selected first reference material as the target video material under the video composition dimension.

[0014] In one optional implementation, after displaying editing instructions on the video creation page, the display module is further configured to: in response to receiving input information under any editing instructions, determine multiple book description dimensions corresponding to the target book list based on the video composition dimension corresponding to the editing instructions, and determine the target book description dimension from the multiple book description dimensions based on the input information; obtain book information under the target book description dimension from the target book list, and associate and display the book information with the input information.

[0015] In one optional implementation, after receiving the target video materials under each of the video composition dimensions, the generation module is further configured to: display the identification information corresponding to the multiple video scripts respectively, and in response to the selection operation of the first identification information, generate a first preview video based on the video script corresponding to the first identification information and the target video materials under each video composition dimension; display the first preview video, and in response to the switching selection operation of the second identification information, switch to displaying the second preview video under the video script corresponding to the second identification information.

[0016] In one optional implementation, when the generation module displays the identification information corresponding to multiple video scripts, it is further configured to: determine the book list type of the target book list, and determine the video style matching the target book list based on the book list type; the book list type of the target book list is determined based on the book types corresponding to multiple books in the target book list; determine the target video script that matches the video style from the multiple video scripts, and recommend and display the identification information corresponding to the target video script.

[0017] In one optional implementation, the video composition dimension of the generation module includes a background audio dimension, and the target video material under the background audio dimension is obtained in the following manner: in response to receiving an editing operation corresponding to the editing instruction information under the background audio dimension, the audio segment to be recorded is selected from the background audio under the recommended video; the currently generated preview video is played, and when the video segment corresponding to the audio segment to be recorded is played in the preview video, the recorded audio for the audio segment to be recorded is received; the recorded audio is used to update the audio segment to be recorded to obtain the video material under the background audio dimension.

[0018] In one optional implementation, after generating the recommended video corresponding to the target book list, the generation module is further configured to: play the recommended video on the video recommendation page, and mark the video segments in the recommended video that are associated with each book in the target book list in the progress bar of the recommended video; in response to the selection operation of any of the video segments, jump to play the recommended video from the starting playback position of the video segment in the recommended video, and display the reading information of the books associated with the video segment.

[0019] Thirdly, an optional implementation of this disclosure also provides a computer device, a processor, and a memory, wherein the memory stores machine-readable instructions executable by the processor, and the processor is configured to execute the machine-readable instructions stored in the memory. When the machine-readable instructions are executed by the processor, the steps of the first aspect above, or any possible implementation of the first aspect, are performed.

[0020] Fourthly, an optional implementation of this disclosure also provides a computer-readable storage medium storing a computer program that, when run, performs the steps of the first aspect or any possible implementation of the first aspect.

[0021] This disclosure provides a video generation method, apparatus, computer device, and storage medium, offering a video creation page for creating recommended videos corresponding to target book lists, facilitating user creation of recommended videos. The video creation page specifically displays editing instructions for multiple video composition dimensions involved in creating recommended videos for target book lists. These instructions guide users in adding video materials used to build the video, eliminating the need for users to manually determine what information to collect. Furthermore, after receiving target video materials under each video composition dimension, the target video materials can be directly integrated using a target video script. This target video script specifically instructs the corresponding display video frames and positions under the recommended videos for the video materials. Therefore, it allows users to directly obtain the arranged recommended videos without needing to arrange the book list content, simplifying user operation.

[0022] To make the above-mentioned objects, features and advantages of this disclosure more apparent and understandable, preferred embodiments are described below in detail with reference to the accompanying drawings. Attached Figure Description

[0023] To more clearly illustrate the technical solutions of the embodiments of this disclosure, the accompanying drawings used in the embodiments will be briefly described below. These drawings are incorporated in and constitute a part of this specification. They illustrate embodiments conforming to this disclosure and, together with the specification, serve to explain the technical solutions of this disclosure. It should be understood that the following drawings only show some embodiments of this disclosure and should not be considered as limiting the scope. Those skilled in the art can obtain other related drawings based on these drawings without creative effort.

[0024] Figure 1 A flowchart illustrating a video generation method provided in an embodiment of this disclosure is shown;

[0025] Figure 2 This illustration shows a diagram corresponding to the icon for generating recommended videos displayed on a topic discussion page, as provided in an embodiment of this disclosure.

[0026] Figure 3 A schematic diagram of a video creation page provided in an embodiment of this disclosure is shown;

[0027] Figure 4 A schematic diagram of another video creation page provided by an embodiment of this disclosure is shown;

[0028] Figure 5 A schematic diagram of a third type of video creation page provided in an embodiment of this disclosure is shown;

[0029] Figure 6A schematic diagram of the fourth video creation page provided in this embodiment of the present disclosure is shown;

[0030] Figure 7 A schematic diagram of multiple video frames under a recommended video provided in an embodiment of this disclosure is shown;

[0031] Figure 8 This diagram illustrates a background sound editing process provided by an embodiment of the present disclosure.

[0032] Figure 9 A schematic diagram of a video recommendation page provided in an embodiment of this disclosure is shown;

[0033] Figure 10 A schematic diagram of a video generation apparatus provided in an embodiment of this disclosure is shown;

[0034] Figure 11 A schematic diagram of a computer device provided in an embodiment of this disclosure is shown. Detailed Implementation

[0035] To make the objectives, technical solutions, and advantages of the embodiments of this disclosure clearer, the technical solutions of the embodiments of this disclosure will be clearly and completely described below with reference to the accompanying drawings. Obviously, the described embodiments are only a part of the embodiments of this disclosure, and not all of them. The components of the embodiments of this disclosure described and shown herein can generally be arranged and designed in various different configurations. Therefore, the following detailed description of the embodiments of this disclosure is not intended to limit the scope of the claimed disclosure, but merely represents selected embodiments of this disclosure. All other embodiments obtained by those skilled in the art based on the embodiments of this disclosure without inventive effort are within the scope of protection of this disclosure.

[0036] Research has found that when generating recommendation videos for book lists, users can use video editing software to incorporate self-collected book list information into video frames. However, this method offers limited tools. Since users can only create videos using video editing tools, and the software only provides basic editing functions, users still need to conceive the video structure and collect and organize the book list content to be included. The process is relatively complex.

[0037] Based on the above research, this disclosure provides a video generation method, offering a video creation page for creating recommended videos corresponding to target book lists, facilitating user creation of these videos. The video creation page specifically displays editing instructions for multiple video composition dimensions involved in creating recommended videos for target book lists. These instructions guide users in adding video materials used to build the video, eliminating the need for users to manually conceive of what information to collect. Furthermore, after receiving the target video materials for each video composition dimension, the target video materials can be directly integrated using a target video script. This target video script specifically instructs the corresponding display video frames and positions under the recommended videos for the video materials. Therefore, it allows users to directly obtain the arranged and organized recommended videos without needing to arrange the book list content, simplifying the user experience.

[0038] The shortcomings of the above solutions are the result of the inventor's practical experience and careful research. Therefore, the discovery process of the above problems and the solutions proposed in this disclosure below should be considered as the inventor's contribution to this disclosure.

[0039] It should be noted that similar labels and letters in the following figures indicate similar items. Therefore, once an item is defined in one figure, it does not need to be further defined and explained in subsequent figures.

[0040] To facilitate understanding of this embodiment, a video generation method disclosed in this disclosure will first be described in detail. The execution entity of the video generation method provided in this disclosure is generally a computer device with certain computing capabilities. This computer device may include, for example, a terminal device, a server, or other processing devices. The terminal device may be a user equipment (UE), mobile device, user terminal, terminal, cellular phone, cordless phone, personal digital assistant (PDA), handheld device, computing device, in-vehicle device, wearable device, etc. In some possible implementations, the video generation method can be implemented by a processor calling computer-readable instructions stored in memory.

[0041] The video generation method provided in this disclosure is described below. Since the video generation method provided in this disclosure involves book lists and video playback, it can be specifically applied to book reading platforms and product sales platforms related to book lists, as well as video playback platforms related to video playback. For example, taking a book reading platform as an example, on the book list recommendation page or on the topic discussion page where book lists can be discussed, an entry point can be provided for editing videos of the book lists involved, and the relevant book lists can be selected as the target book lists for generating recommended videos.

[0042] See Figure 1 The diagram shows a flowchart of a video generation method provided in an embodiment of this disclosure. The method includes steps S101 to S102, wherein:

[0043] S101: In response to generating recommended videos for the target book list, display the video creation page; the video creation page includes editing instructions corresponding to multiple video composition dimensions, and the editing instructions are used to guide the addition of target video materials under each video composition dimension;

[0044] S102: In response to receiving target video materials under each of the video composition dimensions, the target video materials are integrated using a target video script to generate recommended videos corresponding to the target book list; the target video script indicates the target video frames associated with each of the video composition dimensions in the recommended videos, and the display position of the target video materials corresponding to the video composition dimensions in the target video frames.

[0045] Regarding S101 above, taking the book reading platform described above as an example, we will specifically explain the possible situations where recommended videos are generated for the target book list.

[0046] For example, on a book reading platform, specifically on the book list recommendation page or the discussion page for book lists, in addition to displaying the book lists in list format or other possible formats, you can also display instructions to generate recommended videos, such as an icon for generating recommended videos. Specifically, if a user posts a recommended book list on the book reading platform, such as providing multiple books of a certain novel genre through a comment under a novel topic, and this is identified as a book list, then when a book list is detected, the corresponding icon for generating recommended videos can be displayed in the associated location.

[0047] For example, see Figure 2The diagram illustrates an icon for generating recommended videos displayed on a topic discussion page, as provided in this embodiment of the disclosure. Specifically, it shows a user's message: "Highly recommend these novels: Novel A by author A, and Novel B and Novel C by author B, both published last year." The novels "Novel A," "Novel B," and "Novel C" can be used to construct a target book list. The icon for generating recommended videos is displayed in the relevant position to the right of the user's message. Responding to the triggering of this icon, it can be determined that the user's current intention includes generating recommended videos for the target book list, and the video creation page is displayed.

[0048] The displayed video creation page includes editing instructions corresponding to various video composition dimensions. These instructions guide users in adding target video materials for each dimension. Specifically, these instructions are displayed as icons or characters. Video composition dimensions can be divided based on the video's structure, such as dividing it into video segments representing the book list and the individual books within that list, or further subdividing it into cover frames displaying the book list and video frames showcasing each book.

[0049] Within each video component dimension, specific content related to the book list or books can be displayed within the indicated video frames. This content includes, but is not limited to: the book list's name, introductory information, reasons for recommendation, and the titles, authors, and brief descriptions of each book within it.

[0050] Here, when displaying editing instructions corresponding to multiple video composition dimensions, in one possible scenario, the editing instructions for each video composition dimension can be displayed page by page, following the order of the book list and the books. For example, see... Figure 3 The diagram shown is a schematic representation of a video creation page provided in an embodiment of this disclosure. In this approach, a step index prompts the user to add content information corresponding to the object indicated by the index on each page.

[0051] For example, in Figure 3 (a) shows several editing instructions under the "Video Cover" index, guiding users to add information specifically for introducing the book list. These instructions include the prompt "a. Add book list name," followed by an information editing box with the tag "Please enter book list name." After completing the information on the page and turning the page, or if the user actively turns the page, the following instructions will be displayed: Figure 3The page shown in (b) continues to supplement the content when generating a recommended video for the first book in the book list. It is easy to understand that for the two different objects, book lists and books, the corresponding video composition dimensions are different when generating recommended videos, and therefore the corresponding editing instructions are different.

[0052] In this approach, guide words can be used on each page to convey to the user the specific object to which the editing instructions displayed on that page are targeted, so that the user can clearly understand which object is being described and systematically add the corresponding description information of each object as video material.

[0053] In another possible approach, the editing instructions displayed on the topic creation page, across multiple video composition dimensions, could be selectively presented based on user needs. For instance, the user's creative needs within the recommended videos could be inferred from the specific scenario that triggered the generation of those videos.

[0054] For example, in the scenario where a recommended video is generated on the aforementioned topic creation page, based on the user's posted information such as "I strongly recommend these novels: Novel A by author A, and Novel B and Novel C by author B, which were released last year," we can first determine the books the user wants to recommend in this video creation, specifically "Novel A," "Novel B," and "Novel C." Based on these three confirmed books, we can first retrieve available video materials, such as the book titles, covers, and author introductions. However, some information remains uncertain, such as the corresponding book list title and the reasons for recommendation. Furthermore, the user may also add or remove other books from these three lists.

[0055] Based on this, when displaying it on the video creation page, for... Figure 3 For videos with the same structure as the examples, if the corresponding information is available, it will be displayed directly, and the corresponding editing instructions can be used to indicate whether to confirm or change this information. For information that cannot be obtained, it can be processed according to... Figure 3 In a similar manner, the system guides users to fill in further information by providing editing instructions.

[0056] For example, see Figure 4 The diagram shown is a schematic representation of another video creation page provided in an embodiment of this disclosure. Figure 4 In the middle, regarding the first book, "Novel A", and Figure 3Unlike (b), since the book title, cover, and synopsis are available, the corresponding editing instructions can be omitted or replaced with the obtained information and the corresponding confirm or edit buttons. However, for the recommendation reason, this part requires the user to fill in the details themselves; therefore, the displayed editing instructions can be the same as (b). Figure 3 The same as (b).

[0057] Comparing the two methods above, it can be determined that on the video creation page, a better approach is to provide users with specific, selectable information to reduce the burden on users when searching for and collecting information.

[0058] Based on this, the embodiments of this disclosure also select the following two different methods to provide users with optional video materials:

[0059] (1): In response to the triggering operation of any editing instruction information, obtain at least one first reference material associated with the target book list under the video composition dimension corresponding to the editing instruction information; in response to the selection operation of the first reference material, use the selected first reference material as the target video material under the video composition dimension.

[0060] In this case, for example, see Figure 5 The diagram illustrates a third type of video creation page provided in this embodiment. Taking the selected video composition dimension, which includes the cover of video segment 1, as an example, the corresponding editing instructions may include a text box labeled "Please fill in the book synopsis," indicating that a book synopsis will be displayed on the cover. In response to triggering this text box, multiple first reference materials displayed at the bottom of the page can be shown, such as synopsis references for "Novel A" from different web pages. These first reference materials can be directly selected and used as target video materials in this video composition dimension, or they can be selected and edited, with the edited first reference materials used as target video materials.

[0061] (2): In response to receiving input information under any editing instruction information, based on the video composition dimension corresponding to the editing instruction information, determine multiple book description dimensions corresponding to the target book list, and determine the target book description dimension from the multiple book description dimensions based on the input information; obtain the book information under the target book description dimension from the target book list, and associate and display the book information with the input information.

[0062] In this case, we will first illustrate another example of a video creation page that can be provided under the embodiments of this disclosure. In this example, unlike the previous example, the editing instructions can be displayed in a simpler way, rather than the guided editing based on book lists and books as in the previous example. In this way, the flexibility of creating recommended videos can be improved, without being limited to providing multi-dimensional introductory content in the video creation page.

[0063] For example, see Figure 6 The image shown is an example diagram of the fourth type of video creation page provided in this embodiment of the disclosure. First, see... Figure 6 The page illustration shown in (a) illustrates an example where the editing instructions may consist of only a text box with the text "Enter the description you want to add to the book," instructing the user to describe the book below the cover of video segment 1. The user can type in the information they want to describe the book in the recommended video here.

[0064] Here, it can be determined that the editing instructions specifically correspond to the cover of video segment 1. Below the cover of video segment 1, multiple book description dimensions can be pre-defined, including the book title, cover, book synopsis, etc. If the input text includes "First, let me introduce 'Novel A'", then through semantic recognition, it can be determined that the user specifically wants to introduce "Novel A" on the cover, i.e., display the book synopsis. The book synopsis can then be used as the target book description dimension.

[0065] Then, information about the books listed in the target book list, corresponding to the description dimensions of the target books, can be retrieved and displayed in relation to them. Here, in one possible scenario, it can be done as follows: Figure 5 The suggested display method is shown below, or you can also use the following method: Figure 6 As shown in (b), the text is directly linked to the text entered by the user and displayed as a supplement; the details will not be elaborated here.

[0066] In this way, corresponding target video materials can be obtained under each video composition dimension. For example, the target video materials obtained under each video composition dimension may specifically include: video cover - book list name and book list introduction information; one video cover - cover and title of "Novel A", and corresponding video content - introduction of "Novel A"; one video cover - cover and title of "Novel B", and corresponding video content - introduction of "Novel B".

[0067] Regarding S102 above, after receiving the target video materials under each video composition dimension, the target video materials can be integrated according to the target video script to generate recommended videos corresponding to the target book list.

[0068] Here, the target video script specifically indicates the target video frame associated with each video composition dimension in the recommended videos. For example, for the video composition dimension of the video cover mentioned above, the specific associated target video frame is the first frame of the recommended video. In addition, the target video script is also used to indicate the display position of the target video material corresponding to the video composition dimension in the target video frame. For example, for the video cover, it can be displayed in full-screen magnification.

[0069] Thus, for each target video material obtained as described above, the corresponding video frame to be displayed, and its specific position within the video frame, can be determined based on the target video script. For example, see... Figure 7 The diagram shown is a schematic diagram of multiple video frames under the recommended video provided in the embodiments of this disclosure.

[0070] The video begins with the first frame displaying the cover, shown in full screen, featuring the book list title "Recommended Novel Collection" and the description "Three highly recommended novels." Starting with the fifth frame, information about the first book is displayed, including its cover and title, specifically "Book A." From the eighth frame onwards, the corresponding video content is shown, including a brief introduction to "Novel A."

[0071] The above is just one possible video script; in practice, multiple scripts can be chosen. The differences between video scripts can be reflected in the content structure and arrangement. For example, in the example above, the book list is introduced first, followed by a detailed introduction of each book; however, in other possible approaches, multiple books can be compared and introduced without being separated into individual video segments. Furthermore, differences in video scripts can also be reflected in the display position of the target video footage, etc., which will not be elaborated upon here.

[0072] For multiple selectable video scripts, in specific implementation, the identification information corresponding to each of the multiple video scripts can be displayed. In response to the selection operation of the first identification information, a first preview video is generated based on the video script corresponding to the first identification information and the target video material composed of each video dimension. The first preview video is displayed, and in response to the switching selection operation of the second identification information, the second preview video under the video script corresponding to the second identification information is switched to be displayed.

[0073] For example, the differences that may exist between the video scripts described above can result in different styles under different video scripts, such as a reading style that focuses more on showcasing the content of the book, or a commentary style that focuses more on comparison, and so on.

[0074] Before generating recommended videos, users can preview the videos to determine if further modifications are needed, or to generate and publish them. During previewing, users can select one of several video scripts to generate a preview video. When providing multiple video scripts to users, this can be done by displaying identification information corresponding to each video script. Specifically, this identification information can be displayed using textual information such as "reading style" or "commentary style" as described above, or other optional methods; no limitation is made here.

[0075] After selecting any identifier, the selected first identifier can be used to generate and display a first preview video using the target video materials from each video dimension under the corresponding video script. The method for generating the preview video is detailed above and will not be repeated here. Furthermore, if the user is not satisfied with the preview video generated under the selected video script, they can switch to other available video scripts to view the corresponding preview videos.

[0076] Here, in order to enable users to get the desired result after watching fewer preview videos and reduce the computing resources consumed by generating multiple preview videos, we can determine the more suitable video script to select when generating recommended videos based on the target book list obtained in advance, and then recommend and display the corresponding identification information.

[0077] In specific implementation, the book list type of the target book list can be determined, and the video style matching the target book list can be determined based on the book list type; the book list type of the target book list is determined based on the book types corresponding to multiple books in the target book list; from multiple video scripts, a target video script matching the video style is determined, and the identification information corresponding to the target video script is recommended and displayed.

[0078] For example, since the video script is predetermined, the video style can also be predetermined. For instance, video styles could include "ancient style," "modern style," "lighthearted and humorous style," "tragic style," and so on. Different video styles can correspond to different characteristics of the books. For example, under the "ancient style," the characteristics of the books are more likely to match historical novels or ancient romance novels. Therefore, if multiple books on a target book list are of the ancient historical genre, it is more appropriate to select the video script corresponding to the "ancient style" rather than the video script corresponding to the "modern style."

[0079] Therefore, based on the determined target book list, we can first determine the book list type, which can be specifically determined by the book types corresponding to multiple books within it. The specific book type can be derived from its corresponding content characteristics. When determining content characteristics, we can refer to the video styles of the currently available video scripts, specifically which available types they favor. Based on the book list type, we can determine the corresponding video styles, and then determine the target video scripts corresponding to those styles for recommendation and display to users.

[0080] In the above example, the text and image content in the recommended video were specifically explained. However, the video composition dimension can also include a background audio dimension. When obtaining target video material under the background audio dimension, the following method can be used: In response to receiving an editing operation corresponding to the editing instruction information under the background audio dimension, determine the selected audio segment to be recorded from the background audio of the recommended video; play the currently generated preview video, and when playing to the video segment in the preview video corresponding to the audio segment to be recorded, receive the recorded audio for the audio segment to be recorded; update the audio segment to be recorded using the recorded audio to obtain the video material under the background audio dimension.

[0081] One possible approach is to select a piece of music as the background sound for the video, or users can input audio as background sound by recording it.

[0082] During the process of editing background music, users can select a segment of audio and re-record it. For example, see [link to example]. Figure 8 The diagram shown is a schematic representation of background music editing according to an embodiment of this disclosure. In this diagram, a preview video is displayed at the top, and a playback progress bar and total duration of the complete background music corresponding to the preview video are displayed at the bottom.

[0083] When previewing the video, you can access the page shown in the example by editing the instructions, such as the audio recording button. You can then mark the audio segment to be recorded by sliding along the playback progress bar, for example, the audio segment corresponding to 01:00-01:10 in the illustration. To record this audio and keep the video content synchronized with the audio recording instructions, you can first play the currently generated preview video and simultaneously receive the recording audio for the segment to be recorded. Specifically, you can long-press the record button to continuously record audio. Finally, update the recording audio segment with the newly recorded audio to obtain the video footage in the background sound dimension.

[0084] After completing the above steps, you can determine whether to generate a recommended video by checking the preview video.

[0085] In this embodiment of the disclosure, after generating the recommended video, the recommended video can be played on the video recommendation page, and the video segments in the recommended video that are associated with each book in the target book list can be marked in the progress bar of the recommended video; in response to the selection operation of any of the video segments, the recommended video can be played from the starting playback position of the video segment, and the reading information of the books associated with the video segment can be displayed.

[0086] For example, see Figure 9 The image shows a schematic diagram of a video recommendation page provided in this embodiment. Recommended videos are displayed at the top, and a progress bar for the recommended videos is displayed at the bottom. The progress bar is displayed in multiple segments, with the current playback position indicated by a solid black circle. From the example image, it can be determined that the video segment corresponding to "Book A" is currently being displayed. Correspondingly, the relevant reading information is displayed below, including the cover of "Book A," the author's information, and some readable content. A corresponding link to jump to the reading page is also displayed, marked with the text "Click to Read."

[0087] For video segments associated with other books, the user can be redirected to the starting point of the corresponding video segment upon triggering the interaction. Correspondingly, when the user navigates to and displays the video segment for a particular book, the corresponding reading information for that book will also be displayed.

[0088] Here, if you switch to the video segment corresponding to the book list, you can also display multiple books under the book list, as well as the corresponding jump reading links for each book. The specifics can be determined according to the actual situation, and there are no restrictions here.

[0089] Those skilled in the art will understand that, in the above-described method of the specific implementation, the order in which each step is written does not imply a strict execution order and does not constitute any limitation on the implementation process. The specific execution order of each step should be determined by its function and possible internal logic.

[0090] Based on the same inventive concept, this disclosure also provides a video generation apparatus corresponding to the video generation method. Since the principle of the apparatus in this disclosure for solving the problem is similar to the video generation method described above, the implementation of the apparatus can refer to the implementation of the method, and repeated details will not be repeated.

[0091] Reference Figure 10 The diagram shown is a schematic representation of a video generation apparatus according to an embodiment of this disclosure. The apparatus includes: a display module 11 and a generation module 12; wherein,

[0092] Display module 11 is used to display a video creation page in response to generating recommended videos for the target book list; the video creation page includes editing instructions corresponding to multiple video composition dimensions, and the editing instructions are used to guide the addition of target video materials under each video composition dimension;

[0093] The generation module 12 is used to respond to the received target video materials under each of the video composition dimensions, integrate the target video materials using the target video script, and generate the recommended video corresponding to the target book list;

[0094] The target video script indicates the target video frame associated with each of the video composition dimensions in the recommended video, and the display position of the target video material corresponding to the video composition dimension in the target video frame.

[0095] In one optional implementation, after displaying editing instruction information on the video creation page, the display module 11 is further configured to: in response to a trigger operation on any editing instruction information, obtain at least one first reference material associated with the target book list under the video composition dimension corresponding to the editing instruction information; and in response to a selection operation on the first reference material, use the selected first reference material as the target video material under the video composition dimension.

[0096] In one optional implementation, after displaying editing instructions on the video creation page, the display module 11 is further configured to: in response to receiving input information under any editing instructions, determine multiple book description dimensions corresponding to the target book list based on the video composition dimension corresponding to the editing instructions, and determine the target book description dimension from the multiple book description dimensions based on the input information; obtain book information under the target book description dimension from the target book list, and associate and display the book information with the input information.

[0097] In one optional implementation, after receiving the target video materials under each of the video composition dimensions, the generation module 12 is further configured to: display the identification information corresponding to the multiple video scripts respectively, and in response to the selection operation of the first identification information, generate a first preview video based on the video script corresponding to the first identification information and the target video materials under each video composition dimension; display the first preview video, and in response to the switching selection operation of the second identification information, switch to displaying the second preview video under the video script corresponding to the second identification information.

[0098] In one optional implementation, when the generation module 12 displays the identification information corresponding to multiple video scripts, it is further configured to: determine the book list type of the target book list, and determine the video style matching the target book list based on the book list type; the book list type of the target book list is determined based on the book types corresponding to multiple books in the target book list; determine the target video script that matches the video style from the multiple video scripts, and recommend and display the identification information corresponding to the target video script.

[0099] In one optional implementation, the video composition dimension of the generation module 12 includes a background sound dimension, and the target video material under the background sound dimension is obtained in the following manner: in response to receiving an editing operation corresponding to the editing instruction information under the background sound dimension, the audio segment to be recorded is determined from the background audio under the recommended video; the currently generated preview video is played, and when the video segment corresponding to the audio segment to be recorded is played in the preview video, the recorded audio for the audio segment to be recorded is received; the recorded audio is used to update the audio segment to be recorded to obtain the video material under the background sound dimension.

[0100] In one optional implementation, after generating the recommended video corresponding to the target book list, the generation module 12 is further configured to: play the recommended video on the video recommendation page, and mark the video segments in the recommended video that are associated with each book in the target book list in the progress bar of the recommended video; in response to the selection operation of any of the video segments, jump to play the recommended video from the starting playback position of the video segment in the recommended video, and display the reading information of the books associated with the video segment.

[0101] The processing flow of each module in the device and the interaction flow between each module can be referred to the relevant descriptions in the above method embodiments, and will not be detailed here.

[0102] This disclosure also provides a computer device, such as... Figure 11 The diagram shown is a schematic representation of a computer device structure provided in an embodiment of this disclosure, including:

[0103] Processor 10 and memory 20; the memory 20 stores machine-readable instructions executable by processor 10, and processor 10 executes the machine-readable instructions stored in memory 20. When the machine-readable instructions are executed by processor 10, processor 10 performs the following steps:

[0104] In response to generating recommended videos for the target book list, a video creation page is displayed. The video creation page includes editing instructions corresponding to multiple video composition dimensions, which guide the addition of target video materials under each video composition dimension. In response to receiving the target video materials under each video composition dimension, the target video materials are integrated using a target video script to generate recommended videos corresponding to the target book list. The target video script indicates the target video frames associated with each video composition dimension in the recommended videos, and the display position of the target video materials corresponding to each video composition dimension in the target video frames.

[0105] The aforementioned memory 20 includes a main memory 210 and an external memory 220; the main memory 210, also known as internal memory, is used to temporarily store the computational data in the processor 10, as well as the data exchanged with external memory 220 such as a hard disk. The processor 10 exchanges data with the external memory 220 through the main memory 210.

[0106] The specific execution process of the above instructions can be referred to the steps of the video generation method described in the embodiments of this disclosure, and will not be repeated here.

[0107] This disclosure also provides a computer-readable storage medium storing a computer program that, when executed by a processor, performs the steps of the video generation method described in the above-described method embodiments. The storage medium can be a volatile or non-volatile computer-readable storage medium.

[0108] This disclosure also provides a computer program product carrying program code. The program code includes instructions that can be used to execute the steps of the video generation method described in the above method embodiments. For details, please refer to the above method embodiments, which will not be repeated here.

[0109] The aforementioned computer program product can be implemented through hardware, software, or a combination thereof. In one optional embodiment, the computer program product is specifically embodied in a computer storage medium; in another optional embodiment, the computer program product is specifically embodied in a software product, such as a software development kit (SDK), etc.

[0110] Those skilled in the art will clearly understand that, for the sake of convenience and brevity, the specific working processes of the systems and devices described above can be referred to the corresponding processes in the foregoing method embodiments, and will not be repeated here. In the several embodiments provided in this disclosure, it should be understood that the disclosed systems, devices, and methods can be implemented in other ways. The device embodiments described above are merely illustrative. For example, the division of units is only a logical functional division; in actual implementation, there may be other division methods. Furthermore, multiple units or components may be combined or integrated into another system, or some features may be ignored or not executed. Another point is that the displayed or discussed mutual coupling or direct coupling or communication connection may be through some communication interfaces; the indirect coupling or communication connection of devices or units may be electrical, mechanical, or other forms.

[0111] The units described as separate components may or may not be physically separate. The components shown as units may or may not be physical units; that is, they may be located in one place or distributed across multiple network units. Some or all of the units can be selected to achieve the purpose of this embodiment according to actual needs.

[0112] In addition, the functional units in the various embodiments of this disclosure can be integrated into one processing unit, or each unit can exist physically separately, or two or more units can be integrated into one unit.

[0113] If the aforementioned functions are implemented as software functional units and sold or used as independent products, they can be stored in a processor-executable, non-volatile, computer-readable storage medium. Based on this understanding, the technical solution of this disclosure, in essence, or the part that contributes to the prior art, or a portion of the technical solution, can be embodied in the form of a software product. This computer software product is stored in a storage medium and includes several instructions to cause a computer device (which may be a personal computer, server, or network device, etc.) to execute all or part of the steps of the methods described in the various embodiments of this disclosure. The aforementioned storage medium includes various media capable of storing program code, such as USB flash drives, portable hard drives, read-only memory (ROM), random access memory (RAM), magnetic disks, or optical disks.

[0114] Finally, it should be noted that the above-described embodiments are merely specific implementations of this disclosure, used to illustrate the technical solutions of this disclosure, and not to limit it. The protection scope of this disclosure is not limited thereto. Although this disclosure has been described in detail with reference to the foregoing embodiments, those skilled in the art should understand that any person skilled in the art can still modify or easily conceive of changes to the technical solutions described in the foregoing embodiments, or make equivalent substitutions for some of the technical features, within the scope of the technology disclosed in this disclosure. Such modifications, changes, or substitutions do not cause the essence of the corresponding technical solutions to deviate from the spirit and scope of the technical solutions of the embodiments of this disclosure, and should all be covered within the protection scope of this disclosure. Therefore, the protection scope of this disclosure should be determined by the protection scope of the claims.

Claims

1. A method for generating video, characterized in that, include: In response to generating recommended videos for the target book list, a video creation page is displayed; the video creation page includes editing instructions corresponding to multiple video composition dimensions, and the editing instructions are used to guide the addition of target video materials under each video composition dimension. In response to receiving target video materials under each of the video composition dimensions, the target video materials are integrated using the target video script to generate recommended videos corresponding to the target book list; The target video script indicates the target video frame associated with each of the video composition dimensions in the recommended video, and the display position of the target video material corresponding to the video composition dimension in the target video frame; The method further includes, after displaying the editing instructions on the video creation page: In response to receiving input information under any editing instruction, based on the video composition dimension corresponding to the any editing instruction, determine multiple book description dimensions corresponding to the target book list, and determine the target book description dimension from the multiple book description dimensions based on the input information; The book information under the target book description dimension is obtained from the target book list, and the book information is appended to the input information for supplementary display.

2. The method according to claim 1, characterized in that, After displaying editing instructions on the video creation page, the method further includes: In response to a trigger operation on any editing instruction, at least one first reference material associated with the target book list under the video composition dimension corresponding to the editing instruction is obtained; In response to the selection operation of the first reference material, the selected first reference material is used as the target video material in the video composition dimension.

3. The method according to claim 1, characterized in that, In response to receiving target video material in each of the aforementioned video composition dimensions, the method further includes: Display the identification information corresponding to multiple video scripts respectively, and in response to the selection operation of the first identification information, generate a first preview video based on the video script corresponding to the first identification information and the target video material composed of each video dimension; The first preview video is displayed, and in response to the switching selection operation of the second identifier information, the second preview video under the video script corresponding to the second identifier information is switched to be displayed.

4. The method according to claim 3, characterized in that, Displaying the identification information corresponding to multiple video scripts, including: The book list type of the target book list is determined, and a video style matching the target book list is determined based on the book list type; the book list type of the target book list is determined based on the book types corresponding to multiple books in the target book list; From the multiple video scripts, a target video script that matches the video style is determined, and the identification information corresponding to the target video script is recommended and displayed.

5. The method according to claim 1, characterized in that, The video composition dimension includes the background audio dimension, and the target video material under the background audio dimension is obtained using the following method: In response to receiving the editing operation corresponding to the editing instruction information under the background sound dimension, the audio segment to be recorded is determined from the background audio under the recommended video; Play the currently generated preview video, and when the video segment in the preview video corresponds to the audio segment to be recorded, receive the recorded audio for the audio segment to be recorded; The recorded audio is used to update the audio segment to be recorded in order to obtain video material in the background sound dimension.

6. The method according to claim 1, characterized in that, After generating the recommended videos corresponding to the target book list, the method further includes: Play the recommended video on the video recommendation page, and mark the video segments in the recommended video that are associated with each book in the target book list in the progress bar of the recommended video; In response to the selection of any of the video segments, the recommended video is played from the starting position of the recommended video segment, and reading information of the book associated with the video segment is displayed.

7. A video generation apparatus, characterized in that, include: The display module is used to display a video creation page in response to generating recommended videos for the target book list. The video creation page includes editing instructions corresponding to multiple video composition dimensions, which are used to guide the addition of target video materials under each video composition dimension. The generation module is used to respond to the received target video materials under each of the video composition dimensions, integrate the target video materials using the target video script, and generate the recommended videos corresponding to the target book list; The target video script indicates the target video frame associated with each of the video composition dimensions in the recommended video, and the display position of the target video material corresponding to the video composition dimension in the target video frame; Wherein, after displaying the editing instruction information on the video creation page, the display module is further used for: In response to receiving input information under any editing instruction, based on the video composition dimension corresponding to the any editing instruction, determine multiple book description dimensions corresponding to the target book list, and determine the target book description dimension from the multiple book description dimensions based on the input information; The book information under the target book description dimension is obtained from the target book list, and the book information is appended to the input information for supplementary display.

8. A computer device, characterized in that, include: A processor and a memory, the memory storing machine-readable instructions executable by the processor, the processor executing the machine-readable instructions stored in the memory, wherein when the machine-readable instructions are executed by the processor, the processor performs the steps of the video generation method as described in any one of claims 1 to 6.

9. A computer-readable storage medium, characterized in that, The computer-readable storage medium stores a computer program, which, when executed by a computer device, performs the steps of the video generation method as described in any one of claims 1 to 6.