Video note generation method and device and electronic equipment
By setting content generation controls in the video playback window and combining video playback content generation, the problem of users needing to frequently switch video players and note-taking tools is solved, and deep integration of video playback and note-taking processing and efficient note-making are achieved.
Patent Information
- Application Number
- CN202510280607.8
- Authority / Receiving Office
- CN · China
- Patent Type
- Applications(China)
- Current Assignee / Owner
- Filing Date
- 2025-03-10
- Publication Date
- 2025-06-13
AI Technical Summary
In the prior art, users need to frequently switch video players and note-taking tools when watching videos, resulting in low note generation efficiency and lack of deep integration between video players and note-taking tools.
By setting the video playback area and note processing area in the video playback window, and setting content generation controls in the note processing area, the user can combine the video to generate content to obtain the content to be processed according to the operation when the content generation control is detected.
It realizes the deep integration of video playback and note processing, eliminates software isolation between video player and note-taking tool, and improves note-generating efficiency, so that users do not need to frequently switch video player and note-taking tool.
Smart Images

Figure CN120151576A_ABST
Abstract
Description
Technical Field
[0001] The present disclosure relates to the field of artificial intelligence technology, and particularly to the fields of intelligent network disks, video processing, deep learning, natural language processing, computer vision, large models, etc. In particular, it relates to a method, device and electronic device for generating video notes. Background Art
[0002] Currently, when a user is using a video player to watch a video and needs to take notes, they need to rely on a note-taking tool to edit notes in the note-taking tool. However, there is software isolation between the video player and the note-taking tool, which causes the user to need to frequently switch between the video player and the note-taking tool, resulting in low note generation efficiency. Summary of the Invention
[0003] The present disclosure provides a method, device and electronic device for generating video notes.
[0004] According to one aspect of the present disclosure, there is provided a method for generating video notes, the method comprising: displaying a video playback window; a video playback area and a note processing area are provided in the video playback window; at least one content generation control is provided in the note processing area; in the case of detecting a selection operation on a first content generation control among at least one of the content generation controls, performing content generation processing on the video played in the video playback window in combination with the first content generation control to obtain content to be processed; determining a video note corresponding to the video according to an operation on the content to be processed and the content to be processed.
[0005] According to another aspect of the present disclosure, there is provided a device for generating video notes, the device comprising: a window display module for displaying a video playback window; a video playback area and a note processing area are provided in the video playback window; at least one content generation control is provided in the note processing area; a content generation module for performing content generation processing on the video played in the video playback window in combination with the first content generation control in the case of detecting a selection operation on a first content generation control among at least one of the content generation controls to obtain content to be processed; a note determination module for determining a video note corresponding to the video according to an operation on the content to be processed and the content to be processed.
[0006] According to another aspect of the present disclosure, there is provided an electronic device, comprising: at least one processor; and a memory communicatively connected to the at least one processor; wherein, the memory stores instructions executable by the at least one processor, and the instructions are executed by the at least one processor so that the at least one processor can execute the method for generating video notes proposed above in the present disclosure.
[0007] According to another aspect of the present disclosure, there is provided a non-transitory computer-readable storage medium storing computer instructions for causing a computer to execute the method for generating video notes proposed above in the present disclosure.
[0008] According to another aspect of the present disclosure, there is provided a computer program product including a computer program which, when executed by a processor, implements the steps of the method for generating video notes proposed above in the present disclosure.
[0009] It should be understood that the content described in this part is not intended to identify the key or important features of the embodiments of the present disclosure, nor is it used to limit the scope of the present disclosure. Other features of the present disclosure will become easily understood through the following description. Description of the Drawings
[0010] The drawings are used to better understand the solution and do not constitute a limitation to the present disclosure. Among them:
[0011] Figure 1 is a schematic diagram according to the first embodiment of the present disclosure;
[0012] Figure 2A is a schematic diagram according to the second embodiment of the present disclosure;
[0013] Figure 2B is a schematic diagram of a content generation control;
[0014] Figure 2C is a schematic diagram of a processing control;
[0015] Figure 2D is another schematic diagram of a processing control;
[0016] Figure 3 is a schematic diagram according to the third embodiment of the present disclosure;
[0017] Figure 4 is a schematic diagram according to the fourth embodiment of the present disclosure;
[0018] Figure 5 is a schematic diagram according to the fifth embodiment of the present disclosure;
[0019] Figure 6 is a block diagram of an electronic device for implementing the method for generating video notes of the embodiments of the present disclosure. Detailed Embodiments
[0020] The following describes exemplary embodiments of the present disclosure with reference to the accompanying drawings. Various details of the embodiments of the present disclosure are included to assist understanding, and they should be considered merely exemplary. Therefore, those of ordinary skill in the art should recognize that various changes and modifications can be made to the embodiments described herein without departing from the scope and spirit of the present disclosure. Similarly, descriptions of well-known functions and structures are omitted below for clarity and conciseness.
[0021] In the related art, some artificial intelligence (AI) note-taking tools can assist in generating note content through AI technology. For example, some AI note-taking tools support users in generating AI-assisted notes by integrating multiple AI models, and can handle various file formats (such as videos, office (Office) documents, portable document format (PDF), electronic publication (EPUB), etc.), and provide an intelligent organization function to convert fragmented information into structured notes. Some other AI note-taking tools provide the function of generating mind maps by AI to help users quickly sort out the knowledge structure. Specifically, these tools combine natural language processing (NLP) technology to parse the content input by users and generate a visual knowledge structure diagram. There are also some AI note-taking tools that automatically generate quiz questions through AI technology to help users consolidate the learning content. Specifically, these tools can dynamically adjust the question difficulty according to the user's learning progress and content to improve the user's learning effect.
[0022] However, the existing AI note-taking tools have low function integration and relatively scattered functions, and lack deep integration with video players. When users watch videos, they need to switch to the AI note-taking tools to generate AI notes, and cannot directly generate relevant AI notes in the video player.
[0023] In view of the above problems, the present disclosure proposes a method, device, and electronic device for generating video notes.
[0024] Figure 1 It is a schematic diagram according to the first embodiment of the present disclosure. It should be noted that the method for generating video notes in the embodiments of the present disclosure can be applied to a device for generating video notes, and the device can be configured in an electronic device so that the electronic device can execute the function of generating video notes.
[0025] Among them, the electronic device can be any device with computing capabilities, such as a personal computer (PC for short), a mobile terminal, a server, etc. The mobile terminal can be, for example, a vehicle-mounted device, a mobile phone, a tablet computer, a personal digital assistant, a wearable device, a smart speaker, a server, a server cluster, and other hardware devices with various operating systems, touch screens, and / or display screens.
[0026] Among them, the video note generation device can also be software in the electronic device, such as video note generation software, etc. In the following embodiments, the execution subject is taken as an example of an electronic device for description.
[0027] As Figure 1 shown, the video note generation method can include the following steps:
[0028] Step 101, display a video playback window; the video playback window is provided with a video playback area and a note processing area; at least one content generation control is provided in the note processing area.
[0029] In the embodiments of the present disclosure, the electronic device can display a video playback window, which is a graphical user interface element that can be used to display a video; the video playback area is used to play and display the video, and the user can watch the video in the video playback area and perform basic playback controls, such as play, pause, etc.; the note processing area is used to process notes related to the video, and the user can generate note content through the content generation control in the note processing area.
[0030] Among them, the video playback area and the note processing area can refer to two adjacent areas or two non-adjacent areas. It should be noted that the present disclosure does not limit the area sizes of the video playback area and the note processing area. The area sizes of the video playback area and the note processing area can be the same or different. As an example, in order not to affect the user's video viewing experience, the area size of the video playback area can be larger than that of the note processing area.
[0031] In the embodiments of the present disclosure, the content generation control is a clickable control in the note processing area, which is used to trigger the content generation function; after the user clicks the content generation control, the electronic device can automatically generate note content related to the video content.
[0032] Among them, the note content generated by different content generation controls is different. For example, the content generation control can include a text extraction control, a speech recognition control, a summary extraction control, etc., and the corresponding note content can include text, video summary, and other contents.
[0033] Step 102, in the case of detecting a selection operation on a first content generation control among at least one content generation control, perform content generation processing on the video played in the video playback window in combination with the first content generation control to obtain the content to be processed.
[0034] In the embodiments of the present disclosure, the first content generation control may refer to any one of at least one content generation control; the electronic device may detect a selection operation on any content generation control, and perform content generation processing on the video played in the video playback window in combination with the selected first content generation control to obtain the corresponding content to be processed.
[0035] Among them, the insertion position of the content to be processed may be determined according to the position of the cursor in the note processing area. As an example, the content to be processed may be automatically inserted at the position where the cursor is located according to the cursor position indicated by the user; if the cursor loses focus, it may be inserted at the previous cursor position; if there is information in the line where the cursor is located, it will automatically wrap, and if there is no information in the line where the cursor is located, there is no need to wrap.
[0036] Among them, a content generation large model matching the first content generation control may be adopted to perform content generation processing on the video played in the video playback window in combination with the first content generation control to obtain the content to be processed.
[0037] It should be noted that the content generation large models corresponding to each content generation control may be different large models or the same large model.
[0038] In some embodiments, after obtaining the content to be processed, the content to be processed may be displayed in an editing component in the note processing area of the video playback window. Among them, the editing component may be in the form of a card, a pop-up window, etc.
[0039] Among them, the insertion position of the editing component in the note processing area may be determined according to the position of the cursor. Step 103, determine the video note corresponding to the video according to the operation on the content to be processed and the content to be processed.
[0040] In the embodiments of the present disclosure, the operations on the content to be processed may include operations such as moving the position of the content to be processed, adjusting the font size / color of the content to be processed, marking some or all of the content in the content to be processed, etc. The electronic device may combine the operations on the content to be processed and generate the corresponding video note according to the content to be processed.
[0041] The method for generating video notes according to the embodiments of the present disclosure includes displaying a video playback window; a video playback area and a note processing area are provided in the video playback window; at least one content generation control is provided in the note processing area; when a selection operation for a first content generation control among at least one content generation control is detected, content generation processing is performed on the video being played in the video playback window by combining the first content generation control to obtain content to be processed; a video note corresponding to the video is determined according to the operation on the content to be processed and the content to be processed; wherein, by setting the content generation control, the video playback and note processing functions are integrated into one, eliminating the software isolation between the video player and the note tool, and the user does not need to frequently switch between the video player and the note tool, improving the note generation efficiency.
[0042] Among them, in order to meet the diverse needs of users for notes, the electronic device can display at least one note template, and generate a corresponding video note based on the selected note template. As Figure 2A shown, Figure 2A is a schematic diagram according to the second embodiment of the present disclosure, Figure 2A The embodiment shown may include the following steps:
[0043] Step 201, display a video playback window; a video playback area and a note processing area are provided in the video playback window; at least one content generation control is provided in the note processing area.
[0044] In the embodiments of the present disclosure, in order to meet the diverse note processing needs of users, the content generation control may include at least one of the following: a note generation control, a mind map generation control, and an exercise generation control. As Figure 2B shown, Figure 2B is a schematic diagram of the content generation control. The AI note control in the figure is the note generation control, the AI mind map control is the mind map generation control, and the AI question generation control is the exercise generation control.
[0045] Among them, the note generation control is used to generate AI notes, the mind map generation control is used to generate mind maps, and the exercise generation control is used to generate exercises.
[0046] Step 202, when no video note is displayed in the note processing area and no selection operation for any content generation control is detected, display at least one note template; one of the at least one note templates is in a selected state.
[0047] In the embodiments of the present disclosure, at least one note template may be displayed in the note processing area. Among them, the user can preview the note template.
[0048] Among them, when the note processing area is empty and the user does not actively trigger content generation, at least one note template is intelligently presented and one of them is automatically selected, which can guide the user to start taking notes in an intuitive and convenient way.
[0049] To achieve the diversity of note templates and meet the diverse needs of users for note templates, the note templates can include at least one of the following: text note template; image note template; outline note template; text-image note template.
[0050] Among them, the outline note template is used to generate the content to be processed with text, images, and mind maps set, the text note template is used to generate the content to be processed with text set, the image note template is used to generate the content to be processed with images set, and the text-image note template is used to generate the content to be processed with images and text set.
[0051] In the embodiments of the present disclosure, the note template in the selected state may refer to the note template default selected by the electronic device; for example, the note template default selected by the electronic device may refer to the text-image note template.
[0052] Step 203, in the case of detecting a selection operation on the first note template that is not in the selected state among at least one note template, switch the note template in the selected state to the first note template.
[0053] In the embodiments of the present disclosure, the first note template that is not in the selected state may refer to any note template other than the note template in the selected state.
[0054] Among them, through the detection of the selection operation, the content generation requirements of the user can be quickly identified, and then the note template can be switched in a timely manner.
[0055] Step 204, in the case of detecting a selection operation on the first content generation control among at least one content generation control and the first content generation control is a note generation control, obtain the note template in the selected state among at least one note template; perform content generation processing in combination with the video and the note template in the selected state to obtain the content to be processed.
[0056] In the embodiments of the present disclosure, in the case of template switching, the note template in the selected state among at least one note template may refer to the switched first note template; in the case of no template switching, the note template in the selected state among at least one note template may refer to the default selected note template.
[0057] Among them, the note template provides a preset note structure and format. Generating the content to be processed in combination with the selected note template makes the generation of the content to be processed more convenient and efficient, and improves the generation efficiency of the content to be processed.
[0058] Step 205: Determine the video note corresponding to the video according to the operation on the content to be processed and the content to be processed.
[0059] In the embodiments of the present disclosure, if the first content generation control is a note generation control, determining the video note corresponding to the video according to the operation on the content to be processed and the content to be processed may include the following situations:
[0060] (1) When the operation on the content to be processed is an insertion operation, determine the content to be processed as the content in the video note;
[0061] (2) When the operation on the content to be processed is a discard operation, discard determining the content to be processed as the content in the video note;
[0062] (3) When the operation on the content to be processed is a template change operation, discard determining the content to be processed as the content in the video note, and display at least one note template for re-selecting the note template and regenerating the content to be processed.
[0063] Among them, the user can perform flexible operations such as insertion, discard, or template change on the content to be processed according to actual needs. This design enables the user to easily manage the content to be processed, thereby improving the user's operation experience.
[0064] It should be noted that when the operation on the content to be processed is an insertion operation, after inserting the content to be processed into the note, the content to be processed may disappear; when the operation on the content to be processed is a discard operation, discarding the content to be processed may mean deleting the content to be processed.
[0065] In the embodiments of the present disclosure, for the note generation control, at least one of the following processing controls is set for the content to be processed in the note processing area: discard control, copy control, insertion control, and template change control. As Figure 2C shown, Figure 2C is a schematic diagram of the processing control.
[0066] Among them, when the discard control is selected, it is determined that a discard operation on the content to be processed is detected; when the copy control is selected, it is determined that an insertion operation on the content to be processed is detected, and an operation of copying the content in the content to be processed to the clipboard is detected; when the insertion control is selected, it is determined that an insertion operation on the content to be processed is detected; when the template change control is selected, it is determined that a template change operation on the content to be processed is detected.
[0067] In the embodiments of the present disclosure, if the first content generation control is a mind map generation control or an exercise generation control, determining the video note corresponding to the video according to the operation on the content to be processed and the content to be processed may include the following situations:
[0068] (1) When the operation on the content to be processed is an insertion operation, determine the content to be processed as the content in the video note;
[0069] (2) When the operation on the content to be processed is a discard operation, discard the determination of the content to be processed as the content in the video note;
[0070] (3) When the operation on the content to be processed is a regeneration operation, discard the determination of the content to be processed as the content in the video note, and regenerate the content to be processed by recombining the first content generation control.
[0071] Among them, the user can perform flexible operations such as insertion, discard, or regeneration on the content to be processed according to actual needs. This design enables the user to easily manage the content to be processed, thereby enhancing the user's operation experience.
[0072] In the embodiments of the present disclosure, for the mind map generation control or the exercise generation control, at least one of the following processing controls is set for the content to be processed in the note processing area: discard control, copy control, insertion control, and regeneration control. As Figure 2D shown, Figure 2D FIG. is another schematic diagram of the processing control, where the content included in the rectangular frame in the note processing area in the figure is the content to be processed displayed in the card-shaped editing component.
[0073] Among them, when the discard control is selected, it is determined that a discard operation on the content to be processed is detected; when the copy control is selected, it is determined that an insertion operation on the content to be processed is detected, and a copy-to-clipboard operation on the content to be processed is detected; when the insertion control is selected, it is determined that an insertion operation on the content to be processed is detected; when the regeneration control is selected, it is determined that a regeneration operation on the content to be processed is detected.
[0074] In the embodiments of the present disclosure, the video note includes at least one of the following: mind map, exercise, screenshot image of the video frame in the video, relevant text content corresponding to the video frame in the video; the relevant text content is provided with a timestamp; the timestamp indicates the position of the video frame corresponding to the relevant text content in the video.
[0075] Among them, the video note can integrate content such as mind maps, exercises, screenshot images of video frames, and relevant text content of video frames, with rich forms and high flexibility.
[0076] Among them, clicking on the timestamp can jump / to locate to the position of the video frame corresponding to the relevant text content in the video. Through the timestamp, the video frame can be accurately located, which is convenient for the user to quickly review key information, enhance the learning pertinence, and make the note more practical and efficient.
[0077] Among them, when processing notes in the note processing area, after generating corresponding content to be processed based on the selected content and inserting the content to be processed into the note, the user can select content to generate a control again to generate new content to be processed based on the content generated by the control selected again. Finally, based on all the generated content to be processed, the video note is determined.
[0078] For example, the user first selects a note generation control and a graphic note template. After generating the content to be processed based on the graphic note template and inserting the graphic content (screenshot images of video frames and relevant text content corresponding to the video frames) in the content to be processed into the note, the user can select the content generation control again. For example, if the exercise generation control is selected, the corresponding content to be processed can be continuously generated in combination with the exercise generation control, and the exercise in the content to be processed can be inserted into the note. In this case, the video note includes screenshot images of video frames, relevant text content corresponding to the video frames, and exercises.
[0079] It should be noted that the exercise generation strategy corresponding to the exercise generation control can be preset, or the exercise generation strategy corresponding to the exercise generation control can be modified through interaction. Among them, the exercise generation strategy corresponding to the exercise generation control is used to indicate the total number of generated exercises, the types of exercises (such as single-choice / multiple-choice, multiple-choice questions / fill-in-the-blank questions, etc.), the proportion of the number of each type of exercise, and the difficulty of the exercises.
[0080] The method for generating a video note according to an embodiment of the present disclosure includes displaying a video playback window; a video playback area and a note processing area are provided in the video playback window; at least one content generation control is provided in the note processing area; when no video note is displayed in the note processing area and no selection operation for any content generation control is detected, at least one note template is displayed; one of the at least one note template is in a selected state; when a selection operation for a first note template that is not in the selected state among the at least one note template is detected, the note template in the selected state is switched to the first note template; when a selection operation for a first content generation control among the at least one content generation control is detected and the first content generation control is a note generation control, the note template in the selected state among the at least one note template is obtained; content generation processing is performed by combining the video and the note template in the selected state to obtain content to be processed; the video note corresponding to the video is determined according to the operation on the content to be processed and the content to be processed; among them, the content generation requirements of the user can be quickly identified through the selection operation detection, and then the note template can be switched in time; in addition, the note template provides a preset note structure and format, and combining the selected note template to generate the content to be processed makes the generation of the content to be processed more convenient and efficient, and improves the generation efficiency of the content to be processed.
[0081] Among them, in order to avoid the storage resources being occupied by useless content to be processed, when no operation on the content to be processed is detected and a selection operation on the second content generation control among at least one content generation control is detected, the content to be processed is discarded. As Figure 3 shown, Figure 3 is a schematic diagram according to the third embodiment of the present disclosure, Figure 3 The shown embodiment may include the following steps:
[0082] Step 301, display a video playback window; a video playback area and a note processing area are set in the video playback window; at least one content generation control is set in the note processing area.
[0083] Step 302, when a selection operation on the first content generation control among at least one content generation control is detected, perform content generation processing on the video played in the video playback window in combination with the first content generation control to obtain the content to be processed.
[0084] In the embodiments of the present disclosure, before performing content generation processing on the video played in the video playback window in combination with the first content generation control to obtain the content to be processed, it further includes: determining a first playback object corresponding to the video playback window, as well as a shared video note shared by a second playback object to the first playback object and a shared video corresponding to the shared video note; when the shared video includes a video, obtaining the shared video note corresponding to the video; and displaying the shared video note corresponding to the video in the note processing area.
[0085] Among them, the first playback object can easily obtain the video note shared by the second playback object and directly display the shared video note in the note processing area. This design not only realizes the convenient sharing and viewing of video notes, but also enhances the interactivity, making the knowledge transfer more efficient and intuitive.
[0086] It should be noted that the user can, based on the shared video note, obtain the content to be processed through the content generation control, and then adjust the note based on the content to be processed.
[0087] In the embodiments of the present disclosure, before performing content generation processing on the video played in the video playback window in combination with the first content generation control to obtain the content to be processed, it further includes: if no video note is displayed in the note processing area, determining whether the video played in the video playback window is a popular video and whether the popular video has prefabricated AI notes; when the video played in the video playback window is a popular video and the popular video has prefabricated AI notes, displaying the prefabricated AI note corresponding to the video in the note processing area; when the video played in the video playback window is not a popular video, or the popular video does not have prefabricated AI notes, displaying at least one note template.
[0088] It should be noted that for the same popular video, the prefabricated AI notes are the same.
[0089] In the embodiments of the present disclosure, before generating processed content for the video played in the video playback window by combining the first content generation control, the notes displayed in the note processing area can also be the notes previously edited by the user.
[0090] Step 303, in the case where no operation on the processed content is detected and a selection operation on the second content generation control among at least one content generation control is detected, discard the processed content.
[0091] In the embodiments of the present disclosure, after obtaining the processed content, if the user does not perform operations such as insertion or abandonment on the processed content, but instead reselects a second content generation control different from the first content generation control, in this case, the previously obtained processed content can be discarded to avoid the occupied storage resources by useless processed content.
[0092] Step 304, generate processed content for the video played in the video playback window by combining the second content generation control.
[0093] Among them, in the case where the user does not operate on the already generated processed content but selects the second content generation control, the already generated processed content is automatically discarded, and the processed content is regenerated according to the second content generation control, which not only ensures the timeliness of the processed content but also avoids the occupied storage resources by useless processed content.
[0094] Among them, the content of generating processed content for the video played in the video playback window by combining the second content generation control can refer to Figure 1 the relevant description of step 102 in the embodiments shown, and no detailed description will be given here.
[0095] Step 305, determine the video note corresponding to the video according to the operation on the processed content and the processed content.
[0096] In the embodiments of the present disclosure, the video and the corresponding video note of the video can be shared with a second playback object; the second playback object is different from the first playback object corresponding to the video playback window.
[0097] Among them, the video note can refer to the note after adjusting the shared video note, or can refer to a new video note.
[0098] Among them, this design realizes the convenient sharing function of the video and its corresponding video note, which not only improves the convenience of note acquisition but also promotes the interaction and communication among users.
[0099] Among them, it should be noted that for the detailed content of steps 301 to 302, reference can be made to Figure 1 steps 101 to 102 in the illustrated embodiment, and no detailed description will be given here.
[0100] The method for generating a video note according to an embodiment of the present disclosure includes: displaying a video playback window; a video playback area and a note processing area are set in the video playback window; at least one content generation control is set in the note processing area; when a selection operation for a first content generation control among the at least one content generation control is detected, content generation processing is performed on the video being played in the video playback window in combination with the first content generation control to obtain content to be processed; when no operation for the content to be processed is detected and a selection operation for a second content generation control among the at least one content generation control is detected, the content to be processed is discarded; content generation processing is performed on the video being played in the video playback window in combination with the second content generation control to obtain content to be processed; a video note corresponding to the video is determined according to the operation for the content to be processed and the content to be processed; among them, when no operation for the content to be processed is detected and a selection operation for a second content generation control among the at least one content generation control is detected, discarding the content to be processed can avoid useless content to be processed from occupying storage resources.
[0101] Among them, in order to further improve the quality of the video note and enhance the user experience, after obtaining the video note, the content to be processed in the video note can be processed according to a video note processing request to obtain a processed video note. As Figure 4 shown, Figure 4 is a schematic diagram according to the fourth embodiment of the present disclosure, Figure 4 the illustrated embodiment may include the following steps:
[0102] Step 401: Display a video playback window; a video playback area and a note processing area are set in the video playback window.
[0103] Step 402: When a selection operation for a first content generation control among the at least one content generation control is detected, content generation processing is performed on the video being played in the video playback window in combination with the first content generation control to obtain content to be processed.
[0104] In an embodiment of the present disclosure, when a close operation for the video playback window is detected and there is content to be processed displayed in the note processing area, the video playback window is closed and the content to be processed is stored; when a reopen operation for the video playback window is detected, the video playback window is displayed, and the content to be processed is displayed in the note processing area in the video playback window.
[0105] Among them, the method of storing the content to be processed when closing the video playback window ensures that when the user closes the video playback window, the content to be processed in the note processing area is not lost but is automatically stored. When the user re-opens the video playback window, the content to be processed will be displayed in the note processing area again, facilitating the user to continue editing or viewing. This not only improves the continuity and convenience of note processing but also enhances the user experience and avoids content loss caused by the accidental closing of the video playback window.
[0106] It should be noted that when closing the video playback window, the note content that has been generated in the note processing area can also be stored, and when the video playback window is re-opened, the previously generated note content can be displayed.
[0107] In the embodiment of the present disclosure, during the process of performing content generation processing on the video played in the video playback window in combination with the first content generation control, a text prompt during generation is displayed; and / or, in the case of content generation failure, a text indicating generation failure is displayed.
[0108] Among them, the text prompt during generation can be displayed within the editing component in the note processing area, and / or the text indicating generation failure can be displayed within the editing component in the note processing area. For example, the editing component can be in the form of a card, a pop-up window, etc.
[0109] Among them, during the process of performing content generation processing, displaying the text prompt during generation or the text indicating generation failure enables the user to more clearly and intuitively grasp the content generation status and enhances the user experience.
[0110] Among them, the text prompt during generation and the text indicating generation failure can be displayed in the note processing area according to the position of the cursor; or, the text prompt during generation and the text indicating generation failure can be displayed at any position in the note processing area.
[0111] Among them, the following at least one processing control is set for the text prompt during generation in the note processing area: a termination control (after clicking, the text prompt during generation directly disappears and the task is cancelled); the following at least one processing control is set for the text indicating generation failure in the note processing area: a close control (after clicking, the text indicating generation failure disappears), a retry control (changes back to the text prompt during generation and regenerates).
[0112] In an embodiment of the present disclosure, when a close operation on the video playback window is detected and there is generating prompt text displayed in the note processing area, the video playback window is closed, and the content generation process for the video played in the video playback window is continued in combination with the first content generation control to obtain the content to be processed; when a reopen operation on the video playback window is detected and the content generation process is completed, the video playback window is displayed, and the content to be processed is displayed in the note processing area in the video playback window.
[0113] Among them, this design ensures that during the content generation process, even if the video playback window is closed, the content generation process will not be interrupted but will continue in the background. Once the content generation is completed, that is, the content to be processed is generated. When the user reopens the video playback window, the generated content to be processed can be seen displayed in the note processing area, ensuring the continuity of content generation and optimizing the user experience.
[0114] Step 403: Determine the video note corresponding to the video according to the operation on the content to be processed and the content to be processed.
[0115] Step 404: Receive a video note processing request; the video note processing request includes the content to be processed in the video note and the processing type for the content to be processed; process the content to be processed according to the processing type to obtain the processed video note.
[0116] In an embodiment of the present disclosure, the content to be processed may refer to the note content to be optimized or adjusted in the video note; the processing type of the content to be processed is the optimization type or adjustment type of the content to be processed.
[0117] Among them, processing the content to be processed according to the processing type can make the processed video note more in line with the user's needs and enhance the user's satisfaction.
[0118] In some embodiments, the processing type of the content to be processed may include at least one of types such as sentence smoothness adjustment, text style adjustment, and image clarity adjustment.
[0119] In some embodiments, the processing type includes at least one of the following: polishing type, expansion type, abbreviation type, and continuation type.
[0120] Processing the content to be processed according to the processing type in the video note processing request, such as polishing, expanding, abbreviating, or continuing, can optimize the video note, make the note content more rich, refined, or coherent, meet the diverse note processing needs of users, enhance the user's experience, and improve the quality of the video note.
[0121] Among them, it should be noted that for the detailed content of steps 401 to 403, reference can be made to Figure 1 steps 101 to 103 in the illustrated embodiment, and no detailed description will be given here.
[0122] The method for generating video notes according to the embodiments of the present disclosure includes: displaying a video playback window; a video playback area and a note processing area are set in the video playback window; when a selection operation on a first content generation control among at least one content generation control is detected, content generation processing is performed on the video played in the video playback window in combination with the first content generation control to obtain content to be processed; determining a video note corresponding to the video according to the operation on the content to be processed and the content to be processed; receiving a video note processing request; the video note processing request includes the content to be processed in the video note and the processing type for the content to be processed; processing the content to be processed according to the processing type to obtain a processed video note; wherein, processing the content to be processed according to the processing type can make the processed video note more in line with user needs and enhance user satisfaction.
[0123] To clearly illustrate the solution of the present disclosure, the following will be described in conjunction with a note generation control, a mind map generation control, and an exercise generation control:
[0124] Among them, assuming that the content to be processed is displayed in an editing component in the form of a card, for the convenience of understanding, hereinafter, the editing component obtained by clicking the note generation control and including the content to be processed (note) is called an AI note content card, the editing component obtained by clicking the mind map generation control and including the content to be processed (mind map) is called a mind map content card, the editing component obtained by clicking the exercise generation control and including the content to be processed (exercise) is called an exercise content card, the editing component including the text prompt during generation is called a generating card, and the editing component including the text indicating generation failure is called a generation failure card.
[0125] 1. Note generation control:
[0126] When a selection operation on the note generation control is detected, it is judged whether there is an AI note content card, a generating card, or a generation failure card in the video playback window (or the note processing area); if there is an AI note content card, a generating card, or a generation failure card, a toast prompt can be given: there is an existing card, please process the card before using again; if there is no AI note content card, a generating card, or a generation failure card, content generation processing is performed on the video played in the video playback window in combination with the note generation control, and the generating card is displayed, and by default, content generation is performed based on a graphic note template. Among them, a termination control corresponds to the generating card. After clicking, the generating card directly disappears and the task is cancelled.
[0127] After the content generation process is completed, insert the AI note content card at the position of the cursor; if the cursor loses focus, insert it at the previous cursor position. If there is information on the line where the cursor is located, automatically wrap the line; if there is no information on the line where the cursor is located, there is no need to wrap the line. Among them, the AI note content card corresponds to a discard control (delete the card after clicking), a copy control (after clicking, the card content is inserted into the note and the card disappears, and at the same time, the card content is copied to the clipboard, and a toast prompt: the content has been copied to the clipboard), an insert control (after clicking, the card content is inserted into the note and the card disappears), and a change template control (after clicking, a template selection pop-up window is launched). Among them, the user can select the note template and preview the template content in the template selection pop-up window. The template selection pop-up window also includes a generate note control. Click the generate note control, and below the generate note control, there is a small text prompt: It is expected to take 5 to 10 minutes.
[0128] If the user closes the video playback window when there is an AI note content card, a generating card, or a generation failure card displayed in the video playback window, then when the video playback window is reopened, the AI note content card, the generating card, or the generation failure card is still displayed in the video playback window. Among them, if the generating card was displayed before closing and the content generation process has been completed when restarting, then the obtained AI note content card is displayed in the reopened video playback window. It should be noted that the generating AI note content card does not support operations such as copying and cutting.
[0129] In addition, if no video note is displayed in the note processing area (empty), then determine whether the video played in the video playback window is a popular video and whether the popular video has a prefabricated AI note; when the video played in the video playback window is a popular video and the popular video has a prefabricated AI note, display the prefabricated AI note corresponding to the video in the note processing area; when the video played in the video playback window is not a popular video, or the popular video does not have a prefabricated AI note, display at least one note template.
[0130] 2. Mind Map Generation Control
[0131] When a selection operation on the mind map generation control is detected, it is judged whether there is a mind map content card, a generating card or a generation failure card in the video playback window (or note processing area); if there is a mind map content card, a generating card or a generation failure card, a toast prompt can be given: There is already a card. Please process the card before using it; if there is no mind map content card, generating card or generation failure card, the content generation process is performed on the video played in the video playback window in combination with the note generation control, and the generating card is displayed. Among them, the generating card corresponds to a termination control. After clicking, the generating card disappears directly and the task is cancelled; the generation failure card corresponds to a close control (the card disappears after clicking) and a retry control (after clicking, it changes back to the generating card and regenerates the mind map). Among them, if the user clicks the mind map generation control again when there is an exercise content card, a generating card or a generation failure card in the video playback window (or note processing area), it can jump to the corresponding card, and a toast prompt is given: There is already a card. Please process the card before using it.
[0132] After the content generation process is completed, the mind map content card is inserted at the position of the cursor; if the cursor loses focus, it is inserted at the previous cursor position. If there is information in the line where the cursor is located, it will automatically wrap; if there is no information in the line where the cursor is located, there is no need to wrap. Among them, the mind map content card corresponds to a discard control (the card is deleted after clicking), a copy control (after clicking, the card content is inserted into the note and the card disappears, and at the same time the card content is copied to the clipboard, and a toast prompt is given: The content has been copied to the clipboard), an insert control (after clicking, the card content is inserted into the note and the card disappears), and a regenerate control (after clicking, it changes back to the generating card and regenerates the mind map).
[0133] If the user closes the video playback window when there is a mind map content card, a generating card or a generation failure card displayed in the video playback window, when the video playback window is reopened, the mind map content card, generating card or generation failure card is still displayed in the video playback window. Among them, if the generating card is displayed before closing and the content generation process has been completed when restarting, the obtained mind map content card is displayed in the reopened video playback window. It should be noted that the generating mind map content card does not support operations such as copying and cutting.
[0134] 3. Exercise generation control
[0135] When a selection operation on the exercise generation control is detected, it is determined whether there is an exercise content card, a generating card, or a generation failure card in the video playback window (or the note processing area); if there is an exercise content card, a generating card, or a generation failure card, a toast prompt can be given: there is already a card, please process the card before using; if there is no exercise content card, a generating card, or a generation failure card, the content generation process is performed on the video played in the video playback window in combination with the note generation control, and the generating card is displayed. Among them, the generating card corresponds to a termination control, and after clicking, the generating card directly disappears and the task is cancelled; the generation failure card corresponds to a close control (the card disappears after clicking) and a retry control (changes back to the generating card and regenerates the exercise).
[0136] After the content generation process is completed, the exercise content card is inserted at the position of the cursor; if the cursor loses focus, it is inserted at the previous cursor position. If there is information in the line where the cursor is located, it will automatically wrap; if there is no information in the line where the cursor is located, there is no need to wrap. Among them, the exercise content card corresponds to a discard control (deletes the card after clicking), a copy control (the card content is inserted into the note and the card disappears after clicking, and at the same time the card content is copied to the clipboard, and a toast prompt is given: the content has been copied to the clipboard), an insert control (the card content is inserted into the note and the card disappears after clicking), and a regenerate control (changes back to the generating card and regenerates the exercise after clicking).
[0137] If the user closes the video playback window when there is an exercise content card, a generating card, or a generation failure card displayed in the video playback window, when the video playback window is reopened, the exercise content card, the generating card, or the generation failure card is still displayed in the video playback window. Among them, if the generating card is displayed before closing and the content generation process has been completed when restarting, the obtained exercise content card is displayed in the reopened video playback window. It should be noted that the generating exercise content card does not support operations such as copying and cutting.
[0138] Among them, the exercise generation control corresponds to an exercise generation strategy. The exercise generation strategy can be combined to automatically identify knowledge points and randomly generate exercises according to the entire content of the video. For example, 3 questions are generated, and the question types are divided into 2 single-choice questions and 1 multiple-choice question. Among them, when the user clicks the exercise generation control again, the generated exercises are different from the previous ones.
[0139] To implement the above embodiments, the present disclosure also provides a video note generation device. As Figure 5 shown, Figure 5 is a schematic diagram according to the fifth embodiment of the present disclosure. The video note generation device 50 may include: a window display module 501, a content generation module 502, and a note determination module 503.
[0140] Among them, the window display module 501 is used to display a video playback window; a video playback area and a note processing area are set in the video playback window; at least one content generation control is set in the note processing area; the content generation module 502 is used to, when detecting a selection operation on the first content generation control among at least one content generation control, perform content generation processing on the video played in the video playback window in combination with the first content generation control to obtain the content to be processed; the note determination module 503 is used to determine the video note corresponding to the video according to the operation on the content to be processed and the content to be processed.
[0141] As a possible implementation manner of the embodiment of the present disclosure, the content generation control includes at least one of the following: a note generation control, a mind map generation control, and an exercise generation control.
[0142] As a possible implementation manner of the embodiment of the present disclosure, the content generation control includes a note generation control; the device further includes: a template display module, which is used to display at least one note template when no video note is displayed in the note processing area and no selection operation on any content generation control is detected; one of the at least one note templates is in a selected state; the template switching module is used to, when detecting a selection operation on the first note template that is not in the selected state among the at least one note templates, switch the note template in the selected state to the first note template.
[0143] As a possible implementation manner of the embodiment of the present disclosure, the note template includes at least one of the following: a text note template; an image note template; an outline note template; a graphic and text note template; the outline note template is used to generate the content to be processed provided with text, images, and mind maps.
[0144] As a possible implementation manner of the embodiment of the present disclosure, the first content generation control is a note generation control; the content generation module 502 is specifically used to obtain the note template in the selected state among at least one note template; perform content generation processing in combination with the video and the note template in the selected state to obtain the content to be processed.
[0145] As a possible implementation manner of an embodiment of the present disclosure, the first content generation control is a note generation control; specifically, the note determination module 503 is further configured to, when the operation on the content to be processed is an insertion operation, determine the content to be processed as the content in the video note; when the operation on the content to be processed is a discard operation, discard the determination of the content to be processed as the content in the video note; when the operation on the content to be processed is a template change operation, discard the determination of the content to be processed as the content in the video note, and display at least one note template for reselecting the note template and regenerating the content to be processed.
[0146] As a possible implementation manner of an embodiment of the present disclosure, the first content generation control is a mind map generation control or an exercise generation control; specifically, the note determination module 503 is further configured to, when the operation on the content to be processed is an insertion operation, determine the content to be processed as the content in the video note; when the operation on the content to be processed is a discard operation, discard the determination of the content to be processed as the content in the video note; when the operation on the content to be processed is a regeneration operation, discard the determination of the content to be processed as the content in the video note, and regenerate the content to be processed in combination with the first content generation control.
[0147] As a possible implementation manner of an embodiment of the present disclosure, the device further includes: a discard module, configured to discard the content to be processed when no operation on the content to be processed is detected and a selection operation on a second content generation control in at least one content generation control is detected; the content generation module 502 is further configured to generate content for the video played in the video playback window in combination with the second content generation control to obtain the content to be processed.
[0148] As a possible implementation manner of an embodiment of the present disclosure, the device further includes: an object determination module, configured to determine a first playback object corresponding to the video playback window, and a shared video note shared by a second playback object to the first playback object and a shared video corresponding to the shared video note; a note acquisition module, configured to acquire the shared video note corresponding to the video when the shared video includes the video; a note display module, configured to display the shared video note corresponding to the video in the note processing area.
[0149] As a possible implementation manner of an embodiment of the present disclosure, the device further includes: a note sharing module, configured to share the video and the video note corresponding to the video to a second playback object; the second playback object is different from the first playback object corresponding to the video playback window.
[0150] As a possible implementation manner of an embodiment of the present disclosure, the device further includes: a window closing module, configured to close the video playback window and store the content to be processed when a closing operation for the video playback window is detected and there is content to be processed displayed in the note processing area; a window display module 501, further configured to display the video playback window when a reopening operation for the video playback window is detected, and display the content to be processed in the note processing area in the video playback window.
[0151] As a possible implementation manner of an embodiment of the present disclosure, the device further includes: a card display module, configured to display a generating card during the process of performing content generation processing on the video played in the video playback window by combining a first content generation control; and / or, display a generation failure card when the content generation fails.
[0152] As a possible implementation manner of an embodiment of the present disclosure, the device further includes: a window closing module, further configured to close the video playback window when a closing operation for the video playback window is detected and there is a generating card displayed in the note processing area, and continue to perform content generation processing on the video played in the video playback window by combining a first content generation control to obtain the content to be processed; a window display module 501, further configured to display the video playback window and display the content to be processed in the note processing area in the video playback window when a reopening operation for the video playback window is detected and the content generation processing is completed.
[0153] As a possible implementation manner of an embodiment of the present disclosure, the video note includes at least one of the following: a mind map, an exercise, a screenshot image of a video frame in the video, and relevant text content corresponding to the video frame in the video; the relevant text content is provided with a timestamp; the timestamp indicates the position of the video frame corresponding to the relevant text content in the video.
[0154] As a possible implementation manner of an embodiment of the present disclosure, the device further includes: a request receiving module, configured to receive a video note processing request; the video note processing request includes the content to be processed in the video note and the processing type for the content to be processed; a processing module, configured to process the content to be processed according to the processing type to obtain a processed video note.
[0155] As a possible implementation manner of an embodiment of the present disclosure, the processing type includes at least one of the following: a polishing type, an expanding type, a shortening type, and a continuing type.
[0156] The video note generation device according to an embodiment of the present disclosure displays a video playback window; a video playback area and a note processing area are provided in the video playback window; at least one content generation control is provided in the note processing area; in the case of detecting a selection operation on a first content generation control among the at least one content generation control, content generation processing is performed on the video being played in the video playback window by combining the first content generation control to obtain content to be processed; a video note corresponding to the video is determined according to the operation on the content to be processed and the content to be processed; wherein, by setting the content generation control, the video playback and note processing functions are integrated into one, eliminating the software isolation between the video player and the note tool, and the user does not need to frequently switch between the video player and the note tool.
[0157] In the technical solution of the present disclosure, the collection, storage, use, processing, transmission, provision, and disclosure of the user's personal information are all carried out on the premise of obtaining the user's consent, and all comply with the provisions of relevant laws and regulations and do not violate public order and good customs.
[0158] According to an embodiment of the present disclosure, the present disclosure also provides an electronic device, a readable storage medium, and a computer program product.
[0159] Figure 6 FIG. shows a schematic block diagram of an exemplary electronic device 600 that can be used to implement embodiments of the present disclosure. The electronic device is intended to represent various forms of digital computers, such as, for example, a laptop computer, a desktop computer, a workbench, a personal digital assistant, a server, a blade server, a mainframe computer, and other suitable computers. The electronic device can also represent various forms of mobile devices, such as, for example, a personal digital processor, a cellular phone, a smart phone, a wearable device, and other similar computing devices. The components shown herein, their connections and relationships, and their functions are merely exemplary and are not intended to limit the implementation of the present disclosure described and / or claimed herein.
[0160] As Figure 6 shown, the device 600 includes a computing unit 601, which can perform various appropriate actions and processes according to a computer program stored in a read-only memory (ROM) 602 or a computer program loaded from a storage unit 608 into a random access memory (RAM) 603. In the RAM 603, various programs and data required for the operation of the device 600 can also be stored. The computing unit 601, the ROM 602, and the RAM 603 are connected to each other through a bus 604. An input / output (I / O) interface 605 is also connected to the bus 604.
[0161] Multiple components in device 600 are connected to I / O interface 605, including: input unit 606, such as a keyboard, mouse, etc.; output unit 607, such as various types of displays, speakers, etc.; storage unit 608, such as a disk, optical disc, etc.; and communication unit 609, such as a network card, modem, wireless communication transceiver, etc. Communication unit 609 allows device 600 to exchange information / data with other devices via a computer network such as the Internet and / or various telecommunication networks.
[0162] Computing unit 601 can be various general-purpose and / or special-purpose processing components with processing and computing capabilities. Some examples of computing unit 601 include but are not limited to a central processing unit (CPU), a graphics processing unit (GPU), various dedicated artificial intelligence (AI) computing chips, various computing units running machine learning model algorithms, a digital signal processor (DSP), and any suitable processor, controller, microcontroller, etc. Computing unit 601 executes the various methods and processes described above, such as the method for generating video notes. For example, in some embodiments, the method for generating video notes can be implemented as a computer software program that is tangibly contained in a machine-readable medium, such as storage unit 608. In some embodiments, part or all of the computer program can be loaded and / or installed onto device 600 via ROM 602 and / or communication unit 609. When the computer program is loaded into RAM 603 and executed by computing unit 601, one or more steps of the method for generating video notes described above can be executed. Alternatively, in other embodiments, computing unit 601 can be configured to execute the method for generating video notes in any other suitable manner (e.g., by means of firmware).
[0163] Various embodiments of the systems and techniques described above herein can be implemented in digital electronic circuit systems, integrated circuit systems, field-programmable gate arrays (FPGA), application-specific integrated circuits (ASIC), application-specific standard products (ASSP), system-on-chip systems (SOC), complex programmable logic devices (CPLD), computer hardware, firmware, software, and / or combinations thereof. These various embodiments can include: implemented in one or more computer programs that can be executed and / or interpreted on a programmable system including at least one programmable processor, which can be a special-purpose or general-purpose programmable processor, and can receive data and instructions from a storage system, at least one input device, and at least one output device, and transmit the data and instructions to the storage system, the at least one input device, and the at least one output device.
[0164] The program code for implementing the methods of the present disclosure can be written in any combination of one or more programming languages. These program codes can be provided to a processor or controller of a general purpose computer, a special purpose computer, or other programmable data processing device, such that when executed by the processor or controller, the program codes cause the functions / operations specified in the flowchart and / or block diagram to be implemented. The program code can be executed entirely on the machine, partially on the machine, as a stand-alone software package partially on the machine and partially on a remote machine, or entirely on a remote machine or server.
[0165] In the context of the present disclosure, a machine-readable medium can be a tangible medium that can contain or store a program for use by or in connection with an instruction execution system, apparatus, or device. A machine-readable medium can be a machine-readable signal medium or a machine-readable storage medium. A machine-readable medium can include, but is not limited to, electronic, magnetic, optical, electromagnetic, infrared, or semiconductor systems, apparatus, or devices, or any suitable combination of the foregoing. More specific examples of a machine-readable storage medium would include an electrical connection based on one or more wires, a portable computer diskette, a hard disk, a random access memory (RAM), a read-only memory (ROM), an erasable programmable read-only memory (EPROM or Flash memory), an optical fiber, a portable compact disc read-only memory (CD-ROM), an optical storage device, a magnetic storage device, or any suitable combination of the foregoing.
[0166] In order to provide interaction with a user, the systems and techniques described herein can be implemented on a computer having: a display device (e.g., a CRT (cathode ray tube) or LCD (liquid crystal display) monitor) for displaying information to the user; and a keyboard and a pointing device (e.g., a mouse or a trackball) by which the user can provide input to the computer. Other kinds of devices can also be used to provide interaction with the user; for example, the feedback provided to the user can be any form of sensory feedback (e.g., visual feedback, auditory feedback, or tactile feedback); and input from the user can be received in any form (including acoustic input, speech input, or tactile input).
[0167] The systems and techniques described herein can be implemented in a computing system including backend components (e.g., as a data server), or a computing system including middleware components (e.g., an application server), or a computing system including frontend components (e.g., a user computer having a graphical user interface or a web browser through which a user can interact with an implementation of the systems and techniques described herein), or a computing system including any combination of such backend components, middleware components, or frontend components. The components of the system can be interconnected to each other by digital data communication in any form or medium (e.g., a communication network). Examples of communication networks include: local area network (LAN), wide area network (WAN), and the Internet.
[0168] A computer system can include a client and a server. The client and the server are generally far from each other and typically interact through a communication network. The client-server relationship is created by computer programs running on the respective computers and having a client-server relationship with each other. The server can be a cloud server, a server of a distributed system, or a server incorporating blockchain.
[0169] It should be understood that various forms of the processes shown above can be used, with steps reordered, added, or deleted. For example, the steps recited in this disclosure can be executed in parallel, sequentially, or in a different order, as long as the desired results of the technical solution disclosed in this disclosure can be achieved, and this is not limited herein.
[0170] The above specific embodiments do not constitute a limitation on the protection scope of this disclosure. Those skilled in the art should understand that various modifications, combinations, sub-combinations, and substitutions can be made according to design requirements and other factors. Any modifications, equivalent substitutions, and improvements made within the spirit and principles of this disclosure shall be included within the protection scope of this disclosure.
Claims
1. A method for generating a video note, the method comprising: Display the video playback window; The video playback window is provided with a video playback area and a note processing area; At least one content generation control is provided in the note processing area; When a selection operation is detected for a first content generation control among at least one of the content generation controls, performing content generation processing on the video played in the video playback window in combination with the first content generation control to obtain content to be processed; The video notes corresponding to the video are determined according to the operation on the content to be processed and the content to be processed.
2. The method according to claim 1, wherein: The content generation control includes at least one of the following: a note generation control, a mind map generation control, and an exercise generation control.
3. The method according to claim 1 or 2, wherein: The content generation control includes a note generation control; the method further includes: In the case where no video note is displayed in the note processing area and no selection operation is detected for any of the content generation controls, at least one note template is displayed; and one of the at least one note template is in a selected state; When a selection operation is detected for a first note template that is not in a selected state among at least one of the note templates, the selected note template is switched to the first note template.
4. The method according to claim 3, wherein: The note template includes at least one of the following: a text note template; an image note template; an outline note template; a graphic note template; The outline note template is used to generate content to be processed that is provided with text, images, and mind maps.
5. The method according to claim 3, wherein: The first content generation control is a note generation control; the step of performing content generation processing on the video played in the video playback window in combination with the first content generation control to obtain content to be processed includes: Get a selected note template from at least one note template; Content generation processing is performed in combination with the video and the note template in a selected state to obtain the content to be processed.
6. The method according to claim 1 or 2, wherein: The first content generation control is a note generation control; The step of determining the video note corresponding to the video according to the operation on the content to be processed and the content to be processed includes: In a case where the operation on the to-be-processed content is an insert operation, determining the to-be-processed content as the content in the video note; In a case where the operation on the to-be-processed content is a abandonment operation, abandoning determination of the to-be-processed content as the content in the video note; In the case where the operation on the content to be processed is a template change operation, the determination of the content to be processed as the content in the video note is abandoned, and at least one note template is displayed for reselection of the note template and regeneration of the content to be processed.
7. The method according to claim 1 or 2, wherein: The first content generation control is a mind map generation control or an exercise generation control; and the step of determining the video notes corresponding to the video according to the operation on the content to be processed and the content to be processed includes: In a case where the operation on the to-be-processed content is an insert operation, determining the to-be-processed content as the content in the video note; In a case where the operation on the to-be-processed content is a abandonment operation, abandoning determination of the to-be-processed content as the content in the video note; In the case where the operation on the content to be processed is a regeneration operation, the determination of the content to be processed as the content in the video note is abandoned, and the content to be processed is generated by recombining with the first content generation control.
8. The method according to claim 1, wherein: The method further comprises: When no operation on the content to be processed is detected and a selection operation on a second content generation control in at least one of the content generation controls is detected, abandoning the content to be processed; In combination with the second content generation control, content generation processing is performed on the video played in the video playback window to obtain content to be processed.
9. The method according to claim 1, wherein: Before performing content generation processing on the video played in the video playback window in combination with the first content generation control to obtain content to be processed, the method further includes: Determine a first playback object corresponding to the video playback window, a shared video note shared by a second playback object to the first playback object, and a shared video corresponding to the shared video note; In a case where the shared video includes the video, obtaining a shared video note corresponding to the video; The shared video notes corresponding to the video are displayed in the note processing area.
10. The method according to claim 1 or 9, wherein: The method further comprises: The video and the video notes corresponding to the video are shared to a second playback object; the second playback object is different from the first playback object corresponding to the video playback window.
11. The method according to claim 1, wherein: The method further comprises: When a closing operation on the video playback window is detected and there is content to be processed displayed in the note processing area, closing the video playback window and storing the content to be processed; When a reopening operation of the video playback window is detected, the video playback window is displayed, and the to-be-processed content is displayed in a note processing area in the video playback window.
12. The method according to claim 1, wherein: The method further comprises: In the process of performing content generation processing on the video played in the video playback window in combination with the first content generation control, displaying a generating prompt text; and / or, If content generation fails, display the generation failure text.
13. The method according to claim 12, wherein: The method further comprises: When a closing operation on the video playback window is detected and a generating prompt text is displayed in the note processing area, the video playback window is closed, and the video played in the video playback window is continuously subjected to content generation processing in combination with the first content generation control to obtain content to be processed; When a reopening operation of the video playback window is detected and the content generation process is completed, the video playback window is displayed and the to-be-processed content is displayed in the note processing area in the video playback window.
14. The method according to claim 1, wherein: The video notes include at least one of the following: a mind map, exercises, screenshot images of video frames in the video, and relevant text content corresponding to the video frames in the video; The relevant text content is provided with a timestamp; the timestamp indicates the position of the video frame corresponding to the relevant text content in the video.
15. The method according to claim 1, wherein: The method further comprises: receiving a video note processing request; the video note processing request includes the content to be processed in the video note and the processing type for the content to be processed; The content to be processed is processed according to the processing type to obtain a processed video note.
16. The method according to claim 15, wherein: The processing type includes at least one of the following: polishing type, expansion type, abbreviation type, and continuation type.
17. A device for generating video notes, comprising: Window display module, used to display the video playback window; The video playback window is provided with a video playback area and a note processing area; At least one content generation control is provided in the note processing area; A content generation module, configured to, when a selection operation is detected for a first content generation control among at least one of the content generation controls, perform content generation processing on the video played in the video playback window in combination with the first content generation control to obtain content to be processed; The note determination module is used to determine the video notes corresponding to the video according to the operation on the content to be processed and the content to be processed.
18. An electronic device comprising: at least one processor; as well as a memory communicatively connected to the at least one processor; wherein, The memory stores instructions that can be executed by the at least one processor, and the instructions are executed by the at least one processor to enable the at least one processor to perform the method according to any one of claims 1 to 16.
19. A non-transitory computer-readable storage medium storing computer instructions, wherein: The computer instructions are used to cause the computer to execute the method according to any one of claims 1 to 16.
20. A computer program product, comprising a computer program, which, when executed by a processor, implements the method of any one of claims 1 to 16.