Video browsing method and device

By displaying the frame image of the first video and the browsing control of the second video on the electronic device, generating and playing a second video based on TAG in response to user operations, the problem of insufficient video browsing function in the prior art is solved, and better user experience and storage savings are achieved.

CN116708650BActive Publication Date: 2025-06-10HONOR DEVICE CO LTD
View PDF 2 Cites 0 Cited by

Patent Information

Application Number
CN202210188471.4
Authority / Receiving Office
CN · China
Patent Type
Patents(China)
Current Assignee / Owner
Filing Date
2022-02-28
Publication Date
2025-06-10
Estimated Expiration
2042-02-28

AI Technical Summary

Technical Problem

The video browsing function of existing electronic devices is insufficient, making it difficult to effectively display the browsing method of related videos, affecting the user experience.

Method used

A video browsing method is provided, by displaying a frame image of the first video and a browsing control of the second video, generating and playing a TAG-based second video in response to a user operation, thereby realizing browsing of associated videos.

Benefits of technology

Improved video browsing capabilities of electronic devices, provides a new way of browsing, enhances user experience, and saves storage space.

✦ Generated by Eureka AI based on patent content.

Smart Images

  • Figure CN116708650B_ABST
    Figure CN116708650B_ABST
Patent Text Reader

Abstract

An embodiment of the present application provides a method for browsing a video, which is applied to an electronic device. The method includes: displaying a first interface, where the first interface displays at least one frame of an image in a first video and a browsing control for a second video, and the first video includes a tag TAG. In response to an operation on the browsing control for the second video, a browsing interface for the second video is displayed, and the second video generated based on the TAG is played in the browsing interface. Since the second video is generated based on the first video, a browsing method for videos with an associated relationship is provided, improving the browsing function and facilitating a better user experience.
Need to check novelty before this filing date? Find Prior Art

Description

Technical Field

[0001] This application relates to the field of multimedia technologies, and in particular, to a method and apparatus for browsing videos. Background Art

[0002] Shooting videos is a common function of electronic devices. Based on the videos captured by users (referred to as original videos), processed videos can be further obtained. In this case, the video browsing function of electronic devices needs to be improved. Summary of the Invention

[0003] This application provides a method and apparatus for browsing videos, aiming to improve the video browsing function of electronic devices.

[0004] To achieve the above objective, this application provides the following technical solutions:

[0005] A first aspect of this application provides a method for browsing videos, which is applied to an electronic device. The method includes: displaying a first interface, where the first interface displays at least one frame of an image in a first video and a browsing control for a second video, and the first video includes a tag TAG. In response to an operation on the browsing control for the second video, a browsing interface for the second video is displayed, and the second video generated based on the TAG is played in the browsing interface. Since the second video is generated based on the first video, a browsing method for videos with an associated relationship is provided, improving the browsing function and facilitating a better user experience.

[0006] In one implementation, displaying the first interface, where the first interface displays at least one frame of an image in the first video and the browsing control for the second video, includes: in response to an operation on the first frame of the image in the first video downloaded from the cloud, the first interface is displayed, and the first interface displays the first frame of the image in the first video and the browsing control for the second video.

[0007] In one implementation, in response to an operation on the browsing control for the second video, displaying the browsing interface for the second video and playing the second video generated based on the TAG in the browsing interface includes: in response to an operation on the browsing control for the second video, downloading the first video; generating the second video based on the TAG; and playing the second video in the browsing interface for the second video.

[0008] In one implementation, the browsing control for the second video includes: a thumbnail of the first cover of the second video, and the first cover of the second video includes a combination of the first frame of the image in the first video and special effects.

[0009] In one implementation, after downloading the first video, it further includes: generating a second cover for the second video based on the TAG, where the second cover is different from the first cover; and using the first cover as the cover of the second video to keep the covers before and after downloading the original video consistent.

[0010] In one implementation, the process of obtaining the first video includes: collecting a video through a camera; determining that the video meets the first condition or the second condition, where the first condition includes that the duration of the video is greater than a first duration and a wonderful moment is recognized in the video; generating a TAG based on the theme of the video, the wonderful moment, and the scenes recognized in the video; writing the TAG into the video to generate the first video.

[0011] In one implementation, the process of obtaining the first video includes: collecting a video through a camera; determining that the video meets the second condition, where the second condition includes that the duration of the video is greater than a second duration and the video includes video frames that meet the quality condition; generating a TAG based on the theme of the video, the video frames that meet the quality condition, and the scenes recognized in the video; writing the TAG into the video to generate the first video.

[0012] In one implementation, the first interface displays at least one frame image of the first video and browsing controls for the second video, including: a playback area is included in the first interface, and the first video is played in the playback area; the browsing controls for the second video are displayed in the playback area.

[0013] In one implementation, the first interface further includes: a thumbnail area that displays at least a thumbnail of one frame image of the first video.

[0014] In one implementation, the first interface displays at least one frame image of the first video and browsing controls for the second video, including: a thumbnail area and a playback area are included in the first interface, the thumbnail area displays a thumbnail of the cover of the second video that serves as the browsing control for the second video, and the first video is played in the playback area.

[0015] In one implementation, the cover of the second video includes: a combination of a wonderful moment photo in the first video and special effects.

[0016] In one implementation, the thumbnail area further displays: a thumbnail of one frame image of the first video.

[0017] In one implementation, playing the second video includes: displaying a second interface that includes a cover and playback controls; in response to an operation on the playback controls, displaying a browsing interface for the second video and playing the second video in the browsing interface.

[0018] In one implementation, the thumbnail area further includes: a thumbnail of a wonderful moment photo in the first video.

[0019] In one implementation, playing the second video includes: playing the second video in the browsing interface of the second video; the browsing interface of the second video includes: special effect selection controls to facilitate the user to independently select special effects.

[0020] In one implementation, before playing the second video, it further includes: displaying a loading interface for the second video, where the loading interface includes a loading prompt control.

[0021] In one implementation, before displaying the first interface, it further includes: displaying a gallery interface, where a thumbnail of the first video is displayed in the gallery interface, and an identifier is included in the thumbnail of the first video, and the identifier indicates that the first video has an associated second video, so as to achieve a better guiding effect on the user.

[0022] In a second aspect, the present application provides an electronic device, including: one or more processors, a memory, a display screen, and a camera; the memory, the display screen, and the camera are coupled to the one or more processors, the memory is used to store computer program code, the computer program code includes computer instructions, and when the one or more processors execute the computer instructions, the electronic device executes the video browsing method according to any item of the first aspect.

[0023] In a third aspect, the present application provides a computer storage medium for storing a computer program, and when the computer program is executed, it is specifically used to implement the video browsing method according to any item of the first aspect.

[0024] In a fourth aspect, the present application provides a computer program product containing instructions. When the computer program product runs on a computer or a processor, the computer or the processor is caused to execute the video browsing method according to any item of the first aspect above. Description of the Drawings

[0025] Figure 1 It is an example diagram of obtaining a video by the one-record-multiple-gains function provided by the present application;

[0026] Figure 2 It is a GUI legend in the implementation process of a video browsing method disclosed by the present application;

[0027] Figure 3 It is a GUI legend in the implementation process of another video browsing method disclosed by the present application;

[0028] Figure 4 It is a GUI legend in the implementation process of another video browsing method disclosed by the present application;

[0029] Figure 5 It is a GUI legend of a way for a cloud synchronization device disclosed by the present application to browse videos;

[0030] Figure 6 It is a GUI legend of another way for a cloud synchronization device disclosed by the present application to browse videos;

[0031] Figure 7 The hardware structure diagram of the electronic device provided by this application;

[0032] Figure 8 The software architecture diagram of the electronic device provided by this application. Specific embodiments

[0033] Next, the technical solutions in the embodiments of this application will be clearly and completely described in conjunction with the accompanying drawings in the embodiments of this application. The terms used in the following embodiments are only for the purpose of describing specific embodiments, and are not intended to limit this application. As used in the specification and claims of this application, the singular forms "a", "an", "the", "above-mentioned", "this" and "such" are also intended to include the expression form such as "one or more", unless there is a clear opposite indication in the context. It should also be understood that in the embodiments of this application, "one or more" means one, two or more than two; " / ", which describes the association relationship of associated objects, indicates that three relationships can exist; for example, A and / or B can mean: A exists alone, A and B exist at the same time, and B exists alone, where A and B can be singular or plural. The character " / " generally represents an "or" relationship between the associated objects before and after.

[0034] Referring to "one embodiment" or "some embodiments" etc. described in this specification means that in one or more embodiments of this application, specific features, structures or characteristics described in combination with this embodiment are included. Thus, the statements "in one embodiment", "in some embodiments", "in other some embodiments", "in still other embodiments" etc. that appear in different places in this specification do not necessarily refer to the same embodiment, but mean "one or more but not all embodiments", unless otherwise specifically emphasized in other ways. The terms "comprising", "including", "having" and their variants all mean "including but not limited to", unless otherwise specifically emphasized in other ways.

[0035] Multiple involved in the embodiments of this application means greater than or equal to two. It should be noted that in the description of the embodiments of this application, words such as "first" and "second" are only used for the purpose of distinguishing descriptions, and cannot be understood as indicating or implying relative importance, nor can they be understood as indicating or implying order.

[0036] In the following examples, the operation methods of objects such as controls and icons in the graphical user interface (GUI) by the user include but are not limited to touching specific controls on the mobile phone screen, pressing specific physical buttons or button combinations, inputting voices, and air gestures.

[0037] Figure 1An example of a GUI for a user to obtain a video using a mobile phone Figure 1 In it, the user enters the interface of the camera application as shown in Figure 1 (1). After the user clicks on the settings control of the camera in Figure 1 (1), the mobile phone 1 jumps to the settings interface of the camera as shown in Figure 1 (2). In the interface shown in Figure 1 (2), by operating the "One-shot, Multiple-gains" control 101, the "One-shot, Multiple-gains" function of the mobile phone 1 is enabled (enabled by default). Figure 1 For the functions of other settings items in (2), reference can be made to the setting rules or standards of the shooting function, which will not be elaborated here. And it can be understood that for the functions of the controls, texts and other objects that are not labeled or described in the drawings of this application, reference can be made to the setting (and usage) rules or standards of the corresponding interfaces of the mobile phone, which will not be elaborated here. It can be understood that in addition to enabling the One-shot, Multiple-gains function through the camera settings interface, the One-shot, Multiple-gains function can also be enabled through the settings interface of the mobile phone 1, which is not limited here.

[0038] One-shot, Multiple-gains can be understood as the function of obtaining multimedia content by pressing the shooting control (such as Figure 1 (2) shown as 102) once when the user uses the camera application to shoot a video. It can be understood that One-shot, Multiple-gains may also have other names, such as One-key, Multiple-gains, One-key, Multiple-shots, One-key, Output-film, One-key, Blockbuster, AI One-key, Blockbuster, etc.

[0039] Among them, the multimedia content includes the original video and the AI video. The video recorded in response to pressing the "Shoot" control, that is, the video obtained by the video recording function of the camera of the electronic device, is called the original video or the first video.

[0040] During the acquisition process of the original video, the One-shot, Multiple-gains function adds tags (TAGs) to the original video. Some examples of TAGs include: scene TAG, transition (storyboard) TAG, the start and end times corresponding to the scene TAG and the transition (storyboard) TAG, recognition of magic moments (MM), and theme TAG. In some implementation manners, the TAGs are stored in the header of the original video file.

[0041] MM refers to some wonderful picture moments during the recording process of the video (original video). For example, MM can be the best sports moment, the best expression moment, or the best check-in action moment. It can be understood that this application does not limit the term MM, and MM can also be called wonderful moment, magic moment, wonderful instant, decisive moment, or best shot (BS), etc.

[0042] MM tag, i.e., time tag, is used to indicate the position of MM in the recorded video file. For example, in a video file, there is one or more MM tags, which can indicate moments such as the 10th second, 1 minute and 20 seconds in the video file.

[0043] In the function of getting multiple recordings at once, the generation strategy of the AI video is as follows: intercept the segment including MM from the original video according to the TAG. The duration of each segment can be between 2.5 seconds and 6 seconds (2.5 seconds or 6 seconds can be included). Then splice all the segments in the chronological order of the segments and add special effects to obtain the AI video. In some implementation manners, the process of intercepting video segments is to intercept before and after the MM photo, and prevent video segments from crossing shot boundaries based on the scene tag.

[0044] The AI video can also be called a wonderful video, a selected video, a wonderful short video, a selected short video, or a wonderful video, etc. Opposite to the first video, the AI video can also be called the second video. In the following embodiments of the present application, the special effects include objects such as filters, stickers, themes, styles, sound effects, and text that can be added to images. It can be understood that, corresponding to the original video, the AI video can also be called a wonderful video, a selected video, a wonderful short video, a selected short video, or a wonderful video, etc.

[0045] In some implementation manners, the above generation strategy is formed into a configuration file and saved in the mobile phone 1. That is, the configuration file includes: special effect templates, music, and the start and end times of each segment, which are used to generate the AI video subsequently. In some cases, the configuration file also includes the cover information of the AI video. In other implementation manners, instead of saving the AI video configuration file, when generating the AI video, the start and end times of each segment, special effect templates, and music are determined in real time according to the TAG, and the AI video is generated according to the determination.

[0046] However, regardless of whether the AI video configuration file is saved or not, the generation and saving of the AI video are triggered by the user's needs, rather than default generating and saving the AI video to save resources. The specific implementation manners will be described in the following embodiments.

[0047] The multimedia content can also include MM photos. MM photos refer to the image frames corresponding to the MM tags in the recorded video (original video). That is to say, the image frame with the time stamp as the MM tag is the MM photo. It can be understood that in different scenarios, the MM photos can be of different types of pictures. For example, when recording a football game video, the MM photo can be the photo of the moment when the athlete's foot touches the football during a shot or a pass, or the photo of the moment when the football flies into the goal. When recording a video of a person jumping off the ground, the wonderful moment can be the photo of the moment when the person is at the highest point in the air, or the photo of the moment when the person's action is most stretched in the air.

[0048] The original video, MM photos, and the generated AI video obtained from one-shot multi-capture have the same identifier to indicate the association relationship. The identifier can be stored in the database associated with the photo gallery.

[0049] It can be understood that in some implementation manners, when the one-shot multi-capture function is enabled, it is not necessarily the case that every time the video recording function of the camera is used to collect a video, the above-mentioned multimedia content can be obtained. For example, when the duration of the original video is greater than a first duration such as 15 seconds and an MM is recognized, the original video with a TAG and MM photos may be saved, and an AI configuration file may also be saved. Another example is that when the duration of the original video is greater than a second duration such as 30 seconds and there are high-quality video frames, the original video with a TAG may be saved, and an AI configuration file may also be saved. In this case, an AI video can be generated, but there are no MM photos, and the segments in the AI video are intercepted based on the high-quality video frames.

[0050] The embodiments of the present application are described by taking the original video with a TAG, MM photos, and possibly including an AI configuration file obtained based on the one-shot multi-capture function as an example.

[0051] In Figure 1 After the one-shot multi-capture function is enabled as shown in (2), the user can click the Figure 1 return control shown in (2) to return to the interface of the camera application shown in Figure 1 (3), and through the video recording function of the camera application, perform one-shot multi-capture shooting. As shown in Figure 1 (3), in the interface displayed by the camera application, when the user selects the video recording function and clicks the shooting control 102, the mobile phone 1 starts the shooting process of the one-shot multi-capture function.

[0052] Figure 1 (4) shows the shooting process of one-shot multi-capture. Taking the shooting of a football player as an example, Figure 1 The picture captured at 11 seconds of the recording duration is shown in the shooting window 103 in (4).

[0053] Figure 1 The shooting end control 104 is also shown in (4). During the shooting process, when the user clicks the shooting end control 104, the shooting stops, and the camera interface jumps to Figure 1 (5) as shown.

[0054] Figure 1 In (5), the shooting process of the video has ended, and the picture currently captured by the camera is shown in the shooting window 103. The control 105 is the browsing entry for the video and photos obtained from one-shot multi-capture shooting. The user can enter the browsing mode by clicking the control 105. Figure 1In (5), an example of the display style of the control 105 is the first frame image of the original video obtained by "one recording, multiple gains".

[0055] As described above, Figure 1 (1) to Figure 1 The multimedia content that can be browsed by the user in the scenarios shown in (5) includes: the original video, at least one (e.g., 5) MM photos, and the AI video.

[0056] The video browsing method disclosed in the embodiments of the present application aims to realize the browsing of videos with an associated relationship to obtain a better user experience.

[0057] The video browsing method disclosed in the embodiments of the present application is applied to an electronic device, which includes but is not limited to mobile phones, tablet computers, desktop, laptop, notebook computers, ultra-mobile personal computers (UMPCs), handheld computers, netbooks, personal digital assistants (PDAs), wearable electronic devices, smart watches, and other electronic devices with cameras.

[0058] The electronic device has the function of shooting videos through a camera, and after shooting, it obtains videos and / or video profiles that can generate videos.

[0059] The above Figure 1 The "one recording, multiple gains" scenario shown above is an example of the electronic device shooting videos. The videos described in the embodiments of the present application are not limited to the videos and video profiles obtained by the "one recording, multiple gains" function. And, Figure 1 The GUI interface shown above is only an example for implementing the "one recording, multiple gains" function and is not used as a limitation.

[0060] Figure 2 The following is a GUI legend in the implementation process of a video browsing method disclosed in the embodiments of the present application:

[0061] Figure 2 (1) is the interface of the photo gallery on the mobile phone 1. The original video obtained by "one recording, multiple gains" is displayed in this interface in the form of the first thumbnail 106 of the original video. In some implementation manners, the first thumbnail 106 of the original video is any frame in the original video. Taking Figure 2 (1) as an example, any frame is the first frame. The first thumbnail 106 of the original video is the first frame image of the original video, or, as shown by 106 in Figure 2 (1), an image obtained by proportionally reducing the first frame image of the original video, or an image obtained by intercepting the middle area of the first frame image of the original video and then reducing it. Here, the manner of intercepting the middle area is not limited.

[0062] To reflect the differences between the videos obtained by the "One Recording, Multiple Gains" function and the videos and photos obtained by other functions, in some implementation manners, in the first thumbnail 106 of the original video obtained by the "One Recording, Multiple Gains" function, a "One Recording, Multiple Gains" identifier is also displayed, so as to Figure 2 Take the "One Recording, Multiple Gains" identifier 107 in the first thumbnail 106 in (1) as an example.

[0063] The duration of the original video may also be displayed in the first thumbnail 106. Figure 2 Take the duration "00:56" in (1) as an example.

[0064] It can be understood that the display position and style of the "One Recording, Multiple Gains" identifier 107 and the duration of the original video such as "00:56" in the first thumbnail 106 of the original video are only examples and are not intended to be limiting.

[0065] The first thumbnail 106 of the original video is an entry for browsing the original video. In addition, as mentioned above, another entry for browsing the original video is Figure 1 105 shown in (5).

[0066] Figure 2 (2) shows an example of the interface presented by the mobile phone 1 after the user clicks on the first thumbnail 106 of the original video.

[0067] Figure 2 The interface shown in (2) includes four areas: A, B, C, and D.

[0068] Area A displays the shooting time information "April 15, 2021", the location information "Beijing", and the first return control 108 of the original video. After the user clicks on the first return control 108, the display interface of the mobile phone 1 switches to Figure 2 Shown in (1).

[0069] Area B displays the first frame image of the original video, the control "AI One-Click Blockbuster" (abbreviated as the AI video browsing control) 109, which is an entry for browsing the AI video, and the first playback control 110. The first playback control 110 includes a playback status control 1101 and a sound control 1102. It can be understood that showing the first frame image of the original video is only an example, and other frame images of the original video may also be shown. When the image displayed in area B is the same as the image represented by the thumbnail 106, a better user experience can be obtained.

[0070] Region C displays the second thumbnail 111 of the original video and MM photos 112 - 114. It can be understood that in the case where there are 5 MM photos, since it is a vertical screen display and the display width is limited, only a part of MM photo 114 is shown. The user can make MM photo 114 fully displayed by performing a sliding or clicking operation in Region C. Moreover, MM photos 115 and 116 are not shown in Figure 2 (2), and the user can make MM photos 115 and 116 shown in Figure 2 (2) by performing a sliding operation in Region C.

[0071] A pointer control 117 is also shown in Region C. The pointer control 117 is used to indicate the currently selected thumbnail. In some implementation manners, the style of the pointer control 117 is a thickened border, and in some other implementation manners, the style of the pointer control 117 is a pointer located above the thumbnail. The style of the pointer control is not limited in this embodiment.

[0072] As Figure 2 (2) shows, in response to the user's click on the first thumbnail 106 of the original video in Figure 2 (1), after just entering the interface shown in Figure 2 (2), Figure 2 (2) defaults to display the first frame image of the original video, that is, the pointer control 117 in Region C is located at the border of the second thumbnail 111 of the original video, indicating that Region B currently shows the second thumbnail 111 of the original video. It can be understood that when the user selects any one of the thumbnails 112 - 116 of the MM photos in Region C (115 and 116 are not shown in Figure 2 (2)), the pointer control 117 moves to the border of the thumbnail selected by the user.

[0073] It can be understood that the thumbnails of the MM photos shown in Region C are, in some implementation manners, images of the MM photos reduced in proportion, and in some other implementation manners, images of the MM photos after intercepting the middle area and then reducing. The way of intercepting the middle area is not limited here. The second thumbnail 111 of the original video in Region C is an image of the first frame of the original video reduced in proportion or after intercepting the middle area and then reducing. Figure 2 (2) takes the images of thumbnails 111 - 114 in Region C after intercepting the middle area and then reducing as examples.

[0074] Region D shows operation controls, and the operations take "Share", "Favorite", "Edit", "Delete", and "More" (indicating more operations) as examples.

[0075] In order to implement the prompting and guiding functions for the user, Figure 2In (2), regions A, B, and D are all covered with a "mask layer". In this embodiment, the "mask layer" can be understood as a kind of mask layer. Objects in the regions masked by the "mask layer", that is, covered areas, including controls such as 109 and 110, etc., cannot be operated. The video in the region covered by the "mask layer" is in a stopped playback state.

[0076] Since the first playback control control 110 cannot be operated by the user and the original video is not being played, the playback status control 1101 is in a state indicating stopped playback, and the sound control 1102 indicates a mute state.

[0077] In some implementation manners, in order to highlight the "mask layer", the mobile phone 1 displays the "mask layer" in a different style from other regions. For example, Figure 2 in (2), the "mask layer" is displayed as a gray color with a certain transparency.

[0078] In some implementation manners, Figure 2 in (2), a prompt message "One recording, multiple gains. Automatically capture wonderful moments for you." 118 is also displayed to cooperate with the "mask layer" to achieve the prompt effect of the one recording, multiple gains function.

[0079] Figure 2 In (2), taking the prompt message 118 located in region B and above the "mask layer" as an example, in addition, the prompt message 118 can also be located in other positions, and the displayed content is not limited.

[0080] It can be understood that because the function of the "mask layer" is to prompt and guide, and in order not to affect the normal browsing of the original video, in some implementation manners, the "mask layer" disappears after being displayed for a preset duration and jumps to Figure 2 the interface taking (3) as an example. In some other implementation manners, after the user clicks on any position of the "mask layer", the "mask layer" disappears and jumps to Figure 2 the interface taking (3) as an example.

[0081] It can be understood that because region C is not covered by the "mask layer", the user can operate on Figure 2 each thumbnail in region C in (2). For example, after the user clicks on the thumbnail 111 of the original video, the interface of the mobile phone 1 jumps to Figure 2 the interface taking (3) as an example. Another example is that after the user clicks on any one of the thumbnails 112 - 116 of the MM photos, the interface of the mobile phone 1 jumps to the display interface of the MM photo clicked by the user ( Figure 2 not drawn in the figure). An example of the display interface of the MM photo is: replacing the content displayed in region B in Figure 2 (3) with the MM photo selected by the user, and the pointer control 117 is located on the border of the MM photo selected by the user.

[0082] Of course, the "mask layer" can also cover area C, which is not limited in this embodiment.

[0083] Figure 2 (3) shows the display interface of the mobile phone 1 after the "mask layer" disappears. As mentioned before, if it is to display a preset duration, or the user clicks on any position of the "mask layer", or the user clicks on the second thumbnail 111 of the original video, the mobile phone 1 changes from Figure 2 (2) to Figure 2 (3). Figure 2 In (3), the content displayed in areas A, C, and D is the same as that in Figure 2 (2).

[0084] Figure 2 In (3), the prompt message 118 is no longer displayed in area B, but the original video is played. In some implementation manners, the original video starts playing from the "00:00" moment. Figure 2 Taking the example that the original video has been played to the "00:01" moment in (3). Correspondingly, the play status control 1101 in the first play control control 110 indicates the play status, the play time control 1103 shows that the played duration is "00:01", and the duration control 1105 shows that the total duration of the original video is "00:56".

[0085] The sound control 1102 indicates the state of the sound being played or prohibited from playing. In some implementation manners, as Figure 2 (3) shows, the sound control 1102 is in the mute state, indicating the state of the sound being prohibited from playing. That is to say, after the mobile phone 1 starts playing the original video, the original video is muted. The purpose is to reduce the possibility that the original video emits sound when the mobile phone 1 is in an environment where it is not appropriate to play sound, so as to further improve the user experience. In other implementation manners, the sound control 1102 can also be in the play state by default to play the sound of the original video.

[0086] It can be understood that the user can drag the slider 1104 of the first play control control 110 to change the current play progress. Correspondingly, the image frame displayed in area B also changes accordingly. For example, when the user drags the slider 1104 of the first play control control 110 to the "00:10" moment when the original video is played to the "00:01" moment, the image frame corresponding to the "00:10" moment in the original video is displayed in area B, that is, the image frame with the time stamp of "00:10" moment in the original video. Correspondingly, the play time control 1103 shows the moment "00:10" ( Figure 2 not drawn in the figure).

[0087] The user can also click on any thumbnail in area C in Figure 2 (3). In Figure 2When playing the original video in area B as shown in (3), if the user clicks on the second thumbnail 111 of the original video, the original video starts playing from the first frame in area B, that is, starts playing the original video from the "00:00" moment. That is to say, after the original video has been played to "00:01", in response to the user's operation of clicking on the second thumbnail 111 of the original video, the original video starts playing from the beginning in area B, and the pointer control 117 is located at the border of the second thumbnail 111 of the original video.

[0088] When Figure 2 When playing the original video in area B as shown in (3), if the user clicks on any one of the thumbnails 112 - 114 of the MM photos, or first slides to display the MM photos 115 and 116 and then clicks on the MM photo 116, the MM photo clicked by the user is displayed in area B, and the original video stops playing. The pointer control 117 is located at the border of the MM photo clicked by the user. For example, if the user clicks on the MM photo 112, the pointer control 117 moves from the border of the thumbnail 111 of the original video to the border of the MM photo 113, and in area B, the display switches from playing the original video to showing the MM photo 113 ( Figure 2 not shown in the figure).

[0089] The operation control in area D is used to act on the object represented by the thumbnail indicated by the pointer control 117. In some implementation manners, after the user clicks on the second thumbnail 111 of the original video in area C, the pointer control 117 is located at the border of the second thumbnail 111 of the original video. In this case, if the user clicks on the "Delete" control in area D, the mobile phone 1 displays a prompt message asking whether to delete the video. If the user selects "Yes", the mobile phone 1 deletes the original video. In other implementation manners, if the user clicks on the MM photo 114 in area C, the pointer control 117 changes from the border of the second thumbnail 111 of the original video to the border of the MM photo 113. In this case, if the user clicks on the "Delete" control in area D, the mobile phone 1 displays a prompt message asking whether to delete the photo. If the user selects "Yes", the mobile phone 1 deletes the MM photo 114.

[0090] It can be understood that Figure 2 the "mask layer" shown in (2) can be displayed only when the user views the video with the one - record - multiple - benefits identifier 107 for the first time, that is to say Figure 2 shown in (2) when the user first uses the mobile phone 1 to view the original video with an AI video profile, and is not shown in non - first cases.

[0091] Figure 2The selected thumbnail in area C is located in the middle of area C, but this is not a limitation. For example, the thumbnails can also be displayed according to the width of area C. Thumbnail 111 is displayed at the leftmost end of area C. After thumbnail 111 is selected, pointer control 117 is located in the border of thumbnail 111. The position of thumbnail 111 in area C remains unchanged and is still located at the leftmost end of area C instead of being displayed in the center.

[0092] Figure 2 (1)- Figure 2 (3) shows an example of a GUI in the process of browsing an original video. The following is a detailed description of the GUI in the process of browsing an AI video.

[0093] Figure 3 (1)- Figure 3 (3) GUI example of the process of users browsing AI videos:

[0094] Figure 3 (1) Figure 2 (3) Same, user clicks Figure 3 (1) After the AI ​​video browsing control 109, jump to Figure 3 (2) The interface shown.

[0095] Figure 3 (2) is an example of an AI video generation interface, where a loading icon 119, i.e., a circle composed of multiple black dots, indicates that an AI video is being generated (also referred to as loading). It is understood that the style and display location of the icon indicating that an AI video is being generated are not limited.

[0096] In addition to loading icon 119, Figure 3 The interface shown in (2) also displays the first frame of the AI ​​video. Figure 3 (2) The first frame of the AI ​​video is Figure 3 (1) is an example of the first MM photo 112 in area C and the added special effects such as the border 121. It is understandable that, because the AI ​​video is generated and played from the first frame, displaying the first frame of the AI ​​video in the AI ​​video generation process interface can avoid the conflict with the AI ​​video playback interface (such as Figure 3 (3)), thereby further improving the user experience.

[0097] As described above, the generation strategy of the AI video is as follows: intercept the segments including MM from the original video according to the TAG, and the duration of each segment can be between 2.5 seconds and 6 seconds (2.5 seconds or 6 seconds can be included), and then splice all the segments in the chronological order of the segments and add special effects to obtain the AI video. Based on this generation strategy, it can be understood that the image displayed in the first frame of the AI video may not be a photo of MM, or it may be a photo of MM. Therefore, Figure 2 (2) The MM photo 112 described above is only an example.

[0098] Figure 3 (2) The interface shown also includes a second return control 120. After the user clicks the second return control 120, the interface displayed on the mobile phone 1 changes from Figure 3 the one shown in (2) to Figure 3 the one shown in (1).

[0099] It can be understood that because Figure 3 (2) The interface shown is the generation interface of the AI video (also known as the loading interface), indicating that the AI video is being generated. Therefore, in some implementation manners, in addition to the second return control 120, Figure 3 other controls in the interface shown in (2) cannot be operated, that is, other controls do not respond to the user's operations. In other implementation manners, Figure 3 all controls in the interface shown in (2) cannot be operated. In this case, Figure 3 the purpose of displaying the controls in the interface shown in (2) is to be consistent with the display of the browsing interface of the AI video (also known as the playing interface) to further improve the user experience. It can be understood that Figure 3 (2) The interface shown may also only display the first frame of the AI video and the loading icon 119 without displaying other controls.

[0100] It can be understood that Figure 3 (2) The interface shown is only an example. In the case where the generation delay of the AI video is relatively long, such as generating the AI video by real-time decision-making as described above, Figure 3 (2) The interface shown can prompt the user that the AI video is being generated to further improve the user experience. In the case where the generation delay of the AI video is relatively short, such as generating the AI video according to the AI video configuration file as described above, the interface shown in Figure 3 (2) may not be displayed, and after the user clicks the AI video browsing control 109, the browsing interface of the AI video shown in Figure 3 (3) is displayed.

[0101] After the AI video is generated, the mobile phone 1 changes from the interface shown in Figure 3 (2) to Figure 3 the interface shown in (3). Figure 3(3) is the browsing interface (also known as the playback interface) of the AI video.

[0102] Figure 3 After the second return control 120 in (3) is clicked, the display interface of the mobile phone 1 changes from Figure 3 as shown in (3) to Figure 3 as shown in (1).

[0103] From Figure 3 the interface shown in (2) to Figure 3 the interface shown in (3), Figure 3 the AI video starts playing in the interface shown in (3), that is, starting from the first frame image of the AI video. During the playback of the AI video, Figure 3 the second playback control control 124 is also displayed in the interface shown in (3), indicating that the AI video is in the playback state. After the user clicks the second playback control control 124, the AI video stops playing, and the second playback control control 124 indicates that the AI video is in the paused playback state.

[0104] Figure 3 The playback time control 125 and the remaining duration control 127 are also displayed in the interface shown in (3). The playback time control 125 indicates the duration that the AI video has been played. Figure 3 Taking (3) as an example of "00:01", the duration control 127 indicates the total duration of the AI video. Figure 3 Taking (3) as an example of "00:14".

[0105] It can be understood that since the AI video is formed by adding special effects to the segments intercepted from the original video, in some implementation manners, the duration of the AI video is less than the duration of the original video. In other implementation manners, the duration of the AI video can also be equal to or greater than the duration of the original video.

[0106] The user can drag the slider control 126 to adjust the playback progress of the AI video. After the playback progress of the AI video is adjusted, the time displayed by the playback time control 125 is also adjusted accordingly.

[0107] It can be understood that the special effects in the AI video are added by the mobile phone 1 by default, that is, the mobile phone 1 automatically generates the AI video. In some implementation manners, the playback interface of the AI video also displays a control for adding special effects again (i.e., the special effect selection control), such as Figure 3 the music control 122, the editing control 123, and the style control 128 shown in (3).

[0108] After the user clicks the music control 123, the mobile phone 1 displays Figure 3 the interface shown in (4). Figure 3 The music list 1221 is displayed in (4), such as Figure 3As shown in (4), the music list includes music of types such as soothing, romantic, and warm. The user can select any music from the list as the background music for the original video, that is, the music played during the playback of the AI video. After the user selects any music, the display interface of the mobile phone 1 changes from Figure 3 (4) to Figure 3 (3), and plays the AI video generated based on the selected music by the user.

[0109] After the user clicks the editing control 123, the mobile phone 1 displays an editing interface ( Figure 3 not shown in the figure). The editing interface includes, but is not limited to, editing controls such as line cutting, splitting, volume adjustment, and frame adjustment. The user can click any editing control to edit the AI video.

[0110] After the user clicks the style control 128, the mobile phone 1 displays Figure 3 the interface shown in (5). Figure 3 (5) includes a style list 1281. The style list 1281 includes, but is not limited to, styles such as everlasting friendship and warm sun of the day. The user can select any style as the style added to the original video. After the user selects any style, the display interface of the mobile phone 1 changes from Figure 3 (5) to Figure 3 (3), and plays the AI video generated based on the selected style by the user.

[0111] Figure 3 The sharing control 130 is also displayed in the interface shown in (3). After the user clicks the sharing control 130, the mobile phone 1 displays a sharing interface, and realizes the function of sharing the AI video based on the user's operations on the sharing interface.

[0112] Figure 3 The saving control 129 is also displayed in the interface shown in (3). After the user clicks the saving control 129, the mobile phone 1 stores the AI video.

[0113] Figure 2 And Figure 3 the browsing method of the original video and the AI video shown. Displaying the browsing entry of the AI video in the browsing interface of the original video is beneficial to guiding the user to browse the AI video, and can clearly reflect the association relationship between the original video and the AI video, thus obtaining a good user experience.

[0114] Moreover, generating the AI video after the user clicks the control of the AI video is a video generation method based on the user's needs. Compared with the method of generating and saving the AI video after obtaining the original video, it can save storage space.

[0115] Figure 4 This is another video browsing method provided by the embodiments of the present application.

[0116] Figure 4 (1) is the same as Figure 2 (1), for which reference can be made to the description of Figure 2 (1), and will not be elaborated here.

[0117] Figure 4 (2) is an example of the interface displayed on the mobile phone 1 after the user clicks on the first thumbnail 106 of the original video. Figure 4 (2) Compared with Figure 2 (3), the difference in area B is that the AI video browsing control 109 is no longer displayed. That is to say, the entry for browsing the AI video is no longer the control displayed in area B, but becomes the first thumbnail 131 of the AI video cover displayed in area C.

[0118] Figure 4 In area C of (2), the difference compared with Figure 2 area C of (3) is that between the second thumbnail 111 of the original video and the MM photos 112 - 116, the first thumbnail 131 of the AI video cover is added. And, Figure 4 the selected thumbnail in area C of (2) is no longer centered, but remains in its original position, and is only indicated by the pointer control 117. It can be understood that because Figure 4 (2) is in portrait display, and due to the small screen width, not all of the MM photos are displayed. That is, a part of MM photo 115 is displayed, and MM photo 116 is not displayed. The user can make MM photo 116 be displayed on the screen by operating such as swiping the thumbnail. Figure 4 (2) Compared with Figure 2 (3), for the same display content and controls, reference can be made to the description of Figure 2 (3), and will not be elaborated here.

[0119] As mentioned above, the cover of the AI video can be generated according to the cover information in the configuration file, or can be Figure 4 generated in real time according to the cover generation strategy in (2). No matter which method, the cover is a combination of an image and a display effect. The image can be any one of the MM photos, and of course can also be any frame image in the AI video, or any frame in the original video, etc., which is not limited here.

[0120] To increase the attraction to users, the cover of the AI video is a combination of any one of the MM photos and a display effect, because Figure 2The size of the first thumbnail 131 of the cover of the AI video in (4) is small. Therefore, in some implementations, the display special effect in the first thumbnail 131 of the cover of the AI video is a special effect with a more obvious visual effect. It can be understood that a special effect with a more obvious visual effect can be pre-configured as the target special effect. Before displaying the interface shown in Figure 4 (2), add the target special effect to any MM photo to form the cover of the AI video, and then generate the first thumbnail 131 of the cover of the AI video. The thumbnail method for generating the first thumbnail 131 of the cover of the AI video can refer to the thumbnail methods of thumbnails 111-114, which will not be elaborated here.

[0121] Based on the cover and the generation strategy of the AI video, in some implementations, the cover of the AI video is a frame in the AI video. In other implementations, the cover of the AI video is not included in the AI video. For example, according to the preset special effect addition algorithm, an AI video is generated after adding special effects to the MM photo. Therefore, during the generation process of the AI video, it is possible that the special effect added to the MM photo used for generating the cover of the AI video is different from the target special effect, so the cover of the AI video is not included in the AI video.

[0122] Figure 4 In (2), the original video is played in area B. In this case, if the user clicks on the first thumbnail 131 of the cover of the AI video, the interface displayed on the mobile phone 1 jumps to Figure 4 the interface shown in (3).

[0123] Figure 4 In (3), the title "AI Video" and the return control 108 are displayed in area A. After the user clicks on the return control 108, the interface of the mobile phone 1 jumps to Figure 4 that shown in (2). The cover 132 of the AI video is displayed in area B. Area C is different from Figure 4 (2) in that the pointer control 117 is located at the border of the first thumbnail 131 of the cover of the AI video. In area D, only the "Delete" control is displayed, and no other controls are displayed, indicating that only the cover can be deleted.

[0124] It can be understood that since the AI video has not been generated yet, Figure 4 only the cover 132 of the AI video is displayed in (3), and the AI video is not played.

[0125] In some implementations, in order to guide the user to generate an AI video based on the cover 132 of the AI video, a play control 133 is displayed in the cover 132 of the AI video. It can be understood that the display position and style of the play control 133 are only examples and not limitations.

[0126] After the user clicks the play control 133, the mobile phone 1 generates an AI video and displays it. Figure 4 The AI video generation interface (also known as the loading interface) shown in (4). Figure 4 (4) is the same as Figure 3 (2), which will not be elaborated here.

[0127] After the AI video is generated, the interface of the mobile phone 1 jumps to Figure 4 (5) as shown, Figure 4 (5) is the same as Figure 3 (3), which will not be elaborated here.

[0128] Figure 4 For the video browsing method shown, in the original video browsing interface, a thumbnail of the AI video cover is displayed after the second thumbnail of the original video, and the AI video cover has a display effect. Therefore, it can strengthen the guidance for users and more effectively guide users to browse the AI video.

[0129] It can be understood that the files in the photo gallery on the mobile phone 1 can be synchronized to the cloud and the cloud synchronization of the photo gallery can be realized on other electronic devices such as the mobile phone 2, that is, the files in the photo gallery uploaded to the cloud by the mobile phone 1 are downloaded to the mobile phone 2.

[0130] It can be understood that in some implementation methods, the synchronization upload policy is: photos and other images and videos in the photo gallery of the mobile phone 1 are uploaded to the cloud, and the database of the photo gallery is also uploaded to the cloud. The database includes the identifiers of the videos obtained by the one-shot multi-gain function. In other implementation methods, the synchronization upload policy is: photos and other images and videos in the photo gallery of the mobile phone 1 except for MM photos are uploaded to the cloud, and the database of the photo gallery is also uploaded to the cloud. The upload policy can be set by the user on the mobile phone 1. It can be seen that the original videos obtained by the one-shot multi-gain function will be synchronized to the cloud, while in the case where there is an AI video configuration file in the mobile phone 1, the AI video configuration file will not be synchronized to the cloud.

[0131] It can be understood that during the process of the mobile phone 2 realizing the cloud synchronization of the photo gallery, the mobile phone 2 downloads the first frame of the photos and other image videos in the cloud and the database of the photo gallery. The mobile phone 2 downloads the first frame of the video instead of the complete original video, aiming to save the space and transmission resources of the mobile phone 2.

[0132] In the case where the mobile phone 2 only downloads the first frame of the original video, an example of the GUI involved in the browsing method on the mobile phone 2 is as Figure 5 shown:

[0133] In some implementation methods, the interface of the photo gallery of the mobile phone 2 is as Figure 5 (1) shown, Figure 5 (1) is the same as Figure 2(1) The same. It can be understood that, as mentioned above, since the mobile phone 2 synchronizes the database of the photo gallery of the mobile phone 1, the one-record-multiple-gains identifier of the original video can be obtained from the database. Therefore, in Figure (1), the one-record-multiple-gains identifier 107 can also be displayed in the first thumbnail 106 of the original video.

[0134] After the user clicks on the first thumbnail 106 of the first frame of the original video, the interface of the mobile phone 2 jumps to Figure 5 (2) as shown. Figure 5 (2) The difference from Figure 2 (3) is that: since the mobile phone 2 has not downloaded the original video yet, only the first frame of the original video is displayed in area B, and the original video cannot be played. Therefore, the first playback control 110 is not displayed either, but the playback control 150 is displayed in area B. The cloud synchronization identifier 140 is displayed in area A, indicating that there is an undownloaded part of the currently displayed object.

[0135] It can be understood that in the case where the mobile phone 2 synchronizes the MM photos from the cloud, according to the fact that the identifier of the MM photos recorded in the database is the same as that of the original video, in Figure 5 (2) area C, the MM photos are also displayed. But in the case where the mobile phone 2 has not synchronized the MM photos (i.e., the mobile phone 1 has not uploaded them), there are no MM photos in the mobile phone 2. In this case, Figure 5 (2) there is no area C.

[0136] After the user clicks on Figure 5 the playback control 150 in (2), the mobile phone 2 can display a query message such as "This video is from cloud synchronization and needs to be downloaded for use. Do you want to download the video?" ( Figure 5 not drawn in the figure). If the user selects to download the video, the mobile phone 2 downloads the original video from the cloud. After downloading the original video, the mobile phone 2 jumps to the interface shown in Figure 5 (4), and specifically, it can be referred to as shown in Figure 3 (1).

[0137] After the user clicks on Figure 5 the AI video browsing control 109 in (2), the mobile phone 2 can display a query message such as "This video is from cloud synchronization and needs to be downloaded for use. Do you want to download the video?" ( Figure 5 not drawn in the figure). If the user selects to download the video, the mobile phone 2 downloads the original video from the cloud. It can be understood that the mobile phone 2 can also not display the query message, but in response to the user's click on the AI video browsing control 109, download the original video.

[0138] It can be understood that after the mobile phone 2 downloads the original video, in some implementation manners, the mobile phone 2 generates a new AI video configuration file according to the TAG in the original video, and generates an AI video based on the AI video configuration file. In other implementation manners, after the mobile phone 2 downloads the original video, instead of generating an AI video configuration file, an AI video is generated based on the TAG of the original video and the special effect adding strategy. Since the original video is the same as the original video in the mobile phone 1, therefore, when the user clicks Figure 5 the AI video browsing control 109 shown in (2), the AI video played on the mobile phone 2 is the same as the AI video played on the mobile phone 1 (assuming that the special effect adding algorithm of the mobile phone 2 is the same as that of the mobile phone 1). Figure 5 (3) is the playing interface of the AI video, which is the same as Figure 3 (3), and will not be elaborated here. Of course, the special effect adding algorithm of the mobile phone 2 can be different from that of the mobile phone 1. In this case, the AI video (with special effects) played on the mobile phone 2 is different from that on the mobile phone 1.

[0139] It can be understood that when the user clicks Figure 5 the AI video browsing control 109 shown in (2), before the interface shown in (3) is displayed, a loading interface shown in (2) can also be displayed. Figure 5 (3) is displayed, a loading interface shown in (2) can also be displayed. Figure 3 (2).

[0140] It can be understood that after the user downloads the original video by clicking the play control 150, if the user clicks the AI video browsing control 109 again, since the original video has been downloaded, the AI video can be generated based on the original video, and the specific generation method is as described above.

[0141] If the user chooses not to download the video, the mobile phone 2 still displays Figure 5 the interface shown in (2). Since the mobile phone 2 has not downloaded the original video, the original video and the AI video cannot be browsed on the mobile phone 2.

[0142] From Figure 5 it can be seen that the multimedia content obtained by "one recording, multiple gains" can be synchronized to other electronic devices through the cloud and the user's browsing can be realized, which is beneficial to further improving the user experience.

[0143] In the case where the mobile phone 2 only downloads the first frame of the original video, another example of the GUI involved in the browsing method on the mobile phone 2 is as shown in Figure 6 :

[0144] Figure 6 (1) is the same as Figure 5 (1), and will not be elaborated here.

[0145] After the user clicks the first thumbnail 106 of the original video, the mobile phone 2 jumps to Figure 6 the interface shown in (2).Figure 6 (2) The difference from Figure 4 (2) is that since the original video is not downloaded, the original video cannot be played. Therefore, only the first frame of the original video is displayed, and the first playback control is not shown, but the playback control 150 is shown in area B. The cloud synchronization identifier 140 is shown in area A, indicating that there is an undownloaded part of the currently displayed object.

[0146] It can be understood that in the case where the mobile phone 2 synchronizes the MM photos from the cloud, according to the fact that the identifier of the MM photos recorded in the database is the same as that of the original video, in Figure 6 area C of (2), the MM photos are also shown. But in the case where the mobile phone 2 does not synchronize the MM photos (i.e., the mobile phone 1 does not upload), there are no MM photos in the mobile phone 2. In this case, Figure 6 in area C of (2), only the second thumbnail 111 of the original video and the second thumbnail 134 of the cover of the AI video are shown.

[0147] It can be understood that in this embodiment, the cover of the AI video may be different from Figure 4 the cover of the AI video shown. In this embodiment, since only the first frame of the original video exists in the mobile phone 2 and the AI video configuration file in the mobile phone 1 will not be synchronized by the mobile phone 2, therefore, the cover of the AI video is a combination of the first frame of the original video and the display special effects, and is further reduced to the second thumbnail 134 of the cover of the AI video.

[0148] After the user clicks on Figure 6 the playback control 150 in (2), the mobile phone 2 can display inquiry information such as "This video comes from cloud synchronization and needs to be downloaded for use. Do you want to download the video?" ( Figure 6 not shown in the figure), if the user selects to download the video, then the mobile phone 2 downloads the original video from the cloud. After downloading the original video, the mobile phone 2 can jump to the interface shown in Figure 6 (4), Figure 6 (4) The difference from Figure 4 (2) is that the thumbnails 134 and 131 of the cover of the AI video are different.

[0149] After the user clicks on Figure 6 the second thumbnail 134 of the cover of the AI video in (2), the interface of the mobile phone 2 jumps to the one shown in Figure 6 (3). Figure 6 The cover 135 of the AI video is shown in (3), and it can be seen that Figure 4 it is different from the cover 132 of the AI video on the mobile phone 1 shown in (3).

[0150] After the user clicks on Figure 6After the playback control 133 in (3), the mobile phone 2 directly or in response to an instruction from the user to consent to download the original video, downloads the original video from the cloud. After the original video download is completed, an AI video is generated. Since the original video is the same as the original video in the mobile phone 1, the AI video played on the mobile phone 2 is the same as the AI video played on the mobile phone 1 (assuming that the special effect addition algorithm of the mobile phone 2 is the same as that of the mobile phone 1). Figure 6 (4) is the playback interface of the AI video, which is the same as Figure 4 (3), and will not be elaborated here. Of course, the special effect addition algorithm of the mobile phone 2 can be different from that of the mobile phone 1. In this case, the AI video (with special effects) played on the mobile phone 2 is different from that on the mobile phone 1.

[0151] It can be understood that after the AI video is generated, the cover of the AI video may be different from the cover of the AI video before downloading the original video. In this case, for the sake of consistency, the cover of the AI video on the mobile phone 2 before downloading the original video remains unchanged, that is, keeping Figure 6 134 in (2) and Figure 6 135 in (3) the same as before downloading the original video.

[0152] If the user chooses not to download the original video, the mobile phone 2 maintains Figure 6 the interface shown in (2).

[0153] Figure 7 FIG. is a composition example of an electronic device using the video browsing method provided in the embodiments of the present application. Taking a mobile phone as an example, the electronic device 100 may include a processor 110, an internal memory 120, a display screen 130, a camera 140, an antenna 1, an antenna 2, a mobile communication module 150, a wireless communication module 160, and an audio module 170, etc.

[0154] It can be understood that the structure illustrated in this embodiment does not constitute a specific limitation on the electronic device. In other embodiments, the electronic device may include more or fewer components than those shown in the figure, or combine certain components, or split certain components, or have different component arrangements. The illustrated components may be implemented in hardware, software, or a combination of software and hardware.

[0155] The processor 110 may include one or more processing units. For example, the processor 110 may include an application processor (AP), a modem processor, a graphics processing unit (GPU), an image signal processor (ISP), a controller, a video codec, a digital signal processor (DSP), a baseband processor, and / or a neural-network processing unit (NPU), etc. Among them, different processing units may be independent devices or integrated in one or more processors.

[0156] The internal memory 120 may be used to store computer-executable program code, and the executable program code includes instructions. The processor 110 executes various functional applications and data processing of the electronic device 100 by running the instructions stored in the internal memory 110. The internal memory 120 may include a program storage area and a data storage area. Among them, the program storage area may store an operating system and application programs required for at least one function (such as a sound playback function, an image playback function, etc.). The data storage area may store data created during the use of the electronic device (such as audio data, a phone book, etc.). In addition, the internal memory 120 may include high-speed random access memory and may also include non-volatile memory, such as at least one disk storage device, a flash memory device, a universal flash storage (UFS), etc. The processor 110 executes various functional applications and data processing of the electronic device by running the instructions stored in the internal memory 120 and / or the instructions stored in the memory provided in the processor.

[0157] The electronic device realizes the display function through the GPU, the display screen 130, and the application processor, etc. The GPU is a microprocessor for image processing, connected to the display screen 130 and the application processor. The GPU is used to execute mathematical and geometric calculations for graphics rendering. The processor 110 may include one or more GPUs, which execute program instructions to generate or change display information.

[0158] The display screen 130 is used to display images, videos, etc. The display screen 130 includes a display panel. The display panel can be a liquid crystal display (LCD), an organic light-emitting diode (OLED), an active-matrix organic light-emitting diode (AMOLED), a flexible light-emitting diode (FLED), a Miniled, a MicroLed, a Micro-oled, a quantum dot light-emitting diode (QLED), etc. In some embodiments, the electronic device may include one or N display screens 130, where N is a positive integer greater than 1.

[0159] The electronic device 100 can implement the shooting function through the ISP, the camera 140, the video codec, the GPU, the display screen 194, and the application processor, etc.

[0160] The ISP is used to process the data fed back by the camera 140. For example, when taking a photo, the shutter is opened, and the light passes through the lens and is transmitted to the camera photosensitive element. The optical signal is converted into an electrical signal, and the camera photosensitive element transmits the electrical signal to the ISP for processing and converts it into an image visible to the naked eye. The ISP can also optimize the noise, brightness, and skin color of the image through algorithms. The ISP can also optimize parameters such as the exposure and color temperature of the shooting scene. In some embodiments, the ISP can be provided in the camera 140.

[0161] The camera 140 is used to capture still images or videos. The object generates an optical image through the lens and projects it onto the photosensitive element. The photosensitive element can be a charge-coupled device (CCD) or a complementary metal-oxide-semiconductor (CMOS) phototransistor. The photosensitive element converts the optical signal into an electrical signal, and then transmits the electrical signal to the ISP to convert it into a digital image signal. The ISP outputs the digital image signal to the DSP for processing. The DSP converts the digital image signal into an image signal in standard RGB, YUV, etc. formats. In some embodiments, the electronic device 100 may include one or N cameras 140, where N is a positive integer greater than 1.

[0162] In some embodiments, the camera 140 is used to shoot the videos mentioned in the embodiments of the present application.

[0163] The digital signal processor is used to process digital signals. In addition to processing digital image signals, it can also process other digital signals. For example, when the electronic device 100 selects a frequency point, the digital signal processor is used to perform Fourier transform on the frequency point energy, etc.

[0164] The video codec is used to compress or decompress digital videos. The electronic device 100 can support one or more video codecs. In this way, the electronic device 100 can play or record videos in multiple coding formats, such as: Moving Picture Experts Group (MPEG) 1, MPEG2, MPEG3, MPEG4, etc.

[0165] The internal memory 120 can be used to store computer-executable program code, and the executable program code includes instructions. The processor 110 executes various functional applications and data processing of the electronic device 100 by running the instructions stored in the internal memory 120. The internal memory 120 can include a program storage area and a data storage area. Among them, the program storage area can store an operating system, application programs required for at least one function (such as a sound playback function, an image playback function, etc.). The data storage area can store data created during the use of the electronic device 100 (such as audio data, phone book, etc.). In addition, the internal memory 120 can include high-speed random access memory, and can also include non-volatile memory, such as at least one disk storage device, a flash memory device, a Universal Flash Storage (UFS), etc. The processor 110 executes various functional applications and data processing of the electronic device 100 by running the instructions stored in the internal memory 120, and / or the instructions stored in the memory provided in the processor.

[0166] The electronic device 100 can implement audio functions through the audio module 170, the speaker 170A, the microphone 170B, and the application processor, etc. For example, music playback, recording, etc.

[0167] The audio module 170 is used to convert digital audio information into an analog audio signal for output, and is also used to convert analog audio input into digital audio signals. The audio module 170 can also be used to encode and decode audio signals. In some embodiments, the audio module 170 can be provided in the processor 110, or some functional modules of the audio module 170 can be provided in the processor 110.

[0168] The speaker 170A, also known as the "loudspeaker", is used to convert an audio electrical signal into a sound signal. The electronic device 100 can listen to music or hands-free calls through the speaker 170A.

[0169] In some embodiments, the speaker 170A can play the video information with special effects mentioned in the embodiments of the present application.

[0170] The microphone 170B, also known as a "microphone" or "transmitter", is used to convert sound signals into electrical signals. When making a call or sending a voice message, the user can speak close to the microphone 170B with their mouth to input the sound signal into the microphone 170B.

[0171] In some embodiments, the microphone 170B can collect the sounds of the environment where the electronic device is located during the process of the camera shooting video information with special effects.

[0172] The wireless communication function of the electronic device 100 can be implemented by the antenna 1, antenna 2, mobile communication module 150, wireless communication module 160, modulation and demodulation processor, and baseband processor, etc.

[0173] The antenna 1 and antenna 2 are used to transmit and receive electromagnetic wave signals.

[0174] The mobile communication module 150 can provide solutions for wireless communications such as 2G / 3G / 4G / 5G applied to the electronic device 100. The mobile communication module 150 can include at least one filter, switch, power amplifier, low noise amplifier (LNA), etc. The mobile communication module 150 can receive electromagnetic waves by the antenna 1, filter, amplify, etc. the received electromagnetic waves, and transmit them to the modulation and demodulation processor for demodulation. The mobile communication module 150 can also amplify the signal modulated by the modulation and demodulation processor and convert it into electromagnetic waves through the antenna 1 and radiate it out.

[0175] The wireless communication module 160 may provide solutions for wireless communications applied to the electronic device 100, including wireless local area networks (WLANs) (such as wireless fidelity (Wi-Fi) networks), Bluetooth (BT), global navigation satellite systems (GNSS), frequency modulation (FM), near field communication (NFC), infrared technology (IR), etc. The wireless communication module 160 may be one or more devices integrating at least one communication processing module. The wireless communication module 160 receives electromagnetic waves via the antenna 2, performs frequency modulation and filtering processing on the electromagnetic wave signals, and sends the processed signals to the processor 110. The wireless communication module 160 may also receive the signals to be sent from the processor 110, perform frequency modulation and amplification on them, and convert them into electromagnetic waves through the antenna 2 for radiation.

[0176] In addition, an operating system runs on the above components. For example, iOS operating system, Android operating system, Windows operating system, etc. Application programs can be installed and run on the operating system.

[0177] Figure 8 It is a software structure block diagram of the electronic device according to the embodiments of the present application.

[0178] The layered architecture divides the software into several layers, and each layer has a clear role and division of labor. The layers communicate with each other through software interfaces. In some embodiments, the Android system is divided into four layers, from top to bottom, namely the application layer, the application framework layer, the Android runtime and system libraries, and the kernel layer.

[0179] The application layer may include a series of application packages. As Figure 8 shown, the application packages may include applications such as camera, gallery, calendar, call, map, navigation, video editing, etc.

[0180] In some embodiments, the camera is used to shoot the video described in the foregoing embodiments and implement the foregoing multi-recording function.

[0181] The application framework layer provides application programming interfaces (APIs) and programming frameworks for the applications in the application layer. The application framework layer includes some predefined functions. As Figure 8As shown in the figure, the application framework layer may include a Media Library, a window manager, a content provider, a phone manager, a resource manager, a notification manager, a view system, etc.

[0182] In some embodiments, the Media Library is used to store multimedia content.

[0183] The window manager is used to manage window programs. The window manager can obtain the display screen size, determine whether there is a status bar, lock the screen, capture the screen, etc.

[0184] The content provider is used to store and obtain data, and make this data accessible to application programs. The data may include videos, images, audio, incoming and outgoing calls, browsing history and bookmarks, phone books, etc.

[0185] The phone manager is used to provide the communication function of the electronic device. For example, the management of call status (including connection, disconnection, etc.).

[0186] The resource manager provides various resources for application programs, such as localized strings, icons, pictures, layout files, video files, and so on.

[0187] The notification manager enables application programs to display notification information in the status bar. It can be used to convey notification-type messages, which can automatically disappear after a short stay without user interaction. For example, the notification manager is used to notify that the download is complete, message reminders, etc. The notification manager can also be a notification that appears in the system top status bar in the form of a chart or scroll bar text, such as the notification of a background-running application program, or a notification that appears on the screen in the form of a dialogue window. For example, prompt text information in the status bar, emit a prompt tone, the electronic device vibrates, the indicator light flashes, etc.

[0188] The view system includes visible controls, such as controls for displaying text, controls for displaying pictures, etc. The view system can be used to build application programs. The display interface can be composed of one or more views. For example, a display interface including a short message notification icon may include a view for displaying text and a view for displaying pictures.

[0189] Android Runtime includes a core library and a virtual machine. Android runtime is responsible for the scheduling and management of the Android system. In some embodiments of this application, the cold start of the application will run in Android runtime, and Android runtime can obtain the optimized file status parameters of the application from this. Furthermore, Android runtime can determine whether the optimized file is outdated due to system upgrade through the optimized file status parameters, and return the judgment result to the application control module.

[0190] The core library consists of two parts: one part is the functional functions that need to be called by the Java language, and the other part is the core library of Android.

[0191] The application layer and the application framework layer run in the virtual machine. The virtual machine executes the Java files in the application layer and the application framework layer as binary files. The virtual machine is used to perform functions such as management of object life cycles, stack management, thread management, security and exception management, and garbage collection.

[0192] The system library can include multiple functional modules. For example: surface manager, Media Provider, 3D graphics processing library (such as: OpenGL ES), 2D graphics engine (such as: SGL), etc.

[0193] The surface manager is used to manage the display subsystem and provides the fusion of 2D and 3D layers for multiple applications.

[0194] Media Provider can send data to the Media Library. In some embodiments, the Media Library can set a Uniform Resource Identifier (URI) to listen for data from the Media Provider. If the data stored in the Media Provider changes, the Media Provider sends the data to the Media Library.

[0195] The 3D graphics processing library is used to implement 3D graphics drawing, image rendering, synthesis, and layer processing, etc.

[0196] The 2D graphics engine is a drawing engine for 2D drawing.

[0197] The kernel layer is the layer between the hardware and the software. The kernel layer at least includes a display driver, a camera driver, an audio driver, a sensor driver, etc.

[0198] It should be noted that although the embodiments of this application are described by taking the Android system as an example, its basic principles are equally applicable to electronic devices based on operating systems such as iOS and Windows.

[0199] Another embodiment of this application also provides a computer-readable storage medium. Instructions are stored in the computer-readable storage medium. When it runs on a computer or a processor, it causes the computer or the processor to execute one or more steps in any of the above methods.

[0200] Another embodiment of the present application further provides a computer program product containing instructions. When the computer program product runs on a computer or a processor, it causes the computer or the processor to execute one or more steps in any of the above methods.

Claims

1. A method for browsing a video, characterized in that, applied to an electronic device, the method comprising: displaying a first interface, the first interface displaying a first frame image in a first video and a browsing control for a second video, the browsing control for the second video comprising: a thumbnail of a first cover of the second video, the first cover of the second video comprising a combination of the first frame image in the first video and special effects, the first video including a tag TAG, the second video being obtained by splicing segments in chronological order and adding special effects, the segments being segments including exciting picture moments intercepted from the first video based on the TAG; in response to an operation on the browsing control for the second video, downloading the first video; generating the second video and a second cover for the second video based on the TAG, the second cover being different from the first cover; using the first cover as the cover for the second video; displaying a browsing interface for the second video, and playing the second video in the browsing interface for the second video.

2. The method according to claim 1, characterized in that, the displaying of the first interface, the first interface displaying a first frame image in a first video and a browsing control for a second video, comprises: in response to an operation on the first frame image in the first video downloaded from the cloud, displaying the first interface, the first interface displaying the first frame image in the first video and the browsing control for the second video.

3. The method according to claim 1, characterized in that, the process of obtaining the first video comprises: acquiring a video through a camera; determining that the video meets a first condition or a second condition, the first condition comprising that the duration of the video is greater than a first duration and an exciting moment is recognized in the video; generating a TAG based on the theme of the video, the exciting moment, and the scene recognized in the video; writing the TAG into the video to generate the first video.

4. The method according to claim 3, characterized in that, the process of obtaining the first video comprises: acquiring a video through a camera; determining that the video meets a second condition, the second condition comprising that the duration of the video is greater than a second duration and the video includes video frames meeting a quality condition; generating a TAG based on the theme of the video, the video frames meeting the quality condition, and the scene recognized in the video; writing the TAG into the video to generate the first video.

5. The method according to any one of claims 1-4, characterized in that, the first interface displaying a first frame image in a first video and a browsing control for a second video, comprises: the first interface includes a playing area, and the first video is played in the playing area; the browsing control for the second video is displayed in the playing area.

6. The method according to any one of claims 1-4, characterized in that, the first interface further includes: a thumbnail area, the thumbnail area at least displaying thumbnails of one or more frames in the first video.

7. The method according to any one of claims 1-4, characterized in that, The first interface displays the first frame image in the first video and browsing controls for the second video, including: The first interface includes a thumbnail area and a playback area. The thumbnail area displays thumbnails of the covers of the second video, which serve as browsing controls for the second video, and the playback area plays the first video.

8. The method according to claim 7, wherein, The thumbnail area further displays: A thumbnail of a frame image in the first video.

9. The method according to claim 7, wherein, Playing the second video includes: Displaying a second interface, which includes the cover and playback controls; In response to an operation on the playback controls, displaying a browsing interface for the second video and playing the second video in the browsing interface.

10. The method according to any one of claims 1 - 4, wherein, The thumbnail area further includes: Thumbnails of exciting moment photos in the first video.

11. The method according to any one of claims 1 - 4, wherein, Playing the second video includes: Playing the second video in the browsing interface of the second video; The browsing interface of the second video includes: Special effect selection controls.

12. The method according to any one of claims 1 - 4, wherein, Before playing the second video, it further includes: Displaying a loading interface for the second video, and the loading interface includes a loading prompt control.

13. The method according to any one of claims 1 - 4, wherein, Before displaying the first interface, it further includes: Displaying a gallery interface, which displays a thumbnail of the first video, and the thumbnail of the first video includes an identifier indicating that the first video has an associated second video.

14. An electronic device, wherein, It includes: One or more processors, a memory, a display screen, and a camera; The memory, the display screen, and the camera are coupled to the one or more processors. The memory is used to store computer program code, and the computer program code includes computer instructions. When the one or more processors execute the computer instructions, the electronic device executes the video browsing method according to any one of claims 1 to 13.

15. A computer storage medium, wherein, It is used to store a computer program, and when the computer program is executed, it is specifically used to implement the video browsing method according to any one of claims 1 to 13.

Citation Information

Patent Citations

  • Video recording method and video recording device

    CN105979188A

  • Method for processing video file and electronic equipment

    CN111061912A