Video Processing Method, Electronic Device, and Readable Medium
By setting the 'one record and multiple results' mode in the electronic device, the recognition model is used to automatically extract wonderful images in the video stream, which solves the problem that users find it difficult to capture wonderful moment photos when shooting videos, and generate wonderful moment photos and short videos while shooting videos, improving the user experience.
Patent Information
- Application Number
- CN202210187220.4
- Authority / Receiving Office
- CN · China
- Patent Type
- Patents(China)
- Current Assignee / Owner
- Filing Date
- 2022-02-28
- Publication Date
- 2025-07-11
- Estimated Expiration
- 2042-02-28
AI Technical Summary
It is difficult for users to capture memorable photos of wonderful moments when shooting videos, and the existing technology has not effectively solved this need.
Provide a video processing method. By setting the 'one record and multiple results' mode in an electronic device, using the recognition model to automatically extract wonderful images in the video stream and generate wonderful moment photos and short videos. Users can obtain wonderful moment photos and short videos while shooting videos.
It realizes the automatic capture of memorable photos of wonderful moments and generates wonderful short videos while shooting videos, improving the user experience and meeting the users' needs to capture wonderful moments during video shooting.
Smart Images

Figure CN116708649B_ABST
Abstract
Description
Technical Field
[0001] This application relates to the technical field of electronic devices, and in particular, to a video processing method, an electronic device, a program product, and a computer-readable storage medium. Background Art
[0002] Currently, the functions of taking pictures and recording videos have become essential functions of electronic devices. Users' demands and experiences for recording and taking pictures are also constantly increasing. In some application scenarios of video recording, users expect to capture memorable and wonderful instant photos while recording videos.
[0003] Therefore, a method for obtaining memorable and wonderful instant photos while recording videos is needed. Summary of the Invention
[0004] This application provides a video processing method, an electronic device, a program product, and a computer-readable storage medium, aiming to enable users to obtain wonderful instant photos while recording videos.
[0005] To achieve the above object, this application provides the following technical solutions:
[0006] In a first aspect, this application provides a video processing method applied to an electronic device. The video processing method includes: in response to a first operation, shooting a first video; in response to a second operation, displaying a first interface, where the first interface is a details interface of the first video, the first interface includes a first area, a second area, and a first control, or the first interface includes a first area and a second area, or the first interface includes a first area and a first control, the first area is a playback area of the first video, the second area displays a cover thumbnail of the first video, thumbnails of a first image and a second image, the first image is an image of the first video at a first moment, the second image is an image of the first video at a second moment, and the recording process of the first video includes the first moment and the second moment; the first control is used to control the electronic device to generate a second video, the duration of the second video is less than that of the first video, and the second video includes at least images of the first video.
[0007] It can be seen from the above that: the user uses the electronic device to shoot a video, and the electronic device can obtain the shot first video, as well as the first image and the second image, realizing that the user obtains wonderful instant photos while shooting the video.
[0008] In a possible implementation manner, the video processing method further includes: in response to a third operation, displaying a second interface, where the third operation is a touch operation on the first control, and the second interface is a display interface of the second video.
[0009] In a possible implementation, after the user takes a video, the electronic device obtains a first video, a first image, and a second image. The electronic device may also obtain a second video and display it. The duration of the second video is less than that of the first video and includes images of the first video, enabling the user to obtain wonderful instant photos while shooting a video and further obtain a short video of the first video for convenient sharing by the user.
[0010] In a possible implementation, in response to a second operation, a first interface is displayed, including: in response to a fourth operation, a third interface is displayed. The third interface is the interface of the gallery application and includes a cover thumbnail of the first video; in response to a touch operation on the cover thumbnail of the first video, the first interface is displayed.
[0011] In a possible implementation, in response to a second operation, a first interface is displayed, including: in response to a touch operation on a second control, the first interface is displayed. The shooting interface of the electronic device includes the second control, and the second control is used to control the display of the previously captured image or video.
[0012] In a possible implementation, the cover thumbnail of the first video includes a first identifier, and the first identifier is used to indicate that the first video is shot by the electronic device in the one-shot-multiple-results mode.
[0013] In a possible implementation, a mask layer is displayed on the first interface, and a second area is not covered by the mask layer.
[0014] In a possible implementation, the first interface further includes: a first dialog box. The first dialog box is used to prompt the user that the first image and the second image have been generated, and the first dialog box is not covered by the mask layer.
[0015] In a possible implementation, after shooting the first video in response to a first operation, the video processing method further includes: in response to a fifth operation, the shooting interface of the electronic device is displayed. The shooting interface includes: a second dialog box, and the second dialog box is used to prompt the user that the first video and the second video have been generated.
[0016] In a possible implementation, before shooting the first video in response to a first operation, the video processing method further includes: in response to a sixth operation, a fourth interface is displayed. The fourth interface is the shooting settings interface, and the fourth interface includes: an option for one-shot-multiple-results and a text segment. The option for one-shot-multiple-results is used to control the electronic device to turn on or off the one-shot-multiple-results function, and the text segment is used to indicate the function content of the one-shot-multiple-results.
[0017] In one possible embodiment, after shooting the first video in response to the first operation, the video processing method also includes: in response to the seventh operation, displaying the fifth interface, the fifth interface is the interface of the gallery application, the fifth interface includes: a first folder and a second folder, the first folder includes images and videos saved by the electronic device, and the second folder includes the first image and the second image; in response to the eighth operation, displaying the sixth interface, the sixth interface includes a thumbnail of the first image and a thumbnail of the second image, and the eighth operation is a touch operation on the second folder.
[0018] In one possible embodiment, after displaying the second interface in response to the third operation, the video processing method also includes: displaying a seventh interface in response to a ninth operation, the seventh interface being a details interface of the second video, the ninth operation being a touch operation on a third control included in the second interface, the third control being used to control saving of the second video.
[0019] In a possible implementation, the video processing method further includes: in response to the tenth operation, displaying an eighth interface, the eighth interface being an interface of a gallery application, the eighth interface including: a cover thumbnail of the second video and a cover thumbnail of the first video.
[0020] In one possible implementation, after shooting the first video in response to the first operation, the video processing method further includes: in response to the eleventh operation, displaying a first shooting interface of the electronic device, the first shooting interface including a first option and a second option, the first option being used to indicate a photo taking mode, and the second option being used to indicate a video recording mode; in response to an operation on a fourth control of the shooting interface, displaying the first shooting interface of the electronic device, the fourth control being used to start taking photos; in response to an operation on the second option, displaying a second shooting interface of the electronic device, the second shooting interface including a third dialog box, the third dialog box being used to indicate to the user the functional content of "recording in one go".
[0021] In this possible implementation, after the user shoots a video, if the user controls the electronic device to take a photo by touching the fourth control, when the electronic device enters the shooting interface of the electronic device to shoot a video again, a third dialog box can be displayed on the shooting interface to remind the user that the electronic device is configured with a multiple-recording function.
[0022] In a possible implementation, in response to the first operation, during the process of shooting the first video, the video processing method further includes: in response to a twelfth operation, shooting and saving a third image; the twelfth operation is a touch operation on the camera key of the video shooting interface of the electronic device.
[0023] In this possible implementation, when the electronic device is shooting a video, the electronic device also responds to the twelfth operation to shoot and save a third image, thereby configuring a snapshot image function for the electronic device when shooting a video.
[0024] In a possible implementation, the second area further displays a thumbnail of a third image, and the second video includes the third image.
[0025] In this possible implementation, the image captured manually by the electronic device can be used to obtain the second video, so as to implement using the image captured by the user as an image in the second video.
[0026] In a possible implementation, the second area displays a cover thumbnail of the first video, a thumbnail of the first image, and a thumbnail of the second image, including: the second area displays a cover thumbnail of the first video, a thumbnail of the first image, and a thumbnail of the third image; the second video includes at least the first image and the third image.
[0027] In a possible implementation, the generation method of the second video includes: obtaining the first video and the tag data of the first video, where the tag data includes the theme TAG, the storyboard TAG, the first image TAG, and the second image TAG of the first video; determining a style template and music based on the theme TAG of the first video, where the style template includes at least one special effect; obtaining multiple frames of images before and after the first image, and multiple frames of images before and after the second image from the first video based on the storyboard TAG, the first image TAG, and the second image TAG; synthesizing the special effects, music, and target images of the style template to obtain the second video; the target images include at least: the first image and multiple frames of images before and after the first image.
[0028] In a second aspect, the present application provides a video processing method applied to an electronic device. The video processing method includes: in response to a first operation, displaying a first interface and starting to shoot a first video, where the first interface is a preview interface when shooting the first video, and the first interface includes a first control for shooting an image; in response to a second operation, shooting and saving a first image during the process of shooting the first video; the second operation is a touch operation on the first control; after the shooting of the first video is completed, in response to a third operation, displaying a second interface, where the second interface is a details interface of the first video, and the second interface includes a first area, a second area, and a first control, or the second interface includes a first area and a second area, or the second interface includes a first area and a first control; the first area is a playing area of the first video, the second area displays a cover thumbnail of the first video and a thumbnail of the first image, and the first control is used to control the electronic device to generate a second video, where the duration of the second video is less than that of the first video, and the second video includes at least the images of the first video.
[0029] As can be seen from the above: when the user uses the electronic device to shoot a video, the first image can be shot and saved in response to the second operation, realizing the capture while shooting the video. The electronic device can obtain the first video shot, as well as the first image and the second image, realizing the user's capture of an image while shooting the video.
[0030] In a possible implementation, the second area also displays thumbnails of one or more other frames of images, where the one or more other frames of images are images in the first video, and the sum of the number of the first image and the one or more other frames of images is greater than or equal to a preset number, and the preset number is the number of second images automatically recognized by the electronic device during the shooting of the first video.
[0031] In a possible implementation, the second video includes at least one or more frames of images among the following images: the first image, one or more other frames of images.
[0032] In a possible implementation, the video processing method further includes: in response to a fourth operation, displaying a third interface, where the fourth operation is a touch operation on the first control, and the third interface is a display interface of the second video.
[0033] In a possible implementation, after the shooting of the first video is completed, the video processing method further includes: displaying a first shooting interface of the electronic device, where the shooting interface includes a first option and a second option, the first option is used to indicate the photo-taking mode, and the second option is used to indicate the video-recording mode; the first shooting interface is a preview interface when shooting an image; in response to an operation on the second control of the shooting interface, displaying the first shooting interface of the electronic device, where the second control is used to start taking a photo; in response to an operation on the second option, displaying a second shooting interface of the electronic device, where the second shooting interface includes a first dialog box, and the first dialog box is used to indicate the function content of getting multiple recordings for one shot to the user, and the second shooting interface is a preview interface when shooting a video.
[0034] In a possible implementation, after the shooting of the first video is completed, it further includes: in response to a sixth operation, displaying a third interface, where the third interface is an interface of the gallery application, and the third interface includes: a first folder and a second folder, the first folder includes at least the first image, and the second folder includes the second image and the third image, or the second folder includes the second image; in response to a seventh operation, displaying a fourth interface, where the fourth interface includes thumbnails of the second image and the third image, or includes a thumbnail of the second image, and the seventh operation is a touch operation on the second folder.
[0035] In a third aspect, the present application provides an electronic device, including: one or more processors, a memory, a camera, and a display screen; the memory, the camera, and the display screen are coupled to the one or more processors, and the memory is used to store computer program code, the computer program code includes computer instructions, when the one or more processors execute the computer instructions, the electronic device executes the video processing method of any item in the first aspect, or the video processing method of any item in the second aspect.
[0036] In a fourth aspect, the present application provides a computer-readable storage medium for storing a computer program, when the computer program is executed by an electronic device, the electronic device is enabled to implement the video processing method of any item in the first aspect, or the video processing method of any item in the second aspect.
[0037] In a fifth aspect, the present application provides a computer program product, when the computer program product runs on a computer, the computer is enabled to execute the video processing method of any item in the first aspect, or the video processing method of any item in the second aspect. BRIEF DESCRIPTION OF THE DRAWINGS
[0038] Figure 1 It is a hardware structure diagram of the electronic device provided by the present application;
[0039] Figure 2 It is a schematic diagram of enabling "one recording, multiple gains" provided by Embodiment 1 of the present application;
[0040] Figure 3 It is a schematic diagram of a graphical user interface of "one recording, multiple gains" provided by Embodiment 1 of the present application;
[0041] Figure 4 It is another schematic diagram of a graphical user interface of "one recording, multiple gains" provided by Embodiment 1 of the present application;
[0042] Figure 5 It is another schematic diagram of a graphical user interface of "one recording, multiple gains" provided by Embodiment 1 of the present application;
[0043] Figure 6 It is another schematic diagram of a graphical user interface of "one recording, multiple gains" provided by Embodiment 1 of the present application;
[0044] Figure 7 It is another schematic diagram of a graphical user interface of "one recording, multiple gains" provided by Embodiment 1 of the present application;
[0045] Figure 8 It is a flowchart of generating a refined video provided by Embodiment 1 of the present application;
[0046] Figure 9 It is a display diagram of the refined video generated by Embodiment 1 of the present application;
[0047] Figure 10 This is an exemplary display diagram of generating a selected video provided in the first embodiment of the present application;
[0048] Figure 11 This is a schematic diagram of another graphical user interface of "one recording, multiple gains" provided in the first embodiment of the present application;
[0049] Figure 12 This is a schematic diagram of a graphical user interface of "one recording, multiple gains" provided in the second embodiment of the present application;
[0050] Figure 13 This is a schematic diagram of a graphical user interface of "one recording, multiple gains" provided in the third embodiment of the present application;
[0051] Figure 14 This is a schematic diagram of another graphical user interface of "one recording, multiple gains" provided in the third embodiment of the present application. Detailed implementation manners
[0052] Next, the technical solutions in the embodiments of the present application will be clearly and completely described in conjunction with the accompanying drawings in the embodiments of the present application. The terms used in the following embodiments are only for the purpose of describing specific embodiments and are not intended to limit the present application. As used in the specification and the appended claims of the present application, the singular forms "a", "an", "the", "above-mentioned", "said", and "this" are also intended to include, for example, the expression form of "one or more", unless clearly indicated to the contrary in the context. It should also be understood that in the embodiments of the present application, "one or more" means one, two, or more than two; " / ", which describes the association relationship of associated objects, indicates that three relationships can exist; for example, A and / or B can represent: A exists alone, A and B exist simultaneously, and B exists alone, where A and B can be singular or plural. The character " / " generally represents an "or" relationship between the associated objects before and after.
[0053] Referring to "one embodiment" or "some embodiments" described in this specification means that specific features, structures, or characteristics described in conjunction with the embodiment are included in one or more embodiments of the present application. Thus, the statements "in one embodiment", "in some embodiments", "in other some embodiments", "in still other embodiments", etc. that appear in different parts of this specification do not necessarily refer to the same embodiment, but mean "one or more but not all embodiments", unless otherwise specifically emphasized in other ways. The terms "include", "comprise", "have" and their variants all mean "including but not limited to", unless otherwise specifically emphasized in other ways.
[0054] The "multiple" involved in the embodiments of the present application means greater than or equal to two. It should be noted that in the description of the embodiments of the present application, terms such as "first" and "second" are only used for the purpose of distinguishing descriptions, and cannot be understood as indicating or implying relative importance, nor can they be understood as indicating or implying an order.
[0055] Before introducing the embodiments of the present application, some terms or concepts involved in the embodiments of the present application are first explained. It should be understood that the present application does not make specific limitations on the naming of the following terms. The following terms may have other names. The re-named terms still satisfy the following related term explanations.
[0056] 1) One-shot multi-get can be understood as a function that when the user uses the camera application to shoot a video, by pressing the "shoot" icon once, the user can obtain the original video taken, one or more wonderful photos, and one or more selected videos. It can be understood that the duration of the wonderful short video obtained through one-shot multi-get is less than the duration of the entire complete video. For example, if the entire complete video recorded is 1 minute, 5 wonderful moment photos and a wonderful short video with a duration of 15 seconds can be obtained. It can also be understood that one-shot multi-get may have other names, such as one-key multi-get, one-key multi-shot, one-key output, one-key blockbuster, AI one-key blockbuster, etc.
[0057] 2) Wonderful images refer to the pictures of some wonderful moments during the video recording process. For example, wonderful images can be the best sports moment pictures, the best expression moment pictures, or the best check-in action pictures. It should be understood that the present application does not limit the term wonderful images, and wonderful images can also be called beautiful moment images, magical moment images, wonderful instant images, decisive moment images, best shot (BS) images, or AI images, etc. In different scenarios, wonderful images can be different types of instant pictures. For example, when shooting a football game video, the wonderful images can be the images of the moment when the athlete's foot touches the football during a shot or a pass, the image of the football being kicked away by the athlete, the image of the football flying into the goal, or the image of the goalkeeper catching the football. When shooting a video of a person jumping from the ground, the wonderful images can be the image of the person at the highest point in the air, or the image of the person with the most stretched body in the air. When shooting a scenery, the wonderful images can be the images of buildings appearing in the scenery, or the images of the sunset or sunrise.
[0058] 3) Selected videos refer to videos that contain wonderful images. It should be understood that the present application does not limit the term selected videos either, and selected videos can also be called wonderful videos, wonderful short videos, wonderful small videos, or AI videos, etc.
[0059] 4) Tags (TAGs), which can be divided into theme TAGs, scene TAGs, wonderful image TAGs, and shot TAGs, etc.; theme TAGs are used to indicate the style or atmosphere of the video; scene TAGs are used to indicate the scenes of the video; wonderful image TAGs are used to indicate the positions of wonderful images in the captured video, and shot TAGs are used to indicate the positions of scene transitions in the captured video. For example, in a video including one or more wonderful image TAGs, the wonderful image TAGs can indicate that the image frames at the 10th second, 1 minute and 20 seconds, etc. of the video are wonderful images. The video also includes one or more shot TAGs, and the shot TAGs can indicate that the first scene switches to the second scene at the 15th second of the video, and the second scene switches to the third scene at 3 minutes and 43 seconds.
[0060] Currently, the functions of taking photos and recording videos have become essential functions of electronic devices. Users' demands and experiences for recording and taking photos are also continuously increasing. In some video shooting application scenarios, users expect to capture memorable wonderful moment photos while shooting videos. Based on this, this application sets a "multiple gains in one recording" mode in the electronic device, that is, when the electronic device shoots a video in the video recording mode, by analyzing the video stream, wonderful images in the video stream are automatically extracted. And when the video shooting is completed, the electronic device can also generate one or more selected videos based on the wonderful images. In addition, users can view the captured videos, wonderful images, and selected videos in the gallery.
[0061] To support the "multiple gains in one recording" mode of the electronic device, an embodiment of this application provides a video processing method. And a video processing method provided by an embodiment of this application can be applicable to mobile phones, tablet computers, desktop computers, laptop computers, Ultra-mobile Personal Computers (UMPCs), handheld computers, netbooks, Personal Digital Assistants (PDAs), wearable electronic devices, smart watches, etc.
[0062] Taking a mobile phone as an example, Figure 1 This is a composition example of an electronic device provided by an embodiment of this application. As Figure 1 shown, the electronic device 100 may include a processor 110, an internal memory 120, a camera 130, a display screen 140, an antenna 1, an antenna 2, a mobile communication module 150, a wireless communication module 160, an audio module 170, a sensor module 180, and a key 190, etc.
[0063] It can be understood that the structure illustrated in this embodiment does not constitute a specific limitation on the electronic device 100. In other embodiments, the electronic device 100 may include more or fewer components than those illustrated, or combine certain components, or split certain components, or have different component arrangements. The illustrated components may be implemented in hardware, software, or a combination of software and hardware.
[0064] The processor 110 may include one or more processing units. For example, the processor 110 may include an application processor (AP), a modem processor, a graphics processing unit (GPU), an image signal processor (ISP), a controller, a video codec, a digital signal processor (DSP), a baseband processor, a sensor hub, and / or a neural-network processing unit (NPU), etc. Among them, different processing units may be independent devices or integrated in one or more processors.
[0065] A memory may also be provided in the processor 110 for storing instructions and data. In some embodiments, the memory in the processor 110 is a cache memory. This memory may store the instructions or data that the processor 110 has just used or recycled. If the processor 110 needs to use the instruction or data again, it can directly call it from the memory. This avoids repeated accesses, reduces the waiting time of the processor 110, and thus improves the efficiency of the system.
[0066] The internal memory 120 can be used to store computer-executable program codes, and the executable program codes include instructions. The processor 110 executes various functional applications and data processing of the electronic device 100 by running the instructions stored in the internal memory 120. The internal memory 120 can include a program storage area and a data storage area. Among them, the program storage area can store an operating system, application programs required for at least one function (such as a sound playback function, an image playback function, etc.). The data storage area can store data created during the use of the electronic device 100 (such as audio data, a phone book, etc.). In addition, the internal memory 120 can include a high-speed random access memory, and can also include a non-volatile memory, such as at least one disk storage device, a flash memory device, a universal flash storage (UFS), etc. The processor 110 executes various functional applications and data processing of the electronic device 100 by running the instructions stored in the internal memory 120, and / or the instructions stored in the memory provided in the processor.
[0067] In some embodiments, the instructions stored in the internal memory 120 are for executing a video processing method. The processor 110 can control the electronic device to shoot a video in the "one recording, multiple gains" mode by executing the instructions stored in the internal memory 120, so as to obtain the shot video, one or more wonderful photos, and one or more selected video segments.
[0068] The electronic device realizes the display function through the GPU, the display screen 140, and the application processor, etc. The GPU is a microprocessor for image processing, and is connected to the display screen 140 and the application processor. The GPU is used to execute mathematical and geometric calculations for graphics rendering. The processor 110 can include one or more GPUs, which execute program instructions to generate or change display information.
[0069] The display screen 140 is used to display images, videos, etc. The display screen 140 includes a display panel. The display panel can be a liquid crystal display (LCD), an organic light-emitting diode (OLED), an active-matrix organic light-emitting diode (AMOLED), a flexible light-emitting diode (FLED), a MiniLED, a MicroLED, a Micro-OLED, a quantum dot light-emitting diode (QLED), etc. In some embodiments, the electronic device may include one or N display screens 140, where N is a positive integer greater than 1.
[0070] In some embodiments, the electronic device shoots a video in the "one recording, multiple gains" mode to obtain the shot video, one or more wonderful photos, and one or more selected video segments, which are displayed to the user by the display screen 140.
[0071] The electronic device 100 can implement the shooting function through an ISP, a camera 130, a video codec, a GPU, a display screen 140, an application processor, etc.
[0072] The ISP is used to process the data fed back by the camera 130. For example, when taking a photo, the shutter is opened, and the light passes through the lens and is transmitted to the camera sensor. The light signal is converted into an electrical signal, and the camera sensor transmits the electrical signal to the ISP for processing and converts it into an image visible to the naked eye. The ISP can also optimize the noise, brightness, and skin color of the image through algorithms. The ISP can also optimize parameters such as the exposure and color temperature of the shooting scene. In some embodiments, the ISP can be set in the camera 130.
[0073] The camera 130 is used to capture still images or videos. An object generates an optical image through a lens and projects it onto a photosensitive element. The photosensitive element can be a charge coupled device (CCD) or a complementary metal-oxide-semiconductor (CMOS) phototransistor. The photosensitive element converts the optical signal into an electrical signal, and then transmits the electrical signal to the ISP to be converted into a digital image signal. The ISP outputs the digital image signal to the DSP for processing. The DSP converts the digital image signal into an image signal in a standard format such as RGB or YUV. In some embodiments, the electronic device 100 may include one or N cameras 130, where N is a positive integer greater than 1.
[0074] In some embodiments, the camera 130 is used to shoot the videos mentioned in the embodiments of the present application.
[0075] The digital signal processor is used to process digital signals. In addition to processing digital image signals, it can also process other digital signals. For example, when the electronic device 100 selects a frequency point, the digital signal processor is used to perform Fourier transform on the frequency point energy, etc.
[0076] The video codec is used to compress or decompress digital videos. The electronic device 100 may support one or more video codecs. In this way, the electronic device 100 can play or record videos in multiple coding formats, such as: Moving Picture Experts Group (MPEG) 4, MPEG2, MPEG3, MPEG4, etc.
[0077] The wireless communication function of the electronic device 100 can be implemented through antenna 1, antenna 2, mobile communication module 150, wireless communication module 160, modulation and demodulation processor, and baseband processor, etc.
[0078] Antenna 1 and antenna 2 are used to transmit and receive electromagnetic wave signals. Each antenna in the electronic device 100 can be used to cover a single or multiple communication frequency bands. Different antennas can also be multiplexed to improve the utilization rate of the antennas. For example: Antenna 1 can be multiplexed as a diversity antenna for a wireless local area network. In some other embodiments, the antenna can be used in combination with a tuning switch.
[0079] The mobile communication module 150 may provide solutions for wireless communications such as 2G / 3G / 4G / 5G applied to the electronic device 100. The mobile communication module 150 may include at least one filter, switch, power amplifier, low noise amplifier (LNA), etc. The mobile communication module 150 may receive electromagnetic waves through the antenna 1, filter and amplify the received electromagnetic waves, and then transmit them to the modulation and demodulation processor for demodulation. The mobile communication module 150 may also amplify the signal modulated by the modulation and demodulation processor and convert it into electromagnetic waves through the antenna 1 for radiation. In some embodiments, at least some functional modules of the mobile communication module 150 may be provided in the processor 110. In some embodiments, at least some functional modules of the mobile communication module 150 and at least some modules of the processor 110 may be provided in the same device.
[0080] The wireless communication module 160 may provide solutions for wireless communications applied to the electronic device 100, including wireless local area networks (WLAN) (such as wireless fidelity (Wi-Fi) networks), Bluetooth (BT), global navigation satellite system (GNSS), frequency modulation (FM), near field communication (NFC), infrared technology (IR), etc. The wireless communication module 160 may be one or more devices integrating at least one communication processing module. The wireless communication module 160 receives electromagnetic waves through the antenna 2, frequency-modulates and filters the electromagnetic wave signals, and sends the processed signals to the processor 110. The wireless communication module 160 may also receive the signals to be sent from the processor 110, frequency-modulate and amplify them, and convert them into electromagnetic waves through the antenna 2 for radiation.
[0081] The electronic device may implement audio functions through the audio module 170, speaker 170A, receiver 170B, microphone 170C, headphone jack 170D, and the application processor, etc. Such as music playback, recording, etc.
[0082] The audio module 170 is used to convert digital audio information into analog audio signals for output, and is also used to convert analog audio inputs into digital audio signals. The audio module 170 may also be used for encoding and decoding audio signals. In some embodiments, the audio module 170 may be provided in the processor 110, or some functional modules of the audio module 170 may be provided in the processor 110.
[0083] The speaker 170A, also known as the "loudspeaker", is used to convert an audio electrical signal into a sound signal. The electronic device can listen to music or a hands-free call through the speaker 170A.
[0084] The receiver 170B, also known as the "earpiece", is used to convert an audio electrical signal into a sound signal. When the electronic device answers a call or a voice message, the user can listen to the voice by holding the receiver 170B close to the ear.
[0085] The microphone 170C, also known as the "microphone" or "transmitter", is used to convert a sound signal into an electrical signal. When making a call or sending a voice message, the user can speak by bringing the mouth close to the microphone 170C to input the sound signal into the microphone 170C. The electronic device can be provided with at least one microphone 170C. In some other embodiments, the electronic device can be provided with two microphones 170C, which can not only collect sound signals but also implement a noise reduction function. In some other embodiments, the electronic device can also be provided with three, four or more microphones 170C to implement functions such as collecting sound signals, noise reduction, identifying the sound source, and implementing a directional recording function.
[0086] The headphone jack 170D is used to connect a wired headphone. The headphone jack 170D can be a USB jack or a 3.5 mm open mobile terminal platform (OMTP) standard jack or a cellular telecommunications industry association of the USA (CTIA) standard jack.
[0087] In the sensor module 180, the pressure sensor 180A is used to sense a pressure signal and can convert the pressure signal into an electrical signal. In some embodiments, the pressure sensor 180A can be disposed on the display screen 140. There are many types of pressure sensors 180A, such as a resistive pressure sensor, an inductive pressure sensor, a capacitive pressure sensor, etc. The capacitive pressure sensor can include at least two parallel plates with conductive materials. When a force acts on the pressure sensor 180A, the capacitance between the electrodes changes. The electronic device determines the intensity of the pressure according to the change in capacitance. When a touch operation acts on the display screen 140, the electronic device detects the intensity of the touch operation according to the pressure sensor 180A. The electronic device can also calculate the position of the touch according to the detection signal of the pressure sensor 180A. In some embodiments, touch operations with the same touch position but different touch operation intensities can correspond to different operation instructions.
[0088] The touch sensor 180B, also known as the "touch control device". The touch sensor 180B can be disposed on the display screen 140, and the touch sensor 180B and the display screen 140 form a touch screen, also known as the "touch control screen". The touch sensor 180B is used to detect a touch operation acting thereon or nearby. The touch sensor can transmit the detected touch operation to the application processor to determine the type of touch event. Visual output related to the touch operation can be provided through the display screen 140. In some other embodiments, the touch sensor 180B can also be disposed on the surface of the electronic device, at a different position from that of the display screen 140.
[0089] In some embodiments, the pressure sensor 180A and the touch sensor 180B can be used to detect a touch operation of a user on controls, images, icons, videos, etc. displayed on the display screen 140. The electronic device can execute corresponding processes in response to the touch operations detected by the pressure sensor 180A and the touch sensor 180B. For the specific content of the processes executed by the electronic device, reference can be made to the content of the following embodiments.
[0090] The key 190 includes a power-on key, a volume key, etc. The key 190 can be a mechanical key or a touch key. The electronic device can receive a key input and generate a key signal input related to the user settings and function control of the electronic device.
[0091] All the technical solutions involved in the following embodiments can be implemented in the electronic device 100 having the above hardware architecture.
[0092] For ease of understanding, the following embodiments of the present application will take Figure 1 the electronic device having the
[0093] shown structure as an example to specifically elaborate on the video processing method provided in the embodiments of the present application.
[0094] Embodiment 1
[0095] In some embodiments of the present application, the user can manually turn on or off the "multiple recordings in one shot" function provided in the embodiments of the present application. The following will Figure 2 describe the entry of the "multiple recordings in one shot".
[0096] Exemplarily, the user can instruct the mobile phone to start the camera application by touching a specific control on the mobile phone screen, pressing a specific physical key or key combination, inputting voice, performing an air gesture, etc. Figure 2 Figure (a) in Figure 2As shown in (a), the user clicks on the camera application icon displayed on the mobile phone display screen to input an instruction to turn on the camera. After the mobile phone receives the user's instruction to turn on the camera, the mobile phone starts the camera and displays Figure 2 the shooting interface shown in (b) or (c) in the figure.
[0097] Figure 2 The shooting interface shown in (b) in the figure is the shooting interface when the mobile phone is in the photo-taking mode. Figure 2 The shooting interface shown in (c) in the figure is the shooting interface when the mobile phone is in the video-recording mode. Taking Figure 2 the shooting interface shown in (b) in the figure as an example, the shooting interface of the mobile phone includes: a control 201 for turning on or off the flash, a control 202 for settings, a switch list 203, a control 204 for displaying the image taken in the previous shot, a control 205 for controlling shooting, and a control 206 for switching between the front and rear cameras, etc.
[0098] Among them, the control 201 for turning on or off the flash is used to control whether to start the flash when the camera takes pictures.
[0099] The control 202 for settings can be used for setting shooting parameters and shooting functions. For example, setting the photo ratio, setting gesture shooting, setting smiley face capture, setting the video resolution, etc.
[0100] The switch list 203 includes various modes of the camera. The user can slide the switch list left and right to realize the switching operation of various modes of the camera. Exemplarily, Figure 2 the switch list shown in (b) in the figure includes portrait, night scene, photo-taking, video-recording, panorama. Other modes not shown in Figure 2 the figure (b) are in hidden display. The user can display the modes in hidden display by sliding the switch list left and right.
[0101] The control 204 for displaying the image taken in the previous shot is used to display the thumbnail of the image taken in the previous shot by the camera or the cover thumbnail of the video. The user can touch the control 204 for displaying the image taken in the previous shot, and the display screen will display the image or video taken in the previous shot by the camera. Among them, the image or video taken in the previous shot by the camera refers to: the image or video that was taken by the camera before the current shot and the shooting time is the closest to the current shooting time.
[0102] The control 205 for controlling shooting is a control provided for the user to start shooting. In the photo-taking mode of the mobile phone, when the user touches the control 205 for controlling shooting once, the camera can take one frame of image. Of course, the camera can also take multiple frames of images and only select one frame of image for output. In the video-recording mode of the mobile phone, when the user touches the control 205 for controlling shooting, the camera starts recording.
[0103] The control 206 for switching between the front and rear cameras is used to implement the switching operation of multiple cameras of the mobile phone. Generally, a mobile phone includes a camera on the same side as the display screen (referred to as the front camera) and a camera on the outer shell of the mobile phone (referred to as the rear camera). The user can touch the control 206 for switching between the front and rear cameras to implement the switching operation between the front camera and the rear camera of the mobile phone.
[0104] As shown in Figure 2 (b) or Figure 2 (c) in the figure, by clicking the set control 202, the user controls the mobile phone to display the setting interface. Exemplarily, the setting interface can be as shown in Figure 2 (d) in the figure. Figure 2 In the setting interface shown in (d) in the figure, an option 207 for enabling "multiple recordings in one shot" is displayed, which is used to enable the function of multiple recordings in one shot. That is to say, when the user enables this function and the mobile phone is in the video recording mode to shoot a video, the mobile phone will automatically adopt the video processing method provided in the embodiments of the present application to automatically generate wonderful images and short videos while shooting the video. Of course, the user can also manually turn off the function of multiple recordings in one shot in the video recording mode through this option 207.
[0105] Generally, Figure 2 the option 207 for "multiple recordings in one shot" shown in (d) in the figure is in the default enabled state. That is to say, when the mobile phone is powered on for the first time or the system is updated to have the function of "multiple recordings in one shot", Figure 2 the option 207 for "multiple recordings in one shot" in the setting interface shown in (d) in the figure is in the enabled state, and the function of "multiple recordings in one shot" of the mobile phone is started.
[0106] It can be understood that the setting interface may also include controls for other function settings. For example, Figure 2 the controls for photo shooting settings and video shooting settings shown in (d) in the figure, where: the controls for photo shooting settings include: the control for setting the photo ratio, the control for voice-controlled photo shooting, the control for gesture photo shooting, the control for smiley face capture, etc.; the controls for video shooting settings include: the control for setting the video resolution, the control for setting the video frame rate, etc.
[0107] The method for triggering the mobile phone to enter the "multiple recordings in one shot" mode is introduced above, but the present application is not limited to entering "multiple recordings in one shot" in the video recording mode. In some embodiments of the present application, there may be other ways for the user to enable the "multiple recordings in one shot" function.
[0108] After enabling the option for multiple recordings in one shot, the user can control the mobile phone to start shooting a video. Exemplarily, referring to Figure 3 (a) in the figure, the user can click the control 301 for controlling shooting to control the mobile phone to start shooting a video. The mobile phone responds to the user's click operation and starts the camera to shoot a video. Figure 3The interface shown in (b) of [document] shows a picture taken by the user using a mobile phone during a football game. Figure 3 The interface shown in (b) of [document] includes: a stop control 302, a pause control 303, and a capture button 304. During the video shooting process, the user can pause the shooting by clicking the pause control 303, end the shooting by clicking the stop control 302, or manually capture a photo by clicking the capture button 304.
[0109] As Figure 3 In the interface shown in (c) of [document], the user can click the stop control 302 at 56 seconds to end the shooting process and obtain a video with a duration of 56 seconds. After the shooting is completed, the display screen of the mobile phone can enter the shooting interface. And when the "Multiple Gains in One Recording" function of the mobile phone is used for the first time, a generation prompt for Multiple Gains in One Recording is also guided on the shooting interface. Exemplarily, Figure 3 The shooting interface shown in (d) of [document] shows a dialog box with a generation prompt for Multiple Gains in One Recording, and this dialog box shows a text prompt of "Multiple Gains in One Recording wonderful photos have been generated and you can create a blockbuster with one key". And the user can make the dialog box with the generation prompt for Multiple Gains in One Recording disappear by clicking on Figure 3 any area of the interface shown in (d) of [document]. Or, the mobile phone can also be configured to make the dialog box with the generation prompt for Multiple Gains in One Recording disappear automatically after being displayed for a certain duration, such as 5 seconds.
[0110] Of course, after the "Multiple Gains in One Recording" function of the mobile phone is started once, when the user shoots a video again and clicks the stop control to end the shooting, the shooting interface shown on the display screen of the mobile phone will not include a generation prompt for Multiple Gains in One Recording.
[0111] It should be noted that when the duration of the video shot by the mobile phone meets a certain duration requirement, such as 15 seconds, the mobile phone will, according to the requirements of the Multiple Gains in One Recording mode, extract one or more wonderful images from the video shot by the mobile phone and generate a selected video. When the mobile phone can extract one or more wonderful images, Figure 3 the shooting interface shown in (d) of [document] can show a dialog box with a generation prompt for Multiple Gains in One Recording, and this dialog box shows a text prompt of "Multiple Gains in One Recording wonderful photos have been generated and you can create a blockbuster with one key".
[0112] However, the application scenarios for users to shoot videos can also include the following two application scenarios: The first application scenario: The video shooting duration is short and does not meet the video shooting duration requirement of the multi-shot-in-one mode. The second application scenario: The video shooting duration meets the video shooting duration requirement of the multi-shot-in-one mode, but the mobile phone does not recognize excellent images from the shot video. Additionally, in the second application scenario, it is usually required that the shooting time by the user is greater than another duration requirement, which is greater than the video shooting duration of the multi-shot-in-one mode, such as 30 seconds. And although the mobile phone does not recognize excellent images from the shot video, it can recognize images with better quality, and these images with better quality can be used to generate the selected videos proposed below.
[0113] In the first application scenario, when the user shoots a video in the multi-shot-in-one mode for the first time, and as Figure 4 shown in (a) below, when clicking the stop control 302 at the 10th second of shooting, the shooting interface displayed on the mobile phone is as Figure 4 shown in (b) below, and the text in the dialog box for the generation prompt of multi-shot-in-one is: "Multi-shot-in-one" excellent photos have been generated.
[0114] In the second application scenario, when the user shoots a video in the multi-shot-in-one mode for the first time, and as Figure 4 shown in (c) below, when clicking the stop control 302 at the 56th second of shooting, the shooting interface displayed on the mobile phone is as Figure 4 shown in (d) below, and the text in the dialog box for the generation prompt of multi-shot-in-one is: The "Multi-shot-in-one" video can create a blockbuster with one click.
[0115] Additionally, after the mobile phone displays the shooting interface as Figure 4 shown in (b) below, when the mobile phone first determines that the video shot meets the second application scenario, the mobile phone can also, after the user clicks the stop control 302 as Figure 4 shown in (c) below, display the shooting interface as Figure 4 shown in (d) below once. Or, when the mobile phone first determines that the shot video meets the video shooting duration requirement of the multi-shot-in-one mode and excellent images can be recognized, after the user clicks the stop control 302 as Figure 3 shown in (c) below, the mobile phone displays the shooting interface as Figure 3 shown in (d) below once.
[0116] And, after the mobile phone displays the shooting interface as Figure 4 shown in (d) below, when the mobile phone first determines that the video meets the first application scenario, the mobile phone can also, after the user clicks the stop control 302 as Figure 4 shown in (a) below, display the shooting interface as Figure 4The shooting interface shown in (b). Alternatively, when the mobile phone first determines that the shooting duration of the video meets the requirements of the multiple-record-in-one mode and can identify a video with wonderful images, after the user clicks the stop control 302 shown in Figure 3 (c), the shooting interface shown in Figure 3 (d) is displayed once.
[0117] It should be noted that when the user shoots a video in the multiple-record-in-one mode of the mobile phone, the mobile phone can automatically identify wonderful images in the shot video by using an identification model. In some embodiments, the mobile phone is provided with an identification model for wonderful images. When the video is input into the identification model for wonderful images, the identification model can score the images in the input video for wonderfulness to obtain the wonderfulness score value of the images in the video. The mobile phone can use the wonderfulness score value of the images to determine the wonderful images in the video. Generally, the larger the wonderfulness score value of the image, the higher the probability that the image belongs to a wonderful image.
[0118] In addition, during the shooting process of the video, the scene of the video can also be identified. After identifying a video segment of a scene is shot, the video segment of a scene is input into the aforementioned identification model for wonderful images, and the identification model scores the images in the video segment of this scene, including each image, for wonderfulness to obtain the wonderfulness score value of the images in the video. The mobile phone can use the wonderfulness score value of the images to determine the wonderful images in the video. Of course, the larger the wonderfulness score value of the image, the higher the probability that the image belongs to a wonderful image.
[0119] In some embodiments, the mobile phone can be configured to obtain a fixed number of wonderful images, such as 5 wonderful images. Based on this, the mobile phone selects 5 images with higher wonderfulness score values as wonderful images.
[0120] In other embodiments, the mobile phone may not be configured with a limited number of wonderful images. The mobile phone can select images with wonderfulness score values higher than a certain value as wonderful images.
[0121] After the mobile phone finishes shooting the video, the mobile phone obtains the shot video and one or more wonderful images in the video. The mobile phone can also save the shot video and wonderful images to the gallery. In one example, Figure 5 (a) shows the album display interface of the gallery, and this album display interface shows all the photos and videos saved by the mobile phone in the form of folders. Exemplarily, Figure 5 (a) shows that the album display interface includes: a camera folder 401, an all photos folder 402, a video folder 403, and a multiple-record-in-one folder 404. Of course, the album display interface may also include other folders, and the present application does not limit the folders shown in the camera display interface.
[0122] Under normal circumstances, the camera folder 401 includes all photos and videos taken by the camera of the mobile phone. The all photos folder 402 includes all photos and videos saved on the mobile phone. The video folder 403 includes all videos saved on the mobile phone. The multi-shot folder 404 includes all wonderful images taken in the multi-shot mode on the mobile phone.
[0123] Figure 5 Figure (b) shows the display interface of the multi-shot folder, which displays thumbnails of 5 wonderful images in the aforementioned 56-second video taken in the multi-shot mode on the mobile phone. Moreover, the 5 wonderful images are automatically captured by the mobile phone from the 56-second video taken.
[0124] After the mobile phone saves the videos taken and the wonderful images in the videos in the gallery, the user can view the videos and the wonderful images in the videos through the gallery application. Exemplarily, the user clicks the control 501 that displays the image taken in the previous shot in the shooting interface shown in Figure 6 Figure (a), or the user clicks the cover thumbnail of the video 502 in the photo display interface of the gallery shown in Figure 6 Figure (b). In response to the user's click operation, the mobile phone displays the details interface of the video 502 on the display screen of the mobile phone.
[0125] Figure 6 In the lower left corner of the cover thumbnail of the video 502 shown in Figure (b), there is a video exclusive corner mark for the "multi-shot" function. This video exclusive corner mark is used to inform the user that the video 502 corresponding to this cover thumbnail is taken by the mobile phone using the "multi-shot" function.
[0126] Figure 6 The size of the cover thumbnail of the video 502 shown in Figure (b) is different from the thumbnails of Image A and Image B. Image A and Image B may refer to images taken without enabling the multi-shot function on the mobile phone. Exemplarily, Figure 6 the cover thumbnail of the video 502 shown in Figure (b) is larger than the thumbnails of Image A and Image B. Of course, Figure 6 the cover thumbnail of the video 502 shown in Figure (b) may also be the same as the thumbnails of other photos and videos (referring to videos without enabling the multi-shot function) in the gallery. The embodiments of the present application do not limit this.
[0127] When the display screen of the mobile phone first shows the details interface of the video 502 (which can also be called the browsing interface of the video 502), a mask guide is displayed on the details interface. When the display screen of the mobile phone shows the details interface of the video 502 non-first time, no mask is displayed on the details interface. Exemplarily, the details interface of the video 502 with the mask guide shown in Figure 6 Figure (c), and the details interface of the video 502 without the mask guide shown in Figure 6As shown in Figure (d). It should also be noted that the video 502 can be understood as a video obtained by the mobile phone shooting for the first time in the mode of "one recording, multiple gains". Therefore, when the display screen of the mobile phone first shows the details interface of the video 502, it can be understood as the mobile phone first showing the video obtained by the mobile phone shooting for the first time in the mode of "one recording, multiple gains". In this case, a mask guide will be displayed on the details interface of the video. Moreover, when the mobile phone first shows the video obtained by the mobile phone shooting for the first time in the mode of "one recording, multiple gains" and a mask guide is displayed on the details interface of the video, it can play the role of reminding the user of the function of "one recording, multiple gains".
[0128] Figure 6 The details interface of the video 502 shown in Figure (c) includes: a thumbnail area 504 of the wonderful images of the video 502, a control 505, and a play control 506. Among them, the thumbnail area 504 of the wonderful images of the video 502 is exposed and not covered by the mask, and other areas of the display screen are covered by the mask.
[0129] The thumbnail area 504 of the wonderful images of the video 502 includes: the cover thumbnail of the video 502, and the thumbnails of multiple wonderful images of the video 502. The cover thumbnail of the video 502 is usually in the first place. The thumbnails of multiple wonderful images can be arranged according to the shooting time of the wonderful images and are located after the cover thumbnail of the video 502. The wonderful images can be, as described above, the pictures of the wonderful moments automatically recognized in the video when the mobile phone shoots the video and extracted from the video. Moreover, the details interface of the video 502 also includes a reminder dialog box, which displays the text "One recording, multiple gains, intelligently captures multiple wonderful moments for you". The reminder dialog box is usually as Figure 6 shown in Figure (c), above the thumbnail area 504 of the wonderful images of the video 502, and is used to prompt the user of the content displayed in the thumbnail area 504 of the wonderful images of the video 502 to guide the user to view the wonderful images of the video 502. Of course, Figure 6 the text and the setting position shown in the reminder dialog box in Figure (c) are an exemplary display and do not constitute a limitation on the reminder dialog box. The user can click on Figure 6 any area of the details interface of the video 502 shown in Figure (c) to control the disappearance of the reminder dialog box. Or, the mobile phone can also be configured to display the reminder dialog box for a certain duration, such as 5 seconds, and then disappear automatically.
[0130] The control 505 is used to generate a selected video based on the wonderful images of the video 502.
[0131] The play control 506 is used to control the playback of the video 502. Exemplarily, such as Figure 6As shown in (d), the playback control 506 includes: a start / stop control 507, a slidable progress bar 508, and a speaker control 509. The start / stop control 507 is used to control the playback or stop of the video 502; the speaker control 509 is used to select whether to play the video 502 in mute. The slidable progress bar 508 is used to display the playback progress of the video 502. The user can also adjust the playback progress of the video 502 by dragging the circular control on the progress bar left or right.
[0132] The details interface of the video 502 also includes options such as share, favorite, edit, delete, and more. If the user clicks share, the video 502 can be shared; if the user clicks favorite, the video 502 can be saved in a folder; if the user clicks edit, the video 502 can be edited; if the user clicks delete, the video 502 can be deleted; if the user clicks more, other operation functions for the video can be entered (such as move, copy, add note, hide, rename, etc.).
[0133] The details interface of the video 502 also includes the shooting information of the video 502, usually as shown in Figure 6 As shown in (c) or (d), above the video 502. The shooting information of the video 502 includes: the shooting date, shooting time, and shooting address of the video 502. In addition, the details interface of the video 502 can also include a circular control filled with the letter "i" inside. When the user clicks this circular control, the mobile phone can respond to the user's click operation and display the attribute information of the video 502 on the details interface of the video 502. Exemplarily, the attribute information can include the storage path of the video 502, resolution, configuration information of the camera during shooting, etc.
[0134] Figure 6 The mask layer shown in (c) belongs to a kind of mask layer. The mask layer generally refers to the layer mask. The layer mask is a layer of glass sheet layer covered on the layer of the interface displayed on the display screen. And the glass sheet layer is divided into transparent, semi-transparent, and completely opaque. The semi-transparent and completely opaque mask layers can block the light of the display screen, making the interface displayed on the display screen blurred or completely invisible to the user. Figure 6 The mask layer shown in (c) can be understood as a semi-transparent glass sheet layer.
[0135] Figure 6 The mask layer guide shown in (c) is an exemplary display and does not constitute a limitation on the mask layer guide for the details interface of the video shot with one recording and multiple gains for the first time. In some embodiments, the mask layer guide can also be set as a guide with a mask layer and other special effects, such as a guide with a mask layer and bubbles.
[0136] And the user can pass through in Figure 6Perform an input operation on any area of the details interface of video 502 shown in (c) to control the disappearance of the overlay. Of course, the mobile phone can also be set so that the overlay is displayed for a certain duration and disappears automatically after, for example, 3 seconds. Figure 6 After the overlay on the details interface of video 502 shown in (c) disappears and the reminder dialog box disappears, the details interface of this video 502 is as Figure 6 shown in (d).
[0137] It should be noted that Figure 6 Before the overlay on the details interface of video 502 shown in (c) disappears, video 502 is in a static state and will not play. After the overlay disappears, video 502 can play automatically. Under normal circumstances, video 502 can also be played muted. Of course, the user can click on Figure 6 the speaker control 509 shown in (d) to control the mobile phone to play video 502 with sound.
[0138] The user can also Figure 6 perform a left or right swipe operation or a click operation in the thumbnail area 504 of the wonderful images of video 502 shown in (d). In some embodiments, when the user clicks on the thumbnail of a wonderful image in the thumbnail area 504, the mobile phone responds to the user's click operation and displays the wonderful image clicked by the user on the display screen to replace Figure 6 video 502 shown in (d). In another embodiment, when the user performs a left or right swipe operation in the thumbnail area 504, the mobile phone can also respond to the user's swipe operation and display the wonderful images in the thumbnail area 504 on the display screen following the user's swipe direction. The thumbnails of the wonderful images shown in the thumbnail area 504 do not have their corresponding images saved in the photo gallery but are saved in a Yiluoduode album. That is to say, in Figure 6 the interface shown in (b), there are no thumbnails corresponding to the wonderful images. However, when the user clicks on the thumbnail of video 502 to enter the details interface of video 502, thumbnails of the wonderful images associated with video 502 can be displayed below the details interface of video 502.
[0139] The user can also Figure 6 perform a left or right swipe operation on video 502 shown in (d). The mobile phone responds to the user's swipe operation and displays other images or videos saved in the photo gallery of the mobile phone on the display screen. In some embodiments, when the user performs a right swipe operation on video 502 shown in (d), the mobile phone displays the next video or image of video 502 saved in the photo gallery on the display screen. When the user Figure 6 performs a right swipe operation on video 502 shown in (d), the mobile phone displays the next video or image of video 502 saved in the photo gallery on the display screen. When the user Figure 6In Figure (d), when a left - sliding operation is input for the video 502, the mobile phone displays the previous video or image of the video 502 saved in the gallery on the display screen. Herein, the previous video or image refers to the video or image whose shooting time is earlier than that of the video 502 and is the video or image with the closest shooting time to the video 502. The subsequent video or image refers to the video or image whose shooting time is later than that of the video 502 and is the video or image with the closest shooting time to the video 502.
[0140] It should be noted that if the video 502 shot by the user is the video in the first application scenario mentioned above, the detailed interface of the video 502 will be different from that in Figure 6 Figure (c) and Figure 6 Figure (d). The difference lies in that the detailed interface of the video 502 does not include the control 505. If the video 502 shot by the user is the video in the second application scenario mentioned above, the detailed interface of the video 502 will also be different from that in Figure 6 Figure (c) and Figure 6 Figure (d). The difference lies in that the detailed interface of the video 502 does not include the thumbnail area 504.
[0141] It also should be noted that when the user shoots a video in the one - shot - multiple - gains mode of the mobile phone, the mobile phone can obtain the shot video, one or more wonderful images in the video. In addition, the mobile phone can generate a configuration file, and the configuration file can include the tag (TAG) of the video. Or, the mobile phone can obtain tag data, and the tag data includes the tag (TAG) of the video. And this tag data can be added to the video, usually at the beginning of the video.
[0142] The following content will be introduced by taking the mobile phone obtaining the video and the configuration file of the video as an example. Of course, if the tag of the video is stored in the video in the form of tag data, obtaining the configuration file of the video in the following content can be modified to obtaining the tag data of the video.
[0143] In a possible implementation, the tag (TAG) of the video can be set based on the hierarchical information of the video. The hierarchical information of the video can include: the first - level information LV0, the second - level information LV1, the third - level information LV2, and the fourth - level information LV3. Among them:
[0144] The first - level information LV0 is used to represent the theme category of the video and give the style or atmosphere TAG of the whole video.
[0145] The second - level information LV1 is used to represent the scene of the video and give the scene TAG of the video.
[0146] The third-level information LV2 is used to represent that the scene of the video has changed, which can also be understood as the change of the transition storyboard. The information of the third-level information LV2 can give the video transition position (for example, the frame number where the transition occurs), and the transition type (character protagonist switching, rapid camera movement, scene category change, image content change caused by other situations), so as to prevent excessive recommendation of similar scenes. The information of LV2 is used to represent the video scene change (or can also be simply referred to as transition), including but not limited to one or more of the following changes: change of the main character (or protagonist), large change in the composition of the image content, change of the scene at the semantic level, and change of the image brightness or color. The mobile phone can use the third-level information LV2 to add a storyboard TAG to the video when the scene in the video changes.
[0147] The fourth-level information LV3 is used to represent the exciting moment, that is, the shooting moment of the exciting image, and is used to give the exciting image TAG of the video.
[0148] The first-level information LV0, the second-level information LV1, the third-level information LV2, and the fourth-level information LV3 provide decision-making information in the order of coarser to finer granularity to identify the exciting images in the video and generate a selected video.
[0149] The following Table 1 gives examples of the definitions of LV0 and LV1.
[0150] Table 1
[0151]
[0152]
[0153] The mobile phone can use the captured video and the configuration file of the captured video to generate a selected video of the captured video. Of course, when the mobile phone can identify the exciting images from the captured video, the selected video includes the exciting images of the captured video and has some special effects and music. In addition, if the mobile phone fails to identify the exciting images from the captured video but can identify the images with good quality, the mobile phone uses the images with good quality to generate a selected video. Of course, this selected video also has some special effects and music. It should also be noted that the configuration file or tag data of the video will also include the TAG of the images with good quality.
[0154] The images with good quality mentioned in this application refer to: the images are relatively clear, for example, with a higher resolution; or the images are relatively complete.
[0155] The special effects mentioned in this application refer to: those that can be supported by the materials and can present special effects after being added to the video frames, such as: snowflakes, fireworks and other animation effects, as well as filters, stickers, borders, etc. In some embodiments, the special effects can also be referred to as styles or style themes, etc.
[0156] The following content of this application will be introduced by taking the example of a mobile phone generating a selected video using wonderful images.
[0157] Exemplarily, as shown in (a) below, the user clicks on the control 601 of the video details interface. The mobile phone responds to the user's click operation and generates a selected video of a certain duration. Figure 7 As shown in (a) below, the mobile phone responds to the user's click operation and generates a selected video of a certain duration.
[0158] Normally, since generating the selected video of the video takes a certain amount of time, therefore, as shown in (a) below, after the user clicks on the control 601 of the video details interface, the display screen of the mobile phone will display a buffering interface of the selected video 602 as shown in (b) below. After the selected video 602 is generated, the display screen of the mobile phone displays the display interface of the selected video 602. Of course, when the performance of the mobile phone is relatively strong, as shown in (a) below, after the user clicks on the control 601 of the video details interface, the display screen of the mobile phone may not display the buffering interface in (b) below, but directly display the display interface of the selected video 602. Figure 7 As shown in (a) below, after the user clicks on the control 601 of the video details interface, the display screen of the mobile phone will display a buffering interface of the selected video 602 as shown in (b) below. Figure 7 As shown in (b) below, after the selected video 602 is generated, the display screen of the mobile phone displays the display interface of the selected video 602. Of course, when the performance of the mobile phone is relatively strong, as shown in (a) below, after the user clicks on the control 601 of the video details interface, the display screen of the mobile phone may not display the buffering interface in (b) below, but directly display the display interface of the selected video 602. Figure 7 As shown in (a) below, after the user clicks on the control 601 of the video details interface, the display screen of the mobile phone may not display Figure 7 the buffering interface in (b) below, but directly display the display interface of the selected video 602.
[0159] Exemplarily, in the display interface of the selected video 602, the selected video 602 is being played. And the display interface of the selected video 602, as shown in (c) below, includes: a style control 603, a save control 604, a share control 605, a music control 606, an edit control 607, etc. Figure 7 As shown in (c) below, includes: a style control 603, a save control 604, a share control 605, a music control 606, an edit control 607, etc.
[0160] When the user clicks on the style control 603, the mobile phone responds to the user's click operation, and the display screen shows various video styles saved in the mobile phone. The user can select different video styles for the selected video 602. In some embodiments, the video style can be a filter, that is, by applying a filter to perform color adjustment processing on the selected video 602. The filter is a type of video effect used to achieve various special effects of the selected video 602. In some other embodiments, the video style can also be video effects such as fast forward and slow motion. In some other embodiments, the video style can also refer to various themes, and different themes include corresponding filters, music, and other content.
[0161] When the user clicks on the share control 605, the mobile phone responds to the user's click operation and can share the selected video 602.
[0162] When the user clicks on the music control 606, the mobile phone responds to the user's click operation, and the display screen shows an interface for adding different background music to the wonderful video 602. This interface displays multiple background music controls, and the user can click on any of the background music controls to select background music for the wonderful short video, such as soothing, romantic, warm, cozy, serene, etc., to add background music to the selected video 602.
[0163] When the user clicks on the editing control 607, the mobile phone responds to the user's click operation, and the display screen shows the editing interface of the wonderful video 602. The user can input editing operations such as clipping, splitting, volume adjustment, and frame size adjustment for the wonderful video 602 in the editing interface.
[0164] The save control 604 is used to save the selected video 602.
[0165] The following combines Figure 8 , Figure 9 and Figure 10 to introduce the process of generating the selected video 602.
[0166] The method for a mobile phone to generate a selected video provided by the embodiments of this application, refer to Figure 8 , includes the following steps:
[0167] S701. In response to the user's first operation, obtain a video and the configuration file of the video. The configuration file of the video includes a theme TAG, a scene TAG, a shot TAG, and a wonderful image TAG.
[0168] As can be seen from the foregoing content: When the mobile phone shoots a video using the "one-shot, multiple-uses" mode, the mobile phone will identify the content of the shot video and determine the hierarchical information of the video. The mobile phone can also use the hierarchical information of the video to set a theme TAG, a scene TAG, a shot TAG, and a wonderful image TAG for the video, and write the theme TAG, the scene TAG, the shot TAG, and the wonderful image TAG into the configuration file of the video. When the video shooting is completed, the mobile phone can save the shot video and the configuration file of the video.
[0169] In the configuration file of the video, the theme TAG is used to characterize the style or atmosphere of the video. In the Figure 10 shown example, the 60-second video shot by the mobile phone contains tn frames of images, where tn is an integer. During the process of shooting this video by the mobile phone, the mobile phone can continuously identify the style or atmosphere of the video to determine the theme of the video. In some embodiments, the mobile phone can call an identification algorithm to identify the images of the video to determine the style or atmosphere of the video. Of course, the mobile phone can also use the identification algorithm to identify the images of the video after the video shooting is completed to determine the style or atmosphere of the video.
[0170] In Figure 10In the shown example, the mobile phone calls the recognition algorithm to recognize the images of the video and determines that the video belongs to the travel theme.
[0171] The scene TAG is used to characterize the scene of the video. Figure 10 In the shown example, the video is divided into 6 scenes. The first scene contains the video segment from 0 to 5 seconds and belongs to other scenes; the second scene contains the video segment from 5 to 15 seconds and belongs to the people scene; the third scene contains the video segment from 15 to 25 seconds and belongs to the ancient building scene; the fourth scene contains the video segment from 25 to 45 seconds and belongs to the people and ancient building scenes; the fifth scene contains the video segment from 45 to 55 seconds and belongs to the mountain scene; the sixth scene contains the video segment from 55 to 60 seconds and belongs to the people scene.
[0172] The shot TAG is used to characterize the position of the transition scene in the video indicating the shooting of the video. Figure 10 In the shown example, the video includes 6 shot TAGs. The first shot TAG indicates the start of the first scene at 0 seconds, and the second shot TAG indicates the start of the second scene at 5 seconds; the third shot TAG indicates the start of the third scene at 15 seconds, and the fourth shot TAG indicates the start of the fourth scene at 25 seconds; the fifth shot TAG indicates the start of the fifth scene at 45 seconds, and the sixth shot TAG indicates the start of the sixth scene at 55 seconds.
[0173] The wonderful image TAG is used to characterize the position of the wonderful image in the shot video. Figure 10 In the shown example, the video includes 5 wonderful images.
[0174] As Figure 6 shown in (a) below, the user can click on the "AI One - Key Blockbuster" control to input an instruction to generate a selected video to the mobile phone. The mobile phone receives the click operation of the user (which can also be called the first operation) and in response to this click operation, obtains the video saved on the mobile phone and the configuration file of the video.
[0175] S702. Determine the style template and music based on the theme TAG of the video.
[0176] The mobile phone stores multiple style templates and music, and the style template can include multiple special effects.
[0177] See Figure 9 , the clip engine is used to determine the style template and music by using the theme TAG of the video. This style template and music are used to synthesize the selected video of the video. It can be understood that the clip engine belongs to a service or application and can be set in the application layer, application framework layer or system library of the software framework of the mobile phone. The clip engine is used to generate the selected video of the video.
[0178] In some embodiments, the mobile phone stores the correspondence between the theme TAG, the style template, and the music. The mobile phone can determine the style template and music corresponding to the theme TAG based on the theme TAG of the video.
[0179] In other embodiments, the mobile phone can also randomly select the style template and music stored in the mobile phone according to the theme TAG of the video. It should be noted that for the same video, after generating the selected video of the video, if the user edits the selected video and adjusts the style template or music, the mobile phone selects different style templates or music to ensure that the special effects of the selected video generated by the mobile phone for the same video are different each time.
[0180] It should be noted that step S702 can be understood as Figure 9 1. Template selection and 2. Music selection shown.
[0181] S703. Based on the storyboard TAG, determine the storyboard segments in each scene of the video.
[0182] The storyboard TAG of the video can indicate the positions of the transition scenes in the captured video. Therefore, the editing engine can determine the scenes included in the video based on the storyboard TAG. The editing engine can divide the video into multiple storyboard segments according to the scenes by using the storyboard TAG. Of course, the editing engine does not actually split the video into storyboard segments, but marks each scene's storyboard segment of the video by marking the video according to the storyboard TAG.
[0183] In Figure 10 In the example shown, the video includes 6 storyboard TAGs. According to the 6 storyboard TAGs, the storyboard segments in the video include: the storyboard segment of the first scene from 0 to 5 seconds, the storyboard segment of the second scene from 5 seconds to 15 seconds, the storyboard segment of the third scene from 15 seconds to 25 seconds, the storyboard segment of the fourth scene from 25 seconds to 45 seconds, the storyboard segment of the fifth scene from 45 seconds to 55 seconds, and the storyboard segment of the sixth scene from 55 seconds to 60 seconds.
[0184] S704. Based on the TAG of the wonderful image, determine the position of the wonderful image in the storyboard segment.
[0185] The wonderful image TAG is used to indicate the position of the wonderful image in the captured video. Therefore, the position of the wonderful image in the storyboard segment can be determined according to the TAG of the wonderful image. In some embodiments, the storyboard segment of a scene may include one or more wonderful images.
[0186] In Figure 10In the example shown, the video captured by the mobile phone includes 5 wonderful images. The first wonderful image belongs to the storyboard segment of the second scene, the second wonderful image belongs to the storyboard segment of the third scene, the third wonderful image belongs to the storyboard segment of the fourth scene, the fourth wonderful image belongs to the storyboard segment of the fifth scene, and the fifth wonderful image belongs to the storyboard segment of the sixth scene.
[0187] S705. Obtain multiple frames of images before and after a wonderful image from the storyboard segment of the video to obtain the associated images of the wonderful image.
[0188] In some embodiments, the storyboard segment of a scene in the video includes at least one wonderful image. Therefore, in the storyboard segment of the scene to which the wonderful image belongs, the clipping engine obtains several frames of images before and several frames of images after the wonderful image. Exemplarily, 5 frames of images before and 5 frames of images after the wonderful image can be obtained. The clipping engine uses the obtained images as the associated images of the wonderful image.
[0189] In Figure 10 In the example shown, 5 frames of images before and 5 frames of images after the first wonderful image are obtained from the storyboard segment of the second scene; 5 frames of images before and 5 frames of images after the second wonderful image are obtained from the storyboard segment of the third scene; 5 frames of images before and 5 frames of images after the third wonderful image are obtained from the storyboard segment of the fourth scene; 5 frames of images before and 5 frames of images after the fourth wonderful image are obtained from the storyboard segment of the fifth scene; 5 frames of images before and 5 frames of images after the fifth wonderful image are obtained from the storyboard segment of the sixth scene.
[0190] In addition, when the number of wonderful images in the video captured by the mobile phone exceeds the limit of the number of wonderful images that can finally be presented, the wonderful images with higher wonderfulness score values can be preferentially retained, and the wonderful images with lower wonderfulness score values can be discarded, and then step S705 is executed.
[0191] It should be noted that step S703 and step S705 can be understood as Figure 9 shown in 3. Segment selection.
[0192] Among them, a wonderful image and its associated images can be understood as forming a small segment in combination. After obtaining the small segments formed by combining each wonderful image and its associated images, it can also be as Figure 8 shown, execute 8. Content deduplication and segment discretization.
[0193] Content deduplication can be understood as: deleting the small segments that belong to the same content in the small segments formed by combining the wonderful image and its associated images, and only retaining one small segment.
[0194] Fragment discretization can be understood as: discretizing the images contained within small fragments. In some embodiments, transitional images can be inserted between the images contained within the small fragments. The inserted transitional images can be substantially the same as the images within the small fragments and follow the transitional changes from the first few frames of the wonderful image to the wonderful image and then to the subsequent few frames of the wonderful image.
[0195] S706. Splice the wonderful image and its associated images in chronological order to obtain a video clip.
[0196] In some embodiments, as described in step S705, each wonderful image and its associated images can be combined into a small fragment, in which: the wonderful image and its associated images are arranged in chronological order of the photographing time. And, also in chronological order of the shooting time, splice the small fragments to obtain a video clip.
[0197] In other embodiments, it is also possible to splice each wonderful image and the associated image of each wonderful image in chronological order of the shooting time to obtain a video clip.
[0198] In Figure 10 In the shown example, splice in the order of the first 5 images of the first wonderful image, the first wonderful image, the subsequent 5 images of the first wonderful image, the first 5 images of the second wonderful image, the second wonderful image, the subsequent 5 images of the second wonderful image, until the first 5 images of the fifth wonderful image, the fifth wonderful image, and the subsequent 5 images of the fifth wonderful image to obtain a video clip.
[0199] It should be noted that step S706 can be understood as Figure 9 the 5. Fragment splicing strategy shown.
[0200] Similarly referring to Figure 9 , when splicing to obtain a video clip, the rhythm point information of the music determined in step S702 can be referred to to ensure that the spliced video clip fits the rhythm points of the music.
[0201] S707. Synthesize the special effects, music, and video clip provided by the style template to obtain a selected video.
[0202] It should be noted that step S707 can be understood as Figure 8 the 6. Synthesis shown, where the synthesized video is the selected video.
[0203] In addition, if the mobile phone uses high-quality images to generate a selected video, in the method for the mobile phone to generate a selected video mentioned above, just replace the wonderful images with high-quality images. To avoid redundancy, the method for generating a selected video using high-quality images will not be introduced in detail here.
[0204] After generating a selected video using the foregoing content, the selected video can be saved.
[0205] The following will introduce the saving process of the selected video in conjunction with Figure 11 ...
[0206] Exemplarily, the user clicks on the save control 604 of the display interface of the selected video 602 shown in (a) of Figure 11 ... The mobile phone responds to the user's click operation and saves the selected video 602. Exemplarily, as shown in (c) of Figure 11 ..., the mobile phone saves the selected video 602 in the gallery. In some embodiments, the selected video 602 can be saved following the original video 502 of the selected video 602, that is, the selected video 602 and the original video 502 are stored in the same storage area. In other embodiments, the selected video 602 may not be saved following the original video 502 of the selected video 602, that is, the selected video 602 and the original video 502 are stored in different storage areas. It should be noted that the original video 502 refers to the video shot by the user, and the selected video 602 is derived from the original video 502.
[0207] The user clicks on the save control 604 of the display interface of the selected video 602 shown in (a) of Figure 11 ... In addition to the mobile phone responding to the user's click operation and saving the selected video 602, the display screen of the mobile phone can also display the details interface of the selected video 602. Exemplarily, as shown in (b) of Figure 11 ... shows the details interface of the selected video 602. In some embodiments, in the details interface of the selected video 602, the selected video 602 can be automatically played.
[0208] The details interface of the selected video 602 is basically the same as the details interface of the video 502 shown in (d) of Figure 6 ..., the difference being that the selected video 602 played in the details interface of the selected video 602. For the controls included in the details interface of the selected video 602 and their functions, reference can be made to the content of the details interface of the video 502 shown in (d) of Figure 6 ... mentioned above, which will not be elaborated here.
[0209] It should be noted that before the user clicks on the save control 604 shown in (a) of Figure 11 ..., the mobile phone can generate the selected video 602 but does not save it in the internal memory. Only after the user clicks on the save control 604 shown in (a) of Figure 11 ..., the mobile phone will also save the generated selected video 602 in the internal memory.
[0210] Embodiment Two
[0211] During the process of shooting a video in the one - record - multiple - gains mode proposed in the foregoing First Embodiment, wonderful images in the video can be obtained. After the mobile phone finishes shooting the video, the shot video and one or more wonderful images in the video can be obtained. In this way, wonderful moment photos worth commemorating can be captured while shooting the video.
[0212] However, if the user does not know that the mobile phone is set with the one - record - multiple - gains mode, which can support the function of capturing wonderful moment photos while shooting the video, the user will still control the mobile phone to take pictures in the shooting mode when shooting the video. Based on this, when the mobile phone detects the above behavior of the user, it is necessary to guide the user to understand the one - record - multiple - gains function of the mobile phone.
[0213] Exemplarily, Figure 12 Figure (a) shows the video shooting interface of the mobile phone. When the user is shooting a video with the mobile phone and wants to take a picture. The user can click the stop control 1101 on the video shooting interface to stop shooting. In response to the user's click operation, the mobile phone saves the shot video and displays the shooting interface of the mobile phone in the video recording mode as shown in Figure 12 Figure (b). As shown in Figure 12 Figure (b), the user clicks "Take Photo". In response to the user's click operation, the mobile phone enters the shooting interface in the shooting mode as shown in Figure 12 Figure (c).
[0214] The user is in the shooting interface shown in Figure 12 Figure (c) and clicks the control shooting control 1102. In response to the user's click operation, the user takes a picture and saves the taken picture. Then, the user controls the mobile phone to shoot a video again.
[0215] When the mobile phone detects the above operation of the user, it can display a one - record - multiple - gains guiding prompt on the video shooting interface when the user controls the mobile phone to shoot a video. Exemplarily, Figure 12 In the video shooting interface of the mobile phone shown in Figure (d), a dialog box 1103 with a one - record - multiple - gains guiding prompt is displayed. The dialog box 1103 includes the text "When using the rear - camera video recording, wonderful moments are intelligently recognized, and wonderful photos will be automatically generated after the recognition and you can create a blockbuster with one key", and can also prompt the user that during the process of shooting the video, the user can also manually click the control to capture photos.
[0216] To implement the above functions, in a video processing method provided in this embodiment, when the mobile phone determines to shoot a video in the one-shot-multiple-captures mode, the mobile phone can detect whether the user executes an operation of ending the video shooting and immediately taking a photo, and then shooting a video after the photo-taking ends. If the mobile phone detects that the user executes an operation of ending the video shooting and immediately taking a photo, and then shooting a video after the photo-taking ends, when the user finishes taking a photo and then shoots a video, a one-shot-multiple-captures guiding prompt is displayed on the video shooting interface displayed on the mobile phone display screen.
[0217] Of course, during the process of the mobile phone shooting a video in the one-shot-multiple-captures mode, the mobile phone does not detect that the user executes an operation of ending the video shooting and immediately taking a photo, and then shooting a video after the photo-taking ends. If the mobile phone does not detect that the user executes an operation of ending the video shooting and immediately taking a photo, and then shooting a video after the photo-taking ends, the mobile phone can respond to the user's operation according to the normal process.
[0218] The way for the mobile phone to detect whether the user executes an operation of ending the video shooting and immediately taking a photo, and then shooting a video after the photo-taking ends can be:
[0219] The mobile phone traverses the processes of the mobile phone and identifies the state changes of the processes of the mobile phone. By using the result of the state changes of the processes of the mobile phone, it is determined whether the user executes an operation of ending the video shooting and immediately taking a photo, and then shooting a video after the photo-taking ends.
[0220] Among them: If the mobile phone determines that the video recording process is in a running state, and it also determines that within a certain period of time after the video recording process is closed, such as 10 seconds, the photo-taking process is started and in a running state, and within a certain period of time after the photo-taking process is closed, such as 10 seconds, the video recording process is started and in a running state, it indicates that the user executes an operation of ending the video shooting and immediately taking a photo, and then shooting a video after the photo-taking ends.
[0221] In some embodiments, the mobile phone can distinguish different mobile phone processes through process identifiers. The video recording process has a video recording process identifier, and the photo-taking process has a photo-taking process identifier.
[0222] Embodiment III
[0223] During the process of the mobile phone shooting a video using the one-shot-multiple-captures mode, it also supports the manual capture function.
[0224] Exemplarily, Figure 13 The video shooting interface shown in (a) shows a picture during a football game. As Figure 13 shown in (a), the user can click the photo-taking button 1201 to execute the manual capture operation. The mobile phone responds to the user's manual capture operation, calls the camera to take a photo, and saves the taken image to the photo library. Exemplarily, Figure 13In the photo display interface of the photo library shown in (b), a thumbnail of the image 1202 manually captured during video shooting by the mobile phone is displayed.
[0225] During the process of the mobile phone shooting a video using the multiple recordings in one mode, when the user starts the manual capture to shoot an image, it indicates that the image manually captured by the user is a wonderful image that the user deems more worthy of commemoration. Based on this, in response to the user's manual capture operation, in addition to saving the manually captured image in the photo library, the mobile phone can also save it as a wonderful image in the multiple recordings in one folder.
[0226] In some embodiments, the mobile phone can replace the wonderful image recognized by the mobile phone with the manually captured image and save it in the multiple recordings in one for preservation.
[0227] As can be known from the content of the foregoing Embodiment 1: The mobile phone can use the recognition model to score the wonderfulness of the images in the captured video, and use the wonderfulness score value to determine the wonderful images in the video. When the user uses the manual capture function to shoot an image, the mobile phone can discard the wonderful images in the number of captured images in ascending order of the wonderfulness score value. Then the mobile phone saves the remaining wonderful images and the manually captured image as the updated wonderful images in the multiple recordings in one folder.
[0228] Exemplarily, the mobile phone is configured to save 5 wonderful images in the multiple recordings in one folder. Figure 13 In (a), the user starts the manual capture function to shoot an image. Figure 13 In (c), the multiple recordings in one folder is shown. The multiple recordings in one folder includes: thumbnails of 4 wonderful images recognized by the mobile phone from the captured video, and a thumbnail of the image 1202 manually captured by the user.
[0229] Figure 13 The thumbnails of the wonderful images in the multiple recordings in one folder shown in (c) can be arranged in the order of the shooting time of the wonderful images. That is, for the image captured by the mobile phone first, its thumbnail is in the front, and for the image captured by the mobile phone later, its thumbnail is in the back. Of course, this sorting method does not limit the sorting of the image thumbnails in the multiple recordings in one folder.
[0230] In addition, Figure 13 In (d), the details interface of the video captured by the user is shown. The thumbnails of the wonderful images of the video 502 shown in this details interface come from Figure 13 the multiple recordings in one folder shown in (c). Therefore, the thumbnails of the wonderful images of the video 502 include: the cover thumbnail of the video 502, the thumbnails of 4 wonderful images recognized by the mobile phone in the video 502, and the thumbnail of the image 1202 manually captured by the user.
[0231] It should also be noted that when the user captures an image using the manual capture function, after the mobile phone captures the image, the user can also set a TAG for the image to indicate the position of the image in the captured video. The mobile phone can also write the TAG of the manually captured image into the configuration file of the video. In some embodiments, the TAG of the manually captured image by the mobile phone is saved in the configuration file of the video, and at the same time, the TAG of the wonderful image discarded by the mobile phone recorded in the configuration file of the video is deleted.
[0232] When the user Figure 13 clicks on the "AI One-Click Blockbuster" control on the details interface of video 502 shown in (d) of, the mobile phone responds to the user's click operation and generates a selected video of video 502. Since the configuration file of the video saves the TAG of the selected images recognized by the mobile phone and the TAG of the manually captured images. Therefore, the selected video of video 502 generated by using the method for generating a selected video provided in the foregoing Embodiment 1 includes the images manually captured by the user and the wonderful images recognized by the mobile phone and saved in the One-Take-Many-Gains folder.
[0233] Exemplarily, Figure 13 (e) of shows the display interface of the selected video 1203 of video 502. The selected image 1203 plays automatically, and the currently displayed picture is the image manually captured by the user.
[0234] In some other embodiments, the mobile phone can save the manually captured image as a new wonderful image together with the wonderful images recognized by the mobile phone into the One-Take-Many-Gains folder.
[0235] Exemplarily, the mobile phone is configured to save 5 wonderful images in the One-Take-Many-Gains folder. Figure 13 In (a) of, the user activates the manual capture function to capture an image. Figure 14 (a) of shows the One-Take-Many-Gains folder, which includes: thumbnails of 5 wonderful images recognized by the mobile phone from the captured video, and a thumbnail of the image 1202 manually captured by the user.
[0236] Figure 14 The thumbnails of the wonderful images in the One-Take-Many-Gains folder shown in (a) of can be arranged in the order of the shooting time of the wonderful images, and the thumbnail of the image manually captured by the user is located at the last position. Of course, this sorting method does not limit the sorting of the image thumbnails in the One-Take-Many-Gains folder.
[0237] In addition, Figure 14 (b) of shows the details interface of the video captured by the user. The thumbnails of the wonderful images of video 502 shown in this details interface come from Figure 14Therefore, the thumbnails of the wonderful images of video 502 include: the cover thumbnail of video 502, the thumbnails of the five wonderful images in video 502 recognized by the mobile phone, and the thumbnails of the images 1202 manually captured by the user.
[0238] It should also be noted that when a user uses the manual capture function to capture an image, the mobile phone can also set a tag for the image to indicate the position of the image in the captured video. The mobile phone can also write the tag of the manually captured image in the video configuration file.
[0239] User in Figure 14 Click the "Ai One-Click Movie" control on the details interface of video 502 shown in (b), and the mobile phone generates a selected video of video 502 in response to the user's click operation. Because the video configuration file saves the tags of the selected images recognized by the mobile phone and the tags of the manually captured images. Therefore, the selected video of video 502 generated by the method for generating selected videos provided in the above embodiment 1 includes the manually captured images of the user and the wonderful images recognized by the mobile phone and saved in the one-record-multiple-record folder.
[0240] It should also be noted that if the mobile phone fails to identify a wonderful image from the captured video, but identifies a good quality image, the mobile phone can generate a selected video using the user's manually captured images and the good quality images. Of course, the method for generating the selected video can refer to the method for generating a selected video provided in the aforementioned embodiment 1, which will not be described in detail here.
[0241] Another embodiment of the present application further provides a computer-readable storage medium, which stores instructions, and when the computer-readable storage medium is executed on a computer or a processor, the computer or the processor executes one or more steps in any of the above methods.
[0242] The computer-readable storage medium may be a non-transitory computer-readable storage medium, for example, the non-transitory computer-readable storage medium may be a ROM, a random access memory (RAM), a CD-ROM, a magnetic tape, a floppy disk, an optical data storage device, and the like.
[0243] Another embodiment of the present application further provides a computer program product including instructions. When the computer program product is run on a computer or a processor, the computer or the processor executes one or more steps in any of the above methods.
Claims
1. A video processing method, characterized in that, Applied to an electronic device, the video processing method includes: In response to an operation to control the start of video shooting, shoot a first video; In response to a second operation, display a first interface, where the first interface is a details interface of the first video, and the first interface includes a first area, a second area, and a first control, or the first interface includes the first area and the second area, or the first interface includes the first area and the first control, The first area is a playback area of the first video, The second area displays a cover thumbnail of the first video, a thumbnail of a first image, and a thumbnail of a second image. The first image is an image of the first video at a first moment, and the second image is an image of the first video at a second moment. The recording process of the first video includes the first moment and the second moment; The first control is used to control the electronic device to obtain the first video and tag data of the first video. The tag data includes the theme TAG, shot TAG, first image TAG, and second image TAG of the first video; based on the theme TAG of the first video, determine a style template and music. The style template includes at least one special effect; based on the shot TAG, the first image TAG, and the second image TAG, obtain multiple frames of images before and after the first image, and multiple frames of images before and after the second image from the first video; synthesize the special effects, music, and target images of the style template to obtain a second video; the target images at least include: the first image and multiple frames of images before and after the first image; the duration of the second video is less than that of the first video, and the second video at least includes the images of the first video. The tag data of the first video is determined by using the hierarchical information of the first video during shooting.
2. The video processing method according to claim 1, wherein It further includes: In response to a third operation, display a second interface. The third operation is a touch operation on the first control, and the second interface is a display interface of the second video.
3. The video processing method according to claim 1, wherein The step of, in response to the second operation, displaying the first interface includes: In response to a fourth operation, display a third interface. The third interface is an interface of a gallery application, and the third interface includes a cover thumbnail of the first video; In response to a touch operation on the cover thumbnail of the first video, display the first interface.
4. The video processing method according to claim 1, wherein The step of, in response to the second operation, displaying the first interface includes: In response to a touch operation on a second control, display the first interface. The shooting interface of the electronic device includes the second control, and the second control is used to control the display of the image or video captured last time.
5. The video processing method according to claim 3, wherein The cover thumbnail of the first video includes a first identifier, and the first identifier is used to indicate that the first video is shot by the electronic device in a mode of recording multiple videos at once.
6. The video processing method according to any one of claims 1 to 5, characterized in that, A mask layer is displayed on the first interface, and the second area is not covered by the mask layer.
7. The video processing method according to claim 6, wherein The first interface further includes: a first dialog box, and the first dialog box is used to prompt the user that the first image and the second image have been generated. The first dialog box is not covered by the mask layer.
8. The video processing method according to claim 2, wherein After shooting the first video in response to an operation to control the start of video shooting, the method further includes: In response to a fifth operation, display a shooting interface of the electronic device, where the shooting interface includes: a second dialog box for prompting the user that the first video and the second video have been generated.
9. The video processing method according to any one of claims 1 to 5, characterized in that, Before shooting the first video in response to an operation to control the start of video shooting, the method further includes: In response to a sixth operation, display a fourth interface, which is a shooting settings interface. The fourth interface includes: a multi-shot option and a text segment. The multi-shot option is used to control the electronic device to turn on or off the multi-shot function, and the text segment is used to indicate the function content of the multi-shot function.
10. The video processing method according to any one of claims 1 to 5, characterized in that After shooting the first video in response to an operation to control the start of video shooting, the method further includes: In response to a seventh operation, display a fifth interface, which is an interface of a gallery application. The fifth interface includes: a first folder and a second folder. The first folder includes images and the first video saved by the electronic device, and the second folder includes the first image and the second image. In response to an eighth operation, display a sixth interface, which includes thumbnails of the first image and the second image. The eighth operation is a touch operation on the second folder.
11. The video processing method according to claim 2, wherein After displaying the second interface in response to a third operation, the method further includes: In response to a ninth operation, display a seventh interface, which is a details interface of the second video. The ninth operation is a touch operation on a third control included in the second interface, and the third control is used to control the saving of the second video.
12. The video processing method according to claim 11, wherein The method further includes: In response to a tenth operation, display an eighth interface, which is an interface of a gallery application. The eighth interface includes: a cover thumbnail of the second video and a cover thumbnail of the first video.
13. The video processing method according to any one of claims 1 to 5, characterized in that, After shooting the first video in response to an operation to control the start of video shooting, the method further includes: In response to an eleventh operation, display a first shooting interface of the electronic device. The first shooting interface includes a first option for indicating a photo-taking mode and a second option for indicating a video-recording mode. In response to an operation on a fourth control of the shooting interface, display the first shooting interface of the electronic device. The fourth control is used to start photo-taking. In response to an operation on the second option, display a second shooting interface of the electronic device. The second shooting interface includes a third dialog box for indicating the function content of the multi-shot function to the user.
14. The video processing method according to any one of claims 1 to 5, characterized in that, During the process of shooting the first video in response to an operation to control the start of video shooting, the method further includes: In response to a twelfth operation, shoot and save a third image. The twelfth operation is a touch operation on the photo-taking button of the video shooting interface of the electronic device.
15. The video processing method according to claim 14, characterized in that, The second area further displays a thumbnail of the third image, and the second video includes the third image.
16. The video processing method according to claim 14, characterized in that The second area displays a cover thumbnail of the first video, thumbnails of the first image and the second image, including: The second area displays a cover thumbnail of the first video, a thumbnail of the first image, and a thumbnail of the third image; The second video at least includes the first image and the third image.
17. A video processing method, characterized in that, Applied to an electronic device, the video processing method includes: In response to an operation to control the start of video shooting, display a preview interface when shooting the first video and start shooting the first video. The preview interface when shooting the first video includes a control for shooting an image; In response to a touch operation on the control for shooting an image, shoot and save a fourth image during the process of shooting the first video; After completing the shooting of the first video, in response to a thirteenth operation, display a details interface of the first video. The details interface of the first video includes a first area, a second area, and a first control, or the details interface of the first video includes the first area and the second area, or the details interface of the first video includes the first area and the first control; The first area is a playback area of the first video. The second area displays a cover thumbnail of the first video and a thumbnail of the fourth image. The first control is used to control the electronic device to obtain the first video and tag data of the first video. The tag data includes a theme TAG, a storyboard TAG, and a fourth image TAG of the first video. Based on the theme TAG of the first video, determine a style template and music. The style template includes at least one special effect. Based on the storyboard TAG and the fourth image TAG, obtain multiple frames of images before and after the fourth image from the first video. Synthesize the special effects, music, and target images of the style template to obtain a second video. The target images at least include: the fourth image and multiple frames of images before and after the fourth image. The duration of the second video is less than that of the first video. The second video at least includes the images of the first video and the fourth image. The theme TAG and storyboard TAG of the first video are determined by using the hierarchical information of the first video during shooting.
18. The video processing method according to claim 17, wherein The second area further displays thumbnails of one or more other frames of images. The one or more other frames of images are images in the first video. The sum of the number of the fourth image and the one or more other frames of images is greater than or equal to a preset number. The preset number is the number of fifth images automatically recognized by the electronic device during the process of shooting the first video.
19. The video processing method according to claim 18, wherein The second video at least includes one or more frames of images from the following images: the fourth image, the one or more other frames of images.
20. The video processing method according to any one of claims 17 to 19, characterized in that, It further includes: In response to a touch operation on the first control, display a display interface of the second video.
21. The video processing method according to any one of claims 17 to 19, characterized in that After completing the shooting of the first video, the method further includes: Display a first shooting interface of the electronic device. The first shooting interface includes a first option and a second option. The first option is used to indicate the photo shooting mode, and the second option is used to indicate the video recording mode. The first shooting interface is a preview interface when shooting an image; In response to an operation on the first shooting interface for activating the shooting control, display the first shooting interface of the electronic device; In response to an operation on the second option, display the second shooting interface of the electronic device, where the second shooting interface includes a first dialog box for indicating to the user the function content of multi-shot recording, and the second shooting interface is a preview interface during video shooting.
22. The video processing method according to any one of claims 17 to 19, characterized in that, After the shooting of the first video is completed, it further includes: In response to a fourteenth operation, display the interface of the gallery application, where the interface of the gallery application includes: a first folder and a second folder, the first folder at least includes the fourth image, the second folder includes the fourth image and the sixth image, or the second folder includes the fourth image; In response to a touch operation on the second folder, display an interface that includes a thumbnail of the fourth image and a thumbnail of the sixth image, or includes a thumbnail of the fourth image.
23. An electronic device, characterized in that, It includes: One or more processors, a memory, a camera, and a display screen; The memory, the camera, and the display screen are coupled to the one or more processors, and the memory is used to store computer program code. The computer program code includes computer instructions. When the one or more processors execute the computer instructions, the electronic device executes the video processing method according to any one of claims 1 to 16, or the video processing method according to any one of claims 17 to 22.
24. A computer-readable storage medium, characterized in that, For storing a computer program, when the computer program is executed by the electronic device, the electronic device is enabled to implement the video processing method according to any one of claims 1 to 16, or the video processing method according to any one of claims 17 to 22.
Citation Information
Patent Citations
Method for processing video file and electronic equipment
CN111061912A